You are browsing as a guest. Sign up (or log in) to start making projects!

Mock Modular Attention

  • 0 Devlogs
  • 0 Total hours

The Idea was to create an alternative complete attention mechanism that could act as a drop in replacement for softmax. It replaces stochastic, unconstrained heuristic logits with deterministic kernel weightings derived from Ramanujan's third-order mock theta functions and classical q -series and introduces an approximate modular-symmetry inductive bias directly into the forward pass of Transformer architectures.

Delete project?

Are you sure you want to permanently delete this project? This action cannot be undone.

All devlogs, followers, and associated data will be removed.

Followers

Loading…