Model techniques map
Taxonomymodel architecturetoken mixerlinear attention & state space

taxonomy node · level 3

linear attention & state space

16 methods filed at this node or below it, from the sources of 9 models.

model architecture :: token mixer :: linear attention & state space

Matching aids for the classifier: SSM; linear attention; lightning attention.

In this branch 16

Everything filed at this node or below it, with one collapsible heading per child node.

filed here 2

Linear attention core · 2 sources · 2 quotes
Lightning Attention not used · 1 source · 1 quote

Mamba 3

Mamba-2 core · 5 sources · 6 quotes
Mamba core · 2 sources · 2 quotes
Mamba-2 SSM cache optional · 1 source · 1 quote

gated delta network 11

Gated DeltaNet core · 13 sources · 13 quotes
Kimi Delta Attention core · 6 sources · 6 quotes
Gated DeltaNet–sparse MoE hybrid core · 2 sources · 2 quotes
KDA core · 2 sources · 2 quotes
Hybrid linear attention core · 1 source · 1 quote
SimpleGDN evaluated · 1 source · 1 quote

By model

Which of this branch's techniques each model's own documents describe, and how strongly. Under each model: its strongest treatment anywhere in the branch.