hybrid attention
- Papers
- 3
- Released code
- 0
- First seen
- Sept 2026
- Latest
- Sept 2026
3 papers in the last two months, against 0 in the two before.
The papers
Most central to this idea first, not most recent.
- Academiccs.CL
Hybrid backbones can efficiently adapt into diffusion language models
University of Texas at Austin · Sept 2026
- Chinese Techcs.LG
Shared-prefix training accelerates reinforcement learning for hybrid-attention agents
Ant Group · Sept 2026
- Big Techcs.AI
Receiver-specific cache sharing cuts multi-agent costs while improving accuracy
Columbia University, Amazon Web Services · Sept 2026
Concepts are extracted from each paper and reused across the corpus, so this page grows on its own as the desk reads.