A Survey of Efficient Attention Methods: Hardware-Efficient, Sparse, Compact, and Linear Attention
Published in ArXiv preprint, 2025
Chunyu led the linear-attention section, analyzing representative architectures and their quality–efficiency trade-offs.
Recommended citation: Zhang et al., 2025.
Download Paper