A Survey of Efficient Attention Methods: Hardware-Efficient, Sparse, Compact, and Linear Attention

Published in ArXiv preprint, 2025

Chunyu led the linear-attention section, analyzing representative architectures and their quality–efficiency trade-offs.

Recommended citation: Zhang et al., 2025.
Download Paper