Atri RudraFlashAttention: Fast and Memory-Efficient Exact Attention with IO-AwarenessHiPPO: Recurrent Memory with Optimal Polynomial ProjectionsArithmetic Circuits, Structured Matrices and (not so) Deep LearningFlashAttention (GitHub)All names