ELiTA: Linear-Time Attention Done Right
ELiTA: Linear-Time Attention Done Right
A novel Transformer architecture that is much cheaper and faster, while matching and outperforming the standard. Sequence Lengths of 100K+ on 1 GPU. Intuition, evaluation and code available on repository.
Share cardActual performance
Launch Intel predictions
Analyze your own launch →Correct prediction on native model
Similar products
I Wrote an Introduction to Linear Regression in Haskell
Outcrop – Linear for Wikis
I've Translated Linear Algebra Done Right by Axler to Hebrew
Randomize standup order, right inside Linear
Armacmp – compile R linear algebra code to C++
reclaim your time and attention
Pytorch implementation of seq2seq model with attention and pointer
Cybersalience – Guiding user attention using transformer attention
A Rigorous Proofless Approach to Linear Algebra [pdf]
FOMO – See what your community pays attention to