Implementation of the Transformer variant proposed in "Transformer Quality in Linear Time" - View it on GitHub
Star
362
Rank
95840