transformer-tricks

0.3.4
23.47k

A collection of tricks to speed up LLMs, see our transformer-tricks papers on arXiv