Skip to content
r/LocalLLaMA · Communities

A.L.F.R.E.D. – 2B models with template can match 35B Models with 4x less speed

https://preview.redd.it/qd7drw8mqfdh1.png?width=2223&format=png&auto=webp&s=4df7e6f4002cf6c6508463f40ebdb6a39c6d1805 https://preview.redd.it/86s7i9dpqfdh1.png?width=2411&format=png&auto=webp&s=3f69ccc368fe0212fb4392f14b0d60e38261e186 So My idea is to instead of spending bigger models compute on easier tasks, why we jus