Are 1B LLMs Going Away in 2026?
Mirrored from r/LocalLLaMA for archival readability. Support the source by reading on the original site.
I don't know much about llms aside from downloading them through a frontend and running them on my laptop or potato phone.
Google released gemma 4, but unlike gemma 3, there isn't a 1b model this time.
Llama also had a 1b model before, but there doesn't seem to be a new one.
Qwen 3.5 had a 1b (0.8b) class model too, but the latest qwen releases don't seem to be targeting the 1b range anymore.
From my limited experience, gemma 3 1b is still probably the best 1b llm overall. It has good tokens per second, and while there are some nice distilled and finetuned models based on older 1b gemma and qwen models, there doesn't seem to be much that's actually new in this size range.
Bonsai has ternary llms, but in practice i found them to hallucinate a lot and be less reliable than regular llms.
So have ai companies mostly moved away from 1b llms in 2026? Or are they still releasing them and i am just not aware of it?
[link] [comments]
Discussion (0)
Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.
Sign in →No comments yet. Sign in and be the first to say something.