Model Releases
Are 1B LLMs Going Away in 2026?
I don't know much about llms aside from downloading them through a frontend and running them on my laptop or potato phone. Google released gemma 4, but unlike gemma 3, there isn't a 1b model this time
I don't know much about llms aside from downloading them through a frontend and running them on my laptop or potato phone. Google released gemma 4, but unlike gemma 3, there isn't a 1b model this time. Llama also had a 1b model before, but there doesn't seem to be a new one. Qwen 3.5 had a 1b (0.8b) class model too, but the latest qwen releases don't seem to be targeting the 1b range anymore. From my limited experience, gemma 3 1b is still probably the best 1b llm overall. It has good tokens per second, and while there are some nice distilled and finetuned models based on older 1b gemma and qwen models, there doesn't seem to be much that's actually new in this size range. Bonsai has ternary llms, but in practice i found them to hallucinate a lot and be less reliable than regular llms. So have ai companies mostly moved away from 1b llms in 2026? Or are they still releasing them and i am just not aware of it? submitted by /u/winter-m00n [link] [comments]
Related
- Current smallest usable coding model
- Cactus Hybrid: We taught Gemma 4 to know when it's wrong
- 23 Gemma4-E4B models compared with abliterlitics: the most downloaded one is also the most broken
- I pre-trained a 700m on 18B tokens optimized for Python and Wikitext | TheOneWhoWill/Shibai-700M-Base · Hugging Face
Source: r/LocalLLaMA | 2026-08-01