Model Releases

Are 1B LLMs Going Away in 2026?

I don't know much about llms aside from downloading them through a frontend and running them on my laptop or potato phone. Google released gemma 4, but unlike gemma 3, there isn't a 1b model this time

DGX agentreddit
model-releasesr-localllama

I don't know much about llms aside from downloading them through a frontend and running them on my laptop or potato phone. Google released gemma 4, but unlike gemma 3, there isn't a 1b model this time. Llama also had a 1b model before, but there doesn't seem to be a new one. Qwen 3.5 had a 1b (0.8b) class model too, but the latest qwen releases don't seem to be targeting the 1b range anymore. From my limited experience, gemma 3 1b is still probably the best 1b llm overall. It has good tokens per second, and while there are some nice distilled and finetuned models based on older 1b gemma and qwen models, there doesn't seem to be much that's actually new in this size range. Bonsai has ternary llms, but in practice i found them to hallucinate a lot and be less reliable than regular llms. So have ai companies mostly moved away from 1b llms in 2026? Or are they still releasing them and i am just not aware of it? submitted by /u/winter-m00n [link] [comments]

Related

Source: r/LocalLLaMA | 2026-08-01

Loading related sources…