Model Releases
We built a local AI work tool that runs Qwen3.5-35B-A3B on a 16GB Mac
Hi everyone! I’m an intern at Icosa, a startup focused on making local AI accessible. We just released the first version of Zeno, our local AI product, and I thought it might be of interest to this co
Hi everyone! I’m an intern at Icosa, a startup focused on making local AI accessible. We just released the first version of Zeno, our local AI product, and I thought it might be of interest to this community. Zeno is a free agentic AI work tool that runs fully on your Mac - similar to Claude Cowork, but obviously with a smaller model so it can run locally. It works with your files, keeps your data private, and has no usage bills or watermarks. It ships with 4-bit Qwen3.5-35B-A3B. The full model doesn’t fit entirely in the unified memory of a 16GB Mac, so rather than shrinking or pruning it, we built an offloading system. (We’ll share more details on that soon.) We’d love to hear what you think - use cases, performance, bugs, complaints, anything. This is our first version, and it might be a little slow at first. This is our first week, and we'll keep making it better as we learn from people using it. It works best on Macs with 16GB of memory or more. Below are some recent test results. https://preview.redd.it/5oiiu8dcpqkh1.png?width=1178&format=png&auto=webp&s=69e036be60458da823f80c2b20dbd61ccae847fc Download: https://www.icosa.co/zeno Icosa Discord: https://discord.gg/bQRqHRpam submitted by /u/close_Meal6005 [link] [comments]
Related
- Built a Mac app that actually uses Ollama for work, not just chat — with a calibration step for models with flaky tool calling
- Ollama works locally, but my coding-agent integration does not
- claudely: launch Claude Code against Local LLM provider like LM Studio / Ollama / llama.cpp without trashing your real claude config
Source: r/ollama | 2026-08-21