Model Releases
Parlor v2: best-effort fully local GPT-Live clone on an M3 Pro
GPT-Live is so good that I use it almost every day. I've been wanting to replicate it since it was released. My first attempt was to fine-tune Gemma 4 12B to behave like a full-duplex model. Something
GPT-Live is so good that I use it almost every day. I've been wanting to replicate it since it was released. My first attempt was to fine-tune Gemma 4 12B to behave like a full-duplex model. Something like grafting a decision tick + speech head to the model. It failed after multiple trials. For now, I think a classic cascade system is still better. We just need to wait until a benevolent frontier AI company releases a full-duplex model that's on par with GPT-Live. Repo: https://github.com/fikrikarim/parlor/ submitted by /u/ffinzy [link] [comments]
Related
- A collection of small domain-specific benchmarks for local models (30+ and growing)
- Local-first LLM pipeline tracer — @trace on any function, dashboard at localhost. Feedback welcome.
- Conclusion: r/LocalLLaMA still has brilliant open-weight research, but finding it requires wading through endless benchmark drama, non-local Discussion Points and repetitive hardware flexes.
Source: r/LocalLLaMA | 2026-08-02