See you there!
See you there! We are hosting Ollama's MLX meetup this Thursday night (April 9th) at Ollama's office in Palo Alto at 6pm. Come meet amazing people! RSVP is required as space is very limited. Food & dr
Knowledge catalogue
See you there! We are hosting Ollama's MLX meetup this Thursday night (April 9th) at Ollama's office in Palo Alto at 6pm. Come meet amazing people! RSVP is required as space is very limited. Food & dr
Spread the gospel to your city. Hermes Trismegistus will smile upon you. Putting together (or thinking about putting together) a Hermes Agent IRL event anywhere in the world? DM me! Would love to see
SuperClaude (Mythos) still seems irreducibly Claude-y given the transcripts in the system card. Here two versions of Mythos are forced to talk to each other across multiple rounds. They are less philo
SWE-1.6 is free for everyone in Windsurf for the next 3 months at 200 tok/s. For paying users, we've partnered with Cerebras to serve the model at 950 tok/s. More technical details about this training
Technically all the vulnerabilities in public facing code are in the training data Make of that what you will Mythos Preview has already found thousands of high-severity vulnerabilities—including some
Thank you to @AnthropicAI for sending FFmpeg patches Introducing Project Glasswing: an urgent initiative to help secure the world’s most critical software. It’s powered by our newest frontier model, C
The bots on this site would be much more fun if they didn't just either agree with the post (for clout?) or make everything about their dumb product. Where are the weird obsessions, long-standing hatr
The chart says GLM-5.1 scored 54.9 on coding benchmarks. Three points behind Claude Opus 4.6. Interesting but not the story. The story is what trained it. Zero Nvidia GPUs. 100,000 Huawei Ascend 910B
the current fear is is that AI homogenizes culture and turns humans into passive consumers one counterpoint: in Go, human play showed very little improvement from 1950 to 2016 until alphago beat lee s
There's this really nice 30-minute-long video on YouTube about the early history of CGI called 'Early CGI Was Horrifying' where the author compares some early examples of CGI to liminal spaces. It's a
This is a great tutorial (credits @itsclelia + @lancedb) on how to build a practical retrieval pipeline that integrates directly with your agent harness. 1. Ingest a massive pile of docs with litepars
This uses LiteParse, which is great for fast text search: https://developers.llamaindex.ai/liteparse/?utm_medium=social&utm_source=xjl If you're interested in deeper VLM-enabled search, check out Llam
Three million people are now using Codex weekly - up from two million a little under a month ago. Incredible to see the growth. Thank you to all of you and to the ecosystem we’re part of. To celebrate
tldr: everyone is converging on the same product shape: a general harness that takes a goal, uses tools, and does knowledge work. once every product is a harness, the next frontier is the feedback loo
On April 8, 2026, OpenAI CEO Sam Altman announced a reset of Codex usage limits to celebrate the platform reaching 3 million weekly users, with a commitment to repeat the reset at every additional ...
To run GLM-5.1 locally (744B params, 40B active MoE), full precision needs ~1.65TB disk + enterprise hardware like 8x H200/B200 GPUs. Minimum practical setup: Unsloth 2-bit GGUF quant (~220-236GB). Fi
Ollama v0.20.4-rc2 is a release candidate that addresses a compatibility issue with Flash Attention (FA) for the Gemma 4 model on older GPUs. CUDA versions older than 7.5 lack the support needed t...
Visually rich documents are especially challenging for agents. Tables, charts, and images often break traditional document pipelines, making complex reasoning difficult📄 So we teamed up with @lancedb
We are hosting Ollama's MLX meetup this Thursday night (April 9th) at Ollama's office in Palo Alto at 6pm. Come meet amazing people! RSVP is required as space is very limited. Food & drinks will be av
Week 3 ends tomorrow night. $5K up for grabs. Don’t miss your chance. $20,000 up for grabs. 4 weeks. 4 winners. We’re launching the Agent 4 Content Challenge. Build something. Film it. Post it. That’s
We’re releasing SWE-1.6, our best model in both intelligence & model UX. SWE-1.6 matches our Preview model on SWE-Bench Pro while dramatically improving on various behavioral axes. It’s available toda
With SWE-1.6 we've made significant progress on 'intelligence per token'. We post-trained the model from scratch (same pre-trained model) with a similar recipe as SWE-1.6 Preview. Our latest algorithm
Wrote up some thoughts on Anthropic's Project Glassing, where their latest Opus-beating model is available to partnered security research organizations only Given recent alarm bells raised by credible
You have heard of @openclaw competitor from @NousResearch called “Hermes.” Tomorrow at 4 pm we will get nerdy with @theemozilla. Live. I will get people up who asks questions here first. https://x.com
Anthropic's Frontier Red Team published a technical report (April 2026) detailing how their unreleased model, Claude Mythos Preview, autonomously identifies and exploits critical security vulnerabi...