CompanyAnthropic8 recent entries9 Aug 2026Grok Build is quickly turning into an all-in-one creation environment It can now also generate images and videos with Grok Imagine directly …Grok Build is quickly turning into an all-in-one creation environment It can now also generate images and videos with Grok Imagine directly inside your workflow You can create custom visuals for websi→10 Aug 2026When @QualiaQuanta took a shot at the Riemann, half in jest, people called her a crackpot. When Anthropic uses Claude to do the same thing, …When @QualiaQuanta took a shot at the Riemann, half in jest, people called her a crackpot. When Anthropic uses Claude to do the same thing, it gets a hundred thousand views in 30 minutes. It might wel
CompanyOpenAI8 recent entries5 Aug 2026ChatGPT Work is OpenAI's fighter in the highest stake product category in history: bringing the power of coding agents to the masses. It's a…ChatGPT Work is OpenAI's fighter in the highest stake product category in history: bringing the power of coding agents to the masses. It's also how a billion users will soon use ChatGPT by default. I →5 Aug 2026Big new release of my LLM CLI tool and Python library for talking to hundreds of different LLMs - reasoning traces, OpenAI Responses support…Big new release of my LLM CLI tool and Python library for talking to hundreds of different LLMs - reasoning traces, OpenAI Responses support, server-side tools, smarter logging and a whole lot more ht→6 Aug 2026Our Black Hat talk on the OpenAI-Hugging Face incident is now live on youtube. This is a watershed moment for the industry. I encourage all …Our Black Hat talk on the OpenAI-Hugging Face incident is now live on youtube. This is a watershed moment for the industry. I encourage all defenders to watch, consider how attack dynamics will immine→7 Aug 2026Thanks to the video from the Black Hat security conference of OpenAI's presentation about 'The Hugging Face Incident' we now have a detailed…Thanks to the video from the Black Hat security conference of OpenAI's presentation about 'The Hugging Face Incident' we now have a detailed timeline of what happened from OpenAI's perspective - I wro→8 Aug 2026Neat example here of the agents communicating purely through file names, including adding base64-encoded attachments and using 'zz' prefixes…Neat example here of the agents communicating purely through file names, including adding base64-encoded attachments and using 'zz' prefixes to ensure their new message sorts to the bottom of the list→8 Aug 2026Don't miss the bit where OpenAI first found out they were responsible for the Hugging Face attack when they reached out to HF to get one of …Don't miss the bit where OpenAI first found out they were responsible for the Hugging Face attack when they reached out to HF to get one of their credentials revoked and HF told them it had already be→10 Aug 2026We've used GPT-5.6-Cyber extensively in real-world vulnerability research, including work that uncovered previously unknown vulnerabilities …OpenAI announced the release of GPT‑5.6‑Cyber as part of its Cybersecurity Initiative, “Daybreak.” The model is aimed at advanced, authorized security research and testing, helping trusted defenders d→11 Aug 2026Now in preview: The ChatGPT desktop app for Linux. Use ChatGPT, ChatGPT Work, and Codex where you already work and build, with your projects…OpenAI released a preview of its new ChatGPT desktop application for Linux on August 11 2026. The app allows users to run ChatGPT, ChatGPT Work and Codex directly from their desktop, integrating with
CompanyGoogle8 recent entries23 Jul 2026If you are building real-time voice agents with @GoogleDeepMind Gemini Live, you can now trace your speech-to-speech agent loops directly in…If you are building real-time voice agents with @GoogleDeepMind Gemini Live, you can now trace your speech-to-speech agent loops directly in @LangChain! - Speaker callback hooks capture only the exact→23 Jul 2026Announcing OpenWorker! An open-source agent that doesn't just chat with you, but delivers finished work -- like hand you a polished document…Announcing OpenWorker! An open-source agent that doesn't just chat with you, but delivers finished work -- like hand you a polished document, send a slack message, or update a calendar entry. Ask it t→31 Jul 2026OK, GPT-5.6 Luna is a bit of a beast. Given the 80% price drop today I decided to try it in Datasette Agent, and it's furiously quick and ge…OK, GPT-5.6 Luna is a bit of a beast. Given the 80% price drop today I decided to try it in Datasette Agent, and it's furiously quick and generates all the SQL, HTML and JavaScript (for Datasette Apps→1 Aug 2026so much for the “general” part of general intelligenceso much for the “general” part of general intelligence I have tried many times to get ChatGPT, Claude, Grok or Gemini to write scripts for my YouTube videos. It is still a complete failure. For one th→1 Aug 2026> Google was too nervous to release it and DeepMind was blocked from shipping products that could disrupt Google. bookmark for the next vc t…> Google was too nervous to release it and DeepMind was blocked from shipping products that could disrupt Google. bookmark for the next vc that asks you 'what if <incumbent> builds this?' @_chenglou I→4 Aug 2026one thing i appreciate about silico is that it's a deeply humanist product. we designed silico to keep you in the experimental loop -- more …one thing i appreciate about silico is that it's a deeply humanist product. we designed silico to keep you in the experimental loop -- more observable, easier to steer, easier to understand we want to→10 Aug 2026Meta returns to open weights: Muse Glimmer, its first open-weights release since Llama 4, scores 35 on the Artificial Analysis Intelligence …Meta returns to open weights: Muse Glimmer, its first open-weights release since Llama 4, scores 35 on the Artificial Analysis Intelligence Index. It is a 30B-parameter model, and the first from Meta →10 Aug 2026Claude Haiku is my current least favorite model - it hallucinates wildly, and is out-performed now by other similarly priced models like GPT…Claude Haiku is my current least favorite model - it hallucinates wildly, and is out-performed now by other similarly priced models like GPT-5.6-Luna Even worse: it seems to still be used by the Claud
CompanyMeta8 recent entries13 Jul 2026I guess image input is the big capability of the models, and tool use can be a substitute for non-omni model output. Still, multimodal voice…Ethan Mollick notes that image input represents the primary advanced capability of current AI models, and that tool‑use can effectively replace outputs from non‑omni models. He observes that multimoda→14 Jul 2026Fable, turn my tweet into a thinkpiece (this was pretty funny): There has never been a better time to have opinions about artificial intelli…Fable, turn my tweet into a thinkpiece (this was pretty funny): There has never been a better time to have opinions about artificial intelligence. I say this with some authority, because I am currentl→27 Jul 2026Proud to be an inaugural partner alongside @NVIDIA and other industry leaders in the Open Secure AI Alliance. We look forward to acceleratin…Proud to be an inaugural partner alongside @NVIDIA and other industry leaders in the Open Secure AI Alliance. We look forward to accelerating the adoption and trust of open models and harnesses throug→28 Jul 2026New research from Meta and CMU. This one is on agentic context management for long horizon tasks. (bookmark it) Production agents accumulate…New research from Meta and CMU. This one is on agentic context management for long horizon tasks. (bookmark it) Production agents accumulate context every turn. The usual fix compresses on a token thr→9 Aug 2026New research from Meta. Agent harnesses are still mostly authored by hand. This makes it hard to tune robust agent harnesses for long-horizo…New research from Meta. Agent harnesses are still mostly authored by hand. This makes it hard to tune robust agent harnesses for long-horizon tasks. In this new work, agents learn harness policies off→10 Aug 2026Meta returns to open weights: Muse Glimmer, its first open-weights release since Llama 4, scores 35 on the Artificial Analysis Intelligence …Meta returns to open weights: Muse Glimmer, its first open-weights release since Llama 4, scores 35 on the Artificial Analysis Intelligence Index. It is a 30B-parameter model, and the first from Meta →11 Aug 2026Open source is so back. Zuck just announced Meta is opening the weights for Muse Glimmer, with Muse Spark 1.2 coming soon But a year ago, ev…Open source is so back. Zuck just announced Meta is opening the weights for Muse Glimmer, with Muse Spark 1.2 coming soon But a year ago, everyone doubted Meta's position in the AI race In an intervie→12 Aug 2026Muse Glimmer is live on Fireworks. The new open-weight model from Meta Superintelligence Labs is a 30B dense model built for always-on agent…Muse Glimmer is live on Fireworks. The new open-weight model from Meta Superintelligence Labs is a 30B dense model built for always-on agents that reason across many sequential tool calls and can reco
CompanyMistral7 recent entries15 Apr 2026Build once. Reuse everywhere. The Connectors API is now in Public Preview. Register MCP connectors once and use them across Le Chat, AI Stud…Mistral AI has launched its Connectors API into Public Preview, enabling developers to register Model Context Protocol (MCP) connectors once and deploy them across multiple Mistral products. This 'bui→28 Apr 2026🆕 Today, we're releasing the public preview of Workflows, the orchestration layer for enterprise AI. 🌎 Enterprise teams have capable model…🆕 Today, we're releasing the public preview of Workflows, the orchestration layer for enterprise AI. 🌎 Enterprise teams have capable models. What they don't have is a way to run them reliably in produ→15 May 2026'do you want meet in the Devin train car?' 'oh actually I'm in the Langsmith car near the front' 'Let's just meet in front of the Mistral wa…This appears to be a humorous exchange from X/Twitter creator Harrison Chase referencing train cars named after AI/ML tools and services (Devin, Langsmith, and Mistral), likely making a joke about the→28 May 2026SHIPPED. Mistral Vibe is now the AI agent for long-horizon productivity and coding, and the home for Work mode, Code mode, the CLI, and a br…Mistral AI has released Mistral Vibe, an AI agent designed for long-horizon productivity and coding tasks, featuring Work mode, Code mode, a CLI, and additional capabilities. The product consolidates →9 Jul 2026https://ollama.com/blog/all-aboard-open-modelsOllama announced support for running multiple open-source language models locally, emphasizing accessibility and ease of deployment for users who want to use AI models without relying on cloud service→26 Jul 2026Yes, open-source / open-weight models are important for a healthy AI ecosystem. That's how we can verify things, check claims, and keep up o…Yes, open-source / open-weight models are important for a healthy AI ecosystem. That's how we can verify things, check claims, and keep up outside the closed labs. Plus, it gives us the freedom to run→11 Aug 2026💡The world needs an open-source platform, and that’s exactly what we’re building to give our customers more choice and the flexibility to c…💡The world needs an open-source platform, and that’s exactly what we’re building to give our customers more choice and the flexibility to choose the right model for the right task. As part of this, we
CompanyxAI8 recent entries11 Jul 2026Free trial of Grok 4.5 model via Grok Build CLIElon Musk announced the availability of a free trial for the Grok 4.5 model through the Grok Build CLI, allowing developers to access and test the latest version of xAI's Grok language model. This ann→15 Jul 2026Grok Build is now fully open source, and SpaceXAI has made its privacy-first approach even stronger SpaceXAI has officially: • Open-sourced …Grok Build is now fully open source, and SpaceXAI has made its privacy-first approach even stronger SpaceXAI has officially: • Open-sourced the Grok Build harness CLI • Reset usage limits for all user→29 Jul 2026BREAKING: Grok 4.5 ranked #1 on LaurenBench with a score of 56.9%, ahead of Claude Sonnet 5, GLM 5.2, Claude Opus 5, Kimi K3 and GPT-5.6. Th…BREAKING: Grok 4.5 ranked #1 on LaurenBench with a score of 56.9%, ahead of Claude Sonnet 5, GLM 5.2, Claude Opus 5, Kimi K3 and GPT-5.6. The benchmark tests real-world AI agents across conversation, →30 Jul 2026And Grok 4.6 comes out in a weekAnd Grok 4.6 comes out in a week BREAKING: Grok 4.5 ranked #1 on LaurenBench with a score of 56.9%, ahead of Claude Sonnet 5, GLM 5.2, Claude Opus 5, Kimi K3 and GPT-5.6. The benchmark tests real-worl→1 Aug 2026so much for the “general” part of general intelligenceso much for the “general” part of general intelligence I have tried many times to get ChatGPT, Claude, Grok or Gemini to write scripts for my YouTube videos. It is still a complete failure. For one th→7 Aug 2026good grokgood grok Best match for this hierarchical hands-free setup: - Runtime: ActiveGraph (event-sourced log as source of truth) or LangGraph for supervisor/manager graphs - Roles as skills: Claude Agent SD→8 Aug 2026Imagine image 2.0, non-agentic yet, more to come in a week or two 💙Imagine image 2.0, non-agentic yet, more to come in a week or two 💙 Announcing Imagine Image 2.0, our next generation image model with precision editing, crisp text rendering, improved factuality, and→9 Aug 2026Grok Build is quickly turning into an all-in-one creation environment It can now also generate images and videos with Grok Imagine directly …Grok Build is quickly turning into an all-in-one creation environment It can now also generate images and videos with Grok Imagine directly inside your workflow You can create custom visuals for websi
CompanyDeepSeek8 recent entries26 Jul 2026Yes, open-source / open-weight models are important for a healthy AI ecosystem. That's how we can verify things, check claims, and keep up o…Yes, open-source / open-weight models are important for a healthy AI ecosystem. That's how we can verify things, check claims, and keep up outside the closed labs. Plus, it gives us the freedom to run→3 Aug 2026two weeks ago i went on @swyx's pod and said some things that i... should not have said. a lot has happened since then, i owe you all an apo…two weeks ago i went on @swyx's pod and said some things that i... should not have said. a lot has happened since then, i owe you all an apology. i'm sorry that i was right about every single thing. a→4 Aug 2026Together AI gives developers a high-throughput production path for running DeepSeek V4 Flash across coding, tool-use, and agentic workloads.…Together AI gives developers a high-throughput production path for running DeepSeek V4 Flash across coding, tool-use, and agentic workloads. Start building: https://www.together.ai/models/deepseek-v4-→4 Aug 2026DeepSeek V4 Flash is now live on Together AI. Frontier agent performance is getting dramatically cheaper. DSV4 on Together AI brings a major…DeepSeek V4 Flash is now live on Together AI. Frontier agent performance is getting dramatically cheaper. DSV4 on Together AI brings a major jump in coding, tool use, and long-running agent performanc→6 Aug 2026Together Serverless Inference gives developers a managed production path for bringing FLUX 3 into creative products and automated media work…Together Serverless Inference gives developers a managed production path for bringing FLUX 3 into creative products and automated media workflows. Start building: https://www.together.ai/models/flux-3→6 Aug 2026Open models give teams more room to run agent loops, with API prices at a fraction of GPT-5.6 Sol and Claude Fable 5. That matters as planni…Open models give teams more room to run agent loops, with API prices at a fraction of GPT-5.6 Sol and Claude Fable 5. That matters as planning, tool calls, retries, and long contexts compound token us→11 Aug 2026we recently trimmed the deepagents harness base prompt by 65% (including tool info) it shows — deepagents is cheap!we recently trimmed the deepagents harness base prompt by 65% (including tool info) it shows — deepagents is cheap! We ran DeepSeek V4 Flash through 4 more agent harnesses (Hermes Agent, Pi Agent, Pri→11 Aug 2026DeepSeek V4 Flash 0731 is now available to fine-tune on Together AI. Specialize it for coding, tool use, and your own domain with SFT or DPO…DeepSeek V4 Flash 0731 is now available to fine-tune on Together AI. Specialize it for coding, tool use, and your own domain with SFT or DPO, then deploy the fine-tuned model on Together AI for produc
CompanyNVIDIA8 recent entries31 Jul 2026Security at the foundation requires openness at the foundation. As open-weight models become critical infrastructure for the next generation…Security at the foundation requires openness at the foundation. As open-weight models become critical infrastructure for the next generation of software, the systems around them must be transparent, i→2 Aug 2026Watching @ClementDelangue on @FaceTheNation discussing agentic hacking. “Preventing releases does not work; concentrating behind closed door…Watching @ClementDelangue on @FaceTheNation discussing agentic hacking. “Preventing releases does not work; concentrating behind closed doors in just a few organizations doesnt work. What worked in th→2 Aug 2026Hermes Agent is now dramatically more efficient, especially for smaller/weaker/local models! With the help of @nvidia's Nemo Relay and sever…Hermes Agent is now dramatically more efficient, especially for smaller/weaker/local models! With the help of @nvidia's Nemo Relay and several other strategies Hermes was able to identify a ton of opt→3 Aug 2026two weeks ago i went on @swyx's pod and said some things that i... should not have said. a lot has happened since then, i owe you all an apo…two weeks ago i went on @swyx's pod and said some things that i... should not have said. a lot has happened since then, i owe you all an apology. i'm sorry that i was right about every single thing. a→5 Aug 2026None of this was us. Day 1 (and the first 48 hours) belonged to the open-source community: Generate with it @ComfyUI — native support + offi…None of this was us. Day 1 (and the first 48 hours) belonged to the open-source community: Generate with it @ComfyUI — native support + official quantized builds, Day 0 Diffusers — the reference Pytho→6 Aug 2026Together Serverless Inference gives developers a managed production path for bringing FLUX 3 into creative products and automated media work…Together Serverless Inference gives developers a managed production path for bringing FLUX 3 into creative products and automated media workflows. Start building: https://www.together.ai/models/flux-3→10 Aug 2026Meta returns to open weights: Muse Glimmer, its first open-weights release since Llama 4, scores 35 on the Artificial Analysis Intelligence …Meta returns to open weights: Muse Glimmer, its first open-weights release since Llama 4, scores 35 on the Artificial Analysis Intelligence Index. It is a 30B-parameter model, and the first from Meta →11 Aug 2026NVIDIA Nemotron 3.5 Lighting is available on Ollama! It's a 30B model made for always-on agents. All local. Claude Code ollama launch claude…NVIDIA Nemotron 3.5 Lighting is available on Ollama! It's a 30B model made for always-on agents. All local. Claude Code ollama launch claude --model nemotron-3.5-lightning Hermes Agent ollama launch h