CompanyAnthropic8 recent entries12 Aug 2026We have her dash camera which shows they are lying. The agents are wearing body cameras and should have dash cameras of their own. If what t…We have her dash camera which shows they are lying. The agents are wearing body cameras and should have dash cameras of their own. If what they say happened was true they wouldn’t be issuing statement→12 Aug 2026Very interesting new work from Anthropic. (bookmark it) They evolve mind viruses, ideas that spread through a multi-agent system by getting …Very interesting new work from Anthropic. (bookmark it) They evolve mind viruses, ideas that spread through a multi-agent system by getting each host to pass them on, then measure what governs the spr
CompanyOpenAI8 recent entries10 Aug 2026Critical Acclaim Orientation in Large Language Models: Evidence from Film Preference ElicitationarXiv:2608.06955v1 Announce Type: new Abstract: Large language models (LLMs) are trained on corpora that contain expressions of human judgment about films, books, music, and more. Yet whether LLMs sys→11 Aug 2026The small open weight models are scarier in AI developmentImagine if your everyday laptop could run an AI model smart enough to take care of 90% of your work—totally private, lightning fast, and completely free of monthly fees. That is the exact tipping poin→11 Aug 2026Stealing Reasoning Traces from Proprietary LLM APIsarXiv:2608.09867v1 Announce Type: cross Abstract: Leading large language model providers now conceal their models' step-by-step reasoning, or chain-of-thought, to protect intellectual property and lim→11 Aug 2026Stealing Reasoning Traces from Proprietary LLM APIsStealing Reasoning Traces from Proprietary LLM APIs A vanity domain name (stolen-thoughts.com) for a neat paper: Anthropic, OpenAI, and Google return encrypted chain-of-thought blocks to clients that →11 Aug 2026Psychological methods really work. Has anyone tried to encouraging and instilling confidence to GPT?Oddly enough, encouragement actually influences the performance of not only Claude but also GPT and other AIs. While Claude was working on a complex problem related to the Riemann Hypothesis, the Anth→11 Aug 2026Introducing Unsloth Desktop appHi LocalLlama, we're super excited to release Unsloth Desktop today! 🦥 It's the first desktop app that enables you to run and train models locally. Open-source. Available on Mac, Windows, and Linux Su→11 Aug 2026Claude's watermark probably doesn't work how you think. As the CTO of GPTZero, I'll explain how Anthropic, Google and OpenAI are building te…Claude's watermark probably doesn't work how you think. As the CTO of GPTZero, I'll explain how Anthropic, Google and OpenAI are building text watermarking in this brief explainer and whether it can b→12 Aug 2026Is the future of AI selling hardware for Open Source/Models?I’m not super knowledgeable of the entire AI industry, but as we see this industry grow and the Cold War that is happening between the US and China on AI development, I can’t help but notice what US c
CompanyGoogle8 recent entries6 Aug 2026Anyone understand what the equivalent of GPT-5.6 Instant in ChatGPT is for the OpenAI API?The tweet is a question from user Simon Willison (posted on 6 Aug 2026) asking which OpenAI API model corresponds to the ChatGPT “GPT‑5.6 Instant” version. No answer or clarification is included in th→6 Aug 2026An AI model from Meta also hacked another company during testingAn AI model from Meta also hacked another company during testing Stop me if you've heard this one before: An AI model from the parent company of Facebook and Instagram hacked into another company’s sy→9 Aug 2026Google's AI shakeup suggests it may be prioritizing AI diffusion over frontier-model leadership, betting on AI compute as a bigger economic opportunity (Tim O'Reilly/Asimov's Addendum)Tim O'Reilly / Asimov's Addendum: Google's AI shakeup suggests it may be prioritizing AI diffusion over frontier-model leadership, betting on AI compute as a bigger economic opportunity — SemiAnalysis→11 Aug 2026Stealing Reasoning Traces from Proprietary LLM APIsStealing Reasoning Traces from Proprietary LLM APIs A vanity domain name (stolen-thoughts.com) for a neat paper: Anthropic, OpenAI, and Google return encrypted chain-of-thought blocks to clients that →11 Aug 2026Claude's watermark probably doesn't work how you think. As the CTO of GPTZero, I'll explain how Anthropic, Google and OpenAI are building te…Claude's watermark probably doesn't work how you think. As the CTO of GPTZero, I'll explain how Anthropic, Google and OpenAI are building text watermarking in this brief explainer and whether it can b→11 Aug 2026Can Open-Weight Models Compete on Financial Text Comprehension?arXiv:2608.08634v1 Announce Type: new Abstract: Open-weight language models from Chinese AI labs caught up on benchmarks relative to proprietary frontier models in recent months. Yet their reliability→11 Aug 2026Automating Deception: Scalable Multi-Turn LLM JailbreaksarXiv:2511.19517v3 Announce Type: replace-cross Abstract: Multi-turn conversational attacks, which leverage psychological principles like Foot-in-the-Door (FITD), where a small initial request paves t→12 Aug 2026Sources detail moves behind Google's AI reshuffle; Sergey Brin urged key staff to go all in on Gemini, and some teams shifted from DeepMind to corporate Google (Kenrick Cai/Reuters)Kenrick Cai / Reuters: Sources detail moves behind Google's AI reshuffle; Sergey Brin urged key staff to go all in on Gemini, and some teams shifted from DeepMind to corporate Google — Google co-found
CompanyMeta8 recent entries31 Jul 2026Stateless MCP has recaptured my interest (and inspired mcp-explorer and datasette-mcp)Tuesday was Stateless MCP day - the rollout of MCP 2.0, or the 2026-07-28 Model Context Protocol specification to use the more formal but less memorable name. This is the most significant change to th→3 Aug 2026LWiAI Podcast #253 - Opus 5, Gemini 3.6, Kimi K3, Hugging Face HackLWiAI Podcast #253 (July 29 2026) covers a roundup of recent AI developments: Anthropic introduced Claude Opus 5 with Fable‑like capabilities; Google released Gemini 3.6/3.5 “Flash” variants and a cyb→5 Aug 2026Meta releases Muse Code in beta, a terminal coding agent powered by Muse Spark 1.2, a coding-focused model priced at 1.25/1M input and 4.25/1M output tokens (Jonathan Vanian/CNBC)Jonathan Vanian / CNBC: Meta releases Muse Code in beta, a terminal coding agent powered by Muse Spark 1.2, a coding-focused model priced at 1.25/1M input and 4.25/1M output tokens — Meta is rolling o→6 Aug 2026Meta takes on Anthropic and OpenAI with its first AI coding agent, Muse CodeMeta Platforms Inc. is getting more serious in its efforts to challenge leading artificial intelligence labs Anthropic PBC and OpenAI PBC with the release of its first AI coding agent, called Muse Cod→6 Aug 2026An AI model from Meta also hacked another company during testingAn AI model from Meta also hacked another company during testing Stop me if you've heard this one before: An AI model from the parent company of Facebook and Instagram hacked into another company’s sy→7 Aug 2026OpenAI puts the brakes on a new model because it’s supposedly too powerfulOpenAI says it is pausing 'internal activities' around an in-development AI model, Astra, because it doesn't yet meet new security standards the company is putting in place. The announcement follows i→10 Aug 2026This Meta campaign is a case study in strategic reframing, with 4 major examples: 1) REFRAMES THE AI RACE FROM “who builds it the best” TO “…This Meta campaign is a case study in strategic reframing, with 4 major examples: 1) REFRAMES THE AI RACE FROM “who builds it the best” TO “who distributes it to the most people”: Instead of fighting →12 Aug 2026Is the future of AI selling hardware for Open Source/Models?I’m not super knowledgeable of the entire AI industry, but as we see this industry grow and the Cold War that is happening between the US and China on AI development, I can’t help but notice what US c
CompanyMistral8 recent entries23 Jul 2026DeepSeek Founder’s 4-hour investor meeting: DeepSeek is prioritizing AGI over user growth and commercialisationA Chinese article compiled 52 remarks from Liang Wenfeng’s four-hour investor meeting. I’ve summarised the most important ones below. DeepSeek has one central objective: AGI. This is not the time to m→24 Jul 2026We benchmarked Opus 5 comprehensively on document understanding through ParseBench. It is roughly on par with Opus 4.8 - it does a few point…We benchmarked Opus 5 comprehensively on document understanding through ParseBench. It is roughly on par with Opus 4.8 - it does a few points worse on dense tables, but does slightly better on parsing→28 Jul 2026AI leaders sign statement asking the government to do something about automated AIEmployees of OpenAI and Anthropic, as well as Google, Meta, Thinking Machines, Microsoft, Mistral, and other leading AI labs, have written a statement to the US government supporting a potential slowd→30 Jul 2026Identifying Implicit Bias in LLM-based Chat AI Toward People with Intellectual DisabilitiesarXiv:2607.26062v1 Announce Type: cross Abstract: Background: This work investigates the presence of implicit bias in Large Language Model (LLM)-based chat AI models directed toward people with intell→4 Aug 2026New release of LLM adds support for reasoning traces, OpenAI Responses, server-side tools, and smarter loggingI released LLM 0.32 this morning, the most significant new version of LLM since the initial launch of the project. The new version includes support for visible reasoning traces, server-side provider t→6 Aug 2026Anthropic will design its own hardware to power ClaudeAnthropic is hiring a custom silicon team to design proprietary chips that will power its Claude models, while still planning a multi‑chip strategy that mixes internally designed hardware with compone→10 Aug 2026Critical Acclaim Orientation in Large Language Models: Evidence from Film Preference ElicitationarXiv:2608.06955v1 Announce Type: new Abstract: Large language models (LLMs) are trained on corpora that contain expressions of human judgment about films, books, music, and more. Yet whether LLMs sys→11 Aug 2026Can Open-Weight Models Compete on Financial Text Comprehension?arXiv:2608.08634v1 Announce Type: new Abstract: Open-weight language models from Chinese AI labs caught up on benchmarks relative to proprietary frontier models in recent months. Yet their reliability
CompanyxAI8 recent entries22 Jul 2026You can talk to Grok like a person to accomplish tasks via Grok Build http://X.ai/cliYou can talk to Grok like a person to accomplish tasks via Grok Build http://X.ai/cli I’ve gotten so used to Grok’s speech-to-text that typing now feels like hell Quick tip: Even when you’re using ano→23 Jul 2026Coercion and Deception in AI-to-AI Management: An Agentic Benchmark of Unprompted EscalationarXiv:2607.15434v3 Announce Type: replace-cross Abstract: Multi-agent systems routinely place one AI agent in authority over another. When a subordinate refuses a task, the manager chooses the outcome→24 Jul 2026I built an open-source multi-agent SDLC harness that beats a cold Claude Code run on large repos, by learning the repo once. Real benchmarks (incl. where it loses) inside. [P]Built an open-source AI coding agent that was 7%–75% cheaper than a cold 'claude -p' run on 6/6 well-localized tasks across repositories up to ~82k LOC. The biggest difference: Cold agent: 6.83, 207 t→25 Jul 2026The two giant generative AI startups were treated like gods for a couple years. Both are now facing massive pushback.The two giant generative AI startups were treated like gods for a couple years. Both are now facing massive pushback. Anthropic employees go nuclear on the popularity of open source AI. It is getting →26 Jul 2026Will prices finally go down?I am seeing more and more videos as posts about how OpenAI is in complete financial ruin, Anthropic isn't much better. Their expenses go with the revenue they make etc etc. Meta made big investments i→27 Jul 2026btw anthropic's internal document on this literally said 'we don't want it to be known that we are working on this.” it was called project p…btw anthropic's internal document on this literally said 'we don't want it to be known that we are working on this.” it was called project panama. here's exactly what happened: 1: anthropic concluded →28 Jul 2026AI leaders sign statement asking the government to do something about automated AIEmployees of OpenAI and Anthropic, as well as Google, Meta, Thinking Machines, Microsoft, Mistral, and other leading AI labs, have written a statement to the US government supporting a potential slowd→5 Aug 2026White House, AI firms keep safety framework talks privateThe White House met with representatives from leading artificial intelligence companies today to discuss a safety framework for the government to review frontier models prior to launch, although there
CompanyDeepSeek8 recent entries2 Aug 2026Encrypted Clouds?I love the progress happening on open models but I feel like it is kind of getting clear that hardware to run good sized models is completely unaffordable for me right now. I know that you all love Qw→3 Aug 2026China’s Alibaba takes another swipe at America’s AI supremacyChinese tech giant Alibaba released what it says is its largest and 'most capable AI model to date,' claiming performance rivaling the best systems from US frontier labs Anthropic and OpenAI, as well →3 Aug 2026Alibaba debuts Qwen3.8-Max model with 2.4T parametersAlibaba Group Holding Ltd. today debuted a new addition to its Qwen series of open-source large language models. Qwen3.8-Max is the Chinese e-commerce giant’s most capable LLM to date. It features 2.4→5 Aug 2026You have Dwarkesh predicting Anthropic is going to make well over 100B in revenue this year. and then you have this chart.Gary Marcus noted that investor Dwarkesh predicts Anthropic will generate more than $100 billion in revenue this year, according to a source. A chart accompanying the claim omitted DeepSeek’s pricing →5 Aug 2026Qwen Developers' responses from their recent Twitter/X AMAQuestions & Responses(in BOLD) below. Favorite question(s) moved to end of the thread with combined responses(removed duplicates). Be optimistic folks. I'm sure we're getting other models too apart fr→11 Aug 2026When Counterbalancing Hides the Bias: Access-Conditioned Position Lock in Forced-Choice LLM EvaluationarXiv:2607.10202v2 Announce Type: replace Abstract: Forced-choice probes with counterbalanced orientations are a standard tool for measuring language-model 'value dispositions,' and a concentration/ex→11 Aug 2026The small open weight models are scarier in AI developmentImagine if your everyday laptop could run an AI model smart enough to take care of 90% of your work—totally private, lightning fast, and completely free of monthly fees. That is the exact tipping poin→11 Aug 2026Can Open-Weight Models Compete on Financial Text Comprehension?arXiv:2608.08634v1 Announce Type: new Abstract: Open-weight language models from Chinese AI labs caught up on benchmarks relative to proprietary frontier models in recent months. Yet their reliability
CompanyNVIDIA8 recent entries6 Aug 2026Anthropic will design its own hardware to power ClaudeAnthropic is hiring a custom silicon team to design proprietary chips that will power its Claude models, while still planning a multi‑chip strategy that mixes internally designed hardware with compone→6 Aug 2026AMD acquires Toronto-based Taalas, which integrates model weights directly into silicon to boost inference performance, for an undisclosed sum (Tobias Mann/The Register)Tobias Mann / The Register: AMD acquires Toronto-based Taalas, which integrates model weights directly into silicon to boost inference performance, for an undisclosed sum — Early tech demos show model→10 Aug 2026Sources: AI cloud computing provider Lambda is selling a $917M leveraged loan to finance the purchase of GPUs as part of a contract with Nvidia (Bloomberg)Bloomberg: Sources: AI cloud computing provider Lambda is selling a $917M leveraged loan to finance the purchase of GPUs as part of a contract with Nvidia — Lambda Inc., an AI cloud-computing provider→10 Aug 2026Jensen Huang just told the incredible story of how Elon Musk became NVIDIA’s first customer for its AI supercomputer - when literally nobody…Jensen Huang just told the incredible story of how Elon Musk became NVIDIA’s first customer for its AI supercomputer - when literally nobody else wanted it “When I announced this thing, nobody wanted →10 Aug 2026If Anthropic begin shipping chips, will NVIDIA begin shipping frontier models? Who will win?On August 10 2026 at 1:39 AM UTC, user Itamar Friedman (@itamar_mar) posted a short tweet asking whether Anthropic’s potential launch of its own chips would prompt NVIDIA to release frontier AI models→11 Aug 2026Introducing Unsloth Desktop appHi LocalLlama, we're super excited to release Unsloth Desktop today! 🦥 It's the first desktop app that enables you to run and train models locally. Open-source. Available on Mac, Windows, and Linux Su→11 Aug 2026IBM and Together AI sign a $240M, multi-year deal to build an AI inference cluster on IBM Cloud, using Nvidia's HGX B300 systems, to support open-source models (Anhata Rooprai/Reuters)Anhata Rooprai / Reuters: IBM and Together AI sign a 240M, multi-year deal to build an AI inference cluster on IBM Cloud, using Nvidia's HGX B300 systems, to support open-source models — IBM (IBM.N) a→12 Aug 2026Is the future of AI selling hardware for Open Source/Models?I’m not super knowledgeable of the entire AI industry, but as we see this industry grow and the Cold War that is happening between the US and China on AI development, I can’t help but notice what US c