Quoting Matthew Green
Right now we’re in the midst of a historic transition from traditional public-key algorithms based on EC-based cryptography and RSA, moving over to new post-quantum algorithms based on novel problems.
Knowledge catalogue
Right now we’re in the midst of a historic transition from traditional public-key algorithms based on EC-based cryptography and RSA, moving over to new post-quantum algorithms based on novel problems.
arXiv:2607.23386v1 Announce Type: new Abstract: We document a failure class in frontier large language models -- exception chain collapse -- observed in eligibility evaluation under nested conditional
Go ahead and ask yourself, really ask yourself, if all these confident predictors are wrong about jobs and not training radiologists anymore and fast take offs, what else are they wrong about? Then as
Recently I've flipped from being bullish to being bearish about AI. I think I'm updating my bearishness to be more solidly bearish. Early thoughts (which I hope to be disproven in the next year or so,
We recently joined NVIDIA, Microsoft, and others in supporting open-weight AI. Today, we’re putting that belief into the product with model choice on Replit, starting with Kimi K3. The future is the r
You patch security holes by intentionally finding them. If the models refuse to do it, how can companies protect themselves against rogue AIs, whether they are Chinese or OpenAI/Anthropic themselves?
An opinionated guide to which AI to use to do stuff It's interesting watching the evolution of Ethan Mollick's guide over time. A year ago it was still all about chat - ChatGPT, Claude, Gemini - with
I have been trying to do GLM 5.2 as plan / K2.7 as execute, but i hit my usage so fast it's not viable. It's hitting limits much faster than claude code / codex $20 plan. Using in opencode. What are y
Keach Hagey / Wall Street Journal: Nvidia says it made a “substantial” investment in Safe Superintelligence, which will get access to Nvidia GPUs to up computing power by an “order of magnitude” — Par
arXiv:2607.21672v1 Announce Type: cross Abstract: Long source-code contexts consume many text tokens, motivating the proposal to render code as images for vision-language models. Recent work asks whet
but they won’t. AI safety experts say OpenAI’s rogue models may mean the company has already blown past its own internal red lines. That would mean it should pause development until it creates better
A algumas semanas, percebi algo diferente, aparentemente modelos de qualidade, em especial o GLM 5.1 ficou mais burro, e começou a mandar caracteres em mandarim para mim sem eu nem usar eles, e isto n
U.S. lawmakers are preparing to introduce a bill that will give the government the ability to trigger an emergency shutdown of artificial intelligence models that may cause harm to the public. In what
Leo Schwartz / The Information: Meta, Nvidia, Microsoft, a16z, and others sign a letter defending open-source AI; Jensen Huang, in his first X post, says open models strengthen cybersecurity — Many of
We benchmarked Opus 5 comprehensively on document understanding through ParseBench. It is roughly on par with Opus 4.8 - it does a few points worse on dense tables, but does slightly better on parsing
arXiv:2607.19430v1 Announce Type: cross Abstract: Multi-agent LLM applications chain a planner, worker agents, a verifier, and a synthesizer, and every hop between agents is an unmonitored channel thr
A Chinese article compiled 52 remarks from Liang Wenfeng’s four-hour investor meeting. I’ve summarised the most important ones below. DeepSeek has one central objective: AGI. This is not the time to m
I wrote the latest of my occasional guides to which AI to use right now for non-experts who want to get stuff done. The agentic systems available to everyone are getting extremely powerful (even as th
People are wasting their AI subscriptions. They don’t realize how powerful AI can be in achieving their own goals. They’re asking AI to reply to emails with zero goals in mind and nothing laddering up
BREAKING: Inkling by @thinkymachines is 9th overall on Agentic Web App Arena by Design Arena with an Elo of 1257 It's an open-weight model in the same performance band as Claude Opus 4.6 by @Anthropic
“I asked a few AI researchers whether they could name any other real-world software that scales so poorly. None of them could think of any. Even outside the world of software, it’s hard to find a comp
Maxwell Zeff / Wired: OpenAI employees donated a combined 215K+ to Guardrails Alliance, a super PAC seeking stricter AI rules and opposing Greg Brockman-backed Leading the Future — OpenAI employees ha
🆕This Year In Claude https://www.youtube.com/watch?v=uU5Gv2h8-9g @simonw chats with @_catwu and @trq212 about the state of: - @claudeai Code - Claude Fable - @anthropicai culture & product strategy -
Demis Hassabis thinks the world needs an AI watchdog with the power to hit the brakes if frontier models become too dangerous. Writing in a blog post, the Google DeepMind CEO and cofounder said the US
Creating my own interfaces in real-time to do the stuff I want to do is ridiculously empowering. Here's an example: - took ~10 app screenshots of things I wanted to fix - asked Claude to make a feedba
Jarred Sumner / bun.com: Bun's creator says he rewrote Bun from Zig to Rust using a Claude Fable 5 prerelease version in 11 days, noting it would've taken three engineers “about a year” — Disclosure:
OpenAI's latest flagship model hit general availability this morning, and comes in three sizes: Luna, Terra, and Sol (from smallest to largest). The new models are priced per 1M input/output tokens as
Rewriting Bun in Rust Jarred Sumner has been promising this blog post (since May 9th) about his Zig to Rust rewrite of Bun for significantly longer than it took him to finish the rewrite. Honestly, it
We need AI model selection to be MUCH easier ASAP. I want proactive flags from my AI systems suggesting models. I want my AI harness to say 'hey allie, my girl, you keep asking for bar recommendations
I wrote about the sqlite-utils 4.0rc1 release a couple of weeks ago. Since we only have Claude Fable on our Max subscriptions for a few more days, I decided to see if it could help me get to a 4.0 sta
Something about this year’s @aiDotEngineer World’s Fair just hit different. Last year was the year of “let the agents rip.” This year was the year of realizing that autonomy without structure creates
What's new in Claude Sonnet 5 Claude Sonnet 5 came out this morning. I always head straight for the 'what's new' developer docs because they tend to have more actionable information than the official
Claude models are now generally available through Microsoft Foundry on Azure infrastructure, providing Azure customers access to Claude Opus 4.8 and Claude Haiku 4.5. This integration allows enterpris
POD UP! 🚨 Besties are back to discuss: -- SpaceX's record IPO, Cursor deal, and the first trillionaire -- The modern politburo and the new oligarchy (@friedberg cooks) -- Behind the scenes of the Anth
HISTORY LESSON: In 1968 the US, USSR, UK, France, and China signed the Nuclear Non-Proliferation Treaty, declaring nuclear weapons too dangerous for any more countries to build. All five already had t
Easy solution to slow down recursive AI self improvement: The lab with the top-ranked model must agree THEY must not use it for working on frontier AI But everyone else should have access to it. By de
arXiv:2606.08529v1 Announce Type: new Abstract: Published agent capability scores conflate what a model can do with what its scaffold lets it do, and the magnitude of this elicitation gap is not well
From fastest growing company to “worst value among its peers” in 18 months. Why, oh why, stick with the CEO? PitchBook's analysts just ranked OpenAI last for value among its AI peers. Not last for cap
There was an inflection point recently where the tide shifted to model pickers and OSS Mix of tokenmaxxing/cost fatigue, nemotron coalition, harness step functions, brains/claws, etc Long live those w
Uber Caps Usage of AI Tools Like Claude Code to Manage Costs I wrote the other day about Uber blowing its 2026 AI budget in four months, and how that wasn't particularly surprising given they would ha
arXiv:2606.01152v1 Announce Type: cross Abstract: The work of a professional software engineer has begun to consist, increasingly, of directing agents rather than writing code, and the empirical evide
In my Harvard fellowship I study the views of AI accelerationists, safetyists and skeptics. What I have come to realize is that both the Accelerationists and the Safetyists believe that we are creatin
I cannot believe I'm saying this, but getting the literal Pope to canonize your product's specific technical limitations as a spiritual treatise is the single greatest act of vendor lobbying I have ev
The US government absolutely should NOT bail out OpenAI. They have AFAIK been given more funding than anyone in history and there is no evidence they can run a profitable business,and meanwhile many o
This is collectively the largest financial iceberg ever put in front of the stock market. This is $5 trillion in TOTAL AIR. These companies have never made a dime. SpaceX is literally a government wel
arXiv:2605.22714v1 Announce Type: cross Abstract: Large language models are routinely used as automated evaluators: to review code, moderate content, or score outputs, often with many items passing th
I put together these annotated slides from my five minute lightning talk at PyCon US 2026, using the latest iteration of my annotated presentation tool. # I presented this lightning talk at PyCon US 2
BREAKING: The results are in for Slides Arena... @AnthropicAI and @Zai_org models continue to lead the way in soft-verifiable domains 1st: Opus 4.7 by @AnthropicAI 2nd: Opus 4.7 (Thinking) by @Anthrop
In just a short few months, we have witnessed several artificial intelligence technology events that deserve the overused “unprecedented” descriptor: a highly complex supply chain attack by TeamPCP, A
As AI coding agents become deeply embedded in developer workflows, defenders must evolve their definition of malicious files and rethink how to protect against them. Autonomous AI agents operate acros
arXiv:2605.09504v1 Announce Type: cross Abstract: We present swarm-attack, an open-source adversarial testing framework in which multiple lightweight LLM agents coordinate through shared memory, paral
When is the last time a general purpose LLM (putting aside hybrid systems like Claude Code with special purpose symbolic harnesses) last completely blew away all competing prior models? GPT-4 relative
Mythos seems to be a very capable model based on available information, but it is not a cybersecurity model - it is an advanced general purpose model that happens to be good at cyber because it is goo
Zig has one of the most stringent anti-LLM policies of any major open source project: No LLMs for issues. No LLMs for pull requests. No LLMs for comments on the bug tracker, including translation. Eng
About 600 employees at Google LLC have signed a letter urging Chief Executive Sundar Pichai not to make the company’s artificial intelligence tools available to the Pentagon in classified settings. Th
Aditya Soni / Reuters: Microsoft is rolling out Copilot 365 to all of Accenture's roughly 743,000 employees, in the biggest enterprise deal for Copilot, following a 2023 pilot — Microsoft (MSFT.O) is
Bloomberg: A look at Elon Musk's efforts to launch the banking and payments service X Money, delayed by US regulatory concerns, as some industry watchers remain skeptical — More than three years after
OpenAI: Microsoft and OpenAI amend their deal to let OpenAI serve all its products across any cloud provider; Microsoft will no longer pay a revenue share to OpenAI — Amended agreement provides long-t
Stephen Morris / Financial Times: Epoch AI: Google controls ~25% of global AI compute, with ~3.8M TPUs and 1.3M GPUs; Google Cloud CEO Thomas Kurian says demand and revenue justify the spend — Thomas
Makena Kelly / Ars Technica: Palantir Slack logs and staff interviews reveal internal debates over the company's ICE and DOD contracts during Trump's second term, its manifesto, and more — It took jus