AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “simon-willison--x”

GridTimelineEvolution
229 results
22 Apr 2026

Correction: OpenAI are NOT deprecating that model, the announcement was a mistake

ToolsDGX agent

Correction: OpenAI are NOT deprecating that model, the announcement was a mistake Thank you for flagging this, Jeff. This was a mistake: we are not deprecating text-embedding-3-small. We’re looking in

Has anyone seen ANY official communication from Anthropic or an Anthropic staff member about the fact that the checkbox for Claude Code on P…

Model ReleasesDGX agent

Has anyone seen ANY official communication from Anthropic or an Anthropic staff member about the fact that the checkbox for Claude Code on Pro is back to being checked again, or is the only evidence t

I've been mostly prompting the new model directly via the client.images.generate() API, I have no idea if that rewrites my prompts at all or…


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model ReleasesDGX agent

I've been mostly prompting the new model directly via the client.images.generate() API, I have no idea if that rewrites my prompts at all or if it passes them straight to the model - details here http

New sandbox!

AgentsDGX agent

New sandbox! Announcing Cloud Run sandboxes: Secure on-the-fly code execution: Spin up ephemeral, isolated sandboxes from within Cloud Run resources. Safely execute agent-generated code, scripts, or C

Not the first time either - they shut down a bunch of of their original proprietary hosted embedding models in this announcement back in Apr…

Model ReleasesDGX agent

Not the first time either - they shut down a bunch of of their original proprietary hosted embedding models in this announcement back in April 2024 https://openai.com/index/gpt-4-api-general-availabil

Presumably GPT-imagegen-2 (aka ChatGPT Images 2.0 aka gpt-image-2) works as a tool which the models generate prompts for? I wish we could se…

Model ReleasesDGX agent

Presumably GPT-imagegen-2 (aka ChatGPT Images 2.0 aka gpt-image-2) works as a tool which the models generate prompts for? I wish we could see those prompts, like back in the DALL-E 3 days https://simo

Thank you for flagging this, Jeff. This was a mistake: we are not deprecating text-embedding-3-small. We’re looking into where this came fro…

ToolsDGX agent

Thank you for flagging this, Jeff. This was a mistake: we are not deprecating text-embedding-3-small. We’re looking into where this came from now, and we’ll also email users to clarify. Sorry for the

That's from this ChatGPT conversation - https://chatgpt.com/share/69e90ba5-7760-83e8-9e09-f06e25a90930 - my prompt was 'Do a where's Waldo s…

ToolsDGX agent

That's from this ChatGPT conversation - https://chatgpt.com/share/69e90ba5-7760-83e8-9e09-f06e25a90930 - my prompt was 'Do a where's Waldo style image but it's where is the raccoon holding a ham radio

The new Qwen3.6-27B just gave me definitely the best pelican riding a bicycle I've had from a 16.8GB model file! https://simonwillison.net/2…

ToolsDGX agent

This post highlights the Qwen 3.6-27B language model's image generation capabilities, noting that despite its relatively compact 16.8GB file size, it produces high-quality creative outputs like the ex

This is a really big deal - it's easy to run into nasty bills with Cloud Run if your site attracts aggressive scrapers, spending caps make i…

ToolsDGX agent

This is a really big deal - it's easy to run into nasty bills with Cloud Run if your site attracts aggressive scrapers, spending caps make it a whole lot safer to run small projects on Anouncing Spend

This is why I won't use proprietary hosted embedding models myself - I am more than happy to pay for a hosted solution (cheaper, faster and …

ToolsDGX agent

This is why I won't use proprietary hosted embedding models myself - I am more than happy to pay for a hosted solution (cheaper, faster and more convenient than self-hosting) but I want an open weight

... turns out if you dig around in the browser network inspector enough you CAN find the prompt - here's the prompt it used for this image

ToolsDGX agent

This post documents a method for extracting the text prompt used by an image generation AI by examining network traffic in a web browser's developer tools. Simon Willison demonstrates that prompts can

... which keeps things confusing, since it raises an important new question https://x.com/simonw/status/2046798283700617267

Model ReleasesDGX agent

... which keeps things confusing, since it raises an important new question https://x.com/simonw/status/2046798283700617267 @TheAmolAvasare If I sign up for a new $20/month account today and roll the

Wrote up Anthropic's self-own about Claude Code pricing from this afternoon on my blog - it turned out they'd reversed course just as I hit …

Model ReleasesDGX agent

Wrote up Anthropic's self-own about Claude Code pricing from this afternoon on my blog - it turned out they'd reversed course just as I hit publish, so I've tried to update it to reflect the current s

21 Apr 2026

Claude Opus 4.7 with adaptive thinking via the API... am I missing something or is it not possible any more to force it to think? (Prompt ha…

Model ReleasesDGX agent

Claude Opus 4.7 with adaptive thinking via the API... am I missing something or is it not possible any more to force it to think? (Prompt hacks like 'think step by step' don't count here, I mean the e

I came up with a somewhat foolish new benchmark for testing image generation models, to exercise the new ChatGPT Images 2.0: 'Do a where's W…

Model ReleasesDGX agent

I came up with a somewhat foolish new benchmark for testing image generation models, to exercise the new ChatGPT Images 2.0: 'Do a where's Waldo style image but it's where is the raccoon holding a ham

I for one would be delighted to see OpenAI commit to maintaining a tool like this in the long-term, I'm already nervous about mine going sta…

ToolsDGX agent

Simon Willison expresses hope that OpenAI will commit to long-term maintenance of an AI tool, citing concerns about his own tool potentially becoming stale or discontinued. The post reflects broader u

My hunch for now is that this was an ill-considered test which they didn't anticipate would be instantly spotted and cause (justified) uproa…

Model ReleasesDGX agent

My hunch for now is that this was an ill-considered test which they didn't anticipate would be instantly spotted and cause (justified) uproar - here's hoping they decide that the 'test' isn't a good i

OK, here's a resolution - I managed to get it to think using these settings: 'thinking': { 'type': 'adaptive', 'display': 'summarized' }, 'o…

ToolsDGX agent

OK, here's a resolution - I managed to get it to think using these settings: 'thinking': { 'type': 'adaptive', 'display': 'summarized' }, 'output_config': { 'effort': 'max' } Without 'display': 'summa

OpenAI's new Euphony tool works almost exactly the same way as my Codex transcript viewer https://tools.simonwillison.net/codex-timeline?url…

ToolsDGX agent

OpenAI's new Euphony tool works almost exactly the same way as my Codex transcript viewer https://tools.simonwillison.net/codex-timeline?url=https%3A%2F%2Fgist.githubusercontent.com%2Fsimonw%2Fa9eb599

This piece thinks that the reason I got a crap pelican riding a bicycle from Opus 4.7 is that it didn't think about it first, I'm trying to …

ToolsDGX agent

This piece thinks that the reason I got a crap pelican riding a bicycle from Opus 4.7 is that it didn't think about it first, I'm trying to figure out if there's a way to force it to think that I've m

True to form, I've already seen OpenAI themselves refer to the new image model as 'ChatGPT Images 2.0', 'Image gen 2' and 'gpt-image-2'

ToolsDGX agent

OpenAI has been using multiple informal names internally and externally for its new image generation model, including 'ChatGPT Images 2.0,' 'Image gen 2,' and 'gpt-image-2.' The post highlights incons

20 Apr 2026

'I am envisioning a harmonious interplay of form and space, where each element finds its rightful place within the visual narrative. The pel…

AgentsDGX agent

'I am envisioning a harmonious interplay of form and space, where each element finds its rightful place within the visual narrative. The pelican emerges as a central figure, its wings poised in dynami

I tried a 15MB, 30 page text-heavy PDF and Opus 4.7 reported 60,934 tokens while 4.6 reported 56,482 - that's a 1.08x multiplier, significan…

ToolsDGX agent

Claude Opus 4.7 uses approximately 8% more tokens than Opus 4.6 when processing the same text-heavy PDF document, as demonstrated by a test comparing tokenization of a 15MB, 30-page PDF (60,934 tokens

Important to note: that 3x increase for images is entirely due to Opus 4.7 being able to handle higher resolutions. I tried that again with …

ToolsDGX agent

Important to note: that 3x increase for images is entirely due to Opus 4.7 being able to handle higher resolutions. I tried that again with a 682x318 pixel image and it took 314 tokens with Opus 4.7 a

(Just tried community-noting my own tweet to clarify this point, let's see if that works!)

ToolsDGX agent

Simon Willison shared an experience attempting to use X's Community Notes feature to add clarification to his own tweet, experimenting with whether the self-noting functionality would work as intended

New TIL on fetching data from a Datasette instance into Google Sheets using importdata(), named custom functions or Google Apps Script https…

ToolsDGX agent

Simon Willison shared a tutorial on integrating Datasette instances with Google Sheets, demonstrating multiple methods including the importdata() function, custom named functions, and Google Apps Scri

19 Apr 2026

Note to @AnthropicAI - much as I appreciate the public system prompts this would be so much more valuable to me as a Claude power user if yo…

Model ReleasesDGX agent

Note to @AnthropicAI - much as I appreciate the public system prompts this would be so much more valuable to me as a Claude power user if you published the tool descriptions as well Since Anthropic pu

Since Anthropic publish their system prompts we can generate a diff between Claude Opus 4.6 and 4.7 - here are my notes on what's changed ht…

Model ReleasesDGX agent

Simon Willison documents the differences between Anthropic's Claude Opus 4.6 and 4.7 system prompts, analyzing changes that Anthropic made public. The notes likely highlight modifications to model beh

The biggest challenge of using chat-based AI systems is that the details of what they can do are invisible - those tool descriptions are the…

Model ReleasesDGX agent

The biggest challenge of using chat-based AI systems is that the details of what they can do are invisible - those tool descriptions are the missing manual, publishing them would be a huge benefit to

18 Apr 2026

Join us at PyCon US 2026 in Long Beach—we have new AI and security tracks this year https://simonwillison.net/2026/Apr/17/pycon-us-2026/

ToolsDGX agent

PyCon US 2026 will be held in Long Beach and will feature new dedicated tracks focused on AI and security topics. The announcement comes from Simon Willison, highlighting these expanded programming ar

The event is next month - talks are Friday 15th to Sunday 17th of May, with the date before that of tutorials and two days afterwards of spr…

ToolsDGX agent

The event is next month - talks are Friday 15th to Sunday 17th of May, with the date before that of tutorials and two days afterwards of sprints Californians are notoriously last-minute planners so I

17 Apr 2026

Is there still a widespread belief that LLMs and coding agents are good for greenfield development but don't help for maintaining large exis…

ToolsDGX agent

Simon Willison discusses the perception that LLMs and coding agents are primarily useful for greenfield development (starting new projects from scratch) rather than for maintaining and modifying large

16 Apr 2026

A https://claude.ai/ feature I really like is you can tell it to 'clone x/y from GitHub' and it can then answer questions about a repo, or u…

Model ReleasesDGX agent

A https://claude.ai/ feature I really like is you can tell it to 'clone x/y from GitHub' and it can then answer questions about a repo, or use snippets of code from that repo to help build new artifac

Here's Qwen 3.6-35B-A3B v.s. Claude Opus 4.7 for 'Generate an SVG of a flamingo riding a unicycle', in case you thought Qwen might be cheati…

Model ReleasesDGX agent

This post compares the performance of Qwen 3.6-35B-A3B and Claude Opus 4.7 models on a creative task of generating SVG code for a flamingo riding a unicycle, likely demonstrating differences in their

More on my blog, including results from the previously secret 'flamingo on a unicycle' test https://simonwillison.net/2026/Apr/16/qwen-beats…

Model ReleasesDGX agent

Simon Willison discusses results from a 'flamingo on a unicycle' test on his blog, likely comparing AI model performance including Qwen. The post appears to reference previously undisclosed or unconve

15 Apr 2026

I ran the same prompt for a London Estuary accent, a Newcastle accent and an Exeter, Devon accent - all three audio files are now embedded i…

Model ReleasesDGX agent

I ran the same prompt for a London Estuary accent, a Newcastle accent and an Exeter, Devon accent - all three audio files are now embedded in my blog post https://simonwillison.net/2026/Apr/15/gemini-

I'm certain this isn't the message they intended to present, but this comes across to me as a company saying 'we no longer trust in our own …

ToolsDGX agent

I'm certain this isn't the message they intended to present, but this comes across to me as a company saying 'we no longer trust in our own ability to keep your data secure' Open source is dead. That’

@pumfleet @calcom Did you see this piece by @dbreunig? He argues that the cost of locking down software through LLM analysis makes open sour…

ToolsDGX agent

@pumfleet @calcom Did you see this piece by @dbreunig? He argues that the cost of locking down software through LLM analysis makes open source MORE valuable now: https://www.dbreunig.com/2026/04/14/cy

The example prompt for Google's new Gemini Flash TTS text-to-speed model is a lot https://simonwillison.net/2026/Apr/15/gemini-31-flash-tts/

Model ReleasesDGX agent

Google's Gemini 3.1 Flash TTS (text-to-speech) model includes a notably elaborate or extensive example prompt, which Simon Willison highlighted as noteworthy. The post likely comments on the complexit

10 Apr 2026

Finally some good news from @simonw: The Kākāpō parrots of New Zealand are having a fantastic breeding season.

ToolsDGX agent

Finally some good news from @simonw: The Kākāpō parrots of New Zealand are having a fantastic breeding season. Media 'Using coding agents well is taking every inch of my 25 years of experience as a so

I think it's non-obvious to many people that the OpenAI voice mode runs on a much older, much weaker model - it feels like the AI that you c…

ToolsDGX agent

I think it's non-obvious to many people that the OpenAI voice mode runs on a much older, much weaker model - it feels like the AI that you can talk to should be the smartest AI but it really isn't Jud

If you ask ChatGPT voice mode for its knowledge cutoff date it tells you April 2024 - it's a GPT-4o era model

ToolsDGX agent

ChatGPT's voice mode, when directly queried about its knowledge cutoff date, reports April 2024, indicating it is powered by a GPT-4o era model rather than a more recent one. GPT-4o initially had ...

8 Apr 2026

I care because understanding which index helps me understand things like what I need to submit content for crawling to, what search agents I…

ToolsDGX agent

I care because understanding which index helps me understand things like what I need to submit content for crawling to, what search agents I should allow in robots.txt, how frequently I can expect the

Pelicans for Meta's new Muse Spark models - plus I did a bit of a deep dive into the Code Interpreter and fascinating 'container.visual_grou…

ToolsDGX agent

Pelicans for Meta's new Muse Spark models - plus I did a bit of a deep dive into the Code Interpreter and fascinating 'container.visual_grounding' tools in their http://meta.ai chat UI https://simonwi

The feature I most want from AI labs right now is documentation on which underlying search engines they use when their chat tools run a sear…

Model ReleasesDGX agent

The feature I most want from AI labs right now is documentation on which underlying search engines they use when their chat tools run a search OpenAI and Anthropic and Meta AI all have search and I ha

7 Apr 2026

754B parameters, 1.51TB on Hugging Face

ToolsDGX agent

754B parameters, 1.51TB on Hugging Face Introducing GLM-5.1: The Next Level of Open Source - Top-Tier Performance: #1 in open source and #3 globally across SWE-Bench Pro, Terminal-Bench, and NL2Repo.

I'm a big fan of the pelican GLM-5.1 drew me today, it even animated it! https://simonwillison.net/2026/Apr/7/glm-51/

ToolsDGX agent

Simon Willison tested Z.ai's GLM-5.1 using his standard 'pelican on a bicycle' SVG benchmark, and the model stood out by spontaneously generating a full HTML page with both the SVG and a separate ...

Wrote up some thoughts on Anthropic's Project Glassing, where their latest Opus-beating model is available to partnered security research or…

ToolsDGX agent

Wrote up some thoughts on Anthropic's Project Glassing, where their latest Opus-beating model is available to partnered security research organizations only Given recent alarm bells raised by credible

← Previous
1234
Next →