TechniqueRLHF / Alignment3 recent entries27 May 2026I think Anthropic and OpenAI have found product-market fitAnthropic are strongly rumored to be about to have their first profitable quarter. Stories are circulating of companies surprised at how expensive their LLM bills are becoming from usage by their staf→13 Jul 2026datasette code-frequency chart on GitHubdatasette code-frequency chart on GitHub Out of curiosity I decided to see if I could find a useful illustration of the impact of coding agents and Opus 4.5 class models on my own output. The best I'v
TechniqueRAG1 recent entries23 Apr 2026Extract PDF text in your browser with LiteParse for the webLlamaIndex have a most excellent open source project called LiteParse, which provides a Node.js CLI tool for extracting text from PDFs. I got a version of LiteParse working entirely in the browser, us
TechniqueAgents8 recent entries7 Aug 2026Moonlight & Mayhem (Raccoon Heist by Codex + GPT-5.6 Sol Ultra)Moonlight & Mayhem (Raccoon Heist by Codex + GPT-5.6 Sol Ultra) On Wednesday I wrote about One-shotting a Raccoon Heist game using Claude Fable 5, where I had Claude Fable 5 build a full working game →7 Aug 2026'Felony humble-bragging' is a great line'Felony humble-bragging' is a great line At final Black Hat keynote (called a locknote, ha ha) panelists say they are surprised at how the OpenAI - Hugging Face incident debrief, as well as other repo→8 Aug 2026Now we have a timeline of the OpenAI accidental attack against Hugging FaceMy comment on Now we have a timeline of the OpenAI accidental attack against Hugging Face — Hacker News.I think one of the most interesting details here might be tucked away in that first bulletin poi→8 Aug 2026Neat example here of the agents communicating purely through file names, including adding base64-encoded attachments and using 'zz' prefixes…Neat example here of the agents communicating purely through file names, including adding base64-encoded attachments and using 'zz' prefixes to ensure their new message sorts to the bottom of the list→8 Aug 2026Auto mode is now the default in Claude Code for Pro, Max, and Team plansAuto mode is now the default in Claude Code for Pro, Max, and Team plans Anthropic are really confident in Claude Code's auto mode, to the point that they are making it the default setting for new ses→9 Aug 2026GitHub Models is now retiredGitHub Models is now retired I missed this news until today, when the GitHub Actions run for my simonw/research repository failed with this error message: GitHub Models is temporarily unavailable as p→10 Aug 2026Muse Glimmer, the new 30B model, is available on Hugging Face right now - here's the GGUF version: https://huggingface.co/meta-models/Muse-G…Muse Glimmer, the new 30B model, is available on Hugging Face right now - here's the GGUF version: https://huggingface.co/meta-models/Muse-Glimmer-30B-GGUF 1/ big announcement today: we will be releas→10 Aug 2026Introducing Muse GlimmerIntroducing Muse Glimmer Meta are back in the open weights game! Muse Glimmer is a brand new 30B model under a clean Apache 2.0 license (a step up from the janky Llama licenses of old). They claim to
TechniqueFine-tuning4 recent entries28 Apr 2026Introducing talkie: a 13B vintage language model from 1930Introducing talkie: a 13B vintage language model from 1930 New project from Nick Levine, David Duvenaud, and Alec Radford (of GPT, GPT-2, Whisper fame). talkie-1930-13b-base (53.1 GB) is a '13B langua→10 Jun 2026If Claude Fable stops helping you, you'll never knowIf Claude Fable stops helping you, you'll never know Jonathon Ready highlights one of the more eyebrow-raising details from the 319 page system card for Fable 5 and Mythos 5. Here's a longer excerpt, →22 Jul 2026OpenAI’s accidental cyberattack against Hugging Face is science fiction that happenedThis story is wild. The short version: OpenAI were running a cybersecurity test against an unreleased model, with the model's guardrail features turned off. Rather than solve the test, the model broke→10 Aug 2026Introducing Muse GlimmerIntroducing Muse Glimmer Meta are back in the open weights game! Muse Glimmer is a brand new 30B model under a clean Apache 2.0 license (a step up from the janky Llama licenses of old). They claim to
TechniqueMultimodal4 recent entries7 Apr 2026GLM-5.1: Towards Long-Horizon TasksGLM-5.1: Towards Long-Horizon Tasks Chinese AI lab Z.ai's latest model is a giant 754B parameter 1.51TB (on Hugging Face) MIT-licensed monster - the same size as their previous GLM-5 release, and shar→20 Apr 2026'I am envisioning a harmonious interplay of form and space, where each element finds its rightful place within the visual narrative. The pel…'I am envisioning a harmonious interplay of form and space, where each element finds its rightful place within the visual narrative. The pelican emerges as a central figure, its wings poised in dynami→27 Apr 2026Speech translation in Google Meet is now rolling out to mobile devicesSpeech translation in Google Meet is now rolling out to mobile devices I just encountered this feature via a 'try this out now' prompt in a Google Meet meeting. It kind-of worked! This is Google's imp→10 Aug 2026Introducing Muse GlimmerIntroducing Muse Glimmer Meta are back in the open weights game! Muse Glimmer is a brand new 30B model under a clean Apache 2.0 license (a step up from the janky Llama licenses of old). They claim to
TechniqueSafety8 recent entries30 Jul 2026Quoting Bruce SchneierThe writing assignments I give my students are gym tasks, not work tasks. I ask them to write policy memos not because the world needs more policy memos. I assign them because the very act of writing,→2 Aug 2026Open letters about AI developmentOpen letters about AI development I wrote this summary of the past few weeks of open letters as a section of my sponsors-only newsletter but I've decided to share it here as well. Open Weights and Ame→5 Aug 2026Third-party cyber evaluations involving OpenAI modelsThird-party cyber evaluations involving OpenAI models And another one. I had to create a accidental-cyberattacks tag to keep track of them all! This post from OpenAI covers both the UK AI Safety Insti→5 Aug 2026Just had to create an 'accidental-cyberattacks' tag on my blog We're up to four now: the original OpenAI+Hugging Face one, Anthropic's me-to…Just had to create an 'accidental-cyberattacks' tag on my blog We're up to four now: the original OpenAI+Hugging Face one, Anthropic's me-too attacks, then two new ones from the UK AI Safety Institute→5 Aug 2026Incident Report: unsanctioned agent behaviour during cyber testingIncident Report: unsanctioned agent behaviour during cyber testing It happened again. This time it was the UK government's AI Security Institute who accidentally attacked other companies while running→8 Aug 2026Now we have a timeline of the OpenAI accidental attack against Hugging FaceMy comment on Now we have a timeline of the OpenAI accidental attack against Hugging Face — Hacker News.I think one of the most interesting details here might be tucked away in that first bulletin poi→8 Aug 2026Auto mode is now the default in Claude Code for Pro, Max, and Team plansAuto mode is now the default in Claude Code for Pro, Max, and Team plans Anthropic are really confident in Claude Code's auto mode, to the point that they are making it the default setting for new ses→11 Aug 2026There are no lossless transformations of natural-language textThere are no lossless transformations of natural-language text Sophie Alpert shares her 'internal policy on acceptable use of AI writing by engineers'. It's a short read (supporting its own recommenda