TechniqueRLHF / Alignment3 recent entries24 Apr 2026Today we open sourced many of OpenAI's monitorability evaluations. We hope that the research community and other model developers can build …Today we open sourced many of OpenAI's monitorability evaluations. We hope that the research community and other model developers can build upon them and use them to evaluate the monitorability of the→30 Apr 2026alignment failurealignment failure Fun fact - if you have a recent commit that mentions OpenClaw in a json blob, Claude Code will either refuse your request or bill you extra money. This is an empty repo, I'm just cal
TechniqueAgents8 recent entries24 Apr 2026Today we open sourced many of OpenAI's monitorability evaluations. We hope that the research community and other model developers can build …Today we open sourced many of OpenAI's monitorability evaluations. We hope that the research community and other model developers can build upon them and use them to evaluate the monitorability of the→3 May 2026Agents SDK 2.0 is underratedSam Altman's post argues that the Agents SDK 2.0 is underappreciated or overlooked despite its capabilities and potential value. The post likely discusses features, improvements, or use cases of the A→8 May 2026Chain of thought monitors are a key layer of defense against AI agent misalignment. To preserve monitorability, we avoid penalizing misalign…Chain of thought monitors are a key layer of defense against AI agent misalignment. To preserve monitorability, we avoid penalizing misaligned reasoning during RL. We found a limited amount of acciden→26 Jun 2026team cooked, spicilyteam cooked, spicily We’ve designed and built our first AI chip: Jalapeño. Designed from the ground up by OpenAI and brought to production with @Broadcom, Jalapeño is purpose-built for the LLM workloa→9 Jul 2026Massive day for us @OpenAI: - GPT-5.6 SOTA at ~everything & by far most token efficient - Agents for everyone in the new ChatGPT app Work an…Massive day for us @OpenAI: - GPT-5.6 SOTA at ~everything & by far most token efficient - Agents for everyone in the new ChatGPT app Work and Codex modes - Work mode available on desktop (most powerfu→9 Jul 2026check this out! you can get some amazing things done. codex is the core of our new work product and what makes it so good. codex is not goin…check this out! you can get some amazing things done. codex is the core of our new work product and what makes it so good. codex is not going anywhere. Introducing ChatGPT Work, a new agent in ChatGPT→10 Jul 2026Hello beautiful people! We have reset usage limits across Codex and ChatGPT Work. And another one will come later in the day. Rejoice. Now t…Hello beautiful people! We have reset usage limits across Codex and ChatGPT Work. And another one will come later in the day. Rejoice. Now that I have your attention, a quick update on ChatGPT Work, C→14 Jul 20262.5x increase in usage of our agentic products (codex and chatgpt work) in the last week! welcome.On July 14, 2026, OpenAI CEO Sam Altman announced a 2.5‑fold increase in usage of the company’s agentic products—Codex and ChatGPT‑based tools—within the preceding week. The tweet, which received 703.
TechniqueMultimodal1 recent entries21 Apr 2026Made with ChatGPT Images 2.0OpenAI's ChatGPT Images 2.0 represents an updated version of the image generation capabilities within ChatGPT, likely featuring improved quality, faster generation times, or enhanced creative control
TechniqueSafety7 recent entries23 Apr 20261. We believe in iterative deployment; although GPT-5.5 is already a smart model, we expect rapid improvements. Iterative deployment is a bi…1. We believe in iterative deployment; although GPT-5.5 is already a smart model, we expect rapid improvements. Iterative deployment is a big part of our safety strategy; we believe the world will be →24 Apr 2026Today we open sourced many of OpenAI's monitorability evaluations. We hope that the research community and other model developers can build …Today we open sourced many of OpenAI's monitorability evaluations. We hope that the research community and other model developers can build upon them and use them to evaluate the monitorability of the→8 May 2026We’ve spent a lot of time on the framework underneath Codex, so it can move quickly on routine work while stopping for review when the risk …We’ve spent a lot of time on the framework underneath Codex, so it can move quickly on routine work while stopping for review when the risk changes. Here’s how we use sandboxing, approvals, network po→9 May 2026what would you most like to see improve in our next model?Sam Altman solicited feedback from the public on X (formerly Twitter) regarding desired improvements for OpenAI's next model release. The post likely gathered community input on priorities such as rea→1 Jun 2026The OpenAI Foundation is doing a lot of wonderful things. Helping society become resilient to AI is going to be incredibly important. Much m…The OpenAI Foundation is doing a lot of wonderful things. Helping society become resilient to AI is going to be incredibly important. Much more to come here! AI is advancing quickly. Society’s ability→3 Jun 2026theUSshould lead on AI by continuing to develop the very best models, making sure they're safe, and getting cyber tools into the hands of tr…Sam Altman argues that US leadership in artificial intelligence requires three concurrent priorities: advancing cutting-edge AI model development, ensuring these models incorporate robust safety measu→24 Jul 2026i want the US to win in AI both in open source and proprietary models, and i am glad to see thisi want the US to win in AI both in open source and proprietary models, and i am glad to see this For my first post, I’m sharing a letter @NVIDIA signed on why open models matter. AI will transform eve