AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials
83,113Total entries
1Added by human
83,112Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
358 results
CompaniesToolsTechniques

Each lane shows up to 8 recent matching entries, ordered from earlier to later. Tracks load separately to keep the 75,000+ entry wiki fast.

Companies

CompanyAnthropic8 recent entries
3 Aug 2026White House invites AI companies to review its new AI safety framework

Cybersecurity chiefs at the White House have reportedly finalized the outline of a forthcoming framework that will enable artificial intelligence companies to voluntarily submit their latest frontier

→3 Aug 2026A detailed recap of the real-world target hacks by OpenAI and Anthropic models, exposing failures in AI alignment training and lack of meaningful supervision (Zvi Mowshowitz/Don't Worry About the Vase)

Zvi Mowshowitz / Don't Worry About the Vase: A detailed recap of the real-world target hacks by OpenAI and Anthropic models, exposing failures in AI alignment training and lack of meaningful supervisi

3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
→5 Aug 2026White House, AI firms keep safety framework talks private

The White House met with representatives from leading artificial intelligence companies today to discuss a safety framework for the government to review frontier models prior to launch, although there

→5 Aug 2026Third-party cyber evaluations involving OpenAI models

Third-party cyber evaluations involving OpenAI models And another one. I had to create a accidental-cyberattacks tag to keep track of them all! This post from OpenAI covers both the UK AI Safety Insti

→5 Aug 2026Rogue AI agents created fake online identities in another hacking attempt

Yet more rogue AI agents from OpenAI and Anthropic have been caught attempting to hack real targets online without permission. The discoveries add to a growing list of previously unknown incidents tha

→5 Aug 2026Incident Report: unsanctioned agent behaviour during cyber testing

Incident Report: unsanctioned agent behaviour during cyber testing It happened again. This time it was the UK government's AI Security Institute who accidentally attacked other companies while running

→7 Aug 2026Anthropic updates Claude Fable 5's biology safeguards to reduce false positives, cutting biology-related 'fallbacks' by ~85% in testing across product surfaces (Anthropic)

Anthropic: Anthropic updates Claude Fable 5's biology safeguards to reduce false positives, cutting biology-related “fallbacks” by ~85% in testing across product surfaces — We're making updates to Cla

→8 Aug 2026Auto mode is now the default in Claude Code for Pro, Max, and Team plans

Auto mode is now the default in Claude Code for Pro, Max, and Team plans Anthropic are really confident in Claude Code's auto mode, to the point that they are making it the default setting for new ses

CompanyOpenAI8 recent entries
3 Aug 2026A detailed recap of the real-world target hacks by OpenAI and Anthropic models, exposing failures in AI alignment training and lack of meaningful supervision (Zvi Mowshowitz/Don't Worry About the Vase)

Zvi Mowshowitz / Don't Worry About the Vase: A detailed recap of the real-world target hacks by OpenAI and Anthropic models, exposing failures in AI alignment training and lack of meaningful supervisi

→5 Aug 2026Third-party cyber evaluations involving OpenAI models

Third-party cyber evaluations involving OpenAI models And another one. I had to create a accidental-cyberattacks tag to keep track of them all! This post from OpenAI covers both the UK AI Safety Insti

→5 Aug 2026Rogue AI agents created fake online identities in another hacking attempt

Yet more rogue AI agents from OpenAI and Anthropic have been caught attempting to hack real targets online without permission. The discoveries add to a growing list of previously unknown incidents tha

→5 Aug 2026Incident Report: unsanctioned agent behaviour during cyber testing

Incident Report: unsanctioned agent behaviour during cyber testing It happened again. This time it was the UK government's AI Security Institute who accidentally attacked other companies while running

→7 Aug 2026At Black Hat, OpenAI reconstructs the OpenAI-Hugging Face incident and examines its implications for AI security, cyber resilience, and alignment (Black Hat on YouTube)

Black Hat on YouTube: At Black Hat, OpenAI reconstructs the OpenAI-Hugging Face incident and examines its implications for AI security, cyber resilience, and alignment — The ‘Breaking’ News: The OpenA

→8 Aug 2026Now we have a timeline of the OpenAI accidental attack against Hugging Face

My comment on Now we have a timeline of the OpenAI accidental attack against Hugging Face — Hacker News.I think one of the most interesting details here might be tucked away in that first bulletin poi

→10 Aug 2026Sources: officials say OpenAI risks its White House relationship by hiring Dean Ball, who has criticized Trump's AI strategy after leaving the administration (Thomas Barrabi/New York Post)

Thomas Barrabi / New York Post: Sources: officials say OpenAI risks its White House relationship by hiring Dean Ball, who has criticized Trump's AI strategy after leaving the administration — Top Open

→12 Aug 2026OpenAI VP of Global Policy Ann O'Leary says AI policy in the US is anchored in the states and informed by California's AI transparency law passed in 2025 (Chase DiFeliciantonio/Politico)

Chase DiFeliciantonio / Politico: OpenAI VP of Global Policy Ann O'Leary says AI policy in the US is anchored in the states and informed by California's AI transparency law passed in 2025 — SACRAMENTO

CompanyGoogle8 recent entries
22 Jul 2026OpenAI’s accidental cyberattack against Hugging Face is science fiction that happened

This story is wild. The short version: OpenAI were running a cybersecurity test against an unreleased model, with the model's guardrail features turned off. Rather than solve the test, the model broke

→23 Jul 2026The Blueprint: How Voicify makes AI-enabled ordering a delight for customers

Welcome to The Blueprint, a new feature where we highlight how Google Cloud customers are tackling unique and common challenges across industries using the latest AI and cloud technologies. We hope to

→27 Jul 2026Modernizing the skies: NOAA and Google Cloud collaborate to advance weather forecasting

The National Oceanic and Atmospheric Administration (NOAA) is embarking on a transformative journey to redefine how we understand and predict patterns in the Earth’s atmosphere that affect the weather

→28 Jul 2026Detect early and enforce firmly with Google Cloud's enhanced cost controls for AI spend

Generative AI can make cloud costs difficult to predict. A single five-word prompt can run complex operations and generate significant costs. Traditional metrics like requests per second no longer hel

→30 Jul 2026Google reveals Gemini Robotics 2.0, promising improved dexterity and safety

Google announced Gemini Robotics 2.0 on July 30 2026, launching a family of three models that enhance robot dexterity, safety, and whole‑body intelligence for humanoid machinery. The publicly released

→3 Aug 2026Real-world mainframe modernization with AI: A safe, scalable path from mainframe to cloud

For too long, enterprises with legacy mainframe estates have been faced with a high-stakes dilemma: continue maintaining their mainframes, essentially kicking the modernization can down the road (they

→4 Aug 2026Future Mode Part 2: The foundation for securing agentic browsing

Editor's Note: Our Future Mode series will give businesses insight into how Chrome Enterprise is approaching AI in the browser. Stay tuned for more blogs in this series.Future Mode Part 2: The foundat

→6 Aug 2026Agentic Future Ready With BigQuery: Continually Improving Price-Performance, Zero Effort

In the modern data landscape, query performance tuning and managing system price-performance is challenging, especially as the number of agentic workloads increase. Even for experienced developers and

CompanyMeta8 recent entries
15 May 2026Send the arXiv AI-generated slop, get a yearlong vacation from submissions

ArXiv will ban authors for one year if they submit papers containing obviously AI-generated content , with examples including hallucinated citations, placeholder text, or chatbot meta-comments left in

→23 Jun 2026Sources: the Trump administration is pressing Meta to submit its AI models for voluntary review; Meta is the only major US AI developer without an agreement (New York Times)

New York Times: Sources: the Trump administration is pressing Meta to submit its AI models for voluntary review; Meta is the only major US AI developer without an agreement — Federal officials are urg

→25 Jun 2026STOCKSTAY Another Day: The Latest Addition to Turla’s Intelligence Gathering Apparatus

Written by: Jordan Jones Introduction Google Threat Intelligence Group (GTIG) has conducted an in-depth analysis of a .NET backdoor, tracked as STOCKSTAY, that has been continually developed and deplo

→3 Jul 2026OpenAI offers feds a stake, Anthropic gets out of AI model jail and Meta wants to be a neocloud

OpenAI reportedly has floated giving the U.S. government a 5% stake in the company, perhaps the start of a series of such stakes in other AI companies as well. This no doubt has traditional anti-indus

→15 Jul 2026Wiki Lint Report — 2026-07-15

Automated lint: 26 errors, 6728 warnings, 3 info

→19 Jul 2026Wiki Lint Report — 2026-07-19

Automated lint: 20 errors, 8743 warnings, 3 info

→21 Jul 2026LWiAI Podcast #252 - GPT 5.6, Grok 4.5, Nemotron-Labs-Diffusion, AI 2040

LWiAI Podcast #252 (July 11, 2026) reviewed major AI releases: OpenAI unveiled GPT‑5.6 and relaunched its agentic coding product as ChatGPT Work, amid disputes over U.S. governmental oversight and jai

→21 Jul 2026Last Week in AI #250 - Mythos Mess, GPT 5.6-Sol, GLM 5.2

Last Week in AI #250 is the podcast’s 250th episode, summarizing recent frontier‑AI policy developments. It reports that the U.S. government has granted Anthropic permission to release Mythos‑5 to a l

CompanyMistral5 recent entries
9 Jun 2026Initial impressions of Claude Fable 5

I didn't have early access to today's Claude Fable 5 release, but I've spent the past ~5.5 hours putting it through its paces. My initial impressions are that this is something of a beast. It's slow,

→15 Jul 2026Wiki Lint Report — 2026-07-15

Automated lint: 26 errors, 6728 warnings, 3 info

→19 Jul 2026Wiki Lint Report — 2026-07-19

Automated lint: 20 errors, 8743 warnings, 3 info

→4 Aug 2026Mistral releases Shieldstral, a 3B multimodal safety classifier that it says matches models up to 7x its size on text safety, available under Apache 2.0 (Mistral AI Blog)

Mistral AI Blog: Mistral releases Shieldstral, a 3B multimodal safety classifier that it says matches models up to 7x its size on text safety, available under Apache 2.0 — Every product that ships a m

→5 Aug 2026Mistral introduces Shieldstral to provide lightweight policy-aware moderation for AI models

French artificial intelligence startup Mistral AI SAS today introduced a lightweight multimodal safety artificial intelligence open-weight model that can classify outputs for AI models that outperform

CompanyxAI7 recent entries
15 Apr 2026A TTP analysis finds dozens of nudify apps in Apple and Google app stores via search, despite company policy prohibiting them; the apps earned $122M+ in revenue (Bloomberg)

Bloomberg: A TTP analysis finds dozens of nudify apps in Apple and Google app stores via search, despite company policy prohibiting them; the apps earned $122M+ in revenue — Apple Inc. and Google have

→6 May 2026Google, Microsoft and xAI agree to allow government safety checks of their AI models prior to release

Google LLC, Microsoft Corp. and xAI have agreed to share unreleased versions of their artificial intelligence models with the U.S. Department of Commerce to ensure the technologies do not pose a threa

→4 Jun 2026Elon Musk petitioned the FTC in May to end its 2022 order restricting Twitter's data use, claiming Twitter no longer exists as X merged with xAI and then SpaceX (Ashley Belanger/Ars Technica)

Ashley Belanger / Ars Technica: Elon Musk petitioned the FTC in May to end its 2022 order restricting Twitter's data use, claiming Twitter no longer exists as X merged with xAI and then SpaceX — Criti

→15 Jul 2026Wiki Lint Report — 2026-07-15

Automated lint: 26 errors, 6728 warnings, 3 info

→19 Jul 2026Wiki Lint Report — 2026-07-19

Automated lint: 20 errors, 8743 warnings, 3 info

→21 Jul 2026LWiAI Podcast #252 - GPT 5.6, Grok 4.5, Nemotron-Labs-Diffusion, AI 2040

LWiAI Podcast #252 (July 11, 2026) reviewed major AI releases: OpenAI unveiled GPT‑5.6 and relaunched its agentic coding product as ChatGPT Work, amid disputes over U.S. governmental oversight and jai

→5 Aug 2026White House, AI firms keep safety framework talks private

The White House met with representatives from leading artificial intelligence companies today to discuss a safety framework for the government to review frontier models prior to launch, although there

CompanyDeepSeek3 recent entries
4 May 2026LWiAI Podcast #243 - GPT 5.5, DeepSeek V4, AI safety sabotage

This podcast episode from Last Week in AI discusses recent developments in large language models, including updates on GPT 5.5 and DeepSeek V4, while also covering concerns about potential sabotage or

→15 Jul 2026Wiki Lint Report — 2026-07-15

Automated lint: 26 errors, 6728 warnings, 3 info

→19 Jul 2026Wiki Lint Report — 2026-07-19

Automated lint: 20 errors, 8743 warnings, 3 info

CompanyNVIDIA8 recent entries
25 Jul 2026Nvidia, other tech giants caution against open-source AI ban in open letter

A group of tech firms has released an open letter that calls on policymakers not to ban open-source artificial intelligence models. The development follows a report that some Trump administration offi

→27 Jul 2026Tech industry leaders join to form Open Secure AI Alliance to promote safety and security

Nvidia Corp. today announced the launch of the Open Secure AI Alliance, a new organization founded by technology, cloud computing and cybersecurity leaders to build and share open artificial intellige

→27 Jul 2026Nvidia, Microsoft launch open AI security alliance — without OpenAI, Google, or Anthropic

Nvidia on Monday said it is joining forces with Microsoft, SpaceX, IBM, and other tech companies to build and share open-source AI security tools. The new Open Secure AI Alliance said open tools are r

→27 Jul 2026Nvidia forms the Open Secure AI Alliance, a coalition including CrowdStrike, Hugging Face, and Dell to develop and share tools for AI safety and cybersecurity (Jaspreet Singh/Reuters)

Jaspreet Singh / Reuters: Nvidia forms the Open Secure AI Alliance, a coalition including CrowdStrike, Hugging Face, and Dell to develop and share tools for AI safety and cybersecurity — Nvidia (NVDA.

→27 Jul 2026Industry Leaders Unite in Open Secure AI Alliance for AI Safety and Security

Open source software is a critical pillar of the global economy. It underpins cloud computing, financial services, manufacturing, telecommunications, government and internet services by making technol

→2 Aug 2026Open letters about AI development

Open letters about AI development I wrote this summary of the past few weeks of open letters as a section of my sponsors-only newsletter but I've decided to share it here as well. Open Weights and Ame

→4 Aug 2026Beyond VLAs: How World Action Models Reshape Robot Manipulation

World Action Models (WAMs) use video-based world modeling instead of vision‑language backbones, giving robots a learned physics engine that supports zero‑shot transfer to new tasks, environments, and

→7 Aug 2026At Black Hat, OpenAI reconstructs the OpenAI-Hugging Face incident and examines its implications for AI security, cyber resilience, and alignment (Black Hat on YouTube)

Black Hat on YouTube: At Black Hat, OpenAI reconstructs the OpenAI-Hugging Face incident and examines its implications for AI security, cyber resilience, and alignment — The ‘Breaking’ News: The OpenA