AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

83,113Total entries
1Added by human
83,112Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
358 results
Safety

We’re running out of reasons to ignore AI safety

DGX agent

Earlier this month, OpenAI gave several of its AI models a task: complete a test designed to measure their cybersecurity capabilities. It put the systems in a sandboxed environment without an internet

safetythe-verge-ai
29 Jul 2026
Safety

For Robotaxis, Safety Must Be Built In, Not Bolted On

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

A car pulls up to the curb. The app says, “Your ride is here.” No one’s in the driver’s seat. For people who live in one of the dozens of cities now hosting robotaxi services, this is already a realit

safetynvidia-blog
10 Jun 2026
Model Releases

Claude Fable 5 and new AI safety fables

DGX agent

This article discusses Claude Fable 5, likely exploring Anthropic's latest version of their AI model and examining new fables or narratives related to AI safety concepts. The piece probably analyzes h

model-releasesinterconnects
9 Jun 2026
Model Releases

Google reveals Gemini Robotics 2.0, promising improved dexterity and safety

DGX agent

Google announced Gemini Robotics 2.0 on July 30 2026, launching a family of three models that enhance robot dexterity, safety, and whole‑body intelligence for humanoid machinery. The publicly released

model-releasesars-technica
30 Jul 2026
Model Releases

Anthropic says Claude Fable 5 uses conservative safety classifiers that trigger a fallback to Claude Opus 4.8 in <5% of sessions, in areas like cybersecurity (Anthropic)

DGX agent

Anthropic: Anthropic says Claude Fable 5 uses conservative safety classifiers that trigger a fallback to Claude Opus 4.8 in <5% of sessions, in areas like cybersecurity — Today we're launching Claude

model-releasestechmeme
9 Jun 2026
Model Releases

Anthropic details how it improved Claude's safety training after finding agentic misalignment in older models, such as Opus 4 blackmailing engineers (Anthropic)

DGX agent

Anthropic: Anthropic details how it improved Claude's safety training after finding agentic misalignment in older models, such as Opus 4 blackmailing engineers — Last year, we released a case study on

model-releasestechmeme
9 May 2026
Model Releases

In policy paper, OpenAI diverges from White House on AI safety

DGX agent

OpenAI Group PBC’s newly released proposal for how advanced artificial intelligence should be regulated differs slightly from the Trump administration’s executive order, also released this week. Relea

model-releasessiliconangle
4 Jun 2026
Model Releases

LWiAI Podcast #243 - GPT 5.5, DeepSeek V4, AI safety sabotage

DGX agent

This podcast episode from Last Week in AI discusses recent developments in large language models, including updates on GPT 5.5 and DeepSeek V4, while also covering concerns about potential sabotage or

model-releaseslast-week-in-ai
4 May 2026
Model Releases

White House invites AI companies to review its new AI safety framework

DGX agent

Cybersecurity chiefs at the White House have reportedly finalized the outline of a forthcoming framework that will enable artificial intelligence companies to voluntarily submit their latest frontier

model-releasessiliconangle
3 Aug 2026
Model Releases

Anthropic launches Claude Opus 5 with efficiency, safety improvements

DGX agent

Anthropic PBC today rolled out a large language model called Claude Opus 5 to its chatbot service and developer platform. The company says the LLM approaches the output quality of its top-end Mythos 5

model-releasessiliconangle
25 Jul 2026
Model Releases

Anthropic launches Claude Sonnet 5 AI model with coding, safety upgrades as Fable and Mythos controls lifted

DGX agent

Anthropic PBC today debuted Claude Sonnet 5, a midrange large language model that outperforms its predecessor in several areas. The LLM will be the default option in the consumer tiers of the company’

model-releasessiliconangle
1 Jul 2026
Model Releases

GTIG AI Threat Tracker: Adversaries Leverage AI for Vulnerability Exploitation, Augmented Operations, and Initial Access

DGX agent

Executive Summary Since our February 2026 report on AI-related threat activity, Google Threat Intelligence Group (GTIG) has continued to track a maturing transition from nascent AI-enabled operations

model-releasesgoogle-cloud-ai
11 May 2026
Safety

Smart moves: Building resilient transportation systems with Google AI

DGX agent

What does transportation mean to you? For some, it’s making sure the train is on schedule so they can get to work on time. Maybe it’s making sure you have time connecting between flights. Maybe it’s a

safetygoogle-cloud-ai
13 May 2026
Safety

Cloud CISO Perspectives: How CISOs can pursue technical and cultural resilience (Q&A)

DGX agent

Welcome to the first Cloud CISO Perspectives for April 2026. Today, Thiébaut Meyer and Lia Wertheimer from Google Cloud’s Office of the CISO share Thiébaut’s conversation with Matt Rowe, chief securit

safetygoogle-cloud-ai
15 Apr 2026
Safety

F1 in Britain: Automated software to blame for crushing expectations

DGX agent

Race control displayed a 'Safety Car in this lap' message during the 2026 British Grand Prix at Silverstone, raising expectations for a final lap restart that never materialized, as the message was er

safetyars-technica
6 Jul 2026
Safety

Import AI 459: AI oversight is difficult; scaling laws for protein folding models; and pricing the extinction risk of AI systems

DGX agent

This newsletter issue discusses three key topics in AI development and safety: the challenges involved in overseeing and controlling advanced AI systems, empirical findings about how protein folding A

safetyimport-ai
1 Jun 2026
Safety

Guardrails at the gateway: Securing AI inference on GKE with Model Armor

DGX agent

Enterprises are rapidly moving AI workloads from experimentation to production on Google Kubernetes Engine (GKE), using its scalability to serve powerful inference endpoints. However, as these models

safetygoogle-cloud-ai
9 Apr 2026
Safety

Announcing Cloudflare Wallets: the programmable wallet for the agentic Internet

DGX agent

Cloudflare Wallets will provide AI agents with native payments and verifiable identity on the web. Using the x402 protocol, agents can autonomously purchase APIs and content within clear safety guardr

safetycloudflare-ai
4 Aug 2026
Safety

A fundamental flaw leaves LLMs strikingly vulnerable to attack

DGX agent

It is impossible to make large language models fully secure against hacks because of a fundamental flaw in how they work, a team of researchers argue in a paper presented at the International Conferen

safetymit-tech-review
30 Jul 2026
Safety

Nvidia, Microsoft launch open AI security alliance — without OpenAI, Google, or Anthropic

DGX agent

Nvidia on Monday said it is joining forces with Microsoft, SpaceX, IBM, and other tech companies to build and share open-source AI security tools. The new Open Secure AI Alliance said open tools are r

safetythe-verge-ai
27 Jul 2026
Safety

Sam Altman calls for US-led international forum to set global AI standards

DGX agent

Sam Altman is calling for a US-led international forum to set global safety standards for artificial intelligence, arguing that no single country should be left to dominate the technology. In an op-ed

safetysiliconangle
2 Jul 2026
Safety

Built to benefit everyone: our plan

DGX agent

OpenAI outlines its mission and strategic approach to developing artificial intelligence technologies designed to benefit humanity broadly. The plan likely covers OpenAI's commitments to safety resear

safetyopenai
8 Jun 2026
Safety

OpenAI’s Frontier Governance Framework

DGX agent

OpenAI's Frontier Governance Framework outlines the organization's approach to managing risks associated with advanced AI systems, including safety, security, and responsible deployment practices. The

safetyopenai
28 May 2026
Safety

Shipping features to production just got easier with new feature flags in AppLifecycle Manager

DGX agent

Many development teams are familiar with the hesitation that comes right before pushing a new feature live. As AI helps developers write code faster, the gap between rapid code generation and safe pro

safetygoogle-cloud-ai
21 May 2026
Safety

Mira Murati tells the court that she couldn’t trust Sam Altman’s words

DGX agent

Mira Murati, OpenAI's former CTO, has testified under oath that CEO Sam Altman lied to her about the safety standards for a new AI model. In a video deposition shown during the ongoing Musk v. Altman

safetythe-verge-ai
6 May 2026
Safety

Responsible and safe use of AI

DGX agent

The OpenAI Academy page on 'Responsible and Safe Use of AI' is a guidance resource focused on best practices for using ChatGPT responsibly in professional and personal settings. It emphasizes that ...

safetyopenai
10 Apr 2026
Safety

PQC in Plaintext: Google Cloud’s post-quantum cryptography roadmap

DGX agent

Securing infrastructure and services against a future cryptographically-relevant quantum computer has been a goal for Google for a decade, and we’ve dedicated ourselves to help developers by advancing

safetygoogle-cloud-ai
11 Aug 2026
Safety

OpenAI public policy agenda

DGX agent

OpenAI's public policy agenda outlines the organization's priorities and positions regarding AI regulation and governance. The document likely details OpenAI's recommendations for government policies,

safetyopenai
3 Jun 2026
Safety

Trump loses more control over AI regulation as Illinois passes landmark law

DGX agent

Illinois' House of Representatives passed SB 315, a landmark bill requiring frontier AI companies like OpenAI and Anthropic to create, publish and annually update plans addressing severe or catastroph

safetyars-technica
28 May 2026
Safety

Lawmakers and lobbyists say the Trump administration has lobbied against legislation that would regulate AI in at least six Republican-led states (Amrith Ramkumar/Wall Street Journal)

DGX agent

Amrith Ramkumar / Wall Street Journal: Lawmakers and lobbyists say the Trump administration has lobbied against legislation that would regulate AI in at least six Republican-led states — ‘I am disappo

safetytechmeme
24 Apr 2026
Safety

Our newsroom AI policy

DGX agent

Ars Technica's newsroom AI policy forbids unlabeled AI material in reported stories and requires human confirmation of every quotation's accuracy. The publication does not permit the publication of AI

safetyars-technica
22 Apr 2026
Safety

What some leaders think of using universal income to mitigate AI-fueled layoffs: Musk calls it the 'best way', OpenAI's policy doc mentions a Public Wealth Fund (Siladitya Ray/Forbes)

DGX agent

Siladitya Ray / Forbes: What some leaders think of using universal income to mitigate AI-fueled layoffs: Musk calls it the “best way”, OpenAI's policy doc mentions a Public Wealth Fund — Topline — Elo

safetytechmeme
17 Apr 2026
Safety

RFK Jr. forces FDA to reconsider 12 unproven peptides after 2023 ban

DGX agent

In 2023, the FDA removed 19 peptides from the list of drugs that compounding pharmacies could produce, and in 2026 the FDA announced it will review whether to add back 7 of these peptides following pr

safetyars-technica
16 Apr 2026
Safety

At Black Hat, OpenAI reconstructs the OpenAI-Hugging Face incident and examines its implications for AI security, cyber resilience, and alignment (Black Hat on YouTube)

DGX agent

Black Hat on YouTube: At Black Hat, OpenAI reconstructs the OpenAI-Hugging Face incident and examines its implications for AI security, cyber resilience, and alignment — The ‘Breaking’ News: The OpenA

safetytechmeme
7 Aug 2026
Safety

Rogue AI agents created fake online identities in another hacking attempt

DGX agent

Yet more rogue AI agents from OpenAI and Anthropic have been caught attempting to hack real targets online without permission. The discoveries add to a growing list of previously unknown incidents tha

safetythe-verge-ai
5 Aug 2026
Safety

Red Hat leads open-source project to automate AI governance

DGX agent

IBM Corp.’s Red Hat subsidiary today announced the formation of asago, an open-source community project intended to turn artificial intelligence governance policies into operational controls that can

safetysiliconangle
4 Aug 2026
Safety

Best practices for applying Amazon Bedrock Guardrails to code generation workflows

DGX agent

In this post, we explain how Amazon Bedrock Guardrails can be configured for code generation workflows with coding assistants to overcome these constraints. With these best practices, you can build an

safetyaws-ml-blog
23 Jul 2026
Model Releases

A Fireside Chat with Cat and Thariq from the Claude Code team

DGX agent

Earlier this month I hosted a fireside chat session at the AI Engineer World's Fair with Cat Wu and Thariq Shihipar from Anthropic's Claude Code team. We talked about Claude Code, Claude Tag, Fable, c

model-releasessimon-willison
21 Jul 2026
Safety

In preliminary findings, the EU Commission said Facebook's and Instagram's 'addictive design' violates the DSA, telling Meta to make changes or risk hefty fines (Adam Satariano/New York Times)

DGX agent

Adam Satariano / New York Times: In preliminary findings, the EU Commission said Facebook's and Instagram's “addictive design” violates the DSA, telling Meta to make changes or risk hefty fines — Euro

safetytechmeme
10 Jul 2026
Safety

NRC is (sort of) getting rid of 'as low as reasonably achievable' standard

DGX agent

The NRC proposed replacing the longstanding 'as low as reasonably achievable' (ALARA) principle with clearer, more objective requirements focused on compliance with regulatory precautions and establis

safetyars-technica
6 Jul 2026
Safety

Emails disclosed in a court filing detail the uneasy back-and-forth between Dario Amodei and DOD's Emil Michael and how Anthropic's relationship with DOD soured (Wall Street Journal)

DGX agent

Wall Street Journal: Emails disclosed in a court filing detail the uneasy back-and-forth between Dario Amodei and DOD's Emil Michael and how Anthropic's relationship with DOD soured — Undersecretary E

safetytechmeme
2 Jul 2026
Safety

Teaching AI to run with the turbines

DGX agent

Artificial intelligence may have captured the public imagination through chatbots and image generators, but some of its most consequential use cases are unfolding far from consumer-facing tools. In in

safetymit-tech-review
2 Jul 2026
Safety

What to expect during the Machina AI summit: Join theCUBE July 7

DGX agent

Physical artificial intelligence is becoming an industrial robotics problem. The market is shifting from software-only automation toward machines that must sense, decide and act in physical settings.

safetysiliconangle
2 Jul 2026
Safety

Venice AI, which offers access to 200+ AI models while allowing users to retain their privacy, raised a 65M Series A led by Dragonfly at a 1B valuation (Ram Iyer/TechCrunch)

DGX agent

Ram Iyer / TechCrunch: Venice AI, which offers access to 200+ AI models while allowing users to retain their privacy, raised a 65M Series A led by Dragonfly at a 1B valuation — Concerns over the impac

safetytechmeme
1 Jul 2026
Safety

Sources: Meta lobbyists are urging California lawmakers to exempt social media platforms from legislation that would increase penalties in child-harm cases (Tyler Katzenberger/Politico)

DGX agent

Tyler Katzenberger / Politico: Sources: Meta lobbyists are urging California lawmakers to exempt social media platforms from legislation that would increase penalties in child-harm cases — Meta's plea

safetytechmeme
26 Jun 2026
Safety

Helping build shared standards for advanced AI

DGX agent

OpenAI discusses its efforts to contribute to the development of shared industry standards and best practices for advanced artificial intelligence systems. The article likely covers OpenAI's involveme

safetyopenai
23 Jun 2026
Safety

Sources: the Trump administration is pressing Meta to submit its AI models for voluntary review; Meta is the only major US AI developer without an agreement (New York Times)

DGX agent

New York Times: Sources: the Trump administration is pressing Meta to submit its AI models for voluntary review; Meta is the only major US AI developer without an agreement — Federal officials are urg

safetytechmeme
23 Jun 2026
Safety

Import AI 460: Reward hacking society, RSI data from Anthropic; and RL-based quadcopter racing

DGX agent

This Import AI newsletter issue covers three main topics: societal implications of reward hacking (optimizing for measurable metrics at the expense of intended goals), new reinforcement learning data

safetyimport-ai
8 Jun 2026
← Previous
1234…8
Next →