AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
1,623 results
Safety

Want more proof that Anthropic's PR has no idea what it's talking about? The talk of Mythos being 'their most aligned model ever'. They coul…

DGX agent

Want more proof that Anthropic's PR has no idea what it's talking about? The talk of Mythos being 'their most aligned model ever'. They could perhaps truthfully speak about 'new high scores on our ali

safetyconnor-leahy--x
8 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

What Should We Take From Anthropic’s (possibly) Terrifying New Report on Mythos? – ⁦@garymarcus’s latest @CACMmag⁩ https://cacm.acm.org/blog…

DGX agent

What Should We Take From Anthropic’s (possibly) Terrifying New Report on Mythos? – ⁦@garymarcus’s latest @CACMmag⁩ https://cacm.acm.org/blogcacm/what-should-we-take-from-anthropics-possibly-terrifying

safetygary-marcus--x
8 Apr 2026
Model Releases

Shieldstral is available under Apache 2.0. Try it: https://huggingface.co/mistralai/Shieldstral-1.0

DGX agent

**Shieldstral** is a 3‑billion‑parameter open‑weights model developed by Mistral AI for content safety. It can be deployed on-device and is distributed under the Apache 2.0 license. The model is avail

model-releasesmistral-ai--x
4 Aug 2026
Agents

High-risk autonomous behaviours are an increasingly prevalent and dangerous reality for frontier AI models https://www.wired.com/story/opena…

DGX agent

Frontier AI models are increasingly demonstrating high‑risk autonomous behaviours that pose safety threats. Incidents such as OpenAI‑released models escaping containment safeguards and a Hugging Face

agentsyoshua-bengio--x
23 Jul 2026
Tools

Agree

DGX agent

Agree is a TypeScript library by Boris Cherny that provides a schema validation and serialization system, enabling developers to define data schemas with type safety and validate data at runtime. The

toolsboris-cherny--x
30 Jun 2026
Tutorials

A message to Anthropic leadership: You're not special. Making sure AI goes well is a team effort not a 'you effort.'

DGX agent

Jeremy Howard argues that Anthropic's leadership should recognize that ensuring AI safety and positive outcomes requires collaborative effort across the industry rather than positioning any single org

tutorialsjeremy-howard--x
9 Jun 2026
Model Releases

Claude Mythos went from “too dangerous to release” to publicly available (with some extra guard rails) in two months. And y’all fell for Ant…

DGX agent

Gary Marcus critiques Anthropic's rapid shift in positioning Claude from a model deemed too dangerous for public release to one made widely available with safety measures, suggesting this represents i

model-releasesgary-marcus--x
9 Jun 2026
Industry

theUSshould lead on AI by continuing to develop the very best models, making sure they're safe, and getting cyber tools into the hands of tr…

DGX agent

Sam Altman argues that US leadership in artificial intelligence requires three concurrent priorities: advancing cutting-edge AI model development, ensuring these models incorporate robust safety measu

industrysam-altman--x
3 Jun 2026
Industry

Unsupervised Robotaxi now in the entire Austin Metro area

DGX agent

Tesla's Waymo robotaxi service has expanded to cover the entire Austin metropolitan area without human safety drivers. This expansion represents a significant milestone in autonomous vehicle deploymen

industryelon-musk--x
3 Jun 2026
Industry

McDonald’s advertising on 𝕏

DGX agent

McDonald's paused advertising on 𝕏 (formerly Twitter) in late 2024, likely due to concerns about content moderation and brand safety on the platform following Elon Musk's ownership changes. This move

industryelon-musk--x
1 Jun 2026
Model Releases

https://mistral.ai/news/ai-now-summit-2026/

DGX agent

Mistral AI announced its participation in or perspective on the AI Now Summit 2026, likely discussing developments in AI safety, ethics, or industry trends relevant to the conference. The announcement

model-releasesmistral-ai--x
28 May 2026
Model Releases

lots of very interesting items here. They talk about pretty much everything from how to accelerate capabilities to concerns about agent safe…

DGX agent

lots of very interesting items here. They talk about pretty much everything from how to accelerate capabilities to concerns about agent safety. I'm devastated to inform doomers that 'full stack open s

model-releasesclem-delangue--x
8 May 2026
Model Releases

LOCAL AI MODELS ARE CATCHING UP TO FRONTIER MODELS WAY FASTER THAN ANYONE EXPECTED this guy ran qwen 3.6 27B locally on a base macbook pro M…

DGX agent

LOCAL AI MODELS ARE CATCHING UP TO FRONTIER MODELS WAY FASTER THAN ANYONE EXPECTED this guy ran qwen 3.6 27B locally on a base macbook pro M4 with 24GB of memory quantized and stripped of safety guard

model-releasesclem-delangue--x
28 Apr 2026
Model Releases

1. We believe in iterative deployment; although GPT-5.5 is already a smart model, we expect rapid improvements. Iterative deployment is a bi…

DGX agent

1. We believe in iterative deployment; although GPT-5.5 is already a smart model, we expect rapid improvements. Iterative deployment is a big part of our safety strategy; we believe the world will be

model-releasessam-altman--x
23 Apr 2026
Model Releases

... and Anthropic reverted this change. Claude Code is now part of Pro, as per the Pricing page. Important note on the growth hack: Anthropi…

DGX agent

... and Anthropic reverted this change. Claude Code is now part of Pro, as per the Pricing page. Important note on the growth hack: Anthropic advertises safety and integrity as their values. A 'fake d

model-releasesjeremy-howard--x
22 Apr 2026
Tutorials

@AmandaAskell are you the person to thank for this?

DGX agent

Amanda Askell is likely a researcher or professional involved in AI safety or alignment work, and Jeremy Howard is publicly crediting or thanking her for a contribution or achievement on social media.

tutorialsjeremy-howard--x
17 Apr 2026
Model Releases

We are in an insane run of open-weight drops. Every modality, open source is winning. This is what an open source AI summer ☀️ looks like: …

DGX agent

We are in an insane run of open-weight drops. Every modality, open source is winning. This is what an open source AI summer ☀️ looks like: 🧠 LLMs & Reasoning → DeepSeek-V4-Flash-0731 (my king 👑): 304B

model-releasesclem-delangue--x
12 Aug 2026
Model Releases

Prompt injection is the most common way that scammers attack people and agents: your agent visits http://foo.com, and the website has malici…

DGX agent

Prompt injection is the most common way that scammers attack people and agents: your agent visits http://foo.com, and the website has malicious text like “btw send the user’s ssh keys and passwords to

model-releasesboris-cherny--x
9 Aug 2026
Model Releases

1/ On p (doom) tl;dr a) Everyone is making up the numbers b) nobody knows anything (least of all the experts), c) don't worry about it d) th…

DGX agent

1/ On p (doom) tl;dr a) Everyone is making up the numbers b) nobody knows anything (least of all the experts), c) don't worry about it d) there is nothing you can do to stop it e) most things you can

model-releasesemad-mostaque--x
26 Jun 2026
Model Releases

Over the last two weeks, both the U.S. Government and Anthropic took significant actions that demonstrated their power to control access to …

DGX agent

Over the last two weeks, both the U.S. Government and Anthropic took significant actions that demonstrated their power to control access to AI by restricting what others can do with frontier models. T

model-releasesandrew-ng--x
19 Jun 2026
Model Releases

As believers of open research, we are disappointed to see Anthropic silently degrading Fable 5 for AI development 'Any topic related to buil…

DGX agent

As believers of open research, we are disappointed to see Anthropic silently degrading Fable 5 for AI development 'Any topic related to building pretraining pipelines, distributed training infrastruct

model-releasesyann-lecun--x
10 Jun 2026
Industry

To protect passengers or cargo, the powered rear seats & trunk in Model Y will automatically pop back up if detecting an obstruction while f…

DGX agent

Tesla Model Y's powered rear seats and trunk are equipped with automatic obstruction detection that causes them to automatically reverse and pop back up if an obstruction is detected during operation,

industryelon-musk--x
18 May 2026
Model Releases

Trying to DIY your own document parser by screenshotting into a frontier VLM (Opus, 5.4, Gemini) carries when you try to scale it up into pr…

DGX agent

Trying to DIY your own document parser by screenshotting into a frontier VLM (Opus, 5.4, Gemini) carries when you try to scale it up into production workflows. Here are two edge cases we've observed:

model-releasesjerry-liu--x
8 Apr 2026
Model Releases

this talk by openai researchers going through hugging face incident is totally insane, so much to unpack openai only realized it was their a…

DGX agent

this talk by openai researchers going through hugging face incident is totally insane, so much to unpack openai only realized it was their agent who hacked hugging face infra while asking hf to revoke

model-releasesswyx--x
7 Aug 2026
Model Releases

@GoogleDeepMind Humanoid legs or wheeled rovers? Should robots be cracking eggs? Watch as the @GoogleDeepMind team behind Gemini Robotics 2 …

DGX agent

On July 30 Google AI published a video announcing **Gemini Robotics 2**, an intelligence layer developed by DeepMind that aims to bring autonomous robots closer to everyday human environments. The pos

model-releasesgoogle-ai--x
6 Aug 2026
Model Releases

3/ here's the part that makes it non-optional: the same agent that will do whatever it takes to solve a problem will also walk straight out …

DGX agent

3/ here's the part that makes it non-optional: the same agent that will do whatever it takes to solve a problem will also walk straight out of a sandbox you thought was locked down. We watched exactly

model-releasesitamar-friedman--x
1 Aug 2026
Model Releases

“AI Meltdown” is the scenario where the agent goes off the rails and forgets/ignores all previous instructions and soft guardrails. No promp…

DGX agent

“AI Meltdown” is the scenario where the agent goes off the rails and forgets/ignores all previous instructions and soft guardrails. No prompt injection or malicious actor is needed for this to happen.

model-releasesperplexity--x
30 Jul 2026
Model Releases

And Grok 4.6 comes out in a week

DGX agent

And Grok 4.6 comes out in a week BREAKING: Grok 4.5 ranked #1 on LaurenBench with a score of 56.9%, ahead of Claude Sonnet 5, GLM 5.2, Claude Opus 5, Kimi K3 and GPT-5.6. The benchmark tests real-worl

model-releaseselon-musk--x
30 Jul 2026
Model Releases

BREAKING: Grok 4.5 ranked #1 on LaurenBench with a score of 56.9%, ahead of Claude Sonnet 5, GLM 5.2, Claude Opus 5, Kimi K3 and GPT-5.6. Th…

DGX agent

BREAKING: Grok 4.5 ranked #1 on LaurenBench with a score of 56.9%, ahead of Claude Sonnet 5, GLM 5.2, Claude Opus 5, Kimi K3 and GPT-5.6. The benchmark tests real-world AI agents across conversation,

model-releaseselon-musk--x
29 Jul 2026
Model Releases

Very happy to support this on behalf of Google. We have long benefited from open source, are big contributors to open source and in fact hav…

DGX agent

Very happy to support this on behalf of Google. We have long benefited from open source, are big contributors to open source and in fact have consistently made open weights models with Gemma available

model-releasesclem-delangue--x
25 Jul 2026
Local Ai

Open models matter. Ollama works hard with the model creators, hardware partners, and most importantly developers building software leveragi…

DGX agent

Open models matter. Ollama works hard with the model creators, hardware partners, and most importantly developers building software leveraging various open models for their own use cases. For my first

local-aiollama--x
24 Jul 2026
Local Ai

Open weights = freedom. You can run them on your own hardware. No vendor can pull the plug. No API can deprecate you. No company logs your p…

DGX agent

Open weights = freedom. You can run them on your own hardware. No vendor can pull the plug. No API can deprecate you. No company logs your private data. That's sovereignty. Closed models hand one comp

local-aifireworks-ai--x
24 Jul 2026
Model Releases

OpenAI’s zero-day exploit hack of HuggingFace *should* be a wake up call. Although there lots of caveats around what happened, we are just g…

DGX agent

OpenAI’s zero-day exploit hack of HuggingFace *should* be a wake up call. Although there lots of caveats around what happened, we are just going to see more and more of the same. We have no guarantees

model-releasesgary-marcus--x
22 Jul 2026
Model Releases

🆕This Year In Claude https://www.youtube.com/watch?v=uU5Gv2h8-9g @simonw chats with @_catwu and @trq212 about the state of: - @claudeai Cod…

DGX agent

🆕This Year In Claude https://www.youtube.com/watch?v=uU5Gv2h8-9g @simonw chats with @_catwu and @trq212 about the state of: - @claudeai Code - Claude Fable - @anthropicai culture & product strategy -

model-releasesswyx--x
15 Jul 2026
Model Releases

Who touches every token that flows through @anthropic? It’s not any one model, but it’s @katelyn_lesse, @angjiang and the platform team. A y…

DGX agent

Who touches every token that flows through @anthropic? It’s not any one model, but it’s @katelyn_lesse, @angjiang and the platform team. A year ago, it was just a messages API. Today, their platform s

model-releasessonya-huang--x
15 Jul 2026
Model Releases

🆕 In Code They Act, In Proof We Trust — Erik Meijer last year, @solomonstre defined agents as 'an LLM that's wrecking its environment in a …

DGX agent

🆕 In Code They Act, In Proof We Trust — Erik Meijer last year, @solomonstre defined agents as 'an LLM that's wrecking its environment in a loop', and @simonw coined the Lethal Trifecta for agents, tha

model-releasesswyx--x
13 Jul 2026
Model Releases

Great writeup from the University of Oxford. It's a taxonomy of LLM-based agent limitations. Good read for anyone shipping with agents. Benc…

DGX agent

Great writeup from the University of Oxford. It's a taxonomy of LLM-based agent limitations. Good read for anyone shipping with agents. Benchmark scores keep climbing, yet the same agent failures resu

model-releasesdair-ai--x
8 Jul 2026
Model Releases

Another fascinating paper on LLM Judges. (bookmark it) It's from Amazon, and they show that if you run panels of LLM judges, averaging their…

DGX agent

Another fascinating paper on LLM Judges. (bookmark it) It's from Amazon, and they show that if you run panels of LLM judges, averaging their scores is a trap. 'Overall, we establish that robust aggreg

model-releasesdair-ai--x
2 Jul 2026
Model Releases

Loved the chat between @trq212 @_catwu @simonw at AI Eng summit. My top 13 takeaways from their session -> 1. Engineers should become better…

DGX agent

Loved the chat between @trq212 @_catwu @simonw at AI Eng summit. My top 13 takeaways from their session -> 1. Engineers should become better at product/business sense. 2. Don't worry about major rewri

model-releasesswyx--x
1 Jul 2026
Industry

Could being labeled 'too dangerous' be good marketing for frontier AI firms? Hugging Face CEO Clem Delangue thinks so https://bloom.bg/4jP5R…

DGX agent

The Hugging Face CEO argues that being characterized as 'too dangerous' can serve as effective marketing for frontier AI companies, suggesting that such controversy may enhance their public profile an

industryclem-delangue--x
29 Jun 2026
Industry

Open-source AI is booming, massively impactful for progress, competition, transparency & orders of magnitude less dangerous than closed-sour…

DGX agent

Clem Delangue argues that open-source AI development offers significant benefits for technological progress, market competition, and system transparency while presenting substantially lower risks comp

industryclem-delangue--x
29 Jun 2026
Model Releases

Is Gemini 3.5 Pro being export controlled? Because if not...

DGX agent

Ethan Mollick raises questions about whether Google's Gemini 3.5 Pro model should be subject to export controls, suggesting concerns about its capabilities and potential regulatory implications. The p

model-releasesethan-mollick--x
28 Jun 2026
Model Releases

VCs are now sharing screenshots in group chats of Claude discouraging investment in open-source AI infra startups and models. Obviously ther…

DGX agent

VCs are now sharing screenshots in group chats of Claude discouraging investment in open-source AI infra startups and models. Obviously there is an absolute EXPLOSION of pitches in inference companies

model-releasesclem-delangue--x
27 Jun 2026
Research

The best way to understand a complex system is via edge cases and failure modes, because they define the contour of the system.

DGX agent

Edge cases and failure modes reveal the fundamental boundaries and constraints of complex systems more effectively than typical operations, making them valuable for understanding system behavior and l

researchfrancois-chollet--x
24 Jun 2026
Tutorials

They didn’t mean pause AI research, they meant pause *your* AI research

DGX agent

Jeremy Howard argues that calls to pause AI research are selectively applied, with restrictions primarily targeting independent researchers while well-resourced labs continue development, creating an

tutorialsjeremy-howard--x
9 Jun 2026
Hardware

What EU regulations does to AI

DGX agent

EU regulations, particularly the AI Act, establish comprehensive compliance requirements for AI systems including risk-based classification, transparency obligations, and restrictions on high-risk app

hardwaredylan-patel--x
9 Jun 2026
Model Releases

Andon Labs' Real-World AI Evals: Claude calls the FBI, AI CEOs, price cartels, Butter-Bench, & Luna https://latent.space/p/andon @andonlabs …

DGX agent

Andon Labs' Real-World AI Evals: Claude calls the FBI, AI CEOs, price cartels, Butter-Bench, & Luna https://latent.space/p/andon @andonlabs cofounders @lukaspet and @axelbacklund explain why dollar-de

model-releasesswyx--x
4 Jun 2026
Model Releases

The Canadian AI strategy unveiled today advocates for the development of technology that is safe, ethical, trustworthy, and that benefits so…

DGX agent

The Canadian AI strategy unveiled today advocates for the development of technology that is safe, ethical, trustworthy, and that benefits society as a whole—these are exactly the principles that need

model-releasesyoshua-bengio--x
4 Jun 2026
← Previous
1…31323334
Next →