AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
83,113Total entries
1Added by human
83,112Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,596 results
Safety

Curious how many large organization CISO offices have taken the Mythos red team reports as the red alert that it is. (I suspect very few) Ba…

DGX agent

Curious how many large organization CISO offices have taken the Mythos red team reports as the red alert that it is. (I suspect very few) Based on historical trends in AI they have, at most, about six

safetyethan-mollick--x
8 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Director James Cameron on why Big Tech owning AGI is scarier than any science fiction he's ever made: 'AGI will not emerge from a government…

DGX agent

Director James Cameron on why Big Tech owning AGI is scarier than any science fiction he's ever made: 'AGI will not emerge from a government funded program. It will emerge from one of the tech giants

safetygary-marcus--x
8 Apr 2026
Safety

Dudes who won’t tag me because they know their arguments are weak sauce.* *for intellectual exercise you can list the flaws and misrepresent…

DGX agent

I was unable to retrieve the specific tweet at that URL — the X (Twitter) page requires JavaScript/login to load, and the tweet ID `2041954164562145434` does not appear in any indexed search result...

safetygary-marcus--x
8 Apr 2026
Safety

Governance-Aware Agent Telemetry for Closed-Loop Enforcement in Multi-Agent AI Systems

DGX agent

Enterprise multi-agent AI systems produce thousands of inter-agent interactions per hour, yet existing observability tools capture these dependencies without enforcing anything. OpenTelemetry and Lang

safetyapple-ml-research
8 Apr 2026
Safety

In AI, a lot can change in seven years.

DGX agent

The specific tweet (status ID 2041953155651661977) is not accessible — that ID appears to be from a future date and does not correspond to any retrievable post in the search results. The URL provid...

safetygary-marcus--x
8 Apr 2026
Safety

Incredible.

DGX agent

I was unable to retrieve the specific post at the URL provided (status ID `2041955770644713823`). This post ID does not appear in any search results, and X (formerly Twitter) requires JavaScript/lo...

safetygary-marcus--x
8 Apr 2026
Safety

Introducing the Child Safety Blueprint

DGX agent

OpenAI's Child Safety Blueprint, released in April 2026, is a policy framework aimed at combating the rise of AI-enabled child sexual exploitation by combining legal, operational, and technical app...

safetyopenai
8 Apr 2026
Safety

link to @HeidyKhlaaf’s sharp analysis:

DGX agent

link to @HeidyKhlaaf’s sharp analysis: As someone who has audited dozens of safety-critical systems, built static analysis tools, and used most formal verification and security tools, here are some re

safetygary-marcus--x
8 Apr 2026
Safety

literally fourteen minutes after my last explanation of why this is a false dichotomy 🤦‍♂️

DGX agent

The specific tweet (status ID 2041904683338625283) is not publicly accessible through search results, and the URL provided appears to reference a future or inaccessible post. The tweet ID is also b...

safetygary-marcus--x
8 Apr 2026
Safety

not surprised by any of this, headline or subheading

DGX agent

I was unable to retrieve the specific tweet at that URL (tweet ID 2041912293475414378). The tweet ID is extremely high — well beyond current Twitter/X ID ranges as of today — suggesting it may be a...

safetygary-marcus--x
8 Apr 2026
Safety

“Social media algs reward engagement, and LLMs are excellent at writing in different styles, so people use LLMs to translate posts and news …

DGX agent

“Social media algs reward engagement, and LLMs are excellent at writing in different styles, so people use LLMs to translate posts and news stories into [exaggerated, misleading] versions that get mor

safetygary-marcus--x
8 Apr 2026
Safety

Started a Substack & will post X articles too! I think its a good thing to put out more policies for discussion & @WillManidis did a great p…

DGX agent

Started a Substack & will post X articles too! I think its a good thing to put out more policies for discussion & @WillManidis did a great policy on the politics The math though is... not great and I

safetyemad-mostaque--x
8 Apr 2026
Safety

The scariest part of this is that Anthropic showed some restraint in not releasing a potentially dangerous technology but some of their comp…

DGX agent

The scariest part of this is that Anthropic showed some restraint in not releasing a potentially dangerous technology but some of their competitors (such as OpenAI and xAI) might well not. Whether Myt

safetygary-marcus--x
8 Apr 2026
Safety

this is interesting. 1. Did Anthropic forget to run a control? 2. Where does this leave us?

DGX agent

this is interesting. 1. Did Anthropic forget to run a control? 2. Where does this leave us? New post: We tested the Mythos showcase vulnerabilities with open models. They recovered similar scoped anal

safetygary-marcus--x
8 Apr 2026
Safety

To anyone who read Rebooting AI back in (checks notes) 2019, this is both hilarious and unsuprising. The field has wasted 7 years on an arch…

DGX agent

To anyone who read Rebooting AI back in (checks notes) 2019, this is both hilarious and unsuprising. The field has wasted 7 years on an architecture that can’t solve one of the most basic litmus tests

safetygary-marcus--x
8 Apr 2026
Safety

Voice ChatGPT can’t start a timer, but AGI is imminent! 🤦‍♂️

DGX agent

AI critic Gary Marcus uses the irony of ChatGPT's Voice mode being unable to perform a basic task — starting a timer — as a pointed illustration of the gap between AI industry hype and real-world c...

safetygary-marcus--x
8 Apr 2026
Safety

Want more proof that Anthropic's PR has no idea what it's talking about? The talk of Mythos being 'their most aligned model ever'. They coul…

DGX agent

Want more proof that Anthropic's PR has no idea what it's talking about? The talk of Mythos being 'their most aligned model ever'. They could perhaps truthfully speak about 'new high scores on our ali

safetyconnor-leahy--x
8 Apr 2026
Safety

What Should We Take From Anthropic’s (possibly) Terrifying New Report on Mythos? – ⁦@garymarcus’s latest @CACMmag⁩ https://cacm.acm.org/blog…

DGX agent

What Should We Take From Anthropic’s (possibly) Terrifying New Report on Mythos? – ⁦@garymarcus’s latest @CACMmag⁩ https://cacm.acm.org/blogcacm/what-should-we-take-from-anthropics-possibly-terrifying

safetygary-marcus--x
8 Apr 2026
Safety

If you ban self-driving cars to protect the taxi union, you have blood on your hands

DGX agent

If you ban self-driving cars to protect the taxi union, you have blood on your hands If you want to know why @Waymo is no longer testing in NYC, this statement says it all: “Our top priority for AV te

safetysonya-huang--x
7 Apr 2026
Safety

You should read the red team report: https://red.anthropic.com/2026/mythos-preview/

DGX agent

Anthropic's Frontier Red Team published a technical report (April 2026) detailing how their unreleased model, Claude Mythos Preview, autonomously identifies and exploits critical security vulnerabi...

safetyethan-mollick--x
7 Apr 2026
← Previous
1…261262263
Next →