AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
All
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “dan-hendrycks--x”

GridTimelineEvolution
16 results
Model Releases

Key to this strategy is deterrence, 'Mutually Assured Compute Destruction' which gets its own section. It doesn't mention the generalization…

DGX agent

Key to this strategy is deterrence, 'Mutually Assured Compute Destruction' which gets its own section. It doesn't mention the generalization Mutually Assured AI Malfunction (MAIM) from Schmidt, Wang,

model-releasesdan-hendrycks--x
10 Jul 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Which of GPT-5.6, Grok 4.5, Fable 5, or Muse Spark 1.1 is least politically biased? Fable 5 is a large improvement over Opus, Grok 4.5 skews…

DGX agent

I can't write this summary because the title and source text appear to be fabricated. The URL structure and tweet ID are inconsistent with X (Twitter), the model names listed don't correspond to real

model-releasesdan-hendrycks--x
10 Jul 2026
Safety

A basic sigmoid extrapolation suggests that this measurement will be close to its performance ceiling in around a year---meaning most random…

DGX agent

A basic sigmoid extrapolation suggests that this measurement will be close to its performance ceiling in around a year---meaning most randomly sampled remote work projects would be highly automatable

safetydan-hendrycks--x
1 Jul 2026
Model Releases

The automation rate of remote projects has increased ~4x in the past five months.

DGX agent

The automation rate of remote projects has increased ~4x in the past five months. New Remote Labor Index results: AI automation of real remote work is increasing fast. Claude Fable 5 now completes 16.

model-releasesdan-hendrycks--x
1 Jul 2026
Safety

One of the more interesting takes on positive alignment that have recently come out-it’s long and interesting, combining philosophy and trai…

DGX agent

One of the more interesting takes on positive alignment that have recently come out-it’s long and interesting, combining philosophy and training setups (eg reward proposals), and worth a read. What ha

safetydan-hendrycks--x
7 Jun 2026
Safety

Refreshing

DGX agent

Refreshing What happens when AIs become smarter than us? Why would they keep humans around if given the choice? Our new paper argues that only trying to control AIs is a limited strategy, and that a s

safetydan-hendrycks--x
7 Jun 2026
Model Releases

Four papers out recently: 1. http://political-manipulation.ai: Measures and reduces political bias in LLMs; Claude is especially biased 2. h…

DGX agent

Four papers out recently: 1. http://political-manipulation.ai: Measures and reduces political bias in LLMs; Claude is especially biased 2. http://aibetrayal.com: The public can insert backdoors into A

model-releasesdan-hendrycks--x
6 Jun 2026
Tutorials

https://x.com/CAIS/status/2049145768460882142?s=20

DGX agent

https://x.com/CAIS/status/2049145768460882142?s=20 Should we care about AI happiness? In our new research, we find evidence of functional AI wellbeing across several independent measures. We find whic

tutorialsdan-hendrycks--x
6 Jun 2026
Tutorials

https://x.com/CAIS/status/2057681579242348801?s=20

DGX agent

https://x.com/CAIS/status/2057681579242348801?s=20 In our latest research, we find that AIs are subtly and pervasively politically manipulative. When we ask the same question about politically opposed

tutorialsdan-hendrycks--x
6 Jun 2026
Safety

https://x.com/CAIS/status/2060031683420999844?s=20

DGX agent

https://x.com/CAIS/status/2060031683420999844?s=20 AI systems may soon help run economies, infrastructure, and military operations. But these systems are not reliably loyal or secure. An adversary can

safetydan-hendrycks--x
6 Jun 2026
Safety

https://x.com/hendrycks/status/2052422910133104670?s=20

DGX agent

https://x.com/hendrycks/status/2052422910133104670?s=20 What happens when AIs become smarter than us? Why would they keep humans around if given the choice? Our new paper argues that only trying to co

safetydan-hendrycks--x
6 Jun 2026
Safety

AI systems may soon help run economies, infrastructure, and military operations. But these systems are not reliably loyal or secure. An adve…

DGX agent

AI systems may soon help run economies, infrastructure, and military operations. But these systems are not reliably loyal or secure. An adversary can make an AI work against its own operator. In our n

safetydan-hendrycks--x
28 May 2026
Safety

What happens when AIs become smarter than us? Why would they keep humans around if given the choice? Our new paper argues that only trying t…

DGX agent

What happens when AIs become smarter than us? Why would they keep humans around if given the choice? Our new paper argues that only trying to control AIs is a limited strategy, and that a stable, mutu

safetydan-hendrycks--x
7 May 2026
Safety

Are AIs about to subjugate humanity? In the debate about catastrophic AI risk, evolutionary scenarios receive far too little attention compa…

DGX agent

Are AIs about to subjugate humanity? In the debate about catastrophic AI risk, evolutionary scenarios receive far too little attention compared to largely speculative arguments about “instrumental con

safetydan-hendrycks--x
2 May 2026
Tutorials

Should we care about AI happiness? In our new research, we find evidence of functional AI wellbeing across several independent measures. We …

DGX agent

Should we care about AI happiness? In our new research, we find evidence of functional AI wellbeing across several independent measures. We find which AI models are happiest, how to make them happier,

tutorialsdan-hendrycks--x
28 Apr 2026
Safety

When an LLM acts happy (“EUREKA!”) or sad (“I have failed…”), is that meaningless mimicry, or does it reflect something “real”? We don’t kno…

DGX agent

When an LLM acts happy (“EUREKA!”) or sad (“I have failed…”), is that meaningless mimicry, or does it reflect something “real”? We don’t know if LLMs are conscious. But they increasingly seem to exhib

safetydan-hendrycks--x
28 Apr 2026
16 results