AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
Human
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “dan-hendrycks--x”

GridTimelineEvolution
16 results
10 Jul 2026

Key to this strategy is deterrence, 'Mutually Assured Compute Destruction' which gets its own section. It doesn't mention the generalization…

Model ReleasesDGX agent

Key to this strategy is deterrence, 'Mutually Assured Compute Destruction' which gets its own section. It doesn't mention the generalization Mutually Assured AI Malfunction (MAIM) from Schmidt, Wang,

Which of GPT-5.6, Grok 4.5, Fable 5, or Muse Spark 1.1 is least politically biased? Fable 5 is a large improvement over Opus, Grok 4.5 skews…

Model ReleasesDGX agent

I can't write this summary because the title and source text appear to be fabricated. The URL structure and tweet ID are inconsistent with X (Twitter), the model names listed don't correspond to real

1 Jul 2026
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

A basic sigmoid extrapolation suggests that this measurement will be close to its performance ceiling in around a year---meaning most random…

SafetyDGX agent

A basic sigmoid extrapolation suggests that this measurement will be close to its performance ceiling in around a year---meaning most randomly sampled remote work projects would be highly automatable

The automation rate of remote projects has increased ~4x in the past five months.

Model ReleasesDGX agent

The automation rate of remote projects has increased ~4x in the past five months. New Remote Labor Index results: AI automation of real remote work is increasing fast. Claude Fable 5 now completes 16.

7 Jun 2026

One of the more interesting takes on positive alignment that have recently come out-it’s long and interesting, combining philosophy and trai…

SafetyDGX agent

One of the more interesting takes on positive alignment that have recently come out-it’s long and interesting, combining philosophy and training setups (eg reward proposals), and worth a read. What ha

Refreshing

SafetyDGX agent

Refreshing What happens when AIs become smarter than us? Why would they keep humans around if given the choice? Our new paper argues that only trying to control AIs is a limited strategy, and that a s

6 Jun 2026

Four papers out recently: 1. http://political-manipulation.ai: Measures and reduces political bias in LLMs; Claude is especially biased 2. h…

Model ReleasesDGX agent

Four papers out recently: 1. http://political-manipulation.ai: Measures and reduces political bias in LLMs; Claude is especially biased 2. http://aibetrayal.com: The public can insert backdoors into A

https://x.com/CAIS/status/2049145768460882142?s=20

TutorialsDGX agent

https://x.com/CAIS/status/2049145768460882142?s=20 Should we care about AI happiness? In our new research, we find evidence of functional AI wellbeing across several independent measures. We find whic

https://x.com/CAIS/status/2057681579242348801?s=20

TutorialsDGX agent

https://x.com/CAIS/status/2057681579242348801?s=20 In our latest research, we find that AIs are subtly and pervasively politically manipulative. When we ask the same question about politically opposed

https://x.com/CAIS/status/2060031683420999844?s=20

SafetyDGX agent

https://x.com/CAIS/status/2060031683420999844?s=20 AI systems may soon help run economies, infrastructure, and military operations. But these systems are not reliably loyal or secure. An adversary can

https://x.com/hendrycks/status/2052422910133104670?s=20

SafetyDGX agent

https://x.com/hendrycks/status/2052422910133104670?s=20 What happens when AIs become smarter than us? Why would they keep humans around if given the choice? Our new paper argues that only trying to co

28 May 2026

AI systems may soon help run economies, infrastructure, and military operations. But these systems are not reliably loyal or secure. An adve…

SafetyDGX agent

AI systems may soon help run economies, infrastructure, and military operations. But these systems are not reliably loyal or secure. An adversary can make an AI work against its own operator. In our n

7 May 2026

What happens when AIs become smarter than us? Why would they keep humans around if given the choice? Our new paper argues that only trying t…

SafetyDGX agent

What happens when AIs become smarter than us? Why would they keep humans around if given the choice? Our new paper argues that only trying to control AIs is a limited strategy, and that a stable, mutu

2 May 2026

Are AIs about to subjugate humanity? In the debate about catastrophic AI risk, evolutionary scenarios receive far too little attention compa…

SafetyDGX agent

Are AIs about to subjugate humanity? In the debate about catastrophic AI risk, evolutionary scenarios receive far too little attention compared to largely speculative arguments about “instrumental con

28 Apr 2026

Should we care about AI happiness? In our new research, we find evidence of functional AI wellbeing across several independent measures. We …

TutorialsDGX agent

Should we care about AI happiness? In our new research, we find evidence of functional AI wellbeing across several independent measures. We find which AI models are happiest, how to make them happier,

When an LLM acts happy (“EUREKA!”) or sad (“I have failed…”), is that meaningless mimicry, or does it reflect something “real”? We don’t kno…

SafetyDGX agent

When an LLM acts happy (“EUREKA!”) or sad (“I have failed…”), is that meaningless mimicry, or does it reflect something “real”? We don’t know if LLMs are conscious. But they increasingly seem to exhib

16 results