AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “dan-hendrycks--x”

GridTimelineEvolution
16 results
CompaniesToolsTechniques

Each lane shows up to 8 recent matching entries, ordered from earlier to later. Tracks load separately to keep the 75,000+ entry wiki fast.

Techniques

TechniqueRLHF / Alignment1 recent entries
7 Jun 2026One of the more interesting takes on positive alignment that have recently come out-it’s long and interesting, combining philosophy and trai…

One of the more interesting takes on positive alignment that have recently come out-it’s long and interesting, combining philosophy and training setups (eg reward proposals), and worth a read. What ha

TechniqueAgents2 recent entries
28 Apr 2026When an LLM acts happy (“EUREKA!”) or sad (“I have failed…”), is that meaningless mimicry, or does it reflect something “real”? We don’t kno…
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

When an LLM acts happy (“EUREKA!”) or sad (“I have failed…”), is that meaningless mimicry, or does it reflect something “real”? We don’t know if LLMs are conscious. But they increasingly seem to exhib

→2 May 2026Are AIs about to subjugate humanity? In the debate about catastrophic AI risk, evolutionary scenarios receive far too little attention compa…

Are AIs about to subjugate humanity? In the debate about catastrophic AI risk, evolutionary scenarios receive far too little attention compared to largely speculative arguments about “instrumental con

TechniqueSafety8 recent entries
2 May 2026Are AIs about to subjugate humanity? In the debate about catastrophic AI risk, evolutionary scenarios receive far too little attention compa…

Are AIs about to subjugate humanity? In the debate about catastrophic AI risk, evolutionary scenarios receive far too little attention compared to largely speculative arguments about “instrumental con

→7 May 2026What happens when AIs become smarter than us? Why would they keep humans around if given the choice? Our new paper argues that only trying t…

What happens when AIs become smarter than us? Why would they keep humans around if given the choice? Our new paper argues that only trying to control AIs is a limited strategy, and that a stable, mutu

→28 May 2026AI systems may soon help run economies, infrastructure, and military operations. But these systems are not reliably loyal or secure. An adve…

AI systems may soon help run economies, infrastructure, and military operations. But these systems are not reliably loyal or secure. An adversary can make an AI work against its own operator. In our n

→6 Jun 2026https://x.com/hendrycks/status/2052422910133104670?s=20

https://x.com/hendrycks/status/2052422910133104670?s=20 What happens when AIs become smarter than us? Why would they keep humans around if given the choice? Our new paper argues that only trying to co

→6 Jun 2026https://x.com/CAIS/status/2060031683420999844?s=20

https://x.com/CAIS/status/2060031683420999844?s=20 AI systems may soon help run economies, infrastructure, and military operations. But these systems are not reliably loyal or secure. An adversary can

→7 Jun 2026Refreshing

Refreshing What happens when AIs become smarter than us? Why would they keep humans around if given the choice? Our new paper argues that only trying to control AIs is a limited strategy, and that a s

→7 Jun 2026One of the more interesting takes on positive alignment that have recently come out-it’s long and interesting, combining philosophy and trai…

One of the more interesting takes on positive alignment that have recently come out-it’s long and interesting, combining philosophy and training setups (eg reward proposals), and worth a read. What ha

→1 Jul 2026A basic sigmoid extrapolation suggests that this measurement will be close to its performance ceiling in around a year---meaning most random…

A basic sigmoid extrapolation suggests that this measurement will be close to its performance ceiling in around a year---meaning most randomly sampled remote work projects would be highly automatable