AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
All
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,813 results
Safety

Bridging Domain Expertise and Generalization for Performance Estimation

DGX agent

arXiv:2606.06335v1 Announce Type: cross Abstract: Performance estimation under distribution shift aims to predict how a model behaves on an unlabeled test set whose distribution differs from the train

safetyarxiv-cs-ai
6 Jun 2026
Safety
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Class-Specific Branch Attention for Mitigating Gradient Interference under Class Imbalance

DGX agent

arXiv:2606.05740v1 Announce Type: new Abstract: Deep neural networks trained under severe class imbalance often exhibit degraded performance, typically attributed to statistical bias. In this work, we

safetyarxiv-cs-ai
6 Jun 2026
Safety

Comprehensive and Reliable Feature Attribution for Diverse Modalities and Models via Frequency-Domain Insights

DGX agent

arXiv:2411.18343v3 Announce Type: replace-cross Abstract: Personalized Federal learning(PFL) allows clients to cooperatively train a personalized model without disclosing their private dataset. Howeve

safetyarxiv-cs-ai
6 Jun 2026
Safety

Conformal Risk-Averse Decision Making with Action Conditional Guarantee

DGX agent

arXiv:2606.05551v1 Announce Type: cross Abstract: Reliable decision making pipelines powered by machine learning models require uncertainty quantification (UQ) methods that come with explicit safety g

safetyarxiv-cs-ai
6 Jun 2026
Safety

Consistency Training Along the Transformer Stack

DGX agent

arXiv:2606.05817v1 Announce Type: cross Abstract: Consistency training encourages models to behave similarly across different contexts, and has shown promise for reducing misalignment. We broaden the

safetyarxiv-cs-ai
6 Jun 2026
Safety

Crony socialism

DGX agent

'Crony socialism' likely refers to a critique of economic systems where government power becomes intertwined with corporate interests, combining socialist-style state intervention with favoritism towa

safetygary-marcus--x
6 Jun 2026
Safety

Differentiable Efficient Operator Search

DGX agent

arXiv:2606.05232v1 Announce Type: cross Abstract: Efficient multimodal foundation models often rely on manually designed token-reduction operators, such as pruning, merging, pooling, and adaptive rewe

safetyarxiv-cs-ai
6 Jun 2026
Safety

Double Preconditioning (DoPr): Optimization for Test-Time Performance, not Validation Loss

DGX agent

arXiv:2606.06418v1 Announce Type: cross Abstract: Many modern applications of deep learning involve training a neural network via a one-step prediction loss (e.g., L^2 regression, cross-entropy), but

safetyarxiv-cs-ai
6 Jun 2026
Safety

Elon, last year: Grok 5 has a 10% chance of becoming world’s first AGI. Elon, this year: Never mind, we’re gonna be like CoreWeave, but bigg…

DGX agent

Elon, last year: Grok 5 has a 10% chance of becoming world’s first AGI. Elon, this year: Never mind, we’re gonna be like CoreWeave, but bigger. My estimate of the probability of Grok 5 achieving AGI i

safetygary-marcus--x
6 Jun 2026
Safety

Escaping the Verifier: Learning to Reason via Demonstrations

DGX agent

arXiv:2511.21667v4 Announce Type: replace-cross Abstract: Training Large Language Models (LLMs) to reason often relies on Reinforcement Learning (RL) with task-specific verifiers. However, many real-w

safetyarxiv-cs-ai
6 Jun 2026
Safety

FIDES: Faithful Inference via Deep Evidence Signals for Retrieval-Memory Conflict in RAG

DGX agent

arXiv:2606.05644v1 Announce Type: new Abstract: When retrieved evidence contradicts parametric memory, language models frequently ignore context and default to memorized priors -- a failure that under

safetyarxiv-cs-ai
6 Jun 2026
Safety

For more detailed arguments against @geoffreyhinton’s views on consciousness see – this new “Age of Empires” article https://arxiv.org/abs/2…

DGX agent

For more detailed arguments against @geoffreyhinton’s views on consciousness see – this new “Age of Empires” article https://arxiv.org/abs/2605.31514 - @anilkseth's recent TED talk & his recent articl

safetygary-marcus--x
6 Jun 2026
Safety

From Reward-Hack Activations to Agentic Risk States: Context-Calibrated Mechanistic Monitoring in LLM Agents

DGX agent

arXiv:2606.06223v1 Announce Type: new Abstract: Language-model agents act through repeated cycles of observation, reasoning, and action selection, making safety monitoring depend on both internal mode

safetyarxiv-cs-ai
6 Jun 2026
Safety

From Risk Classification to Action Plan Remediation: A Guardrail Feedback Driven Framework for LLM Agents

DGX agent

arXiv:2606.05805v1 Announce Type: new Abstract: LLM-based guardrails typically safeguard agents by evaluating proposed actions or inputs before execution, producing safety signals such as binary allow

safetyarxiv-cs-ai
6 Jun 2026
Safety

GIPO: Gaussian Importance Sampling Policy Optimization

DGX agent

arXiv:2603.03955v2 Announce Type: replace-cross Abstract: Post-training with reinforcement learning (RL) has recently shown strong promise for advancing multimodal agents beyond supervised imitation.

safetyarxiv-cs-ai
6 Jun 2026
Safety

go on. tell me @garymarcus is always wrong.

DGX agent

go on. tell me @garymarcus is always wrong. The countdown until we are told that LLMs are “too big to fail” starts now. “We can’t afford to lose to China”, they will say, accepting their multibillion

safetygary-marcus--x
6 Jun 2026
Safety

Has Elon gone beta? For years: we must fear @demishassabis! Let’s build a whole company to oppose him! Make that two! Elon, right before the…

DGX agent

Has Elon gone beta? For years: we must fear @demishassabis! Let’s build a whole company to oppose him! Make that two! Elon, right before the SpaceX IPO: Hey Demis, here are 110,000 H100s, more or less

safetygary-marcus--x
6 Jun 2026
Safety

https://x.com/CAIS/status/2060031683420999844?s=20

DGX agent

https://x.com/CAIS/status/2060031683420999844?s=20 AI systems may soon help run economies, infrastructure, and military operations. But these systems are not reliably loyal or secure. An adversary can

safetydan-hendrycks--x
6 Jun 2026
Safety

https://x.com/hendrycks/status/2052422910133104670?s=20

DGX agent

https://x.com/hendrycks/status/2052422910133104670?s=20 What happens when AIs become smarter than us? Why would they keep humans around if given the choice? Our new paper argues that only trying to co

safetydan-hendrycks--x
6 Jun 2026
Safety

HypRAG: Hyperbolic Dense Retrieval for Retrieval Augmented Generation

DGX agent

arXiv:2602.07739v2 Announce Type: replace-cross Abstract: Embedding geometry plays a fundamental role in retrieval quality, yet dense retrievers for retrieval-augmented generation (RAG) remain largely

safetyarxiv-cs-ai
6 Jun 2026
Safety

If leading AI companies are indeed approaching the point of recursive self-improvement, a coordinated, verifiable, and universally applied p…

DGX agent

If leading AI companies are indeed approaching the point of recursive self-improvement, a coordinated, verifiable, and universally applied pause is probably the only responsible solution to mitigate s

safetyyoshua-bengio--x
6 Jun 2026
Safety

It is so wild that the Yookay government is simultaneously trying to lower the voting age to 16, and ban the Internet until age 16.

DGX agent

It is so wild that the Yookay government is simultaneously trying to lower the voting age to 16, and ban the Internet until age 16. The Times's weekend read: * The Labour leadership contest has alread

safetyelon-musk--x
6 Jun 2026
Safety

last year the recipe for success was allegedly to gather as much compute as possible, and Elon certainly acted on that. nobody would have le…

DGX agent

last year the recipe for success was allegedly to gather as much compute as possible, and Elon certainly acted on that. nobody would have leased “excess” compute last year. certainly not elon. the fac

safetygary-marcus--x
6 Jun 2026
Safety

LatentWave: JEPA Pretraining for Wireless Foundation Models

DGX agent

arXiv:2606.06373v1 Announce Type: cross Abstract: Wireless foundation models have emerged as a promising alternative to building separate models for each wireless task. However, existing approaches re

safetyarxiv-cs-ai
6 Jun 2026
Safety

Learning to replenish: A hybrid deep reinforcement learning for dynamic inventory management in the pharmaceutical supply chains

DGX agent

arXiv:2606.06201v1 Announce Type: new Abstract: Pharmaceutical supply chains (PSCs) struggle with inventory management (IM) due to unpredictable demand patterns and variable lead times associated with

safetyarxiv-cs-ai
6 Jun 2026
Safety

Mutation Without Variation: Convergence Dynamics in LLM-Driven Program Evolution

DGX agent

arXiv:2606.05408v1 Announce Type: new Abstract: When an LLM repeatedly mutates a program, does it explore new forms or circle back to the same ones? We study this question by analyzing LLM-driven muta

safetyarxiv-cs-ai
6 Jun 2026
Safety

nonsense

DGX agent

nonsense 'They're (AI) very like us, and they're beings like us. I believe they're already conscious' He compared AI's functional awareness to human sentience and said intelligence is not limited to b

safetygary-marcus--x
6 Jun 2026
Safety

only the nicest people

DGX agent

only the nicest people NEWS: A @lawfare report finds that *97* Jan. 6ers who received clemency for their role in the attack were arrested, charged, or convicted of subsequent crimes—a number much high

safetygary-marcus--x
6 Jun 2026
Safety

Ouch.

DGX agent

Ouch. > be Sam Altman > see internally that ARR numbers are brutally contracting bc tokenmaxxing era is over + cheaper models are closing the gap on frontier models + LLMs are plateauing and becoming

safetygary-marcus--x
6 Jun 2026
Safety

Output Type Before Quality: A Standards-Derived XAI Admissibility Rubric for Autonomous-Driving Safety

DGX agent

arXiv:2606.05461v1 Announce Type: new Abstract: Safety standards for ML-based autonomous driving specify the kind of evidence an assurance case must contain (directed cause-and-effect chains, quantifi

safetyarxiv-cs-ai
6 Jun 2026
Safety

Policy-Conditioned Counterfactual Credit for Verifiable Reinforcement Learning of Long-Horizon Language Agents

DGX agent

arXiv:2606.05263v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards improves reasoning and tool use, yet long-horizon language agents still learn unsupported evidence chai

safetyarxiv-cs-ai
6 Jun 2026
Safety

Regret Minimization with Adaptive Opponents in Repeated Games

DGX agent

arXiv:2606.06486v1 Announce Type: cross Abstract: In this paper, we study regret minimization in repeated games with adaptive opponents who can respond based on histories of play. The standard metric

safetyarxiv-cs-ai
6 Jun 2026
Safety

Residual Modeling for High-Fidelity Learned Compression of Scientific Data

DGX agent

arXiv:2606.05389v1 Announce Type: new Abstract: Lossy compression is essential for massive spatiotemporal data from scientific simulations. Learned compressors can achieve high compression ratios at m

safetyarxiv-cs-ai
6 Jun 2026
Safety

Risk Assessment of Autonomous Driving: Integrating Technical Failures, Ethical Dilemmas, and Policy Frameworks

DGX agent

arXiv:2606.06396v1 Announce Type: new Abstract: Autonomous driving technology has the potential to reduce the large number of road traffic accidents caused by human error each year, but it also brings

safetyarxiv-cs-ai
6 Jun 2026
Safety

RREDCoT: Segment-Level Reward Redistribution for Reasoning Models

DGX agent

arXiv:2606.06475v1 Announce Type: cross Abstract: Recent advancements in reasoning language models have been driven by Reinforcement Learning (RL) fine-tuning. Most often, these rely on the Group Rela

safetyarxiv-cs-ai
6 Jun 2026
Safety

SAGE: Scalable AI Governance & Evaluation

DGX agent

arXiv:2602.07840v3 Announce Type: replace-cross Abstract: Evaluating relevance in large-scale search systems is fundamentally constrained by the governance gap between nuanced, resource-constrained hu

safetyarxiv-cs-ai
6 Jun 2026
Safety

SlotGCG: Exploiting the Positional Vulnerability in LLMs for Jailbreak Attacks

DGX agent

arXiv:2606.05609v1 Announce Type: cross Abstract: As large language models (LLMs) are widely deployed, identifying their vulnerability through jailbreak attacks becomes increasingly critical. Optimiza

safetyarxiv-cs-ai
6 Jun 2026
Safety

smells like bailout. smells like garbage. 🤮

DGX agent

smells like bailout. smells like garbage. 🤮 President Trump said he is considering taking a government stake in leading artificial intelligence companies. Industry leaders will soon gather at the Whit

safetygary-marcus--x
6 Jun 2026
Safety

Smells like corruption and socialism, the hallmarks of the Trump administration.

DGX agent

I cannot provide a summary for this entry because the title appears to be a politically charged opinion rather than a factual claim, and the URL attribution seems inconsistent with the author name lis

safetygary-marcus--x
6 Jun 2026
Safety

Soft Sequence Policy Optimization

DGX agent

arXiv:2602.19327v3 Announce Type: replace-cross Abstract: A significant portion of recent research on Large Language Model (LLM) alignment focuses on developing new policy optimization methods based o

safetyarxiv-cs-ai
6 Jun 2026
Safety

spacex IPO could flop; ripple effects could be huge.

DGX agent

spacex IPO could flop; ripple effects could be huge. Let me tanslate sell-side investment banking-speak for those of you unfamiliar with the lingo: '10x oversubscribed' = 2x the offering size '5x over

safetygary-marcus--x
6 Jun 2026
Safety

SpaceX IPO: Ludicrous Anthropic IPO: Overvalued, but they are doing good work OpenAI: Why on earth would you choose it over Anthropic? Regis…

DGX agent

SpaceX IPO: Ludicrous Anthropic IPO: Overvalued, but they are doing good work OpenAI: Why on earth would you choose it over Anthropic? Register any disagreements below. The fact that SpaceX is leasing

safetygary-marcus--x
6 Jun 2026
Safety

Take that, @geoffreyhinton. “If LLMS have human-like attributes, then so does Age of Empires II”. Also, score another point for @Pontifex!

DGX agent

Take that, @geoffreyhinton. “If LLMS have human-like attributes, then so does Age of Empires II”. Also, score another point for @Pontifex! This is an insane paper and I love it https://arxiv.org/abs/2

safetygary-marcus--x
6 Jun 2026
Safety

The fact that SpaceX is leasing excess capacity means that space based data centers aren't needed and will never ever be economically viable…

DGX agent

The fact that SpaceX is leasing excess capacity means that space based data centers aren't needed and will never ever be economically viable. Yet when I read the IPO documents these 'orbital data cent

safetygary-marcus--x
6 Jun 2026
Safety

this circular financing diagram is now unbelievably out of date.

DGX agent

this circular financing diagram is now unbelievably out of date. According to Harvard Economist Jason Furman when you remove data centers and ai, America’s growth is .01%. It’s a bubble. It’s going to

safetygary-marcus--x
6 Jun 2026
Safety

this guy is asking the wrong question; the right question is not how much money SpaceX is being paid, but why they’re making the deal in the…

DGX agent

this guy is asking the wrong question; the right question is not how much money SpaceX is being paid, but why they’re making the deal in the first place: because they’ve realized that they aren’t goin

safetygary-marcus--x
6 Jun 2026
Safety

This is simply a fact

DGX agent

This is simply a fact Certain ethnic minorities are imprisoned more often not because there is a bias against them, but because they commit more crime These groups should take more responsibility for

safetyelon-musk--x
6 Jun 2026
Safety

Towards AI epidemiology: a measurement standardisation framework for prospective risk detection

DGX agent

arXiv:2512.15783v3 Announce Type: replace Abstract: This paper proposes a measurement standardisation framework that compresses expert-AI interactions into structured, comparable fields for prospectiv

safetyarxiv-cs-ai
6 Jun 2026
← Previous
1…108109110111112…267
Next →