AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,490 results
Safety

World Action Verifier: Self-Improving World Models via Forward-Inverse Asymmetry

DGX agent

arXiv:2604.01985v2 Announce Type: replace-cross Abstract: General-purpose world models promise scalable policy evaluation, optimization, and planning, yet achieving the required level of robustness re

safetyarxiv-cs-ai
1 Jun 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

World2Act: Latent Action Post-Training from World Model Dynamics

DGX agent

arXiv:2603.10422v2 Announce Type: replace Abstract: World Models (WMs) offer a promising mechanism for post-training Vision-Language-Action (VLA) policies by providing dynamics priors that improve gen

safetyarxiv-cs-cv
1 Jun 2026
Safety

Your Teacher Can't Help You Here: Combating Supervision Fidelity Decay in On-Policy Distillation

DGX agent

arXiv:2605.30833v1 Announce Type: cross Abstract: On-policy distillation transfers reasoning capabilities by training a student model on its own generated trajectories using token-level feedback from

safetyarxiv-cs-ai
1 Jun 2026
Safety

ZAPS-DA: Zero-Phase Action Policy Smoothing with Decoupled Actor for Continuous Control in Reinforcement Learning

DGX agent

arXiv:2605.30612v1 Announce Type: cross Abstract: Continuous control policies trained with off-policy reinforcement learning frequently exhibit high-frequency action jitter, rendering direct deploymen

safetyarxiv-cs-lg
1 Jun 2026
Safety

Zero Collapse: A Failure Mode of Policy Gradient Methods in Discontinuous Reward Environments

DGX agent

arXiv:2605.30896v1 Announce Type: new Abstract: Bidding in repeated auctions is a central challenge for reinforcement learning (RL), combining continuous control with the strategic complexities of dig

safetyarxiv-cs-lg
1 Jun 2026
Safety

Anyone remember how I said in January 2025 that AI was going to be stumble and be dubbed “too big to fail”, along with cries for bailouts? I…

DGX agent

Anyone remember how I said in January 2025 that AI was going to be stumble and be dubbed “too big to fail”, along with cries for bailouts? If that call was correct – which increasingly seems likely, a

safetygary-marcus--x
31 May 2026
Safety

further discussion here: https://open.substack.com/pub/garymarcus/p/the-pope-appears-to-understand-ai?r=8tdk6&utm_campaign=post-expanded-sha…

DGX agent

Gary Marcus discusses the Pope's understanding and perspective on artificial intelligence, examining statements or positions the religious leader has taken regarding AI technology and its implications

safetygary-marcus--x
31 May 2026
Safety

I am absolutely with @GaryMarcus and the Pope on this. (Not something you expect to say everyday). We are not creating beings. Systems don’t…

DGX agent

I am absolutely with @GaryMarcus and the Pope on this. (Not something you expect to say everyday). We are not creating beings. Systems don’t exp. grief nor hope, hold a value construct, or the ability

safetygary-marcus--x
31 May 2026
Safety

Listen, I’m a big “the index is the index” guy, but they are openly looting the coffers. This is 100% fraud.

DGX agent

Listen, I’m a big “the index is the index” guy, but they are openly looting the coffers. This is 100% fraud. Rule changes for the SpaceX SPCX IPO: Index providers waived the profitability requirement

safetygary-marcus--x
31 May 2026
Safety

@ParValue26 @Hedgeye Let’s be clear what this actually is: the administration and the world’s richest man working together to screw over ord…

DGX agent

@ParValue26 @Hedgeye Let’s be clear what this actually is: the administration and the world’s richest man working together to screw over ordinary investors to goose returns on the IPO for the wealthie

safetygary-marcus--x
31 May 2026
Safety

serious accusation. does this fit with people’s experience?

DGX agent

serious accusation. does this fit with people’s experience? Is Anthropic altering model performance to force costly upgrades? Chapter Co-Founder and CEO @CobiBGantz outlines a shift his team recently

safetygary-marcus--x
31 May 2026
Safety

SpaceX being rammed into indices with no profit requirements, seasoning, and generally looser constraints is economic terrorism. Index track…

DGX agent

SpaceX being rammed into indices with no profit requirements, seasoning, and generally looser constraints is economic terrorism. Index trackers will eat the loss when reality catches up and retail inv

safetygary-marcus--x
31 May 2026
Safety

The backlash against AI - generated content is so visceral - expect the 'human-authored' certification gain momentum, esp for fiction. https…

DGX agent

Gary Marcus argues that consumer backlash against AI-generated content will be increasingly visceral, particularly in creative fields like fiction, leading to growing demand for 'human-authored' certi

safetygary-marcus--x
31 May 2026
Safety

Three companies are about to IPO at a higher combined value than all 2,600 dot-com IPOs from 1995 to 2000 combined. – SpaceX, OpenAI, Anthro…

DGX agent

Three companies are about to IPO at a higher combined value than all 2,600 dot-com IPOs from 1995 to 2000 combined. – SpaceX, OpenAI, Anthropic (2026): ~3.75 trillion – Every dot-com IPO from 1995–200

safetygary-marcus--x
31 May 2026
Safety

Weird how the Pope seems to understand AI better than @geoffreyhinton, but I am 100% with the Pope on this. We are NOT creating beings. The …

DGX agent

Weird how the Pope seems to understand AI better than @geoffreyhinton, but I am 100% with the Pope on this. We are NOT creating beings. The Pope is right. We are creating interactive fiction that is t

safetygary-marcus--x
31 May 2026
Safety

weird the way this tweet was getting a ton of traffic and then just stopped. 🤷‍♂️

DGX agent

weird the way this tweet was getting a ton of traffic and then just stopped. 🤷‍♂️ I honestly think Elon’s best days are behind him: BYD is crushing Tesla in EVs. Waymo is crushing Tesla in AVs. Anthro

safetygary-marcus--x
31 May 2026
Safety

calling someone a retard when you don’t how apostrophes work 🙄

DGX agent

This post likely critiques the irony of someone insulting another person's intelligence while making a basic grammatical error (omitting an apostrophe in 'don't'). Gary Marcus, a cognitive scientist a

safetygary-marcus--x
30 May 2026
Safety

genius reply, full of reasoned intellectual argument, from the kind of retail investor who is probably going to get burned on the SpaceX IPO…

DGX agent

genius reply, full of reasoned intellectual argument, from the kind of retail investor who is probably going to get burned on the SpaceX IPO. Going to MIT and NYU may have meant something in the past,

safetygary-marcus--x
30 May 2026
Safety

i am only blocking tesla supporters that come at me with insult rather than argument. but that’s a lot of them. and a not great sign for the…

DGX agent

Gary Marcus discusses his moderation approach on social media, noting that he blocks Tesla supporters primarily when they resort to insults rather than substantive arguments, and observes that this oc

safetygary-marcus--x
30 May 2026
Safety

I honestly think Elon’s best days are behind him: BYD is crushing Tesla in EVs. Waymo is crushing Tesla in AVs. Anthropic, Openai, and Googl…

DGX agent

I honestly think Elon’s best days are behind him: BYD is crushing Tesla in EVs. Waymo is crushing Tesla in AVs. Anthropic, Openai, and Google are crushing Xai on AI. The SpaceX S-1 is so ridiculous th

safetygary-marcus--x
30 May 2026
Safety

“I'm more concerned about the lack of intellectual diversity within the frontier AI commentariat/research world. This improved a lot over th…

DGX agent

“I'm more concerned about the lack of intellectual diversity within the frontier AI commentariat/research world. This improved a lot over the last two years, but we're still far from a healthy ecosyst

safetygary-marcus--x
30 May 2026
Safety

is having a four month lead a sustainable multitrillion dollar business model?

DGX agent

is having a four month lead a sustainable multitrillion dollar business model? We took another look at the capability gap between open-weight and proprietary models. Since the start of the year, open-

safetygary-marcus--x
30 May 2026
Safety

it is the single most clever move Elon ever pulled off

DGX agent

it is the single most clever move Elon ever pulled off Rule changes for the SpaceX SPCX IPO: Index providers waived the profitability requirement and cut the seasoning window from 90 days to 5. This f

safetygary-marcus--x
30 May 2026
Safety

More of my thoughts on the topic in @Corriere: https://www.corriere.it/esteri/26_maggio_29/yoshua-bengio-ai-pope-6aa95de2-33f7-4f73-8905-34b…

DGX agent

More of my thoughts on the topic in @Corriere: https://www.corriere.it/esteri/26_maggio_29/yoshua-bengio-ai-pope-6aa95de2-33f7-4f73-8905-34b1abedbxlk.shtml “Like nuclear energy, AI must be at the serv

safetyyoshua-bengio--x
30 May 2026
Safety

My feed is suddenly filled w Elon supporters who can’t understand the difference between an insult and a genuine counterargument. As always,…

DGX agent

Gary Marcus expresses frustration about an influx of Elon Musk supporters on his social media feed who conflate personal insults with substantive counterarguments. The post appears to critique a lack

safetygary-marcus--x
30 May 2026
Safety

so much for multitrillion dollar candy companies

DGX agent

so much for multitrillion dollar candy companies Note to all staff: Turns out that super tasty AI candy we've been putting out in large bowls isn't free and it isn't necessarily leading to better prod

safetygary-marcus--x
30 May 2026
Safety

this won’t end well. it will end with a bailout.

DGX agent

this won’t end well. it will end with a bailout. Cash flow no longer covers the AI capex bill, so hyperscalers are funding it with record debt: Hyperscaler bond issuance has soared to $150 billion YTD

safetygary-marcus--x
30 May 2026
Safety

what a time to be alive

DGX agent

what a time to be alive wow, Opus 4.8 is very... argument-happy? it picked a fight with me about my usage of the word 'ontology', and when we eventually got back on the same page philosophically, told

safetygary-marcus--x
30 May 2026
Safety

Where is the power/value of money physically located? It's clearly not in the actual physical bills. Why are we more scared of an elderly ma…

DGX agent

Where is the power/value of money physically located? It's clearly not in the actual physical bills. Why are we more scared of an elderly mafia boss than their much more physically dangerous underling

safetyconnor-leahy--x
30 May 2026
Safety

Why LLMs rarely payoff—and what I have been saying literally for 7 years—confirmed yet again: LLMs can’t handle the truth. (Nor apparently c…

DGX agent

Why LLMs rarely payoff—and what I have been saying literally for 7 years—confirmed yet again: LLMs can’t handle the truth. (Nor apparently can my critics, who keep saying I am “always wrong”, when I h

safetygary-marcus--x
30 May 2026
Safety

A 25B fund just refused SpaceX at any price. It says the company can't be worth more than 1T, half the $1.8T IPO target, and Musk's 85% co…

DGX agent

A 25B fund just refused SpaceX at any price. It says the company can't be worth more than 1T, half the $1.8T IPO target, and Musk's 85% control makes it impossible to fix from inside. https://thenextw

safetygary-marcus--x
29 May 2026
Safety

A Fully Convolutional Approach to Denoising Structural Dynamics Data from X-Ray Photon Correlation Spectroscopy

DGX agent

arXiv:2605.29975v1 Announce Type: new Abstract: We present a fully convolutional denoising autoencoder (FC-DAE) for denoising two-time intensity-intensity correlation functions (C_2) in X-ray photon c

safetyarxiv-cs-lg
29 May 2026
Safety

A Geometric View of SRC: Learning Representations for Stable Residual Inference

DGX agent

arXiv:2605.29673v1 Announce Type: cross Abstract: Reconstruction-based inference assigns a class by comparing class-wise reconstruction residuals; Sparse Representation Classification (SRC) is a canon

safetyarxiv-cs-cv
29 May 2026
Safety

A Heterogeneous Architecture for Robot RL Beyond GPU-Dominant Paradigms

DGX agent

arXiv:2605.30313v1 Announce Type: new Abstract: Simulation-based RL for contemporary robot control is increasingly organized around GPU-resident simulation: physics, rollout collection, and learning a

safetyarxiv-cs-ro
29 May 2026
Safety

A Modular Architecture for Typologically Controlled Lexicon Generation

DGX agent

arXiv:2605.28824v1 Announce Type: new Abstract: Constructing artificial lexicons that are pronounceable, typologically plausible, and semantically structured remains an open challenge in computational

safetyarxiv-cs-cl
29 May 2026
Safety

A Predictive Law for On-Policy Self-Distillation From World Feedback

DGX agent

arXiv:2605.30070v1 Announce Type: cross Abstract: Moving beyond simple scalar rewards toward richer world feedback is a natural path to more scalable RL post-training. On-policy self-distillation (OPS

safetyarxiv-cs-ai
29 May 2026
Safety

ActTraitBench: Quantifying the Knowledge-Decision Gap in Large Language Models via Human-Grounded Behavioral Validation

DGX agent

arXiv:2605.29791v1 Announce Type: new Abstract: While Large Language Models (LLMs) can convincingly simulate personas in explicit self-reports, they often deviate in implicit behavioral decisions, rev

safetyarxiv-cs-cl
29 May 2026
Safety

Adaptive Interviewing for Persona Simulation in LLMs: Evidence-Grounded Reasoning Improves Decision Alignment

DGX agent

arXiv:2605.29458v1 Announce Type: cross Abstract: Accurately simulating the decisions of a specific individual remains challenging for large language models (LLMs), partly because persona information

safetyarxiv-cs-ai
29 May 2026
Safety

AG-REPA: Causal Layer Selection for Representation Alignment in Audio Flow Matching

DGX agent

arXiv:2603.01006v2 Announce Type: replace-cross Abstract: REPresentation Alignment (REPA) improves the training of generative flow models by aligning intermediate hidden states with pretrained teacher

safetyarxiv-cs-ai
29 May 2026
Safety

AIRGuard: Guarding Agent Actions with Runtime Authority Control

DGX agent

arXiv:2605.28914v1 Announce Type: cross Abstract: Tool-using language agents turn model decisions into external side effects: they read files, run scripts, call APIs, send messages, and invoke Model C

safetyarxiv-cs-ai
29 May 2026
Safety

AliMark: Enhancing Robustness of Sentence-Level Watermarking Against Text Paraphrasing

DGX agent

arXiv:2605.29434v1 Announce Type: cross Abstract: Existing sentence-level watermarking methods enhance robustness to paraphrasing by anchoring watermarks in sentence semantics. However, their prefix-b

safetyarxiv-cs-ai
29 May 2026
Safety

Auditing Training Data in Generative Music Models via Black-Box Membership Inference

DGX agent

arXiv:2605.29202v1 Announce Type: new Abstract: Recent advances in text-to-music generation enable high-fidelity synthesis of structured musical audio, raising growing concerns about data provenance,

safetyarxiv-cs-lg
29 May 2026
Safety

BEAMS: Benchmarking and Evaluating AI for Modeling and Simulation

DGX agent

arXiv:2605.28994v1 Announce Type: new Abstract: AI tools to support real world decision making must be able to build simulation models that inform their recommendations and render them interpretable.

safetyarxiv-cs-ai
29 May 2026
Safety

Behavior-Induced Mirror-Prox Temporal-Difference Learning for Faster Off-Policy Prediction

DGX agent

arXiv:2605.28849v1 Announce Type: new Abstract: Gradient temporal-difference methods provide stable off-policy prediction with linear function approximation, but their practical performance is strongl

safetyarxiv-cs-ai
29 May 2026
Safety

Beyond Bilingual Transfer: Multilingual Code-Switching in Instruction Tuning

DGX agent

arXiv:2605.29414v1 Announce Type: cross Abstract: Recent studies have shown that code-switching data (CSD), in which multiple languages are mixed within the same context, can improve cross-lingual tra

safetyarxiv-cs-ai
29 May 2026
Safety

Beyond Trajectory Rewards: Step-level Credit Assignment for Agentic Search via Graph Modeling

DGX agent

arXiv:2605.29697v1 Announce Type: new Abstract: In Agentic Search, trajectory-level outcome rewards fail to quantify the behavioral contributions of individual steps, while existing step-level reward

safetyarxiv-cs-ai
29 May 2026
Safety

bold set of counterpredictions, from @scaling01:

DGX agent

bold set of counterpredictions, from @scaling01: Cold take on what comes next: - OpenAI will flourish - Anthropic will continue to be profitable - Google will not catch up to Anthropic or OpenAI - no

safetygary-marcus--x
29 May 2026
Safety

BORA: Bridging Offline Reinforcement Learning and Online Residual Adaptation for Real-World Dexterous VLA Models

DGX agent

arXiv:2605.30226v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have emerged as a promising paradigm for grounding visual-language understanding into real-world robotic manipulat

safetyarxiv-cs-ai
29 May 2026
← Previous
1…170171172173174…302
Next →