AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,812 results
25 May 2026

Reflex: Reinforcement Learning with Reflection Symmetry Exploitation in State-Based Continuous Control

SafetyDGX agent

arXiv:2605.23415v1 Announce Type: cross Abstract: Reinforcement learning has long struggled with poor sample efficiency. One promising approach to mitigate this problem is leveraging group-invariant M

Reinforcement Learning with Verifiable yet Noisy Rewards under Imperfect Verifiers

SafetyDGX agent

arXiv:2510.00915v4 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) replaces costly human labeling with automated verifiers. To reduce verifier hacking, man

Relevant Walk Search for Explaining Graph Neural Networks

SafetyDGX agent

arXiv:2605.23673v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) have become important machine learning tools for graph analysis, and its explainability is crucial for safety, fairness, an


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Remote Teleoperation of Endovascular Intervention Robots: A Systematic Review

SafetyDGX agent

arXiv:2605.22889v1 Announce Type: new Abstract: Remote robotic-assisted endovascular intervention offers a promising approach to reduce clinician radiation exposure and physical strain, while extendin

Robotic Strawberry Harvesting with Robust Vision and Deep Reinforcement Learning based Sim-to-Real Control

SafetyDGX agent

arXiv:2605.23863v1 Announce Type: new Abstract: This study presents a closed-loop robotic strawberry harvesting system that combines a robust vision module, simulation-trained deep reinforcement learn

Safe Reinforcement Learning with Preference-based Constraint Inference

SafetyDGX agent

arXiv:2603.23565v2 Announce Type: replace-cross Abstract: Safe reinforcement learning (RL) is a standard paradigm for safety-critical decision making. However, real-world safety constraints can be com

SafeHarbor: Hierarchical Memory-Augmented Guardrail for LLM Agent Safety

SafetyDGX agent

arXiv:2605.05704v2 Announce Type: replace-cross Abstract: Recent advances in foundation models have transformed LLMs from passive conversational systems into autonomous agents capable of reasoning and

Sample-wise Targeted Adversarial Attacks on Test-time Adaptation

SafetyDGX agent

arXiv:2605.23411v1 Announce Type: cross Abstract: Test-time adaptation (TTA) effectively counters distribution shifts but exposes models to adversarial manipulation via the unlabeled test stream. Exis

Score-Based One-step MeanFlow Policy Optimization

SafetyDGX agent

arXiv:2605.23365v1 Announce Type: cross Abstract: Diffusion and flow matching have emerged as expressive policy classes in reinforcement learning, but their reliance on multi-step denoising imposes su

SCRIPT: Scalable Diffusion Policy with Multi-stage Training for Language-driven Physics-Based Humanoid Control

SafetyDGX agent

arXiv:2605.22894v1 Announce Type: cross Abstract: Controlling physics-based humanoids from natural-language instructions is a critical step toward general-purpose embodied agents. However, existing me

SeedER: Seed-and-Expand Retrieval from Knowledge Graphs

SafetyDGX agent

arXiv:2605.23753v1 Announce Type: new Abstract: Knowledge graphs (KGs) offer a rich representation for relational knowledge, but their irregular structure makes retrieval challenging: ego-graph expans

smdh

SafetyDGX agent

smdh Chamath Lays Out the Case for SpaceX at $2 Trillion – Starlink: the most important internet infra project since the internet itself – Rockets: underlying platform that allows everything else to h

Spatio-Temporal Similarity Volume Aggregation for Open-Vocabulary Action Recognition

SafetyDGX agent

arXiv:2605.23288v1 Announce Type: new Abstract: Recent Open-Vocabulary Action Recognition (OVAR) methods typically aggregate visual features into a global representation before computing text alignmen

SpinFlow: A Physics-Informed Spin Field Framework for Traffic Phase Inference and Transition Detection

SafetyDGX agent

arXiv:2605.23306v1 Announce Type: cross Abstract: Active traffic management (ATM) is frequently hindered by traditional macroscopic models and rigid empirical thresholds that fail to capture metastabl

sure but that shouldn’t affect the valuation of spacex as it pivots from rockets to AI at all, right? right?

SafetyDGX agent

Gary Marcus expresses skepticism about SpaceX's pivot from rockets to AI, questioning whether this strategic shift should impact the company's valuation. The post uses rhetorical questioning to sugges

TactileReflex: Noise-Statistics-Driven Vision-Tactile Reflex Control for Force-Sensitive Manipulation

SafetyDGX agent

arXiv:2605.23568v1 Announce Type: new Abstract: Manipulating fragile deformable containers, such as disposable plastic cups filled with liquid, demands real-time grip-force adaptation within an extrem

Task-Awareness Improves LLM Generations and Uncertainty

SafetyDGX agent

arXiv:2601.21500v2 Announce Type: replace Abstract: In many applications of LLMs, natural language responses often have an underlying structure such as representing discrete labels, numerical values,

Test-Time Training Undermines Safety Guardrails

SafetyDGX agent

arXiv:2605.22984v1 Announce Type: cross Abstract: Test-Time Training (TTT) is an emerging paradigm that enables models to adapt their parameters during inference, improving performance on tasks such a

The CEO of BlackRock, Larry Fink, admits that the trillions of dollars being used to build data centers and power grids will come from ordin…

SafetyDGX agent

The CEO of BlackRock, Larry Fink, admits that the trillions of dollars being used to build data centers and power grids will come from ordinary people’s savings accounts and pension funds, and says it

The Frenchman (@ylecun) seems incapable of giving credit where credit is due. @GaryMarcus is correct about LeCun's lack of integrity. Here h…

SafetyDGX agent

The Frenchman (@ylecun) seems incapable of giving credit where credit is due. @GaryMarcus is correct about LeCun's lack of integrity. Here he is rephrasing something that everyone else in the business

The FTC settles with Cox, MindSift, and 1010 Digital Works for $930K over claims they falsely said they could use phone mics to spy on users for ad targeting (Adi Robertson/The Verge)

SafetyDGX agent

Adi Robertson / The Verge: The FTC settles with Cox, MindSift, and 1010 Digital Works for $930K over claims they falsely said they could use phone mics to spy on users for ad targeting — More specific

The Implicit Bias of Depth: From Neural Collapse to Softmax Codes

SafetyDGX agent

arXiv:2605.23087v1 Announce Type: new Abstract: Neural collapse (NC) describes the structured geometry that emerges in the features and weights of trained classifiers. Recent theory suggests NC can be

The physics of AI weather models

SafetyDGX agent

arXiv:2605.23778v1 Announce Type: cross Abstract: Could it be that AI weather models are solving physical equations, although they may not be the equations used by conventional NWP models? We compute

The Pope rightly warns that AI must serve human dignity, not become a tool of domination or exclusion. But if we hand governments sweeping p…

SafetyDGX agent

The Pope rightly warns that AI must serve human dignity, not become a tool of domination or exclusion. But if we hand governments sweeping power over AI development in the name of safety, how do we pr

the tech bros think they know better than me they think they know better than @michaeljburry they think they know better than the Nobelist @…

SafetyDGX agent

the tech bros think they know better than me they think they know better than @michaeljburry they think they know better than the Nobelist @DAcemogluMIT they think they know better than @sundarpichai

this bit on physics is deep and actually resonates with the fact that LLMs are equally comfortable learning any world, whereas humans are bu…

SafetyDGX agent

this bit on physics is deep and actually resonates with the fact that LLMs are equally comfortable learning any world, whereas humans are built for our world. See also Chomsky’s 2023 conversation with

This creates problems for the @Kevinroose narrative that AGI is near.

SafetyDGX agent

This creates problems for the @Kevinroose narrative that AGI is near. Finally, a big name has the courage to tell it: we are nowhere near AGI. Demis Hassabis, CEO of Google DeepMind and Nobel laureate

Through the Stealth Lens: Attention-Aware Defenses Against Poisoning in RAG

SafetyDGX agent

arXiv:2506.04390v2 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) systems are vulnerable to attacks that inject poisoned passages into the retrieved context, even at low c

time to get out of index funds? they will soon be saddled with the IPOs of three giant companies that barely make a profit (or don’t make pr…

SafetyDGX agent

time to get out of index funds? they will soon be saddled with the IPOs of three giant companies that barely make a profit (or don’t make profits at all). I don’t personally want to be part of that. (

Transform-Invariant Generative Ray Path Sampling for Efficient Radio Propagation Modeling

SafetyDGX agent

arXiv:2603.01655v2 Announce Type: replace Abstract: Ray tracing has become a standard for accurate radio propagation modeling, but suffers from exponential computational complexity, as the number of c

UfM*: Uncertainty from Motion* for DNN Depth Estimation Using Gaussians

SafetyDGX agent

arXiv:2605.23098v1 Announce Type: new Abstract: Reliable uncertainty estimation is critical for deploying monocular depth deep neural networks (DNNs) in safety-critical robotic systems. Conventional u

Understanding Goal Generalisation in Sequential Reinforcement Learning

SafetyDGX agent

arXiv:2605.23565v1 Announce Type: cross Abstract: Reinforcement learning agents often exhibit unintended goal-directed behaviour outside their training distribution, but we currently lack a principled

UniEmo: Unifying Emotional Understanding and Generation with Learnable Expert Queries

SafetyDGX agent

arXiv:2507.23372v2 Announce Type: replace Abstract: Emotional understanding and generation are often treated as separate tasks, yet they are inherently complementary and can mutually enhance each othe

UniReg: A Universal Model for Controllable CT Image Registration

SafetyDGX agent

arXiv:2503.12868v2 Announce Type: replace Abstract: Learning-based medical image registration has matched the accuracy of conventional methods while offering superior computational efficiency. However

V-VLAPS: Value-Guided Planning for Vision-Language-Action Models

SafetyDGX agent

arXiv:2601.00969v2 Announce Type: replace-cross Abstract: Vision-language-action (VLA) models provide strong action priors for robotic manipulation, but their reactive behavior can fail under distribu

VI-CuRL: Stabilizing Verifier-Independent RL Reasoning via Confidence-Guided Variance Reduction

SafetyDGX agent

arXiv:2602.12579v2 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has emerged as a dominant paradigm for enhancing Large Language Models (LLMs) reasoning,

Vision Transformers Need Better Token Interaction

SafetyDGX agent

arXiv:2605.23868v1 Announce Type: new Abstract: Vision Transformers (ViTs) can learn strong image-level representations while their patch representations become less effective for dense prediction dur

What I don’t get is why so few people care that LeCun is basically a thief. I really don’t get it.

SafetyDGX agent

What I don’t get is why so few people care that LeCun is basically a thief. I really don’t get it. Dr. LeCun's heavily promoted Joint Embedding Predictive Architecture (JEPA, 2022) [5] is the heart of

When Determinants Are Not Enough: Private Rare Switching

SafetyDGX agent

arXiv:2605.23131v1 Announce Type: new Abstract: In this note, I would like to share a small research moment where Codex helped me find the right way to adapt rare switching to the private setting. The

Whose Good, Whose Place? The Moral Geography of Agentic AI for Social Good

SafetyDGX agent

arXiv:2605.22995v1 Announce Type: cross Abstract: Agentic AI systems are increasingly proposed for social-good domains, often invoking the United Nations Sustainable Development Goals (SDGs) as a voca

woah @Noahpinion is attacking @demishassabis here — calling what Demis said not “credible” — but misinterpreting what Hassabis is *actually*…

SafetyDGX agent

woah @Noahpinion is attacking @demishassabis here — calling what Demis said not “credible” — but misinterpreting what Hassabis is *actually* saying. What Demis actually said is both completely credibl

24 May 2026

AGI is certainly not here by the definitions I have repeatedly laid out. or by criteria that @hendrycks @Yoshua_Bengio and I and others rece…

SafetyDGX agent

AGI is certainly not here by the definitions I have repeatedly laid out. or by criteria that @hendrycks @Yoshua_Bengio and I and others recently laid about at http://agidefinition.AI i don’t think any

alas the humor and irony in this post seems to have been lost; let me spell it out:

SafetyDGX agent

Gary Marcus posted a clarification about humor and irony in a previous post that was apparently misunderstood by readers, indicating he felt the need to make his intended meaning more explicit. The po

Can't agree more. After trying out ChatGPT 5.4/5.5 for two months I feel it's far from replacing human researchers, even though it's enablin…

SafetyDGX agent

Can't agree more. After trying out ChatGPT 5.4/5.5 for two months I feel it's far from replacing human researchers, even though it's enabling me to do things not possible before. The effect on how out

clarifying: the issue is that alignment instructions and don’t pass, and the emotional weight that some people attach to LLMs can cause chal…

SafetyDGX agent

Gary Marcus discusses how alignment instructions in large language models often fail to work as intended, and argues that the emotional attachment some people develop toward LLMs can create additional

even @geohotz is starting to sound like me 🤣

SafetyDGX agent

Gary Marcus humorously notes that George Hotz, an AI researcher and entrepreneur, is beginning to echo Marcus's own views or criticisms, likely regarding AI safety, limitations, or technical concerns.

fascinating - the X algorithm literally gives no weight to being truthful or well-sourced. which kind of shows.

SafetyDGX agent

fascinating - the X algorithm literally gives no weight to being truthful or well-sourced. which kind of shows. So I spent some time studying the new Twitter/X algorithm today since the latest version

Gary has been wrong about plenty of things, and right about plenty more. That's what happens when you say things publicly about the future. …

SafetyDGX agent

Gary has been wrong about plenty of things, and right about plenty more. That's what happens when you say things publicly about the future. I disagree with him about lots of his expectations about the

@GaryMarcus Dr Marcus you've been right on pretty much all of this. The degree of hate is an indication of that.

SafetyDGX agent

This appears to be a social media comment acknowledging Gary Marcus's previous critical positions on AI, with the commenter suggesting that negative reactions directed at Marcus validate the accuracy

@GaryMarcus has won again.😂 (I'm from Hong Kong, so the pictures are in Traditional Chinese)

SafetyDGX agent

This post from Gary Marcus appears to reference a humorous or notable moment related to his work, with image attachments in Traditional Chinese that provide context to the claim that he 'has won again

@GaryMarcus @hendrycks @Yoshua_Bengio We recently tested all of the major LLMs with tic-tac-toe, modified chess, and a novel game -- even th…

SafetyDGX agent

@GaryMarcus @hendrycks @Yoshua_Bengio We recently tested all of the major LLMs with tic-tac-toe, modified chess, and a novel game -- even the top models all failed: illegal moves, claiming a win when

GOOGLE CEO: “THERE IS SOME IRRATIONALITY IN THE CURRENT AI BOOM. THE GROWTH OF AI INVESTMENT HAS BEEN AN EXTRAORDINARY MOMENT.” ASKED WHAT H…

SafetyDGX agent

GOOGLE CEO: “THERE IS SOME IRRATIONALITY IN THE CURRENT AI BOOM. THE GROWTH OF AI INVESTMENT HAS BEEN AN EXTRAORDINARY MOMENT.” ASKED WHAT HAPPENS IF THE AI BUBBLE POPS, HE SAID: “I THINK NO COMPANY I

have always wondered about this: is it that the true believers willfully simply don’t want to understand my actual positions or do they lack…

SafetyDGX agent

have always wondered about this: is it that the true believers willfully simply don’t want to understand my actual positions or do they lack the cognitive chops to follow them? @GaryMarcus This, sadly

here’s one sample, below, not even the only one I saw today; another claimed I don’t do research anymore when I had published articles in tw…

SafetyDGX agent

here’s one sample, below, not even the only one I saw today; another claimed I don’t do research anymore when I had published articles in two top journals (Science and the Proceedings of The Royal Soc

he’s not dumb and certainly his IQ is above average but as I wrote somewhere before “I think that Musk has consistently made a dumb set of c…

SafetyDGX agent

he’s not dumb and certainly his IQ is above average but as I wrote somewhere before “I think that Musk has consistently made a dumb set of choices - largely stemming from a mix of four cognitive error

hilarious how all the objections to this post about how elon understands engineering better than science point to elon’s engineering accompl…

SafetyDGX agent

hilarious how all the objections to this post about how elon understands engineering better than science point to elon’s engineering accomplishments, rather than his (meager) record of scientific inno

⚠️⚠️⚠️i don’t think most people understand the implications of the mood shift below, so I will spell them out. they are serious, and eventua…

SafetyDGX agent

⚠️⚠️⚠️i don’t think most people understand the implications of the mood shift below, so I will spell them out. they are serious, and eventually will affect the global economy. when a serious coder as

> I’m now in the LeCun/Marcus camp on LLMs > real programming agents will need world models > not some RLVR shit it’s over

SafetyDGX agent

> I’m now in the LeCun/Marcus camp on LLMs > real programming agents will need world models > not some RLVR shit it’s over The Eternal Sloptember https://geohot.github.io//blog/jekyll/update/2026/05/2

Interesting point … but are there any critics of mine that are “authentic” who have laid out a systematic critique, without relying on straw…

SafetyDGX agent

Interesting point … but are there any critics of mine that are “authentic” who have laid out a systematic critique, without relying on strawmen? @GaryMarcus While addressing some critics can be helpfu

interestingly (although it is bit apples and oranges) the number below is a similar to the fraction of cars involved in car accidents (mostl…

SafetyDGX agent

interestingly (although it is bit apples and oranges) the number below is a similar to the fraction of cars involved in car accidents (mostly nonfatal) in a week. but, crucially, we have tons of measu

← Previous
1…121122123124125…214
Next →