AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,349 results
10 Apr 2026

RoSHI: A Versatile Robot-oriented Suit for Human Data In-the-Wild

SafetyDGX agent

arXiv:2604.07331v1 Announce Type: cross Abstract: Scaling up robot learning will likely require human data containing rich and long-horizon interactions in the wild. Existing approaches for collecting

Rotation Equivariant Convolutions in Deformable Registration of Brain MRI

SafetyDGX agent

arXiv:2604.08034v1 Announce Type: new Abstract: Image registration is a fundamental task that aligns anatomical structures between images. While CNNs perform well, they lack rotation equivariance - a

Seeing but Not Thinking: Routing Distraction in Multimodal Mixture-of-Experts

SafetyDGX agent

arXiv:2604.08541v1 Announce Type: cross Abstract: Multimodal Mixture-of-Experts (MoE) models have achieved remarkable performance on vision-language tasks. However, we identify a puzzling phenomenon t

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Seeing Like an AI: How LLMs Apply (and Misapply) Wikipedia Neutrality Norms

SafetyDGX agent

arXiv:2407.04183v4 Announce Type: replace Abstract: Large language models (LLMs) are trained on broad corpora and then used in communities with specialized norms. Is providing LLMs with community rule

Self-Debias: Self-correcting for Debiasing Large Language Models

SafetyDGX agent

arXiv:2604.08243v1 Announce Type: new Abstract: Although Large Language Models (LLMs) demonstrate remarkable reasoning capabilities, inherent social biases often cascade throughout the Chain-of-Though

Self-Distilled RLVR

SafetyDGX agent

arXiv:2604.03128v2 Announce Type: replace Abstract: On-policy distillation (OPD) has become a popular training paradigm in the LLM community. This paradigm selects a larger model as the teacher to pro

Semantic-Aware UAV Command and Control for Efficient IoT Data Collection

SafetyDGX agent

arXiv:2604.08153v1 Announce Type: new Abstract: Unmanned Aerial Vehicles (UAVs) have emerged as a key enabler technology for data collection from Internet of Things (IoT) devices. However, effective d

Semantic Noise Reduction via Teacher-Guided Dual-Path Audio-Visual Representation Learning

SafetyDGX agent

arXiv:2604.08147v1 Announce Type: cross Abstract: Recent advances in audio-visual representation learning have shown the value of combining contrastive alignment with masked reconstruction. However, j

SeMoBridge: Semantic Modality Bridge for Efficient Few-Shot Adaptation of CLIP

SafetyDGX agent

arXiv:2509.26036v3 Announce Type: replace Abstract: While Contrastive Language-Image Pretraining (CLIP) excels at zero-shot tasks by aligning image and text embeddings, its performance in few-shot cla

Shortcut Learning in Glomerular AI: Adversarial Penalties Hurt, Entropy Helps

SafetyDGX agent

arXiv:2604.07936v1 Announce Type: new Abstract: Stain variability is a pervasive source of distribution shift and potential shortcut learning in renal pathology AI. We ask whether lupus nephritis glom

SIM1: Physics-Aligned Simulator as Zero-Shot Data Scaler in Deformable Worlds

SafetyDGX agent

arXiv:2604.08544v1 Announce Type: cross Abstract: Robotic manipulation with deformable objects represents a data-intensive regime in embodied learning, where shape, contact, and topology co-evolve in

SMPL-GPTexture: Dual-View 3D Human Texture Estimation using Text-to-Image Generation Models

SafetyDGX agent

arXiv:2504.13378v2 Announce Type: replace-cross Abstract: Generating high-quality, photorealistic textures for 3D human avatars remains a fundamental yet challenging task in computer vision and multim

Soft-Quantum Algorithms

SafetyDGX agent

arXiv:2604.06523v1 Announce Type: cross Abstract: Quantum operations on pure states can be fully represented by unitary matrices. Variational quantum circuits, also known as quantum neural networks, e

Spike-based alignment learning solves the weight transport problem

SafetyDGX agent

arXiv:2503.02642v3 Announce Type: replace-cross Abstract: In both machine learning and in computational neuroscience, plasticity in functional neural networks is frequently expressed as gradient desce

Sumo: Dynamic and Generalizable Whole-Body Loco-Manipulation

SafetyDGX agent

arXiv:2604.08508v1 Announce Type: new Abstract: This paper presents a sim-to-real approach that enables legged robots to dynamically manipulate large and heavy objects with whole-body dexterity. Our k

SYN-DIGITS: A Synthetic Control Framework for Calibrated Digital Twin Simulation

SafetyDGX agent

arXiv:2604.07513v1 Announce Type: cross Abstract: AI-based persona simulation -- often referred to as digital twin simulation -- is increasingly used for market research, recommender systems, and soci

Synthetic Data for any Differentiable Target

SafetyDGX agent

arXiv:2604.08423v1 Announce Type: new Abstract: What are the limits of controlling language models via synthetic training data? We develop a reinforcement learning (RL) primitive, the Dataset Policy G

Temporal Inversion for Learning Interval Change in Chest X-Rays

SafetyDGX agent

arXiv:2604.04563v2 Announce Type: replace-cross Abstract: Recent advances in vision--language pretraining have enabled strong medical foundation models, yet most analyze radiographs in isolation, over

The Defense Trilemma: Why Prompt Injection Defense Wrappers Fail?

SafetyDGX agent

arXiv:2604.06436v2 Announce Type: cross Abstract: We prove that no continuous, utility-preserving wrapper defense-a function D: Xo X that preprocesses inputs before the model sees them-can make al

The End of the Foundation Model Era: Open-Weight Models, Sovereign AI, and Inference as Infrastructure

SafetyDGX agent

arXiv:2604.06217v1 Announce Type: cross Abstract: The foundation model era -- roughly 2020 to 2025 -- is over. The forces that defined it have inverted. Open source models have reached frontier perfor

The Master Key Hypothesis: Unlocking Cross-Model Capability Transfer via Linear Subspace Alignment

SafetyDGX agent

arXiv:2604.06377v1 Announce Type: cross Abstract: We investigate whether post-trained capabilities can be transferred across models without retraining, with a focus on transfer across different model

The Sustainability Gap in Robotics: A Large-Scale Survey of Sustainability Awareness in 50,000 Research Articles

SafetyDGX agent

arXiv:2604.07921v1 Announce Type: new Abstract: We present a large-scale survey of sustainability communication and motivation in robotics research. Our analysis covers nearly 50,000 open-access paper

The Theorems of Dr. David Blackwell and Their Contributions to Artificial Intelligence

SafetyDGX agent

arXiv:2604.06621v1 Announce Type: cross Abstract: Dr. David Blackwell was a mathematician and statistician of the first rank, whose contributions to statistical theory, game theory, and decision theor

this from @Kasparov63 applies to the AI autocrats as well. “he would never sell my personal conversations to the government.” oh, yes, he wo…

SafetyDGX agent

this from @Kasparov63 applies to the AI autocrats as well. “he would never sell my personal conversations to the government.” oh, yes, he would. I will point out preemptively that one of the autocrat'

This is despicable. Satanic maybe. So AI can talk your kids into suicide and there's nothing you can do about it. Join MAMA and be a part of…

SafetyDGX agent

This is despicable. Satanic maybe. So AI can talk your kids into suicide and there's nothing you can do about it. Join MAMA and be a part of the future we all want and deserve. Mothers Against Media A

To a degree that may surprise some people, I agree with much of this* from @deanwball and would only add that you don’t have to believe that…

SafetyDGX agent

To a degree that may surprise some people, I agree with much of this* from @deanwball and would only add that you don’t have to believe that AGI is remotely close to want to find—ASAP—a regulatory reg

Today we visited Japan's hottest AI startup @SakanaAILabs🎏🇯🇵! We met their research scientists and discussed the implications and impact …

SafetyDGX agent

Today we visited Japan's hottest AI startup @SakanaAILabs🎏🇯🇵! We met their research scientists and discussed the implications and impact of some their works like 'The AI scientist' and 'Continous Thou

Towards Privacy-Preserving Large Language Model: Text-free Inference Through Alignment and Adaptation

SafetyDGX agent

arXiv:2604.06831v1 Announce Type: cross Abstract: Current LLM-based services typically require users to submit raw text regardless of its sensitivity. While intuitive, such practice introduces substan

TwinLoop: Simulation-in-the-Loop Digital Twins for Online Multi-Agent Reinforcement Learning

SafetyDGX agent

arXiv:2604.06610v1 Announce Type: cross Abstract: Decentralised online learning enables runtime adaptation in cyber-physical multi-agent systems, but when operating conditions change, learned policies

URMF: Uncertainty-aware Robust Multimodal Fusion for Multimodal Sarcasm Detection

SafetyDGX agent

arXiv:2604.06728v1 Announce Type: cross Abstract: Multimodal sarcasm detection (MSD) aims to identify sarcastic intent from semantic incongruity between text and image. Although recent methods have im

ViVa: A Video-Generative Value Model for Robot Reinforcement Learning

SafetyDGX agent

arXiv:2604.08168v1 Announce Type: new Abstract: Vision-language-action (VLA) models have advanced robot manipulation through large-scale pretraining, but real-world deployment remains challenging due

We are not getting to the G in Artificial General Intelligence; we are getting to (impressive) advances in particular areas where particular…

SafetyDGX agent

We are not getting to the G in Artificial General Intelligence; we are getting to (impressive) advances in particular areas where particular (verifiable) techniques can be used, on problems with advan

WebExpert: domain-aware web agents with critic-guided expert experience for high-precision search

SafetyDGX agent

arXiv:2604.06177v1 Announce Type: cross Abstract: Specialized web tasks in finance, biomedicine, and pharmaceuticals remain challenging due to missing domain priors: queries drift, evidence is noisy,

What Drives Representation Steering? A Mechanistic Case Study on Steering Refusal

SafetyDGX agent

arXiv:2604.08524v1 Announce Type: cross Abstract: Applying steering vectors to large language models (LLMs) is an efficient and effective model alignment technique, but we lack an interpretable explan

What Makes an Ideal Quote? Recommending 'Unexpected yet Rational' Quotations via Novelty

SafetyDGX agent

arXiv:2602.22220v2 Announce Type: replace-cross Abstract: Quotation recommendation aims to enrich writing by suggesting quotes that complement a given context, yet existing systems mostly optimize sur

What They Saw, Not Just Where They Looked: Semantic Scanpath Similarity via VLMs and NLP metric

SafetyDGX agent

arXiv:2604.08494v1 Announce Type: cross Abstract: Scanpath similarity metrics are central to eye-movement research, yet existing methods predominantly evaluate spatial and temporal alignment while neg

When Numbers Speak: Aligning Textual Numerals and Visual Instances in Text-to-Video Diffusion Models

SafetyDGX agent

arXiv:2604.08546v1 Announce Type: new Abstract: Text-to-video diffusion models have enabled open-ended video synthesis, but often struggle with generating the correct number of objects specified in a

9 Apr 2026

A different sense of the word “bubble”

SafetyDGX agent

A different sense of the word “bubble” @GaryMarcus It's actually kinda hilarious that @OpenAI thinks that buying a niche tech podcast followed mostly by Bay Area insiders is going to solve their colos

Also not a sign of someone who has any clue about the realities of current neuroscience.

SafetyDGX agent

Also not a sign of someone who has any clue about the realities of current neuroscience. Sam Altman has admitted he is on a waitlist for a procedure that would digitize his brain. The procedure would

Folks, you can relax. Mythos is not some off-trend exponential gain. And the gains weren’t about recursive self-improvement. Good thread fro…

SafetyDGX agent

Folks, you can relax. Mythos is not some off-trend exponential gain. And the gains weren’t about recursive self-improvement. Good thread from @ramez. Anthropic's Mythos does not appear to show any acc

My considered opinion is that The Mythos stuff was mostly a myth. Take it as a serious warning sign that we need to get our act together wit…

SafetyDGX agent

My considered opinion is that The Mythos stuff was mostly a myth. Take it as a serious warning sign that we need to get our act together with respect to cybersecurity. But don’t take the details serio

Perhaps because professional investors would look more carefully at the numbers? Wouldn’t want that!

SafetyDGX agent

Perhaps because professional investors would look more carefully at the numbers? Wouldn’t want that! OpenAI intends to set aside a share allocation for retail investors when the company goes public, t

The AI industry’s race for profits is now existential

SafetyDGX agent

Today on Decoder, let’s talk about the looming AI monetization cliff, and whether some of the biggest companies in the space can become real, profitable businesses before they careen right off it. My

There are plenty of people who are awed by AI who are not coders, I think the argument that AI impresses programmers most is, in part, selec…

SafetyDGX agent

There are plenty of people who are awed by AI who are not coders, I think the argument that AI impresses programmers most is, in part, selection bias on X, which is heavy on coders and people making f

This – “The CEO of Google DeepMind (@demishassabis) just admitted that if the decision had been his, we would've cured cancer before anyone …

SafetyDGX agent

This – “The CEO of Google DeepMind (@demishassabis) just admitted that if the decision had been his, we would've cured cancer before anyone ever used ChatGPT.” is exactly what i am trying to say in my

8 Apr 2026

1. If true (and it does fit with my perceptions FWIW), this is an amazing and incredibly damning graph 2. Can anyone find the source on whic…

SafetyDGX agent

1. If true (and it does fit with my perceptions FWIW), this is an amazing and incredibly damning graph 2. Can anyone find the source on which it is based? 'Anthropic, OpenAl and Google release their n

A few weeks ago I had a conversation with an American who genuinely believed Europe and Canada would help the United States in its war with …

SafetyDGX agent

A few weeks ago I had a conversation with an American who genuinely believed Europe and Canada would help the United States in its war with Iran. I asked him why he thought that, given that Trump had

Curious how many large organization CISO offices have taken the Mythos red team reports as the red alert that it is. (I suspect very few) Ba…

SafetyDGX agent

Curious how many large organization CISO offices have taken the Mythos red team reports as the red alert that it is. (I suspect very few) Based on historical trends in AI they have, at most, about six

Director James Cameron on why Big Tech owning AGI is scarier than any science fiction he's ever made: 'AGI will not emerge from a government…

SafetyDGX agent

Director James Cameron on why Big Tech owning AGI is scarier than any science fiction he's ever made: 'AGI will not emerge from a government funded program. It will emerge from one of the tech giants

Governance-Aware Agent Telemetry for Closed-Loop Enforcement in Multi-Agent AI Systems

SafetyDGX agent

Enterprise multi-agent AI systems produce thousands of inter-agent interactions per hour, yet existing observability tools capture these dependencies without enforcing anything. OpenTelemetry and Lang

“Social media algs reward engagement, and LLMs are excellent at writing in different styles, so people use LLMs to translate posts and news …

SafetyDGX agent

“Social media algs reward engagement, and LLMs are excellent at writing in different styles, so people use LLMs to translate posts and news stories into [exaggerated, misleading] versions that get mor

Started a Substack & will post X articles too! I think its a good thing to put out more policies for discussion & @WillManidis did a great p…

SafetyDGX agent

Started a Substack & will post X articles too! I think its a good thing to put out more policies for discussion & @WillManidis did a great policy on the politics The math though is... not great and I

The scariest part of this is that Anthropic showed some restraint in not releasing a potentially dangerous technology but some of their comp…

SafetyDGX agent

The scariest part of this is that Anthropic showed some restraint in not releasing a potentially dangerous technology but some of their competitors (such as OpenAI and xAI) might well not. Whether Myt

this is interesting. 1. Did Anthropic forget to run a control? 2. Where does this leave us?

SafetyDGX agent

this is interesting. 1. Did Anthropic forget to run a control? 2. Where does this leave us? New post: We tested the Mythos showcase vulnerabilities with open models. They recovered similar scoped anal

To anyone who read Rebooting AI back in (checks notes) 2019, this is both hilarious and unsuprising. The field has wasted 7 years on an arch…

SafetyDGX agent

To anyone who read Rebooting AI back in (checks notes) 2019, this is both hilarious and unsuprising. The field has wasted 7 years on an architecture that can’t solve one of the most basic litmus tests

Want more proof that Anthropic's PR has no idea what it's talking about? The talk of Mythos being 'their most aligned model ever'. They coul…

SafetyDGX agent

Want more proof that Anthropic's PR has no idea what it's talking about? The talk of Mythos being 'their most aligned model ever'. They could perhaps truthfully speak about 'new high scores on our ali

What Should We Take From Anthropic’s (possibly) Terrifying New Report on Mythos? – ⁦@garymarcus’s latest @CACMmag⁩ https://cacm.acm.org/blog…

SafetyDGX agent

What Should We Take From Anthropic’s (possibly) Terrifying New Report on Mythos? – ⁦@garymarcus’s latest @CACMmag⁩ https://cacm.acm.org/blogcacm/what-should-we-take-from-anthropics-possibly-terrifying

13 Aug 2026

Uncertainty-Aware Compositional Localization and Placement Assessment of Catheters and Tubes in Chest X-Rays

Local AiDGX agent

arXiv:2608.11288v1 Announce Type: cross Abstract: Assessing catheter and tube placement on chest X-rays is safety-critical yet tedious and error-prone. Current deep learning methods either classify pl

12 Aug 2026

Mistral says its platform will support third-party open models, starting with Z.ai's GLM-5.2, and run them on the same infrastructure as its own models (Mistral AI Blog)

Model ReleasesDGX agent

Mistral AI Blog: Mistral says its platform will support third-party open models, starting with Z.ai's GLM-5.2, and run them on the same infrastructure as its own models — At Mistral, we believe every

Rethinking LLM Verification: Evidence Structure, Uncertainty, and Selective Refinement

Model ReleasesDGX agent

arXiv:2608.10725v1 Announce Type: new Abstract: Large language models (LLMs) often rely on shortcuts rather than systematic reasoning, raising safety concerns in medical applications. Allowing models

← Previous
1…218219220221222…240
Next →