AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,704 results
28 Apr 2026

Vision-Language-Action Safety: Threats, Challenges, Evaluations, and Mechanisms

SafetyDGX agent

arXiv:2604.23775v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models are emerging as a unified substrate for embodied intelligence. This shift raises a new class of safety challenges, s

Voxify3D: Pixel Art Meets Volumetric Rendering

SafetyDGX agent

arXiv:2512.07834v2 Announce Type: replace Abstract: Voxel art is a distinctive stylization widely used in games and digital media, yet automated generation from 3D meshes remains challenging due to co

When an LLM acts happy (“EUREKA!”) or sad (“I have failed…”), is that meaningless mimicry, or does it reflect something “real”? We don’t kno…

SafetyDGX agent

When an LLM acts happy (“EUREKA!”) or sad (“I have failed…”), is that meaningless mimicry, or does it reflect something “real”? We don’t know if LLMs are conscious. But they increasingly seem to exhib


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

When Annotators Agree but Labels Disagree: The Projection Problem in Stance Detection

SafetyDGX agent

arXiv:2603.24231v2 Announce Type: replace Abstract: Stance detection is nearly always formulated as classifying text into Favor, Against, or Neutral. This convention was inherited from debate analysis

When Policies Cannot Be Retrained: A Unified Closed-Form View of Post-Training Steering in Offline Reinforcement Learning

SafetyDGX agent

arXiv:2604.22873v1 Announce Type: cross Abstract: Offline reinforcement learning (RL) can learn effective policies from fixed datasets, but deployment objectives may change after training, and in many

When you project exponential growth – and miss. I stand by my call that OpenAI is likely to someday be seen as the WeWork of AI.

SafetyDGX agent

When you project exponential growth – and miss. I stand by my call that OpenAI is likely to someday be seen as the WeWork of AI. Exclusive: OpenAI recently missed its own targets for new users and rev

World-R1: Reinforcing 3D Constraints for Text-to-Video Generation

SafetyDGX agent

arXiv:2604.24764v1 Announce Type: new Abstract: Recent video foundation models demonstrate impressive visual synthesis but frequently suffer from geometric inconsistencies. While existing methods atte

Wow. Musk’s lawyers will capitalize on this.

SafetyDGX agent

Wow. Musk’s lawyers will capitalize on this. 🚨 Microsoft's lawyer just told the jury Microsoft 'knew nothing'. OpenAI's non-profit board approved all funding, every step of the way and Microsoft did i

XGRAG: A Graph-Native Framework for Explaining KG-based Retrieval-Augmented Generation

SafetyDGX agent

arXiv:2604.24623v1 Announce Type: new Abstract: Graph-based Retrieval-Augmented Generation (GraphRAG) extends traditional RAG by using knowledge graphs (KGs) to give large language models (LLMs) a str

Your Students Don't Use LLMs Like You Wish They Did

SafetyDGX agent

arXiv:2604.23486v1 Announce Type: new Abstract: Educational NLP systems are typically evaluated using engagement metrics and satisfaction surveys, which are at best a proxy for meeting pedagogical goa

Z^2-Sampling: Zero-Cost Zigzag Trajectories for Semantic Alignment in Diffusion Models

SafetyDGX agent

arXiv:2604.23536v1 Announce Type: new Abstract: Diffusion models have achieved unprecedented success in text-aligned generation, largely driven by Classifier-Free Guidance (CFG). However, standard CFG

ZenBrain: A Neuroscience-Inspired 7-Layer Memory Architecture for Autonomous AI Systems

SafetyDGX agent

arXiv:2604.23878v1 Announce Type: new Abstract: Despite a century of empirical memory research, existing AI agent memory systems rely on system-engineering metaphors (virtual-memory paging, flat LLM s

Zoom In, Reason Out: Efficient Far-field Anomaly Detection in Expressway Surveillance Videos via Focused VLM Reasoning Guided by Bayesian Inference

SafetyDGX agent

arXiv:2604.23724v1 Announce Type: cross Abstract: Expressway video anomaly detection is essential for safety management. However, identifying anomalies across diverse scenes remains challenging, parti

27 Apr 2026

A Co-Evolutionary Theory of Human-AI Coexistence: Mutualism, Governance, and Dynamics in Complex Societies

SafetyDGX agent

arXiv:2604.22227v1 Announce Type: cross Abstract: Classical robot ethics is often framed around obedience, most famously through Asimov's laws. This framing is too narrow for contemporary AI systems,

AdaFair-MARL: Enforcing Adaptive Fairness Constraints in Multi-Agent Reinforcement Learning

SafetyDGX agent

arXiv:2511.14135v2 Announce Type: replace-cross Abstract: Fair workload enforcement in heterogeneous multi-agent systems that pursue shared objectives remains challenging. Fixed fairness penalties oft

Adaptive vs. Static Robot-to-Human Handover: A Study on Orientation and Approach Direction

SafetyDGX agent

arXiv:2604.22378v1 Announce Type: new Abstract: Robot-to-human handovers often rely on static, open-loop strategies (or, at best, approaches that adapt only the position), which generally do not consi

AgentBound: Securing Execution Boundaries of AI Agents

SafetyDGX agent

arXiv:2510.21236v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have evolved into AI agents that interact with external tools and environments to perform complex tasks. The Mode

Algorithmic Feature Highlighting for Human-AI Decision-Making

SafetyDGX agent

arXiv:2604.22236v1 Announce Type: cross Abstract: Human decision-makers often face choices about complex cases with many potentially relevant features, but limited bandwidth to inspect and integrate a

“Always keep a separate backup.” For sure, every experienced programmer should know this. 𝘉𝘶𝘵 𝘸𝘦 𝘫𝘶𝘴𝘵 𝘤𝘳𝘦𝘢𝘵𝘦𝘥 𝘢 𝘸𝘩𝘰𝘭𝘦 …

SafetyDGX agent

“Always keep a separate backup.” For sure, every experienced programmer should know this. 𝘉𝘶𝘵 𝘸𝘦 𝘫𝘶𝘴𝘵 𝘤𝘳𝘦𝘢𝘵𝘦𝘥 𝘢 𝘸𝘩𝘰𝘭𝘦 𝘨𝘦𝘯𝘦𝘳𝘢𝘵𝘪𝘰𝘯 𝘰𝘧 𝘷𝘪𝘣𝘦 𝘤𝘰𝘥𝘦𝘳𝘴 𝘸𝘩𝘰 𝘥𝘰𝘯’𝘵. And we now have a legion of synthetic coding

An Integrated Framework for Explainable, Fair, and Observable Hospital Readmission Prediction: Development and Validation on MIMIC-IV

SafetyDGX agent

arXiv:2604.22535v1 Announce Type: new Abstract: Objective: To propose and retrospectively validate an integrated framework addressing three barriers to clinical translation of readmission prediction:

Beyond Chain-of-Thought: Rewrite as a Universal Interface for Generative Multimodal Embeddings

SafetyDGX agent

arXiv:2604.22280v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have emerged as a promising foundation for universal multimodal embeddings. Recent studies have shown that reas

btw polling software is broken! votes aren’t showing, as many folks have noted. eg I see none:

SafetyDGX agent

Gary Marcus reported on X that polling software is experiencing a malfunction where votes are not displaying, a problem he notes has been observed by multiple users. He confirmed the issue from his ow

Calibrated Principal Component Regression

SafetyDGX agent

arXiv:2510.19020v2 Announce Type: replace-cross Abstract: We propose a new method for statistical inference in generalized linear models. In the overparameterized regime, Principal Component Regressio

CGC: Compositional Grounded Contrast for Fine-Grained Multi-Image Understanding

SafetyDGX agent

arXiv:2604.22498v1 Announce Type: cross Abstract: Although Multimodal Large Language Models (MLLMs) have advanced rapidly, they still face notable challenges in fine-grained multi-image understanding,

Clutter-Robust Vision-Language-Action Models through Object-Centric and Geometry Grounding

SafetyDGX agent

arXiv:2512.22519v2 Announce Type: replace Abstract: Recent Vision-Language-Action (VLA) models have made impressive progress toward general-purpose robotic manipulation by post-training large Vision-L

CognitiveTwin: Robust Multi-Modal Digital Twins for Predicting Cognitive Decline in Alzheimer's Disease

SafetyDGX agent

arXiv:2604.22428v1 Announce Type: new Abstract: Predicting individual cognitive decline in Alzheimer's disease (AD) is difficult due to the heterogeneity of disease progression. Reliable clinical tool

Concave Statistical Utility Maximization Bandits via Influence-Function Gradients

SafetyDGX agent

arXiv:2604.22140v1 Announce Type: cross Abstract: We study stochastic multi-armed bandits in which the objective is a statistical functional of the long-run reward distribution, rather than expected r

Conditional Diffusion Posterior Alignment for Sparse-View CT Reconstruction

SafetyDGX agent

arXiv:2604.21960v1 Announce Type: cross Abstract: Computed Tomography (CT) is a widely used imaging modality in medical and industrial applications. To limit radiation exposure and measurement time, t

Context-Fidelity Boosting: Enhancing Faithful Generation through Watermark-Inspired Decoding

SafetyDGX agent

arXiv:2604.22335v1 Announce Type: new Abstract: Large language models (LLMs) often produce content that contradicts or overlooks information provided in the input context, a phenomenon known as faithf

Controllable Spoken Dialogue Generation: An LLM-Driven Grading System for K-12 Non-Native English Learners

SafetyDGX agent

arXiv:2604.22542v1 Announce Type: cross Abstract: Large language models (LLMs) often fail to meet the pedagogical needs of K-12 English learners in non-native contexts due to a proficiency mismatch. T

Cross-Session Decoding of Neural Spiking Data via Task-Conditioned Latent Alignment

SafetyDGX agent

arXiv:2601.19963v2 Announce Type: replace-cross Abstract: Training a high-performing neural decoder can be difficult when only limited data are available from a recording session. To address this chal

Data-Free Contribution Estimation in Federated Learning using Gradient von Neumann Entropy

SafetyDGX agent

arXiv:2604.22562v1 Announce Type: cross Abstract: Client contribution estimation in Federated Learning is necessary for identifying clients' importance and for providing fair rewards. Current methods

Don't try to build a self-improving AI agent without evals. You are just wasting time and compute. An agent can't improve from traces it can…

SafetyDGX agent

Don't try to build a self-improving AI agent without evals. You are just wasting time and compute. An agent can't improve from traces it can't evaluate. This is why it's exciting to see @FutureAGI_ go

dWorldEval: Scalable Robotic Policy Evaluation via Discrete Diffusion World Model

SafetyDGX agent

arXiv:2604.22152v1 Announce Type: new Abstract: Evaluating robotics policies across thousands of environments and thousands of tasks is infeasible with existing approaches. This motivates the need for

Emergent Strategic Reasoning Risks in AI: A Taxonomy-Driven Evaluation Framework

SafetyDGX agent

arXiv:2604.22119v1 Announce Type: new Abstract: As reasoning capacity and deployment scope grow in tandem, large language models (LLMs) gain the capacity to engage in behaviors that serve their own ob

Estimating Tail Risks in Language Model Output Distributions

SafetyDGX agent

arXiv:2604.22167v1 Announce Type: cross Abstract: Language models are increasingly capable and are being rapidly deployed on a population-level scale. As a result, the safety of these models is increa

Ethics Testing: Proactive Identification of Generative AI System Harms

SafetyDGX agent

arXiv:2604.22089v1 Announce Type: cross Abstract: Generative Artificial Intelligence (GAI) systems that can automatically generate content in the form of source code or other contents (e.g., images) h

FlowAnchor: Stabilizing the Editing Signal for Inversion-Free Video Editing

SafetyDGX agent

arXiv:2604.22586v1 Announce Type: new Abstract: We propose FlowAnchor, a training-free framework for stable and efficient inversion-free, flow-based video editing. Inversion-free editing methods have

@GaryMarcus

SafetyDGX agent

Gary Marcus is a prominent AI researcher, cognitive scientist, and entrepreneur who frequently discusses artificial intelligence, machine learning, and cognitive science on X (formerly Twitter). His p

@GaryMarcus Never ever use generative AI for anything critical. The technology is probablistic all the way down and therefore inherently unr…

SafetyDGX agent

@GaryMarcus Never ever use generative AI for anything critical. The technology is probablistic all the way down and therefore inherently unreliable. Why is this so hard to understand? It’s wild how ma

@GaryMarcus These coding tools are intellectual chain saws. Powerful in the hands of a caring professional, extremely dangerous to the opera…

SafetyDGX agent

@GaryMarcus These coding tools are intellectual chain saws. Powerful in the hands of a caring professional, extremely dangerous to the operator when used improperly. Implying AGI-adjacency has created

Generative AI as the Hindenberg

SafetyDGX agent

Generative AI as the Hindenberg 🦔 Michael Wooldridge, professor of AI at Oxford, is warning that the race to market AI has raised the risk of a 'Hindenburg moment' that could shatter global confidence

Handling Missing Modalities in Multimodal Survival Prediction for Non-Small Cell Lung Cancer

SafetyDGX agent

arXiv:2601.10386v2 Announce Type: replace-cross Abstract: Accurate survival prediction in Non-Small Cell Lung Cancer (NSCLC) requires integrating clinical, radiological, and histopathological data. Mu

How Large Language Models Balance Internal Knowledge with User and Document Assertions

SafetyDGX agent

arXiv:2604.22193v1 Announce Type: new Abstract: Large language models (LLMs) often need to balance their internal parametric knowledge with external information, such as user beliefs and content from

How Many Visual Levers Drive Urban Perception? Interventional Counterfactuals via Multiple Localised Edits

SafetyDGX agent

arXiv:2604.22103v1 Announce Type: cross Abstract: Street-view perception models predict subjective attributes such as safety at scale, but remain correlational: they do not identify which localized vi

How Vulnerable Is My Learned Policy? Universal Adversarial Perturbation Attacks On Modern Behavior Cloning Policies

SafetyDGX agent

arXiv:2502.03698v4 Announce Type: replace Abstract: Learning from demonstrations is a popular approach to train AI models; however, their vulnerability to adversarial attacks remains underexplored. We

I don’t always agree with @ElonMusk. I often disagree. Sometimes loudly. But he’s basically right here, as far I can see.

SafetyDGX agent

I don’t always agree with @ElonMusk. I often disagree. Sometimes loudly. But he’s basically right here, as far I can see. Scam Altman and Greg Stockman stole a charity. Full stop. Greg got tens of bil

Identifying and typifying demographic unfairness in phoneme-level embeddings of self-supervised speech recognition models

SafetyDGX agent

arXiv:2604.22631v1 Announce Type: new Abstract: Modern automatic speech recognition (ASR) systems have been observed to function better for certain speaker groups (SGs) than others, despite recent gai

If extremely violent criminals are not imprisoned, eventually they will murder innocent people

SafetyDGX agent

If extremely violent criminals are not imprisoned, eventually they will murder innocent people This bodega owner told ABC a year ago that he fears for his safety in NY Last night, he was kiIIed by a s

Improving Driver Drowsiness Detection via Personalized EAR/MAR Thresholds and CNN-Based Classification

SafetyDGX agent

arXiv:2604.22479v1 Announce Type: new Abstract: Driver drowsiness is a major cause of traffic accidents worldwide, posing a serious threat to public safety. Vision-based driver monitoring systems ofte

Is it me or X starting to look like a vibe coded mess? Polls are broken. Accounts are getting hacked. My DMs are full of phishing scams. Bas…

SafetyDGX agent

Gary Marcus discusses technical and security issues affecting the X platform, including malfunctioning polls, compromised accounts, and increased phishing scams in direct messages. The post appears to

Learning-augmented robotic automation for real-world manufacturing

SafetyDGX agent

arXiv:2604.22235v1 Announce Type: cross Abstract: Industrial robots are widely used in manufacturing, yet most manipulation still depends on fixed waypoint scripts that are brittle to environmental ch

Learning Control Policies to Provably Satisfy Hard Affine Constraints for Black-Box Hybrid Dynamical Systems

SafetyDGX agent

arXiv:2604.22244v1 Announce Type: new Abstract: Ensuring safety for black-box hybrid dynamical systems presents significant challenges due to their instantaneous state jumps and unknown explicit nonli

Learning Evidence Highlighting for Frozen LLMs

SafetyDGX agent

arXiv:2604.22565v1 Announce Type: cross Abstract: Large Language Models (LLMs) can reason well, yet often miss decisive evidence when it is buried in long, noisy contexts. We introduce HiLight, an Evi

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form

SafetyDGX agent

arXiv:2408.16286v5 Announce Type: replace Abstract: Designing a safe policy for uncertain environments is crucial in real-world control systems. However, this challenge remains inadequately addressed

Near-Optimal Regret for the Safe Learning-based Control of the Constrained Linear Quadratic Regulator

SafetyDGX agent

arXiv:2604.22158v1 Announce Type: cross Abstract: We study the problem of adaptive control of the stochastic linear quadratic regulator (LQR) with constraints that must be satisfied at every time step

On Star Price's new @hitpausepod podcast, ControlAI's US Director Connor Leahy (@NPCollapse) explains how AIs can already tell when they're …

SafetyDGX agent

On Star Price's new @hitpausepod podcast, ControlAI's US Director Connor Leahy (@NPCollapse) explains how AIs can already tell when they're being tested. Currently, we can still catch them cheating, b

On the Properties of Feature Attribution for Supervised Contrastive Learning

SafetyDGX agent

arXiv:2604.22540v1 Announce Type: cross Abstract: Most Neural Networks (NNs) for classification are trained using Cross-Entropy as a loss function. This approach requires the model to have an explicit

People keep blaming the users for the growing number of vibe-coded reasoners. Doing so is *half* right. Users *are*screwing up - by letting …

SafetyDGX agent

People keep blaming the users for the growing number of vibe-coded reasoners. Doing so is *half* right. Users *are*screwing up - by letting vibe coded stuff access their files, without proper backups,

PermaFrost-Attack: Stealth Pretraining Seeding(SPS) for planting Logic Landmines During LLM Training

SafetyDGX agent

arXiv:2604.22117v1 Announce Type: cross Abstract: Aligned large language models(LLMs) remain vulnerable to adversarial manipulation, and their dependence on web-scale pretraining creates a subtle but

← Previous
1…178179180181182…212
Next →