AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,704 results
Safety

Zoom In, Reason Out: Efficient Far-field Anomaly Detection in Expressway Surveillance Videos via Focused VLM Reasoning Guided by Bayesian Inference

DGX agent

arXiv:2604.23724v1 Announce Type: cross Abstract: Expressway video anomaly detection is essential for safety management. However, identifying anomalies across diverse scenes remains challenging, parti

safetyarxiv-cs-ai
28 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

A Co-Evolutionary Theory of Human-AI Coexistence: Mutualism, Governance, and Dynamics in Complex Societies

DGX agent

arXiv:2604.22227v1 Announce Type: cross Abstract: Classical robot ethics is often framed around obedience, most famously through Asimov's laws. This framing is too narrow for contemporary AI systems,

safetyarxiv-cs-ai
27 Apr 2026
Safety

AdaFair-MARL: Enforcing Adaptive Fairness Constraints in Multi-Agent Reinforcement Learning

DGX agent

arXiv:2511.14135v2 Announce Type: replace-cross Abstract: Fair workload enforcement in heterogeneous multi-agent systems that pursue shared objectives remains challenging. Fixed fairness penalties oft

safetyarxiv-cs-ai
27 Apr 2026
Safety

Adaptive vs. Static Robot-to-Human Handover: A Study on Orientation and Approach Direction

DGX agent

arXiv:2604.22378v1 Announce Type: new Abstract: Robot-to-human handovers often rely on static, open-loop strategies (or, at best, approaches that adapt only the position), which generally do not consi

safetyarxiv-cs-ro
27 Apr 2026
Safety

AgentBound: Securing Execution Boundaries of AI Agents

DGX agent

arXiv:2510.21236v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have evolved into AI agents that interact with external tools and environments to perform complex tasks. The Mode

safetyarxiv-cs-ai
27 Apr 2026
Safety

Algorithmic Feature Highlighting for Human-AI Decision-Making

DGX agent

arXiv:2604.22236v1 Announce Type: cross Abstract: Human decision-makers often face choices about complex cases with many potentially relevant features, but limited bandwidth to inspect and integrate a

safetyarxiv-cs-lg
27 Apr 2026
Safety

“Always keep a separate backup.” For sure, every experienced programmer should know this. 𝘉𝘶𝘵 𝘸𝘦 𝘫𝘶𝘴𝘵 𝘤𝘳𝘦𝘢𝘵𝘦𝘥 𝘢 𝘸𝘩𝘰𝘭𝘦 …

DGX agent

“Always keep a separate backup.” For sure, every experienced programmer should know this. 𝘉𝘶𝘵 𝘸𝘦 𝘫𝘶𝘴𝘵 𝘤𝘳𝘦𝘢𝘵𝘦𝘥 𝘢 𝘸𝘩𝘰𝘭𝘦 𝘨𝘦𝘯𝘦𝘳𝘢𝘵𝘪𝘰𝘯 𝘰𝘧 𝘷𝘪𝘣𝘦 𝘤𝘰𝘥𝘦𝘳𝘴 𝘸𝘩𝘰 𝘥𝘰𝘯’𝘵. And we now have a legion of synthetic coding

safetygary-marcus--x
27 Apr 2026
Safety

An Integrated Framework for Explainable, Fair, and Observable Hospital Readmission Prediction: Development and Validation on MIMIC-IV

DGX agent

arXiv:2604.22535v1 Announce Type: new Abstract: Objective: To propose and retrospectively validate an integrated framework addressing three barriers to clinical translation of readmission prediction:

safetyarxiv-cs-lg
27 Apr 2026
Safety

Beyond Chain-of-Thought: Rewrite as a Universal Interface for Generative Multimodal Embeddings

DGX agent

arXiv:2604.22280v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have emerged as a promising foundation for universal multimodal embeddings. Recent studies have shown that reas

safetyarxiv-cs-cv
27 Apr 2026
Safety

btw polling software is broken! votes aren’t showing, as many folks have noted. eg I see none:

DGX agent

Gary Marcus reported on X that polling software is experiencing a malfunction where votes are not displaying, a problem he notes has been observed by multiple users. He confirmed the issue from his ow

safetygary-marcus--x
27 Apr 2026
Safety

Calibrated Principal Component Regression

DGX agent

arXiv:2510.19020v2 Announce Type: replace-cross Abstract: We propose a new method for statistical inference in generalized linear models. In the overparameterized regime, Principal Component Regressio

safetyarxiv-cs-lg
27 Apr 2026
Safety

CGC: Compositional Grounded Contrast for Fine-Grained Multi-Image Understanding

DGX agent

arXiv:2604.22498v1 Announce Type: cross Abstract: Although Multimodal Large Language Models (MLLMs) have advanced rapidly, they still face notable challenges in fine-grained multi-image understanding,

safetyarxiv-cs-ai
27 Apr 2026
Safety

Clutter-Robust Vision-Language-Action Models through Object-Centric and Geometry Grounding

DGX agent

arXiv:2512.22519v2 Announce Type: replace Abstract: Recent Vision-Language-Action (VLA) models have made impressive progress toward general-purpose robotic manipulation by post-training large Vision-L

safetyarxiv-cs-ro
27 Apr 2026
Safety

CognitiveTwin: Robust Multi-Modal Digital Twins for Predicting Cognitive Decline in Alzheimer's Disease

DGX agent

arXiv:2604.22428v1 Announce Type: new Abstract: Predicting individual cognitive decline in Alzheimer's disease (AD) is difficult due to the heterogeneity of disease progression. Reliable clinical tool

safetyarxiv-cs-ai
27 Apr 2026
Safety

Concave Statistical Utility Maximization Bandits via Influence-Function Gradients

DGX agent

arXiv:2604.22140v1 Announce Type: cross Abstract: We study stochastic multi-armed bandits in which the objective is a statistical functional of the long-run reward distribution, rather than expected r

safetyarxiv-cs-lg
27 Apr 2026
Safety

Conditional Diffusion Posterior Alignment for Sparse-View CT Reconstruction

DGX agent

arXiv:2604.21960v1 Announce Type: cross Abstract: Computed Tomography (CT) is a widely used imaging modality in medical and industrial applications. To limit radiation exposure and measurement time, t

safetyarxiv-cs-cv
27 Apr 2026
Safety

Context-Fidelity Boosting: Enhancing Faithful Generation through Watermark-Inspired Decoding

DGX agent

arXiv:2604.22335v1 Announce Type: new Abstract: Large language models (LLMs) often produce content that contradicts or overlooks information provided in the input context, a phenomenon known as faithf

safetyarxiv-cs-cl
27 Apr 2026
Safety

Controllable Spoken Dialogue Generation: An LLM-Driven Grading System for K-12 Non-Native English Learners

DGX agent

arXiv:2604.22542v1 Announce Type: cross Abstract: Large language models (LLMs) often fail to meet the pedagogical needs of K-12 English learners in non-native contexts due to a proficiency mismatch. T

safetyarxiv-cs-ai
27 Apr 2026
Safety

Cross-Session Decoding of Neural Spiking Data via Task-Conditioned Latent Alignment

DGX agent

arXiv:2601.19963v2 Announce Type: replace-cross Abstract: Training a high-performing neural decoder can be difficult when only limited data are available from a recording session. To address this chal

safetyarxiv-cs-ai
27 Apr 2026
Safety

Data-Free Contribution Estimation in Federated Learning using Gradient von Neumann Entropy

DGX agent

arXiv:2604.22562v1 Announce Type: cross Abstract: Client contribution estimation in Federated Learning is necessary for identifying clients' importance and for providing fair rewards. Current methods

safetyarxiv-cs-ai
27 Apr 2026
Safety

Don't try to build a self-improving AI agent without evals. You are just wasting time and compute. An agent can't improve from traces it can…

DGX agent

Don't try to build a self-improving AI agent without evals. You are just wasting time and compute. An agent can't improve from traces it can't evaluate. This is why it's exciting to see @FutureAGI_ go

safetydair-ai--x
27 Apr 2026
Safety

dWorldEval: Scalable Robotic Policy Evaluation via Discrete Diffusion World Model

DGX agent

arXiv:2604.22152v1 Announce Type: new Abstract: Evaluating robotics policies across thousands of environments and thousands of tasks is infeasible with existing approaches. This motivates the need for

safetyarxiv-cs-ro
27 Apr 2026
Safety

Emergent Strategic Reasoning Risks in AI: A Taxonomy-Driven Evaluation Framework

DGX agent

arXiv:2604.22119v1 Announce Type: new Abstract: As reasoning capacity and deployment scope grow in tandem, large language models (LLMs) gain the capacity to engage in behaviors that serve their own ob

safetyarxiv-cs-ai
27 Apr 2026
Safety

Estimating Tail Risks in Language Model Output Distributions

DGX agent

arXiv:2604.22167v1 Announce Type: cross Abstract: Language models are increasingly capable and are being rapidly deployed on a population-level scale. As a result, the safety of these models is increa

safetyarxiv-cs-ai
27 Apr 2026
Safety

Ethics Testing: Proactive Identification of Generative AI System Harms

DGX agent

arXiv:2604.22089v1 Announce Type: cross Abstract: Generative Artificial Intelligence (GAI) systems that can automatically generate content in the form of source code or other contents (e.g., images) h

safetyarxiv-cs-ai
27 Apr 2026
Safety

FlowAnchor: Stabilizing the Editing Signal for Inversion-Free Video Editing

DGX agent

arXiv:2604.22586v1 Announce Type: new Abstract: We propose FlowAnchor, a training-free framework for stable and efficient inversion-free, flow-based video editing. Inversion-free editing methods have

safetyarxiv-cs-cv
27 Apr 2026
Safety

@GaryMarcus

DGX agent

Gary Marcus is a prominent AI researcher, cognitive scientist, and entrepreneur who frequently discusses artificial intelligence, machine learning, and cognitive science on X (formerly Twitter). His p

safetygary-marcus--x
27 Apr 2026
Safety

@GaryMarcus Never ever use generative AI for anything critical. The technology is probablistic all the way down and therefore inherently unr…

DGX agent

@GaryMarcus Never ever use generative AI for anything critical. The technology is probablistic all the way down and therefore inherently unreliable. Why is this so hard to understand? It’s wild how ma

safetygary-marcus--x
27 Apr 2026
Safety

@GaryMarcus These coding tools are intellectual chain saws. Powerful in the hands of a caring professional, extremely dangerous to the opera…

DGX agent

@GaryMarcus These coding tools are intellectual chain saws. Powerful in the hands of a caring professional, extremely dangerous to the operator when used improperly. Implying AGI-adjacency has created

safetygary-marcus--x
27 Apr 2026
Safety

Generative AI as the Hindenberg

DGX agent

Generative AI as the Hindenberg 🦔 Michael Wooldridge, professor of AI at Oxford, is warning that the race to market AI has raised the risk of a 'Hindenburg moment' that could shatter global confidence

safetygary-marcus--x
27 Apr 2026
Safety

Handling Missing Modalities in Multimodal Survival Prediction for Non-Small Cell Lung Cancer

DGX agent

arXiv:2601.10386v2 Announce Type: replace-cross Abstract: Accurate survival prediction in Non-Small Cell Lung Cancer (NSCLC) requires integrating clinical, radiological, and histopathological data. Mu

safetyarxiv-cs-ai
27 Apr 2026
Safety

How Large Language Models Balance Internal Knowledge with User and Document Assertions

DGX agent

arXiv:2604.22193v1 Announce Type: new Abstract: Large language models (LLMs) often need to balance their internal parametric knowledge with external information, such as user beliefs and content from

safetyarxiv-cs-cl
27 Apr 2026
Safety

How Many Visual Levers Drive Urban Perception? Interventional Counterfactuals via Multiple Localised Edits

DGX agent

arXiv:2604.22103v1 Announce Type: cross Abstract: Street-view perception models predict subjective attributes such as safety at scale, but remain correlational: they do not identify which localized vi

safetyarxiv-cs-cv
27 Apr 2026
Safety

How Vulnerable Is My Learned Policy? Universal Adversarial Perturbation Attacks On Modern Behavior Cloning Policies

DGX agent

arXiv:2502.03698v4 Announce Type: replace Abstract: Learning from demonstrations is a popular approach to train AI models; however, their vulnerability to adversarial attacks remains underexplored. We

safetyarxiv-cs-lg
27 Apr 2026
Safety

I don’t always agree with @ElonMusk. I often disagree. Sometimes loudly. But he’s basically right here, as far I can see.

DGX agent

I don’t always agree with @ElonMusk. I often disagree. Sometimes loudly. But he’s basically right here, as far I can see. Scam Altman and Greg Stockman stole a charity. Full stop. Greg got tens of bil

safetygary-marcus--x
27 Apr 2026
Safety

Identifying and typifying demographic unfairness in phoneme-level embeddings of self-supervised speech recognition models

DGX agent

arXiv:2604.22631v1 Announce Type: new Abstract: Modern automatic speech recognition (ASR) systems have been observed to function better for certain speaker groups (SGs) than others, despite recent gai

safetyarxiv-cs-cl
27 Apr 2026
Safety

If extremely violent criminals are not imprisoned, eventually they will murder innocent people

DGX agent

If extremely violent criminals are not imprisoned, eventually they will murder innocent people This bodega owner told ABC a year ago that he fears for his safety in NY Last night, he was kiIIed by a s

safetyelon-musk--x
27 Apr 2026
Safety

Improving Driver Drowsiness Detection via Personalized EAR/MAR Thresholds and CNN-Based Classification

DGX agent

arXiv:2604.22479v1 Announce Type: new Abstract: Driver drowsiness is a major cause of traffic accidents worldwide, posing a serious threat to public safety. Vision-based driver monitoring systems ofte

safetyarxiv-cs-cv
27 Apr 2026
Safety

Is it me or X starting to look like a vibe coded mess? Polls are broken. Accounts are getting hacked. My DMs are full of phishing scams. Bas…

DGX agent

Gary Marcus discusses technical and security issues affecting the X platform, including malfunctioning polls, compromised accounts, and increased phishing scams in direct messages. The post appears to

safetygary-marcus--x
27 Apr 2026
Safety

Learning-augmented robotic automation for real-world manufacturing

DGX agent

arXiv:2604.22235v1 Announce Type: cross Abstract: Industrial robots are widely used in manufacturing, yet most manipulation still depends on fixed waypoint scripts that are brittle to environmental ch

safetyarxiv-cs-ai
27 Apr 2026
Safety

Learning Control Policies to Provably Satisfy Hard Affine Constraints for Black-Box Hybrid Dynamical Systems

DGX agent

arXiv:2604.22244v1 Announce Type: new Abstract: Ensuring safety for black-box hybrid dynamical systems presents significant challenges due to their instantaneous state jumps and unknown explicit nonli

safetyarxiv-cs-ro
27 Apr 2026
Safety

Learning Evidence Highlighting for Frozen LLMs

DGX agent

arXiv:2604.22565v1 Announce Type: cross Abstract: Large Language Models (LLMs) can reason well, yet often miss decisive evidence when it is buried in long, noisy contexts. We introduce HiLight, an Evi

safetyarxiv-cs-ai
27 Apr 2026
Safety

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form

DGX agent

arXiv:2408.16286v5 Announce Type: replace Abstract: Designing a safe policy for uncertain environments is crucial in real-world control systems. However, this challenge remains inadequately addressed

safetyarxiv-cs-lg
27 Apr 2026
Safety

Near-Optimal Regret for the Safe Learning-based Control of the Constrained Linear Quadratic Regulator

DGX agent

arXiv:2604.22158v1 Announce Type: cross Abstract: We study the problem of adaptive control of the stochastic linear quadratic regulator (LQR) with constraints that must be satisfied at every time step

safetyarxiv-cs-lg
27 Apr 2026
Safety

On Star Price's new @hitpausepod podcast, ControlAI's US Director Connor Leahy (@NPCollapse) explains how AIs can already tell when they're …

DGX agent

On Star Price's new @hitpausepod podcast, ControlAI's US Director Connor Leahy (@NPCollapse) explains how AIs can already tell when they're being tested. Currently, we can still catch them cheating, b

safetyconnor-leahy--x
27 Apr 2026
Safety

On the Properties of Feature Attribution for Supervised Contrastive Learning

DGX agent

arXiv:2604.22540v1 Announce Type: cross Abstract: Most Neural Networks (NNs) for classification are trained using Cross-Entropy as a loss function. This approach requires the model to have an explicit

safetyarxiv-cs-ai
27 Apr 2026
Safety

People keep blaming the users for the growing number of vibe-coded reasoners. Doing so is *half* right. Users *are*screwing up - by letting …

DGX agent

People keep blaming the users for the growing number of vibe-coded reasoners. Doing so is *half* right. Users *are*screwing up - by letting vibe coded stuff access their files, without proper backups,

safetygary-marcus--x
27 Apr 2026
Safety

PermaFrost-Attack: Stealth Pretraining Seeding(SPS) for planting Logic Landmines During LLM Training

DGX agent

arXiv:2604.22117v1 Announce Type: cross Abstract: Aligned large language models(LLMs) remain vulnerable to adversarial manipulation, and their dependence on web-scale pretraining creates a subtle but

safetyarxiv-cs-ai
27 Apr 2026
← Previous
1…223224225226227…265
Next →