AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
All
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,707 results
Safety

Towards Neuro-symbolic Causal Rule Synthesis, Verification, and Evaluation Grounded in Legal and Safety Principles

DGX agent

arXiv:2604.28087v1 Announce Type: cross Abstract: Rule-based systems remain central in safety-critical domains but often struggle with scalability, brittleness, and goal misspecification. These limita

safetyarxiv-cs-ai
1 May 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

TRUST: A Framework for Decentralized AI Service v.0.1

DGX agent

arXiv:2604.27132v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) and Multi-Agent Systems (MAS) in high-stakes domains demand reliable verification, yet centralized approaches suffer four

safetyarxiv-cs-ai
1 May 2026
Safety

Vision-Language Models Mistake Head Orientation for Gaze Direction: Nonverbal Conversation Cues

DGX agent

arXiv:2506.05412v3 Announce Type: replace-cross Abstract: Where someone looks is a nonverbal communication cue that children and adults readily use. How well can Vision-Language Models (VLMs) infer ga

safetyarxiv-cs-cl
1 May 2026
Safety

When Does Structure Matter in Continual Learning? Dimensionality Controls When Modularity Shapes Representational Geometry

DGX agent

arXiv:2604.27656v1 Announce Type: cross Abstract: To preserve previously learned representations, continual learning systems must strike a balance between plasticity, the ability to acquire new knowle

safetyarxiv-cs-ai
1 May 2026
Safety

who's ready to dance? 😆🕺🎶 uploaded 11 new favorite @suno songs to @spotify [search: 'captain yohei'] 💿 Captain Yohei Vol 1. 💿 ♫ Don't G…

DGX agent

who's ready to dance? 😆🕺🎶 uploaded 11 new favorite @suno songs to @spotify [search: 'captain yohei'] 💿 Captain Yohei Vol 1. 💿 ♫ Don't Go in the Office [hard rock] ♫ Nine-Nine-Six [hip-hop/rap] ♫ Fog F

safetyyohei-nakajima--x
1 May 2026
Safety

A Multimodal Pre-trained Network for Integrated EEG-Video Seizure Detection

DGX agent

arXiv:2604.26379v1 Announce Type: new Abstract: Reliable seizure detection in mouse models is essential for preclinical epilepsy research, yet manual review of synchronized video-EEG recordings is lab

safetyarxiv-cs-cv
30 Apr 2026
Safety

A Scaled Three-Vehicle Platooning Platform

DGX agent

arXiv:2604.25963v1 Announce Type: new Abstract: Vehicle platooning has attracted increasing attention as a promising approach to improve traffic efficiency, energy consumption, and roadway safety thro

safetyarxiv-cs-ro
30 Apr 2026
Safety

A Scoping Review of LLM-as-a-Judge in Healthcare and the MedJUDGE Framework

DGX agent

arXiv:2604.25933v1 Announce Type: cross Abstract: As large language models (LLMs) increasingly generate and process clinical text, scalable evaluation has become critical. LLM-as-a-Judge (LaaJ), which

safetyarxiv-cs-ai
30 Apr 2026
Safety

A Survey of Process Reward Models: From Outcome Signals to Process Supervisions for Large Language Models

DGX agent

arXiv:2510.08049v3 Announce Type: replace-cross Abstract: Although Large Language Models (LLMs) exhibit advanced reasoning ability, conventional alignment remains largely dominated by outcome reward m

safetyarxiv-cs-ai
30 Apr 2026
Safety

A Survey of Safe Reinforcement Learning and Constrained MDPs: A Technical Survey on Single-Agent and Multi-Agent Safety

DGX agent

arXiv:2505.17342v2 Announce Type: replace Abstract: Safe Reinforcement Learning (SafeRL) is the subfield of reinforcement learning that explicitly deals with safety constraints during the learning and

safetyarxiv-cs-lg
30 Apr 2026
Safety

A Survey on the Safety and Security Threats of Computer-Using Agents: JARVIS or Ultron?

DGX agent

arXiv:2505.10924v4 Announce Type: replace-cross Abstract: Recently, AI-driven interactions with computing devices have advanced from basic prototype tools to sophisticated, LLM-based systems that emul

safetyarxiv-cs-ai
30 Apr 2026
Safety

Accelerating RL Post-Training Rollouts via System-Integrated Speculative Decoding

DGX agent

arXiv:2604.26779v1 Announce Type: cross Abstract: RL post-training of frontier language models is increasingly bottlenecked by autoregressive rollout generation, making rollout acceleration a central

safetyarxiv-cs-cl
30 Apr 2026
Safety

Adversarial Robustness of NTK Neural Networks

DGX agent

arXiv:2604.25965v1 Announce Type: cross Abstract: Deep learning models are widely deployed in safety-critical domains, but remain vulnerable to adversarial attacks. In this paper, we study the adversa

safetyarxiv-cs-lg
30 Apr 2026
Safety

alignment in 2016: obviously any real AI will be made inside a faraday cage magnetically suspended in a 10×10×10 cube of telekill alloy alig…

DGX agent

alignment in 2016: obviously any real AI will be made inside a faraday cage magnetically suspended in a 10×10×10 cube of telekill alloy alignment in 2026: yeah we can not make it stop talking about go

safetyconnor-leahy--x
30 Apr 2026
Safety

ATLAS: An Annotation Tool for Long-horizon Robotic Action Segmentation

DGX agent

arXiv:2604.26637v1 Announce Type: cross Abstract: Annotating long-horizon robotic demonstrations with precise temporal action boundaries is crucial for training and evaluating action segmentation and

safetyarxiv-cs-ai
30 Apr 2026
Safety

Atomic-Probe Governance for Skill Updates in Compositional Robot Policies

DGX agent

arXiv:2604.26689v1 Announce Type: cross Abstract: Skill libraries in deployed robotic systems are continually updated through fine-tuning, fresh demonstrations, or domain adaptation, yet existing type

safetyarxiv-cs-ai
30 Apr 2026
Safety

Attribution-Guided Multimodal Deepfake Detection via Cross-Modal Forensic Fingerprints

DGX agent

arXiv:2604.26453v1 Announce Type: new Abstract: Audio-visual deepfakes have reached a level of realism that makes perceptual detection unreliable, threatening media integrity and biometric security. W

safetyarxiv-cs-cv
30 Apr 2026
Safety

Benchmarking the Safety of Large Language Models for Robotic Health Attendant Control

DGX agent

arXiv:2604.26577v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly considered for deployment as the control component of robotic health attendants, yet their safety in this

safetyarxiv-cs-ai
30 Apr 2026
Safety

Beyond Shortcuts: Mitigating Visual Illusions in Frozen VLMs via Qualitative Reasoning

DGX agent

arXiv:2604.26250v1 Announce Type: new Abstract: While Vision-Language Models (VLMs) have achieved state-of-the-art performance in general visual tasks, their perceptual robustness remains remarkably b

safetyarxiv-cs-cv
30 Apr 2026
Safety

Big Tech’s $700 billion spending on AI this year is called the ‘greatest capital misallocation in history’ https://trib.al/qUQIbrJ

DGX agent

Big Tech companies are projected to spend approximately $700 billion on AI infrastructure and development in the current year, a figure that AI researcher and entrepreneur Gary Marcus has criticized a

safetygary-marcus--x
30 Apr 2026
Safety

Classification of Public Opinion on the Free Nutritional Meal Program on YouTube Media Using the LSTM Method

DGX agent

arXiv:2604.26312v1 Announce Type: new Abstract: Public opinion towards the Free Nutritious Meal Program (MBG) on YouTube social media reflects diverse community responses. This study applies the Long

safetyarxiv-cs-cl
30 Apr 2026
Safety

Co-Learning Port-Hamiltonian Systems and Optimal Energy-Shaping Control

DGX agent

arXiv:2604.26172v1 Announce Type: cross Abstract: We develop a physics-informed learning framework for energy-shaping control of port-Hamiltonian (pH) systems from trajectory data. The proposed approa

safetyarxiv-cs-ai
30 Apr 2026
Safety

Consist-Retinex: One-Step Noise-Emphasized Consistency Training Accelerates High-Quality Retinex Enhancement

DGX agent

arXiv:2512.08982v2 Announce Type: replace-cross Abstract: Retinex-based low-light image enhancement benefits from separating reflectance and illumination, yet recent generative approaches often rely o

safetyarxiv-cs-ai
30 Apr 2026
Safety

Correcting Performance Estimation Bias in Imbalanced Classification with Minority Subconcepts

DGX agent

arXiv:2604.26024v1 Announce Type: cross Abstract: Class-level evaluation can conceal substantial performance disparities across subconcepts within the same class, causing models that perform well on a

safetyarxiv-cs-ai
30 Apr 2026
Safety

Crime Hotspot Prediction Using Deep Graph Convolutional Networks

DGX agent

arXiv:2506.13116v2 Announce Type: replace-cross Abstract: Crime hotspot prediction is critical for ensuring urban safety and effective law enforcement, it remains challenging due to complex spatial de

safetyarxiv-cs-cl
30 Apr 2026
Safety

Cuando sucede un problema en un puesto automatizado por la IA... ¿Quién es el responsable?

DGX agent

This post discusses the liability and accountability questions that arise when problems occur in AI-automated workplaces, exploring who bears responsibility—whether the AI developer, the employer, the

safetygary-marcus--x
30 Apr 2026
Safety

Culturally Aware GenAI Risks for Youth: Perspectives from Youth, Parents, and Teachers in a Non-Western Context

DGX agent

arXiv:2604.26494v1 Announce Type: cross Abstract: Generative AI tools are widely used by youth and have introduced new privacy and safety challenges. While prior research has explored youth's safety i

safetyarxiv-cs-ai
30 Apr 2026
Safety

Data-Centric Foundation Models in Computational Healthcare: A Survey

DGX agent

arXiv:2401.02458v3 Announce Type: replace-cross Abstract: The advent of foundation models (FMs) as an emerging suite of AI techniques has struck a wave of opportunities in computational healthcare. Th

safetyarxiv-cs-ai
30 Apr 2026
Safety

DC-Ada: Reward-Only Decentralized Sensor Adaptation for Heterogeneous Multi-Robot Teams

DGX agent

arXiv:2604.03905v2 Announce Type: replace-cross Abstract: Heterogeneity is a defining feature of deployed multi-robot teams: platforms often differ in sensing modalities, ranges, fields of view, and f

safetyarxiv-cs-ai
30 Apr 2026
Safety

Dear @elonmusk, If you still genuinely care about AI safety, you can’t let the Trump administration leave the AI industry almost entirely un…

DGX agent

Dear @elonmusk, If you still genuinely care about AI safety, you can’t let the Trump administration leave the AI industry almost entirely unregulated. You just can’t. - Gary The judge just instructed

safetygary-marcus--x
30 Apr 2026
Safety

Delta Score Matters! Spatial Adaptive Multi Guidance in Diffusion Models

DGX agent

arXiv:2604.26503v1 Announce Type: new Abstract: Diffusion models have achieved remarkable success in synthesizing complex static and temporal visuals, a breakthrough largely driven by Classifier-Free

safetyarxiv-cs-cv
30 Apr 2026
Safety

DORA: A Scalable Asynchronous Reinforcement Learning System for Language Model Training

DGX agent

arXiv:2604.26256v1 Announce Type: new Abstract: Reinforcement learning (RL) has become a critical paradigm for LLM post-training, yet the rollout phase -- accounting for 50--80% of total step time --

safetyarxiv-cs-lg
30 Apr 2026
Safety

Efficient and Interpretable Transformer for Counterfactual Fairness

DGX agent

arXiv:2604.26188v1 Announce Type: new Abstract: The growing reliance of machine learning models in high-stakes, highly regulated domains such as finance and insurance has created a growing tension bet

safetyarxiv-cs-lg
30 Apr 2026
Safety

Evaluating Strategic Reasoning in Forecasting Agents

DGX agent

arXiv:2604.26106v1 Announce Type: new Abstract: Forecasting benchmarks produce accuracy leaderboards but little insight into why some forecasters are more accurate than others. We introduce Bench to t

safetyarxiv-cs-ai
30 Apr 2026
Safety

Evaluating the Alignment Between GeoAI Explanations and Domain Knowledge in Satellite-Based Flood Mapping

DGX agent

arXiv:2604.26051v1 Announce Type: cross Abstract: The increasing number of satellites has improved the temporal resolution of Earth observation, making satellite-based flood mapping a promising approa

safetyarxiv-cs-ai
30 Apr 2026
Safety

EvoSelect: Data-Efficient LLM Evolution for Targeted Task Adaptation

DGX agent

arXiv:2604.26170v1 Announce Type: new Abstract: Adapting large language models (LLMs) to a targeted task efficiently and effectively remains a fundamental challenge. Such adaptation often requires ite

safetyarxiv-cs-cl
30 Apr 2026
Safety

FedPF: Accurate Target Privacy Preserving Federated Learning Balancing Fairness and Utility

DGX agent

arXiv:2510.26841v2 Announce Type: replace-cross Abstract: Federated Learning (FL) enables collaborative model training without data sharing, yet participants face a fundamental challenge, e.g., simult

safetyarxiv-cs-ai
30 Apr 2026
Safety

FLARE: Fully Integration of Vision-Language Representations for Deep Cross-Modal Understanding

DGX agent

arXiv:2504.09925v3 Announce Type: replace Abstract: We introduce FLARE, a family of vision language models (VLMs) with a fully vision-language alignment and integration paradigm. Unlike existing appro

safetyarxiv-cs-cv
30 Apr 2026
Safety

For better or worse, regulation for closed-source models served by a few (quite large) companies is easy. It is not as easy to imagine how y…

DGX agent

For better or worse, regulation for closed-source models served by a few (quite large) companies is easy. It is not as easy to imagine how you regulate open-source models that can be served by a range

safetyethan-mollick--x
30 Apr 2026
Safety

From Prompt Risk to Response Risk: Paired Analysis of Safety Behavior of Large Language Model

DGX agent

arXiv:2604.26052v1 Announce Type: new Abstract: Safety evaluations of large language models (LLMs) typically report binary outcomes such as attack success rate, refusal rate, or harmful/not-harmful re

safetyarxiv-cs-cl
30 Apr 2026
Safety

Fundamental Physics, Existential Risks and Human Futures

DGX agent

arXiv:2604.26530v1 Announce Type: cross Abstract: Over the past 25 years, I have been involved in some intriguing developments in the foundations of physics, exploring the quantum reality problem, the

safetyarxiv-cs-ai
30 Apr 2026
Safety

Generative Bid Shading in Real-Time Bidding Advertising

DGX agent

arXiv:2508.06550v3 Announce Type: replace-cross Abstract: Bid shading plays a crucial role in Real-Time Bidding (RTB) by adaptively adjusting the bid to avoid advertisers overspending. Existing mainst

safetyarxiv-cs-lg
30 Apr 2026
Safety

Glance-or-Gaze: Incentivizing LMMs to Adaptively Focus Search via Reinforcement Learning

DGX agent

arXiv:2601.13942v2 Announce Type: replace-cross Abstract: Large Multimodal Models (LMMs) have achieved remarkable success in visual understanding, yet they struggle with knowledge-intensive queries in

safetyarxiv-cs-ai
30 Apr 2026
Safety

GNC-Pose: Geometry-Aware GNC-PnP for Accurate 6D Pose Estimation

DGX agent

arXiv:2512.06565v2 Announce Type: replace Abstract: We present GNC-Pose, a fully learning-free monocular 6D object pose estimation pipeline for textured objects that combines rendering-based initializ

safetyarxiv-cs-cv
30 Apr 2026
Safety

Heterogeneous Adaptive Policy Optimization: Tailoring Optimization to Every Token's Nature

DGX agent

arXiv:2509.16591v2 Announce Type: replace Abstract: Using entropy as a measure of heterogeneity to guide optimization has emerged as a crucial research direction in Reinforcement Learning for LLMs. Ho

safetyarxiv-cs-cl
30 Apr 2026
Safety

Hierarchical Multi-Persona Induction from User Behavioral Logs: Learning Evidence-Grounded and Truthful Personas

DGX agent

arXiv:2604.26120v1 Announce Type: new Abstract: Behavioral logs provide rich signals for user modeling, but are noisy and interleaved across diverse intents. Recent work uses LLMs to generate interpre

safetyarxiv-cs-ai
30 Apr 2026
Safety

HiPAN: Hierarchical Posture-Adaptive Navigation for Quadruped Robots in Unstructured 3D Environments

DGX agent

arXiv:2604.26504v1 Announce Type: new Abstract: Navigating quadruped robots in unstructured 3D environments poses significant challenges, requiring goal-directed motion, effective exploration to escap

safetyarxiv-cs-ro
30 Apr 2026
Safety

Improving Bayesian Optimization for Portfolio Management with an Adaptive Scheduling

DGX agent

arXiv:2504.13529v4 Announce Type: replace Abstract: Existing black-box portfolio management systems are prevalent in the financial industry due to commercial and safety constraints, though their perfo

safetyarxiv-cs-lg
30 Apr 2026
← Previous
1…214215216217218…265
Next →