AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,314 results
Safety

Transformer-Based Autonomous Driving Models and Deployment-Oriented Compression: A Survey

DGX agent

arXiv:2304.10891v2 Announce Type: replace-cross Abstract: Transformer-based models are becoming a central paradigm in autonomous driving because they can capture long-range spatial dependencies, multi

safetyarxiv-cs-cv
13 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Causal Explanations from the Geometric Properties of ReLU Neural Networks

DGX agent

arXiv:2605.10396v1 Announce Type: new Abstract: Neural networks have proved an effective means of learning control policies for autonomous systems, but these learned policies are difficult to understa

safetyarxiv-cs-lg
12 May 2026
Safety

Code Mixologist : A Practitioner's Guide to Building Code-Mixed LLMs

DGX agent

arXiv:2602.11181v2 Announce Type: replace Abstract: Code-mixing and code-switching (CSW) remain challenging phenomena for large language models (LLMs). Despite recent advances in multilingual modeling

safetyarxiv-cs-cl
12 May 2026
Safety

Constraint-Aware Diffusion Priors for High-Fidelity and Versatile Quadruped Locomotion

DGX agent

arXiv:2605.08804v1 Announce Type: new Abstract: Reinforcement learning combined with imitation learning has significantly advanced biomimetic quadrupedal locomotion. However, scaling these frameworks

safetyarxiv-cs-ro
12 May 2026
Safety

EROAS: 3D Efficient Reactive Obstacle Avoidance System for Autonomous Underwater Vehicles using 2.5D Forward-Looking Sonar

DGX agent

arXiv:2411.05516v3 Announce Type: replace Abstract: Autonomous Underwater Vehicles (AUVs) have advanced significantly in obstacle detection and path planning through sonar, cameras, and learning-based

safetyarxiv-cs-ro
12 May 2026
Safety

Hierarchical End-to-End Taylor Bounds for Complete Neural Network Verification

DGX agent

arXiv:2605.10621v1 Announce Type: new Abstract: Reachability analysis of neural networks, which seeks to compute or bound the set of outputs attainable over a given input domain, is central to certify

safetyarxiv-cs-lg
12 May 2026
Safety

Mechanism Design Is Not Enough: Prosocial Agents for Cooperative AI

DGX agent

arXiv:2605.08426v1 Announce Type: cross Abstract: Ensuring that AI agents behave safely and beneficially when interacting with other parties has emerged as one of the central challenges of modern AI s

safetyarxiv-cs-ai
12 May 2026
Safety

NEXUS: Continual Learning of Symbolic Constraints for Safe and Robust Embodied Planning

DGX agent

arXiv:2605.09387v1 Announce Type: new Abstract: While Large Language Models (LLMs) have catalyzed progress in embodied intelligence, a fundamental gap between their inherent probabilistic uncertainty

safetyarxiv-cs-ai
12 May 2026
Safety

Reinforcement Learning for Scalable and Trustworthy Intelligent Systems

DGX agent

arXiv:2605.08378v1 Announce Type: cross Abstract: Reinforcement learning has become a powerful paradigm for improving the capability of intelligent systems, but its practical deployment faces two cent

safetyarxiv-cs-ai
12 May 2026
Safety

RUBEN: Rule-Based Explanations for Retrieval-Augmented LLM Systems

DGX agent

arXiv:2605.10862v1 Announce Type: new Abstract: This paper demonstrates RUBEN, an interactive tool for discovering minimal rules to explain the outputs of retrieval-augmented large language models (LL

safetyarxiv-cs-cl
12 May 2026
Safety

Verification Mirage: Mapping the Reliability Boundary of Self-Verification in Medical VQA

DGX agent

arXiv:2605.10850v1 Announce Type: new Abstract: Self-verification, re-invoking the same vision language model (VLM) in a fresh context to check its own generated answer, is increasingly used as a defa

safetyarxiv-cs-cv
12 May 2026
Model Releases

GLiGuard: Schema-Conditioned Classification for LLM Safeguard

DGX agent

arXiv:2605.07982v1 Announce Type: new Abstract: Ensuring safe, policy-compliant outputs from large language models requires real-time content moderation that can scale across multiple safety dimension

model-releasesarxiv-cs-cl
11 May 2026
Safety

GPO-V: Jailbreak Diffusion Vision Language Model by Global Probability Optimization

DGX agent

arXiv:2605.07399v1 Announce Type: new Abstract: Diffusion Vision-Language Models (dVLMs), built upon the non-causal foundations of Diffusion Large Language Models (dLLMs), have demonstrated remarkable

safetyarxiv-cs-cv
11 May 2026
Safety

MORPH-U: Multi-Objective Resilient Motion Planning for V2X-Enabled Autonomous Driving in High-Uncertainty Environments via Simulation

DGX agent

arXiv:2605.07370v1 Announce Type: cross Abstract: V2X can warn an autonomous vehicle about hazards beyond line-of-sight, but it also brings uncertainty: messages may be delayed, dropped, or even forge

safetyarxiv-cs-ai
11 May 2026
Safety

Operating Within the Operational Design Domain: Zero-Shot Perception with Vision-Language Models

DGX agent

arXiv:2605.07649v1 Announce Type: cross Abstract: Over the last few years, research on autonomous systems has matured to such a degree that the field is increasingly well-positioned to translate resea

safetyarxiv-cs-ai
11 May 2026
Safety

OrchJail: Jailbreaking Tool-Calling Text-to-Image Agents by Orchestration-Guided Fuzzing

DGX agent

arXiv:2605.07414v1 Announce Type: cross Abstract: Tool-calling text-to-image (T2I) agents can plan and execute multi-step tool chains to accomplish complex generation and editing queries. However, thi

safetyarxiv-cs-ai
11 May 2026
Safety

Sparse Autoencoders as Plug-and-Play Firewalls for Adversarial Attack Detection in VLMs

DGX agent

arXiv:2605.07447v1 Announce Type: cross Abstract: Vision-language models (VLMs) have advanced rapidly and are increasingly deployed in real-world applications, especially with the rise of agent-based

safetyarxiv-cs-ai
11 May 2026
Safety

LLM-Based Human-Agent Collaboration and Interaction Systems: A Survey

DGX agent

arXiv:2505.00753v5 Announce Type: replace Abstract: Recent advances in large language models (LLMs) have sparked growing interest in building fully autonomous agents. However, fully autonomous LLM-bas

safetyarxiv-cs-cl
7 May 2026
Safety

SafeRedir: Prompt Embedding Redirection for Robust Unlearning in Image Generation Models

DGX agent

arXiv:2601.08623v2 Announce Type: replace Abstract: Image generation models (IGMs), while capable of producing impressive and creative content, often memorize a wide range of undesirable concepts from

safetyarxiv-cs-cv
7 May 2026
Safety

Software Engineering for Self-Adaptive Robotics: A Research Agenda

DGX agent

arXiv:2505.19629v3 Announce Type: replace-cross Abstract: Self-adaptive robotic systems operate autonomously in dynamic and uncertain environments, requiring robust real-time monitoring and adaptive b

safetyarxiv-cs-ro
7 May 2026
Safety

Learning Reactive Dexterous Grasping via Hierarchical Task-Space RL Planning and Joint-Space QP Control

DGX agent

arXiv:2605.03363v1 Announce Type: new Abstract: In this work, we propose a hybrid hierarchical control framework for reactive dexterous grasping that explicitly decouples high-level spatial intent fro

safetyarxiv-cs-ro
6 May 2026
Safety

MAGE: Safeguarding LLM Agents against Long-Horizon Threats via Shadow Memory

DGX agent

arXiv:2605.03228v1 Announce Type: cross Abstract: As large language model (LLM)-powered agents are increasingly deployed to perform complex, real-world tasks, they face a growing class of attacks that

safetyarxiv-cs-cl
6 May 2026
Safety

Analyzing Adversarial Inputs in Deep Reinforcement Learning

DGX agent

arXiv:2402.05284v2 Announce Type: replace Abstract: In recent years, Deep Reinforcement Learning (DRL) has become a popular paradigm in machine learning due to its successful applications to real-worl

safetyarxiv-cs-lg
5 May 2026
Safety

Cut-In Gap Acceptance Toward Autonomous vs. Human-Driven Vehicles: Evidence from the Waymo Open Motion Dataset

DGX agent

arXiv:2605.01485v1 Announce Type: new Abstract: Autonomous vehicles (AVs) are widely known to follow conservative, rule-based motion policies that surrounding drivers can learn to anticipate. A direct

safetyarxiv-cs-ro
5 May 2026
Safety

Reliability-Oriented Multilingual Orthopedic Diagnosis: A Domain-Adaptive Modeling and a Conceptual Validation Framework

DGX agent

arXiv:2605.02266v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly proposed for clinical decision support including multilingual diagnosis in low-resource settings. However,

safetyarxiv-cs-cl
5 May 2026
Safety

SAGA: A Robust Self-Attention and Goal-Aware Anchor-based Planner for Safe UAV Autonomous Navigation

DGX agent

arXiv:2605.02301v1 Announce Type: new Abstract: Agile unmanned aerial vehicle (UAV) navigation in cluttered environments demands a planning architecture that is both computationally efficient and stru

safetyarxiv-cs-ro
5 May 2026
Safety

Visibility-Aware Mobile Grasping in Dynamic Environments

DGX agent

arXiv:2605.02487v1 Announce Type: new Abstract: This paper addresses the problem of mobile grasping in dynamic, unknown environments where a robot must operate under a limited field-of-view. The funda

safetyarxiv-cs-ro
5 May 2026
Safety

Zero-Shot, Safe and Time-Efficient UAV Navigation via Potential-Based Reward Shaping, Control Lyapunov and Barrier Functions

DGX agent

arXiv:2605.01787v1 Announce Type: cross Abstract: Autonomous navigation and obstacle avoidance remain a core challenge of modern Unmanned Aerial Vehicles (UAVs). While traditional control methods stru

safetyarxiv-cs-lg
5 May 2026
Model Releases

Jailbreaking Vision-Language Models Through the Visual Modality

DGX agent

arXiv:2605.00583v1 Announce Type: new Abstract: The visual modality of vision-language models (VLMs) is an underexplored attack surface for bypassing safety alignment. We introduce four jailbreak atta

model-releasesarxiv-cs-cv
4 May 2026
Safety

ReLay: Personalized LLM-Generated Plain-Language Summaries for Better Understanding, but at What Cost?

DGX agent

arXiv:2605.00468v1 Announce Type: new Abstract: Plain Language Summaries (PLS) aim to make research accessible to lay readers, but they are typically written in a one-size-fits-all style that ignores

safetyarxiv-cs-cl
4 May 2026
Safety

Detecting Clinical Discrepancies in Health Coaching Agents: A Dual-Stream Memory and Reconciliation Architecture

DGX agent

arXiv:2604.27045v1 Announce Type: cross Abstract: As Large Language Model (LLM) agents transition from single-session tools to persistent systems managing longitudinal healthcare journeys, their memor

safetyarxiv-cs-ai
1 May 2026
Safety

Mechanized Foundations of Structural Governance: Machine-Checked Proofs for Governed Intelligence

DGX agent

arXiv:2604.27289v1 Announce Type: new Abstract: We present five results in the theory of structural governance for cognitive workflow systems. Three are mechanized in Coq 8.19 using the Interaction Tr

safetyarxiv-cs-ai
1 May 2026
Safety

Test Before You Deploy: Governing Updates in the LLM Supply Chain

DGX agent

arXiv:2604.27789v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used as core dependencies in software systems. However, the hosted LLM services evolve continuously thro

safetyarxiv-cs-ai
1 May 2026
Safety

A Scoping Review of LLM-as-a-Judge in Healthcare and the MedJUDGE Framework

DGX agent

arXiv:2604.25933v1 Announce Type: cross Abstract: As large language models (LLMs) increasingly generate and process clinical text, scalable evaluation has become critical. LLM-as-a-Judge (LaaJ), which

safetyarxiv-cs-ai
30 Apr 2026
Safety

Rule-based High-Level Coaching for Goal-Conditioned Reinforcement Learning in Search-and-Rescue UAV Missions Under Limited-Simulation Training

DGX agent

arXiv:2604.26833v1 Announce Type: cross Abstract: This paper presents a hierarchical decision-making framework for unmanned aerial vehicle (UAV) missions motivated by search-and-rescue (SAR) scenarios

safetyarxiv-cs-ai
30 Apr 2026
Safety

Walk With Me: Long-Horizon Social Navigation for Human-Centric Outdoor Assistance

DGX agent

arXiv:2604.26839v1 Announce Type: new Abstract: Assisting humans in open-world outdoor environments requires robots to translate high-level natural-language intentions into safe, long-horizon, and soc

safetyarxiv-cs-ro
30 Apr 2026
Safety

Generative AI Carries Non-Democratic Biases and Stereotypes: Representation of Women, Black Individuals, Age Groups, and People with Disability in AI-Generated Images across Occupations

DGX agent

arXiv:2409.13869v2 Announce Type: replace-cross Abstract: In this study, I investigate how generative artificial intelligence (AI) systems reproduce and reinforce societal biases, with a specific focu

safetyarxiv-cs-cl
29 Apr 2026
Safety

Learning from Medical Entity Trees: An Entity-Centric Medical Data Engineering Framework for MLLMs

DGX agent

arXiv:2604.25296v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have shown transformative potential in medical applications, yet their performance is hindered by conventional

safetyarxiv-cs-cl
29 Apr 2026
Safety

Sustained Gradient Alignment Mediates Subliminal Learning in a Multi-Step Setting: Evidence from MNIST Auxiliary Logit Distillation Experiment

DGX agent

arXiv:2604.25779v1 Announce Type: new Abstract: In the MNIST auxiliary logit distillation experiment, a student can acquire an unintended teacher trait despite distilling only on no-class logits throu

safetyarxiv-cs-lg
29 Apr 2026
Safety

Vocabulary Dropout for Curriculum Diversity in LLM Co-Evolution

DGX agent

arXiv:2604.03472v2 Announce Type: replace Abstract: Co-evolutionary self-play, where one language model generates problems and another solves them, promises autonomous curriculum learning without huma

safetyarxiv-cs-cl
29 Apr 2026
Safety

Adaptive Multi-Subspace Representation Steering for Attribute Alignment in Large Language Models

DGX agent

arXiv:2508.10599v4 Announce Type: replace Abstract: Activation steering offers a promising approach to controlling the behavior of Large Language Models by directly manipulating their internal activat

safetyarxiv-cs-ai
28 Apr 2026
Safety

AMAVA: Adaptive Motion-Aware Video-to-Audio Framework for Visually-Impaired Assistance

DGX agent

arXiv:2604.23909v1 Announce Type: new Abstract: Navigational aids for blind and low vision individuals struggle conveying dynamic real-world environments, leading to cognitive overload from continuous

safetyarxiv-cs-cv
28 Apr 2026
Safety

An empirical evaluation of the risks of AI model updates using clinical data: stability, arbitrariness, and fairness

DGX agent

arXiv:2604.23954v1 Announce Type: new Abstract: Artificial Intelligence and Machine Learning (AI/ML) models used in clinical settings are increasingly deployed to support clinical decision-making. How

safetyarxiv-cs-ai
28 Apr 2026
Safety

Analytica: Soft Propositional Reasoning for Robust and Scalable LLM-Driven Analysis

DGX agent

arXiv:2604.23072v1 Announce Type: new Abstract: Large language model (LLM) agents are increasingly tasked with complex real-world analysis (e.g., in financial forecasting, scientific discovery), yet t

safetyarxiv-cs-ai
28 Apr 2026
Safety

ArgRE: Formal Argumentation for Conflict Resolution in Multi-Agent Requirements Negotiation

DGX agent

arXiv:2604.23124v1 Announce Type: cross Abstract: As software systems grow in complexity, they must satisfy an increasing number of competing quality attributes, making it essential to balance them in

safetyarxiv-cs-ai
28 Apr 2026
Safety

Autocorrelation Reintroduces Spectral Bias in KANs for Time Series Forecasting

DGX agent

arXiv:2604.23518v1 Announce Type: cross Abstract: Existing theory suggests that Kolmogorov-Arnold Networks (KANs) can overcome the spectral bias commonly observed in neural networks under the assumpti

safetyarxiv-cs-ai
28 Apr 2026
Model Releases

Beyond Context: Large Language Models' Failure to Grasp Users' Intent

DGX agent

arXiv:2512.21110v3 Announce Type: replace Abstract: Current Large Language Models (LLMs) safety approaches focus on explicitly harmful content while overlooking a critical vulnerability: the inability

model-releasesarxiv-cs-ai
28 Apr 2026
Safety

Context-Aware Hospitalization Forecasting Evaluations for Decision Support using LLMs

DGX agent

arXiv:2604.23949v1 Announce Type: new Abstract: Medical and public health experts must make real-time resource decisions, such as expanding hospital bed capacity, based on projected hospitalization tr

safetyarxiv-cs-ai
28 Apr 2026
← Previous
1…3334353637…257
Next →