AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,314 results
Safety

Adversarially Robust Control of Conditional Value-at-Risk via Rockafellar-Uryasev Conformal Inference

DGX agent

arXiv:2606.00320v1 Announce Type: new Abstract: We present an online, distribution-free framework for controlling the Conditional Value-at-Risk (CVaR), extending conformal tail risk control to non-sta

safetyarxiv-cs-lg
2 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Agent Operating Systems (AOS): Integrating Agentic Control Planes into, and Beyond, Traditional Operating Systems

DGX agent

arXiv:2606.01508v1 Announce Type: cross Abstract: Traditional operating systems were designed around deterministic programs, explicit control flow, and human initiated workflows. Their core abstractio

safetyarxiv-cs-ai
2 Jun 2026
Safety

Balancing Accuracy and Efficiency: Adaptive Dynamics Orchestration for Model Predictive Control

DGX agent

arXiv:2606.00085v1 Announce Type: new Abstract: Model Predictive Control (MPC) for autonomous navigation faces a fundamental trade-off between model accuracy and real-time efficiency. High-fidelity dy

safetyarxiv-cs-ro
2 Jun 2026
Safety

Bridging the Last Mile of Time Series Forecasting with LLM Agents

DGX agent

arXiv:2606.02497v1 Announce Type: new Abstract: Time series forecasting has advanced rapidly, especially with the emergence of foundation models that show strong zero-shot performance on numerical ext

safetyarxiv-cs-ai
2 Jun 2026
Safety

Collaborative Space Object Detection with Multi-Satellite Viewpoints in LEO Constellations

DGX agent

arXiv:2606.01895v1 Announce Type: cross Abstract: With the growing number of satellites in low Earth orbit (LEO) constellations, the near-Earth space environment has become increasingly congested, mak

safetyarxiv-cs-ai
2 Jun 2026
Safety

DeepIPCv3: Event-Aware Multi-Modal Sensor Fusion for Sudden Pedestrian Crossing Avoidance

DGX agent

arXiv:2606.01277v1 Announce Type: cross Abstract: Current end-to-end autonomous driving systems predominantly rely on frame-based sensors, which suffer from inherent perception latency and motion blur

safetyarxiv-cs-ai
2 Jun 2026
Safety

Dive into Ambiguity: A*-Inspired Multi-Agents Commonsense Obfuscation Attack on LLM Prompts

DGX agent

arXiv:2606.01441v1 Announce Type: new Abstract: Large language models (LLMs) excel in reasoning and knowledge-intensive tasks but remain vulnerable to prompt-level adversarial attacks that preserve in

safetyarxiv-cs-ai
2 Jun 2026
Safety

Efficient LLM Moderation with Multi-Layer Latent Prototypes

DGX agent

arXiv:2502.16174v4 Announce Type: replace-cross Abstract: Although modern LLMs are aligned with human values during post-training, robust moderation remains essential to prevent harmful outputs at dep

safetyarxiv-cs-ai
2 Jun 2026
Safety

Emergent Collaborative Deliberation in Multi-Model AI Systems: A BFT-Derived Protocol for Epistemic Synthesis

DGX agent

arXiv:2606.00005v1 Announce Type: new Abstract: We present the Consilium Protocol, a Byzantine Fault Tolerance-derived architecture for structured multi-model AI deliberation that treats inter-model d

safetyarxiv-cs-ai
2 Jun 2026
Safety

Interpretable Multimodal Gesture Recognition for Drone and Mobile Robot Teleoperation via Log-Likelihood Ratio Fusion

DGX agent

arXiv:2602.23694v3 Announce Type: replace-cross Abstract: Human operators are still frequently exposed to hazardous environments such as disaster zones and industrial facilities, where intuitive and r

safetyarxiv-cs-ai
2 Jun 2026
Safety

Jailbreaking Multimodal Large Language Models using Multi-Clip Video

DGX agent

arXiv:2606.02111v1 Announce Type: cross Abstract: As multimodal large language models (MLLMs) have advanced to process video inputs, concerns have emerged about their potential for malicious misuse. P

safetyarxiv-cs-ai
2 Jun 2026
Safety

MidSteer: Optimal Affine Framework for Steering Generative Models

DGX agent

arXiv:2605.05220v2 Announce Type: replace-cross Abstract: Steering intermediate representations has emerged as a powerful strategy for controlling generative models, particularly in post-deployment al

safetyarxiv-cs-ai
2 Jun 2026
Safety

NormEval: A Unified Multi-Metric Framework for Evaluating Semantic Fidelity in Text Normalization

DGX agent

arXiv:2511.20409v2 Announce Type: replace Abstract: Text normalization methods such as stemming and lemmatization are fundamental components of NLP pipelines. As new normalization tools are developed

safetyarxiv-cs-cl
2 Jun 2026
Safety

Probabilistic Performance Guarantees for Multi-Task Reinforcement Learning

DGX agent

arXiv:2602.02098v2 Announce Type: replace-cross Abstract: Multi-task reinforcement learning trains generalist policies that can execute multiple tasks. While recent years have seen significant progres

safetyarxiv-cs-ai
2 Jun 2026
Safety

Real-Time Sensing of Inaccessible Physical Fields via an Edge-Deployable Hardware-Portable Graph Neural Operator

DGX agent

arXiv:2604.01802v2 Announce Type: replace Abstract: Real-time inference of inaccessible interior physical fields from sparse boundary observations is a fundamental but unresolved problem in scientific

safetyarxiv-cs-lg
2 Jun 2026
Safety

RESBev: Making BEV Perception More Robust

DGX agent

arXiv:2603.09529v2 Announce Type: replace Abstract: Bird's-eye-view (BEV) perception has emerged as a cornerstone of autonomous driving systems, providing a structured, ego-centric representation crit

safetyarxiv-cs-cv
2 Jun 2026
Safety

Robust Integrated Planning and Control for Quadrotors in Dynamic Environments via NMPC with CBF Penalties

DGX agent

arXiv:2606.01038v1 Announce Type: new Abstract: This paper presents a new robust integrated planning and control (IPC) strategy for multirotor uncrewed aerial vehicles. We propose a nonlinear model pr

safetyarxiv-cs-ro
2 Jun 2026
Safety

SafeMCP: Proactive Power Regulation for LLM Agent Defense via Environment-Grounded Look-Ahead Reasoning

DGX agent

arXiv:2606.01991v1 Announce Type: new Abstract: As Large Language Model (LLM) agents increasingly leverage the Model Context Protocol (MCP) to operate in complex environments, the expansion of their a

safetyarxiv-cs-ai
2 Jun 2026
Safety

SurrogateSHAP: Training-Free Contributor Attribution for Text-to-Image (T2I) Models

DGX agent

arXiv:2601.22276v2 Announce Type: replace-cross Abstract: As Text-to-Image (T2I) diffusion models are increasingly used in real-world creative workflows, a principled framework for valuing contributor

safetyarxiv-cs-cv
2 Jun 2026
Safety

THRD: A Training-Free Multi-Turn Defense Framework for Jailbreak Attacks on Large Language Models

DGX agent

arXiv:2606.01738v1 Announce Type: cross Abstract: Multi-turn jailbreak attacks pose a growing threat to LLMs by exploiting conversational dynamics such as gradual escalation and cross-turn coordinatio

safetyarxiv-cs-ai
2 Jun 2026
Safety

Visual Persuasion: What Influences Decisions of Vision-Language Models?

DGX agent

arXiv:2602.15278v2 Announce Type: replace-cross Abstract: The web is littered with images, once created for human consumption and now increasingly interpreted by agents using vision-language models (V

safetyarxiv-cs-ai
2 Jun 2026
Safety

What to Test Next: Interpretable Coverage Gap Discovery in Driving VLMs

DGX agent

arXiv:2606.01624v1 Announce Type: new Abstract: Driving vision-language models (VLMs) must accurately understand scenes across diverse conditions defined by Operational Design Domains (ODDs), yet veri

safetyarxiv-cs-cv
2 Jun 2026
Safety

World Models: A Comprehensive Survey of Architectures, Methodologies, Reasoning Paradigms, and Applications

DGX agent

arXiv:2606.00133v1 Announce Type: new Abstract: World models, internal simulators that learn the structure and dynamics of an environment, have emerged as a central paradigm in the pursuit of artifici

safetyarxiv-cs-lg
2 Jun 2026
Safety

X-Foresight: A Joint Vision-Action Causal Forecasting Network via Predictive World Modeling

DGX agent

arXiv:2605.24892v2 Announce Type: replace Abstract: Physical world knowledge resides mainly in videos. Equipping Vision-Language-Action (VLA) models with such knowledge is fundamental for safe and gen

safetyarxiv-cs-cv
2 Jun 2026
Safety

A study on a Real-Time VR-Based Teleoperation Framework for Manipulator in Dynamic Environment

DGX agent

arXiv:2605.30989v1 Announce Type: new Abstract: Robot teleoperation enables safe, non-contact task execution in hazardous environments where direct human access is difficult, and its application has e

safetyarxiv-cs-ro
1 Jun 2026
Safety

Adaptive Artificial Time-Delay Control with Barrier Lyapunov Constraints for Euler-Lagrange Robots

DGX agent

arXiv:2605.31405v1 Announce Type: new Abstract: This paper addresses the challenge of simultaneously compensating for state-dependent uncertainties and enforcing time-varying state constraints in Eule

safetyarxiv-cs-ro
1 Jun 2026
Safety

Benchmarking Machine Learning Uncertainty Quantification Methodologies for Predicting Turbine Gas Temperature Degradation

DGX agent

arXiv:2605.30585v1 Announce Type: cross Abstract: Effective prognostics and health management of modern engines relies on accurate turbine gas temperature predictions and robust uncertainty quantifica

safetyarxiv-cs-ai
1 Jun 2026
Safety

Counterfactual Evaluation Reveals Hidden Capability Profiles in Clinical LLMs and Agents

DGX agent

arXiv:2605.30590v1 Announce Type: cross Abstract: Two clinical AI systems can score nearly identically on coverage-based rubrics yet behave radically differently when their patient inputs change: one

safetyarxiv-cs-ai
1 Jun 2026
Model Releases

CSULoRA: Closest Safe Update Low-Rank Adaptation

DGX agent

arXiv:2605.30640v1 Announce Type: cross Abstract: Low-rank adaptation has become a standard method for parameter-efficient fine-tuning of large language models, but even small amounts of unsafe or adv

model-releasesarxiv-cs-cl
1 Jun 2026
Safety

dashi: A Python library for Dataset Shift Characterization to Support Trustworthy AI Development and Deployment

DGX agent

arXiv:2605.31360v1 Announce Type: cross Abstract: The Artificial Intelligence (AI) life cycle requires a thorough understanding of the underlying data dynamics for robust, safe and cost-effective AI d

safetyarxiv-cs-ai
1 Jun 2026
Safety

Decoding the Surgical Scene: A Scoping Review of Scene Graphs in Surgery

DGX agent

arXiv:2509.20941v2 Announce Type: replace Abstract: As surgical AI transitions from pixel-level detection to complex reasoning, Scene Graphs (SGs) offer the structured, relational representations nece

safetyarxiv-cs-cv
1 Jun 2026
Safety

Does Visual Information Play a Decisive Role in Vision-Language-Action Model Driving Behavior?

DGX agent

arXiv:2605.31041v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have demonstrated promising capability in autonomous driving, highlighting the potential of unified multimodal arc

safetyarxiv-cs-ai
1 Jun 2026
Safety

Exploiting Chordal Sparsity for Globally Optimal Estimation with Factor Graphs

DGX agent

arXiv:2605.30617v1 Announce Type: new Abstract: Robust and efficient state estimation is crucial for perception, navigation, and control in robotics. State estimation problems are conveniently modeled

safetyarxiv-cs-ro
1 Jun 2026
Safety

From Internal Diagnosis to External Auditing: A VLM-Driven Paradigm for Data-Free Online Backdoor Defense

DGX agent

arXiv:2601.19448v2 Announce Type: replace Abstract: Deep Neural Networks remain inherently vulnerable to backdoor attacks. Traditional test-time defenses largely operate under the paradigm of internal

safetyarxiv-cs-lg
1 Jun 2026
Safety

Geometry-Aware Control Barrier Functions for Collision Avoidance via Bernstein Polynomial Approximations

DGX agent

arXiv:2605.30696v1 Announce Type: new Abstract: Safe navigation often relies on well-defined conditions based on the shape of robots and obstacles, and can be challenging when they have irregular geom

safetyarxiv-cs-ro
1 Jun 2026
Safety

GSAM: A Generalizable and Safe Robotic Framework for Articulated Object Manipulation

DGX agent

arXiv:2605.30740v1 Announce Type: cross Abstract: Articulated object manipulation is a unique challenge for service robots. Existing methods employ end-to-end policy learning, visionmotion planning, a

safetyarxiv-cs-ai
1 Jun 2026
Safety

Hybrid Energy-Aware Reward Shaping: A Unified Lightweight Physics-Guided Methodology for Policy Optimization

DGX agent

arXiv:2603.11600v2 Announce Type: replace Abstract: Deep reinforcement learning for continuous control often suffers from high variance, low energy efficiency, and poor generalization under distributi

safetyarxiv-cs-lg
1 Jun 2026
Safety

LiftNav: Path Planning via Semantic Lifting in TSDF-Guided Gaussian Splatting

DGX agent

arXiv:2605.31376v1 Announce Type: cross Abstract: Autonomous robots in unknown indoor environments require both reliable collision avoidance and object-level understanding. Classical representations s

safetyarxiv-cs-cv
1 Jun 2026
Safety

Reinforcement Learning Amplifies Emergent Misalignment from Harmless Rewards

DGX agent

arXiv:2605.31328v1 Announce Type: new Abstract: Emergent misalignment (EM) is the surprising tendency of language models to become broadly misaligned after fine-tuning on narrowly misaligned examples.

safetyarxiv-cs-cl
1 Jun 2026
Safety

Safeguarding Text-to-Image Generation via Inference-Time Prompt-Noise Optimization

DGX agent

arXiv:2412.03876v2 Announce Type: replace Abstract: Text-to-Image (T2I) diffusion models are widely recognized for their ability to generate high-quality and diverse images based on text prompts. Howe

safetyarxiv-cs-cv
1 Jun 2026
Safety

Used Car Salesbots? Honesty and Credulity of LLMs as Bargaining Agents under Partial Information

DGX agent

arXiv:2605.31445v1 Announce Type: cross Abstract: In this work we study agents in simulated bargaining scenarios, where a buyer and a seller communicate through a text channel and attempt to negotiate

safetyarxiv-cs-ai
1 Jun 2026
Safety

Agora: Toward Autonomous Bug Detection in Production-Level Consensus Protocols with LLM Agents

DGX agent

arXiv:2605.29910v1 Announce Type: cross Abstract: Consensus protocols form the backbone of distributed systems and blockchains, where implementation bugs can cause data corruption and financial losses

safetyarxiv-cs-ai
29 May 2026
Safety

Anytime-Valid Federated Conformal RAG for LLM Swarms

DGX agent

arXiv:2605.29139v1 Announce Type: cross Abstract: Federated Conformal RAG (FC-RAG) provides distribution-free coverage for a bandwidth-limited swarm of weak language models, but only at a fixed horizo

safetyarxiv-cs-lg
29 May 2026
Safety

Audio Jailbreaks in Large Audio-Language Models: Taxonomy, Attack-Defense Analysis, and Cost-Aware Evaluation

DGX agent

arXiv:2605.30031v1 Announce Type: cross Abstract: Large Audio Language Models (LALMs) expand jailbreak risks from token-level prompting to the full speech perception-to-reasoning pipeline, where unsaf

safetyarxiv-cs-ai
29 May 2026
Safety

DLM-SWAI: Steering Diffusion Language Models Before They Unmask

DGX agent

arXiv:2605.29626v1 Announce Type: cross Abstract: Steering language model generation toward desired textual properties is essential for practical deployment, and inference-time methods are particularl

safetyarxiv-cs-ai
29 May 2026
Safety

From General Vision to Reliable Traversability Estimation: Adapting Vision Foundation Models for Unstructured Outdoor Environments

DGX agent

arXiv:2605.29565v1 Announce Type: new Abstract: Vision-based approaches have become the dominant paradigm for traversability estimation in unstructured outdoor environments, typically adapting vision

safetyarxiv-cs-cv
29 May 2026
Safety

Harmonizing Real-Time Constraints and Long-Horizon Reasoning: An Asynchronous Agentic Framework for Dynamic Scheduling

DGX agent

arXiv:2605.29262v1 Announce Type: new Abstract: The Dynamic Flexible Job Shop Scheduling Problem (DFJSP) necessitates a trade-off between instant reaction to stochastic disturbances and global optimiz

safetyarxiv-cs-ai
29 May 2026
Safety

Jailbreaking and Mitigation of Vulnerabilities in Large Language Models

DGX agent

arXiv:2410.15236v4 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have transformed artificial intelligence by advancing natural language understanding and generation, enabling app

safetyarxiv-cs-ai
29 May 2026
← Previous
1…4647484950…257
Next →