AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,314 results
Safety

SONG: A Photorealistic 3D Gaussian Simulation Platform for Benchmarking Social Navigation

DGX agent

arXiv:2607.25219v1 Announce Type: new Abstract: Social navigation has progressed from simplified 2D environments toward a more general vision-based setting, in which a robot needs to achieve socially

safetyarxiv-cs-ro
29 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

A Computational Ethical Framework for Financial Digital Phenotyping for Mental Health

DGX agent

arXiv:2607.24275v1 Announce Type: cross Abstract: Ethical governance of AI-driven systems is often expressed through high-level principles and static documentation, creating a gap between regulatory r

safetyarxiv-cs-ai
28 Jul 2026
Safety

Adapter Merging Reactivates Latent Reasoning Traces: A Mechanism Analysis

DGX agent

arXiv:2601.18350v5 Announce Type: replace-cross Abstract: Large language models fine-tuned via a two-stage pipeline (domain adaptation followed by instruction alignment) can exhibit non-trivial interf

safetyarxiv-cs-ai
28 Jul 2026
Safety

ARdena: Scenario-driven control of real-time LLM agents

DGX agent

arXiv:2607.22651v1 Announce Type: new Abstract: Large language models (LLMs) have enabled increasingly capable conversational agents, but reliably controlling their behavior in real-time interactive e

safetyarxiv-cs-ai
28 Jul 2026
Safety

Frustratingly Simple Black-Box Adaptation of Language Models via Logit Bias

DGX agent

arXiv:2607.22837v1 Announce Type: cross Abstract: Many organizations aim to adapt language models for internal use, both to improve performance on domain-specific tasks and to address privacy concerns

safetyarxiv-cs-ai
28 Jul 2026
Safety

How LLM Task-Adaptation Reshapes Alignment: A Multi-dimensional Study of Behavioral and Representational Drift

DGX agent

arXiv:2607.22676v1 Announce Type: new Abstract: Post-training is a key mechanism for adapting large language models to downstream tasks. While prior work suggests that task adaptation can alter a mode

safetyarxiv-cs-ai
28 Jul 2026
Safety

IJCB-AFMFR 2026: Competition on Adapting Foundation Models for Face Recognition Using Synthetic Training Data

DGX agent

arXiv:2607.24422v1 Announce Type: new Abstract: This paper presents a summary of the Competition on Adapting Foundation Models for Face Recognition Using Synthetic Training Data (AFMFR), held at the 2

safetyarxiv-cs-cv
28 Jul 2026
Safety

Position: AI/ML Deepfake Research is Misaligned with AI-Generated Non-Consensual Intimate Imagery (AIG-NCII)

DGX agent

arXiv:2607.18263v2 Announce Type: replace Abstract: AI-generated non-consensual intimate imagery (AIG-NCII) is not adequately addressed in AI/ML literature regarding AI-generated media, commonly refer

safetyarxiv-cs-ai
28 Jul 2026
Safety

RM-Distiller: Exploiting Generative LLM for Reward Model Distillation

DGX agent

arXiv:2601.14032v2 Announce Type: replace Abstract: Reward models (RMs) play a pivotal role in aligning large language models (LLMs) with human preferences. Due to the difficulty of obtaining high-qua

safetyarxiv-cs-cl
28 Jul 2026
Safety

Self-Authored Verification Is Unreliable in Heuristic Self-Improving Agents

DGX agent

arXiv:2607.24300v1 Announce Type: new Abstract: Self-improving agents accumulate capability by repeatedly rewriting procedural policies, controllers, or heuristic rules. They typically rely on self-au

safetyarxiv-cs-cl
28 Jul 2026
Safety

STAIF: A Stage-wise Optimization for Complex Instruction Following

DGX agent

arXiv:2607.22649v1 Announce Type: new Abstract: Following complex instructions with multiple explicit constraints remains a fundamental challenge for large language models (LLMs). Existing alignment m

safetyarxiv-cs-ai
28 Jul 2026
Safety

Tailored untruths: How personalisation challenges LLM safeguards

DGX agent

arXiv:2510.12993v3 Announce Type: replace Abstract: Large Language Models (LLMs) can generate highly persuasive disinformation, yet little is known about how effectively they personalise it across lan

safetyarxiv-cs-cl
28 Jul 2026
Safety

Unequal Trips, Unequal Places: Diagnosing and Mitigating Delay Inequity in Autonomous Vehicle Fleet Coordination

DGX agent

arXiv:2607.24336v1 Announce Type: new Abstract: City-scale autonomous vehicle fleet coordinators are typically optimized for aggregate travel time, yet fleet averages conceal how delay is distributed

safetyarxiv-cs-ai
28 Jul 2026
Safety

Interaction Dynamics Modeling and Predictive Control for Safe Steerable Catheter--Tissue Interaction

DGX agent

arXiv:2607.20939v1 Announce Type: cross Abstract: Safe steerable catheter control is fundamentally a problem of interaction dynamics: the tip must follow a planned motion, remain compliant against mov

safetyarxiv-cs-ai
24 Jul 2026
Safety

Robust Critics: Defending LLMs Against Multi-Turn Attacks

DGX agent

arXiv:2607.20472v1 Announce Type: new Abstract: When a user asks a language model something harmful, is it a genuine attack or a misunderstood but well-meaning question? This ambiguity is one of the c

safetyarxiv-cs-ai
24 Jul 2026
Safety

Safe and Scalable Multi-Drone Payload Transport via CBF-based Reinforcement Learning with Zero-Shot Sim-to-Real Transfer

DGX agent

arXiv:2607.20665v1 Announce Type: new Abstract: Multi-drone payload transportation has emerged as a promising research paradigm with potential applications in construction, logistics, and disaster res

safetyarxiv-cs-ro
24 Jul 2026
Safety

Formal Foundations for Known Good Reliable Die Screening in Chiplet-Based AI Systems-on-Chip

DGX agent

arXiv:2607.20141v1 Announce Type: cross Abstract: The rapid growth of chiplet-based artificial intelligence systems-on-chip (SoCs) has exposed a fundamental gap in semiconductor test methodology. Exis

safetyarxiv-cs-ai
23 Jul 2026
Safety

No Training, Better Flights: Test-Time Scaled VLMs for UAV Navigation

DGX agent

arXiv:2607.19288v1 Announce Type: new Abstract: Test-time scaling offers a promising method to improve the inference performance of Vision-Language Models (VLMs) without additional training. Existing

safetyarxiv-cs-cv
23 Jul 2026
Safety

OPIUM: Mitigating Steering Externalities and Over-Refusal via Dual Objective Latent Optimization

DGX agent

arXiv:2607.19806v1 Announce Type: cross Abstract: Activation steering provides a lightweight mechanism for controlling large language models at inference time, but steering vectors can have unintended

safetyarxiv-cs-ai
23 Jul 2026
Model Releases

A Comparative Evaluation of Large Vision-Language Models for 2D Object Detection under SOTIF Conditions

DGX agent

arXiv:2601.22830v2 Announce Type: replace Abstract: Reliable environmental perception remains one of the main obstacles for safe operation of automated vehicles. Safety of the Intended Functionality (

model-releasesarxiv-cs-cv
16 Jul 2026
Model Releases

Ask Before You Diagnose: Safe-Psych, a Sequential Evaluation Benchmark for LLMs in Psychiatry

DGX agent

arXiv:2607.13036v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used for decision support in healthcare, but clinical evidence is often incomplete or evolving. When the

model-releasesarxiv-cs-ai
16 Jul 2026
Safety

Distributionally Robust and Safe Imitation Learning

DGX agent

arXiv:2607.13436v1 Announce Type: new Abstract: Imitation learning (IL) has achieved remarkable success in complex decision-making tasks. However, its performance is highly sensitive to distribution s

safetyarxiv-cs-lg
16 Jul 2026
Safety

Earthquaker-AI: A Retrieval-Augmented Generation Framework with Rubric-Based Assessment for Primary School Earthquake Education

DGX agent

arXiv:2607.14046v1 Announce Type: new Abstract: This paper presents Earthquaker-AI, a hybrid educational framework building upon a previously implemented educational robotics project by integrating a

safetyarxiv-cs-ai
16 Jul 2026
Model Releases

Safeguard-Conditioned Uplift: Measuring Utility-Risk Frontiers for Dual-Use Biology Assistants

DGX agent

arXiv:2607.13039v1 Announce Type: cross Abstract: Safety evaluations for dual-use biology assistants often measure base-model capability, refusal behavior, or jailbreak success. These metrics miss a d

model-releasesarxiv-cs-ai
16 Jul 2026
Safety

A Neurosymbolic Approach to Natural Language Formalization and Verification

DGX agent

arXiv:2511.09008v2 Announce Type: replace-cross Abstract: Large Language Models perform well at natural language interpretation and reasoning, but their lack of formal correctness guarantees limits th

safetyarxiv-cs-ai
15 Jul 2026
Safety

Can Induced Emotion Bias LLM Behaviors in Sequential Decision Making?

DGX agent

arXiv:2607.12631v1 Announce Type: cross Abstract: As Large Language Models (LLMs) are increasingly deployed as autonomous agents in high-stakes domains, understanding contextual factors that may modul

safetyarxiv-cs-ai
15 Jul 2026
Safety

Expert Knowledge-driven Reinforcement Learning for Autonomous Racing via Trajectory Guidance and Dynamics Constraints

DGX agent

arXiv:2603.05842v2 Announce Type: replace Abstract: Reinforcement learning has demonstrated significant potential in the field of autonomous driving. However, it suffers from defects such as training

safetyarxiv-cs-ro
15 Jul 2026
Safety

From Sentiment to Actionable Insights: Public Sentiment Analysis of Advanced Air Mobility

DGX agent

arXiv:2606.20751v2 Announce Type: replace Abstract: Advanced Air Mobility (AAM) is an emerging low-altitude transportation system whose successful deployment depends on both technological progress and

safetyarxiv-cs-cl
15 Jul 2026
Safety

StratMamba: Strategic and Reactive Stream Partitioning for Path-Efficient LiDAR-Based Obstacle Avoidance

DGX agent

arXiv:2607.12370v1 Announce Type: new Abstract: This paper proposes StratMamba, a dual-stream Mamba-based temporal modeling architecture, to more efficiently capture long-horizon temporal dependencies

safetyarxiv-cs-ro
15 Jul 2026
Safety

Alignment Plausibility: A New Standard for Assuring AI in Healthcare

DGX agent

arXiv:2607.07766v1 Announce Type: new Abstract: Large language models (LLMs) have become significant providers of mental health support, yet they remain products of an attention economy whose operatio

safetyarxiv-cs-ai
10 Jul 2026
Safety

Post-Training in End-to-End Autonomous Driving

DGX agent

arXiv:2607.08072v1 Announce Type: new Abstract: End-to-end models that map multimodal inputs directly to future trajectories/maneuvers have emerged as an increasingly prominent research paradigm in au

safetyarxiv-cs-cv
10 Jul 2026
Safety

Securing Autonomous Vehicle Systems via Twin-Aware Federated Reinforcement Learning

DGX agent

arXiv:2607.08137v1 Announce Type: cross Abstract: Federated reinforcement learning (FRL) is crucial for enabling collaborative learning across multiple agents without sharing raw data, thereby enhanci

safetyarxiv-cs-lg
10 Jul 2026
Safety

Online Data Selection Is Implicit Alignment

DGX agent

arXiv:2607.07023v1 Announce Type: new Abstract: Supervised fine-tuning (SFT) is often treated as a capability-adaptation step, while alignment is attributed to later preference optimization or reinfor

safetyarxiv-cs-lg
9 Jul 2026
Safety

Progressive Crystallization: Turning Agent Exploration into Deterministic, Lower-Cost Workflows in Production

DGX agent

arXiv:2607.07052v1 Announce Type: cross Abstract: AI agents deployed for IT operations are typically permanent cost centers because every execution requires full LLM inference, even for previously sol

safetyarxiv-cs-ai
9 Jul 2026
Safety

Agentic AI for Commercial Insurance Underwriting with Adversarial Self-Critique

DGX agent

arXiv:2602.13213v2 Announce Type: replace Abstract: Commercial insurance underwriting is a labor-intensive process that requires manual review of extensive documentation to assess risk and determine p

safetyarxiv-cs-ai
8 Jul 2026
Safety

Closed-loop vs. Open-loop Kalman Filter Architectures in Airborne Aided Inertial Navigation

DGX agent

arXiv:2607.03338v1 Announce Type: new Abstract: Closed-loop (or feedback) error-state Kalman filters with their relatives and offspring are the state-of-the-art in modern aided inertial navigation res

safetyarxiv-cs-ro
7 Jul 2026
Safety

Directional Curvature from Armijo Backtracking: A Low-Cost Sharpness Probe and a Calibration-Free Learning-Rate Safeguard for Adam

DGX agent

arXiv:2607.03998v1 Announce Type: new Abstract: The local sharpness of the loss, the top Hessian eigenvalue lambda_1, determines the largest stable gradient step, but measuring it normally requires La

safetyarxiv-cs-lg
7 Jul 2026
Safety

Integrating Physics-Informed Neural Networks for Safe Reinforcement Learning in a 1-DoF Helicopter System

DGX agent

arXiv:2607.03125v1 Announce Type: new Abstract: Deep reinforcement learning (DRL) offers powerful control for industrial cyber-physical systems (ICPSs), but its 'black-box' exploration risks violating

safetyarxiv-cs-lg
7 Jul 2026
Safety

Interaction Dynamics for Dexterous Manipulation

DGX agent

arXiv:2606.14606v2 Announce Type: replace Abstract: Dexterous manipulation is fundamentally a problem of interaction dynamics: the hand must track precise finger trajectories, regulate the contact for

safetyarxiv-cs-ro
7 Jul 2026
Safety

NavEYE: Vision-Centered Multi-Sensor Fusion-Based Situational Awareness System for Intelligent Surface Vehicles

DGX agent

arXiv:2607.03915v1 Announce Type: new Abstract: With the rapid development of sensor and artificial intelligence (AI) technologies, intelligent surface vehicles (ISVs) have gained increasing attention

safetyarxiv-cs-cv
7 Jul 2026
Safety

Pretraining Curricula Enable Selective Fine-tuning

DGX agent

arXiv:2607.04846v1 Announce Type: cross Abstract: Transformers follow implicit curricula whereby some tasks are learned before others. However, how explicit pretraining curricula influence learning, g

safetyarxiv-cs-ai
7 Jul 2026
Safety

Safe Inference-Time Alignment via Lagrangian Reward Augmentation

DGX agent

arXiv:2607.02781v1 Announce Type: cross Abstract: Inference-time alignment steers a frozen language model during decoding using auxiliary reward signals, avoiding the cost of repeated weight updates.

safetyarxiv-cs-ai
7 Jul 2026
Model Releases

Toward Trustworthy Large Language Model Agents in Healthcare

DGX agent

arXiv:2607.05055v1 Announce Type: new Abstract: Healthcare appointment scheduling remains a persistent operational bottleneck, driven by manual coordination, fragmented legacy systems, and high admini

model-releasesarxiv-cs-ai
7 Jul 2026
Safety

Distributed Multi Robot Lunar Cargo Transportation via Phase Decomposed Reinforcement Learning

DGX agent

arXiv:2607.00160v1 Announce Type: new Abstract: Modular reconfigurable robotic systems provide a scalable solution for cooperative surface operations in future lunar missions. However, cooperative car

safetyarxiv-cs-ro
2 Jul 2026
Safety

NeHMO: Neural Hamilton-Jacobi Reachability Learning for Decentralized Safe Multi-Arm Motion Planning

DGX agent

arXiv:2607.00326v1 Announce Type: new Abstract: Safe multi-arm motion planning is a challenging problem in robotics due to its high dimensionality, coupled configuration space, and complex collision c

safetyarxiv-cs-ro
2 Jul 2026
Safety

Automating Cause-Effect Specification with Knowledge Graphs and Large Language Models

DGX agent

arXiv:2606.31614v1 Announce Type: cross Abstract: Engineering specifications such as interlocks, alarm rationalization tables, and cause-and-effect (C&E) matrices remain central to process control and

safetyarxiv-cs-ai
1 Jul 2026
Safety

Verification-Gated Agentic Mission-State Governance for Intelligent Industrial Multi-Robot Systems

DGX agent

arXiv:2606.31339v1 Announce Type: new Abstract: Agentic artificial intelligence is increasingly used to decompose industrial tasks, propose robot actions, and adapt execution plans in dynamic cyber-ph

safetyarxiv-cs-ro
1 Jul 2026
Safety

Carolina Guide: A Multi-Agent RAG System with Institutional Guardrails for Academic Policy Assistance

DGX agent

arXiv:2606.28360v1 Announce Type: cross Abstract: University students often struggle to navigate complex academic policies, leading to advising bottlenecks and delayed access to critical information.

safetyarxiv-cs-ai
30 Jun 2026
← Previous
1…3031323334…257
Next →