AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,314 results
Safety

LP-NavOA: Integrated Local Navigation and Obstacle Avoidance for Humanoid Robots under Limited Perception

DGX agent

arXiv:2606.23249v1 Announce Type: new Abstract: Humanoid local navigation in cluttered environments must jointly resolve obstacle avoidance, sparse-goal recovery, and stable whole-body locomotion unde

safetyarxiv-cs-ro
23 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Noise is Signal: Density-Based Outliers as Leading Indicators of Occupational Emergence in Labor Market Text

DGX agent

arXiv:2606.22769v1 Announce Type: new Abstract: Standard NLP pipelines for occupational clustering discard the 10-15% of job postings that density-based methods assign to noise. We argue this is an er

safetyarxiv-cs-lg
23 Jun 2026
Safety

OFMU: Optimization-Driven Framework for Machine Unlearning

DGX agent

arXiv:2509.22483v2 Announce Type: replace Abstract: Large language models deployed in sensitive applications increasingly require the ability to unlearn specific knowledge, such as user requests, copy

safetyarxiv-cs-lg
23 Jun 2026
Safety

On the Limits of Prompt-Conditioned Language Models as General-Purpose Learners

DGX agent

arXiv:2606.23668v1 Announce Type: new Abstract: Large Language Models (LLMs) are frequently portrayed as general-purpose solvers capable of solving arbitrary tasks. We argue that this view overlooks a

safetyarxiv-cs-lg
23 Jun 2026
Safety

Platooning Connected, Autonomous, and Human-Driven Vehicles: A Deep Reinforcement Learning-based Approach

DGX agent

arXiv:2606.20648v1 Announce Type: cross Abstract: Conventionally, existing vehicle platooning approaches are designed for connected vehicles, typically including connected autonomous vehicles and conn

safetyarxiv-cs-lg
23 Jun 2026
Safety

Real-Time Multimodal Activity-Aware Error Detection in Robot-Assisted Surgery

DGX agent

arXiv:2606.23593v1 Announce Type: cross Abstract: Robot-assisted minimally invasive surgery improves surgical precision but introduces complexity, making technical error detection essential for ensuri

safetyarxiv-cs-cv
23 Jun 2026
Model Releases

SingGuard: A Policy-Adaptive Multimodal LLM Guardrail with Dynamic Reasoning

DGX agent

arXiv:2606.22873v1 Announce Type: new Abstract: Vision-language models (VLMs) are increasingly deployed in consumer, medical, financial, and enterprise applications. This broad deployment expands the

model-releasesarxiv-cs-cv
23 Jun 2026
Safety

TIP-Search: Time-Predictable Inference Scheduling for Market Prediction under Uncertain Load

DGX agent

arXiv:2506.08026v3 Announce Type: replace-cross Abstract: Real-time market prediction services need correct predictions before a decision deadline; a correct prediction delivered late is not a usable

safetyarxiv-cs-lg
23 Jun 2026
Safety

What if? Emulative Simulation with World Models for Situated Reasoning

DGX agent

arXiv:2603.06445v2 Announce Type: replace Abstract: Situated reasoning often relies on active exploration, yet in many real-world scenarios such exploration is infeasible due to physical constraints o

safetyarxiv-cs-cv
23 Jun 2026
Safety

When Confidence Lacks Concepts: Interpretable OOD Detection via Representation Perturbations

DGX agent

arXiv:2606.16196v2 Announce Type: replace-cross Abstract: Deep neural networks have achieved remarkable performance across medical imaging tasks, yet their tendency to overgeneralize under distributio

safetyarxiv-cs-cv
23 Jun 2026
Safety

3D-CBM: A Framework for Concept-Based Interpretability in Generative 3D Modeling

DGX agent

arXiv:2606.11446v1 Announce Type: new Abstract: This research introduces a framework for incorporating Concept Bottleneck Models (CBMs) into 3D generative architectures to address the inherent 'semant

safetyarxiv-cs-cv
11 Jun 2026
Safety

AI Researchers Must Help Lead Arms Control to Mitigate Military AI Risks

DGX agent

arXiv:2606.11533v1 Announce Type: cross Abstract: The advancement of AI capabilities compels researchers and the public to be more aware of its potential worldwide impact. A pressing near-term concern

safetyarxiv-cs-ai
11 Jun 2026
Safety

Automating Geometry-Intensive Compliance Checking in BIM: Graph-Based Semantic Reasoning Framework

DGX agent

arXiv:2606.12065v1 Announce Type: new Abstract: Automating compliance check for geometry-intensive regulations remains a significant technical bottleneck in Building Information Modeling (BIM), primar

safetyarxiv-cs-ai
11 Jun 2026
Safety

ConsistencyPlanner: Real-time Planning with Fast-Sampling Consistency Models

DGX agent

arXiv:2606.11569v1 Announce Type: cross Abstract: Closed-loop planning in complex, real-world driving scenarios presents a critical challenge for autonomous driving systems. While traditional rule-bas

safetyarxiv-cs-ai
11 Jun 2026
Safety

Designing AI-Supported Focus Groups: A Role x Modality Playbook

DGX agent

arXiv:2606.11835v1 Announce Type: cross Abstract: Collecting participants' lived experiences is central to design research. Focus groups are uniquely valuable because participants not only share indiv

safetyarxiv-cs-ai
11 Jun 2026
Safety

EKF-Based Depth Camera and Deep Learning Fusion for UAV-Person Distance Estimation and Following in SAR Operations

DGX agent

arXiv:2602.20958v2 Announce Type: replace-cross Abstract: Vision-based Unmanned Aerial Vehicles (UAVs) frameworks aid human search tasks by detecting and recognizing specific individuals, then trackin

safetyarxiv-cs-ai
11 Jun 2026
Safety

From Consumption to Reflection: Designing Human-AI Relations for Stable Reasoning

DGX agent

arXiv:2606.11195v1 Announce Type: cross Abstract: Large language models (LLMs) have transformed how humans access information, but not how we reason with it. Their fluency accelerates consumption whil

safetyarxiv-cs-ai
11 Jun 2026
Safety

Intermittent time series forecasting: local vs global models

DGX agent

arXiv:2601.14031v2 Announce Type: replace-cross Abstract: Forecasting intermittent time series, which contain zeros, is a crucial challenge in supply chains as inventory policies require probabilistic

safetyarxiv-cs-lg
11 Jun 2026
Safety

Mahalanobis-Guided Latent OOD Detection for Hybrid ES-DRL Control in Time-Varying Systems

DGX agent

arXiv:2606.11474v1 Announce Type: new Abstract: In this paper, we study Mahalanobis-guided latent out-of-distribution (OOD) detection for test-time RL controller switching in nonlinear time-varying sy

safetyarxiv-cs-lg
11 Jun 2026
Safety

Performance Analysis of YOLOv11 and YOLOv8 for Mixed Traffic Object Detection under Adverse Weather Conditions in Developing Countries

DGX agent

arXiv:2606.12066v1 Announce Type: new Abstract: In modern vehicular systems, robust performance under harsh conditions has become a critical problem of autonomous driving. Our study delivers a compreh

safetyarxiv-cs-cv
11 Jun 2026
Safety

Risk Under Pressure: Compute-Aware Evaluation of Adversarial Robustness in Language Models

DGX agent

arXiv:2606.11409v1 Announce Type: cross Abstract: Adversarial robustness evaluations of large language models (LLMs) typically report attack success rate (ASR) under fixed query budgets, implicitly tr

safetyarxiv-cs-ai
11 Jun 2026
Safety

Scenario-based Probing and Steering Cultural Values in Large Language Models--Extended Version

DGX agent

arXiv:2606.11399v1 Announce Type: new Abstract: Large Language Models (LLMs) are deployed across cultural contexts but often reflect homogenized values inherited from training data. Evaluations of cul

safetyarxiv-cs-cl
11 Jun 2026
Safety

Semantically-Aware Diver Activity Recognition Framework for Effective Underwater Multi-Human-Robot Collaboration

DGX agent

arXiv:2606.12374v1 Announce Type: cross Abstract: Effective multi-human-robot collaboration is essential for expanding human-led operations in the challenging and high-risk underwater environment. For

safetyarxiv-cs-cv
11 Jun 2026
Safety

The Algorithm Is Not the Behavior: Learned Priors Override Look-Ahead in a Chess-Playing Neural Network

DGX agent

arXiv:2508.21380v3 Announce Type: replace-cross Abstract: Recent mechanistic work has uncovered learned algorithms within neural networks, from modular arithmetic to search and planning in game-playin

safetyarxiv-cs-ai
11 Jun 2026
Safety

Traceable Virtual Sea Trials in the Marine Robotics Unity Simulator for Manoeuvring Assessment of Unmanned Surface Vehicles

DGX agent

arXiv:2606.12349v1 Announce Type: new Abstract: Accurate identification of hydrodynamic derivatives is essential for control and navigation of Unmanned Surface Vehicles (USVs), but high-fidelity manoe

safetyarxiv-cs-ro
11 Jun 2026
Safety

Act on What You See: Unlocking Safe Social Navigation in Vision-Language-Action Models

DGX agent

arXiv:2606.10495v1 Announce Type: new Abstract: Safe social navigation requires robots to distinguish people from ordinary obstacles and to react before danger becomes imminent. We show that pretraine

safetyarxiv-cs-ro
10 Jun 2026
Safety

Conformal Prediction for Neural Operators: Distribution-Free Uncertainty Quantification in Physics Simulation

DGX agent

arXiv:2606.09923v1 Announce Type: cross Abstract: Neural operators such as the Fourier Neural Operator (FNO) have emerged as powerful surrogates for solving partial differential equations (PDEs), achi

safetyarxiv-cs-ai
10 Jun 2026
Safety

Diffusion Forcing Planner: History-Annealed Planning with Time-Dependent Guidance for Autonomous Driving

DGX agent

arXiv:2606.11019v1 Announce Type: cross Abstract: Learning-based motion planners, despite recent progress, often suffer from temporal inconsistency. Small perturbations across frames can accumulate in

safetyarxiv-cs-ai
10 Jun 2026
Safety

Does Reasoning Preserve Alignment? On the Trustworthiness of Large Reasoning Models

DGX agent

arXiv:2606.11046v1 Announce Type: new Abstract: Instruction-tuned LLMs are increasingly converted into reasoning models through post-training to improve multi-step task performance. This conversion is

safetyarxiv-cs-cl
10 Jun 2026
Safety

Exploring Accurate and Transparent Domain Adaptation in Predictive Healthcare via Concept-Grounded Orthogonal Inference

DGX agent

arXiv:2602.12542v2 Announce Type: replace-cross Abstract: Deep learning models for clinical event prediction on electronic health records (EHR) often suffer performance degradation when deployed under

safetyarxiv-cs-ai
10 Jun 2026
Safety

Gradient-Guided Reward Optimization for Inference-time Alignment

DGX agent

arXiv:2606.09635v1 Announce Type: cross Abstract: Ensuring the reliability of Large Language Models (LLMs) under distribution drift requires inference-time adaptation. While inference-time alignment m

safetyarxiv-cs-lg
10 Jun 2026
Safety

Human-AI Coordination Zones: A Framework for Designing Human-in-the-Loop Experiences with Agentic AI

DGX agent

arXiv:2606.09848v1 Announce Type: cross Abstract: As generative and agentic AI becomes embedded in everyday products, practitioners face a persistent challenge: how to design human-AI coordination --

safetyarxiv-cs-ai
10 Jun 2026
Safety

IMPACT: Learning Internal-Model Predictive Control for Forceful Robotic Manipulation

DGX agent

arXiv:2606.10818v1 Announce Type: cross Abstract: Real-world robotic manipulation tasks often involve forceful interactions with the environment, such as using tools of varying weights, transporting o

safetyarxiv-cs-cv
10 Jun 2026
Safety

MARCH: Model-Assisted Reinforcement Learning for the Perceptive Control of Humanoids over Sparse Footholds

DGX agent

arXiv:2606.10288v1 Announce Type: new Abstract: Perceptive bipedal locomotion over sparse terrain remains a difficult challenge: model-based methods are precise but brittle to uncertainty, while model

safetyarxiv-cs-ro
10 Jun 2026
Safety

Mechanistic Analysis of Alignment Algorithms in Language Models

DGX agent

arXiv:2606.09850v1 Announce Type: cross Abstract: Post-training alignment algorithms are predominantly evaluated as black boxes, obscuring how they reshape language models' internal computations. We p

safetyarxiv-cs-cl
10 Jun 2026
Safety

ParaBridge: Bridging Paralinguistic Perception and Dialogue Behavior in Speech Language Models

DGX agent

arXiv:2606.10581v1 Announce Type: new Abstract: Speech carries more information than just words: a child's voice, a fearful tone, or a noisy background should all lead a sufficiently competent spoken-

safetyarxiv-cs-cl
10 Jun 2026
Model Releases

The Interlocutor Effect: Why LLMs Leak More Personal Data to Agents Than Humans

DGX agent

arXiv:2606.09844v1 Announce Type: cross Abstract: Large Language Models (LLMs) alter their privacy behavior based on the perceived identity of their interlocutor. While safety mechanisms typically pre

model-releasesarxiv-cs-ai
10 Jun 2026
Safety

Uncovering Vulnerability of Vision-Language-Action Models under Joint-Level Physical Faults

DGX agent

arXiv:2606.10501v1 Announce Type: new Abstract: Deploying Vision-Language-Action (VLA) models in real robotic systems requires robustness not only to semantic and perceptual variations, but also to em

safetyarxiv-cs-ro
10 Jun 2026
Safety

When the Chain of Thought Knows Better: Failure Modes in Multi-Turn Reasoning Models

DGX agent

arXiv:2606.10740v1 Announce Type: new Abstract: Failures in multi-turn reasoning models are largely invisible to terminal-score evaluation. A model can lock onto an unsafe stance early in a long dialo

safetyarxiv-cs-ai
10 Jun 2026
Safety

6G Empowering Future Robotics: A Vision for Next-Generation Autonomous Systems

DGX agent

arXiv:2602.12246v2 Announce Type: replace-cross Abstract: The convergence of robotics and next-generation communication is a critical driver of technological advancement. As the world transitions from

safetyarxiv-cs-ro
9 Jun 2026
Safety

A practical probabilistic framework for deformable image registration uncertainty in radiotherapy dose propagation

DGX agent

arXiv:2606.09253v1 Announce Type: new Abstract: Deformable image registration (DIR) is widely used in radiotherapy for dose propagation and accumulation, but uncertainty in the underlying deformation

safetyarxiv-cs-cv
9 Jun 2026
Safety

AeroSpectra Sentinel: An Auditable LLM Prompt-Chaining Decision-Support Workflow for Acute Asthma Risk Assessment from Respiratory Sounds and Clinical Signals

DGX agent

arXiv:2606.08247v1 Announce Type: cross Abstract: Acute asthma risk assessment requires rapid interpretation of respiratory sounds, oxygenation, airflow limitation, speech ability, work of breathing,

safetyarxiv-cs-ai
9 Jun 2026
Safety

Baichuan-M4: A Clinical-Grade Medical Agent System for Continuous Care

DGX agent

arXiv:2606.08982v1 Announce Type: new Abstract: Baichuan-M4 is Baichuan Intelligence's clinical-grade medical large model, designed for continuous care rather than single-turn medical question answeri

safetyarxiv-cs-ai
9 Jun 2026
Safety

Beyond Accuracy: Interpreting Topic Representation in Suicide Ideation Detection Models

DGX agent

arXiv:2606.07714v1 Announce Type: cross Abstract: Suicide ideation detection models are typically evaluated using aggregate performance metrics, yet little is known about how they internally represent

safetyarxiv-cs-ai
9 Jun 2026
Safety

Beyond Linear Activation Steering: Invertible Latent Transformations for Controlling LLM Behavior

DGX agent

arXiv:2606.08454v1 Announce Type: new Abstract: Activation steering provides a lightweight inference-time mechanism for controlling large language models (LLMs) by modifying their internal activation

safetyarxiv-cs-lg
9 Jun 2026
Safety

Distilling LLM Reasoning into an Interpretable Policy Tree for Human-AI Collaboration

DGX agent

arXiv:2606.08596v1 Announce Type: new Abstract: Constructing efficient and reliable policies to assist humans is indispensable for human-AI collaboration. Existing methods mainly follow two lines of w

safetyarxiv-cs-ai
9 Jun 2026
Model Releases

Distilling Safe LLM Systems via Soft Prompts for On Device Settings

DGX agent

arXiv:2606.09388v1 Announce Type: new Abstract: Deploying safe large language models (LLMs) on resource-constrained edge devices presents a critical challenge: while dual-model systems combining LLMs

model-releasesarxiv-cs-lg
9 Jun 2026
Safety

Learning to Attack and Defend: Adaptive Red Teaming of Language Models via GRPO

DGX agent

arXiv:2606.09701v1 Announce Type: cross Abstract: AI red teaming must continually adapt to evolving attackers and defenders. Reinforcement learning offers a promising approach to discovering novel att

safetyarxiv-cs-ai
9 Jun 2026
← Previous
1…4445464748…257
Next →