AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,356 results
Safety

Helping build shared standards for advanced AI

DGX agent

OpenAI discusses its efforts to contribute to the development of shared industry standards and best practices for advanced artificial intelligence systems. The article likely covers OpenAI's involveme

safetyopenai
23 Jun 2026
Safety

HoloAgent-0: A Unified Embodied Agent Framework with 3D Spatial Memory

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2606.23565v1 Announce Type: cross Abstract: LLM agents follow a practical execution loop in digital environments: they reason over structured states, invoke tools, inspect feedback, and revise a

safetyarxiv-cs-cv
23 Jun 2026
Safety

Iterative Convex Optimization with Control Barrier Functions for Obstacle Avoidance among Polytopes

DGX agent

arXiv:2603.05916v2 Announce Type: replace Abstract: Obstacle avoidance of polytopic obstacles by polytopic robots is a challenging problem in optimization-based control and trajectory planning. Many e

safetyarxiv-cs-ro
23 Jun 2026
Safety

LP-NavOA: Integrated Local Navigation and Obstacle Avoidance for Humanoid Robots under Limited Perception

DGX agent

arXiv:2606.23249v1 Announce Type: new Abstract: Humanoid local navigation in cluttered environments must jointly resolve obstacle avoidance, sparse-goal recovery, and stable whole-body locomotion unde

safetyarxiv-cs-ro
23 Jun 2026
Safety

Noise is Signal: Density-Based Outliers as Leading Indicators of Occupational Emergence in Labor Market Text

DGX agent

arXiv:2606.22769v1 Announce Type: new Abstract: Standard NLP pipelines for occupational clustering discard the 10-15% of job postings that density-based methods assign to noise. We argue this is an er

safetyarxiv-cs-lg
23 Jun 2026
Safety

OFMU: Optimization-Driven Framework for Machine Unlearning

DGX agent

arXiv:2509.22483v2 Announce Type: replace Abstract: Large language models deployed in sensitive applications increasingly require the ability to unlearn specific knowledge, such as user requests, copy

safetyarxiv-cs-lg
23 Jun 2026
Safety

On the Limits of Prompt-Conditioned Language Models as General-Purpose Learners

DGX agent

arXiv:2606.23668v1 Announce Type: new Abstract: Large Language Models (LLMs) are frequently portrayed as general-purpose solvers capable of solving arbitrary tasks. We argue that this view overlooks a

safetyarxiv-cs-lg
23 Jun 2026
Safety

Platooning Connected, Autonomous, and Human-Driven Vehicles: A Deep Reinforcement Learning-based Approach

DGX agent

arXiv:2606.20648v1 Announce Type: cross Abstract: Conventionally, existing vehicle platooning approaches are designed for connected vehicles, typically including connected autonomous vehicles and conn

safetyarxiv-cs-lg
23 Jun 2026
Safety

Real-Time Multimodal Activity-Aware Error Detection in Robot-Assisted Surgery

DGX agent

arXiv:2606.23593v1 Announce Type: cross Abstract: Robot-assisted minimally invasive surgery improves surgical precision but introduces complexity, making technical error detection essential for ensuri

safetyarxiv-cs-cv
23 Jun 2026
Model Releases

SingGuard: A Policy-Adaptive Multimodal LLM Guardrail with Dynamic Reasoning

DGX agent

arXiv:2606.22873v1 Announce Type: new Abstract: Vision-language models (VLMs) are increasingly deployed in consumer, medical, financial, and enterprise applications. This broad deployment expands the

model-releasesarxiv-cs-cv
23 Jun 2026
Safety

Sources: the Trump administration is pressing Meta to submit its AI models for voluntary review; Meta is the only major US AI developer without an agreement (New York Times)

DGX agent

New York Times: Sources: the Trump administration is pressing Meta to submit its AI models for voluntary review; Meta is the only major US AI developer without an agreement — Federal officials are urg

safetytechmeme
23 Jun 2026
Safety

TIP-Search: Time-Predictable Inference Scheduling for Market Prediction under Uncertain Load

DGX agent

arXiv:2506.08026v3 Announce Type: replace-cross Abstract: Real-time market prediction services need correct predictions before a decision deadline; a correct prediction delivered late is not a usable

safetyarxiv-cs-lg
23 Jun 2026
Safety

What if? Emulative Simulation with World Models for Situated Reasoning

DGX agent

arXiv:2603.06445v2 Announce Type: replace Abstract: Situated reasoning often relies on active exploration, yet in many real-world scenarios such exploration is infeasible due to physical constraints o

safetyarxiv-cs-cv
23 Jun 2026
Safety

When Confidence Lacks Concepts: Interpretable OOD Detection via Representation Perturbations

DGX agent

arXiv:2606.16196v2 Announce Type: replace-cross Abstract: Deep neural networks have achieved remarkable performance across medical imaging tasks, yet their tendency to overgeneralize under distributio

safetyarxiv-cs-cv
23 Jun 2026
Safety

this is actually buddhism btw

DGX agent

This post likely discusses how certain contemporary concepts, practices, or ideologies are actually rooted in or aligned with Buddhist philosophy, possibly challenging common misconceptions about thei

safetyconnor-leahy--x
21 Jun 2026
Safety

It's crazy how badly the tech industry wants this guy @AlexBores defeated, all because he passed a relatively tame AI transparency law

DGX agent

It's crazy how badly the tech industry wants this guy @AlexBores defeated, all because he passed a relatively tame AI transparency law AI billionaires have spent $8,027,347 trying to beat one New York

safetygary-marcus--x
20 Jun 2026
Safety

3D-CBM: A Framework for Concept-Based Interpretability in Generative 3D Modeling

DGX agent

arXiv:2606.11446v1 Announce Type: new Abstract: This research introduces a framework for incorporating Concept Bottleneck Models (CBMs) into 3D generative architectures to address the inherent 'semant

safetyarxiv-cs-cv
11 Jun 2026
Safety

AI Researchers Must Help Lead Arms Control to Mitigate Military AI Risks

DGX agent

arXiv:2606.11533v1 Announce Type: cross Abstract: The advancement of AI capabilities compels researchers and the public to be more aware of its potential worldwide impact. A pressing near-term concern

safetyarxiv-cs-ai
11 Jun 2026
Safety

Automating Geometry-Intensive Compliance Checking in BIM: Graph-Based Semantic Reasoning Framework

DGX agent

arXiv:2606.12065v1 Announce Type: new Abstract: Automating compliance check for geometry-intensive regulations remains a significant technical bottleneck in Building Information Modeling (BIM), primar

safetyarxiv-cs-ai
11 Jun 2026
Safety

ConsistencyPlanner: Real-time Planning with Fast-Sampling Consistency Models

DGX agent

arXiv:2606.11569v1 Announce Type: cross Abstract: Closed-loop planning in complex, real-world driving scenarios presents a critical challenge for autonomous driving systems. While traditional rule-bas

safetyarxiv-cs-ai
11 Jun 2026
Safety

Designing AI-Supported Focus Groups: A Role x Modality Playbook

DGX agent

arXiv:2606.11835v1 Announce Type: cross Abstract: Collecting participants' lived experiences is central to design research. Focus groups are uniquely valuable because participants not only share indiv

safetyarxiv-cs-ai
11 Jun 2026
Safety

EKF-Based Depth Camera and Deep Learning Fusion for UAV-Person Distance Estimation and Following in SAR Operations

DGX agent

arXiv:2602.20958v2 Announce Type: replace-cross Abstract: Vision-based Unmanned Aerial Vehicles (UAVs) frameworks aid human search tasks by detecting and recognizing specific individuals, then trackin

safetyarxiv-cs-ai
11 Jun 2026
Safety

From Consumption to Reflection: Designing Human-AI Relations for Stable Reasoning

DGX agent

arXiv:2606.11195v1 Announce Type: cross Abstract: Large language models (LLMs) have transformed how humans access information, but not how we reason with it. Their fluency accelerates consumption whil

safetyarxiv-cs-ai
11 Jun 2026
Safety

Intermittent time series forecasting: local vs global models

DGX agent

arXiv:2601.14031v2 Announce Type: replace-cross Abstract: Forecasting intermittent time series, which contain zeros, is a crucial challenge in supply chains as inventory policies require probabilistic

safetyarxiv-cs-lg
11 Jun 2026
Safety

Mahalanobis-Guided Latent OOD Detection for Hybrid ES-DRL Control in Time-Varying Systems

DGX agent

arXiv:2606.11474v1 Announce Type: new Abstract: In this paper, we study Mahalanobis-guided latent out-of-distribution (OOD) detection for test-time RL controller switching in nonlinear time-varying sy

safetyarxiv-cs-lg
11 Jun 2026
Safety

Performance Analysis of YOLOv11 and YOLOv8 for Mixed Traffic Object Detection under Adverse Weather Conditions in Developing Countries

DGX agent

arXiv:2606.12066v1 Announce Type: new Abstract: In modern vehicular systems, robust performance under harsh conditions has become a critical problem of autonomous driving. Our study delivers a compreh

safetyarxiv-cs-cv
11 Jun 2026
Safety

Risk Under Pressure: Compute-Aware Evaluation of Adversarial Robustness in Language Models

DGX agent

arXiv:2606.11409v1 Announce Type: cross Abstract: Adversarial robustness evaluations of large language models (LLMs) typically report attack success rate (ASR) under fixed query budgets, implicitly tr

safetyarxiv-cs-ai
11 Jun 2026
Safety

Scenario-based Probing and Steering Cultural Values in Large Language Models--Extended Version

DGX agent

arXiv:2606.11399v1 Announce Type: new Abstract: Large Language Models (LLMs) are deployed across cultural contexts but often reflect homogenized values inherited from training data. Evaluations of cul

safetyarxiv-cs-cl
11 Jun 2026
Safety

Semantically-Aware Diver Activity Recognition Framework for Effective Underwater Multi-Human-Robot Collaboration

DGX agent

arXiv:2606.12374v1 Announce Type: cross Abstract: Effective multi-human-robot collaboration is essential for expanding human-led operations in the challenging and high-risk underwater environment. For

safetyarxiv-cs-cv
11 Jun 2026
Safety

The Algorithm Is Not the Behavior: Learned Priors Override Look-Ahead in a Chess-Playing Neural Network

DGX agent

arXiv:2508.21380v3 Announce Type: replace-cross Abstract: Recent mechanistic work has uncovered learned algorithms within neural networks, from modular arithmetic to search and planning in game-playin

safetyarxiv-cs-ai
11 Jun 2026
Safety

Traceable Virtual Sea Trials in the Marine Robotics Unity Simulator for Manoeuvring Assessment of Unmanned Surface Vehicles

DGX agent

arXiv:2606.12349v1 Announce Type: new Abstract: Accurate identification of hydrodynamic derivatives is essential for control and navigation of Unmanned Surface Vehicles (USVs), but high-fidelity manoe

safetyarxiv-cs-ro
11 Jun 2026
Safety

Act on What You See: Unlocking Safe Social Navigation in Vision-Language-Action Models

DGX agent

arXiv:2606.10495v1 Announce Type: new Abstract: Safe social navigation requires robots to distinguish people from ordinary obstacles and to react before danger becomes imminent. We show that pretraine

safetyarxiv-cs-ro
10 Jun 2026
Safety

Anthropic has long advocated for transparency requirements for frontier AI, because the risks weren't yet clear enough to regulate precisely…

DGX agent

Anthropic's leadership, through CEO Dario Amodei, has advocated for transparency requirements in frontier AI development as a regulatory approach, arguing that the risks were insufficiently understood

safetydario-amodei--x
10 Jun 2026
Safety

Conformal Prediction for Neural Operators: Distribution-Free Uncertainty Quantification in Physics Simulation

DGX agent

arXiv:2606.09923v1 Announce Type: cross Abstract: Neural operators such as the Fourier Neural Operator (FNO) have emerged as powerful surrogates for solving partial differential equations (PDEs), achi

safetyarxiv-cs-ai
10 Jun 2026
Safety

Diffusion Forcing Planner: History-Annealed Planning with Time-Dependent Guidance for Autonomous Driving

DGX agent

arXiv:2606.11019v1 Announce Type: cross Abstract: Learning-based motion planners, despite recent progress, often suffer from temporal inconsistency. Small perturbations across frames can accumulate in

safetyarxiv-cs-ai
10 Jun 2026
Safety

Does Reasoning Preserve Alignment? On the Trustworthiness of Large Reasoning Models

DGX agent

arXiv:2606.11046v1 Announce Type: new Abstract: Instruction-tuned LLMs are increasingly converted into reasoning models through post-training to improve multi-step task performance. This conversion is

safetyarxiv-cs-cl
10 Jun 2026
Safety

Exploring Accurate and Transparent Domain Adaptation in Predictive Healthcare via Concept-Grounded Orthogonal Inference

DGX agent

arXiv:2602.12542v2 Announce Type: replace-cross Abstract: Deep learning models for clinical event prediction on electronic health records (EHR) often suffer performance degradation when deployed under

safetyarxiv-cs-ai
10 Jun 2026
Safety

Gradient-Guided Reward Optimization for Inference-time Alignment

DGX agent

arXiv:2606.09635v1 Announce Type: cross Abstract: Ensuring the reliability of Large Language Models (LLMs) under distribution drift requires inference-time adaptation. While inference-time alignment m

safetyarxiv-cs-lg
10 Jun 2026
Safety

Human-AI Coordination Zones: A Framework for Designing Human-in-the-Loop Experiences with Agentic AI

DGX agent

arXiv:2606.09848v1 Announce Type: cross Abstract: As generative and agentic AI becomes embedded in everyday products, practitioners face a persistent challenge: how to design human-AI coordination --

safetyarxiv-cs-ai
10 Jun 2026
Safety

IMPACT: Learning Internal-Model Predictive Control for Forceful Robotic Manipulation

DGX agent

arXiv:2606.10818v1 Announce Type: cross Abstract: Real-world robotic manipulation tasks often involve forceful interactions with the environment, such as using tools of varying weights, transporting o

safetyarxiv-cs-cv
10 Jun 2026
Safety

MARCH: Model-Assisted Reinforcement Learning for the Perceptive Control of Humanoids over Sparse Footholds

DGX agent

arXiv:2606.10288v1 Announce Type: new Abstract: Perceptive bipedal locomotion over sparse terrain remains a difficult challenge: model-based methods are precise but brittle to uncertainty, while model

safetyarxiv-cs-ro
10 Jun 2026
Safety

Mechanistic Analysis of Alignment Algorithms in Language Models

DGX agent

arXiv:2606.09850v1 Announce Type: cross Abstract: Post-training alignment algorithms are predominantly evaluated as black boxes, obscuring how they reshape language models' internal computations. We p

safetyarxiv-cs-cl
10 Jun 2026
Safety

ParaBridge: Bridging Paralinguistic Perception and Dialogue Behavior in Speech Language Models

DGX agent

arXiv:2606.10581v1 Announce Type: new Abstract: Speech carries more information than just words: a child's voice, a fearful tone, or a noisy background should all lead a sufficiently competent spoken-

safetyarxiv-cs-cl
10 Jun 2026
Model Releases

The Interlocutor Effect: Why LLMs Leak More Personal Data to Agents Than Humans

DGX agent

arXiv:2606.09844v1 Announce Type: cross Abstract: Large Language Models (LLMs) alter their privacy behavior based on the perceived identity of their interlocutor. While safety mechanisms typically pre

model-releasesarxiv-cs-ai
10 Jun 2026
Safety

Uncovering Vulnerability of Vision-Language-Action Models under Joint-Level Physical Faults

DGX agent

arXiv:2606.10501v1 Announce Type: new Abstract: Deploying Vision-Language-Action (VLA) models in real robotic systems requires robustness not only to semantic and perceptual variations, but also to em

safetyarxiv-cs-ro
10 Jun 2026
Safety

When the Chain of Thought Knows Better: Failure Modes in Multi-Turn Reasoning Models

DGX agent

arXiv:2606.10740v1 Announce Type: new Abstract: Failures in multi-turn reasoning models are largely invisible to terminal-score evaluation. A model can lock onto an unsafe stance early in a long dialo

safetyarxiv-cs-ai
10 Jun 2026
Safety

6G Empowering Future Robotics: A Vision for Next-Generation Autonomous Systems

DGX agent

arXiv:2602.12246v2 Announce Type: replace-cross Abstract: The convergence of robotics and next-generation communication is a critical driver of technological advancement. As the world transitions from

safetyarxiv-cs-ro
9 Jun 2026
Safety

A practical probabilistic framework for deformable image registration uncertainty in radiotherapy dose propagation

DGX agent

arXiv:2606.09253v1 Announce Type: new Abstract: Deformable image registration (DIR) is widely used in radiotherapy for dose propagation and accumulation, but uncertainty in the underlying deformation

safetyarxiv-cs-cv
9 Jun 2026
← Previous
1…4950515253…300
Next →