AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,356 results
23 Jun 2026

Helping build shared standards for advanced AI

SafetyDGX agent

OpenAI discusses its efforts to contribute to the development of shared industry standards and best practices for advanced artificial intelligence systems. The article likely covers OpenAI's involveme

HoloAgent-0: A Unified Embodied Agent Framework with 3D Spatial Memory

SafetyDGX agent

arXiv:2606.23565v1 Announce Type: cross Abstract: LLM agents follow a practical execution loop in digital environments: they reason over structured states, invoke tools, inspect feedback, and revise a

Iterative Convex Optimization with Control Barrier Functions for Obstacle Avoidance among Polytopes

SafetyDGX agent

arXiv:2603.05916v2 Announce Type: replace Abstract: Obstacle avoidance of polytopic obstacles by polytopic robots is a challenging problem in optimization-based control and trajectory planning. Many e

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

LP-NavOA: Integrated Local Navigation and Obstacle Avoidance for Humanoid Robots under Limited Perception

SafetyDGX agent

arXiv:2606.23249v1 Announce Type: new Abstract: Humanoid local navigation in cluttered environments must jointly resolve obstacle avoidance, sparse-goal recovery, and stable whole-body locomotion unde

Noise is Signal: Density-Based Outliers as Leading Indicators of Occupational Emergence in Labor Market Text

SafetyDGX agent

arXiv:2606.22769v1 Announce Type: new Abstract: Standard NLP pipelines for occupational clustering discard the 10-15% of job postings that density-based methods assign to noise. We argue this is an er

OFMU: Optimization-Driven Framework for Machine Unlearning

SafetyDGX agent

arXiv:2509.22483v2 Announce Type: replace Abstract: Large language models deployed in sensitive applications increasingly require the ability to unlearn specific knowledge, such as user requests, copy

On the Limits of Prompt-Conditioned Language Models as General-Purpose Learners

SafetyDGX agent

arXiv:2606.23668v1 Announce Type: new Abstract: Large Language Models (LLMs) are frequently portrayed as general-purpose solvers capable of solving arbitrary tasks. We argue that this view overlooks a

Platooning Connected, Autonomous, and Human-Driven Vehicles: A Deep Reinforcement Learning-based Approach

SafetyDGX agent

arXiv:2606.20648v1 Announce Type: cross Abstract: Conventionally, existing vehicle platooning approaches are designed for connected vehicles, typically including connected autonomous vehicles and conn

Real-Time Multimodal Activity-Aware Error Detection in Robot-Assisted Surgery

SafetyDGX agent

arXiv:2606.23593v1 Announce Type: cross Abstract: Robot-assisted minimally invasive surgery improves surgical precision but introduces complexity, making technical error detection essential for ensuri

SingGuard: A Policy-Adaptive Multimodal LLM Guardrail with Dynamic Reasoning

Model ReleasesDGX agent

arXiv:2606.22873v1 Announce Type: new Abstract: Vision-language models (VLMs) are increasingly deployed in consumer, medical, financial, and enterprise applications. This broad deployment expands the

Sources: the Trump administration is pressing Meta to submit its AI models for voluntary review; Meta is the only major US AI developer without an agreement (New York Times)

SafetyDGX agent

New York Times: Sources: the Trump administration is pressing Meta to submit its AI models for voluntary review; Meta is the only major US AI developer without an agreement — Federal officials are urg

TIP-Search: Time-Predictable Inference Scheduling for Market Prediction under Uncertain Load

SafetyDGX agent

arXiv:2506.08026v3 Announce Type: replace-cross Abstract: Real-time market prediction services need correct predictions before a decision deadline; a correct prediction delivered late is not a usable

What if? Emulative Simulation with World Models for Situated Reasoning

SafetyDGX agent

arXiv:2603.06445v2 Announce Type: replace Abstract: Situated reasoning often relies on active exploration, yet in many real-world scenarios such exploration is infeasible due to physical constraints o

When Confidence Lacks Concepts: Interpretable OOD Detection via Representation Perturbations

SafetyDGX agent

arXiv:2606.16196v2 Announce Type: replace-cross Abstract: Deep neural networks have achieved remarkable performance across medical imaging tasks, yet their tendency to overgeneralize under distributio

21 Jun 2026

this is actually buddhism btw

SafetyDGX agent

This post likely discusses how certain contemporary concepts, practices, or ideologies are actually rooted in or aligned with Buddhist philosophy, possibly challenging common misconceptions about thei

20 Jun 2026

It's crazy how badly the tech industry wants this guy @AlexBores defeated, all because he passed a relatively tame AI transparency law

SafetyDGX agent

It's crazy how badly the tech industry wants this guy @AlexBores defeated, all because he passed a relatively tame AI transparency law AI billionaires have spent $8,027,347 trying to beat one New York

11 Jun 2026

3D-CBM: A Framework for Concept-Based Interpretability in Generative 3D Modeling

SafetyDGX agent

arXiv:2606.11446v1 Announce Type: new Abstract: This research introduces a framework for incorporating Concept Bottleneck Models (CBMs) into 3D generative architectures to address the inherent 'semant

AI Researchers Must Help Lead Arms Control to Mitigate Military AI Risks

SafetyDGX agent

arXiv:2606.11533v1 Announce Type: cross Abstract: The advancement of AI capabilities compels researchers and the public to be more aware of its potential worldwide impact. A pressing near-term concern

Automating Geometry-Intensive Compliance Checking in BIM: Graph-Based Semantic Reasoning Framework

SafetyDGX agent

arXiv:2606.12065v1 Announce Type: new Abstract: Automating compliance check for geometry-intensive regulations remains a significant technical bottleneck in Building Information Modeling (BIM), primar

ConsistencyPlanner: Real-time Planning with Fast-Sampling Consistency Models

SafetyDGX agent

arXiv:2606.11569v1 Announce Type: cross Abstract: Closed-loop planning in complex, real-world driving scenarios presents a critical challenge for autonomous driving systems. While traditional rule-bas

Designing AI-Supported Focus Groups: A Role x Modality Playbook

SafetyDGX agent

arXiv:2606.11835v1 Announce Type: cross Abstract: Collecting participants' lived experiences is central to design research. Focus groups are uniquely valuable because participants not only share indiv

EKF-Based Depth Camera and Deep Learning Fusion for UAV-Person Distance Estimation and Following in SAR Operations

SafetyDGX agent

arXiv:2602.20958v2 Announce Type: replace-cross Abstract: Vision-based Unmanned Aerial Vehicles (UAVs) frameworks aid human search tasks by detecting and recognizing specific individuals, then trackin

From Consumption to Reflection: Designing Human-AI Relations for Stable Reasoning

SafetyDGX agent

arXiv:2606.11195v1 Announce Type: cross Abstract: Large language models (LLMs) have transformed how humans access information, but not how we reason with it. Their fluency accelerates consumption whil

Intermittent time series forecasting: local vs global models

SafetyDGX agent

arXiv:2601.14031v2 Announce Type: replace-cross Abstract: Forecasting intermittent time series, which contain zeros, is a crucial challenge in supply chains as inventory policies require probabilistic

Mahalanobis-Guided Latent OOD Detection for Hybrid ES-DRL Control in Time-Varying Systems

SafetyDGX agent

arXiv:2606.11474v1 Announce Type: new Abstract: In this paper, we study Mahalanobis-guided latent out-of-distribution (OOD) detection for test-time RL controller switching in nonlinear time-varying sy

Performance Analysis of YOLOv11 and YOLOv8 for Mixed Traffic Object Detection under Adverse Weather Conditions in Developing Countries

SafetyDGX agent

arXiv:2606.12066v1 Announce Type: new Abstract: In modern vehicular systems, robust performance under harsh conditions has become a critical problem of autonomous driving. Our study delivers a compreh

Risk Under Pressure: Compute-Aware Evaluation of Adversarial Robustness in Language Models

SafetyDGX agent

arXiv:2606.11409v1 Announce Type: cross Abstract: Adversarial robustness evaluations of large language models (LLMs) typically report attack success rate (ASR) under fixed query budgets, implicitly tr

Scenario-based Probing and Steering Cultural Values in Large Language Models--Extended Version

SafetyDGX agent

arXiv:2606.11399v1 Announce Type: new Abstract: Large Language Models (LLMs) are deployed across cultural contexts but often reflect homogenized values inherited from training data. Evaluations of cul

Semantically-Aware Diver Activity Recognition Framework for Effective Underwater Multi-Human-Robot Collaboration

SafetyDGX agent

arXiv:2606.12374v1 Announce Type: cross Abstract: Effective multi-human-robot collaboration is essential for expanding human-led operations in the challenging and high-risk underwater environment. For

The Algorithm Is Not the Behavior: Learned Priors Override Look-Ahead in a Chess-Playing Neural Network

SafetyDGX agent

arXiv:2508.21380v3 Announce Type: replace-cross Abstract: Recent mechanistic work has uncovered learned algorithms within neural networks, from modular arithmetic to search and planning in game-playin

Traceable Virtual Sea Trials in the Marine Robotics Unity Simulator for Manoeuvring Assessment of Unmanned Surface Vehicles

SafetyDGX agent

arXiv:2606.12349v1 Announce Type: new Abstract: Accurate identification of hydrodynamic derivatives is essential for control and navigation of Unmanned Surface Vehicles (USVs), but high-fidelity manoe

10 Jun 2026

Act on What You See: Unlocking Safe Social Navigation in Vision-Language-Action Models

SafetyDGX agent

arXiv:2606.10495v1 Announce Type: new Abstract: Safe social navigation requires robots to distinguish people from ordinary obstacles and to react before danger becomes imminent. We show that pretraine

Anthropic has long advocated for transparency requirements for frontier AI, because the risks weren't yet clear enough to regulate precisely…

SafetyDGX agent

Anthropic's leadership, through CEO Dario Amodei, has advocated for transparency requirements in frontier AI development as a regulatory approach, arguing that the risks were insufficiently understood

Conformal Prediction for Neural Operators: Distribution-Free Uncertainty Quantification in Physics Simulation

SafetyDGX agent

arXiv:2606.09923v1 Announce Type: cross Abstract: Neural operators such as the Fourier Neural Operator (FNO) have emerged as powerful surrogates for solving partial differential equations (PDEs), achi

Diffusion Forcing Planner: History-Annealed Planning with Time-Dependent Guidance for Autonomous Driving

SafetyDGX agent

arXiv:2606.11019v1 Announce Type: cross Abstract: Learning-based motion planners, despite recent progress, often suffer from temporal inconsistency. Small perturbations across frames can accumulate in

Does Reasoning Preserve Alignment? On the Trustworthiness of Large Reasoning Models

SafetyDGX agent

arXiv:2606.11046v1 Announce Type: new Abstract: Instruction-tuned LLMs are increasingly converted into reasoning models through post-training to improve multi-step task performance. This conversion is

Exploring Accurate and Transparent Domain Adaptation in Predictive Healthcare via Concept-Grounded Orthogonal Inference

SafetyDGX agent

arXiv:2602.12542v2 Announce Type: replace-cross Abstract: Deep learning models for clinical event prediction on electronic health records (EHR) often suffer performance degradation when deployed under

Gradient-Guided Reward Optimization for Inference-time Alignment

SafetyDGX agent

arXiv:2606.09635v1 Announce Type: cross Abstract: Ensuring the reliability of Large Language Models (LLMs) under distribution drift requires inference-time adaptation. While inference-time alignment m

Human-AI Coordination Zones: A Framework for Designing Human-in-the-Loop Experiences with Agentic AI

SafetyDGX agent

arXiv:2606.09848v1 Announce Type: cross Abstract: As generative and agentic AI becomes embedded in everyday products, practitioners face a persistent challenge: how to design human-AI coordination --

IMPACT: Learning Internal-Model Predictive Control for Forceful Robotic Manipulation

SafetyDGX agent

arXiv:2606.10818v1 Announce Type: cross Abstract: Real-world robotic manipulation tasks often involve forceful interactions with the environment, such as using tools of varying weights, transporting o

MARCH: Model-Assisted Reinforcement Learning for the Perceptive Control of Humanoids over Sparse Footholds

SafetyDGX agent

arXiv:2606.10288v1 Announce Type: new Abstract: Perceptive bipedal locomotion over sparse terrain remains a difficult challenge: model-based methods are precise but brittle to uncertainty, while model

Mechanistic Analysis of Alignment Algorithms in Language Models

SafetyDGX agent

arXiv:2606.09850v1 Announce Type: cross Abstract: Post-training alignment algorithms are predominantly evaluated as black boxes, obscuring how they reshape language models' internal computations. We p

ParaBridge: Bridging Paralinguistic Perception and Dialogue Behavior in Speech Language Models

SafetyDGX agent

arXiv:2606.10581v1 Announce Type: new Abstract: Speech carries more information than just words: a child's voice, a fearful tone, or a noisy background should all lead a sufficiently competent spoken-

The Interlocutor Effect: Why LLMs Leak More Personal Data to Agents Than Humans

Model ReleasesDGX agent

arXiv:2606.09844v1 Announce Type: cross Abstract: Large Language Models (LLMs) alter their privacy behavior based on the perceived identity of their interlocutor. While safety mechanisms typically pre

Uncovering Vulnerability of Vision-Language-Action Models under Joint-Level Physical Faults

SafetyDGX agent

arXiv:2606.10501v1 Announce Type: new Abstract: Deploying Vision-Language-Action (VLA) models in real robotic systems requires robustness not only to semantic and perceptual variations, but also to em

When the Chain of Thought Knows Better: Failure Modes in Multi-Turn Reasoning Models

SafetyDGX agent

arXiv:2606.10740v1 Announce Type: new Abstract: Failures in multi-turn reasoning models are largely invisible to terminal-score evaluation. A model can lock onto an unsafe stance early in a long dialo

9 Jun 2026

6G Empowering Future Robotics: A Vision for Next-Generation Autonomous Systems

SafetyDGX agent

arXiv:2602.12246v2 Announce Type: replace-cross Abstract: The convergence of robotics and next-generation communication is a critical driver of technological advancement. As the world transitions from

A practical probabilistic framework for deformable image registration uncertainty in radiotherapy dose propagation

SafetyDGX agent

arXiv:2606.09253v1 Announce Type: new Abstract: Deformable image registration (DIR) is widely used in radiotherapy for dose propagation and accumulation, but uncertainty in the underlying deformation

AeroSpectra Sentinel: An Auditable LLM Prompt-Chaining Decision-Support Workflow for Acute Asthma Risk Assessment from Respiratory Sounds and Clinical Signals

SafetyDGX agent

arXiv:2606.08247v1 Announce Type: cross Abstract: Acute asthma risk assessment requires rapid interpretation of respiratory sounds, oxygenation, airflow limitation, speech ability, work of breathing,

alienating your customers before you IPO is maybe not the best idea, @AnthropicAI

SafetyDGX agent

alienating your customers before you IPO is maybe not the best idea, @AnthropicAI Brilliant idea! Next up: Apple randomly reboots your Mac if you're building competing tech, Gmail silently edits your

Baichuan-M4: A Clinical-Grade Medical Agent System for Continuous Care

SafetyDGX agent

arXiv:2606.08982v1 Announce Type: new Abstract: Baichuan-M4 is Baichuan Intelligence's clinical-grade medical large model, designed for continuous care rather than single-turn medical question answeri

Beyond Accuracy: Interpreting Topic Representation in Suicide Ideation Detection Models

SafetyDGX agent

arXiv:2606.07714v1 Announce Type: cross Abstract: Suicide ideation detection models are typically evaluated using aggregate performance metrics, yet little is known about how they internally represent

Beyond Linear Activation Steering: Invertible Latent Transformations for Controlling LLM Behavior

SafetyDGX agent

arXiv:2606.08454v1 Announce Type: new Abstract: Activation steering provides a lightweight inference-time mechanism for controlling large language models (LLMs) by modifying their internal activation

Brilliant idea! Next up: Apple randomly reboots your Mac if you're building competing tech, Gmail silently edits your email if you mention r…

SafetyDGX agent

Brilliant idea! Next up: Apple randomly reboots your Mac if you're building competing tech, Gmail silently edits your email if you mention rival platforms, and Tesla Autopilot swerves if it detects yo

Distilling LLM Reasoning into an Interpretable Policy Tree for Human-AI Collaboration

SafetyDGX agent

arXiv:2606.08596v1 Announce Type: new Abstract: Constructing efficient and reliable policies to assist humans is indispensable for human-AI collaboration. Existing methods mainly follow two lines of w

Distilling Safe LLM Systems via Soft Prompts for On Device Settings

Model ReleasesDGX agent

arXiv:2606.09388v1 Announce Type: new Abstract: Deploying safe large language models (LLMs) on resource-constrained edge devices presents a critical challenge: while dual-model systems combining LLMs

Learning to Attack and Defend: Adaptive Red Teaming of Language Models via GRPO

SafetyDGX agent

arXiv:2606.09701v1 Announce Type: cross Abstract: AI red teaming must continually adapt to evolving attackers and defenders. Reinforcement learning offers a promising approach to discovering novel att

Memetic Capture: A Pluralistic Policy Framework for Governing AI-Driven Cultural Disempowerment

SafetyDGX agent

arXiv:2606.07802v1 Announce Type: cross Abstract: Culture is the most insidious vector of gradual human disempowerment by AI: unlike economic or political displacement, cultural displacement attacks t

Neuro-Symbolic Injection of LTLf Constraints in Autoregressive Reinforcement Learning Policies

SafetyDGX agent

arXiv:2606.08312v1 Announce Type: new Abstract: In this work we study offline reinforcement learning (RL) under temporally extended task constraints expressed in Linear Temporal Logic over finite trac

Overcoming the Regulatory Bottleneck via Agent-to-Agent Protocols: A Nuclear Case Study

SafetyDGX agent

arXiv:2606.07866v1 Announce Type: new Abstract: Regulatory review of advanced nuclear reactor designs routinely spans more than three years and consumes hundreds of millions of dollars in combined reg

← Previous
1…3940414243…240
Next →