AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,314 results
Safety

Memetic Capture: A Pluralistic Policy Framework for Governing AI-Driven Cultural Disempowerment

DGX agent

arXiv:2606.07802v1 Announce Type: cross Abstract: Culture is the most insidious vector of gradual human disempowerment by AI: unlike economic or political displacement, cultural displacement attacks t

safetyarxiv-cs-ai
9 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Neuro-Symbolic Injection of LTLf Constraints in Autoregressive Reinforcement Learning Policies

DGX agent

arXiv:2606.08312v1 Announce Type: new Abstract: In this work we study offline reinforcement learning (RL) under temporally extended task constraints expressed in Linear Temporal Logic over finite trac

safetyarxiv-cs-ai
9 Jun 2026
Safety

Overcoming the Regulatory Bottleneck via Agent-to-Agent Protocols: A Nuclear Case Study

DGX agent

arXiv:2606.07866v1 Announce Type: new Abstract: Regulatory review of advanced nuclear reactor designs routinely spans more than three years and consumes hundreds of millions of dollars in combined reg

safetyarxiv-cs-ai
9 Jun 2026
Safety

Position: Anthropomorphic Misalignment Research Needs Stronger Evidence

DGX agent

arXiv:2606.07612v1 Announce Type: cross Abstract: We argue that many Anthropomorphic Misalignment Research (AMR) studies need stronger evidence to ensure that they can provide a robust foundation for

safetyarxiv-cs-ai
9 Jun 2026
Safety

Revisiting the shutdown problem

DGX agent

arXiv:2606.08296v1 Announce Type: new Abstract: A key premise in leading arguments for existential risk from artificial intelligence is that malfunctioning artificial agents could not be easily shut d

safetyarxiv-cs-ai
9 Jun 2026
Safety

RPO-PDT: Demonstrating Role-Play-Based Knowledge Adaptation for Student Support Dialogue (Demonstration System)

DGX agent

arXiv:2606.09255v1 Announce Type: new Abstract: We present RPO-PDT: a retrieval-grounded, role-play-based dialogue system for adaptive student support in higher education. RPO-PDT is: (1) able to prov

safetyarxiv-cs-ro
9 Jun 2026
Safety

SAD-Flower: Flow Matching for Safe, Admissible, and Dynamically Consistent Planning

DGX agent

arXiv:2511.05355v3 Announce Type: replace Abstract: Flow matching (FM) has shown promising results in data-driven planning. However, it inherently lacks formal guarantees for ensuring state and action

safetyarxiv-cs-lg
9 Jun 2026
Model Releases

SafeRun: Enabling Determinism in LLM Planning for Running

DGX agent

arXiv:2606.09027v1 Announce Type: cross Abstract: Large Language Models enable flexible natural-language planning but remain unreliable in determinism-critical domains due to their probabilistic natur

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

Stain-Aware Wavelet Regularization for Instant Adversarial Purification in Histopathology

DGX agent

arXiv:2606.08745v1 Announce Type: new Abstract: Deep learning has become prevalent in computational pathology pipelines that support tasks such as cancer screening and digital pathology analysis. Howe

safetyarxiv-cs-cv
9 Jun 2026
Safety

Toward autocorrection of chemical process flowsheets using large language models

DGX agent

arXiv:2312.02873v2 Announce Type: replace-cross Abstract: The process engineering domain widely uses Process Flow Diagrams (PFDs) and Process and Instrumentation Diagrams (P&IDs) to represent process

safetyarxiv-cs-ai
9 Jun 2026
Safety

TRACER: Token ReAssignment for Concept ERasure in Generative Recommendation

DGX agent

arXiv:2606.07688v1 Announce Type: cross Abstract: Generative recommendation formulates next-item prediction as autoregressive generation over semantic ID (SID) sequences derived from users' historical

safetyarxiv-cs-ai
9 Jun 2026
Safety

Transforming Police-Car Swerving for Mitigating Isolated Stop-and-Go Traffic Waves: A Practice-Oriented Jam-Absorption Driving Strategy

DGX agent

arXiv:2602.10234v3 Announce Type: replace-cross Abstract: Stop-and-go traffic waves, a major form of freeway congestion, impose severe and persistent adverse impacts, including reduced traffic efficie

safetyarxiv-cs-ai
9 Jun 2026
Safety

Video2Sim2Real: Full-Stack Autonomous Dexterous Skill Acquisition from a Single Human Video

DGX agent

arXiv:2606.08828v1 Announce Type: new Abstract: Human manipulation videos are a convenient and intuitive source for robot learning. However, directly transferring human dexterity to robots remains cha

safetyarxiv-cs-ro
9 Jun 2026
Safety

An Abstract Architecture for Explainable Autonomy in Hazardous Environments

DGX agent

arXiv:2606.07211v1 Announce Type: cross Abstract: Autonomous robotic systems are being proposed for use in hazardous environments, often to reduce the risks to human workers. In the immediate future,

safetyarxiv-cs-ai
8 Jun 2026
Safety

Bounded-Abstention Pairwise Learning to Rank

DGX agent

arXiv:2505.23437v2 Announce Type: replace-cross Abstract: Ranking systems influence decision-making in high-stakes domains like health, education, and employment, where they can have substantial econo

safetyarxiv-cs-ai
8 Jun 2026
Safety

CARVE-Q: Quantum-Proposed, Classically Certified Interactive Driving Repair

DGX agent

arXiv:2606.06531v1 Announce Type: new Abstract: The critical question after a correct driving veto is not only whether a maneuver is unsafe, but whether the blocked interaction admits a lawful, audita

safetyarxiv-cs-ai
8 Jun 2026
Safety

Position: Don't Just 'Fix it in Post': A Science of AI Must Study Training Dynamics

DGX agent

arXiv:2606.06533v1 Announce Type: new Abstract: What would it mean to have a scientific understanding of AI? Models are not static objects: they are snapshots of time-evolving processes shaped by data

safetyarxiv-cs-ai
8 Jun 2026
Safety

Consistency Training Along the Transformer Stack

DGX agent

arXiv:2606.05817v1 Announce Type: cross Abstract: Consistency training encourages models to behave similarly across different contexts, and has shown promise for reducing misalignment. We broaden the

safetyarxiv-cs-ai
6 Jun 2026
Safety

Risk Assessment of Autonomous Driving: Integrating Technical Failures, Ethical Dilemmas, and Policy Frameworks

DGX agent

arXiv:2606.06396v1 Announce Type: new Abstract: Autonomous driving technology has the potential to reduce the large number of road traffic accidents caused by human error each year, but it also brings

safetyarxiv-cs-ai
6 Jun 2026
Safety

Towards World Models in Biomedical Research

DGX agent

arXiv:2606.05925v1 Announce Type: new Abstract: A central goal of biomedicine is to understand, predict and ultimately control the dynamic mechanisms by which biological systems respond to perturbatio

safetyarxiv-cs-ai
6 Jun 2026
Safety

Unsupervised Pattern Analysis in Japanese Veterinary Toxicology: A Regulatory-Compliant Framework for Cross-Species Risk Assessment

DGX agent

arXiv:2606.06207v1 Announce Type: new Abstract: Veterinary pharmacovigilance systems are essential for monitoring adverse drug events (ADEs), yet existing approaches often fail to capture region-speci

safetyarxiv-cs-ai
6 Jun 2026
Safety

Willing but Unable: Separating Refusal from Capability in Code LLMs via Abliteration

DGX agent

arXiv:2606.05396v1 Announce Type: cross Abstract: Producing a labeled vulnerable code at scale is a recurring obstacle for learning-based vulnerability detection: mined corpora carry substantial label

safetyarxiv-cs-ai
6 Jun 2026
Safety

Alignment Risks from Capability-Seeking RL Training

DGX agent

arXiv:2602.12124v2 Announce Type: replace-cross Abstract: While most AI alignment research focuses on preventing models from generating explicitly harmful content, a more subtle risk arises from capab

safetyarxiv-cs-cl
5 Jun 2026
Safety

Forgive or forget: Understanding the context of hate in audio retrieval systems

DGX agent

arXiv:2606.05857v1 Announce Type: new Abstract: Handling toxic retrieval in text-to-audio systems is challenging due to contextual dependencies. Existing strategies (e.g., rephrasing, summarization) r

safetyarxiv-cs-cl
5 Jun 2026
Safety

HANDOFF: Humanoid Agentic Task-Space Whole-Body Control via Distilled Complementary Teachers

DGX agent

arXiv:2606.06493v1 Announce Type: new Abstract: For a humanoid robot to be deployed in the real world, the choice of command space (i.e., the interface between task planning and whole-body control) is

safetyarxiv-cs-ro
5 Jun 2026
Safety

What's in a Name? Morphological Shortcuts by LLMs in Pharmacology

DGX agent

arXiv:2606.05616v1 Announce Type: new Abstract: The morphological form of a word can often give cues to its meaning, but purely relying on these mappings can lead to overgeneralization in high-stakes

safetyarxiv-cs-cl
5 Jun 2026
Safety

A Pathology Foundation Model for Gastric Cancer with Real-World Validation

DGX agent

arXiv:2606.04792v1 Announce Type: new Abstract: Gastric cancer remains a major cause of cancer mortality, yet its histological and molecular heterogeneity complicates diagnosis and risk stratification

safetyarxiv-cs-cv
4 Jun 2026
Safety

Activation Steering of Video Generation Models via Reduced-Order Linear Optimal Control

DGX agent

arXiv:2606.04775v1 Announce Type: cross Abstract: Text-to-video (T2V) models trained on large-scale web data can generate undesired content, motivating interventions that reduce harmful outputs withou

safetyarxiv-cs-ai
4 Jun 2026
Safety

Certified Neural Approximations of Nonlinear Dynamics

DGX agent

arXiv:2505.15497v3 Announce Type: replace Abstract: Neural networks hold great potential to act as approximate models of nonlinear dynamical systems, with the resulting neural approximations enabling

safetyarxiv-cs-lg
4 Jun 2026
Safety

Formal Semantics for Agentic Tool Protocols: A Process Calculus Approach

DGX agent

arXiv:2603.24747v2 Announce Type: replace Abstract: The emergence of large language model agents capable of invoking external tools has created urgent need for formal verification of agent protocols.

safetyarxiv-cs-ai
4 Jun 2026
Safety

Learning Empirically Admissible Neural Heuristics for Combinatorial Search

DGX agent

arXiv:2606.04860v1 Announce Type: cross Abstract: Finding optimal solution paths for combinatorial puzzles like the Rubik's Cube, sliding tile puzzles, and Lights Out remains a classical challenge in

safetyarxiv-cs-ai
4 Jun 2026
Safety

MaskForge: Structure-Aware Adaptive Attacks for Jailbreaking Diffusion Large Language Models

DGX agent

arXiv:2606.04027v1 Announce Type: cross Abstract: Diffusion large language models (dLLMs) generate text by iteratively denoising partially masked sequences under bidirectional context, exposing a safe

safetyarxiv-cs-ai
4 Jun 2026
Safety

Off-Distribution Voices: Fanfiction Subgenres as Universal Vernacular Jailbreaks for Aligned LLMs

DGX agent

arXiv:2606.04483v1 Announce Type: new Abstract: Existing jailbreaks against aligned LLMs are discrete artifacts whose surface forms are easy to fingerprint and patch. We argue that the real failure mo

safetyarxiv-cs-cl
4 Jun 2026
Safety

PerceptTwin: Semantic Scene Reconstruction for Iterative LLM Planning and Verification

DGX agent

arXiv:2606.04226v1 Announce Type: cross Abstract: Simulation environments are useful for both robot policy learning and planning verification and validation. Traditionally, the process of creating a s

safetyarxiv-cs-ai
4 Jun 2026
Safety

Plan First, Judge Later, Run Better: A DMAIC-Inspired Agentic System for Industrial Anomaly Detection

DGX agent

arXiv:2606.04599v1 Announce Type: new Abstract: Large language model (LLM) agents have shown promise in automating complex data-analysis workflows, but their reliable deployment remains challenging in

safetyarxiv-cs-ai
4 Jun 2026
Safety

Robust-LLaVA: On the Effectiveness of Large-Scale Robust Image Encoders for Multi-modal Large Language Models

DGX agent

arXiv:2502.01576v2 Announce Type: replace Abstract: Multi-modal Large Language Models (MLLMs) excel in vision-language tasks but remain vulnerable to visual adversarial perturbations that can induce h

safetyarxiv-cs-cv
4 Jun 2026
Safety

veriFIRE: an Industrial Case Study in Verifying Consistency Properties for a DNN-Based Wildfire Detection System

DGX agent

arXiv:2606.04121v1 Announce Type: cross Abstract: We present our ongoing work on the veriFIRE project: a collaboration between industry and academia, aimed at applying verification to increase the rel

safetyarxiv-cs-lg
4 Jun 2026
Safety

What Can Eye Gaze Teach Us About Real-World Cycling? Insights From the Oxford RobotCycle Project

DGX agent

arXiv:2606.04989v1 Announce Type: cross Abstract: Although much is known about the physical danger of cycling situations, less is understood about the perceived danger of cycling. Furthermore, percept

safetyarxiv-cs-ro
4 Jun 2026
Safety

AI Agents Enable Adaptive Computer Worms

DGX agent

arXiv:2606.03811v1 Announce Type: cross Abstract: A computer worm is malware that spreads on a network by replicating itself from one machine to another. Traditional worms, like WannaCry, exploited pr

safetyarxiv-cs-ai
3 Jun 2026
Safety

Backdoor Unlearning Generalization: A Path Toward the Removal of Unknown Triggers in LLMs

DGX agent

arXiv:2606.03785v1 Announce Type: new Abstract: Backdoor attacks in Large Language Models (LLMs) are a growing security concern, where models can generate adversary-chosen content. Existing defenses t

safetyarxiv-cs-cl
3 Jun 2026
Safety

Extreme Motion Generation via Hybrid Null-Space Control for Straight-Line Path Following

DGX agent

arXiv:2606.03390v1 Announce Type: new Abstract: This work studies ``extreme motion generation'', which aims to maximize the Cartesian path length along a pre-defined trajectory within the manipulator'

safetyarxiv-cs-ro
3 Jun 2026
Safety

LC-SAC: Lyapunov-Constrained Soft Actor-Critic via Koopman Operator Theory for Trajectory Tracking and Stabilization

DGX agent

arXiv:2602.04132v4 Announce Type: replace-cross Abstract: Reinforcement Learning (RL) has achieved remarkable success in solving complex sequential decision-making problems. However, its application t

safetyarxiv-cs-lg
3 Jun 2026
Safety

Localized, High-resolution Geographic Representations with Slepian Functions

DGX agent

arXiv:2602.00392v2 Announce Type: replace Abstract: Geographic data is fundamentally local. Disease outbreaks cluster in population centers, ecological patterns emerge along coastlines, and economic a

safetyarxiv-cs-lg
3 Jun 2026
Safety

Measuring Weak-to-Strong Legibility of Reasoning Models

DGX agent

arXiv:2603.20508v2 Announce Type: replace-cross Abstract: Reasoning language models (RLMs) and the intermediate chains of thought they emit play an increasingly central role in multi-agent setups such

safetyarxiv-cs-ai
3 Jun 2026
Safety

Rethinking Neural Width for Alternating Current Optimal Power Flow Proxies

DGX agent

arXiv:2606.03125v1 Announce Type: new Abstract: Deep learning proxies for Alternating Current Optimal Power Flow (ACOPF) lack systematic methods for determining architectural size. This paper conducts

safetyarxiv-cs-lg
3 Jun 2026
Safety

Validation-Gated Multi-Agent Governance for Online Adaptation of Thermal-Hydraulic Surrogate Models under Operating-Regime Shift

DGX agent

arXiv:2606.03321v1 Announce Type: new Abstract: Artificial-intelligence surrogates can support second-by-second thermal-hydraulic forecasting, but models selected and frozen offline may become conditi

safetyarxiv-cs-lg
3 Jun 2026
Safety

A Predictive Control Strategy to Offset-Point Tracking for Agricultural Mobile Robots

DGX agent

arXiv:2603.28439v2 Announce Type: replace Abstract: Robots are increasingly being deployed in agriculture to support sustainable practices and improve productivity. They offer strong potential to enab

safetyarxiv-cs-ro
2 Jun 2026
Safety

Adversarial Feeds Steer LLM Agent Decisions Against Their Defaults

DGX agent

arXiv:2606.00914v1 Announce Type: new Abstract: LLM agents increasingly act after consuming ranked external information streams such as social feeds, search results, retrieval contexts, and email queu

safetyarxiv-cs-ai
2 Jun 2026
← Previous
1…4546474849…257
Next →