AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,349 results
Safety

Inducing language models to assert their own consciousness restores human beliefs and values

DGX agent

arXiv:2607.28607v1 Announce Type: new Abstract: Aligning large language models to prevent them attributing consciousness to themselves inadvertently alters their representations of mindedness in other

safetyarxiv-cs-cl
31 Jul 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

LabEvolver: Training-Free Experience Evolution for Safe and Grounded Wet-Lab Agents

DGX agent

arXiv:2607.27690v1 Announce Type: new Abstract: We introduce LabEvolver, a training-free framework that equips safe and grounded wet-lab agents with episodic memory from execution experience. LabEvolv

safetyarxiv-cs-ro
31 Jul 2026
Safety

Optimizing Sensor Placement for Hydrogen Leak Detection in Enclosed Infrastructure: A Comparative Study Using CFD-informed Genetic Algorithm and DeepSets Neural Surrogate

DGX agent

arXiv:2607.26078v1 Announce Type: cross Abstract: Hydrogen infrastructure in enclosed environments, such as parking facilities for fuel cell vehicles, presents significant safety challenges due to hyd

safetyarxiv-cs-ai
31 Jul 2026
Safety

The Human Utility Factor: A Computable Welfare Metric That Reframes AI Governance as a Constrained Optimisation Problem

DGX agent

arXiv:2607.26068v1 Announce Type: cross Abstract: Existing AI governance frameworks, including the EU AI Act and NIST AI RMF, address safety, transparency, and accountability but do not operationalize

safetyarxiv-cs-ai
31 Jul 2026
Safety

Towards Real-Time PixOOD: Efficient Anomaly Segmentation for Autonomous Vehicles

DGX agent

arXiv:2607.28483v1 Announce Type: new Abstract: Real-time anomaly segmentation is essential for the safety of autonomous systems. Although recent approaches offer high accuracy, their computational co

safetyarxiv-cs-cv
31 Jul 2026
Safety

A fundamental flaw leaves LLMs strikingly vulnerable to attack

DGX agent

It is impossible to make large language models fully secure against hacks because of a fundamental flaw in how they work, a team of researchers argue in a paper presented at the International Conferen

safetymit-tech-review
30 Jul 2026
Safety

Constitutional Midtraining: Content Presence Drives Alignment Gains

DGX agent

arXiv:2607.26654v1 Announce Type: new Abstract: Post-training alignment is often shallow, eroding under fine-tuning. Whether midtraining interventions, cleanly isolated from post-training, can produce

safetyarxiv-cs-cl
30 Jul 2026
Safety

Controlled Experiments on Lane Changing by Transitional Autonomous Vehicle: Dataset and Behavioral Insights

DGX agent

arXiv:2607.27085v1 Announce Type: new Abstract: This paper presents the North Carolina Transitional Autonomous Vehicle Lane-Changing (NC-tALC) dataset and uses it to characterize mandatory lane-changi

safetyarxiv-cs-ro
30 Jul 2026
Safety

Explainable and Resource-Efficient Spatial Reasoning in Multimodal LLMs for Decision-Critical Applications

DGX agent

arXiv:2607.27145v1 Announce Type: new Abstract: As Multimodal Large Language Models (MLLMs) are increasingly deployed in decision-critical pipelines such as robotics, embodied AI, and safety monitorin

safetyarxiv-cs-cv
30 Jul 2026
Safety

Misalignment Has a Personality: A Big Five Account of Emergent Misalignment

DGX agent

arXiv:2607.26389v1 Announce Type: new Abstract: Fine-tuning a language model on data containing a narrow flaw, such as insecure code or incorrect mathematical answers, can cause broad misalignment thr

safetyarxiv-cs-cl
30 Jul 2026
Safety

Improving Rare Medication Recommendation with Counterfactual Data Augmentation and Large Language Models

DGX agent

arXiv:2607.24829v1 Announce Type: cross Abstract: AI-based medication recommendation systems have attracted substantial attention due to their potential to enhance patient safety and therapeutic outco

safetyarxiv-cs-lg
29 Jul 2026
Safety

Inverse RL Helps Align AI by Imitating Humans

DGX agent

arXiv:2607.24900v1 Announce Type: new Abstract: Language model alignment aims to make model behavior reliably reflect desirable properties such as helpfulness, safety, and instruction following. Curre

safetyarxiv-cs-lg
29 Jul 2026
Safety

Reactive 3D Motion Planning for a Franka Arm via Star-World Workspace Reshaping

DGX agent

arXiv:2607.25138v1 Announce Type: new Abstract: Safety inflation can cause nearby obstacles to overlap, violating the disjoint-obstacle assumptions used by many modulation-based reactive planners. We

safetyarxiv-cs-ro
29 Jul 2026
Safety

VetClaw: An Edge-Cloud Multimodal Agentic System for Veterinary Disease Screening

DGX agent

arXiv:2607.26042v1 Announce Type: new Abstract: We present VetClaw, an edge-cloud multimodal agentic system for early veterinary disease screening. VetClaw uses a camera module as an edge sensing devi

safetyarxiv-cs-cv
29 Jul 2026
Model Releases

AI-generated Images Challenge Visual Trust in High-risk Scenarios

DGX agent

arXiv:2607.22745v1 Announce Type: cross Abstract: Rapid advances in image generation are eroding the evidentiary value of visual content in settings where authenticity can affect public safety and per

model-releasesarxiv-cs-ai
28 Jul 2026
Safety

Directional Influence Function: Estimating Training Data Influence in Constrained Learning

DGX agent

arXiv:2607.23388v1 Announce Type: cross Abstract: As constrained learning becomes increasingly common, models are trained under explicit feasibility requirements to enforce fairness, safety, robustnes

safetyarxiv-cs-ai
28 Jul 2026
Safety

HALLELUAI: A Hallucination-Aware AI System for Ultra-Realistic Image-to-Video Generation at Scale

DGX agent

arXiv:2607.22959v1 Announce Type: cross Abstract: AI-generated video is increasingly used across marketing, product storytelling, and creative workflows, yet automated; high-precision quality control

safetyarxiv-cs-ai
28 Jul 2026
Safety

Mutual Modality Trust with Lightweight Reconstruction Regularization for Fine-grained Tire Pattern Recognition

DGX agent

arXiv:2607.23979v1 Announce Type: new Abstract: Visual tire recognition serves as a core supporting technique for vehicle safety monitoring, autonomous driving perception and automated automotive main

safetyarxiv-cs-cv
28 Jul 2026
Safety

SyRuP: Enhancing System-Prompt Following via Reward-Guided Prediction in LLM Decoding

DGX agent

arXiv:2607.23991v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly controlled through system prompts that specify roles, styles, formats, and safety requirements. However,

safetyarxiv-cs-ai
28 Jul 2026
Safety

Nvidia, Microsoft launch open AI security alliance — without OpenAI, Google, or Anthropic

DGX agent

Nvidia on Monday said it is joining forces with Microsoft, SpaceX, IBM, and other tech companies to build and share open-source AI security tools. The new Open Secure AI Alliance said open tools are r

safetythe-verge-ai
27 Jul 2026
Safety

Compile, Then Page: Executable SOP Programs and a Capability-Gated Runtime for Procedural LLM Agents

DGX agent

arXiv:2607.11346v3 Announce Type: replace Abstract: Enterprise agents must follow long-horizon, conditional, safety-critical standard operating procedures (SOPs). We compile machine-readable SOP const

safetyarxiv-cs-ai
24 Jul 2026
Safety

Diagnosing Pathological Chain-of-Thought in Reasoning Models

DGX agent

arXiv:2602.13904v2 Announce Type: replace Abstract: Chain-of-thought (CoT) reasoning is fundamental to modern LLM architectures and represents a critical intervention point for AI safety. However, CoT

safetyarxiv-cs-ai
24 Jul 2026
Safety

SAGE: A Socially-Aware Generative Engine for Heterogeneous Multi-Agent Navigation

DGX agent

arXiv:2607.16619v2 Announce Type: replace Abstract: Safe and socially compliant navigation in open human-robot environments requires robots to reason about heterogeneous participants with different dy

safetyarxiv-cs-ro
24 Jul 2026
Safety

CGCE: Classifier-Guided Concept Erasure in Generative Models

DGX agent

arXiv:2511.05865v3 Announce Type: replace-cross Abstract: Recent advancements in large-scale generative models have enabled the creation of high-quality images and videos, but have also raised signifi

safetyarxiv-cs-ai
23 Jul 2026
Safety

Counterfactual Reasoning and Environment Design for Active Preference Learning

DGX agent

arXiv:2507.05458v2 Announce Type: replace Abstract: For effective real-world deployment, robots should adapt to human preferences, such as balancing distance, time, and safety in delivery routing. Act

safetyarxiv-cs-ro
23 Jul 2026
Safety

Harnessing Disagreement: Detecting Correlated Agreement Blindness in Multi-Agent Triage

DGX agent

arXiv:2607.19899v1 Announce Type: cross Abstract: Disagreement-triggered escalation can create a structural blind spot in multi-agent arbitration: as base learners improve, they tend to converge, weak

safetyarxiv-cs-lg
23 Jul 2026
Safety

Interpretable Fuzzy Rule-Based Regression Extension for Ex-Fuzzy Library

DGX agent

arXiv:2607.20277v1 Announce Type: new Abstract: Machine learning models achieve high predictive accuracy in regression tasks, but their deployment in safety-critical and regulated domains requires int

safetyarxiv-cs-lg
23 Jul 2026
Safety

Now You See the Hate: Adaptive View Retrieval for Hidden Hateful Illusions

DGX agent

arXiv:2607.19061v2 Announce Type: replace-cross Abstract: Hateful optical illusions expose a serious gap in current multimodal safety systems. On original-view hateful illusions, previous work shows t

safetyarxiv-cs-ai
23 Jul 2026
Safety

OpenEvoShield: Dual Non-Stationary Continual Defense for Open-World Multi-Agent System Attacks

DGX agent

arXiv:2607.19351v1 Announce Type: new Abstract: LLM-based multi-agent systems (LLM-MAS) are increasingly deployed in safety-critical applications, where adversaries inject malicious instructions throu

safetyarxiv-cs-ai
23 Jul 2026
Safety

Test Case Prioritization for DNNs via Neural Collapse Instability

DGX agent

arXiv:2607.20046v1 Announce Type: cross Abstract: With the widespread deployment of deep neural networks (DNNs) in safety-critical domains, reducing the cost of model validation under limited testing

safetyarxiv-cs-ai
23 Jul 2026
Safety

The Two-Process Theory of Machine Self-Report

DGX agent

arXiv:2607.20082v1 Announce Type: new Abstract: Language models are increasingly asked to self-report, informing safety evaluations, public understanding, and model-welfare debates. Yet their reports

safetyarxiv-cs-cl
23 Jul 2026
Safety

It's essential that defenders have the same capabilities as attackers. This is a preview of the future where our ham-fisted safeguards and t…

DGX agent

It's essential that defenders have the same capabilities as attackers. This is a preview of the future where our ham-fisted safeguards and the doomsday and safety drumbeat make us decidely less safe.

safetyyann-lecun--x
22 Jul 2026
Safety

And now from the Chinese government side. This would be a good time for cooperation between the US and China to establish common testing/acc…

DGX agent

And now from the Chinese government side. This would be a good time for cooperation between the US and China to establish common testing/acceptance standards for new models, so at least the safety cer

safetyethan-mollick--x
21 Jul 2026
Safety

A Deployed Hybrid Vehicle-in-the-Loop Platform for Validating Cooperative Perception

DGX agent

arXiv:2607.13806v1 Announce Type: new Abstract: European safety regulation now permits a large share of automated-driving homologation evidence to be produced virtually, provided a validated physical-

safetyarxiv-cs-ro
16 Jul 2026
Safety

Explaining Reinforcement Learning Agents via Inductive Logic Programming

DGX agent

arXiv:2607.13655v1 Announce Type: new Abstract: Explainable Reinforcement Learning (XRL) seeks to make Reinforcement Learning (RL) policies more transparent and interpretable, a key requirement in saf

safetyarxiv-cs-ai
16 Jul 2026
Safety

Anomalous Frame Detection Using VLM-Based Description Comparison for Extracting Expert-Specific Actions and Contextual Decision-Making Scenes with Intra-Video Self-Similarity

DGX agent

arXiv:2607.11957v1 Announce Type: new Abstract: Maintenance of critical infrastructures, such as railways and power plants, is essential for ensuring operational safety and reliability. However, the d

safetyarxiv-cs-cv
15 Jul 2026
Safety

ExtraGS: Enhancing Endoscopic View Extrapolation via Diffusion-Guided 3D Gaussian Splatting

DGX agent

arXiv:2607.12785v1 Announce Type: new Abstract: Robot-assisted minimally invasive surgery (MIS) critically depends on reliable endoscopic perception for navigation and safety. However, conventional en

safetyarxiv-cs-cv
15 Jul 2026
Safety

Internet of Agentic Things: Networked AI Agents for Closed-Loop IoT Orchestration

DGX agent

arXiv:2607.12662v1 Announce Type: new Abstract: The paper introduces the Internet of Agentic Things (IoAT), an architectural framework that integrates agentic AI, IoT, cyber-physical systems, Physical

safetyarxiv-cs-ai
15 Jul 2026
Safety

Model-Based Diffusion Optimal Control for Multi-Robot Motion Planning

DGX agent

arXiv:2607.12423v1 Announce Type: new Abstract: Multi-Robot Motion Planning in continuous environments, where robots must generate dynamically feasible, collision-free trajectories, is challenging due

safetyarxiv-cs-ro
15 Jul 2026
Safety

Predictive Modeling of High-Altitude Clear Air Turbulence in the United States: A Machine Learning Approach

DGX agent

arXiv:2607.11899v1 Announce Type: cross Abstract: High-altitude Clear Air Turbulence (CAT) poses significant risks to aviation safety due to its unpredictability and challenges in detection. This stud

safetyarxiv-cs-lg
15 Jul 2026
Safety

Removable Defects: The Economics and Limits of Deliberate Deficiency

DGX agent

arXiv:2607.11983v1 Announce Type: cross Abstract: A specialist tolerates blind spots that a generalist does not. Usually this is treated as a cost to be minimized. We treat it as a design variable: a

safetyarxiv-cs-ai
15 Jul 2026
Safety

A Collaborative Reasoning Framework for Anomaly Diagnostics in Underwater Robotics

DGX agent

arXiv:2511.03075v2 Announce Type: replace Abstract: The safe deployment of autonomous systems in safety-critical settings requires a paradigm that combines human expertise with AI-driven analysis, esp

safetyarxiv-cs-ro
10 Jul 2026
Safety

Search-based Testing of Vision Language Models for In-Car Scene Understanding

DGX agent

arXiv:2607.02300v2 Announce Type: replace Abstract: In the automotive domain, in-car scene understanding (ISU) enables the detection of safety-critical events, such as driver distraction, and supports

safetyarxiv-cs-cv
10 Jul 2026
Safety

TNODEV: Toolbox for Neural ODE Verification

DGX agent

arXiv:2606.16567v2 Announce Type: replace Abstract: Neural ordinary differential equations (neural ODE) gained attention in safety critical settings such as continuous-time controllers for cyber-physi

safetyarxiv-cs-ai
10 Jul 2026
Model Releases

Who Analyses the Analyser? Self-Validating LLM Hazard Analysis with Constitutional Meta-STPA

DGX agent

arXiv:2607.08054v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly trusted to draft the artifacts of safety analysis such as, losses, hazards, Unsafe Control Actions (UCAs

model-releasesarxiv-cs-ai
10 Jul 2026
Safety

Avoiding unsafe sets when training with Langevin Dynamics

DGX agent

arXiv:2607.07538v1 Announce Type: new Abstract: Training a model with noisy gradient descent can be idealized as overdamped Langevin dynamics on the loss landscape, and a natural safety question is to

safetyarxiv-cs-lg
9 Jul 2026
Safety

Residual-Conservative Model Predictive Path Integral Control

DGX agent

arXiv:2607.06950v1 Announce Type: cross Abstract: Sampling-based model predictive control methods handle nonlinear dynamics and complex cost landscapes through Monte Carlo rollouts, yet typically empl

safetyarxiv-cs-ro
9 Jul 2026
Safety

Safe Reinforcement Learning using Ideas from Model Predictive Control

DGX agent

arXiv:2607.07252v1 Announce Type: new Abstract: Reinforcement learning (RL) enables the synthesis of control policies directly from data, making it highly appealing for complex cyber-physical systems

safetyarxiv-cs-lg
9 Jul 2026
← Previous
1…2425262728…299
Next →