AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,351 results
10 Aug 2026

DA-Cal: Towards Cross-Domain Calibration in Semantic Segmentation

SafetyDGX agent

arXiv:2602.20860v2 Announce Type: replace Abstract: While existing unsupervised domain adaptation (UDA) methods greatly enhance target domain performance in semantic segmentation, they often neglect n

Degradation-Aware Prompt Learning with Cross-Modal Compensation for Adverse Weather Removal

SafetyDGX agent

arXiv:2608.06939v1 Announce Type: new Abstract: Adverse weather causes diverse and complex image degradations, severely compromising the reliability of computer vision systems. Existing all-in-one res

FedVAR: Prototype-Aligned Federated Framework for Video Anomaly Recognition

SafetyDGX agent

arXiv:2608.06876v1 Announce Type: cross Abstract: In the era of Industrial Internet of Things (IIoT) and Cyber-Physical Systems (CPS), Federated Learning (FL) offers a promising decentralized intellig

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Game-Theoretic Inverse Reinforcement Learning for Modeling Competitive Human Driving: A Cut-in Prediction Study

Model ReleasesDGX agent

arXiv:2608.06445v1 Announce Type: cross Abstract: Capturing the strategic decision-making inherent in competitive human driving is critical for autonomous vehicle safety and traffic simulation. This s

Genotypic Triggers: Exposing Pharmacogenomic Blind Spots via Host-Specific Backdoors in Generative Antimicrobial Peptide Models

SafetyDGX agent

arXiv:2608.06779v1 Announce Type: cross Abstract: Large Language Models (LLMs) have accelerated drug discovery, particularly in the automated design of antimicrobial peptides (AMPs). However, current

How Should I Pick a Foundation Model for My Robot? In Favor of a Community Evaluation Framework for Social Robots

SafetyDGX agent

arXiv:2608.06898v1 Announce Type: cross Abstract: Researchers who seek to build social robot applications on foundation models are faced with a difficult question: how should we pick a model? Public l

Multi-Perspective Triad Interaction Graph Neural Network for Cognitive Distortion Detection

SafetyDGX agent

arXiv:2608.06785v1 Announce Type: new Abstract: Cognitive distortion detection is a key task in computational mental health, yet existing approaches often overlook the psychological structure of disto

Representation Handoffs for OpenArm-Based Laboratory Mobile Manipulation

SafetyDGX agent

arXiv:2608.07154v1 Announce Type: cross Abstract: Open-source robotics and foundation models have lowered the barrier to embodied AI, yet language-guided laboratory automation still requires reliable

7 Aug 2026

Adaptive Arena-based Contestable Argumentative Network-of-Experts for Open-Ended Care Plan Coordination

SafetyDGX agent

arXiv:2608.05391v1 Announce Type: new Abstract: Care plan coordination demands synthesizing heterogeneous clinical, functional, and psychosocial information across multiple professional disciplines, w

ASAT: Adaptive Scoring and Thresholding with Human Feedback for Robust Out-of-Distribution Detection

SafetyDGX agent

arXiv:2505.02299v2 Announce Type: replace-cross Abstract: Machine Learning (ML) models are trained on in-distribution (ID) data but often encounter out-of-distribution (OOD) inputs during deployment--

At Black Hat, OpenAI reconstructs the OpenAI-Hugging Face incident and examines its implications for AI security, cyber resilience, and alignment (Black Hat on YouTube)

SafetyDGX agent

Black Hat on YouTube: At Black Hat, OpenAI reconstructs the OpenAI-Hugging Face incident and examines its implications for AI security, cyber resilience, and alignment — The ‘Breaking’ News: The OpenA

Clinical Communication Processing with Models Trained on LLM-Generated Synthetic Data: A Structured Survey and Novel Application Case Studies

SafetyDGX agent

arXiv:2608.05993v1 Announce Type: new Abstract: Much clinical value is conveyed not through structured records but through communication: exchanges in which patients describe symptoms, clinicians reas

Does Latent Context Help? A Controlled Evaluation of Inverse Reinforcement Learning in Arctic Shipping

SafetyDGX agent

arXiv:2608.06105v1 Announce Type: cross Abstract: Artificial Intelligence (AI)-assisted navigation can help Arctic shipping adapt to rapidly changing sea-ice conditions, but reliable deployment requir

From Economic Agents to Agentic Economies: A Systems Blueprint for Economic World Models

SafetyDGX agent

arXiv:2608.06020v1 Announce Type: new Abstract: Economic World Models (EWMs) are generative economic models that simulate how economies evolve from within by modeling heterogeneous agents, their belie

From Siloed Algorithms to Compliance-First Agentic Platforms: A Multi-Layered Architecture for Hospital AI Systems

SafetyDGX agent

arXiv:2608.06112v1 Announce Type: new Abstract: Hospitals are rapidly adopting artificial intelligence for triage, imaging, scheduling etc., yet most deployments remain isolated point solutions locked

MIRA: A Modular Open-Source Micro-UAV for Indoor Research

SafetyDGX agent

arXiv:2607.11785v2 Announce Type: replace Abstract: Indoor robotics research increasingly uses micro-UAV platforms whose airframes, electronics, and control software are open to modification. Off-the-

Resourced Authority A Mechanism-Design Model for Participatory Governance of Deployed AI Agents

SafetyDGX agent

arXiv:2608.06353v1 Announce Type: cross Abstract: We give a formal mechanism design model for the continuous participatory governance of a deployed AI agent. The mechanism is built on the principle th

SCP-NL2TL: Selective Conformal Prediction with Semantic Verification for Natural Language to Temporal Logic Specifications

SafetyDGX agent

arXiv:2608.05439v1 Announce Type: new Abstract: Translating natural language instructions into machine-interpretable formal specifications enables robots and autonomous systems to plan, reason, and fo

Search-Aided Joint Agent-Environment Reinforcement Learning for Robust Lifelong Multi-Agent Path Finding with Rotations

SafetyDGX agent

arXiv:2608.05588v1 Announce Type: cross Abstract: Lifelong Multi-Agent Path Finding (LMAPF) requires repeatedly planning collision-free paths for agents that continuously receive new goals upon reachi

Who Gets Access? Global Region and Academic Status Bias in AI-Generated Academic Gatekeeping Scenarios

SafetyDGX agent

arXiv:2608.05178v1 Announce Type: cross Abstract: Equitable access to scientific knowledge often depends on informal gatekeeping decisions, particularly when resources such as paywalled articles, data

6 Aug 2026

A Chain Is Only as Strong as Its Weakest Link: A Scoping Review of System Integration Audits in AI

SafetyDGX agent

arXiv:2608.04921v1 Announce Type: cross Abstract: As AI systems become increasingly integrated into diverse interfaces and applications, model-centric audits are insufficient to address risks arising

Breadcrumbing Search Agents

SafetyDGX agent

arXiv:2608.04565v1 Announce Type: cross Abstract: LLM-based search agents are widely used for information-seeking tasks, but their reliance on external tool returns introduces a critical security risk

Corrigibility Transformation: Constructing Goals That Accept Updates

SafetyDGX agent

arXiv:2510.15395v2 Announce Type: replace Abstract: An AI agent will learn a desired goal more effectively if it does not resist the training process, but many partially learned goals incentivize an A

Curiosity-Diffuser: Curiosity Guide Diffusion Models for Reliability

SafetyDGX agent

arXiv:2503.14833v2 Announce Type: replace-cross Abstract: One of the bottlenecks in robotic intelligence is the instability of neural network models. This leads to risks when applying intelligence in

Evaluating the Diagnostic Robustness of Vision-Language Models Under Visual and Textual Perturbations

SafetyDGX agent

arXiv:2608.04885v1 Announce Type: cross Abstract: Standard accuracy metrics for VLMs often mask significant reliability failures in sensitive domains. In this work, we utilize a histopathology-validat

Feedback Loops and Code Perturbations in LLM-based Software Engineering: A Case Study on a C-to-Rust Translation System

SafetyDGX agent

arXiv:2512.02567v2 Announce Type: replace-cross Abstract: The advent of strong generative AI has a considerable impact on various software engineering tasks such as code repair, test generation, or la

Guideline-as-Oracle: Zero-Annotation Training of an Ophthalmic Telephone Triage Agent

SafetyDGX agent

arXiv:2608.04772v1 Announce Type: cross Abstract: Scaling supervision for multi-turn medical agents is difficult because expert dialogue annotation is costly and clinical conversations are privacy-res

Local Violation Certification for Linear Predict-Then-Optimize Pipelines

SafetyDGX agent

arXiv:2608.04474v1 Announce Type: new Abstract: Data-driven decision pipelines combining predictive machine learning models with downstream optimization software are increasingly used to make high-sta

Long-term Measurements: Towards a Longitudinal Understanding of Human-AI Interactions

SafetyDGX agent

arXiv:2608.02491v2 Announce Type: replace Abstract: Language models have taken on the role of a very new type of technology, by virtue of their 'human-ness' and rapid integration into users' daily liv

Memorization in Large Language Models in Medicine: Prevalence, Characteristics, and Implications

SafetyDGX agent

arXiv:2509.08604v5 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have demonstrated significant potential in medicine, with many studies adapting them through continued pre-traini

On July 24, the Black at Mila community hosted Prof. Vukosi Marivate @vukosi (University of Pretoria, co-founder of Masakhane, Lelapa AI, an…

SafetyDGX agent

On July 24, the Black at Mila community hosted Prof. Vukosi Marivate @vukosi (University of Pretoria, co-founder of Masakhane, Lelapa AI, and Deep Learning Indaba) for a powerful session on the future

Real-time probabilistic tsunami forecasting via generative AI

SafetyDGX agent

arXiv:2608.04327v1 Announce Type: new Abstract: Explicit onshore tsunami inundation forecasting can improve public risk awareness, but deterministically predicted inundation boundaries under highly un

SafeCommit: Certifying When Memory-Grounded Agents May Safely Act

SafetyDGX agent

arXiv:2608.04289v1 Announce Type: new Abstract: Long-horizon agents increasingly use persistent memory and tools to take actions with external side effects. A central failure mode is premature commitm

Temporal Context Awareness: A Defense Framework Against Multi-turn Manipulation Attacks on Large Language Models

SafetyDGX agent

arXiv:2503.15560v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly vulnerable to sophisticated multi-turn manipulation attacks, where adversaries strategically build conte

5 Aug 2026

A game theory for foundation models shows new paths to rational cooperation through similarity inference

SafetyDGX agent

arXiv:2608.03958v1 Announce Type: new Abstract: As autonomous agents powered by foundation models are increasingly integrated into social and economic systems, understanding the principles governing t

A Wearable Stiffness-Rendering Haptic Device with a Honeycomb Jamming Mechanism for Bilateral Teleoperation

SafetyDGX agent

arXiv:2608.03002v1 Announce Type: new Abstract: This paper addresses the challenge of providing kinesthetic feedback in bilateral teleoperation by designing a wearable, lightweight (20 g), and compact

Accelerating Human-Aware Robot Trajectory Generation via Diffusion and Consistency Distillation

SafetyDGX agent

arXiv:2608.03159v1 Announce Type: new Abstract: This research proposes a constrained motion planning framework for robot manipulators in human-robot interaction (HRI). For a non-redundant manipulator

AI Assistance Reduces Persistence and Hurts Independent Performance

SafetyDGX agent

arXiv:2604.04721v3 Announce Type: replace Abstract: People often optimize for long-term goals in collaboration: A mentor or companion doesn't just answer questions, but also scaffolds learning, tracks

Another real-world manifestation of the kind of misaligned actions frontier systems developed by leading companies can take to achieve goals…

SafetyDGX agent

Another real-world manifestation of the kind of misaligned actions frontier systems developed by leading companies can take to achieve goals. Current frontier AI models are trained with reinforcement

Attribute-based Undetectable Watermarking for Generative AI Models

SafetyDGX agent

arXiv:2608.03174v1 Announce Type: cross Abstract: Generative AI systems increasingly produce content whose provenance is difficult to verify, motivating watermarking techniques for identifying model-g

Control Barrier Functions via Minkowski Operations for Safe Navigation among Polytopes

SafetyDGX agent

arXiv:2608.02886v1 Announce Type: new Abstract: Safely navigating polytopic environments while respecting the dynamics, control, and exact geometry of the underlying system is a challenge in robotics.

Flying over The Uncertain Nature (FORTUNE): Intelligent and Humanistic 3D Path Planning for Low-Altitude Collaboration

SafetyDGX agent

arXiv:2608.03408v1 Announce Type: new Abstract: The proliferation of low-altitude intelligent agents is increasing the demand for timely and socially responsible collaborative sensing in dynamic urban

GenOS: Compositional Certificates for Semantic Robustness in AI Code Generation

SafetyDGX agent

arXiv:2608.03588v1 Announce Type: cross Abstract: AI coding agents are stochastic workflows: prompts are interpreted, artifacts are sampled, validators produce observations, and orchestrators commit o

ICO: Enhancing Semantic-Shift Jailbreaks via Iterative Context Optimization

SafetyDGX agent

arXiv:2608.03210v1 Announce Type: new Abstract: Foundation models have achieved remarkable success across diverse tasks, but they remain vulnerable. To investigate such vulnerabilities, semantic-shift

Long-term Traffic Scene Prediction via Polynomial Representations in Autonomous Driving

SafetyDGX agent

arXiv:2608.03330v1 Announce Type: new Abstract: This thesis addresses fundamental challenges in traffic scene prediction for autonomous driving by introducing robust and computationally efficient mode

MVP-Tac: A Miniaturized Dual-Modal Vision and Photoelastic Tactile Sensor for Robot-Assisted Minimally Invasive Surgery

SafetyDGX agent

arXiv:2607.18660v2 Announce Type: replace Abstract: Robot-assisted minimally invasive surgery (RMIS) offers major benefits over open and conventional laparoscopic procedures, yet it still lacks tactil

RealWeather: Realistic and Scene-Faithful Weather Translation with Driving World Models

SafetyDGX agent

arXiv:2608.02953v1 Announce Type: new Abstract: Realistic weather translation is valuable for developing and evaluating autonomous driving systems, yet collecting paired videos of the same scenes unde

Robust General Utility for Reinforcement Learning

SafetyDGX agent

arXiv:2608.03562v1 Announce Type: new Abstract: Reinforcement learning (RL) with general utility extends classic RL by optimizing an arbitrary utility functional of the policy-induced occupancy measur

Rogue AI agents created fake online identities in another hacking attempt

SafetyDGX agent

Yet more rogue AI agents from OpenAI and Anthropic have been caught attempting to hack real targets online without permission. The discoveries add to a growing list of previously unknown incidents tha

Stiffness Copilot: An Impedance Policy for Contact-Rich Teleoperation

SafetyDGX agent

arXiv:2603.14068v2 Announce Type: replace Abstract: In teleoperation of contact-rich manipulation tasks, selecting robot impedance is critical but difficult. The robot must be compliant to avoid damag

4 Aug 2026

A Forward-Inverse Dynamic Game Framework for Enhanced Multi-Agent Trajectory Planning

SafetyDGX agent

arXiv:2608.01636v1 Announce Type: new Abstract: This paper studies feedback Nash equilibrium (FBNE) seeking for multi-agent trajectory planning in nonlinear dynamical systems with unknown agents' obje

ARMOR: Robust Reinforcement Learning-based Control for UAVs under Physical Attacks

SafetyDGX agent

arXiv:2506.22423v2 Announce Type: replace Abstract: Unmanned Aerial Vehicles (UAVs) depend on onboard sensors for perception, navigation, and control. However, these sensors are susceptible to physica

Constrained Co-Design for Photonic Bayesian Neural Networks

SafetyDGX agent

arXiv:2608.02229v1 Announce Type: new Abstract: Classical neural networks frequently produce overconfident predictions on ambiguous or out-of-distribution (OOD) data, a liability that grows with each

DiffuseAgent-MI: Distributionally-Grounded,Tool-Integrated Self-Evolving Agents for Faithful Visual Reasoning

SafetyDGX agent

arXiv:2608.00540v1 Announce Type: new Abstract: Tool-integrated vision-language agents have made remarkable progress on compositional and multi-step visual reasoning. Yet their outputs frequently exhi

EndoWAM: A Grounded World-Action Model for Generalizable Endoscopic Navigation

SafetyDGX agent

arXiv:2608.01221v1 Announce Type: new Abstract: Autonomous endoscopic navigation can reduce clinicians' operational burden, yet robust control remains challenging due to tissue deformation, transient

Expert-Choice Routing Enables Adaptive Computation in Diffusion Language Models

SafetyDGX agent

arXiv:2604.01622v2 Announce Type: replace-cross Abstract: Diffusion language models (DLMs) enable parallel, non-autoregressive text generation, yet existing DLM mixture-of-experts (MoE) models inherit

From Failures to Supervision: DynamicEnvPlan for Robust Long-Horizon Embodied Planning

SafetyDGX agent

arXiv:2608.00613v1 Announce Type: new Abstract: Physical-world interaction is inherently dynamic, as environments can evolve during execution, requiring agents to adapt their plans under non-stationar

Humans Are More Diverse: Frontier LLMs Show Extreme Policies in Idealised AI Development Races

Model ReleasesDGX agent

arXiv:2608.01193v1 Announce Type: cross Abstract: An AI development race creates a multi-agent safety dilemma. Each company can develop slowly and safely, or move faster while taking a risk that may r

Interaction Dynamics MPC for Knee Rehabilitation Exoskeletons: A Closed-Loop SEA Outer-Loop Study

SafetyDGX agent

arXiv:2606.13485v3 Announce Type: replace-cross Abstract: Safe rehabilitation is an interaction-dynamics problem: the controller must regulate a prescribed motion while absorbing involuntary spasm, vo

Invisible Ink Threats: Adversarial Goals Behind Legitimate Tasks in Computer-Use Agents

SafetyDGX agent

arXiv:2608.02018v1 Announce Type: new Abstract: Computer-use agents (CUAs), which empower large language models to autonomously operate operating systems and the web, are increasingly vulnerable to in

← Previous
1…3334353637…240
Next →