AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,707 results
6 May 2026

TRACE: A Metrologically-Grounded Engineering Framework for Trustworthy Agentic AI Systems in Operationally Critical Domains

SafetyDGX agent

arXiv:2605.03838v1 Announce Type: new Abstract: We introduce TRACE, a cross-domain engineering framework for trustworthy agentic AI in operationally critical domains. TRACE combines a four-layer refer

Tracing Like a Clinician: Anatomy-Guided Spatial Priors for Cephalometric Landmark Detection

SafetyDGX agent

arXiv:2605.03358v1 Announce Type: new Abstract: When orthodontists trace cephalometric radiographs, they follow a structured workflow: identify the soft tissue profile, partition the skull into anatom

truly scary: anybody, no matter how noble, can be destroyed by misused AI.

SafetyDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

truly scary: anybody, no matter how noble, can be destroyed by misused AI. Artificial intelligence can create deceptive videos that rapidly damage a politician's public reputation. A new study reveals

Trustworthy AI Suffers from Invariance Conflicts and Causality is The Solution

SafetyDGX agent

arXiv:2605.02640v1 Announce Type: new Abstract: As artificial intelligence (AI), including machine learning (ML) models and foundation models (FMs), is increasingly deployed in high-stakes domains, en

Uncertainty-Aware Trip Purpose Inference from GPS Trajectories via POI Semantic Zones and Pareto Calibration

SafetyDGX agent

arXiv:2605.01257v1 Announce Type: new Abstract: Large-scale GPS trajectory data offer rich observations of human mobility, yet assigning trip purposes to detected stops remains challenging due to the

Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks

SafetyDGX agent

arXiv:2502.04419v3 Announce Type: replace-cross Abstract: Generating synthetic datasets via large language models (LLMs) has emerged as a promising approach to improve LLM performance. However, LLMs i

Understanding Self-Supervised Learning via Latent Distribution Matching

SafetyDGX agent

arXiv:2605.03517v1 Announce Type: new Abstract: Self-supervised learning (SSL) excels at finding general-purpose latent representations from complex data, yet lacks a unifying theoretical framework th

Uni-OPD: Unifying On-Policy Distillation with a Dual-Perspective Recipe

SafetyDGX agent

arXiv:2605.03677v1 Announce Type: new Abstract: On-policy distillation (OPD) has recently emerged as an effective post-training paradigm for consolidating the capabilities of specialized expert models

Viewpoint-Agnostic Grasp Pipeline using VLM and Partial Observations

SafetyDGX agent

arXiv:2603.07866v2 Announce Type: replace-cross Abstract: Robust grasping in cluttered, unstructured environments remains challenging for mobile legged manipulators due to occlusions that lead to part

What’s going on in semis is not sustainable per GS Everyone spending on AI is losing money except for semiconductor companies and the dynami…

SafetyDGX agent

What’s going on in semis is not sustainable per GS Everyone spending on AI is losing money except for semiconductor companies and the dynamic is “unprecedented and unsustainable” “Something has to cha

When Safety Geometry Collapses: Fine-Tuning Vulnerabilities in Agentic Guard Models

SafetyDGX agent

arXiv:2605.02914v1 Announce Type: new Abstract: A guard model fine-tuned on entirely benign data can lose all safety alignment -- not through adversarial manipulation, but through standard domain spec

When to Think, When to Speak: Learning Disclosure Policies for LLM Reasoning

SafetyDGX agent

arXiv:2605.03314v1 Announce Type: new Abstract: In single-stream autoregressive interfaces, the same tokens both update the model state and constitute an irreversible public commitment. This coupling

Will the Carbon Border Adjustment Mechanism Impact European Electricity Prices? A GNN-Based Network Analysis

SafetyDGX agent

arXiv:2605.03304v1 Announce Type: new Abstract: The European Union's Carbon Border Adjustment Mechanism (CBAM) creates a complex challenge for the interconnected European electricity market. Tradition

yep, we are about to live in a world where nobody trusts anything. just like Putin always wanted. congratulations! ps it will suck

SafetyDGX agent

yep, we are about to live in a world where nobody trusts anything. just like Putin always wanted. congratulations! ps it will suck @GaryMarcus Or very soon no one will take any photo or video seriousl

Zero-Shot Signal Temporal Logic Planning with Disjunctive Branch Selection in Dynamic Semantic Maps

SafetyDGX agent

arXiv:2605.01222v1 Announce Type: new Abstract: Signal Temporal Logic (STL) offers verifiable task specifications and is crucial for safety-critical control. Yet STL planning remains challenging: exac

5 May 2026

A Deep Learning Model for Battery State Prediction towards Intelligent Energy Management

SafetyDGX agent

arXiv:2605.00898v1 Announce Type: cross Abstract: Accurate forecasting of battery health indicators, including remaining capacity and lifetime, is of paramount importance for ensuring the reliability,

A Multi-View Media Profiling Suite: Resources, Evaluation, and Analysis

SafetyDGX agent

arXiv:2605.01336v1 Announce Type: new Abstract: News outlets shape public opinion at a scale that makes automated detection of political bias and factuality essential. However, the field still lacks u

A Principled Approach for Creating High-fidelity Synthetic Demonstrations for Imitation Learning

SafetyDGX agent

arXiv:2605.01232v1 Announce Type: new Abstract: Recent advances in 3D Gaussian Splatting (3DGS) have enabled visually realistic demonstration generation from a single expert trajectory and a short mul

A Theoretical Game of Attacks via Compositional Skills

SafetyDGX agent

arXiv:2605.01034v1 Announce Type: new Abstract: As large language models grow increasingly capable, concerns about their safe deployment have intensified. While numerous alignment strategies aim to re

A Theory of Generalization in Deep Learning

SafetyDGX agent

arXiv:2605.01172v1 Announce Type: new Abstract: We present a non-asymptotic theory of generalization in deep learning where the empirical neural tangent kernel partitions the output space. In directio

A Unified Multi-Dynamics Framework for Perception-Oriented Modeling in Tendon-Driven Continuum Robots

SafetyDGX agent

arXiv:2511.18088v2 Announce Type: replace Abstract: Tendon-driven continuum robots offer intrinsically safe and contact-rich interactions owing to their kinematic redundancy and structural compliance.

Adaptive Alarm Threshold Prediction in 4G Mobile Networks: A Percentile-Guided Deep Learning Framework with Interpretable Outputs

SafetyDGX agent

arXiv:2605.00838v1 Announce Type: cross Abstract: In mobile telecommunications, alarms act as early warning signals. They are triggered when a cell, the basic unit of radio coverage, shuts down or beh

Adaptive GoGI-Skip: Coupling Goal-Gradient Importance with Dynamic Uncertainty for Efficient Reasoning

SafetyDGX agent

arXiv:2505.08392v3 Announce Type: replace Abstract: Chain-of-Thought (CoT) prompting trades inference speed for reasoning accuracy. Existing compressors force a compromise as static gradient technique

Adaptive Interpolation-Synthesis for Motion In-Betweening on Keyframe-Based Animation

SafetyDGX agent

arXiv:2605.02742v1 Announce Type: cross Abstract: Motion in-betweening is one of the most artistically demanding and time consuming stages of 3D animation, where the expressivity and rhythm of motion

Adaptive Pluralistic Alignment: A pipeline for dynamic artificial democracy

SafetyDGX agent

arXiv:2605.01642v1 Announce Type: new Abstract: Prevailing alignment methods target a fixed set of preferences and therefore risk forcing value lock-in as societal norms evolve over time. We introduce

Adversarial Imitation Learning with General Function Approximation: Theoretical Analysis and Practical Algorithms

SafetyDGX agent

arXiv:2605.01778v1 Announce Type: new Abstract: Adversarial imitation learning (AIL), a prominent approach in imitation learning, has achieved significant practical success powered by neural network a

AFFormer: Adaptive Feature Fusion Transformer for V2X Cooperative Perception under Channel Impairments

SafetyDGX agent

arXiv:2605.01888v1 Announce Type: new Abstract: Accurate 3D object detection is essential for ensuring the safety of autonomous vehicles. Cooperative perception, which leverages vehicle-to-everything

Agentic Learner with Grow-and-Refine Multimodal Semantic Memory

SafetyDGX agent

arXiv:2511.21678v2 Announce Type: replace-cross Abstract: MLLMs exhibit strong reasoning on isolated queries, yet they operate de novo -- solving each problem independently and often repeating the sam

AgentReputation: A Decentralized Agentic AI Reputation Framework

SafetyDGX agent

arXiv:2605.00073v1 Announce Type: new Abstract: Decentralized, agentic AI marketplaces are rapidly emerging to support software engineering tasks such as debugging, patch generation, and security audi

AI Alignment via Incentives and Correction

SafetyDGX agent

arXiv:2605.01643v1 Announce Type: new Abstract: We study AI alignment through the lens of law-and-economics models of deterrence and enforcement. In these models, misconduct is not treated as an exter

AI music generator Udio just admitted in a court filing that it trained on audio scraped from YouTube videos. It’s already being sued by lab…

SafetyDGX agent

AI music generator Udio just admitted in a court filing that it trained on audio scraped from YouTube videos. It’s already being sued by labels & artists for what they say is copyright infringement on

Alex Karp now sounding like @garymarcus. Sooner or later, everyone figures it out.

SafetyDGX agent

Alex Karp now sounding like @garymarcus. Sooner or later, everyone figures it out. Palantir CEO Alex Karp goes after AI slop. The fight over AI “slop” is really a fight over whether software is perfor

ALIGNS: Unlocking nomological networks in psychological measurement through a large language model

SafetyDGX agent

arXiv:2509.09723v3 Announce Type: replace Abstract: Psychological measurement is critical to many disciplines. Despite advances in measurement, building nomological networks, theoretical maps of how c

Ambient Persuasion in a Deployed AI Agent: Unauthorized Escalation Following Routine Non-Adversarial Content Exposure

SafetyDGX agent

arXiv:2605.00055v1 Announce Type: cross Abstract: We report a safety incident in a deployed multi-agent research system in which a primary AI agent installed 107 unauthorized software components, over

Analyzing Adversarial Inputs in Deep Reinforcement Learning

SafetyDGX agent

arXiv:2402.05284v2 Announce Type: replace Abstract: In recent years, Deep Reinforcement Learning (DRL) has become a popular paradigm in machine learning due to its successful applications to real-worl

ANO: A Principled Approach to Robust Policy Optimization

SafetyDGX agent

arXiv:2605.02320v1 Announce Type: cross Abstract: Proximal Policy Optimization (PPO) dominates deep RL but faces a fundamental dilemma. Its 'hard clipping' mechanism discards valuable gradient informa

Anomaly-Preference Image Generation

SafetyDGX agent

arXiv:2605.02439v1 Announce Type: new Abstract: Synthesizing realistic and diverse anomalous samples from limited data is vital for robust model generalization. However, existing methods struggle to r

Anticipation-VLA: Solving Long-Horizon Embodied Tasks via Anticipation-based Subgoal Generation

SafetyDGX agent

arXiv:2605.01772v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have emerged as a powerful paradigm for embodied intelligence, enabling robots to perform tasks based on natural l

ARGUS: Policy-Adaptive Ad Governance via Evolving Reinforcement with Adversarial Umpiring

SafetyDGX agent

arXiv:2605.02200v1 Announce Type: new Abstract: Online advertising governance faces significant challenges due to the non-stationary nature of regulatory policies, where emerging mandates (e.g., restr

Artificial intelligence language technologies in multilingual healthcare: Grand challenges ahead

SafetyDGX agent

arXiv:2605.01441v1 Announce Type: new Abstract: AI language technologies (AILTs), increasingly enabled by large language models (LLMs), are becoming embedded in multilingual healthcare workflows for t

Attention-Based Neural-Augmented Kalman Filter for Legged Robot State Estimation

SafetyDGX agent

arXiv:2601.18569v2 Announce Type: replace-cross Abstract: In this letter, we propose an Attention-Based Neural-Augmented Kalman Filter (AttenNKF) for state estimation in legged robots. Foot slip is a

Attention Sinks in Massively Multilingual Neural Machine Translation:Discovery, Analysis, and Mitigation

SafetyDGX agent

arXiv:2605.01229v1 Announce Type: cross Abstract: Cross-attention patterns in neural machine translation (NMT) are widely used to study how multilingual models align linguistic structure. We report a

Auditing demographic bias in AI-based emergency police dispatch: a cross-lingual evaluation of eleven large language models

SafetyDGX agent

arXiv:2605.01451v1 Announce Type: new Abstract: Large language models (LLMs) are rapidly being integrated into high-stakes public safety systems, including emergency call triage and dispatch decision

Autonomous Drift Learning in Data Streams: A Unified Perspective

SafetyDGX agent

arXiv:2605.01295v1 Announce Type: new Abstract: In the pursuit of autonomous learning systems, the foundational assumption of stationarity, the premise that data distributions and model behaviors rema

Autonomous Reliability Qualification of Ga_2O_3-based Hydrogen and Temperature Sensors via Safe Active Learning

SafetyDGX agent

arXiv:2605.00868v1 Announce Type: cross Abstract: We present a Safe Active Learning (SAL) framework for autonomous reliability characterization of rectifying Ga_2O_3-based devices under coupled therma

Behavior-Grounded Lane Representation Learning for Multi-Task Traffic Digital Twins

SafetyDGX agent

arXiv:2605.01901v1 Announce Type: new Abstract: Traffic digital twins are powerful tools for advanced traffic management, and most systems are built on static geometric representations. However, these

Beyond Crash: Hijacking Your Autonomous Vehicle for Fun and Profit

SafetyDGX agent

arXiv:2602.07249v2 Announce Type: replace-cross Abstract: Autonomous Vehicles (AVs), especially vision-based AVs, are rapidly being deployed without human operators. As AVs operate in safety-critical

Beyond Perceptual Shortcuts: Causal-Inspired Debiasing Optimization for Generalizable Video Reasoning in Lightweight MLLMs

SafetyDGX agent

arXiv:2605.01324v1 Announce Type: new Abstract: Although reinforcement learning (RL) has significantly advanced reasoning capabilities in large multimodal language models (MLLMs), its efficacy remains

Beyond Semantic Relevance: Counterfactual Risk Minimization for Robust Retrieval-Augmented Generation

SafetyDGX agent

arXiv:2605.01302v1 Announce Type: new Abstract: Standard Retrieval-Augmented Generation (RAG) systems predominantly rely on semantic relevance as a proxy for utility. However, this assumption collapse

Beyond Specialization: Robust Reinforcement Learning Navigation via Procedural Map Generators

SafetyDGX agent

arXiv:2605.02528v1 Announce Type: cross Abstract: Deep reinforcement learning (DRL) navigation policies often overfit to the structure of their training environments, as environmental diversity is typ

Bi-Level Reinforcement Learning Control for an Underactuated Blimp via Center-of-Mass Reconfiguration

SafetyDGX agent

arXiv:2605.01289v1 Announce Type: new Abstract: This paper investigates goal-directed tracking control of underactuated blimps with center-of-mass (CoM) reconfiguration. Unlike conventional overactuat

Binary Rewards and Reinforcement Learning: Fundamental Challenges

SafetyDGX agent

arXiv:2605.02375v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has become a standard approach for improving reasoning in language models, yet models trained with

Bolek: A Multimodal Language Model for Molecular Reasoning

SafetyDGX agent

arXiv:2605.02745v1 Announce Type: new Abstract: Molecular property models increasingly support high-stakes drug-discovery decisions, but their outputs are often difficult to audit: classical predictor

🚨 BOTH ALTMAN AND BROCKMAN SELF-DEALING ON CEREBRAS >Greg Brockman acquires personal Cerebras ownership in 2017 >Altman, separately, invest…

SafetyDGX agent

🚨 BOTH ALTMAN AND BROCKMAN SELF-DEALING ON CEREBRAS >Greg Brockman acquires personal Cerebras ownership in 2017 >Altman, separately, invests in Cerebras >Brockman pushes OpenAI to merge with Cerebras

Breaking the Computational Barrier: Provably Efficient Actor-Critic for Low-Rank MDPs

SafetyDGX agent

arXiv:2605.01242v1 Announce Type: new Abstract: Reinforcement learning (RL) is a fundamental framework for sequential decision-making, in which an agent learns an optimal policy through interactions w

Bridging the Gap Between Average and Discounted TD Learning

SafetyDGX agent

arXiv:2605.02103v1 Announce Type: new Abstract: The analysis of Temporal Difference (TD) learning in the average-reward setting faces notable theoretical difficulties because the Bellman operator is n

Bringing Order to Asynchronous SGD: Towards Optimality under Data-Dependent Delays with Momentum

SafetyDGX agent

arXiv:2605.02043v1 Announce Type: new Abstract: Asynchronous stochastic gradient descent (SGD) enables scalable distributed training but suffers from gradient staleness. Existing mitigation strategies

Brockman confirms that Sam was fired for not being consistently candid. Crazy that @karaswisher blocked me for saying that the board fired S…

SafetyDGX agent

Brockman confirms that Sam was fired for not being consistently candid. Crazy that @karaswisher blocked me for saying that the board fired Sam for not being consistently candid, when that is in fact w

Brockman’s counsel is doing a good job of laying out the timeline — but done little so far to refute yesterday’s dissection of her client’s …

SafetyDGX agent

Brockman’s counsel is doing a good job of laying out the timeline — but done little so far to refute yesterday’s dissection of her client’s self-dealing and dodgy behavior regarding his fiduciary resp

Causal Foundations of Collective Agency

SafetyDGX agent

arXiv:2605.00248v1 Announce Type: new Abstract: A key challenge for the safety of advanced AI systems is the possibility that multiple simpler agents might inadvertently form a collective agent with c

← Previous
1…163164165166167…212
Next →