AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,708 results
Safety

Towards Safer Large Reasoning Models by Promoting Safety Decision-Making before Chain-of-Thought Generation

DGX agent

arXiv:2603.17368v2 Announce Type: replace Abstract: Large reasoning models (LRMs) achieved remarkable performance via chain-of-thought (CoT), but recent studies showed that such enhanced reasoning cap

safetyarxiv-cs-ai
6 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

TRACE: A Metrologically-Grounded Engineering Framework for Trustworthy Agentic AI Systems in Operationally Critical Domains

DGX agent

arXiv:2605.03838v1 Announce Type: new Abstract: We introduce TRACE, a cross-domain engineering framework for trustworthy agentic AI in operationally critical domains. TRACE combines a four-layer refer

safetyarxiv-cs-cl
6 May 2026
Safety

Tracing Like a Clinician: Anatomy-Guided Spatial Priors for Cephalometric Landmark Detection

DGX agent

arXiv:2605.03358v1 Announce Type: new Abstract: When orthodontists trace cephalometric radiographs, they follow a structured workflow: identify the soft tissue profile, partition the skull into anatom

safetyarxiv-cs-cv
6 May 2026
Safety

truly scary: anybody, no matter how noble, can be destroyed by misused AI.

DGX agent

truly scary: anybody, no matter how noble, can be destroyed by misused AI. Artificial intelligence can create deceptive videos that rapidly damage a politician's public reputation. A new study reveals

safetygary-marcus--x
6 May 2026
Safety

Trustworthy AI Suffers from Invariance Conflicts and Causality is The Solution

DGX agent

arXiv:2605.02640v1 Announce Type: new Abstract: As artificial intelligence (AI), including machine learning (ML) models and foundation models (FMs), is increasingly deployed in high-stakes domains, en

safetyarxiv-cs-ai
6 May 2026
Safety

Uncertainty-Aware Trip Purpose Inference from GPS Trajectories via POI Semantic Zones and Pareto Calibration

DGX agent

arXiv:2605.01257v1 Announce Type: new Abstract: Large-scale GPS trajectory data offer rich observations of human mobility, yet assigning trip purposes to detected stops remains challenging due to the

safetyarxiv-cs-ai
6 May 2026
Safety

Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks

DGX agent

arXiv:2502.04419v3 Announce Type: replace-cross Abstract: Generating synthetic datasets via large language models (LLMs) has emerged as a promising approach to improve LLM performance. However, LLMs i

safetyarxiv-cs-cl
6 May 2026
Safety

Understanding Self-Supervised Learning via Latent Distribution Matching

DGX agent

arXiv:2605.03517v1 Announce Type: new Abstract: Self-supervised learning (SSL) excels at finding general-purpose latent representations from complex data, yet lacks a unifying theoretical framework th

safetyarxiv-cs-lg
6 May 2026
Safety

Uni-OPD: Unifying On-Policy Distillation with a Dual-Perspective Recipe

DGX agent

arXiv:2605.03677v1 Announce Type: new Abstract: On-policy distillation (OPD) has recently emerged as an effective post-training paradigm for consolidating the capabilities of specialized expert models

safetyarxiv-cs-lg
6 May 2026
Safety

Viewpoint-Agnostic Grasp Pipeline using VLM and Partial Observations

DGX agent

arXiv:2603.07866v2 Announce Type: replace-cross Abstract: Robust grasping in cluttered, unstructured environments remains challenging for mobile legged manipulators due to occlusions that lead to part

safetyarxiv-cs-lg
6 May 2026
Safety

What’s going on in semis is not sustainable per GS Everyone spending on AI is losing money except for semiconductor companies and the dynami…

DGX agent

What’s going on in semis is not sustainable per GS Everyone spending on AI is losing money except for semiconductor companies and the dynamic is “unprecedented and unsustainable” “Something has to cha

safetygary-marcus--x
6 May 2026
Safety

When Safety Geometry Collapses: Fine-Tuning Vulnerabilities in Agentic Guard Models

DGX agent

arXiv:2605.02914v1 Announce Type: new Abstract: A guard model fine-tuned on entirely benign data can lose all safety alignment -- not through adversarial manipulation, but through standard domain spec

safetyarxiv-cs-lg
6 May 2026
Safety

When to Think, When to Speak: Learning Disclosure Policies for LLM Reasoning

DGX agent

arXiv:2605.03314v1 Announce Type: new Abstract: In single-stream autoregressive interfaces, the same tokens both update the model state and constitute an irreversible public commitment. This coupling

safetyarxiv-cs-cl
6 May 2026
Safety

Will the Carbon Border Adjustment Mechanism Impact European Electricity Prices? A GNN-Based Network Analysis

DGX agent

arXiv:2605.03304v1 Announce Type: new Abstract: The European Union's Carbon Border Adjustment Mechanism (CBAM) creates a complex challenge for the interconnected European electricity market. Tradition

safetyarxiv-cs-lg
6 May 2026
Safety

yep, we are about to live in a world where nobody trusts anything. just like Putin always wanted. congratulations! ps it will suck

DGX agent

yep, we are about to live in a world where nobody trusts anything. just like Putin always wanted. congratulations! ps it will suck @GaryMarcus Or very soon no one will take any photo or video seriousl

safetygary-marcus--x
6 May 2026
Safety

Zero-Shot Signal Temporal Logic Planning with Disjunctive Branch Selection in Dynamic Semantic Maps

DGX agent

arXiv:2605.01222v1 Announce Type: new Abstract: Signal Temporal Logic (STL) offers verifiable task specifications and is crucial for safety-critical control. Yet STL planning remains challenging: exac

safetyarxiv-cs-ai
6 May 2026
Safety

A Deep Learning Model for Battery State Prediction towards Intelligent Energy Management

DGX agent

arXiv:2605.00898v1 Announce Type: cross Abstract: Accurate forecasting of battery health indicators, including remaining capacity and lifetime, is of paramount importance for ensuring the reliability,

safetyarxiv-cs-lg
5 May 2026
Safety

A Multi-View Media Profiling Suite: Resources, Evaluation, and Analysis

DGX agent

arXiv:2605.01336v1 Announce Type: new Abstract: News outlets shape public opinion at a scale that makes automated detection of political bias and factuality essential. However, the field still lacks u

safetyarxiv-cs-cl
5 May 2026
Safety

A Principled Approach for Creating High-fidelity Synthetic Demonstrations for Imitation Learning

DGX agent

arXiv:2605.01232v1 Announce Type: new Abstract: Recent advances in 3D Gaussian Splatting (3DGS) have enabled visually realistic demonstration generation from a single expert trajectory and a short mul

safetyarxiv-cs-ro
5 May 2026
Safety

A Theoretical Game of Attacks via Compositional Skills

DGX agent

arXiv:2605.01034v1 Announce Type: new Abstract: As large language models grow increasingly capable, concerns about their safe deployment have intensified. While numerous alignment strategies aim to re

safetyarxiv-cs-cl
5 May 2026
Safety

A Theory of Generalization in Deep Learning

DGX agent

arXiv:2605.01172v1 Announce Type: new Abstract: We present a non-asymptotic theory of generalization in deep learning where the empirical neural tangent kernel partitions the output space. In directio

safetyarxiv-cs-lg
5 May 2026
Safety

A Unified Multi-Dynamics Framework for Perception-Oriented Modeling in Tendon-Driven Continuum Robots

DGX agent

arXiv:2511.18088v2 Announce Type: replace Abstract: Tendon-driven continuum robots offer intrinsically safe and contact-rich interactions owing to their kinematic redundancy and structural compliance.

safetyarxiv-cs-ro
5 May 2026
Safety

Adaptive Alarm Threshold Prediction in 4G Mobile Networks: A Percentile-Guided Deep Learning Framework with Interpretable Outputs

DGX agent

arXiv:2605.00838v1 Announce Type: cross Abstract: In mobile telecommunications, alarms act as early warning signals. They are triggered when a cell, the basic unit of radio coverage, shuts down or beh

safetyarxiv-cs-lg
5 May 2026
Safety

Adaptive GoGI-Skip: Coupling Goal-Gradient Importance with Dynamic Uncertainty for Efficient Reasoning

DGX agent

arXiv:2505.08392v3 Announce Type: replace Abstract: Chain-of-Thought (CoT) prompting trades inference speed for reasoning accuracy. Existing compressors force a compromise as static gradient technique

safetyarxiv-cs-cl
5 May 2026
Safety

Adaptive Interpolation-Synthesis for Motion In-Betweening on Keyframe-Based Animation

DGX agent

arXiv:2605.02742v1 Announce Type: cross Abstract: Motion in-betweening is one of the most artistically demanding and time consuming stages of 3D animation, where the expressivity and rhythm of motion

safetyarxiv-cs-lg
5 May 2026
Safety

Adaptive Pluralistic Alignment: A pipeline for dynamic artificial democracy

DGX agent

arXiv:2605.01642v1 Announce Type: new Abstract: Prevailing alignment methods target a fixed set of preferences and therefore risk forcing value lock-in as societal norms evolve over time. We introduce

safetyarxiv-cs-lg
5 May 2026
Safety

Adversarial Imitation Learning with General Function Approximation: Theoretical Analysis and Practical Algorithms

DGX agent

arXiv:2605.01778v1 Announce Type: new Abstract: Adversarial imitation learning (AIL), a prominent approach in imitation learning, has achieved significant practical success powered by neural network a

safetyarxiv-cs-lg
5 May 2026
Safety

AFFormer: Adaptive Feature Fusion Transformer for V2X Cooperative Perception under Channel Impairments

DGX agent

arXiv:2605.01888v1 Announce Type: new Abstract: Accurate 3D object detection is essential for ensuring the safety of autonomous vehicles. Cooperative perception, which leverages vehicle-to-everything

safetyarxiv-cs-cv
5 May 2026
Safety

Agentic Learner with Grow-and-Refine Multimodal Semantic Memory

DGX agent

arXiv:2511.21678v2 Announce Type: replace-cross Abstract: MLLMs exhibit strong reasoning on isolated queries, yet they operate de novo -- solving each problem independently and often repeating the sam

safetyarxiv-cs-lg
5 May 2026
Safety

AgentReputation: A Decentralized Agentic AI Reputation Framework

DGX agent

arXiv:2605.00073v1 Announce Type: new Abstract: Decentralized, agentic AI marketplaces are rapidly emerging to support software engineering tasks such as debugging, patch generation, and security audi

safetyarxiv-cs-ai
5 May 2026
Safety

AI Alignment via Incentives and Correction

DGX agent

arXiv:2605.01643v1 Announce Type: new Abstract: We study AI alignment through the lens of law-and-economics models of deterrence and enforcement. In these models, misconduct is not treated as an exter

safetyarxiv-cs-lg
5 May 2026
Safety

AI music generator Udio just admitted in a court filing that it trained on audio scraped from YouTube videos. It’s already being sued by lab…

DGX agent

AI music generator Udio just admitted in a court filing that it trained on audio scraped from YouTube videos. It’s already being sued by labels & artists for what they say is copyright infringement on

safetygary-marcus--x
5 May 2026
Safety

Alex Karp now sounding like @garymarcus. Sooner or later, everyone figures it out.

DGX agent

Alex Karp now sounding like @garymarcus. Sooner or later, everyone figures it out. Palantir CEO Alex Karp goes after AI slop. The fight over AI “slop” is really a fight over whether software is perfor

safetygary-marcus--x
5 May 2026
Safety

ALIGNS: Unlocking nomological networks in psychological measurement through a large language model

DGX agent

arXiv:2509.09723v3 Announce Type: replace Abstract: Psychological measurement is critical to many disciplines. Despite advances in measurement, building nomological networks, theoretical maps of how c

safetyarxiv-cs-cl
5 May 2026
Safety

Ambient Persuasion in a Deployed AI Agent: Unauthorized Escalation Following Routine Non-Adversarial Content Exposure

DGX agent

arXiv:2605.00055v1 Announce Type: cross Abstract: We report a safety incident in a deployed multi-agent research system in which a primary AI agent installed 107 unauthorized software components, over

safetyarxiv-cs-ai
5 May 2026
Safety

Analyzing Adversarial Inputs in Deep Reinforcement Learning

DGX agent

arXiv:2402.05284v2 Announce Type: replace Abstract: In recent years, Deep Reinforcement Learning (DRL) has become a popular paradigm in machine learning due to its successful applications to real-worl

safetyarxiv-cs-lg
5 May 2026
Safety

ANO: A Principled Approach to Robust Policy Optimization

DGX agent

arXiv:2605.02320v1 Announce Type: cross Abstract: Proximal Policy Optimization (PPO) dominates deep RL but faces a fundamental dilemma. Its 'hard clipping' mechanism discards valuable gradient informa

safetyarxiv-cs-lg
5 May 2026
Safety

Anomaly-Preference Image Generation

DGX agent

arXiv:2605.02439v1 Announce Type: new Abstract: Synthesizing realistic and diverse anomalous samples from limited data is vital for robust model generalization. However, existing methods struggle to r

safetyarxiv-cs-cv
5 May 2026
Safety

Anticipation-VLA: Solving Long-Horizon Embodied Tasks via Anticipation-based Subgoal Generation

DGX agent

arXiv:2605.01772v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have emerged as a powerful paradigm for embodied intelligence, enabling robots to perform tasks based on natural l

safetyarxiv-cs-lg
5 May 2026
Safety

ARGUS: Policy-Adaptive Ad Governance via Evolving Reinforcement with Adversarial Umpiring

DGX agent

arXiv:2605.02200v1 Announce Type: new Abstract: Online advertising governance faces significant challenges due to the non-stationary nature of regulatory policies, where emerging mandates (e.g., restr

safetyarxiv-cs-cl
5 May 2026
Safety

Artificial intelligence language technologies in multilingual healthcare: Grand challenges ahead

DGX agent

arXiv:2605.01441v1 Announce Type: new Abstract: AI language technologies (AILTs), increasingly enabled by large language models (LLMs), are becoming embedded in multilingual healthcare workflows for t

safetyarxiv-cs-cl
5 May 2026
Safety

Attention-Based Neural-Augmented Kalman Filter for Legged Robot State Estimation

DGX agent

arXiv:2601.18569v2 Announce Type: replace-cross Abstract: In this letter, we propose an Attention-Based Neural-Augmented Kalman Filter (AttenNKF) for state estimation in legged robots. Foot slip is a

safetyarxiv-cs-lg
5 May 2026
Safety

Attention Sinks in Massively Multilingual Neural Machine Translation:Discovery, Analysis, and Mitigation

DGX agent

arXiv:2605.01229v1 Announce Type: cross Abstract: Cross-attention patterns in neural machine translation (NMT) are widely used to study how multilingual models align linguistic structure. We report a

safetyarxiv-cs-cl
5 May 2026
Safety

Auditing demographic bias in AI-based emergency police dispatch: a cross-lingual evaluation of eleven large language models

DGX agent

arXiv:2605.01451v1 Announce Type: new Abstract: Large language models (LLMs) are rapidly being integrated into high-stakes public safety systems, including emergency call triage and dispatch decision

safetyarxiv-cs-cl
5 May 2026
Safety

Autonomous Drift Learning in Data Streams: A Unified Perspective

DGX agent

arXiv:2605.01295v1 Announce Type: new Abstract: In the pursuit of autonomous learning systems, the foundational assumption of stationarity, the premise that data distributions and model behaviors rema

safetyarxiv-cs-lg
5 May 2026
Safety

Autonomous Reliability Qualification of Ga_2O_3-based Hydrogen and Temperature Sensors via Safe Active Learning

DGX agent

arXiv:2605.00868v1 Announce Type: cross Abstract: We present a Safe Active Learning (SAL) framework for autonomous reliability characterization of rectifying Ga_2O_3-based devices under coupled therma

safetyarxiv-cs-lg
5 May 2026
Safety

Behavior-Grounded Lane Representation Learning for Multi-Task Traffic Digital Twins

DGX agent

arXiv:2605.01901v1 Announce Type: new Abstract: Traffic digital twins are powerful tools for advanced traffic management, and most systems are built on static geometric representations. However, these

safetyarxiv-cs-cv
5 May 2026
Safety

Beyond Crash: Hijacking Your Autonomous Vehicle for Fun and Profit

DGX agent

arXiv:2602.07249v2 Announce Type: replace-cross Abstract: Autonomous Vehicles (AVs), especially vision-based AVs, are rapidly being deployed without human operators. As AVs operate in safety-critical

safetyarxiv-cs-lg
5 May 2026
← Previous
1…204205206207208…265
Next →