AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,357 results
6 May 2026

SoDa2: Single-Stage Open-Set Domain Adaptation via Decoupled Alignment for Cross-Scene Hyperspectral Image Classification

SafetyDGX agent

arXiv:2605.03371v1 Announce Type: new Abstract: Cross-scene hyperspectral image (HSI) classification stands as a fundamental research topic in remote sensing, with extensive applications spanning vari

Stream-R1: Reliability-Perplexity Aware Reward Distillation for Streaming Video Generation

SafetyDGX agent

arXiv:2605.03849v1 Announce Type: new Abstract: Distillation-based acceleration has become foundational for making autoregressive streaming video diffusion models practical, with distribution matching

Structured Diffusion Bridges: Inductive Bias for Denoising Diffusion Bridges

SafetyDGX agent

arXiv:2605.02973v1 Announce Type: new Abstract: Modality translation is inherently under-constrained, as multiple cross-modal mappings may yield the same marginals. Recent work has shown that diffusio

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

T^2PO: Uncertainty-Guided Exploration Control for Stable Multi-Turn Agentic Reinforcement Learning

SafetyDGX agent

arXiv:2605.02178v1 Announce Type: new Abstract: Recent progress in multi-turn reinforcement learning (RL) has significantly improved reasoning LLMs' performances on complex interactive tasks. Despite

Talk is Cheap, Communication is Hard: Dynamic Grounding Failures and Repair in Multi-Agent Negotiation

SafetyDGX agent

arXiv:2605.01750v1 Announce Type: cross Abstract: Grounding is the collaborative process of establishing mutual belief sufficient for the current communicative purpose. While static grounding maps lan

TeamUp: Semantic Project Matching and Team Formation for Learning at Scale

SafetyDGX agent

arXiv:2605.03237v1 Announce Type: cross Abstract: Project-based learning improves student engagement and learning outcomes, yet allocating students to appropriately challenging projects while forming

The Design and Composition of Structural Causal Decision Processes

SafetyDGX agent

arXiv:2605.02681v1 Announce Type: cross Abstract: We present two new classes of causal models of decision-making agents. Our approach is motivated by the needs of modeling the economics of computing s

The Garden of Forking Paths: Narrative Arc-Conditioned Gameplay Planning

SafetyDGX agent

arXiv:2605.01245v1 Announce Type: cross Abstract: Narrative archetypes (e.g., Hero's Journey, Three-act structure) provide universal story structures that resonate across cultures and media and are im

TRACE: A Metrologically-Grounded Engineering Framework for Trustworthy Agentic AI Systems in Operationally Critical Domains

SafetyDGX agent

arXiv:2605.03838v1 Announce Type: new Abstract: We introduce TRACE, a cross-domain engineering framework for trustworthy agentic AI in operationally critical domains. TRACE combines a four-layer refer

Tracing Like a Clinician: Anatomy-Guided Spatial Priors for Cephalometric Landmark Detection

SafetyDGX agent

arXiv:2605.03358v1 Announce Type: new Abstract: When orthodontists trace cephalometric radiographs, they follow a structured workflow: identify the soft tissue profile, partition the skull into anatom

truly scary: anybody, no matter how noble, can be destroyed by misused AI.

SafetyDGX agent

truly scary: anybody, no matter how noble, can be destroyed by misused AI. Artificial intelligence can create deceptive videos that rapidly damage a politician's public reputation. A new study reveals

Trustworthy AI Suffers from Invariance Conflicts and Causality is The Solution

SafetyDGX agent

arXiv:2605.02640v1 Announce Type: new Abstract: As artificial intelligence (AI), including machine learning (ML) models and foundation models (FMs), is increasingly deployed in high-stakes domains, en

Uncertainty-Aware Trip Purpose Inference from GPS Trajectories via POI Semantic Zones and Pareto Calibration

SafetyDGX agent

arXiv:2605.01257v1 Announce Type: new Abstract: Large-scale GPS trajectory data offer rich observations of human mobility, yet assigning trip purposes to detected stops remains challenging due to the

Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks

SafetyDGX agent

arXiv:2502.04419v3 Announce Type: replace-cross Abstract: Generating synthetic datasets via large language models (LLMs) has emerged as a promising approach to improve LLM performance. However, LLMs i

Understanding Self-Supervised Learning via Latent Distribution Matching

SafetyDGX agent

arXiv:2605.03517v1 Announce Type: new Abstract: Self-supervised learning (SSL) excels at finding general-purpose latent representations from complex data, yet lacks a unifying theoretical framework th

Uni-OPD: Unifying On-Policy Distillation with a Dual-Perspective Recipe

SafetyDGX agent

arXiv:2605.03677v1 Announce Type: new Abstract: On-policy distillation (OPD) has recently emerged as an effective post-training paradigm for consolidating the capabilities of specialized expert models

What’s going on in semis is not sustainable per GS Everyone spending on AI is losing money except for semiconductor companies and the dynami…

SafetyDGX agent

What’s going on in semis is not sustainable per GS Everyone spending on AI is losing money except for semiconductor companies and the dynamic is “unprecedented and unsustainable” “Something has to cha

When to Think, When to Speak: Learning Disclosure Policies for LLM Reasoning

SafetyDGX agent

arXiv:2605.03314v1 Announce Type: new Abstract: In single-stream autoregressive interfaces, the same tokens both update the model state and constitute an irreversible public commitment. This coupling

Will the Carbon Border Adjustment Mechanism Impact European Electricity Prices? A GNN-Based Network Analysis

SafetyDGX agent

arXiv:2605.03304v1 Announce Type: new Abstract: The European Union's Carbon Border Adjustment Mechanism (CBAM) creates a complex challenge for the interconnected European electricity market. Tradition

yep, we are about to live in a world where nobody trusts anything. just like Putin always wanted. congratulations! ps it will suck

SafetyDGX agent

yep, we are about to live in a world where nobody trusts anything. just like Putin always wanted. congratulations! ps it will suck @GaryMarcus Or very soon no one will take any photo or video seriousl

5 May 2026

A Multi-View Media Profiling Suite: Resources, Evaluation, and Analysis

SafetyDGX agent

arXiv:2605.01336v1 Announce Type: new Abstract: News outlets shape public opinion at a scale that makes automated detection of political bias and factuality essential. However, the field still lacks u

A Principled Approach for Creating High-fidelity Synthetic Demonstrations for Imitation Learning

SafetyDGX agent

arXiv:2605.01232v1 Announce Type: new Abstract: Recent advances in 3D Gaussian Splatting (3DGS) have enabled visually realistic demonstration generation from a single expert trajectory and a short mul

A Theoretical Game of Attacks via Compositional Skills

SafetyDGX agent

arXiv:2605.01034v1 Announce Type: new Abstract: As large language models grow increasingly capable, concerns about their safe deployment have intensified. While numerous alignment strategies aim to re

A Theory of Generalization in Deep Learning

SafetyDGX agent

arXiv:2605.01172v1 Announce Type: new Abstract: We present a non-asymptotic theory of generalization in deep learning where the empirical neural tangent kernel partitions the output space. In directio

A Unified Multi-Dynamics Framework for Perception-Oriented Modeling in Tendon-Driven Continuum Robots

SafetyDGX agent

arXiv:2511.18088v2 Announce Type: replace Abstract: Tendon-driven continuum robots offer intrinsically safe and contact-rich interactions owing to their kinematic redundancy and structural compliance.

Adaptive Alarm Threshold Prediction in 4G Mobile Networks: A Percentile-Guided Deep Learning Framework with Interpretable Outputs

SafetyDGX agent

arXiv:2605.00838v1 Announce Type: cross Abstract: In mobile telecommunications, alarms act as early warning signals. They are triggered when a cell, the basic unit of radio coverage, shuts down or beh

Adaptive GoGI-Skip: Coupling Goal-Gradient Importance with Dynamic Uncertainty for Efficient Reasoning

SafetyDGX agent

arXiv:2505.08392v3 Announce Type: replace Abstract: Chain-of-Thought (CoT) prompting trades inference speed for reasoning accuracy. Existing compressors force a compromise as static gradient technique

Adaptive Interpolation-Synthesis for Motion In-Betweening on Keyframe-Based Animation

SafetyDGX agent

arXiv:2605.02742v1 Announce Type: cross Abstract: Motion in-betweening is one of the most artistically demanding and time consuming stages of 3D animation, where the expressivity and rhythm of motion

Adaptive Pluralistic Alignment: A pipeline for dynamic artificial democracy

SafetyDGX agent

arXiv:2605.01642v1 Announce Type: new Abstract: Prevailing alignment methods target a fixed set of preferences and therefore risk forcing value lock-in as societal norms evolve over time. We introduce

Adversarial Imitation Learning with General Function Approximation: Theoretical Analysis and Practical Algorithms

SafetyDGX agent

arXiv:2605.01778v1 Announce Type: new Abstract: Adversarial imitation learning (AIL), a prominent approach in imitation learning, has achieved significant practical success powered by neural network a

Agentic Learner with Grow-and-Refine Multimodal Semantic Memory

SafetyDGX agent

arXiv:2511.21678v2 Announce Type: replace-cross Abstract: MLLMs exhibit strong reasoning on isolated queries, yet they operate de novo -- solving each problem independently and often repeating the sam

AI Alignment via Incentives and Correction

SafetyDGX agent

arXiv:2605.01643v1 Announce Type: new Abstract: We study AI alignment through the lens of law-and-economics models of deterrence and enforcement. In these models, misconduct is not treated as an exter

AI music generator Udio just admitted in a court filing that it trained on audio scraped from YouTube videos. It’s already being sued by lab…

SafetyDGX agent

AI music generator Udio just admitted in a court filing that it trained on audio scraped from YouTube videos. It’s already being sued by labels & artists for what they say is copyright infringement on

Alex Karp now sounding like @garymarcus. Sooner or later, everyone figures it out.

SafetyDGX agent

Alex Karp now sounding like @garymarcus. Sooner or later, everyone figures it out. Palantir CEO Alex Karp goes after AI slop. The fight over AI “slop” is really a fight over whether software is perfor

ALIGNS: Unlocking nomological networks in psychological measurement through a large language model

SafetyDGX agent

arXiv:2509.09723v3 Announce Type: replace Abstract: Psychological measurement is critical to many disciplines. Despite advances in measurement, building nomological networks, theoretical maps of how c

ANO: A Principled Approach to Robust Policy Optimization

SafetyDGX agent

arXiv:2605.02320v1 Announce Type: cross Abstract: Proximal Policy Optimization (PPO) dominates deep RL but faces a fundamental dilemma. Its 'hard clipping' mechanism discards valuable gradient informa

Anomaly-Preference Image Generation

SafetyDGX agent

arXiv:2605.02439v1 Announce Type: new Abstract: Synthesizing realistic and diverse anomalous samples from limited data is vital for robust model generalization. However, existing methods struggle to r

Anticipation-VLA: Solving Long-Horizon Embodied Tasks via Anticipation-based Subgoal Generation

SafetyDGX agent

arXiv:2605.01772v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have emerged as a powerful paradigm for embodied intelligence, enabling robots to perform tasks based on natural l

ARGUS: Policy-Adaptive Ad Governance via Evolving Reinforcement with Adversarial Umpiring

SafetyDGX agent

arXiv:2605.02200v1 Announce Type: new Abstract: Online advertising governance faces significant challenges due to the non-stationary nature of regulatory policies, where emerging mandates (e.g., restr

Attention-Based Neural-Augmented Kalman Filter for Legged Robot State Estimation

SafetyDGX agent

arXiv:2601.18569v2 Announce Type: replace-cross Abstract: In this letter, we propose an Attention-Based Neural-Augmented Kalman Filter (AttenNKF) for state estimation in legged robots. Foot slip is a

Attention Sinks in Massively Multilingual Neural Machine Translation:Discovery, Analysis, and Mitigation

SafetyDGX agent

arXiv:2605.01229v1 Announce Type: cross Abstract: Cross-attention patterns in neural machine translation (NMT) are widely used to study how multilingual models align linguistic structure. We report a

Autonomous Drift Learning in Data Streams: A Unified Perspective

SafetyDGX agent

arXiv:2605.01295v1 Announce Type: new Abstract: In the pursuit of autonomous learning systems, the foundational assumption of stationarity, the premise that data distributions and model behaviors rema

Behavior-Grounded Lane Representation Learning for Multi-Task Traffic Digital Twins

SafetyDGX agent

arXiv:2605.01901v1 Announce Type: new Abstract: Traffic digital twins are powerful tools for advanced traffic management, and most systems are built on static geometric representations. However, these

Beyond Perceptual Shortcuts: Causal-Inspired Debiasing Optimization for Generalizable Video Reasoning in Lightweight MLLMs

SafetyDGX agent

arXiv:2605.01324v1 Announce Type: new Abstract: Although reinforcement learning (RL) has significantly advanced reasoning capabilities in large multimodal language models (MLLMs), its efficacy remains

Beyond Specialization: Robust Reinforcement Learning Navigation via Procedural Map Generators

SafetyDGX agent

arXiv:2605.02528v1 Announce Type: cross Abstract: Deep reinforcement learning (DRL) navigation policies often overfit to the structure of their training environments, as environmental diversity is typ

Bi-Level Reinforcement Learning Control for an Underactuated Blimp via Center-of-Mass Reconfiguration

SafetyDGX agent

arXiv:2605.01289v1 Announce Type: new Abstract: This paper investigates goal-directed tracking control of underactuated blimps with center-of-mass (CoM) reconfiguration. Unlike conventional overactuat

Binary Rewards and Reinforcement Learning: Fundamental Challenges

SafetyDGX agent

arXiv:2605.02375v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has become a standard approach for improving reasoning in language models, yet models trained with

Bolek: A Multimodal Language Model for Molecular Reasoning

SafetyDGX agent

arXiv:2605.02745v1 Announce Type: new Abstract: Molecular property models increasingly support high-stakes drug-discovery decisions, but their outputs are often difficult to audit: classical predictor

🚨 BOTH ALTMAN AND BROCKMAN SELF-DEALING ON CEREBRAS >Greg Brockman acquires personal Cerebras ownership in 2017 >Altman, separately, invest…

SafetyDGX agent

🚨 BOTH ALTMAN AND BROCKMAN SELF-DEALING ON CEREBRAS >Greg Brockman acquires personal Cerebras ownership in 2017 >Altman, separately, invests in Cerebras >Brockman pushes OpenAI to merge with Cerebras

Breaking the Computational Barrier: Provably Efficient Actor-Critic for Low-Rank MDPs

SafetyDGX agent

arXiv:2605.01242v1 Announce Type: new Abstract: Reinforcement learning (RL) is a fundamental framework for sequential decision-making, in which an agent learns an optimal policy through interactions w

Bridging the Gap Between Average and Discounted TD Learning

SafetyDGX agent

arXiv:2605.02103v1 Announce Type: new Abstract: The analysis of Temporal Difference (TD) learning in the average-reward setting faces notable theoretical difficulties because the Bellman operator is n

Bringing Order to Asynchronous SGD: Towards Optimality under Data-Dependent Delays with Momentum

SafetyDGX agent

arXiv:2605.02043v1 Announce Type: new Abstract: Asynchronous stochastic gradient descent (SGD) enables scalable distributed training but suffers from gradient staleness. Existing mitigation strategies

Brockman confirms that Sam was fired for not being consistently candid. Crazy that @karaswisher blocked me for saying that the board fired S…

SafetyDGX agent

Brockman confirms that Sam was fired for not being consistently candid. Crazy that @karaswisher blocked me for saying that the board fired Sam for not being consistently candid, when that is in fact w

Brockman’s counsel is doing a good job of laying out the timeline — but done little so far to refute yesterday’s dissection of her client’s …

SafetyDGX agent

Brockman’s counsel is doing a good job of laying out the timeline — but done little so far to refute yesterday’s dissection of her client’s self-dealing and dodgy behavior regarding his fiduciary resp

Colinearity Decay: Training Quantization-Friendly ViTs with Outlier Decay

SafetyDGX agent

arXiv:2605.01330v1 Announce Type: new Abstract: Low-bit quantization is a practical route for efficiently deploying vision Transformers, yet activation outliers complicate fully quantized deployment.

Combining Trained Models in Reinforcement Learning

SafetyDGX agent

arXiv:2605.02159v1 Announce Type: new Abstract: Deep reinforcement learning (DRL) has delivered strong results in domains such as Atari and Go, but it still suffers from high sample cost and weak tran

Compared to What? Baselines and Metrics for Counterfactual Prompting

SafetyDGX agent

arXiv:2605.01048v1 Announce Type: new Abstract: Counterfactual prompting (i.e., perturbing a single factor and measuring output change) is widely used to evaluate things like LLM bias and CoT faithful

Compliance-Aware Agentic Payments on Stablecoin Rails

SafetyDGX agent

arXiv:2605.00071v1 Announce Type: cross Abstract: Agentic payment systems extend delegated action to financial transfers, but scaling them on stablecoin rails in regulated settings requires safeguards

Contrastive Residual Energy Test-time Adaptation

SafetyDGX agent

arXiv:2505.19607v2 Announce Type: replace Abstract: Test-time adaptation (TTA) enhances model robustness by enabling adaptation to target distributions that differ from training distributions, improvi

CUE: Concept-Aware Multi-Label Expansion to Mitigate Concept Confusion in Long-Tailed Learning

SafetyDGX agent

arXiv:2605.01309v1 Announce Type: new Abstract: Long-tailed distributions are common in real-world recognition tasks, where a few head classes have many samples while most tail classes have very few.

← Previous
1…184185186187188…240
Next →