AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,357 results
11 Aug 2026

The Transparency Trap: How AI Disclaimers Create Overconfidence in High-Stakes Decisions

SafetyDGX agent

arXiv:2608.07493v1 Announce Type: cross Abstract: Current AI disclaimers often fail to function as intended due to warning habituation and a transparency paradox. As AI-generated information becomes p

The Voiceprint Fallacy: Why Voices Are Not Unique Biometric Imprints

SafetyDGX agent

arXiv:2608.07980v1 Announce Type: cross Abstract: In recent years, the term voiceprint has regained attention, particularly in technological applications and policy-making contexts, often carrying the

There are no lossless transformations of natural-language text

SafetyDGX agent

There are no lossless transformations of natural-language text Sophie Alpert shares her 'internal policy on acceptable use of AI writing by engineers'. It's a short read (supporting its own recommenda

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Three Necessary Principles for Self-Supervised Visual Representation Learning

SafetyDGX agent

arXiv:2608.08309v1 Announce Type: cross Abstract: We argue that learning visual representations without labels requires a training signal jointly complete across three non-overlapping objectives: sema

Topographic Constraints Shape Brain-Like Component Structure in Auditory Models

SafetyDGX agent

arXiv:2509.24039v2 Announce Type: replace-cross Abstract: If topography is a fundamental feature of the brain, it should influence both how neurons are arranged in space (i.e. explain brain maps) and

Towards Expressive and Faithful Audio-to-Image Generation: A Unified Multimodal Dataset and Synthesis Framework

SafetyDGX agent

arXiv:2608.09529v1 Announce Type: new Abstract: As an important subfield of cross-modal generation, synthesizing static visual content in the form of images from audio, namely audio-to-image (A2I) gen

Triple Expert Learning from Noisy Labels for Semi-Supervised Vision Foundation Model Adaptation

SafetyDGX agent

arXiv:2608.09052v1 Announce Type: cross Abstract: Semi-supervised adaptation of vision foundation models (VFMs) commonly freezes the pretrained backbone and updates lightweight modules such as LoRA. H

Unaccountable Delegation, Fading Skills: Mapping the Risks of Workplace AI Agents

SafetyDGX agent

arXiv:2608.08601v1 Announce Type: new Abstract: To anticipate socio-technical risks from AI agents, organizations need taxonomies to classify them. However, existing AI risk taxonomies focus on broad

Uncertainty-Aware Variational Reward Factorization via Probabilistic Preference Bases for LLM Personalization

SafetyDGX agent

arXiv:2604.00997v2 Announce Type: replace Abstract: Reward factorization personalizes large language models (LLMs) by decomposing rewards into shared basis functions and user-specific weights. Yet, ex

Understanding Reasoning from Pretraining to Post-Training

SafetyDGX agent

arXiv:2607.16097v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) has become central to improving large language models (LLMs) on complex reasoning tasks, yet RL post-training is l

Unimodality-Promoting Regularized Learning for Ordinal Regression

SafetyDGX agent

arXiv:2608.08359v1 Announce Type: new Abstract: Ordinal regression, also called ordinal classification, is classification of ordinal data, in which the underlying target variable is categorical and co

VADER: Adaptive Debiasing for Hallucination Mitigation in Video Large Language Models

SafetyDGX agent

arXiv:2608.08622v1 Announce Type: new Abstract: Large vision-language models (LVLMs) have demonstrated strong performance in open-ended video understanding, yet they remain prone to fluent responses u

VANE: Reliable Test-Time Training for Vision-Language-Action Models via Future Visual Representation Prediction

SafetyDGX agent

arXiv:2608.09448v1 Announce Type: cross Abstract: Test-time training (TTT) offers a lightweight way to adapt vision--language--action (VLA) policies from unlabeled deployment streams, but it remains d

Vid2WAM: Distilling Video Diffusion Priors into World Action Models

SafetyDGX agent

arXiv:2608.08558v1 Announce Type: new Abstract: World Action Models (WAMs) improve robot policy learning by jointly modeling future visual dynamics and actions. However, their scalability and generali

VoxZip: Semantic-Anchored Temporal KV Cache Compression for Long-Context Audio Inference

SafetyDGX agent

arXiv:2608.08569v1 Announce Type: new Abstract: Recent advancements in Speech Large Language Models have demonstrated remarkable capabilities in understanding complex audio tasks. Despite this progres

WDL-OPD: Weak-Driven On-Policy Distillation via Mixture-Constrained Co-Training

SafetyDGX agent

arXiv:2608.09447v1 Announce Type: cross Abstract: On-policy distillation (OPD) aligns a student with a teacher on trajectories sampled from the student itself, reducing the train-test state mismatch o

When Does Trace-Driven Evaluation Mislead MoE Expert Caching? Replay Semantics, Workload Contamination, and Operating Regimes

SafetyDGX agent

arXiv:2608.07911v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) models have outgrown accelerator memory, and offloading expert weights to host memory is now standard. This makes expert cache

Who Built This Model? Tracing LLM Lineage via Spectral Fingerprints in Weight Space

SafetyDGX agent

arXiv:2608.07786v1 Announce Type: new Abstract: Open-weight large language models (LLMs) are increasingly developed through complex, multi-stage pipelines, leading to intricate lineage relationships t

World Tokens: Enhancing Embodied Policies with Training-Time World Modeling

SafetyDGX agent

arXiv:2608.09730v1 Announce Type: new Abstract: Vision-language-action (VLA) models are a widely adopted paradigm for embodied policies. They excel at efficient closed-loop control but do not explicit

10 Aug 2026

A Practical Evaluation Method for Long-Form Simultaneous Speech-to-Speech Translation

SafetyDGX agent

arXiv:2606.15059v2 Announce Type: replace Abstract: Simultaneous speech-to-speech translation (SimulS2ST) enables real-time cross-lingual communication, but existing evaluation has focused largely on

An AI4AI Framework for Visual Token Pruning

SafetyDGX agent

arXiv:2608.07193v1 Announce Type: cross Abstract: Visual-token pruning can substantially reduce the inference cost of multimodal large language models (MLLMs), yet existing methods largely rely on fix

AutoIntervene: Calibrated Intervention for Action-Chunking Imitation Learning Policies

SafetyDGX agent

arXiv:2608.07065v1 Announce Type: cross Abstract: Action-chunking visuomotor policies learn from demonstrations and improve temporal consistency by predicting short action sequences rather than single

Automated item evaluation: Predicting item acceptance and rejection using LLM-generated critiques

SafetyDGX agent

arXiv:2608.06609v1 Announce Type: new Abstract: Automated item evaluation (AIE) refers to the use of computational methods to assess item quality without requiring manual expert review or field testin

Automated Terminal-to-Housing Assembly System for Flat Ribbon Cable Harness

SafetyDGX agent

arXiv:2608.06996v1 Announce Type: new Abstract: This paper presents a sensor-minimal automated assembly system for bidirectional single-row flat ribbon cable harnesses (FRCHs). Unlike conventional peg

Beyond Post-Hoc Temperature Scaling: Bilevel Optimization for LLM Calibration

SafetyDGX agent

arXiv:2608.07419v1 Announce Type: new Abstract: Preference alignment often makes large language models (LLMs) overconfident and poorly calibrated. Traditional post-hoc temperature scaling is inherentl

Bootstrap-Conditioned Action Selection with Tabular Foundation Models

SafetyDGX agent

arXiv:2608.06559v1 Announce Type: new Abstract: Contextual bandits offer a natural framework for sample-efficient personalization, but practical deployment remains difficult under sparse, biased inter

Bypassing Krum: Selection-Aware Backdoor Attacks in Federated Learning

SafetyDGX agent

arXiv:2608.06637v1 Announce Type: cross Abstract: Robust aggregation methods are widely used in federated learning to mitigate the impact of adversarial client behavior. Distance-based aggregation rul

Calibrating WEAT Against Anisotropy: ZCA Whitening as a Geometric Pre-Processing Step for Embedding Association Tests

SafetyDGX agent

arXiv:2608.06908v1 Announce Type: cross Abstract: We propose Zero-phase Component Analysis (ZCA) whitening as a geometric pre-processing step for the Word Embedding Association Test (WEAT). WEAT is a

Cascade: Exploiting SLO-Aware latency budget for fair and high goodput LLM inference serving

SafetyDGX agent

arXiv:2608.06557v1 Announce Type: cross Abstract: The reasoning and agentic capabilities of large language models have expanded the range of applications they support, from short interactive exchanges

CEDAR: Agent-Orchestrated Tree Search for Goal-Directed Optimization of Complex Systems

SafetyDGX agent

arXiv:2608.06871v1 Announce Type: new Abstract: Complex systems, core objects of study in artificial life, model diverse phenomena through nonlinear, feedback-driven interactions that produce emergent

CHIME: A Case for Efficient Long-Context Attention-FC Disaggregated Inference with DIMM-PIM

SafetyDGX agent

arXiv:2504.17584v2 Announce Type: replace-cross Abstract: Attention-FC Disaggregated (AFD) LLM inference systems offload memory-bound Attention operations to memory-rich accelerators (e.g., CPUs, HBM-

Confidence Estimation for Financial Vision-Language Models in Chart and Document Understanding

SafetyDGX agent

arXiv:2608.06532v1 Announce Type: new Abstract: LVLMs are increasingly used to read financial charts, tables, and documents, where a single misread figure can move a decision and the most authoritativ

Counterfactual Shapley Credit Assignment

SafetyDGX agent

arXiv:2607.16999v2 Announce Type: replace-cross Abstract: The Credit Assignment Problem (CAP) is fundamental to developing efficient and explainable Reinforcement Learning (RL) agents. Existing framew

CreativeInstruct: Scalably Teaching LLMs to Balance Quality, Creativity, and Diversity

SafetyDGX agent

arXiv:2608.07460v1 Announce Type: cross Abstract: While post-training improves the capabilities of large language models (LLMs), it generally lowers their output diversity and creativity, negatively i

CrystalGRPO: Target-Aligned and Coverage-Preserving Reinforcement Learning for Flow-Based Crystal Structure Prediction

SafetyDGX agent

arXiv:2608.06582v1 Announce Type: new Abstract: Flow-based generative models can efficiently produce candidate structures for crystal structure prediction (CSP), but their pretrained objectives do not

Density-Functional Excited-State Gradients and Nonadiabatic Couplings on a Consumer GPU from a Contraction-DAG

SafetyDGX agent

arXiv:2608.06536v1 Announce Type: cross Abstract: Nonadiabatic dynamics needs an excited-state gradient and an interstate nonadiabatic coupling matrix element (NACME) at every nuclear geometry, and a

DiDPO: Diff-in-Diff Policy Optimization for Coding Agent Training

SafetyDGX agent

arXiv:2608.07147v1 Announce Type: new Abstract: Reinforcement learning with Verifiable Reward (RLVR) has emerged as a powerful paradigm for training coding agents, where the execution feedback from co

Do 3D Medical Foundation Models See Through MRI Artifacts? A Controlled Study of Representation Robustness

SafetyDGX agent

arXiv:2608.06613v1 Announce Type: cross Abstract: Self-supervised 3D medical foundation models are increasingly used as general-purpose feature extractors, yet their sensitivity to MRI artifacts remai

Does Splitting a Triage Decision Across Agents Hide Bias or Help Catch It? A Multi-Agent Simulation Study of LLM-Based Resource Allocation Under Audit Capacity Constraints

SafetyDGX agent

arXiv:2608.06949v1 Announce Type: new Abstract: Prior benchmarking work has shown that a single large language model (LLM), forced to make life-or-death resource-allocation decisions, exhibits measura

Dual-Space Modality Consistency Learning for Universal Cross-Modal Re-Identification

SafetyDGX agent

arXiv:2608.06943v1 Announce Type: new Abstract: Cross-modal Re-Identification (ReID) aims to retrieve the same identity across heterogeneous imaging modalities and has been widely studied in visible-i

Explore or Converge? Stage-Guided Per-Step Optimization for Diffusion Models

SafetyDGX agent

arXiv:2608.06768v1 Announce Type: new Abstract: Diffusion models have strong generative capabilities. However, their maximum likelihood training objective only focuses on reconstructing the data distr

From Optimal Actions to World Models: Identifiability of Transition Kernels in Discounted MDPs

SafetyDGX agent

arXiv:2608.07301v1 Announce Type: new Abstract: We study what can be recovered about the transition probabilities of a Markov decision process from optimal actions alone. This is closely related to th

Gated-BEPO: Confidence-Gated Bellman Credit Assignment for Large Language Model Agents

SafetyDGX agent

arXiv:2608.06861v1 Announce Type: new Abstract: Training large language model agents in long-horizon environments requires assigning credit from sparse terminal outcomes to individual actions. Existin

GOPI: Generation-Oriented 3D Pose Inference for Furniture Insertion from Single-View RGB-D Indoor Scenes

SafetyDGX agent

arXiv:2608.06836v1 Announce Type: new Abstract: We study the problem of inserting new furniture into indoor scene images. Under masked single-view 2D image-plane conditioning, however, the physical sc

Grad-ECLIP: Gradient-based Visual and Textual Explanations for CLIP

SafetyDGX agent

arXiv:2502.18816v3 Announce Type: replace Abstract: Significant progress has been achieved on the improvement and downstream usages of the Contrastive Language-Image Pre-training (CLIP) vision-languag

Graph Machine: Exploring Edge Mechanisms as an Inductive Bias

SafetyDGX agent

arXiv:2608.06834v1 Announce Type: new Abstract: Transformers provide a powerful architecture for global content-based matching, but reasoning problems may benefit from a stronger inductive bias toward

High-dimensional ridgeless least squares interpolation under spiked covariance structures

SafetyDGX agent

arXiv:2608.07281v1 Announce Type: cross Abstract: This paper investigates the asymptotic behavior of the out-of-sample prediction risk of the high-dimensional ridgeless least-squares estimator when th

How Much, Then Where: Credit-Conserving Action-to-Token Allocation for Multi-Turn Agent Reinforcement Learning

SafetyDGX agent

arXiv:2608.07118v1 Announce Type: new Abstract: Credit assignment in multi-turn agent reinforcement learning operates at two levels: assigning trajectory-level credit to actions and distributing each

How WPP operationalizes platform and data engineering for AI marketing

SafetyDGX agent

Between chaotic levels of market fragmentation and economic volatility, marketing and communications agencies can no longer rely on the human intuition they’ve traditionally used to win clients and op

Identity as Presence: Towards Appearance and Voice Personalized Joint Audio-Video Generation

SafetyDGX agent

arXiv:2603.17889v4 Announce Type: replace Abstract: Recent advances in video synthesis have enabled realistic integration of real individuals, driving demand for identity-aware generation. While emerg

Improving Performance of Spike-based Deep Q-Learning using Ternary Neurons

SafetyDGX agent

arXiv:2506.03392v2 Announce Type: replace Abstract: We propose a new ternary spiking neuron model to improve the representation capacity of binary spiking neurons in deep Q-learning. Although a ternar

Interpretable reinforcement learning with decision-tree pruning

SafetyDGX agent

arXiv:2608.07151v1 Announce Type: cross Abstract: Reinforcement learning policies are difficult to inspect, but interpreting them is a prerequisite for trustworthiness. Converting a trained policy int

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models

SafetyDGX agent

arXiv:2608.06799v1 Announce Type: cross Abstract: Learning structured and control-relevant latent representations remains a key challenge for world models. Recent JEPA-based world models learn action-

KReF: Training-Free Retrieval for Long-Term Time-Series Forecasting and Predictive Uncertainty

SafetyDGX agent

arXiv:2608.06748v1 Announce Type: cross Abstract: Probabilistic long-term time-series forecasting commonly relies on trained models. Training-free conformal methods typically construct intervals aroun

Learning Suffers More Than the Policy Class Under Partial Observability: A Closed-Form Analysis

SafetyDGX agent

arXiv:2608.07228v1 Announce Type: new Abstract: When a reinforcement learning agent cannot observe the full state, we usually blame its policies: it cannot see enough to represent a good one. We show

Learning to Walk With Less: A Dyna-Style Approach to Quadrupedal Locomotion

SafetyDGX agent

arXiv:2509.06296v2 Announce Type: replace-cross Abstract: Traditional on-policy reinforcement learning (RL) controllers for quadrupedal locomotion often suffer from low data efficiency, requiring mill

Let's Unlearn Stereotypes Before Decision-Making: Assessing the Impact of Intrinsic Bias Mitigation on Downstream Fairness in LLMs

SafetyDGX agent

arXiv:2509.16462v2 Announce Type: replace Abstract: Large Language Models (LLMs) are increasingly used in high-stakes decision-making systems, where biased predictions can reinforce social and economi

LMM Modality Transfer: A Pre-requisite for Autonomous GIS Agents

SafetyDGX agent

arXiv:2608.06948v1 Announce Type: new Abstract: AI models are becoming increasingly adept at understanding and processing spatial information, thereby facilitating agentic problem-solving in spatial t

LyEvO: Lyapunov-Guided Evolutionary Optimization for Safe and Robust Sim-to-Real Policy Learning

SafetyDGX agent

arXiv:2608.06481v1 Announce Type: cross Abstract: Training controllers that are safe and robust in simulation, and systematically assessing their readiness for real-world deployment, remain key challe

MaskFlow: Precise, Consistent and Seamless Regional Image Editing

SafetyDGX agent

arXiv:2608.06929v1 Announce Type: cross Abstract: Regional image editing has attracted considerable attention for its spatial controllability. Although instruction-based and mask-reference-based editi

← Previous
1…5657585960…240
Next →