AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,356 results
24 Apr 2026

Supervised Learning Has a Necessary Geometric Blind Spot: Theory, Consequences, and Minimal Repair

SafetyDGX agent

arXiv:2604.21395v1 Announce Type: cross Abstract: We prove that empirical risk minimisation (ERM) imposes a necessary geometric constraint on learned representations: any encoder that minimises superv

Tempered Sequential Monte Carlo for Trajectory and Policy Optimization with Differentiable Dynamics

SafetyDGX agent

arXiv:2604.21456v1 Announce Type: new Abstract: We propose a sampling-based framework for finite-horizon trajectory and policy optimization under differentiable dynamics by casting controller design a

Temporal Prototyping and Hierarchical Alignment for Unsupervised Video-based Visible-Infrared Person Re-Identification

SafetyDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.21324v1 Announce Type: new Abstract: Visible-infrared person re-identification (VI-ReID) enables cross-modality identity matching for all-day surveillance, yet existing methods predominantl

The Effect of Idea Elaboration on the Automatic Assessment of Idea Originality

SafetyDGX agent

arXiv:2604.20569v1 Announce Type: cross Abstract: Automatic systems are increasingly used to assess the originality of responses in creative tasks. They offer a potential solution to key limitations o

'This Wasn't Made for Me': Recentering User Experience and Emotional Impact in the Evaluation of ASR Bias

SafetyDGX agent

arXiv:2604.21148v1 Announce Type: new Abstract: Studies on bias in Automatic Speech Recognition (ASR) tend to focus on reporting error rates for speakers of underrepresented dialects, yet less researc

Time, Causality, and Observability Failures in Distributed AI Inference Systems

SafetyDGX agent

arXiv:2604.21361v1 Announce Type: new Abstract: Distributed AI inference pipelines rely heavily on timestamp-based observability to understand system behavior. This work demonstrates that even small c

Trust-SSL: Additive-Residual Selective Invariance for Robust Aerial Self-Supervised Learning

SafetyDGX agent

arXiv:2604.21349v1 Announce Type: cross Abstract: Self-supervised learning (SSL) is a standard approach for representation learning in aerial imagery. Existing methods enforce invariance between augme

UniGenDet: A Unified Generative-Discriminative Framework for Co-Evolutionary Image Generation and Generated Image Detection

SafetyDGX agent

arXiv:2604.21904v1 Announce Type: new Abstract: In recent years, significant progress has been made in both image generation and generated image detection. Despite their rapid, yet largely independent

VFM-VAE: Vision Foundation Models Can Be Good Tokenizers for Latent Diffusion Models

SafetyDGX agent

arXiv:2510.18457v3 Announce Type: replace Abstract: The performance of Latent Diffusion Models (LDMs) is critically dependent on the quality of their visual tokenizers. While recent works have explore

VFM^{4}SDG: Unveiling the Power of VFMs for Single-Domain Generalized Object Detection

SafetyDGX agent

arXiv:2604.21502v1 Announce Type: new Abstract: In real-world scenarios, continual changes in weather, illumination, and imaging conditions cause significant domain shifts, leading detectors trained o

VLA-Forget: Vision-Language-Action Unlearning for Embodied Foundation Models

SafetyDGX agent

arXiv:2604.03956v2 Announce Type: replace-cross Abstract: Vision-language-action (VLA) models are emerging as embodied foundation models for robotic manipulation, but their deployment introduces a new

When Bigger Isn't Better: A Comprehensive Fairness Evaluation of Political Bias in Multi-News Summarisation

SafetyDGX agent

arXiv:2604.21309v1 Announce Type: new Abstract: Multi-document news summarisation systems are increasingly adopted for their convenience in processing vast daily news content, making fairness across d

Who Defines Fairness? Target-Based Prompting for Demographic Representation in Generative Models

SafetyDGX agent

arXiv:2604.21036v1 Announce Type: new Abstract: Text-to-image(T2I) models like Stable Diffusion and DALL-E have made generative AI widely accessible, yet recent studies reveal that these systems often

Why are all LLMs Obsessed with Japanese Culture? On the Hidden Cultural and Regional Biases of LLMs

SafetyDGX agent

arXiv:2604.21751v1 Announce Type: cross Abstract: LLMs have been showing limitations when it comes to cultural coverage and competence, and in some cases show regional biases such as amplifying Wester

XtraGPT: Context-Aware and Controllable Academic Paper Revision via Human-AI Collaboration

SafetyDGX agent

arXiv:2505.11336v4 Announce Type: replace Abstract: Despite the growing adoption of large language models (LLMs) in academic workflows, their capabilities remain limited in supporting high-quality sci

23 Apr 2026

A lot of institutions and companies won’t die because of AI. They’ll die because of their own internal inability to learn. Someone recently …

SafetyDGX agent

A lot of institutions and companies won’t die because of AI. They’ll die because of their own internal inability to learn. Someone recently told me that at a big company, legal had to review and appro

A Survey of Scaling in Large Language Model Reasoning

SafetyDGX agent

arXiv:2504.02181v2 Announce Type: replace Abstract: The rapid advancements in large Language models (LLMs) have significantly enhanced their reasoning capabilities, driven by various strategies such a

A Synchronized Audio-Visual Multi-View Capture System

SafetyDGX agent

arXiv:2603.23089v2 Announce Type: replace Abstract: Multi-view capture systems have been an important tool in research for recording human motion under controlling conditions. Most existing systems ar

AdaTracker: Learning Adaptive In-Context Policy for Cross-Embodiment Active Visual Tracking

SafetyDGX agent

arXiv:2604.20305v1 Announce Type: new Abstract: Realizing active visual tracking with a single unified model across diverse robots is challenging, as the physical constraints and motion dynamics vary

AI models of unstable flow exhibit hallucination

SafetyDGX agent

arXiv:2604.20372v1 Announce Type: cross Abstract: We report the first systematic evidence of hallucination in AI models of fluid dynamics, demonstrated in the canonical problem of hydrodynamically uns

All Languages Matter: Understanding and Mitigating Language Bias in Multilingual RAG

SafetyDGX agent

arXiv:2604.20199v1 Announce Type: new Abstract: Multilingual Retrieval-Augmented Generation (mRAG) leverages cross-lingual evidence to ground Large Language Models (LLMs) in global knowledge. However,

Auto-ART: Structured Literature Synthesis and Automated Adversarial Robustness Testing

SafetyDGX agent

arXiv:2604.20704v1 Announce Type: cross Abstract: Adversarial robustness evaluation underpins every claim of trustworthy ML deployment, yet the field suffers from fragmented protocols and undetected g

Avoiding Overthinking and Underthinking: Curriculum-Aware Budget Scheduling for LLMs

SafetyDGX agent

arXiv:2604.19780v1 Announce Type: new Abstract: Scaling test-time compute via extended reasoning has become a key paradigm for improving the capabilities of large language models (LLMs). However, exis

Best Policy Learning from Trajectory Preference Feedback

SafetyDGX agent

arXiv:2501.18873v4 Announce Type: replace Abstract: Reinforcement Learning from Human Feedback (RLHF) has emerged as a powerful approach for aligning generative models, but its reliance on learned rew

Bias in the Tails: How Name-conditioned Evaluative Framing in Resume Summaries Destabilizes LLM-based Hiring

SafetyDGX agent

arXiv:2604.19984v1 Announce Type: cross Abstract: Research has documented LLMs' name-based bias in hiring and salary recommendations. In this paper, we instead consider a setting where LLMs generate c

Breaking the Assistant Mold: Modeling Behavioral Variation in LLM Based Procedural Character Generation

SafetyDGX agent

arXiv:2601.03396v2 Announce Type: replace Abstract: Procedural content generation has enabled vast virtual worlds through levels, maps, and quests, but large-scale character generation remains underex

CAST: Achieving Stable LLM-based Text Analysis for Data Analytics

SafetyDGX agent

arXiv:2602.15861v2 Announce Type: replace-cross Abstract: Text analysis of tabular data relies on two core operations: summarization for corpus-level theme extraction and tagging for row-level labelin

ChipCraftBrain: Validation-First RTL Generation via Multi-Agent Orchestration

SafetyDGX agent

arXiv:2604.19856v1 Announce Type: cross Abstract: Large Language Models (LLMs) show promise for generating Register-Transfer Level (RTL) code from natural language specifications, but single-shot gene

CLIP-RD: Relative Distillation for Efficient CLIP Knowledge Distillation

SafetyDGX agent

arXiv:2603.25383v3 Announce Type: replace Abstract: CLIP aligns image and text embeddings via contrastive learning and demonstrates strong zero-shot generalization. Its large-scale architecture requir

CodeRL+: Improving Code Generation via Reinforcement with Execution Semantics Alignment

SafetyDGX agent

arXiv:2510.18471v2 Announce Type: replace-cross Abstract: While Large Language Models (LLMs) excel at code generation by learning from vast code corpora, a fundamental semantic gap remains between the

Cognitive Alignment At No Cost: Inducing Human Attention Biases For Interpretable Vision Transformers

SafetyDGX agent

arXiv:2604.20027v1 Announce Type: cross Abstract: For state-of-the-art image understanding, Vision Transformers (ViTs) have become the standard architecture but their processing diverges substantially

CoRe: Joint Optimization with Contrastive Learning for Medical Image Registration

SafetyDGX agent

arXiv:2603.23694v2 Announce Type: replace Abstract: Medical image registration is a fundamental task in medical image analysis, enabling the alignment of images from different modalities or time point

Coverage, Not Averages: Semantic Stratification for Trustworthy Retrieval Evaluation

SafetyDGX agent

arXiv:2604.20763v1 Announce Type: cross Abstract: Retrieval quality is the primary bottleneck for accuracy and robustness in retrieval-augmented generation (RAG). Current evaluation relies on heuristi

CubeDAgger: Interactive Imitation Learning for Dynamic Systems with Efficient yet Low-risk Interaction

SafetyDGX agent

arXiv:2505.04897v2 Announce Type: replace-cross Abstract: Interactive imitation learning makes an agent's control policy robust by stepwise supervisions from an expert. The recent algorithms mostly em

Decentralized Machine Learning with Centralized Performance Guarantees via Gibbs Algorithms

SafetyDGX agent

arXiv:2604.20492v1 Announce Type: cross Abstract: In this paper, it is shown, for the first time, that centralized performance is achievable in decentralized learning without sharing the local dataset

Distributional Inverse Reinforcement Learning

SafetyDGX agent

arXiv:2510.03013v3 Announce Type: replace Abstract: We propose a distributional framework for offline Inverse Reinforcement Learning (IRL) that jointly models uncertainty over reward functions and ful

DRIV-EX: Counterfactual Explanations for Driving LLMs

SafetyDGX agent

arXiv:2603.00696v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly used as reasoning engines in autonomous driving, yet their decision-making remains opaque. We propose

Dubious predictions that have yet to be fulfilled have given dubious people like Sam Altman and Elon Musk extraordinary power over society. …

SafetyDGX agent

Dubious predictions that have yet to be fulfilled have given dubious people like Sam Altman and Elon Musk extraordinary power over society. @carissaveliz's lecture @@TEDTalks (which I was lucky enough

Efficient Reinforcement Learning using Linear Koopman Dynamics for Nonlinear Robotic Systems

SafetyDGX agent

arXiv:2604.19980v1 Announce Type: new Abstract: This paper presents a model-based reinforcement learning (RL) framework for optimal closed-loop control of nonlinear robotic systems. The proposed appro

Efficiently Closing Loops in LiDAR-Based SLAM Using Point Cloud Density Maps

SafetyDGX agent

arXiv:2501.07399v2 Announce Type: replace Abstract: Consistent maps are key for most autonomous mobile robots, and they often use SLAM approaches to build such maps. Loop closures via place recognitio

EmbodiedMidtrain: Bridging the Gap between Vision-Language Models and Vision-Language-Action Models via Mid-training

SafetyDGX agent

arXiv:2604.20012v1 Announce Type: cross Abstract: Vision-Language-Action Models (VLAs) inherit their visual and linguistic capabilities from Vision-Language Models (VLMs), yet most VLAs are built from

Epistemic Constitutionalism Or: how to avoid coherence bias

SafetyDGX agent

arXiv:2601.14295v3 Announce Type: replace Abstract: Large language models increasingly function as artificial reasoners: they evaluate arguments, assign credibility, and express confidence. Yet their

Epistemology gives a Future to Complementarity in Human-AI Interactions

SafetyDGX agent

arXiv:2601.09871v2 Announce Type: replace Abstract: Human-AI complementarity is the claim that a human supported by an AI system can outperform either alone in a decision-making process. Since its int

ETac: A Lightweight and Efficient Tactile Simulation Framework for Learning Dexterous Manipulation

SafetyDGX agent

arXiv:2604.20295v1 Announce Type: new Abstract: Tactile sensors are increasingly integrated into dexterous robotic manipulators to enhance contact perception. However, learning manipulation policies t

Evaluating Black-Box Vulnerabilities with Wasserstein-Constrained Data Perturbations

SafetyDGX agent

arXiv:2603.15867v2 Announce Type: replace Abstract: The growing use of Machine Learning (ML) tools comes with critical challenges, such as limited model explainability. We propose a global explainabil

Explainable Speech Emotion Recognition: Weighted Attribute Fairness to Model Demographic Contributions to Social Bias

SafetyDGX agent

arXiv:2604.19763v1 Announce Type: cross Abstract: Speech Emotion Recognition (SER) systems have growing applications in sensitive domains such as mental health and education, where biased predictions

Fairness-Aware Multi-Group Target Detection in Online Discussion

SafetyDGX agent

arXiv:2407.11933v4 Announce Type: replace Abstract: Target-group detection is the task of detecting which group(s) a piece of content is ``directed at or about''. Applications include targeted marketi

Faster Fixed-Point Methods for Multichain MDPs

SafetyDGX agent

arXiv:2506.20910v2 Announce Type: replace-cross Abstract: We study value-iteration (VI) algorithms for solving general (a.k.a. multichain) Markov decision processes (MDPs) under the average-reward cri

FingerEye: Continuous and Unified Vision-Tactile Sensing for Dexterous Manipulation

SafetyDGX agent

arXiv:2604.20689v1 Announce Type: new Abstract: Dexterous robotic manipulation requires comprehensive perception across all phases of interaction: pre-contact, contact initiation, and post-contact. Su

FluSplat: Sparse-View 3D Editing without Test-Time Optimization

SafetyDGX agent

arXiv:2604.20038v1 Announce Type: new Abstract: Recent advances in text-guided image editing and 3D Gaussian Splatting (3DGS) have enabled high-quality 3D scene manipulation. However, existing pipelin

Frictionless Love: Associations Between AI Companion Roles and Behavioral Addiction

SafetyDGX agent

arXiv:2604.20011v1 Announce Type: cross Abstract: AI companion chatbots increasingly shape how people seek social and emotional connection, sometimes substituting for relationships with romantic partn

FurnSet: Exploiting Repeats for 3D Scene Reconstruction

SafetyDGX agent

arXiv:2604.20093v1 Announce Type: new Abstract: Single-view 3D scene reconstruction involves inferring both object geometry and spatial layout. Existing methods typically reconstruct objects independe

Generative Flow Networks for Model Adaptation in Digital Twins of Natural Systems

SafetyDGX agent

arXiv:2604.20707v1 Announce Type: new Abstract: Digital twins of natural systems must remain aligned with physical systems that evolve over time, are only partially observed, and are typically modeled

GRPO-VPS: Enhancing Group Relative Policy Optimization with Verifiable Process Supervision for Effective Reasoning

SafetyDGX agent

arXiv:2604.20659v1 Announce Type: cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has advanced the reasoning capabilities of Large Language Models (LLMs) by leveraging direct out

Hidden Reliability Risks in Large Language Models: Systematic Identification of Precision-Induced Output Disagreements

SafetyDGX agent

arXiv:2604.19790v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed under diverse numerical precision configurations, including standard floating-point formats (e.g.

Hybrid Latent Reasoning with Decoupled Policy Optimization

SafetyDGX agent

arXiv:2604.20328v1 Announce Type: new Abstract: Chain-of-Thought (CoT) reasoning significantly elevates the complex problem-solving capabilities of multimodal large language models (MLLMs). However, a

Hybrid Policy Distillation for LLMs

SafetyDGX agent

arXiv:2604.20244v1 Announce Type: cross Abstract: Knowledge distillation (KD) is a powerful paradigm for compressing large language models (LLMs), whose effectiveness depends on intertwined choices of

i-WiViG: Interpretable Window Vision GNN

SafetyDGX agent

arXiv:2503.08321v2 Announce Type: replace Abstract: Vision graph neural networks have emerged as a popular approach for modeling the global and spatial context for image recognition. However, a signif

I’m in Madrid this week for the first in-person meeting of the United Nations Independent International Scientific Panel on AI. I'm thrilled…

SafetyDGX agent

I’m in Madrid this week for the first in-person meeting of the United Nations Independent International Scientific Panel on AI. I'm thrilled to be part of this distinguished group of 40 international

Kalman Filter Enhanced GRPO for Reinforcement Learning-Based Language Model Reasoning

SafetyDGX agent

arXiv:2505.07527v5 Announce Type: replace Abstract: The advantage function is a central concept in RL that helps reduce variance in policy gradient estimates. For language modeling, Group Relative Pol

← Previous
1…198199200201202…240
Next →