AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,435 results
Safety

Causal Direct Preference Optimization for Distributionally Robust Generative Recommendation

DGX agent

arXiv:2603.22335v2 Announce Type: replace-cross Abstract: Direct Preference Optimization (DPO) guides large language models (LLMs) to generate recommendations aligned with user historical behavior dis

safetyarxiv-cs-ai
28 May 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Causal Machine Learning: A Survey and Open Problems

DGX agent

arXiv:2206.15475v3 Announce Type: replace Abstract: Causal Machine Learning (CausalML) is an umbrella term for machine learning methods that formalize the data-generation process as a structural causa

safetyarxiv-cs-lg
28 May 2026
Safety

CIRF: Tokenizing Chain-of-Thoughts into Reusable Functional Units for Efficient Latent Reasoning in Large Language Models

DGX agent

arXiv:2605.28292v1 Announce Type: new Abstract: Implicit Chain-of-Thought (CoT) reduces the inference cost of large language models by internalizing the explicit rationales. However, existing approach

safetyarxiv-cs-cl
28 May 2026
Safety

CodeGENCAT: Generative Computerized Adaptive Testing for Open-ended Coding Problems

DGX agent

arXiv:2602.20020v2 Announce Type: replace Abstract: Existing Computerized Adaptive Testing (CAT) frameworks typically select questions based on the predicted likelihood that the student will answer co

safetyarxiv-cs-cl
28 May 2026
Safety

Commit to the Bit: Reactive Reinforcement Learning Done Right

DGX agent

arXiv:2605.28276v1 Announce Type: new Abstract: Reinforcement learning algorithms are commonly analyzed (and designed) under the Markov assumption. This is unrealistic, as most environments encountere

safetyarxiv-cs-lg
28 May 2026
Safety

Compositional Text-to-Image Generation Via Region-aware Bimodal Direct Preference Optimization

DGX agent

arXiv:2605.28615v1 Announce Type: new Abstract: Despite the rapid progress of text-to-image (T2I) models, generating images that accurately reflect complex compositional prompts (covering attribute bi

safetyarxiv-cs-cv
28 May 2026
Safety

Con-DSO: Learning Short-Horizon Consistency Priors for RGB-D Direct Sparse Odometry

DGX agent

arXiv:2605.27952v1 Announce Type: new Abstract: Visual odometry (VO) is a fundamental component in robotics and augmented reality. RGB-D direct VO benefits from metric depth measurements, but it can d

safetyarxiv-cs-cv
28 May 2026
Safety

Counterfactually Fair Regression via Optimal Transport

DGX agent

arXiv:2605.28251v1 Announce Type: cross Abstract: We consider the problem of learning a counterfactually fair regressor. We adopt a causal uncertainty view in which counterfactual fairness is defined

safetyarxiv-cs-lg
28 May 2026
Safety

CPPO: Contrastive Perception Policy Optimization for VLM Agents

DGX agent

arXiv:2601.00501v2 Announce Type: replace Abstract: We introduce CPPO, a Contrastive Perception Policy Optimization method for finetuning vision--language models (VLMs). Reliable perception is a core

safetyarxiv-cs-cv
28 May 2026
Safety

Cross-Entropy Games and Frost Training

DGX agent

arXiv:2605.27701v1 Announce Type: new Abstract: We present Frost Training, a method for improving Monte Carlo-based policy optimization for a large family of LLM-as-a-judge tasks called Cross-Entropy

safetyarxiv-cs-ai
28 May 2026
Safety

Cyberbullying Governance on Social Media: A Unified Framework from Content Identification to Intervention

DGX agent

arXiv:2605.27584v1 Announce Type: new Abstract: The proliferation of social media platforms and online communities has inadvertently catalyzed the spread of cyberbullying, hate speech, and other forms

safetyarxiv-cs-ai
28 May 2026
Safety

DebFilter: Eradicating Biases Stashed in Value

DGX agent

arXiv:2605.28167v1 Announce Type: new Abstract: Text-to-image diffusion models, which are theoretically equivalent to score-based generative models, generate images through a multi-step denoising proc

safetyarxiv-cs-cv
28 May 2026
Safety

Deconstructing Spatial Complexity: Hierarchical Decomposition for LLM Spatial Reasoning

DGX agent

arXiv:2605.28144v1 Announce Type: new Abstract: LLMs have shown remarkable proficiency in general language understanding and reasoning. However, they consistently underperform in spatial reasoning tha

safetyarxiv-cs-ai
28 May 2026
Safety

DenoiseRL: Bootstrapping Reasoning Models to Recover from Noisy Prefixes

DGX agent

arXiv:2605.28421v1 Announce Type: new Abstract: Reinforcement learning has become a central paradigm for advancing reasoning in large language models, yet most existing methods still depend on stronge

safetyarxiv-cs-ai
28 May 2026
Safety

Diagnosing Live Within-Policy Instruction Conflicts in LLM Agents with Witnessed Resolution Profiles

DGX agent

arXiv:2605.27784v1 Announce Type: new Abstract: LLM agents are governed by long-lived natural-language prompt policies, but individually reasonable standing rules can interact in uninspected ways. We

safetyarxiv-cs-ai
28 May 2026
Safety

Diffusion-Augmented Markov Decision Processes for Maximum Entropy Reinforcement Learning

DGX agent

arXiv:2512.02019v3 Announce Type: replace-cross Abstract: Diffusion models excel at sampling from complex, unnormalized distributions. In this work, we extend Maximum Entropy Reinforcement Learning (M

safetyarxiv-cs-ai
28 May 2026
Safety

Diffusion Large Language Models for Visual Speech Recognition

DGX agent

arXiv:2605.28456v1 Announce Type: new Abstract: Existing Visual Speech Recognition (VSR) systems commonly rely on left-to-right autoregressive decoding, which can force premature decisions on visually

safetyarxiv-cs-ai
28 May 2026
Safety

DiscoForcing: A Unified Framework for Real-Time Audio-Driven Character Control with Diffusion Forcing

DGX agent

arXiv:2605.28491v1 Announce Type: new Abstract: We study real-time audio-responsive character control as a deployment-faithful problem: strictly causal, bounded-latency streaming that must generate co

safetyarxiv-cs-cv
28 May 2026
Safety

DREAM-R: Multimodal Speculative Reasoning with RL-Based Refined Drafting, Precise Verification, and Fully Parallel Execution

DGX agent

arXiv:2605.28678v1 Announce Type: new Abstract: Speculative reasoning has recently been proposed as a means to accelerate reasoning-intensive generation in large multimodal models, but its effectivene

safetyarxiv-cs-ai
28 May 2026
Safety

EAPO: Entropy-Driven Adaptive Positive-Negative Sample Weighting for Policy Optimization in Open-Ended QA

DGX agent

arXiv:2605.27846v1 Announce Type: new Abstract: Large Reasoning Models are typically trained via reinforcement learning from verifiable rewards (RLVR). However, existing approaches adopt fixed weights

safetyarxiv-cs-ai
28 May 2026
Safety

ECHO: Entropy-Confidence Hybrid Optimization for Test-Time Reinforcement Learning

DGX agent

arXiv:2602.02150v2 Announce Type: replace-cross Abstract: Test-time reinforcement learning generates multiple candidate answers via repeated rollouts and performs online updates using pseudo-labels co

safetyarxiv-cs-ai
28 May 2026
Safety

Emerging Extrinsic Dexterity in Cluttered Scenes via Dynamics-aware Policy Learning

DGX agent

arXiv:2603.09882v2 Announce Type: replace-cross Abstract: Extrinsic dexterity leverages environmental contact to overcome the limitations of prehensile manipulation. However, achieving such dexterity

safetyarxiv-cs-ai
28 May 2026
Safety

EntroAD: Structural Entropy-Guided Prompt Adaptation for Zero-Shot Anomaly Detection

DGX agent

arXiv:2605.28630v1 Announce Type: new Abstract: Zero-Shot Anomaly Detection (ZSAD) aims to detect anomalies in unseen domains without target-domain adaptation. Recent CLIP-based methods have shown pro

safetyarxiv-cs-cv
28 May 2026
Safety

Escape the Language Prior: Mitigating Late-Stage Modality Collapse in Audio Reasoning via Modality-Aware Policy Optimization

DGX agent

arXiv:2605.27741v1 Announce Type: new Abstract: Audio and omni-modal large language models exhibit impressive cross-modal reasoning capabilities. However, applying standard reinforcement learning post

safetyarxiv-cs-cl
28 May 2026
Safety

Evaluating the Realism of LLM-powered Social Agents: A Case Study of Reactions to Spanish Online News

DGX agent

arXiv:2605.28598v1 Announce Type: cross Abstract: LLM-powered social agents are increasingly used to simulate online social behavior, yet their realism remains difficult to validate. Existing work has

safetyarxiv-cs-ai
28 May 2026
Safety

Examining Agents' Bias Amplification versus Suppression in Multi-Agent Systems

DGX agent

arXiv:2605.28098v1 Announce Type: new Abstract: Multi-agent systems are increasingly deployed to support various tasks where agents interact to achieve individual and collective objectives. Although t

safetyarxiv-cs-ai
28 May 2026
Safety

FABSVer: Faster Training and Better Self-Verification for LLM Mathematical Reasoning

DGX agent

arXiv:2605.28389v1 Announce Type: new Abstract: While large language models have made significant progress in mathematical reasoning, they remain unreliable at judging the correctness of their own sol

safetyarxiv-cs-cl
28 May 2026
Safety

FedEHR-Gen: Federated Synthetic Time-Series EHR Generation via Latent Space Alignment and Distribution-Aware Aggregation

DGX agent

arXiv:2605.27892v1 Announce Type: new Abstract: Synthetic Electronic Health Record (EHR) generation provides a promising avenue for data augmentation and cross-hospital modeling in privacy-constrained

safetyarxiv-cs-lg
28 May 2026
Safety

From Affect to Complex Behavior: Advancing Multimodal Human-Centered AI at the 10th ABAW Workshop & Competition

DGX agent

arXiv:2605.27451v1 Announce Type: new Abstract: The 10th Affective & Behavior Analysis in-the-Wild (ABAW) Workshop and Competition, held at CVPR 2026, continues to advance research on modelling, analy

safetyarxiv-cs-cv
28 May 2026
Safety

From AR to Diffusion: Efficiently Adapting Large Language Models with Strictly Causal and Elastic Horizons

DGX agent

arXiv:2605.27387v1 Announce Type: cross Abstract: Diffusion models promise efficient parallel text generation but rely on bidirectional attention, creating a structural mismatch with pre-trained Autor

safetyarxiv-cs-ai
28 May 2026
Safety

From Learning Resources to Competencies: LLM-Based Tagging with Evidence and Graph Constraints

DGX agent

arXiv:2605.28483v1 Announce Type: new Abstract: Linking learning resources to a structured competency framework is key to enabling competency-based search and curriculum analytics in Learning Manageme

safetyarxiv-cs-ai
28 May 2026
Safety

From Pixels to Words -- Towards Native One-Vision Models at Scale

DGX agent

arXiv:2605.28820v1 Announce Type: new Abstract: Current vision-language models (VLMs) typically stitch together separate image encoders and language decoders via multi-stage alignment, a modular frame

safetyarxiv-cs-cv
28 May 2026
Safety

GE-Sim 2.0: A Roadmap Towards Comprehensive Closed-loop Video World Simulators for Robotic Manipulation

DGX agent

arXiv:2605.27491v1 Announce Type: new Abstract: We introduce GE-Sim 2.0 (Genie Envisioner World Simulator 2.0), a closed-loop video world simulator for robotic manipulation. Building on the action-con

safetyarxiv-cs-ro
28 May 2026
Safety

GeneralThinker: Domain-General Reasoning through Likelihood-Guided Answer-Conditioned Optimization

DGX agent

arXiv:2605.27934v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards improves language model reasoning, but its reliance on domain-specific verifiers, sparse outcome rewards,

safetyarxiv-cs-cl
28 May 2026
Safety

Geometry of Human Perceptual Domains Emerges Transiently in LLM Representations

DGX agent

arXiv:2605.27970v1 Announce Type: new Abstract: While large language models (LLMs) are trained purely on textual data, prior work has shown that their internal representations can exhibit rich geometr

safetyarxiv-cs-ai
28 May 2026
Safety

Geometry of Relaxed Fair Regression: A Unified Framework for Aware and Unaware Settings

DGX agent

arXiv:2605.28233v1 Announce Type: cross Abstract: Fairness-accuracy trade-offs are a central concern in the deployment of fairness-aware machine learning methods. When sensitive attributes are unavail

safetyarxiv-cs-lg
28 May 2026
Safety

Grimlock: Guarding High-Agency Systems with eBPF and Attested Channels

DGX agent

arXiv:2605.27488v1 Announce Type: cross Abstract: Agentic systems increasingly run user-authored orchestration code that invokes tools, spawns subtasks, and delegates work across machines and clouds.

safetyarxiv-cs-ai
28 May 2026
Safety

GS-FUSE: Granger-Supervised Gated Fusion and Multi-Granularity Alignment for Event-Driven Financial Forecasting

DGX agent

arXiv:2605.28520v1 Announce Type: new Abstract: Accurately forecasting the impact of salient financial events on markets is critical for investors and policymakers. However, existing multimodal time-s

safetyarxiv-cs-ai
28 May 2026
Safety

Guaranteed Optimal Compositional Explanations for Neurons

DGX agent

arXiv:2511.20934v2 Announce Type: replace Abstract: Compositional explanations are a family of methods that aim to describe the spatial alignment between neurons' receptive field activations and conce

safetyarxiv-cs-ai
28 May 2026
Safety

Heterogeneous Causal Discovery of Repeated Undesirable Health Outcomes

DGX agent

arXiv:2503.11477v2 Announce Type: replace Abstract: Understanding the factors that trigger or prevent undesirable health outcomes across patient subpopulations is essential for designing targeted inte

safetyarxiv-cs-ai
28 May 2026
Safety

Hierarchical Synthetic Tabular Data Generation: A Hybrid Top-Down and Bottom-Up Framework

DGX agent

arXiv:2605.28198v1 Announce Type: new Abstract: Existing approaches for synthetic tabular data generation are based on either purely generative models or LLMs, both of which struggle with data heterog

safetyarxiv-cs-lg
28 May 2026
Safety

HiRQA: Hierarchical Ranking and Quality Alignment for Opinion-Unaware Image Quality Assessment

DGX agent

arXiv:2508.15130v2 Announce Type: replace Abstract: Despite significant progress in no-reference image quality assessment (NR-IQA), dataset biases and reliance on subjective labels continue to hinder

safetyarxiv-cs-cv
28 May 2026
Safety

How the Optimizer Shapes Learned Solutions in Equivariant Neural Networks

DGX agent

arXiv:2605.27662v1 Announce Type: cross Abstract: Equivariant neural networks encode geometric symmetries by construction, yet they are often difficult to optimize and can underperform less constraine

safetyarxiv-cs-ai
28 May 2026
Safety

Human-like in-group bias in instruction-tuned language model agents

DGX agent

arXiv:2605.28114v1 Announce Type: new Abstract: As autonomous AI agents are deployed in persistent, interacting networks -- coordinating tasks, routing resources, and accumulating reputational histori

safetyarxiv-cs-ai
28 May 2026
Safety

ICG: Improving Cover Image Generation via MLLM-based Prompting and Personalized Preference Alignment

DGX agent

arXiv:2605.27374v1 Announce Type: new Abstract: Recent advances in multimodal large language models (MLLMs) and diffusion models (DMs) have opened new possibilities for AI-generated content. Yet, pers

safetyarxiv-cs-cl
28 May 2026
Safety

Identifying and Mitigating Bottlenecks in Role-Playing Agents: A Systematic Study of Disentangling Character Profile Axes

DGX agent

arXiv:2601.04716v3 Announce Type: replace Abstract: While Large Language Model (LLM) role-playing agents have advanced rapidly, it remains unclear which profile elements genuinely drive role-playing q

safetyarxiv-cs-cl
28 May 2026
Safety

Imitating and Finetuning Model Predictive Control for Robust and Symmetric Quadrupedal Locomotion

DGX agent

arXiv:2311.02304v3 Announce Type: replace Abstract: Control of legged robots is a challenging problem that has been investigated by different approaches, such as model-based control and learning algor

safetyarxiv-cs-ro
28 May 2026
Safety

IMU Propagation as Preintegration

DGX agent

arXiv:2605.28279v1 Announce Type: new Abstract: IMU preintegration is widely used in factor-graph-based visual--inertial, lidar--inertial, and radar--inertial state estimation, yet it is often treated

safetyarxiv-cs-ro
28 May 2026
← Previous
1…153154155156157…260
Next →