AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
14 May 2026

Towards Robust Federated Multimodal Graph Learning under Modality Heterogeneity

ApplicationsDGX agent

arXiv:2605.12584v1 Announce Type: cross Abstract: Recently, multimodal graph learning (MGL) has garnered significant attention for integrating diverse modality information and structured context to su

Towards Unified Surgical Scene Understanding:Bridging Reasoning and Grounding via MLLMs

ApplicationsDGX agent

arXiv:2605.13530v1 Announce Type: cross Abstract: Surgical scene understanding is a cornerstone of computer-assisted intervention. While recent advances, particularly in surgical image segmentation, h

Tracing Persona Vectors Through LLM Pretraining

SafetyDGX agent

arXiv:2605.13329v1 Announce Type: cross Abstract: How large language models internally represent high-level behaviors is a core interpretability question with direct relevance to AI safety: it determi


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Training Large Language Models to Predict Clinical Events

Model ReleasesDGX agent

arXiv:2605.12817v1 Announce Type: cross Abstract: Longitudinal clinical notes contain rich evidence of how patients evolve over time, but converting this signal into training supervision for clinical

Training LLMs with Reinforcement Learning for Intent-Aware Personalized Question Answering

Model ReleasesDGX agent

arXiv:2605.12645v1 Announce Type: cross Abstract: Effective personalized question answering (PQA) in language models requires grounding responses in the user's underlying intent, where intent refers t

TRIAGE: Evaluating Prospective Metacognitive Control in LLMs under Resource Constraints

AgentsDGX agent

arXiv:2605.13414v1 Announce Type: new Abstract: Deploying language models as autonomous agents requires more than per-task accuracy: when an agent faces a queue of problems under a finite token budget

Uncovering Latent Pathological Signatures in Pulmonary CT via Cross-Window Knowledge Distillation

TutorialsDGX agent

arXiv:2605.12562v1 Announce Type: cross Abstract: Multi-window CT imaging captures complementary pathological information across anatomical structures of differing densities, yet existing deep learnin

Uncovering Symmetry Transfer in Large Language Models via Layer-Peeled Optimization

ResearchDGX agent

arXiv:2605.12756v1 Announce Type: cross Abstract: Large language models (LLMs) are pretrained by minimizing the cross-entropy loss for next-token prediction. In this paper, we study whether this optim

Understanding and Accelerating the Training of Masked Diffusion Language Models

Model ReleasesDGX agent

arXiv:2605.13026v1 Announce Type: cross Abstract: Masked diffusion models (MDMs) have emerged as a promising alternative to autoregressive models (ARMs) for language modeling. However, MDMs are known

UniJEPA: Enhancing Robot Policy via Unified Continuous and Discrete Representation Learning

SafetyDGX agent

arXiv:2510.10642v3 Announce Type: replace-cross Abstract: Building generalist robot policies that can handle diverse tasks in open-ended environments is a central challenge in robotics. To leverage kn

Unweighted ranking for value-based decision making with uncertainty

SafetyDGX agent

arXiv:2605.13601v1 Announce Type: new Abstract: As intelligent systems are increasingly implemented in our society to make autonomous decisions, their commitment to human values raises serious concern

Useful Memories Become Faulty When Continuously Updated by LLMs

Model ReleasesDGX agent

arXiv:2605.12978v1 Announce Type: new Abstract: Learning from past experience benefits from two complementary forms of memory: episodic traces -- raw trajectories of what happened -- and consolidated

Utility-Oriented Visual Evidence Selection for Multimodal Retrieval-Augmented Generation

ResearchDGX agent

arXiv:2605.13277v1 Announce Type: cross Abstract: Visual evidence selection is a critical component of multimodal retrieval-augmented generation (RAG), yet existing methods typically rely on semantic

VERA-MH Concept Paper

Model ReleasesDGX agent

arXiv:2510.15297v4 Announce Type: replace-cross Abstract: We introduce VERA-MH (Validation of Ethical and Responsible AI in Mental Health), an automated evaluation of the safety of AI chatbots used in

VERA-MH: Validation of Ethical and Responsible AI in Mental Health

SafetyDGX agent

arXiv:2605.13318v1 Announce Type: new Abstract: Chatbot usage has increased, including in fields for which they were never developed for--notably mental health support. To that end, we introduce Valid

VideoSEAL: Mitigating Evidence Misalignment in Agentic Long Video Understanding by Decoupling Answer Authority

SafetyDGX agent

arXiv:2605.12571v1 Announce Type: cross Abstract: Long video question answering requires locating sparse, time-scattered visual evidence within highly redundant content. Although current MLLMs perform

VIPO: Value Function Inconsistency Penalized Offline Reinforcement Learning

TutorialsDGX agent

arXiv:2504.11944v3 Announce Type: replace-cross Abstract: Offline reinforcement learning (RL) learns effective policies from pre-collected datasets, offering a practical solution for applications wher

Visual Accommodation: Rethinking Image Scale as a Learnable Variable for Object Detection

TutorialsDGX agent

arXiv:2412.06341v2 Announce Type: replace-cross Abstract: We propose Ciliary-DETR (previous name: Elastic-DETR), a framework for test-time resolution adjustment analogous to biological accommodation.

Visual Aesthetic Benchmark: Can Frontier Models Judge Beauty?

Model ReleasesDGX agent

arXiv:2605.12684v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) are now routinely deployed for visual understanding, generation, and curation. A substantial fraction of thes

Vividh-ASR: A Complexity-Tiered Benchmark and Optimization Dynamics for Robust Indic Speech Recognition

Model ReleasesDGX agent

arXiv:2605.13087v1 Announce Type: cross Abstract: Fine-tuning multilingual ASR models like Whisper for low-resource languages often improves read speech but degrades spontaneous audio performance, a p

WARDEN: Endangered Indigenous Language Transcription and Translation with 6 Hours of Training Data

ResearchDGX agent

arXiv:2605.13846v1 Announce Type: cross Abstract: This paper introduces WARDEN, an early language model system capable of transcribing and translating Wardaman, an endangered Australian indigenous lan

Watermarking Should Be Treated as a Monitoring Primitive

SafetyDGX agent

arXiv:2605.13095v1 Announce Type: cross Abstract: Watermarking is widely proposed for provenance, attribution, and safety monitoring in generative models, yet is typically evaluated only under adversa

Weakly Supervised Segmentation as Semantic-Based Regularization

ResearchDGX agent

arXiv:2605.13674v1 Announce Type: cross Abstract: Weakly supervised semantic segmentation (WSSS) trains dense pixel-level segmentation models from partial or coarse annotations such as bounding boxes,

Weakly-Supervised Spatiotemporal Anomaly Detection

ResearchDGX agent

arXiv:2605.13746v1 Announce Type: cross Abstract: In this paper, we explore a weakly supervised method for anomaly detection. Since annotating videos is time-consuming, we only look at weak video-leve

What Do You Think I Think? Accounting for Human Beliefs Using Second-Order Theory of Mind

AgentsDGX agent

arXiv:2605.12745v1 Announce Type: cross Abstract: Discrepancies between an agent's actual knowledge and what a person thinks the agent knows can hinder interactions. If an agent could detect such disc

What Limits Vision-and-Language Navigation ?

AgentsDGX agent

arXiv:2605.13328v1 Announce Type: cross Abstract: Vision-and-Language Navigation (VLN) is a cornerstone of embodied intelligence. However, current agents often suffer from significant performance degr

What properties of reasoning supervision are associated with improved downstream model quality?

SafetyDGX agent

arXiv:2605.13290v1 Announce Type: new Abstract: Validating training data for reasoning models typically requires expensive trial-and-error fine-tuning cycles. In this work, we investigate whether the

When Absolute State Fails: Evaluating Proprioceptive Encodings for Robust Manipulation

ApplicationsDGX agent

arXiv:2605.13067v1 Announce Type: cross Abstract: As end-to-end robotic policies are progressively deployed in the real world to solve real tasks, they face a gap between the training and inference co

When Attention Closes: How LLMs Lose the Thread in Multi-Turn Interaction

Model ReleasesDGX agent

arXiv:2605.12922v1 Announce Type: new Abstract: Large language models can follow complex instructions in a single turn, yet over long multi-turn interactions they often lose the thread of instructions

When Diffusion Breaks Constraints: Sequential Autoregressive Generation with RL and MCTS

ResearchDGX agent

arXiv:2512.01242v3 Announce Type: replace-cross Abstract: Data-driven generative models excel in language and vision, but diffusion models often fail in constrained planning and design tasks, exhibiti

When Does Hierarchy Help? Benchmarking Agent Coordination in Event-Driven Industrial Scheduling

Model ReleasesDGX agent

arXiv:2605.13172v1 Announce Type: cross Abstract: Recent advances in agent and multi-agent systems have shown strong performance on tool use, reasoning, and collaborative tasks. However, existing benc

When Should an AI Workflow Release? Always-Valid Inference for Black-Box Generate-Verify Systems

AgentsDGX agent

arXiv:2605.12947v1 Announce Type: cross Abstract: LLM-enabled AI workflows increasingly produce outputs through iterative generate-evaluate-revise loops. Each iteration can improve the candidate, but

When to Think Fast and Slow? AMOR: Adaptive Entropy Gate for Hybrid Models

ResearchDGX agent

arXiv:2602.13215v2 Announce Type: replace Abstract: Recurrent-attention hybrids aim to combine the efficiency of recurrence with the expressivity of attention, but existing approaches typically apply

Where Does Reasoning Break? Step-Level Hallucination Detection via Hidden-State Transport Geometry

Local AiDGX agent

arXiv:2605.13772v1 Announce Type: cross Abstract: Large language models hallucinate during multi-step reasoning, but most existing detectors operate at the trace level: they assign one confidence scor

Why the Unfinished Keeps Returning: Canxianization and the Dynamics of Conscious Priority

ResearchDGX agent

arXiv:2605.12543v1 Announce Type: cross Abstract: Some conscious contents disappear after access; others return repeatedly, long after their triggering conditions have ceased. We propose Canxianizatio

WriteSAE: Sparse Autoencoders for Recurrent State

ResearchDGX agent

arXiv:2605.12770v1 Announce Type: cross Abstract: We introduce WriteSAE, the first sparse autoencoder that decomposes and edits the matrix cache write of state-space and hybrid recurrent language mode

X-Restormer++: 1st Place Solution for the UG2+ CVPR 2026 All-Weather Restoration Challenge

ResearchDGX agent

arXiv:2605.13258v1 Announce Type: cross Abstract: In this work, we present our winning solution for the 8th UG2+ Challenge (CVPR 2026) Track 1: Image Restoration under All-weather Conditions. Our meth

12 May 2026

A Cognitively Grounded Bayesian Framework for Misinformation Susceptibility

Model ReleasesDGX agent

arXiv:2605.09483v1 Announce Type: cross Abstract: In this (work in progress) paper, we present Bounded Pragmatic Listener (or BPL), a cognitively grounded Bayesian framework for modelling susceptibili

A Cold Diffusion Approach for Percussive Dereverberation

ApplicationsDGX agent

arXiv:2605.10256v1 Announce Type: cross Abstract: Most recent advances in audio dereverberation focus almost exclusively on speech, leaving percussive and drum signals largely unexplored despite their

A Communication-Theoretic Framework for LLM Agents: Cost-Aware Adaptive Reliability

Model ReleasesDGX agent

arXiv:2605.09121v1 Announce Type: cross Abstract: Agents built on large language models (LLMs) rely on a range of reliability techniques, including retry, majority voting, and self-consistency, that h

A Comparative Study of Machine Learning and Deep Learning for Out-of-Distribution Detection

ApplicationsDGX agent

arXiv:2605.10181v1 Announce Type: cross Abstract: Out-of-distribution (OOD) detection is essential for building reliable AI systems, as models that produce outputs for invalid inputs cannot be trusted

A Game Theoretic Free Energy Analysis of Higher Order Synergy in Attention Heads of Large Language Models

Model ReleasesDGX agent

arXiv:2605.09515v1 Announce Type: new Abstract: Large language models rely on multihead attention, but interactions among heads remain poorly understood. We apply the Game Theoretic Free Energy Princi

A Geometric Perspective on Next-Token Prediction in Large Language Models: Three Emerging Phases

Model ReleasesDGX agent

arXiv:2605.09011v1 Announce Type: cross Abstract: We investigate the geometry of predictive information across the layers of large language models (LLMs). We repurpose representation lenses-learned af

A meshfree exterior calculus for generalizable and data-efficient learning of physics from point clouds

Model ReleasesDGX agent

arXiv:2605.08436v1 Announce Type: cross Abstract: We introduce a meshfree exterior calculus (MEEC) for learning structure-preserving descriptions of physics on point clouds, and use it to build MEEC-N

A Minimum Variance Path Principle for Accurate and Stable Score-Based Density Ratio Estimation

ResearchDGX agent

arXiv:2602.00834v4 Announce Type: replace-cross Abstract: Score-based methods are powerful across machine learning, but they face a paradox: theoretically path-independent, yet practically path-depend

A Paired Point-of-Care Ultrasound Dataset for Image Quality Enhancement and Benchmarking via a cGAN Baseline

ResearchDGX agent

arXiv:2605.08282v1 Announce Type: cross Abstract: Purpose: We aim to enhance the image quality of point-of-care ultrasound (POCUS) devices using deep learning and a novel paired dataset of POCUS and h

A Physical Theory of Backpropagation: Exact Gradients from the Least-Action Principle

Local AiDGX agent

arXiv:2602.02281v2 Announce Type: replace-cross Abstract: Backpropagation is typically presented as a symbolic procedure: a backward pass topologically distinct from inference, with non-local error si

A probabilistic framework for crystal structure denoising, phase classification, and order parameters

Local AiDGX agent

arXiv:2512.11077v3 Announce Type: replace-cross Abstract: Atomistic simulations generate large volumes of noisy structural data, yet extracting phase labels and continuous order parameters (OPs) in a

A Prompt-Aware Structuring Framework for Reliable Reuse of AI-Generated Content in the Agentic Web

AgentsDGX agent

arXiv:2605.09283v1 Announce Type: new Abstract: The evolution of Large Language Models (LLMs) and the software agents built on them (AI agents) marks a turning point in the transition from a human-cen

A Qualitative Test-Risk Mechanism for Scaling Behavior in Normalized Residual Networks

ResearchDGX agent

arXiv:2605.08297v1 Announce Type: cross Abstract: The scaling behavior, in which test performance often improves as model size and data increase, is a central empirical phenomenon in modern deep learn

A Quantum Inspired Variational Kernel and Explainable AI Framework for Cross Region Solar and Wind Energy Forecasting

Model ReleasesDGX agent

arXiv:2605.09032v1 Announce Type: cross Abstract: Reliable short horizon forecasting of solar and wind generation is a structural prerequisite of any modern power system yet most published forecasters

A Reconfigurable Multiplier Architecture for Error-Resilient Applications in RISC-V Core

Local AiDGX agent

arXiv:2605.08785v1 Announce Type: cross Abstract: Neural Networks (NNs) have been widely adopted due to their outstanding efficacy and adaptability across computer vision and deep learning application

A Recursive Decomposition Framework for Causal Structure Learning in the Presence of Latent Variables

ApplicationsDGX agent

arXiv:2605.10651v1 Announce Type: cross Abstract: Constraint-based causal discovery is widely used for learning causal structures, but heavy reliance on conditional independence (CI) testing makes it

A Reflective Storytelling Agent for Older Adults: Integrating Argumentation Schemes and Argument Mining in LLM-Based Personalised Narratives

AgentsDGX agent

arXiv:2605.10531v1 Announce Type: new Abstract: This work investigates whether knowledge-driven large language model (LLM)-based storytelling can support purposeful narrative interaction with a digita

A Resilient Solution for Sewer Overflow Monitoring across Cloud and Edge

ResearchDGX agent

arXiv:2605.10592v1 Announce Type: new Abstract: Aging combined sewer systems in many historical cities are increasingly stressed by extreme rainfall events, which can trigger combined sewer overflows

A Robust Out-of-Distribution Detection Framework via Synergistic Smoothing

ResearchDGX agent

arXiv:2605.08191v1 Announce Type: cross Abstract: Reliable out-of-distribution (OOD) detection is a critical requirement for the safe deployment of machine learning systems. Despite recent progress, s

A Scalable Entity-Based Framework for Auditing Bias in LLMs

SafetyDGX agent

arXiv:2601.12374v2 Announce Type: replace-cross Abstract: Existing approaches to bias evaluation in large language models (LLMs) trade ecological validity for statistical control, relying either on ar

A Semantic-Sampling Framework for Evaluating Calibration in Open-Ended Question Answering

TutorialsDGX agent

arXiv:2605.08432v1 Announce Type: cross Abstract: Calibration measures whether a model's predicted confidence aligns with its empirical accuracy, and is central to the reliable deployment of large lan

A Single Neuron Is Sufficient to Bypass Safety Alignment in Large Language Models

SafetyDGX agent

arXiv:2605.08513v1 Announce Type: cross Abstract: Safety alignment in language models operates through two mechanistically distinct systems: refusal neurons that gate whether harmful knowledge is expr

A Unified Pair-GRPO Family: From Implicit to Explicit Preference Constraints for Stable and General RL Alignment

Local AiDGX agent

arXiv:2605.06375v1 Announce Type: cross Abstract: Large language model (LLM) alignment via reinforcement learning from human preferences (RLHF) suffers from unstable policy updates, ambiguous gradient

← Previous
1…261262263264265…358
Next →