AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
19 May 2026

QuantFPFlow: Quantum Amplitude Estimation for Fokker--Planck Policy Optimisation in Continuous Reinforcement Learning

SafetyDGX agent

arXiv:2605.16429v1 Announce Type: cross Abstract: We introduce extbf{QuantFPFlow}, a reinforcement learning framework that integrates quantum amplitude estimation into the Fokker--Planck~(FP) formulat

Quantum Sidecar Architectures for Hybrid AI Training and Inference: Stateful Protected Registers, Stateless Reset-and-Reprepare Circuits and Quantum Weight-State Outlook

ResearchDGX agent

arXiv:2605.18031v1 Announce Type: cross Abstract: We propose a quantum sidecar architecture family for future hybrid AI training and inference. The central idea is not to store an entire Transformer i

Query-Conditioned Knowledge Alignment for Reliable Cross-System Medical Reasoning


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety
DGX agent

arXiv:2605.18570v1 Announce Type: new Abstract: Cross-domain knowledge alignment is essential for integrating heterogeneous medical systems, yet existing approaches typically treat entity alignment as

Qumus: Realization of An Embodied AI Quantum Material Experimentalist

AgentsDGX agent

arXiv:2605.18407v1 Announce Type: cross Abstract: While modern Large Language Models (LLMs) and agentic artificial intelligence (AI) have demonstrated transformative capabilities in digital domains, t

RaBiT: Residual-Aware Binarization Training for Accurate and Efficient LLMs

TutorialsDGX agent

arXiv:2602.05367v2 Announce Type: replace Abstract: Efficient deployment of large language models (LLMs) requires extreme quantization, forcing a critical trade-off between low-bit efficiency and perf

RadGame: An AI-Powered Platform for Radiology Education

Local AiDGX agent

arXiv:2509.13270v2 Announce Type: replace-cross Abstract: We introduce RadGame, an AI-powered gamified platform for radiology education that targets two core skills: localizing findings and generating

RAG-based EEG-to-Text Translation Using Deep Learning and LLMs

ResearchDGX agent

arXiv:2605.17503v1 Announce Type: new Abstract: The decoding of linguistic information from electroencephalography (EEG) signals remains an extremely challenging problem in brain-computer interface (B

RAGA: Reading-And-Graph-building-Agent for Autonomous Knowledge Graph Construction and Retrieval-Augmented Generation

AgentsDGX agent

arXiv:2605.17072v1 Announce Type: new Abstract: Existing LLM-driven knowledge graph (KG) construction methods predominantly employ stateless batch processing pipelines, exhibiting structural deficienc

Randomized Advantage Transformation (RAT): Computing Natural Policy Gradients via Direct Backpropagation

SafetyDGX agent

arXiv:2605.18591v1 Announce Type: cross Abstract: Natural policy gradients improve optimization by accounting for the geometry of distribution space, but their practical use is limited by the cost of

RAP: Runtime Adaptive Pruning for LLM Inference

Model ReleasesDGX agent

arXiv:2505.17138v5 Announce Type: replace-cross Abstract: Large language models (LLMs) excel at language understanding and generation, but their enormous computational and memory requirements hinder d

RAPT: Retrieval-Augmented Post-hoc Thresholding for Multi-Label Classification

Local AiDGX agent

arXiv:2605.16535v1 Announce Type: cross Abstract: Industrial multi-label document understanding pipelines score candidate labels and threshold or rank them to form a label set per document. This early

Real-Time Aligned Reward Model beyond Semantics

SafetyDGX agent

arXiv:2601.22664v4 Announce Type: replace Abstract: Reinforcement Learning from Human Feedback (RLHF) is a pivotal technique for aligning large language models (LLMs) with human preferences, yet it is

Reasoning Before Diagnosis: Physician-Inspired Structured Thinking for ECG Classification

SafetyDGX agent

arXiv:2605.17308v1 Announce Type: new Abstract: Electrocardiogram (ECG) diagnosis in clinical practice relies on structured reasoning over multiple hierarchical aspects, including cardiac rhythm, cond

Reasoning Can Be Restored by Correcting a Few Decision Tokens

TutorialsDGX agent

arXiv:2605.16874v1 Announce Type: new Abstract: Large reasoning models (LRMs) substantially outperform their base LLM counterparts on challenging reasoning benchmarks, yet it remains poorly understood

Recall Isn't Enough: Bounding Commitments in Personalized Language Systems

ResearchDGX agent

arXiv:2605.16712v1 Announce Type: new Abstract: Long-context and memory systems usually treat personalization as a recall problem. In practice, many failures occur later, when a system commits: it tur

Reconciling Contradictory Views on the Effectiveness of SFT in LLMs: An Interaction Perspective

ResearchDGX agent

arXiv:2605.17967v1 Announce Type: new Abstract: This paper explores a scientific question in supervised fine-tuning (SFT): why SFT is broadly effective for small-scale deep neural networks, yet can pr

Reducing Credit Assignment Variance via Counterfactual Reasoning Paths

SafetyDGX agent

arXiv:2605.16302v1 Announce Type: cross Abstract: Reinforcement learning for multi-step reasoning with large language models (LLMs) often relies on sparse terminal rewards, leading to poor credit assi

Reducing Hallucination in Vision-Language Models via Stage-wise Preference Optimization under Distribution Shift

ApplicationsDGX agent

arXiv:2605.16411v1 Announce Type: cross Abstract: Hallucination remains a fundamental challenge in vision-language models (VLMs), where autoregressive generation may produce linguistically plausible y

Reliability and Effectiveness of Autonomous AI Agents in Supply Chain Management

SafetyDGX agent

arXiv:2605.17036v1 Announce Type: new Abstract: This paper studies autonomous generative AI agents in multi-echelon supply chains using the MIT Beer Game. We identify four inference-time levers that s

Remembering More, Risking More: Longitudinal Safety Risks in Memory-Equipped LLM Agents

SafetyDGX agent

arXiv:2605.17830v1 Announce Type: new Abstract: Safety evaluations of memory-equipped LLM agents typically measure within-task safety: whether an agent completes a single scenario safely, often under

Representational Alignment with Chemical Induced Fit for Molecular Relational Learning

SafetyDGX agent

arXiv:2502.07027v2 Announce Type: replace-cross Abstract: Molecular Relational Learning (MRL) is widely applied in natural sciences to predict relationships between molecular pairs by extracting struc

Response-free item difficulty modelling for multiple-choice items with fine-tuned transformers: Component-wise representation and multi-task learning

ResearchDGX agent

arXiv:2605.16991v1 Announce Type: cross Abstract: Response-free item difficulty modelling promises to reduce reliance on response-based calibration but is intrinsically difficult on reading-comprehens

Responsible Agentic AI Requires Explicit Provenance

Model ReleasesDGX agent

arXiv:2605.17169v1 Announce Type: new Abstract: Agentic AI is rapidly proliferating across diverse real-world domains such as software engineering, yet public trust has not kept pace. The central reas

ReTAMamba: Reliability-Aware Temporal Aggregation with Mamba for Irregular Clinical Time Series Prediction

ResearchDGX agent

arXiv:2605.16380v1 Announce Type: cross Abstract: Clinical time-series data are difficult to model with methods designed for regular sequences because they exhibit irregular sampling, frequent missing

Rethinking Code Review in the Age of AI: A Vision for Agentic Code Review

SafetyDGX agent

arXiv:2605.17548v1 Announce Type: cross Abstract: Code review has evolved for decades, from informal peer checking to today's pull request (PR) workflows, yet it remains a largely manual, uneven, and

Rethinking GNNs and Missing Features: Challenges, Evaluation and a Robust Solution

Model ReleasesDGX agent

arXiv:2601.04855v2 Announce Type: replace-cross Abstract: Handling missing node features is a key challenge for deploying Graph Neural Networks (GNNs) in real-world domains such as healthcare and sens

Retrieval and competition: how a protein foundation model starts a protein

SafetyDGX agent

arXiv:2605.16331v1 Announce Type: cross Abstract: Protein language models are increasingly used to guide experimental and clinical decisions, yet it is often unclear whether a confident prediction ref

Reversa: A Reverse Documentation Engineering Framework for Converting Legacy Software into Operational Specifications for AI Agents

AgentsDGX agent

arXiv:2605.18684v1 Announce Type: cross Abstract: Legacy systems concentrate business rules, architectural decisions, and operational exceptions that often remain implicit in code, data, configuration

Reverse-Engineering Model Editing on Language Models

Model ReleasesDGX agent

arXiv:2602.10134v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are pretrained on corpora containing trillions of tokens and, therefore, inevitably memorize sensitive informatio

Revisiting Long-term Time Series Forecasting: An Investigation on Linear Mapping

ApplicationsDGX agent

arXiv:2305.10721v2 Announce Type: replace-cross Abstract: Introduction: Long-term time series forecasting (LTSF) has gained significant attention in recent years. While various specialized designs exi

RGB-only Active 3D Scene Graph Generation for Indoor Mobile Robots

ResearchDGX agent

arXiv:2605.18197v1 Announce Type: cross Abstract: Current approaches to 3D scene graph generation rely on dedicated depth sensors, such as LiDAR or RGB-D cameras, for metric 3D reconstruction. This li

RLBFF: Binary Flexible Feedback to bridge between Human Feedback & Verifiable Rewards

Model ReleasesDGX agent

arXiv:2509.21319v3 Announce Type: replace-cross Abstract: Reinforcement Learning with Human Feedback (RLHF) and Reinforcement Learning with Verifiable Rewards (RLVR) are the main RL paradigms used in

RoboMME: Benchmarking and Understanding Memory for Robotic Generalist Policies

Model ReleasesDGX agent

arXiv:2603.04639v2 Announce Type: replace-cross Abstract: Memory is critical for long-horizon and history-dependent robotic manipulation. Such tasks often involve counting repeated actions or manipula

Rover: Context-aware Conflict Resolution with LLM

TutorialsDGX agent

arXiv:2605.17279v1 Announce Type: cross Abstract: Code merging is a significant challenge, particularly in large-scale projects. Existing solutions, including program analysis and machine learning, sh

S-Bus: Automatic Read-Set Reconstruction for Multi-Agent LLM State Coordination

AgentsDGX agent

arXiv:2605.17076v1 Announce Type: cross Abstract: Concurrent LLM agents sharing mutable natural-language state produce Structural Race Conditions (SRCs): write-write and cross-shard stale-read conflic

SaaSBench: Exploring the Boundaries of Coding Agents in Long-Horizon Enterprise SaaS Engineering

Model ReleasesDGX agent

arXiv:2605.17526v1 Announce Type: cross Abstract: As autonomous coding agents become capable of handling increasingly long-horizon tasks, they have gradually demonstrated the potential to complete end

SAFE-SVD: Sensitivity-Aware Fidelity-Enforcing SVD for Physics Foundation Models

ResearchDGX agent

arXiv:2605.17985v1 Announce Type: cross Abstract: We propose a new method for compressing physics foundation models (PFMs) which is a new trend in AI for Science. While model compression is essential

Safety Geometry Collapse in Multimodal LLMs and Adaptive Drift Correction

SafetyDGX agent

arXiv:2605.18104v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) often fail to transfer safety capabilities learned in the text modality to semantically equivalent non-text inp

SAME: A Semantically-Aligned Music Autoencoder

Model ReleasesDGX agent

arXiv:2605.18613v1 Announce Type: cross Abstract: Latent representations are at the heart of the majority of modern generative models. In the audio domain they are typically produced by a neural-audio

Same Signal, Different Semantics: A Cross-Framework Behavioral Analysis of Software Engineering Agents

AgentsDGX agent

arXiv:2605.18332v1 Announce Type: cross Abstract: Behavioral studies of LLM-based software engineering agents extract operational rules about which trajectory shapes correlate with higher resolution r

SAPO: Step-Aligned Policy Optimization for Reasoning-Based Generative Recommendation

SafetyDGX agent

arXiv:2605.17648v1 Announce Type: new Abstract: Generative recommendation treats next-item prediction as autoregressive item-identifier generation. Specifically, items are encoded as semantic identifi

SAS: Semantic-aware Sampling for Generative Dataset Distillation

ResearchDGX agent

arXiv:2605.18012v1 Announce Type: cross Abstract: Deep neural networks have achieved impressive performance across a wide range of tasks, but this success often comes with substantial computational an

Scalable Environments Drive Generalizable Agents

ResearchDGX agent

arXiv:2605.18181v1 Announce Type: new Abstract: Generalizable agents should adapt to diverse tasks and unseen environments beyond their training distribution. This position paper argues that such gene

Scalable Uncertainty Reasoning in Knowledge Graphs

ApplicationsDGX agent

arXiv:2605.16568v1 Announce Type: new Abstract: Knowledge Graphs are pivotal for semantic data integration. The real-world data they model is often inherently uncertain. Within knowledge graphs, uncer

Scales++: Compute Efficient Evaluation Subset Selection with Cognitive Scales Embeddings

Model ReleasesDGX agent

arXiv:2510.26384v2 Announce Type: replace Abstract: The prohibitive cost of evaluating large language models (LLMs) on comprehensive benchmarks necessitates the creation of small yet representative da

Scheduling That Speaks: An Interpretable Programmatic Reinforcement Learning Framework

Model ReleasesDGX agent

arXiv:2605.18454v1 Announce Type: cross Abstract: Deep reinforcement learning (DRL) has recently emerged as a promising approach to solve combinatorial optimization problems such as job shop schedulin

SCICONVBENCH: Benchmarking LLMs on Multi-Turn Clarification for Task Formulation in Computational Science

Model ReleasesDGX agent

arXiv:2605.18630v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed as scientific AI as- sistants, and a growing body of benchmarks evaluates their capabilities acro

Scientific Logicality Enriched Methodology for LLM Reasoning: A Practice in Physics

ResearchDGX agent

arXiv:2605.17104v1 Announce Type: new Abstract: With the continuous advancement of reasoning abilities in Large Language Models (LLMs), their application to scientific reasoning tasks has gained signi

SD-Search: On-Policy Hindsight Self-Distillation for Search-Augmented Reasoning

SafetyDGX agent

arXiv:2605.18299v1 Announce Type: new Abstract: Search-augmented reasoning agents interleave internal reasoning with calls to an external retriever, and their performance relies on the quality of each

See What I Mean: Aligning Vision and Language Representations for Video Fine-grained Object Understanding

Local AiDGX agent

arXiv:2605.18018v1 Announce Type: cross Abstract: We present SWIM (See What I Mean), a novel training strategy that aligns vision and language representations to enable fine-grained object understandi

Self-Evolving Spatial Reasoning in Vision Language Models via Geometric Logic Consistency

SafetyDGX agent

arXiv:2605.18162v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have made striking progress, yet their spatial reasoning remains fragile: models that answer an original input correctly

Self-Play Only Evolves When Self-Synthetic Pipeline Ensures Learnable Information Gain

Model ReleasesDGX agent

arXiv:2603.02218v2 Announce Type: replace-cross Abstract: Large language models (LLMs) make it plausible to build systems that improve through self-evolving loops, but many existing proposals are bett

Self-Supervised Bootstrapping of Action-Predictive Embodied Reasoning

AgentsDGX agent

arXiv:2602.08167v2 Announce Type: replace-cross Abstract: Embodied Chain-of-Thought (CoT) reasoning has significantly enhanced Vision-Language-Action (VLA) models, yet current methods rely on rigid te

Self-supervised Hierarchical Visual Reasoning with World Model

Model ReleasesDGX agent

arXiv:2605.17537v1 Announce Type: new Abstract: 3D open-world environments with adversarial opponents remain a core challenge for reinforcement learning due to their vast state spaces. Effective reaso

SEMA-RAG: A Self-Evolving Multi-Agent Retrieval-Augmented Generation Framework for Medical Reasoning

AgentsDGX agent

arXiv:2605.17101v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) is widely employed to mitigate risks such as hallucinations and knowledge obsolescence in medical question answer

Semantic Generative Tuning for Unified Multimodal Models

ResearchDGX agent

arXiv:2605.18714v1 Announce Type: cross Abstract: Unified multimodal models (UMMs) strive to consolidate visual understanding and visual generation within a single architecture. However, prevailing tr

Semantic Smoothing via Novel View Synthesis for Robust SAR Image Classification

SafetyDGX agent

arXiv:2605.16440v1 Announce Type: cross Abstract: Deep neural networks are vulnerable to adversarial perturbations, limiting deployment in safety-critical applications such as synthetic aperture radar

SENSE: Satellite-based ENergy Synthesis for Sustainable Environment

ResearchDGX agent

arXiv:2605.18101v1 Announce Type: cross Abstract: Urban Building Energy Modeling plays a critical role in achieving the United Nations' Sustainable Development Goals 7 and 11. Although existing studie

ShareChat: A Dataset of Chatbot Conversations in the Wild

Model ReleasesDGX agent

arXiv:2512.17843v4 Announce Type: replace-cross Abstract: By evaluating Large Language Models (LLMs) through uniform, text-only interfaces, current academic benchmarks obscure how the unique designs a

Shared Backbone PPO for Multi-UAV Communication Coverage with Connection Preservation

SafetyDGX agent

arXiv:2605.17999v1 Announce Type: new Abstract: This paper proposes a Shared Backbone Proximal Policy Optimization (Shared Backbone PPO) algorithm. By sharing the base module between the Actor and Cri

← Previous
1…241242243244245…358
Next →