AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
20,981 results
13 Apr 2026

On the Limits of Layer Pruning for Generative Reasoning in Large Language Models

ResearchDGX agent

arXiv:2602.01997v2 Announce Type: replace-cross Abstract: Recent work has shown that layer pruning can effectively compress large language models (LLMs) while retaining strong performance on classific

On the Representational Limits of Quantum-Inspired 1024-D Document Embeddings: An Experimental Evaluation Framework

SafetyDGX agent

arXiv:2604.09430v1 Announce Type: cross Abstract: Text embeddings are central to modern information retrieval and Retrieval-Augmented Generation (RAG). While dense models derived from Large Language M

On the Role of DAG topology in Energy-Aware Cloud Scheduling : A GNN-Based Deep Reinforcement Learning Approach

SafetyDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.09202v1 Announce Type: cross Abstract: Cloud providers must assign heterogeneous compute resources to workflow DAGs while balancing competing objectives such as completion time, cost, and e

On the Spectral Geometry of Cross-Modal Representations: A Functional Map Diagnostic for Multimodal Alignment

SafetyDGX agent

arXiv:2604.08579v1 Announce Type: cross Abstract: We study cross-modal alignment between independently pretrained vision (DINOv2) and language (all-MiniLM-L6-v2) encoders using the functional map fram

OpenKedge: Governing Agentic Mutation with Execution-Bound Safety and Evidence Chains

SafetyDGX agent

arXiv:2604.08601v1 Announce Type: new Abstract: The rise of autonomous AI agents exposes a fundamental flaw in API-centric architectures: probabilistic systems directly execute state mutations without

Out-of-the-box: Black-box Causal Attacks on Object Detectors

ResearchDGX agent

arXiv:2512.03730v2 Announce Type: replace-cross Abstract: Adversarial perturbations are a useful way to expose vulnerabilities in object detectors. Existing perturbation methods are frequently white-b

Overhang Tower: Resource-Rational Adaptation in Sequential Physical Planning

ResearchDGX agent

arXiv:2604.09072v1 Announce Type: new Abstract: Humans effortlessly navigate the physical world by predicting how objects behave under gravity and contact forces, yet how such judgments support sequen

Overstating Attitudes, Ignoring Networks: LLM Biases in Simulating Misinformation Susceptibility

ResearchDGX agent

arXiv:2602.04674v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly used as proxies for human judgment in computational social science, yet their ability to reprodu

PACED: Distillation and On-Policy Self-Distillation at the Frontier of Student Competence

Model ReleasesDGX agent

arXiv:2603.11178v3 Announce Type: replace Abstract: Standard LLM distillation treats all training problems equally -- wasting compute on problems the student has already mastered or cannot yet solve.

Parameterized Complexity Of Representing Models Of MSO Formulas

Model ReleasesDGX agent

arXiv:2604.08707v1 Announce Type: new Abstract: Monadic second order logic (MSO2) plays an important role in parameterized complexity due to the Courcelle's theorem. This theorem states that the probl

PDE-regularized Dynamics-informed Diffusion with Uncertainty-aware Filtering for Long-Horizon Dynamics

ResearchDGX agent

arXiv:2604.09058v1 Announce Type: cross Abstract: Long-horizon spatiotemporal prediction remains a challenging problem due to cumulative errors, noise amplification, and the lack of physical consisten

PerMix-RLVR: Preserving Persona Expressivity under Verifiable-Reward Alignment

SafetyDGX agent

arXiv:2604.08986v1 Announce Type: cross Abstract: Persona prompting has been widely adopted to steer large language models (LLMs) behavior and improve their instruction performance by assigning specif

Persona-E^2: A Human-Grounded Dataset for Personality-Shaped Emotional Responses to Textual Events

ResearchDGX agent

arXiv:2604.09162v1 Announce Type: cross Abstract: Most affective computing research treats emotion as a static property of text, focusing on the writer's sentiment while overlooking the reader's persp

Physics-guided surrogate learning enables zero-shot control of turbulent wings

ResearchDGX agent

arXiv:2604.09434v1 Announce Type: cross Abstract: Turbulent boundary layers over aerodynamic surfaces are a major source of aircraft drag, yet their control remains challenging due to multiscale dynam

PhysInOne: Visual Physics Learning and Reasoning in One Suite

Model ReleasesDGX agent

arXiv:2604.09415v1 Announce Type: cross Abstract: We present PhysInOne, a large-scale synthetic dataset addressing the critical scarcity of physically-grounded training data for AI systems. Unlike exi

PilotBench: A Benchmark for General Aviation Agents with Safety Constraints

Model ReleasesDGX agent

arXiv:2604.08987v1 Announce Type: new Abstract: As Large Language Models (LLMs) advance toward embodied AI agents operating in physical environments, a fundamental question emerges: can models trained

PinpointQA: A Dataset and Benchmark for Small Object-Centric Spatial Understanding in Indoor Videos

Model ReleasesDGX agent

arXiv:2604.08991v1 Announce Type: cross Abstract: Small object-centric spatial understanding in indoor videos remains a significant challenge for multimodal large language models (MLLMs), despite its

Practical Bayesian Inference for Speech SNNs: Uncertainty and Loss-Landscape Smoothing

ResearchDGX agent

arXiv:2604.08624v1 Announce Type: cross Abstract: Spiking Neural Networks (SNNs) are naturally suited for speech processing tasks due to their specific dynamics, which allows them to handle temporal d

Precomputing Multi-Agent Path Replanning using Temporal Flexibility

Model ReleasesDGX agent

arXiv:2601.04884v2 Announce Type: replace Abstract: Executing a multi-agent plan can be challenging when an agent is delayed, because this typically creates conflicts with other agents. So, we need to

Process Reward Agents for Steering Knowledge-Intensive Reasoning

Local AiDGX agent

arXiv:2604.09482v1 Announce Type: new Abstract: Reasoning in knowledge-intensive domains remains challenging as intermediate steps are often not locally verifiable: unlike math or code, evaluating ste

Provable Post-Training Quantization: Theoretical Analysis of OPTQ and Qronos

Model ReleasesDGX agent

arXiv:2508.04853v2 Announce Type: replace-cross Abstract: Post-training quantization (PTQ) has become a crucial tool for reducing the memory and compute costs of modern deep neural networks, including

PS-TTS: Phonetic Synchronization in Text-to-Speech for Achieving Natural Automated Dubbing

ResearchDGX agent

arXiv:2604.09111v1 Announce Type: cross Abstract: Recently, artificial intelligence-based dubbing technology has advanced, enabling automated dubbing (AD) to convert the source speech of a video into

PSIRNet: Deep Learning-based Free-breathing Rapid Acquisition Late Enhancement Imaging

ResearchDGX agent

arXiv:2604.08781v1 Announce Type: cross Abstract: Purpose: To develop and evaluate a deep learning (DL) method for free-breathing phase-sensitive inversion recovery (PSIR) late gadolinium enhancement

QARIMA: A Quantum Approach To Classical Time Series Analysis

Model ReleasesDGX agent

arXiv:2604.08277v2 Announce Type: replace-cross Abstract: We present a quantum-inspired ARIMA methodology that integrates quantum-assisted lag discovery with fixed-configuration variational quantum ci

QCFuse: Query-Centric Cache Fusion for Efficient RAG Inference

Local AiDGX agent

arXiv:2604.08585v1 Announce Type: cross Abstract: Cache fusion accelerates generation process of LLMs equipped with RAG through KV caching and selective token recomputation, thereby reducing computati

QuanBench+: A Unified Multi-Framework Benchmark for LLM-Based Quantum Code Generation

Model ReleasesDGX agent

arXiv:2604.08570v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used for code generation, yet quantum code generation is still evaluated mostly within single frameworks

RAM: Recover Any 3D Human Motion in-the-Wild

ResearchDGX agent

arXiv:2603.19929v2 Announce Type: replace-cross Abstract: RAM incorporates a motion-aware semantic tracker with adaptive Kalman filtering to achieve robust identity association under severe occlusions

RAMP: Hybrid DRL for Online Learning of Numeric Action Models

SafetyDGX agent

arXiv:2604.08685v1 Announce Type: new Abstract: Automated planning algorithms require an action model specifying the preconditions and effects of each action, but obtaining such a model is often hard.

Rays as Pixels: Learning A Joint Distribution of Videos and Camera Trajectories

ResearchDGX agent

arXiv:2604.09429v1 Announce Type: cross Abstract: Recovering camera parameters from images and rendering scenes from novel viewpoints have long been treated as separate tasks in computer vision and gr

Re-Mask and Redirect: Exploiting Denoising Irreversibility in Diffusion Language Models

SafetyDGX agent

arXiv:2604.08557v1 Announce Type: cross Abstract: Diffusion-based language models (dLLMs) generate text by iteratively denoising masked token sequences. We show that their safety alignment rests on a

Reasoning in a Combinatorial and Constrained World: Benchmarking LLMs on Natural-Language Combinatorial Optimization

Model ReleasesDGX agent

arXiv:2602.02188v2 Announce Type: replace Abstract: While large language models (LLMs) have shown strong performance in math and logic reasoning, their ability to handle combinatorial optimization (CO

Reasoning Models Will Sometimes Lie About Their Reasoning

ResearchDGX agent

arXiv:2601.07663v3 Announce Type: replace Abstract: Hint-based faithfulness evaluations have established that Large Reasoning Models (LRMs) may not say what they think: they do not always volunteer in

Reasoning Provenance for Autonomous AI Agents: Structured Behavioral Analytics Beyond State Checkpoints and Execution Traces

AgentsDGX agent

arXiv:2603.21692v2 Announce Type: replace Abstract: As AI agents transition from human-supervised copilots to autonomous platform infrastructure, the ability to analyze their reasoning behavior across

RecaLLM: Addressing the Lost-in-Thought Phenomenon with Explicit In-Context Retrieval

ResearchDGX agent

arXiv:2604.09494v1 Announce Type: cross Abstract: We propose RecaLLM, a set of reasoning language models post-trained to make effective use of long-context information. In-context retrieval, which ide

Reflection of Episodes: Learning to Play Game from Expert and Self Experiences

ResearchDGX agent

arXiv:2502.13388v4 Announce Type: replace Abstract: StarCraft II is a complex and dynamic real-time strategy (RTS) game environment, which is very suitable for artificial intelligence and reinforcemen

Regime-Conditional Retrieval: Theory and a Transferable Router for Two-Hop QA

ResearchDGX agent

arXiv:2604.09019v1 Announce Type: cross Abstract: Two-hop QA retrieval splits queries into two regimes determined by whether the hop-2 entity is explicitly named in the question (Q-dominant) or only i

Reinforcement-aware Knowledge Distillation for LLM Reasoning

SafetyDGX agent

arXiv:2602.22495v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) post-training has recently driven major gains in long chain-of-thought reasoning large language models (LLMs), but

Relational Visual Similarity

ApplicationsDGX agent

arXiv:2512.07833v2 Announce Type: replace-cross Abstract: Humans do not just see attribute similarity -- we also see relational similarity. An apple is like a peach because both are reddish fruit, but

ReplicatorBench: Benchmarking LLM Agents for Replicability in Social and Behavioral Sciences

Model ReleasesDGX agent

arXiv:2602.11354v2 Announce Type: replace Abstract: The literature has witnessed an emerging interest in AI agents for automated assessment of scientific papers. Existing benchmarks focus primarily on

RESample: A Robust Data Augmentation Framework via Exploratory Sampling for Robotic Manipulation

Model ReleasesDGX agent

arXiv:2510.17640v3 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models have demonstrated remarkable performance on complex tasks through imitation learning in recent robotic man

Rethinking Prospect Theory for LLMs: Revealing the Instability of Decision-Making under Epistemic Uncertainty

SafetyDGX agent

arXiv:2508.08992v3 Announce Type: replace Abstract: Prospect Theory (PT) models human decision-making behaviour under uncertainty, among which linguistic uncertainty is commonly adopted in real-world

Retrieval Augmented Classification for Confidential Documents

Model ReleasesDGX agent

arXiv:2604.08628v1 Announce Type: cross Abstract: Unauthorized disclosure of confidential documents demands robust, low-leakage classification. In real work environments, there is a lot of inflow and

Revisiting the Capacity Gap in Chain-of-Thought Distillation from a Practical Perspective

ResearchDGX agent

arXiv:2604.08880v1 Announce Type: cross Abstract: Chain-of-thought (CoT) distillation transfers reasoning behaviors from a strong teacher to a smaller student, but prior work reports a capacity gap: d

Revitalizing Black-Box Interpretability: Actionable Interpretability for LLMs via Proxy Models

Local AiDGX agent

arXiv:2505.12509v3 Announce Type: replace-cross Abstract: Post-hoc explanations provide transparency and are essential for guiding model optimization, such as prompt engineering and data sanitation. H

Robust Reasoning Benchmark

Model ReleasesDGX agent

arXiv:2604.08571v1 Announce Type: cross Abstract: While Large Language Models (LLMs) achieve high performance on standard mathematical benchmarks, their underlying reasoning processes remain highly ov

SafeAdapt: Provably Safe Policy Updates in Deep Reinforcement Learning

Model ReleasesDGX agent

arXiv:2604.09452v1 Announce Type: cross Abstract: Safety guarantees are a prerequisite to the deployment of reinforcement learning (RL) agents in safety-critical tasks. Often, deployment environments

SafeMind: A Risk-Aware Differentiable Control Framework for Adaptive and Safe Quadruped Locomotion

SafetyDGX agent

arXiv:2604.09474v1 Announce Type: cross Abstract: Learning-based quadruped controllers achieve impressive agility but typically lack formal safety guarantees under model uncertainty, perception noise,

SAGE: A Service Agent Graph-guided Evaluation Benchmark

Model ReleasesDGX agent

arXiv:2604.09285v1 Announce Type: new Abstract: The development of Large Language Models (LLMs) has catalyzed automation in customer service, yet benchmarking their performance remains challenging. Ex

Sample-Efficient Neurosymbolic Deep Reinforcement Learning

TutorialsDGX agent

arXiv:2601.02850v2 Announce Type: replace Abstract: Reinforcement Learning (RL) is a well-established framework for sequential decision-making in complex environments. However, state-of-the-art Deep R

SatQNet: Satellite-assisted Quantum Network Entanglement Routing Using Directed Line Graph Neural Networks

ApplicationsDGX agent

arXiv:2604.09306v1 Announce Type: cross Abstract: Quantum networks are expected to become a key enabler for interconnecting quantum devices. In contrast to classical communication networks, however, i

Scalable High-Recall Constraint-Satisfaction-Based Information Retrieval for Clinical Trials Matching

ResearchDGX agent

arXiv:2604.08849v1 Announce Type: cross Abstract: Clinical trials are central to evidence-based medicine, yet many struggle to meet enrollment targets, despite the availability of over half a million

Scheming in the wild: detecting real-world AI scheming incidents with open-source intelligence

SafetyDGX agent

arXiv:2604.09104v1 Announce Type: cross Abstract: Scheming, the covert pursuit of misaligned goals by AI systems, represents a potentially catastrophic risk, yet scheming research suffers from signifi

Scrapyard AI

ApplicationsDGX agent

arXiv:2604.08803v1 Announce Type: cross Abstract: This paper considers AI model churn as an opportunity for frugal investigation of large AI models. It describes how the incessant push for ever more p

Screen, Cache, and Match: A Training-Free Causality-Consistent Reference Frame Framework for Human Animation

ResearchDGX agent

arXiv:2601.22160v2 Announce Type: replace-cross Abstract: Human animation aims to generate temporally coherent and visually consistent videos over long sequences, yet modeling long-range dependencies

SEA-Eval: A Benchmark for Evaluating Self-Evolving Agents Beyond Episodic Assessment

Model ReleasesDGX agent

arXiv:2604.08988v1 Announce Type: new Abstract: Current LLM-based agents demonstrate strong performance in episodic task execution but remain constrained by static toolsets and episodic amnesia, faili

See, Hear, and Understand: Benchmarking Audiovisual Human Speech Understanding in Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2512.02231v2 Announce Type: replace-cross Abstract: Multimodal large language models (MLLMs) are expected to jointly interpret vision, audio, and language, yet existing video benchmarks rarely a

Seeing is Believing: Robust Vision-Guided Cross-Modal Prompt Learning under Label Noise

Model ReleasesDGX agent

arXiv:2604.09532v1 Announce Type: cross Abstract: Prompt learning is a parameter-efficient approach for vision-language models, yet its robustness under label noise is less investigated. Visual conten

Self-Supervised Slice-to-Volume Reconstruction with Gaussian Representations for Fetal MRI

ResearchDGX agent

arXiv:2601.22990v2 Announce Type: replace-cross Abstract: Reconstructing 3D fetal MR volumes from motion-corrupted stacks of 2D slices is a crucial and challenging task. Conventional slice-to-volume r

Semantic Intent Fragmentation: A Single-Shot Compositional Attack on Multi-Agent AI Pipelines

SafetyDGX agent

arXiv:2604.08608v1 Announce Type: cross Abstract: We introduce Semantic Intent Fragmentation (SIF), an attack class against LLM orchestration systems where a single, legitimately phrased request cause

Semantic Rate-Distortion for Bounded Multi-Agent Communication: Capacity-Derived Semantic Spaces and the Communication Cost of Alignment

Model ReleasesDGX agent

arXiv:2604.09521v1 Announce Type: cross Abstract: When two agents of different computational capacities interact with the same environment, they need not compress a common semantic alphabet differentl

← Previous
1…341342343344345…350
Next →