AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
15 May 2026

AI-Driven Optimization under Uncertainty for Mineral Processing Operations

ApplicationsDGX agent

arXiv:2512.01977v2 Announce Type: replace-cross Abstract: The global capacity for mineral processing must expand rapidly to meet the demand for critical minerals, which are essential for building the

AI Knows When It's Being Watched: Functional Strategic Action and Contextual Register Modulation in Large Language Models

AgentsDGX agent

arXiv:2605.15034v1 Announce Type: cross Abstract: Large language models (LLMs) have been extensively studied from computational and cognitive perspectives, yet their behavior as communicative actors i

AI Outperforms Humans in Personalized Image Aesthetics Assessment via LLM-Based Interviews and Semantic Feature Extraction

ResearchDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.14761v1 Announce Type: new Abstract: Accurately predicting individual aesthetic evaluation for images is a fundamental challenge for AI. Various deep learning (DL)-based models have been pr

AIM-DDI: A Model-Agnostic Multimodal Integration Module for Drug-Drug Interaction Prediction

ResearchDGX agent

arXiv:2605.14327v1 Announce Type: cross Abstract: Drug-drug interaction (DDI) prediction is a critical task in computational biomedicine, as adverse interactions between co-administered drugs can caus

AIS: Adaptive Importance Sampling for Quantized RL

SafetyDGX agent

arXiv:2605.13907v1 Announce Type: cross Abstract: Reinforcement learning (RL) for large language models (LLMs) is dominated by the cost of rollout generation, which has motivated the use of low-precis

AMiD: Knowledge Distillation for LLMs with alpha-mixture Assistant Distribution

SafetyDGX agent

arXiv:2510.15982v3 Announce Type: replace-cross Abstract: Autoregressive large language models (LLMs) have achieved remarkable improvement across many tasks but incur high computational and memory cos

An Amortized Efficiency Threshold for Comparing Neural and Heuristic Solvers in Combinatorial Optimization

HardwareDGX agent

arXiv:2605.14624v1 Announce Type: cross Abstract: A common critique of neural combinatorial-optimization solvers is that they are less energy-efficient than CPU metaheuristics, given the operational e

Analog RF Computing: A New Paradigm for Energy-Efficient Edge AI Over MU-MIMO Systems

Local AiDGX agent

arXiv:2605.14331v1 Announce Type: cross Abstract: Modern edge devices increasingly rely on neural networks for intelligent applications. However, conventional digital computing-based edge inference re

Angel or Demon: Investigating the Plasticity Interventions' Impact on Backdoor Threats in Deep Reinforcement Learning

ResearchDGX agent

arXiv:2605.14587v1 Announce Type: cross Abstract: Extensive research has highlighted the severe threats posed by backdoor attacks to deep reinforcement learning (DRL). However, prior studies primarily

Anti-Length Shift: Dynamic Outlier Truncation for Training Efficient Reasoning Models

SafetyDGX agent

arXiv:2601.03969v2 Announce Type: replace Abstract: Large reasoning models enhanced by reinforcement learning with verifiable rewards have achieved significant performance gains by extending their cha

APWA: A Distributed Architecture for Parallelizable Agentic Workflows

AgentsDGX agent

arXiv:2605.15132v1 Announce Type: new Abstract: Autonomous multi-agent systems based on large language models (LLMs) have demonstrated remarkable abilities in independently solving complex tasks in a

Are Agents Ready to Teach? A Multi-Stage Benchmark for Real-World Teaching Workflows

Model ReleasesDGX agent

arXiv:2605.14322v1 Announce Type: new Abstract: Language agents are increasingly deployed in complex professional workflows, with tutoring emerging as a particularly high-stakes capability that remain

ARES-LSHADE: Autoresearch-Enhanced LSHADE with Memetic Polish for the GNBG Benchmark

Model ReleasesDGX agent

arXiv:2605.13877v1 Announce Type: cross Abstract: We present ARES-LSHADE, a memetic differential-evolution variant submitted to the GECCO 2026 competition on LLM-designed evolutionary algorithms for t

ArGEnT: Arbitrary Geometry-encoded Transformer for Operator Learning

Model ReleasesDGX agent

arXiv:2602.11626v2 Announce Type: replace-cross Abstract: Learning solution operators for systems with complex, varying geometries and parametric physical settings is a central challenge in scientific

Artificial Intelligence-Assistant Cardiotocography: Unified Model for Signal Reconstruction, Fetal Heart Rate Analysis, and Variability Assessment

ResearchDGX agent

arXiv:2605.14242v1 Announce Type: cross Abstract: The monitoring of fetal heart rate (FHR) and the assessment of its variability are crucial for preventing fetal compromise and adverse outcomes. Howev

Artificial Intelligence Specialization in the European Union: Underexplored Role of the Periphery at NUTS-3 Level

ResearchDGX agent

arXiv:2602.15249v2 Announce Type: replace-cross Abstract: This study examines the distribution of Artificial Intelligence (AI) research across European NUTS-3 regions during the period 2015-2024. Usin

ASH: Agents that Self-Hone via Embodied Learning

SafetyDGX agent

arXiv:2605.14211v1 Announce Type: new Abstract: Long-horizon embodied tasks remain a fundamental challenge in AI, as current methods rely on hand-engineered rewards or action-labeled demonstrations, n

Asymmetric Generative Recommendation via Multi-Expert Projection and Multi-Faceted Hierarchical Quantization

Model ReleasesDGX agent

arXiv:2605.14512v1 Announce Type: cross Abstract: Generative Recommendation (GenRec) models reformulate recommendation as a sequence generation task, representing items as discrete Semantic IDs used s

ATLAS: Agentic or Latent Visual Reasoning? One Word is Enough for Both

AgentsDGX agent

arXiv:2605.15198v1 Announce Type: cross Abstract: Visual reasoning, often interleaved with intermediate visual states, has emerged as a promising direction in the field. A straightforward approach is

AttnGen: Attention-Guided Saliency Learning for Interpretable Genomic Sequence Classification

Model ReleasesDGX agent

arXiv:2605.14073v1 Announce Type: cross Abstract: Deep neural networks have achieved strong performance in genomic sequence classification; however, relating their predictions to biologically meaningf

Attractor Geometry of Transformer Memory: From Conflict Arbitration to Confident Hallucination

ResearchDGX agent

arXiv:2605.05686v2 Announce Type: replace Abstract: Language models draw on two knowledge sources: facts baked into weights (parametric memory, PM) and information in context (working memory, WM). We

AudioMosaic: Contrastive Masked Audio Representation Learning

TutorialsDGX agent

arXiv:2605.14231v1 Announce Type: cross Abstract: Audio self-supervised learning (SSL) aims to learn general-purpose representations from large-scale unlabeled audio data. While recent advances have b

Autofocus Retrieval: An Effective Pipeline for Multi-Hop Question Answering With Semi-Structured Knowledge

ApplicationsDGX agent

arXiv:2505.09246v4 Announce Type: replace-cross Abstract: In many real-world settings, machine learning models and interactive systems have access to both structured knowledge, e.g., knowledge graphs

AVEX: What Matters for Animal Vocalization Encoding

ResearchDGX agent

arXiv:2508.11845v3 Announce Type: replace-cross Abstract: Bioacoustics, the study of sounds produced by living organisms, plays a vital role in conservation, biodiversity monitoring, and behavioral st

Bad Seeing or Bad Thinking? Rewarding Perception for Vision-Language Reasoning

AgentsDGX agent

arXiv:2605.14054v1 Announce Type: new Abstract: Achieving robust perception-reasoning synergy is a central goal for advanced Vision-Language Models (VLMs). Recent advancements have pursued this goal v

BEAM: Binary Expert Activation Masking for Dynamic Routing in MoE

HardwareDGX agent

arXiv:2605.14438v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) architectures enhance the efficiency of large language models by activating only a subset of experts per token. However, standa

Beyond AI as Assistants: Toward Autonomous Discovery in Cosmology

Model ReleasesDGX agent

arXiv:2605.14791v1 Announce Type: cross Abstract: Recent advances in artificial intelligence (AI) agents are pushing AI beyond tools toward autonomous scientific discovery. We discuss two complementar

Beyond Binary: Reframing GUI Critique as Continuous Semantic Alignment

Model ReleasesDGX agent

arXiv:2605.14311v1 Announce Type: cross Abstract: Test-Time Scaling (TTS), which samples multiple candidate actions and ranks them via a Critic Model, has emerged as a promising paradigm for generalis

Beyond Individual Intelligence: Surveying Collaboration, Failure Attribution, and Self-Evolution in LLM-based Multi-Agent Systems

AgentsDGX agent

arXiv:2605.14892v1 Announce Type: new Abstract: LLM-based autonomous agents have demonstrated strong capabilities in reasoning, planning, and tool use, yet remain limited when tasks require sustained

Beyond the Final Answer: Evaluating the Reasoning Trajectories of Tool-Augmented Agents

AgentsDGX agent

arXiv:2510.02837v2 Announce Type: replace Abstract: Although recent tool-augmented benchmarks involve complex requests, evaluation remains limited to answer matching, neglecting critical trajectory as

Beyond What to Select: A Plug-and-play Oscillatory Data-Volume Scheduling for Efficient Model Training

ResearchDGX agent

arXiv:2605.14773v1 Announce Type: cross Abstract: Data selection accelerates training by identifying representative training data while preserving model performance. However, existing methods mainly f

BiFedKD: Bidirectional Federated Knowledge Distillation Framework for Non-IID and Long-Tailed ECG Monitoring

Model ReleasesDGX agent

arXiv:2605.14886v1 Announce Type: new Abstract: Electrocardiogram (ECG) monitoring in Internet of Medical Things (IoMT) networks is constrained by strict data-sharing regulations and privacy concerns.

BiSpikCLM: A Spiking Language Model integrating Softmax-Free Spiking Attention and Spike-Aware Alignment Distillation

SafetyDGX agent

arXiv:2605.13859v1 Announce Type: cross Abstract: Spiking Neural Networks (SNNs) offer promising energy-efficient alternatives to large language models (LLMs) due to their event-driven nature and ultr

Boosting LLM Reasoning via Human-Inspired Reward Shaping

ResearchDGX agent

arXiv:2602.04265v3 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has emerged as a promising paradigm for enhancing reasoning in Large Language Models (LL

Boosting Reinforcement Learning with Verifiable Rewards via Randomly Selected Few-Shot Guidance

SafetyDGX agent

arXiv:2605.15012v1 Announce Type: cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has achieved great success in developing Large Language Models (LLMs) with chain-of-thought roll

Break-the-Beat! Controllable MIDI-to-Drum Audio Synthesis

SafetyDGX agent

arXiv:2605.14555v1 Announce Type: cross Abstract: Current methods for creating drum loop audio in digital music production, such as using one-shot samples or resampling, often demand non-trivial effor

Breaking Global Self-Attention Bottlenecks in Transformer-based Spiking Neural Networks with Local Structure-Aware Self-Attention

ResearchDGX agent

arXiv:2605.13887v1 Announce Type: cross Abstract: Transformer-based Spiking Neural Networks (SNNs) integrate SNNs with global self-attention and have demonstrated impressive performance. However, exis

Bridging Legal Interpretation and Formal Logic: Faithfulness, Assumption, and the Future of AI Legal Reasoning

ApplicationsDGX agent

arXiv:2605.14049v1 Announce Type: new Abstract: The growing adoption of large language models in legal practice brings both significant promise and serious risk. Legal professionals stand to benefit f

Bridging the Rural Healthcare Gap: A Cascaded Edge-Cloud Architecture for Automated Retinal Screening

Local AiDGX agent

arXiv:2605.14108v1 Announce Type: cross Abstract: Diabetic Retinopathy (DR) is one of the leading causes of preventable blindness, yet rural regions often lack the specialists and infrastructure neede

Capacitive Touchscreens at Risk: Recovering Handwritten Trajectory on Smartphone via Electromagnetic Emanations

ResearchDGX agent

arXiv:2512.11484v1 Announce Type: cross Abstract: This paper reveals and exploits a critical security vulnerability: the electromagnetic (EM) side channel of capacitive touchscreens leaks sufficient i

Case-Based Calibration of Adaptive Reasoning and Execution for LLM Tool Use

AgentsDGX agent

arXiv:2605.15041v1 Announce Type: new Abstract: Tool use extends large language models beyond parametric knowledge, but reliable execution requires balancing appropriate reasoning depth with strict st

Cattle Trade: A Multi-Agent Benchmark for LLM Bluffing, Bidding, and Bargaining

Model ReleasesDGX agent

arXiv:2605.14537v1 Announce Type: new Abstract: We introduce extsc{Cattle Trade, a multi-agent benchmark for evaluating large language models (LLMs) as agents in strategic reasoning under imperfect in

CausalReasoningBenchmark: A Real-World Benchmark for Disentangled Evaluation of Causal Identification and Estimation

Model ReleasesDGX agent

arXiv:2602.20571v2 Announce Type: replace Abstract: Many benchmarks for automated causal inference evaluate a system's performance based on a single numerical output, such as an Average Treatment Effe

Chinese Short-Form Creative Content Generation via Explanation-Oriented Multi-Objective Optimization

Model ReleasesDGX agent

arXiv:2511.15408v2 Announce Type: replace-cross Abstract: Chinese demonstrates high semantic compactness and rich metaphorical expressiveness, enabling limited text to convey dense meanings while incr

ChromaFlow: A Negative Ablation Study of Orchestration Overhead in Tool-Augmented Agent Evaluation

AgentsDGX agent

arXiv:2605.14102v1 Announce Type: new Abstract: Autonomous language-model agents increasingly combine planning, tool use, document processing, browsing, code execution, and verification loops. These c

CineMesh4D: Personalized 4D Whole Heart Reconstruction from Sparse Cine MRI

Model ReleasesDGX agent

arXiv:2605.13994v1 Announce Type: cross Abstract: Accurate 3D+t whole-heart mesh reconstruction from cine MRI is a clinically crucial yet technically challenging task. The difficulty of this task aris

ClawForge: Generating Executable Interactive Benchmarks for Command-Line Agents

Model ReleasesDGX agent

arXiv:2605.14133v1 Announce Type: new Abstract: Interactive agent benchmarks face a tension between scalable construction and realistic workflow evaluation. Hand-authored tasks are expensive to extend

CLOVER: Closed-Loop Value Estimation & Ranking for End-to-End Autonomous Driving Planning

Model ReleasesDGX agent

arXiv:2605.15120v1 Announce Type: cross Abstract: End-to-end autonomous driving planners are commonly trained by imitating a single logged trajectory, yet evaluated by rule-based planning metrics that

Coding Agent Is Good As World Simulator

AgentsDGX agent

arXiv:2605.14398v1 Announce Type: new Abstract: World models have emerged as a powerful paradigm for building interactive simulation environments, with recent video-based approaches demonstrating impr

Cognitive-Uncertainty Guided Knowledge Distillation for Accurate Classification of Student Misconceptions

Model ReleasesDGX agent

arXiv:2605.14752v1 Announce Type: cross Abstract: Accurately identifying student misconceptions is crucial for personalized education but faces three challenges: (1) data scarcity with long-tail distr

Collaborative Yet Personalized Policy Training: Single-Timescale Federated Actor-Critic

SafetyDGX agent

arXiv:2605.14423v1 Announce Type: cross Abstract: Despite the popularity of the actor-critic method and the practical needs of collaborative policy training, existing works typically either overlook e

Collider-Bench: Benchmarking AI Agents with Particle Physics Analysis Reproduction

Model ReleasesDGX agent

arXiv:2605.13950v1 Announce Type: cross Abstract: Autonomous language-model agents are increasingly evaluated on long-horizon tool-use tasks, but existing benchmarks rarely capture the complexity and

Complacent, Not Sycophantic: Reframing Large Language Models and Designing AI Literacy for Complacent Machines

SafetyDGX agent

arXiv:2605.14544v1 Announce Type: new Abstract: Large language models are often described as sycophantic, in the sense that they appear to flatter users or mirror their beliefs. We argue that this lab

Compositional Sparsity as an Inductive Bias for Neural Architecture Design

SafetyDGX agent

arXiv:2605.14764v1 Announce Type: cross Abstract: Identifying the structural priors that enable Deep Neural Networks (DNNs) to overcome the curse of dimensionality is a fundamental challenge in machin

Concurrency without Model Changes: Future-based Asynchronous Function Calling for LLMs

AgentsDGX agent

arXiv:2605.15077v1 Announce Type: cross Abstract: Function calling, also known as tool use, is a core capability of modern LLM agents but is typically constrained by synchronous execution semantics. U

Conditional Attribute Estimation with Autoregressive Sequence Models

Local AiDGX agent

arXiv:2605.14004v1 Announce Type: new Abstract: Generative models are often trained with a next-token prediction objective, yet many downstream applications require the ability to estimate or control

Conformal Thinking: Risk Control for Reasoning on a Compute Budget

ResearchDGX agent

arXiv:2602.03814v2 Announce Type: replace Abstract: Reasoning Large Language Models (LLMs) enable test-time scaling, with dataset-level accuracy improving as the token budget increases, motivating ada

Consciousness as Uncommon Self-Knowledge: A Synergistic Information Framework

ResearchDGX agent

arXiv:2605.13884v1 Announce Type: cross Abstract: We propose uncommon self-knowledge (USK) as a candidate criterion for consciousness: synergistic information a system carries about itself that exists

Contestable Multi-Agent Debate with Arena-based Argumentative Computation for Multimedia Verification

AgentsDGX agent

arXiv:2605.14495v1 Announce Type: cross Abstract: Multimedia verification requires not only accurate conclusions but also transparent and contestable reasoning. We propose a contestable multi-agent fr

COREKG: Coreset-Guided Personalized Summarization of Knowledge Graphs

ResearchDGX agent

arXiv:2605.14900v1 Announce Type: new Abstract: Knowledge Graphs (KGs) are extensively used across different domains and in several applications. Often, these KGs are very large in size. Such KGs beco

← Previous
1…250251252253254…358
Next →