AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
29 May 2026

Ultra-Reduced-Impact-Encased-Logging (URIEL): propose a new method for selective sustainable logging and post-harvest silvicultural treatment in tropical forest using airborne robotics systems

ResearchDGX agent

arXiv:2605.28883v1 Announce Type: new Abstract: Tropical forests worldwide are under intense deforestation pressure driven by economic and political interests, and scientific evidence suggests this de

Uncertainty-Aware Transfer Learning for Cross-Building Energy Forecasting: Toward Robust and Scalable District-Level Energy Management

Model ReleasesDGX agent

arXiv:2605.29733v1 Announce Type: new Abstract: Scaling data-driven energy forecasting to district level requires models that can be re-used across buildings with minimal target-domain data and honest

Unifying Temporal and Structural Credit Assignment in LLM-Based Multi-Agent Prompt Optimization


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Local AiDGX agent

arXiv:2605.30227v1 Announce Type: cross Abstract: While Multi-Agent Systems (MAS) empower Large Language Models to tackle complex reasoning tasks through collaborative interaction, optimizing their dy

unix-ctf: Procedural Environments for Unix-Competence Reinforcement Learning

AgentsDGX agent

arXiv:2605.29115v1 Announce Type: cross Abstract: Unix competence is the ability to use shell and operating-system primitives as first-class tools, not merely to write programs through a terminal. Cur

Unlocking the Working Memory of Large Language Models for Latent Reasoning

ResearchDGX agent

arXiv:2605.30343v1 Announce Type: cross Abstract: To improve the reasoning capabilities of large language models, test-time compute is typically scaled by generating intermediate tokens before the fin

Unveiling Multi-regime Patterns in SciML: Distinct Failure Modes and Regime-specific Optimization

ResearchDGX agent

arXiv:2605.29153v1 Announce Type: cross Abstract: Neural networks trained under different hyperparameter settings can fall into distinct training 'regimes,' with consistent behavior within regimes and

VFEAgent: A Multimodal Agent Framework for End-to-End Automated Finite Element Analysis

AgentsDGX agent

arXiv:2605.28978v1 Announce Type: new Abstract: Finite Element Analysis (FEA) serves as the cornerstone of modern engineering design. However, its workflow is inherently complex and relies heavily on

VideoMLA: Low-Rank Latent KV Cache for Minute-Scale Autoregressive Video Diffusion

HardwareDGX agent

arXiv:2605.30351v1 Announce Type: cross Abstract: Long-rollout causal video diffusion has converged on a fixed-size sliding-window KV cache, with recent progress innovating within this layout by chang

VikingMem: A Memory Base Management System for Stateful LLM-based Applications

AgentsDGX agent

arXiv:2605.29640v1 Announce Type: new Abstract: Large Language Models have revolutionized interactive applications; however, their finite context windows pose a critical data management challenge for

VisualThink-VLA: Visual Intermediate Reasoning for Effective and Low-Latency Vision-Language-Action Policies

AgentsDGX agent

arXiv:2605.30011v1 Announce Type: cross Abstract: Recent work has begun to equip vision-language-action (VLA) policies with explicit intermediate reasoning. In embodied control, however, textual chain

VitalAgent: A Tool-Augmented Agent for Reactive and Proactive Physiological Monitoring over Wearable Health Data

Model ReleasesDGX agent

arXiv:2605.29483v1 Announce Type: new Abstract: Wearable devices enable continuous monitoring of physiological signals such as ECG and PPG, but existing mHealth systems are largely limited to task-spe

VLA-Pro: Cross-Task Procedural Memory Transfer for Vision-Language-Action Models

ApplicationsDGX agent

arXiv:2605.29562v1 Announce Type: cross Abstract: Vision-Language-Action~(VLA) models have shown strong potential for general-purpose robotic manipulation, yet they still struggle to generalize to uns

VLA-Trace: Diagnosing Vision-Language-Action Models through Representation and Behavior Tracing

SafetyDGX agent

arXiv:2605.30117v1 Announce Type: new Abstract: Understanding how Vision-Language-Action (VLA) models transform multimodal knowledge into embodied control remains an open challenge. We present VLA-Tra

Wait! There's a Way Out: A Decision Mechanism for Forecasting Conversational Derailment

ResearchDGX agent

arXiv:2605.29243v1 Announce Type: cross Abstract: Forecasting conversational derailment is the task of predicting, as the conversation unfolds, whether it will eventually derail into personal attacks.

Weakly Supervised Detection and Temporal Localization of Whale Calls in Long-Duration Bioacoustic Data

ResearchDGX agent

arXiv:2502.20838v3 Announce Type: replace-cross Abstract: Passive acoustic monitoring (PAM) systems generate continuous recordings spanning months, yet automated bioacoustic analysis of whale calls re

What drives performance in molecular MPNNs? An operator-level factorial benchmark

Model ReleasesDGX agent

arXiv:2605.30195v1 Announce Type: cross Abstract: Message-passing neural networks (MPNNs) are widely used for molecular property prediction, but their deployment as monolithic architectures makes it d

When and How Human Curation Backfires: Preference Alignment under Multi-Model Self-Consuming Loop

SafetyDGX agent

arXiv:2605.29267v1 Announce Type: new Abstract: Foundation models are increasingly trained on synthetic data generated by prior model iterations rather than exclusively on real data. This self-consumi

When and How Long? The Readout-Mediator Angle in Temporal Reasoning

SafetyDGX agent

arXiv:2605.29126v1 Announce Type: cross Abstract: A linear probe can decode a representation almost perfectly and yet be completely irrelevant to how the model uses it. On calendar-date duration reaso

When Cloud Agents Meet Device Agents: Lessons from Hybrid Multi-Agent Systems

Local AiDGX agent

arXiv:2605.30102v1 Announce Type: cross Abstract: The design space of agentic AI inference spans two extremes: frontier large language models (LLMs), typically hosted in the cloud and offering strong

When Does Persona Prompting Actually Help? A Retrieval and Metric Analysis of Expert Role Injection in LLMs

ApplicationsDGX agent

arXiv:2605.29420v1 Announce Type: new Abstract: Persona prompting is widely used to steer large language models, yet its practical value remains unclear. Prior work often evaluates persona prompting u

When Models Disagree: Rethinking LLM Evaluation for Public Comment Analysis

ResearchDGX agent

arXiv:2605.29025v1 Announce Type: new Abstract: Federal agencies are deploying large language models (LLMs) to categorize public comment corpora, where the model's organization of the record shapes wh

When Models Learn to Ask Why: Adaptive Causal Reasoning for Trustworthy Medical Vision-Language Models

TutorialsDGX agent

arXiv:2603.23085v2 Announce Type: replace Abstract: Vision-Language Models (VLMs) have enabled interpretable medical diagnosis by integrating visual perception with linguistic reasoning. Yet, existing

When Should a Robot Think? Resource-Aware Reasoning via Reinforcement Learning for Embodied Robotic Decision-Making

Model ReleasesDGX agent

arXiv:2603.16673v4 Announce Type: replace-cross Abstract: Embodied robotic systems increasingly rely on large language model (LLM)-based agents to support high-level reasoning, planning, and decision-

When Should Models Change Their Minds? Contextual Belief Management in Large Language Models

Model ReleasesDGX agent

arXiv:2605.30219v1 Announce Type: new Abstract: Long-horizon interactions require language models to manage accumulating information: when to update their state, when to preserve their state, and what

Who can we trust? LLM-as-a-jury for Comparative Assessment

Model ReleasesDGX agent

arXiv:2602.16610v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly applied as automatic evaluators for natural language generation assessment often using pairwise

Why Specialist Models Still Matter: A Heterogeneous Multi-Agent Paradigm for Medical Artificial Intelligence

Model ReleasesDGX agent

arXiv:2605.29744v1 Announce Type: new Abstract: The impressive performance of generalist large language models (LLMs) such as GPT and Claude in healthcare raises a critical question: will domain-speci

Xetrieval: Mechanistically Explaining Dense Retrieval

ResearchDGX agent

arXiv:2605.29507v1 Announce Type: new Abstract: Explaining why dense retrievers assign high relevance scores remains challenging because retrieval decisions are made through opaque high-dimensional em

xModel-KD: Cross-modal Knowledge Distillation for 3D Scene Perception using LiDAR

SafetyDGX agent

arXiv:2605.30111v1 Announce Type: cross Abstract: Point cloud segmentation is a fundamental task in 3D scene understanding. Its progress is constrained by the high cost and time required for dense 3D

28 May 2026

A Comparative Study of Rule-Based and Data-Driven Approaches in Industrial Monitoring

SafetyDGX agent

arXiv:2509.15848v2 Announce Type: replace Abstract: Industrial monitoring systems, especially when deployed in Industry 4.0 environments, are experiencing a shift in paradigm from traditional rule-bas

A Conflict-Aware Penalty and Statistical Loss Framework for Balancing Modalities and Enhancing Stability in Multimodal Sentiment Analysis

ResearchDGX agent

arXiv:2605.28575v1 Announce Type: new Abstract: Multimodal Sentiment Analysis (MSA) fuses text, acoustic, and visual streams to infer sentiment. Because pre-trained text encoders are far more expressi

A Fixed-Budget, Cluster-Aware Standard for LLM-as-a-Judge Evaluation: A Multi-Hop RAG Stress Test

ResearchDGX agent

arXiv:2605.27789v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) systems are often compared by asking a large language model (LLM) judge which answer is better. For multi-hop RAG,

A Fresh Look at Lamarckian Evolution and the Baldwin Effect

Model ReleasesDGX agent

arXiv:2605.28703v1 Announce Type: cross Abstract: Baldwinian and Lamarckian evolution have existed for a long time in evolutionary algorithms (EAs) without ever dominating the academic literature or p

A Matter of TASTE: Improving Coverage and Difficulty of Agent Benchmarks

Model ReleasesDGX agent

arXiv:2605.28556v1 Announce Type: new Abstract: As agent capabilities advance, existing benchmarks, such as au^2-Bench, are becoming increasingly saturated. Yet constructing new benchmark tasks remain

A Multi-dimensional Framework for Evaluating Generalization in EEG Foundation Models

Model ReleasesDGX agent

arXiv:2605.28563v1 Announce Type: cross Abstract: Evaluating foundation models under appropriate adaptation settings is essential for understanding the quality and transferability of the learned repre

A Policy-Driven Runtime Layer for Agentic LLM Serving

SafetyDGX agent

arXiv:2605.27744v1 Announce Type: new Abstract: Multi-agent LLM systems have become the dominant production workload, but the serving stack was not built for them. The agent framework above knows agen

A Query Engine for the Agents

Model ReleasesDGX agent

arXiv:2605.27785v1 Announce Type: new Abstract: The fastest-growing data in production today is unstructured text: agent traces, chat logs, reasoning chains, model outputs. People want to analyze it,

A Sheaf-Theoretic and Topological Perspective on Complex Network Modeling and Attention Mechanisms in Graph Neural Models

ApplicationsDGX agent

arXiv:2601.21207v3 Announce Type: replace-cross Abstract: Combinatorial and topological structures, such as graphs, simplicial complexes, and cell complexes, form the foundation of geometric and topol

A Systematic Evaluation of Retrieval-Augmented Generation and Language Models for Space Operations

ResearchDGX agent

arXiv:2605.27444v1 Announce Type: cross Abstract: The rapid expansion of space activities has led to an unprecedented accumulation of technical documentation, operational guidelines, and scientific li

A Unified Framework for the Evaluation of LLM Agentic Capabilities

Model ReleasesDGX agent

arXiv:2605.27898v1 Announce Type: new Abstract: As LLMs are increasingly deployed as agents, reliable assessment of their agentic capabilities has become essential. However, reported benchmark scores

AdaMerge: Salience-Aware Adaptive Token Merging for Training-Free Acceleration of Vision Transformers

ResearchDGX agent

arXiv:2605.27465v1 Announce Type: cross Abstract: The quadratic cost of self-attention in Vision Transformers (ViTs) constitutes a fundamental bottleneck for practical deployment, motivating a vibrant

Adapting, Fast and Slow: On Few-Shot Transportability of Compositions

ResearchDGX agent

arXiv:2512.22777v2 Announce Type: replace-cross Abstract: Generalization across domains requires stable structure that links the source and target distributions. Building on causal transportability th

Adapting the Interface, Not the Model: Runtime Harness Adaptation for Deterministic LLM Agents

AgentsDGX agent

arXiv:2605.22166v2 Announce Type: replace Abstract: LLM agents are shaped not only by their language models, but also by the runtime harness that mediates observation, tool use, action execution, feed

Adaptive Multimodal Agents-Based Framework for Automatic Workflow Execution

AgentsDGX agent

arXiv:2605.28607v1 Announce Type: new Abstract: Modern information systems require autonomous agents capable of navigating complex workflows, yet current methodologies often struggle with the transiti

Adaptive Reservoir Computing for Multi-Scenario Chaotic System Forecasting

Model ReleasesDGX agent

arXiv:2605.28145v1 Announce Type: new Abstract: We present an adaptive reservoir computing framework for the CTF-4-Science Lorenz benchmark, which evaluates machine learning models across twelve disti

Advancing Direct Training for Spiking Neural Networks with Circulate-Firing Neurons and Learnable Gradients

ResearchDGX agent

arXiv:2605.27412v1 Announce Type: cross Abstract: Spiking Neural Networks (SNNs) have emerged with promising energy-efficient property, yet a substantial performance gap persists compared to Artificia

ADWIN: Adaptive Windows for Horizon-Aware On-Policy Distillation

SafetyDGX agent

arXiv:2605.28396v1 Announce Type: cross Abstract: On-policy distillation (OPD) transfers reasoning behavior by training a student on teacher feedback along student-generated trajectories, but standard

AgensFlow: A Coordination-Policy Substrate for Multi-Agent Systems

SafetyDGX agent

arXiv:2605.27466v1 Announce Type: cross Abstract: Multi-agent systems built on large language models (LLMs) require many coordination choices that are difficult to fix a priori: which skill protocol t

Agentic Active Omni-Modal Perception for Multi-Hop Audio-Visual Reasoning

Model ReleasesDGX agent

arXiv:2605.28192v1 Announce Type: new Abstract: Multi-hop audio-visual reasoning remains challenging for Omni-LLMs, as relevant evidence is often sparse, temporally dispersed, and distributed across b

Agentic Literacy Debt: A Structural Problem the AI Literacy Field Has Not Yet Named

AgentsDGX agent

arXiv:2605.27396v1 Announce Type: cross Abstract: Autonomous AI agents now plan, decide, and act on behalf of users across healthcare, financial services, and workplace contexts, often without step-by

Agyn: An Open-Source Platform for AI Agents with Scalable On-Demand Execution, Agent Definition as a Code, and Zero-Trust Access

AgentsDGX agent

arXiv:2605.27575v1 Announce Type: new Abstract: As organizations move toward production deployments of AI agents, which execute non-deterministic workflows, maintain stateful sessions, and often opera

AI in the Workplace: The Impact of AI on Perceived Job Decency and Meaningfulness

ApplicationsDGX agent

arXiv:2605.28680v1 Announce Type: cross Abstract: The proliferation of Artificial Intelligence (AI) in workplaces is transforming how we work. While existing research on human-AI collaboration at work

AI, Take the Wheel: What Drives Delegation and Trust in Human-Computer Cooperative Question Answering?

SafetyDGX agent

arXiv:2605.28255v1 Announce Type: new Abstract: AI systems are fallible, and humans can make mistakes in deciding whether to trust AI over their own judgment. Thus, improving human-AI collaboration re

AIBuildAI-2: A Knowledge-Enhanced Agent for Automatically Building AI Models

AgentsDGX agent

arXiv:2605.27873v1 Announce Type: new Abstract: AI models underpin data-centric applications from image and text processing to scientific discovery in biology, physics, and chemistry. Yet developing t

Aligning Language Model Benchmarks with Pairwise Preferences

Model ReleasesDGX agent

arXiv:2602.02898v2 Announce Type: replace Abstract: Language model benchmarks are pervasive and computationally-efficient proxies for real-world performance. However, many recent works find that bench

Aligning LLMs with Human Uncertainty: A Beta-Bernoulli Calibrator for LLM Forecasting

TutorialsDGX agent

arXiv:2605.27668v1 Announce Type: cross Abstract: Probabilistic forecasting estimates the likelihood of uncertain future events. To improve LLM forecasting, existing methods typically learn from binar

AlphaForgeBench: Benchmarking End-to-End Trading Strategy Design with Large Language Models

Model ReleasesDGX agent

arXiv:2602.18481v2 Announce Type: replace-cross Abstract: The rapid advancement of Large Language Models (LLMs) has led to a surge of financial benchmarks, evolving from static knowledge evaluation to

AlphaTransit: Learning to Design City-scale Transit Routes

Model ReleasesDGX agent

arXiv:2605.28730v1 Announce Type: new Abstract: Designing a transit network requires many sequential route extension decisions, but their quality is often visible only after the full network is assemb

An Empirical Audit of k-NAF Budget Accounting for Anchored Decoding

ResearchDGX agent

arXiv:2605.28001v1 Announce Type: new Abstract: We empirically audit the k-NAF budget-accounting mechanism in Anchored Decoding using (i) a fixed, class-stratified workload (approximately 8,500 random

An Enhanced Large Neighborhood Search Approach for the Capacitated Facility Location Problem with Incompatible Customers

Model ReleasesDGX agent

arXiv:2605.28337v1 Announce Type: new Abstract: A new variant of the classic capacitated facility location problem, which considers incompatibilities between customers, has recently been introduced in

An LLM-Based Assistance System for Intuitive and Flexible Capability-Based Planning

AgentsDGX agent

arXiv:2605.28666v1 Announce Type: new Abstract: In modern industry, dynamic environments and the complexity of modular and reconfigurable resources require automated planning of process sequences. Cap

← Previous
1…195196197198199…358
Next →