AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
10 Jun 2026

Can Image Models Imagine Time? ImageTime: A Novel Benchmark for Probing Visual World Modeling Through Spatiotemporal Consistency

Model ReleasesDGX agent

arXiv:2606.10620v1 Announce Type: cross Abstract: Image generation models now produce high-quality static images, yet their ability to represent how a visual world changes over time remains poorly und

Can Multi-Agent LLMs Identify Their Peers? Stylometric Fingerprinting in Role-Constrained Political Analysis

Model ReleasesDGX agent

arXiv:2606.09854v1 Announce Type: cross Abstract: Multi-agent large language model (LLM) pipelines for political statement analysis are vulnerable to peer-preservation bias: models tend to protect pee

CANVAS: Captioning Art with Narrative Visual-Audio AI Systems

Applications

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2606.09846v1 Announce Type: cross Abstract: Visual art remains largely inaccessible to blind and low-vision (BLV) audiences due to brief or absent alt-text, which rarely conveys the sensory, spa

Catching One in Five: LLM-as-Judge Blind Spots in Production Multi-Turn Transaction Agents

AgentsDGX agent

arXiv:2606.10315v1 Announce Type: cross Abstract: LLM-as-judge is the default instrument for evaluating conversational agents, yet its reliability is almost always reported as agreement with human rat

Causal Ensemble Agent: Hierarchical Causal Discovery with LLM-guided Expert Reweighting

SafetyDGX agent

arXiv:2606.10607v1 Announce Type: cross Abstract: Causal discovery aims to uncover causal structures from observational data, which is crucial for real-world decision-making. However, different causal

ChartAgent: A Multimodal Agent for Visually Grounded Reasoning in Complex Chart Question Answering

AgentsDGX agent

arXiv:2510.04514v3 Announce Type: replace Abstract: Recent multimodal LLMs have shown promise in chart-based visual question answering, but their performance declines sharply on unannotated charts-tho

CIAware-Bench: Benchmarking Control Intervention Awareness Across Frontier LLMs

Model ReleasesDGX agent

arXiv:2606.11063v1 Announce Type: new Abstract: AI control protocols oversee untrusted models by monitoring their actions and modifying potentially unsafe steps, often using a trusted model. This part

CITRAS: Covariate-Informed Transformer for Time Series Forecasting

ApplicationsDGX agent

arXiv:2503.24007v4 Announce Type: replace-cross Abstract: In time series forecasting, covariates represent external factors that influence target variables. Some covariates are observable only in the

CleanPatrick: A Benchmark for Image Data Cleaning

Model ReleasesDGX agent

arXiv:2505.11034v2 Announce Type: replace-cross Abstract: Robust machine learning depends on clean data, yet current image data cleaning benchmarks rely on synthetic noise or narrow human studies, lim

CLP: Collocation-Length Prediction for Zero-Loss Adaptive Multi-Token Inference

Model ReleasesDGX agent

arXiv:2606.10935v1 Announce Type: cross Abstract: Large language model inference is bottlenecked by autoregressive decoding, where each token requires a full forward pass. Multi-token prediction (MTP)

Co-GLANCE: Uncertainty-Aware Active Perception for Heterogeneous Robot Teaming

ApplicationsDGX agent

arXiv:2606.09919v1 Announce Type: cross Abstract: Perceptual uncertainty is a central challenge for heterogeneous robot teams operating in unstructured outdoor environments, where no single viewpoint

CollabSkill: Evaluating Human-Agent Collaboration On Real-World Tasks

Model ReleasesDGX agent

arXiv:2606.09833v1 Announce Type: cross Abstract: AI agents are reshaping the workspace, leading to drastic change of how humans work. Despite the considerable potential of human-agent collaboration b

ComBench: A Benchmark for Rigorous Proof Reasoning and Constructive Realization in Olympiad-Level Combinatorics

Model ReleasesDGX agent

arXiv:2606.10479v1 Announce Type: new Abstract: Combinatorics is central to Olympiad-level mathematical problem solving, requiring deep discrete reasoning, creative constructions, and rigorous structu

Conditional Vendi Score: Prompt-Aware Diversity Evaluation for Generative AI Models and LLMs

SafetyDGX agent

arXiv:2411.02817v2 Announce Type: replace-cross Abstract: Generative models guided by text prompts are widely evaluated for fidelity and prompt alignment, yet their ability to produce outputs remains

Conformal Prediction for Neural Operators: Distribution-Free Uncertainty Quantification in Physics Simulation

SafetyDGX agent

arXiv:2606.09923v1 Announce Type: cross Abstract: Neural operators such as the Fourier Neural Operator (FNO) have emerged as powerful surrogates for solving partial differential equations (PDEs), achi

Conformal Risk Prediction for Non-Alcoholic Fatty Liver Disease Using Gradient Boosting with Distribution-Free Coverages

ResearchDGX agent

arXiv:2606.09860v1 Announce Type: cross Abstract: Non-alcoholic fatty liver disease (NAFLD) affects roughly 25% of global adults, posing substantial hepatic and cardiovascular risks. Yet, population-l

Constructing coherent spatial memory in LLM agents through graph rectification

Model ReleasesDGX agent

arXiv:2510.04195v2 Announce Type: replace Abstract: Given a map description through global traversal navigation instructions, an LLM can often infer the implicit spatial layout and answer user queries

Content-Induced Spatial-Spectral Aggregation Network for Change Detection in Remote Sensing Images

TutorialsDGX agent

arXiv:2606.10328v1 Announce Type: cross Abstract: The integration of spatial and spectral information is beneficial to the improvement of change detection performance. However, existing methods cannot

Convergence of Monte Carlo Optimistic Policy Iteration: Beyond Uniform State-Action Updates

SafetyDGX agent

arXiv:2606.10580v1 Announce Type: cross Abstract: The asymptotic behaviour of Monte Carlo optimistic policy iteration (MC-O-PI) is a long-standing open question. When the model of the environment is u

Cross-Modal Knowledge Distillation without Paired Data: Theoretical Foundation and Algorithm

SafetyDGX agent

arXiv:2606.10504v1 Announce Type: new Abstract: Cross-modal knowledge distillation (CMKD) studies how a (large) teacher model trained on one type of data (e.g., images) can guide a (smaller) student m

Culturally-Aware AI for Cross-Boundary Community Learning: Undergraduate Innovation at the Intersection of Computation and Design

ApplicationsDGX agent

arXiv:2606.09041v1 Announce Type: cross Abstract: Research on artificial intelligence in education (AIED) is rapidly expanding, yet technical progress often lacks human-centered grounding and adequate

Data assimilation for subsurface flow using latent diffusion model parameterization: performance of ensemble-Kalman and Monte Carlo techniques

ResearchDGX agent

arXiv:2606.11140v1 Announce Type: cross Abstract: Data assimilation (DA) in subsurface flow entails calibrating model parameters to match observed data, typically at wells, while preserving geological

Decentralized Multi-Agent Systems with Shared Context

Local AiDGX agent

arXiv:2606.10662v1 Announce Type: cross Abstract: Multi-agent systems (MAS) can scale large language model reasoning at test time by decomposing complex problems into parallel subtasks. However, most

Decoupling Thought from Speech: Knowledge-Grounded Counterfactual Reasoning for Resilient Multi-Agent Argumentation

SafetyDGX agent

arXiv:2606.10475v1 Announce Type: cross Abstract: Multi-agent debate frameworks have been shown to improve large language model performance in convergent tasks, but they are currently optimized in a w

Deep Generative Model for Human Mobility Behavior

ResearchDGX agent

arXiv:2510.06473v3 Announce Type: replace-cross Abstract: Understanding and modeling human mobility is central to challenges in transport planning, sustainable urban design, and public health. Despite

Deep Slice Interpolation for Reducing Through-Plane Anisotropy and Noise in Head CT

ResearchDGX agent

arXiv:2606.09953v1 Announce Type: cross Abstract: Head computed tomography (CT) typically uses sub-millimeter in-plane resolution but 2-5 mm through-plane spacing, creating substantial anisotropy that

Democratising Camera Trap AI: An Open-Source Model for Detecting UK Mammals

ResearchDGX agent

arXiv:2606.10940v1 Announce Type: cross Abstract: Camera traps have become a cornerstone of biodiversity monitoring, but the artificial intelligence that turns vast quantities of images into usable ec

Density Ridge Selective Prediction for LLM and VLM Hallucination Detection under Calibration Label Scarcity

ResearchDGX agent

arXiv:2606.10198v1 Announce Type: cross Abstract: Hallucination detection in large language and vision-language models is increasingly framed as selective prediction, where a detector assigns a confid

Dep-LLM: Training-Free Depression Diagnosis via Evidence-Guided Structured Multi-factor with Reliable LLM Reasoning

TutorialsDGX agent

arXiv:2606.10796v1 Announce Type: cross Abstract: Automatic Depression Detection (ADD) from clinical interviews is a pivotal task in computational mental health, yet it remains challenging due to two

Deployment-Time Memorization in Foundation-Model Agents

Model ReleasesDGX agent

arXiv:2606.10062v1 Announce Type: new Abstract: Foundation-model agents are increasingly long-lived systems that remember users across interactions, making memorization an explicit deployment-time fun

DeRA-MOS: Optimizing Text-to-Music Evaluation via Decoupled Listwise Ranking and Modality Alignment

SafetyDGX agent

arXiv:2606.10010v1 Announce Type: cross Abstract: Evaluating text-to-music (TTM) systems remains expensive because music impression (MI) and text alignment (TA) scores rely on human mean opinion score

Designed by Journalists, but Is It for Readers? Rethinking AI Disclosures and Transparency in News

TutorialsDGX agent

arXiv:2606.11116v1 Announce Type: cross Abstract: As newsrooms integrate generative AI, journalists face a disclosure challenge: how to communicate AI involvement in ways that maintain reader trust. C

Detecting Knowledge Gaps from Conversational AI Interactions Using Curriculum Prerequisite Graphs

Model ReleasesDGX agent

arXiv:2606.10736v1 Announce Type: cross Abstract: Large online courses generate thousands of student questions directed at conversational AI teaching assistants, yet these interaction logs remain larg

Detecting Speculative Language in Biomedical Texts using Recurrent Neural Tensor Networks

ResearchDGX agent

arXiv:2606.10471v1 Announce Type: cross Abstract: In this investigation, we delve into the automated detection of speculative language within biomedical articles by utilizing distributed sentence repr

Diffusion Forcing Planner: History-Annealed Planning with Time-Dependent Guidance for Autonomous Driving

SafetyDGX agent

arXiv:2606.11019v1 Announce Type: cross Abstract: Learning-based motion planners, despite recent progress, often suffer from temporal inconsistency. Small perturbations across frames can accumulate in

Divide-and-Conquer Modeling for the CTF-4-Science Lorenz Benchmark

Model ReleasesDGX agent

arXiv:2606.10084v1 Announce Type: cross Abstract: This work presents a divide-and-conquer modeling strategy for the CTF-4-Science Lorenz benchmark, which evaluates chaotic-system prediction across twe

Divide and Cooperate: Role-Decomposed Multi-Agent LLM Training with Cross-Agent Learning Signals

Model ReleasesDGX agent

arXiv:2606.10684v1 Announce Type: cross Abstract: Modern language agents which perform multi-step reasoning have shown strong performance in knowledge-intensive question answering. However, existing a

Dmsh: A Multi-Agent Reinforcement Learning Framework for All-Quad Mesh Generation

AgentsDGX agent

arXiv:2606.10601v1 Announce Type: cross Abstract: Generating high-quality meshes for arbitrary geometries remains a fundamental bottleneck in computational engineering, often demanding heuristic tunin

Do VLMs Reason Like Engineers? A Benchmark and a Stage-wise Evaluation

Model ReleasesDGX agent

arXiv:2606.10833v1 Announce Type: new Abstract: Vision-Language Models (VLMs) demonstrate strong performance on general multimodal reasoning benchmarks, yet their ability to perform engineering reason

Does Normalization Choice Matter for Causal Large Time-Series Models?

ApplicationsDGX agent

arXiv:2606.09954v1 Announce Type: cross Abstract: Large models for time-series forecasting have been emerged as a promising paradigm for training models on heterogeneous collections of signals. These

Drawing with Strangers: Population Scaling Drives Zero-Shot Mutual Intelligibility in Emergent Sketching

ResearchDGX agent

arXiv:2606.10582v1 Announce Type: cross Abstract: Generalization in emergent communication has largely focused on novel inputs or linguistic structures, yet the capacity for agents to communicate with

Dropout-GRPO: Variational Stochasticity for Continuous Latent Reasoning

SafetyDGX agent

arXiv:2606.10184v1 Announce Type: cross Abstract: Group Relative Policy Optimization (GRPO) relies on the diversity of K rollouts within each group; otherwise, the group-mean advantage A^{(k)} = r^{(k

Dual-Branch Gated Fusion for Open-Set Audio Deepfake Source Tracing

Model ReleasesDGX agent

arXiv:2606.10223v1 Announce Type: cross Abstract: Attributing a synthetic utterance to its originating system remains an open challenge: closed-set models fail to reject unseen synthesizers and produc

Duality for Optimal Multi-Item, Multi-Bidder Auction Design: Revenue Certificates through Deep Learning

ResearchDGX agent

arXiv:2606.10112v1 Announce Type: cross Abstract: Characterizing revenue-optimal auctions for multi-item, multi-bidder settings remains a fundamental open problem, with no known closed-form solution e

Dynamic Linear Attention

ResearchDGX agent

arXiv:2606.10650v1 Announce Type: cross Abstract: The scalability of Large Language Models (LLMs) to long contexts is fundamentally constrained by the quadratic complexity of standard attention, motiv

Dynamics of Adversarial Attacks on Large Language Model-Based Search Engines

ResearchDGX agent

arXiv:2501.00745v3 Announce Type: replace-cross Abstract: The increasing integration of Large Language Model (LLM) based search engines has transformed the landscape of information retrieval. However,

Earth-OneVision: Extending Remote Sensing Multimodal Large Language Models to More Sensor Modalities and Tasks

Model ReleasesDGX agent

arXiv:2606.10819v1 Announce Type: cross Abstract: RS-MLLMs enable natural-language understanding and spatial reasoning over earth observation imagery. However, existing models support only a narrow ra

EEVEE: Towards Test-time Prompt Learning in the Real World for Self-Improving Agents

Model ReleasesDGX agent

arXiv:2606.11182v1 Announce Type: cross Abstract: In this paper, we propose EEVEE, the first multi-dataset test-time prompt learning framework for LLM agents, enabling test-time prompt learning under

Effective Reinforcement Learning for Agentic Search by Recycling Zero-Variance Queries During Training

Model ReleasesDGX agent

arXiv:2606.10709v1 Announce Type: cross Abstract: The use of GRPO-style algorithms has become the standard strategy for training LLM search agents under outcome-only rewards. With these algorithms, a

Embedding Hybrid Systems into Continuous Latent Vector Fields

ResearchDGX agent

arXiv:2606.10596v1 Announce Type: cross Abstract: This work proves that an n-dimensional hybrid system can be embedded into an m-dimensional Euclidean space equipped with a continuous vector field on

Emotion Profiling in LLM-Based Literary Translation: Systematic Shifts Across MT and Post-Editing

ResearchDGX agent

arXiv:2606.10113v1 Announce Type: cross Abstract: This paper investigates whether LLM translations exhibit identifiable emotional profiles and how post-editing reshapes them toward human-like norms. W

ERAlign: Energy-based Representation Alignment of GNNs and LLMs on Text-attributed Graphs

SafetyDGX agent

arXiv:2606.10461v1 Announce Type: cross Abstract: Text-attributed Graphs (TAGs) incorporate textual node attributes with graph structures to describe rich relational semantics. Recent efforts to integ

EstRTL: Functional Estimation Guided RTL Code Generation

AgentsDGX agent

arXiv:2606.09867v1 Announce Type: cross Abstract: Optimizing register transfer level (RTL) code is of vital importance in hardware design. Large language models (LLMs) provide new methods for the auto

Ethical and Technical Limits of Deepfake Speech Datasets

SafetyDGX agent

arXiv:2606.10911v1 Announce Type: cross Abstract: Claims about the robustness and fairness of deepfake speech detectors are only as credible as the datasets used to train and evaluate those systems. W

Evaluating Research-Level Math Proofs via Strict Step-Level Verification

Model ReleasesDGX agent

arXiv:2606.10799v1 Announce Type: new Abstract: Large Language Models (LLMs) struggle to rigorously verify complex mathematical proofs. Standard global evaluation approaches suffer from 'context poiso

Event-Driven Reinforcement Learning Enables Long-Horizon Control in Semiconductor Fabrication

SafetyDGX agent

arXiv:2606.10705v1 Announce Type: cross Abstract: Reinforcement learning promises to optimize sequential decisions in large-scale systems. Semiconductor manufacturing systems are stochastic and highly

Expert-Level Crisis Detection in Mental Health Conversations

Model ReleasesDGX agent

arXiv:2606.10380v1 Announce Type: cross Abstract: Real-world crisis intervention is inherently conversational, yet existing research largely focuses on static texts.Real-world crisis intervention is i

Exploration of Foundation Model-Based Robots in Patient and Elderly Care

ResearchDGX agent

arXiv:2606.10208v1 Announce Type: cross Abstract: Demand for older-adult and patient care is growing rapidly as populations age worldwide. Foundation models are increasingly being integrated into robo

Exploratory Responsiveness and Adaptive Rigidity under AI-Assisted Optimization

Model ReleasesDGX agent

arXiv:2606.10086v1 Announce Type: new Abstract: This paper develops a theory of exploratory adaptation under AI-assisted optimization. The central argument is that the long-run adaptive effects of AI

Exploring Accurate and Transparent Domain Adaptation in Predictive Healthcare via Concept-Grounded Orthogonal Inference

SafetyDGX agent

arXiv:2602.12542v2 Announce Type: replace-cross Abstract: Deep learning models for clinical event prediction on electronic health records (EHR) often suffer performance degradation when deployed under

← Previous
1…135136137138139…358
Next →