AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlog
85,136Total entries
1Added by human
85,135Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,688 results
Research

CoHyDE: Iterative Co-Training of LLM Rewriter & Dense Encoder for Tool Retrieval

DGX agent

arXiv:2605.29271v1 Announce Type: new Abstract: Tool retrieval over large API catalogs is a core bottleneck for LLM agents: user queries arrive in colloquial, often underspecified language, while the

researcharxiv-cs-ai
29 May 2026
Model Releases
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Combating Data Laundering in LLM Training

DGX agent

arXiv:2604.01904v2 Announce Type: replace-cross Abstract: Data rights owners can detect unauthorized data use in large language model (LLM) training by querying with proprietary samples. Often, superi

model-releasesarxiv-cs-ai
29 May 2026
Research

COMET: Concept Space Dissection of the Modality Gap in Audio-Text Multimodal Contrastive Embeddings

DGX agent

arXiv:2605.29628v1 Announce Type: cross Abstract: Contrastive Language-Audio Pretraining (CLAP) models are widely used for audio understanding and support modality-agnostic condition swapping in many

researcharxiv-cs-ai
29 May 2026
Research

Comparing Post-Hoc Explainable AI Methods for Interpreting Black-Box EEG Models in Depression Detection

DGX agent

arXiv:2605.28977v1 Announce Type: cross Abstract: Recent advances in deep learning have enabled increasingly accurate electroencephalography (EEG)-based classification of Major Depressive Disorder (MD

researcharxiv-cs-ai
29 May 2026
Agents

Compass: Navigating Global Marine Lead Data Integration through Expert-Guided LLM Agent

DGX agent

arXiv:2605.29966v1 Announce Type: new Abstract: Marine lead (Pb) and its isotopes are critical tracers for ocean circulation and anthropogenic pollution, yet in-situ observations remain costly and spa

agentsarxiv-cs-ai
29 May 2026
Model Releases

Composing Non-Conjugate Factor Graphs with Closed-Form Variational Inference

DGX agent

arXiv:2605.29467v1 Announce Type: cross Abstract: Stacking probabilistic building blocks into deeper architectures typically breaks closed-form inference. We show that closed-form inference can be pre

model-releasesarxiv-cs-ai
29 May 2026
Research

Compute Allocation in Evolutionary Search: From Depth-Breadth to Multi-Armed Bandits

DGX agent

arXiv:2605.29268v1 Announce Type: cross Abstract: LLM-guided evolutionary search (Evolve systems) has reached state-of-the-art results on mathematical and combinatorial tasks, yet most existing system

researcharxiv-cs-ai
29 May 2026
Research

Conf-Gen: Conformal Uncertainty Quantification for Generative Models

DGX agent

arXiv:2605.28920v1 Announce Type: cross Abstract: Conformal prediction (CP) and its extension, conformal risk control (CRC), are established frameworks for quantifying uncertainty in supervised machin

researcharxiv-cs-ai
29 May 2026
Research

Conformal Certification of Reasoning Trace Prefixes

DGX agent

arXiv:2605.30085v1 Announce Type: new Abstract: Language model reasoning traces are rarely all-or-nothing; they frequently contain valid intermediate steps before a critical error occurs. Existing unc

researcharxiv-cs-ai
29 May 2026
Model Releases

ConMoE: Expert-Pool Consolidation via Prototype Reassignment for MoE Compression

DGX agent

arXiv:2605.29350v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) language models reduce per-token computation but still require storing and serving all experts, making deployment memory-intens

model-releasesarxiv-cs-ai
29 May 2026
Research

Context Distillation as Latent Memory Management

DGX agent

arXiv:2605.28889v1 Announce Type: cross Abstract: Context distillation compresses contextual information into model parameters, yet existing methods often ignore how multiple distilled latent memories

researcharxiv-cs-ai
29 May 2026
Research

Continuity and Ordinality Matter: Constraining Time Series Tokens for Effective Time Series Analysis with Large Language Models

DGX agent

arXiv:2605.28866v1 Announce Type: cross Abstract: Token-based time series large language models (TS-LLMs) have emerged as a promising direction for time series analysis and reasoning. However, prior s

researcharxiv-cs-ai
29 May 2026
Research

Controlling the Risk of Corrupted Contexts for Language Models via Early-Exiting

DGX agent

arXiv:2510.02480v3 Announce Type: replace Abstract: Large language models (LLMs) can be influenced by harmful or irrelevant context, which can significantly harm model performance on downstream tasks.

researcharxiv-cs-ai
29 May 2026
Model Releases

Cookie-Bench: Continuous On-screen Key Interaction Evaluation for Web Generation

DGX agent

arXiv:2605.30000v1 Announce Type: new Abstract: Front-end web code has become a core product surface for every frontier LLM release, yet evaluating these interactive applications at development speed

model-releasesarxiv-cs-ai
29 May 2026
Research

CORE-T: COherent REtrieval of Tables for Text-to-SQL

DGX agent

arXiv:2601.13111v2 Announce Type: replace-cross Abstract: Realistic text-to-SQL workflows often require joining multiple tables. As a result, accurately retrieving the relevant set of tables becomes a

researcharxiv-cs-ai
29 May 2026
Model Releases

CosmicFish-HRM: Adaptive Reasoning via Hierarchical Recurrent Mechanisms in Compact Language Models

DGX agent

arXiv:2605.28919v1 Announce Type: cross Abstract: Large language models have achieved strong reasoning capabilities, though often at the cost of massive parameter counts and expensive inference. In th

model-releasesarxiv-cs-ai
29 May 2026
Safety

Crafting Desirable Climate Trajectories with RL Explored Socio-Environmental Simulations

DGX agent

arXiv:2410.07287v2 Announce Type: replace-cross Abstract: Climate change poses an existential threat, necessitating effective climate policies to enact impactful change. Decisions in this domain are i

safetyarxiv-cs-ai
29 May 2026
Safety

CRITIC-R1: Learning Structured Critics for Retrieval-Augmented Generation

DGX agent

arXiv:2605.29886v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) improves knowledge-intensive question answering by incorporating external evidence. However, existing RAG methods

safetyarxiv-cs-ai
29 May 2026
Agents

Croissant Tasks: A Metadata Format for Reproducible Machine Learning Evaluations

DGX agent

arXiv:2605.29786v1 Announce Type: new Abstract: Reproducibility is fundamental to the scientific method, yet remains a critical challenge in machine learning. Contributing factors include underspecifi

agentsarxiv-cs-ai
29 May 2026
Model Releases

CrystalXRD-Bench: Benchmarking Vision-Language Models for XRD Peak Indexing Across Diverse Crystalline Materials

DGX agent

arXiv:2605.29446v1 Announce Type: new Abstract: Miller-index identification from powder XRD patterns requires capabilities untested by existing multimodal benchmarks: the model must read a narrow peak

model-releasesarxiv-cs-ai
29 May 2026
Safety

DAMEL: Dual-Axis Multi-Expert Learning for Class-Imbalanced Learning

DGX agent

arXiv:2605.30135v1 Announce Type: cross Abstract: Various algorithms have been proposed to address the challenges posed by class-imbalanced learning from real-world data with long-tailed distributions

safetyarxiv-cs-ai
29 May 2026
Research

Data filtering methods for training language models

DGX agent

arXiv:2605.29807v1 Announce Type: cross Abstract: Data quality is a critical factor in the effectiveness of machine learning models. Label errors, present even in widely used benchmarks, introduce noi

researcharxiv-cs-ai
29 May 2026
Safety

DeepSurvey: Enhancing Analytical Depth and Citation Reliability in Automated Survey Generation

DGX agent

arXiv:2605.29522v1 Announce Type: new Abstract: As scientific literature grows rapidly, automated survey generation has become a key capability for AI scientists and human researchers. However, existi

safetyarxiv-cs-ai
29 May 2026
Research

DeepTool: Scaling Interleaved Deliberation in Tool-Integrated Reasoning via Process-Supervised Reinforcement Learning

DGX agent

arXiv:2605.29568v1 Announce Type: new Abstract: Tool-Integrated Reasoning (TIR) extends LLM capabilities by leveraging external environments. However, existing methods lack the deliberation during seq

researcharxiv-cs-ai
29 May 2026
Hardware

DELOS: Detecting Shallow Transits in Kepler Photometry Using a Contrastive-Learning Framework

DGX agent

arXiv:2605.29428v1 Announce Type: cross Abstract: We present DEtection in phase-folded Light curves with cOntrastive Scoring (DELOS), a contrastive-learning-based framework designed to search for shal

hardwarearxiv-cs-ai
29 May 2026
Local Ai

Demystifying Data Organization for Enhanced LLM Training

DGX agent

arXiv:2605.30334v1 Announce Type: new Abstract: Large Language Models (LLMs) have revolutionized various fields, yet their training efficiency is heavily reliant on effective data curation. While data

local-aiarxiv-cs-ai
29 May 2026
Model Releases

DenseSteer: Steering Small Language Models towards Dense Math Reasoning

DGX agent

arXiv:2605.29247v1 Announce Type: new Abstract: Large language models (LLMs) demonstrate strong chain-of-thought (CoT) reasoning abilities, while smaller models (<= 3B parameters) significantly underp

model-releasesarxiv-cs-ai
29 May 2026
Research

Diagnosing Harmful Continuation in Answer-Correct Long-CoT Training Traces

DGX agent

arXiv:2605.29288v1 Announce Type: new Abstract: Long chain-of-thought (CoT) traces are widely used as supervision for reasoning-oriented LLM SFT, yet answer-correct traces can still lead to markedly d

researcharxiv-cs-ai
29 May 2026
Model Releases

Differentiable Belief-based Opponent Shaping

DGX agent

arXiv:2605.29042v1 Announce Type: new Abstract: Human coordination often relies on the ability to influence the beliefs of others through strategic action. In multi-agent reinforcement learning, oppon

model-releasesarxiv-cs-ai
29 May 2026
Safety

Discovering Cooperative Pipelines: Autoresearch for Sequential Social Dilemmas

DGX agent

arXiv:2605.30003v1 Announce Type: cross Abstract: We study two-level autoresearch for cooperation: an outer-loop AI agent autonomously redesigns the inner-loop pipeline of an LLM policy-synthesis syst

safetyarxiv-cs-ai
29 May 2026
Agents

Dissociative Identity: Language Model Agents Lack Grounding for Reputation Mechanisms

DGX agent

arXiv:2605.30169v1 Announce Type: cross Abstract: As autonomous language model agents proliferate, forming an emerging agentic web with real-world consequences, what credibility signals can you use to

agentsarxiv-cs-ai
29 May 2026
Safety

DLM-SWAI: Steering Diffusion Language Models Before They Unmask

DGX agent

arXiv:2605.29626v1 Announce Type: cross Abstract: Steering language model generation toward desired textual properties is essential for practical deployment, and inference-time methods are particularl

safetyarxiv-cs-ai
29 May 2026
Research

Do Language Models Track Entities Across State Changes?

DGX agent

arXiv:2605.30233v1 Announce Type: cross Abstract: Entity tracking (ET), the ability to keep track of states, is a fundamental skill that underlies complex reasoning. An increasing amount of work inves

researcharxiv-cs-ai
29 May 2026
Model Releases

Do Physics Foundation Models Learn Generalizable Physics? A Bias-Aware Benchmark Across Physical Regimes and Distribution Shifts

DGX agent

arXiv:2605.29283v1 Announce Type: cross Abstract: Recent physics foundation models claim general spatiotemporal forecasting ability, yet their evaluations often collapse performance into a single aver

model-releasesarxiv-cs-ai
29 May 2026
Local Ai

Do Proactive Agents Really Need an LLM to Decide When to Wake and What to Anchor?

DGX agent

arXiv:2605.30152v1 Announce Type: cross Abstract: Proactive agents read user activity as text and call an LLM on every event to decide whether to act. But user activity is not natively text: it is a s

local-aiarxiv-cs-ai
29 May 2026
Research

Does Distributed Training Undermine Compute Governance?

DGX agent

arXiv:2605.29359v1 Announce Type: cross Abstract: Compute governance proposals often rely on the assumption that frontier AI training requires large, detectable computing clusters. However, recent adv

researcharxiv-cs-ai
29 May 2026
Agents

Does The Way You Plan Matter? An Empirical Study of Planning Representations for LLM Web Agents

DGX agent

arXiv:2605.29927v1 Announce Type: cross Abstract: Despite recent advances, LLM-based web agents still struggle with limited exploration, omission of critical steps, and sensitivity to task constraints

agentsarxiv-cs-ai
29 May 2026
Research

Domain-Informed Representation for Evolutionary Sieving in Integral and Module Lattices

DGX agent

arXiv:2605.29169v1 Announce Type: cross Abstract: Traditional cryptography, rooted in problems, e.g., integer factorisation or discrete log, is inevitably vulnerable to a fully operational quantum com

researcharxiv-cs-ai
29 May 2026
Tutorials

Domain-Specific Data Synthesis for LLMs via Minimal Sufficient Representation Learning

DGX agent

arXiv:2605.30039v1 Announce Type: new Abstract: Large Language Models have demonstrated remarkable progress in general-purpose capabilities and can achieve strong performance in specific domains throu

tutorialsarxiv-cs-ai
29 May 2026
Applications

Double-Edged Sword or Sharp Tool? Designing and Evaluating Triadic LLM-Teacher Collaboration for K-12 Writing at Scale

DGX agent

arXiv:2605.30200v1 Announce Type: new Abstract: The double-edged sword of integrating Large Language Models (LLMs) requires an effective triadic collaboration mechanism among LLMs, teachers and studen

applicationsarxiv-cs-ai
29 May 2026
Safety

Dynamics Within Latent Chain-of-Thought: An Empirical Study of Causal Structure

DGX agent

arXiv:2602.08783v3 Announce Type: replace Abstract: Latent or continuous chain-of-thought methods replace explicit textual rationales with a number of internal latent steps, but these intermediate com

safetyarxiv-cs-ai
29 May 2026
Model Releases

DynSess: Dynamic Session-Level Evaluation and Optimization Framework for Role-Playing Agents

DGX agent

arXiv:2605.29256v1 Announce Type: cross Abstract: Role-playing with large language models is fundamentally a session-level task, requiring agents to sustain character identity and interaction quality

model-releasesarxiv-cs-ai
29 May 2026
Agents

E-valuator: Reliable Agent Verifiers with Sequential Hypothesis Testing

DGX agent

arXiv:2512.03109v2 Announce Type: replace-cross Abstract: Agentic AI systems execute a sequence of actions, such as reasoning steps or tool calls, in response to a user prompt. To evaluate the success

agentsarxiv-cs-ai
29 May 2026
Safety

EAPO: Enhancing Policy Optimization with On-Demand Expert Assistance

DGX agent

arXiv:2509.23730v2 Announce Type: replace Abstract: Large language models (LLMs) have recently advanced in reasoning when optimized with reinforcement learning (RL) under verifiable rewards. Existing

safetyarxiv-cs-ai
29 May 2026
Safety

Emergent Semantic Representations in World Models through Physical Interaction without Linguistic Supervision

DGX agent

arXiv:2605.28865v1 Announce Type: cross Abstract: What does a world model learn from physical exploration, without any linguistic supervision? We argue the answer is organized by a single principle: t

safetyarxiv-cs-ai
29 May 2026
Model Releases

Empathic Prompting: Non-Verbal Context Integration for Multimodal LLM Conversations

DGX agent

arXiv:2510.20743v2 Announce Type: replace-cross Abstract: We present Empathic Prompting, a novel framework for multimodal human-AI interaction that enriches Large Language Model (LLM) conversations wi

model-releasesarxiv-cs-ai
29 May 2026
Applications

Energy-Aware NECO for Single-Pass Pixel-wise Out-of-Distribution Detection in Semantic Segmentation

DGX agent

arXiv:2605.29773v1 Announce Type: cross Abstract: Reliable semantic segmentation for mobile robots requires both accurate dense prediction and robust uncertainty estimation under distribution shift. S

applicationsarxiv-cs-ai
29 May 2026
Agents

Enhancing Multi-Agent Communication through Attention Steering with Context Relevance

DGX agent

arXiv:2605.30136v1 Announce Type: new Abstract: LLM-based multi-agent systems have demonstrated remarkable performance on complex tasks through collaborative reasoning. However, these systems tend to

agentsarxiv-cs-ai
29 May 2026
← Previous
1…241242243244245…452
Next →