AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Model Releases

vAttention: Verified Sparse Attention

DGX agent

arXiv:2510.05688v2 Announce Type: replace-cross Abstract: State-of-the-art sparse attention methods for reducing decoding latency fall into two main categories: approximate top-k (and its extension, t

model-releasesarxiv-cs-ai
26 May 2026
Applications

VectorArk: Learning Practical Image Vectorization with Rounded Polygon Representation

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2605.24398v1 Announce Type: cross Abstract: Recent vision-language model (VLM)-based approaches have achieved impressive results on image vectorization tasks. However, they are typically evaluat

applicationsarxiv-cs-ai
26 May 2026
Safety

VEN-VL: A Visual Ensemble MoE Framework for Effective and Efficient Multi-Modal Understanding

DGX agent

arXiv:2605.25952v1 Announce Type: cross Abstract: Despite the remarkable progress achieved by recent efficient methods in accelerating multimodal understanding, they still suffer from noticeable perfo

safetyarxiv-cs-ai
26 May 2026
Research

Verified SHAP: Provable Bounds for Exact Shapley Values of Neural Networks

DGX agent

arXiv:2605.24084v1 Announce Type: cross Abstract: Shapley additive explanations (SHAP) are widely recognised as computationally intractable for neural networks, since they induce an exponential search

researcharxiv-cs-ai
26 May 2026
Model Releases

VeriTrace: Evolving Mental Models for Deep Research Agents

DGX agent

arXiv:2605.26081v1 Announce Type: new Abstract: Deep research agents face vast, interdependent, and pervasively uncertain information. Existing systems explore what evolving intermediate representatio

model-releasesarxiv-cs-ai
26 May 2026
Agents

VineLM: Trie-Based Fine-Grained Control for Agentic Workflows

DGX agent

arXiv:2605.23914v1 Announce Type: cross Abstract: Agentic workflows interleave configurable LLM stages with tool stages and often include retries or refinement loops. Existing workflow managers profil

agentsarxiv-cs-ai
26 May 2026
Model Releases

VisualOverload: Probing Visual Understanding of VLMs in Really Dense Scenes

DGX agent

arXiv:2509.25339v3 Announce Type: replace-cross Abstract: Is basic visual understanding really solved in state-of-the-art VLMs? We present VisualOverload, a slightly different visual question answerin

model-releasesarxiv-cs-ai
26 May 2026
Research

Voting with the Graph: Stable RLAIF via Topological Consistency Maximization

DGX agent

arXiv:2510.15514v3 Announce Type: replace Abstract: Reinforcement Learning from AI Feedback (RLAIF) relies on LLM judges as preference measurement instruments, yet these instruments are fundamentally

researcharxiv-cs-ai
26 May 2026
Tutorials

Weakly Supervised Camouflaged Object Detection Based on the SAM Model and Mask Guidance

DGX agent

arXiv:2605.25385v1 Announce Type: cross Abstract: Camouflaged object detection (COD) from a single image is a challenging task due to the high similarity between objects and their surroundings. Existi

tutorialsarxiv-cs-ai
26 May 2026
Safety

What Gets Cited: Competitive GEO in AI Answer Engines

DGX agent

arXiv:2605.25517v1 Announce Type: new Abstract: AI answer engines generate answers from retrieved pages but cite only a few sources. This makes visibility depend not just on ranking, but on being cite

safetyarxiv-cs-ai
26 May 2026
Applications

What Happens Next? Anticipating Future Motion by Generating Point Trajectories

DGX agent

arXiv:2509.21592v2 Announce Type: replace-cross Abstract: We consider the problem of forecasting motion from a single image, i.e., predicting how objects in the world are likely to move, without the a

applicationsarxiv-cs-ai
26 May 2026
Model Releases

When Can We Trust Early Warnings? Leakage-Excluded Early Outcome Prediction from LMS Interaction Logs

DGX agent

arXiv:2605.25794v1 Announce Type: new Abstract: Early-warning models built from Learning Management System (LMS) logs aim to predict end-of-course outcomes early enough to enable timely learner suppor

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

When Correct Beliefs Collapse: Epistemic Resilience of LLMs under Clinical Pressure

DGX agent

arXiv:2605.23932v1 Announce Type: new Abstract: Despite strong medical benchmark accuracy, LLMs can exhibit severe multi-turn sycophancy in clinical dialogue, abandoning initial correct diagnosis unde

model-releasesarxiv-cs-ai
26 May 2026
Safety

When Does Multi-Agent RL Improve LLM Workflows? Workflow, Scale, and Policy-Sharing Tradeoffs

DGX agent

arXiv:2605.24202v1 Announce Type: new Abstract: Multi-agent LLM workflows route inference through specialized roles to lift end-task accuracy, but jointly training those roles with reinforcement learn

safetyarxiv-cs-ai
26 May 2026
Research

When Does Synthetic Patent Data Help? Volume-Fidelity Trade-offs in Low-Resource Multi-Label Classification

DGX agent

arXiv:2605.24296v1 Announce Type: new Abstract: We study when LLM-generated synthetic data helps low-resource multi-label patent classification, separating true synthetic value from the confound that

researcharxiv-cs-ai
26 May 2026
Research

When Gradients Collide: Failure Modes of Multi-Objective Prompt Optimization for LLM Judges

DGX agent

arXiv:2605.26046v1 Announce Type: cross Abstract: Customizing an LLM judge to a specific task or domain often involves optimizing its prompt across multiple evaluation criteria simultaneously. Textual

researcharxiv-cs-ai
26 May 2026
Model Releases

When Mean CE Fails: Median CE Can Better Track Language Model Quality

DGX agent

arXiv:2605.24667v1 Announce Type: new Abstract: Mean cross-entropy is the standard validation metric for language models, but it can fail to track model quality during training. We examine this in two

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

When Reasoning Hurts: Source-Aware Evaluation of Frontier LLMs for Clinical SOAP Note Generation

DGX agent

arXiv:2605.24902v1 Announce Type: cross Abstract: Reasoning-enabled LLMs perform strongly on medical reasoning benchmarks, but it remains unclear whether these gains transfer to structured clinical do

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

When Search Becomes Memory: Turning Robot Design Trials into Transferable Skills

DGX agent

arXiv:2605.25832v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as proposal generators for evolutionary robot design, yet most loops remain memoryless: simulator r

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

When the Manual Lies: A Realistic Benchmark to Evaluate MCP Poisoning Attacks for LLM Agents

DGX agent

arXiv:2605.24069v1 Announce Type: cross Abstract: The rise of tool-using Large Language Model (LLM) agents, standardized by protocols like the Model Context Protocol (MCP), has unlocked unprecedented

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Who judges the judges? Governance from metrics: a runtime framework for continuous LLM compliance monitoring

DGX agent

arXiv:2605.24737v1 Announce Type: cross Abstract: Current approaches to AI compliance treat conformity as a binary, audit-time verdict rather than a continuous, measurable property of production syste

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Whose Alignment? Comparing LLM Process Alignment Across Diverse Organizational Decision Contexts

DGX agent

arXiv:2605.25256v1 Announce Type: new Abstract: Aligning AI systems with organizational decision-making is typically framed as a single-target problem: make the model behave like the organization. We

model-releasesarxiv-cs-ai
26 May 2026
Agents

Why We Need World Models for AGI: Where LLMs Fail and How World Models May Outperform

DGX agent

arXiv:2605.23972v1 Announce Type: new Abstract: Large language models achieve strong performance in language generation and knowledge-intensive tasks, yet remain limited in settings requiring causal r

agentsarxiv-cs-ai
26 May 2026
Agents

Why Your Deep Research Agent Fails? On Hallucination Evaluation in Full Research Trajectory

DGX agent

arXiv:2601.22984v2 Announce Type: replace Abstract: Diagnosing failure patterns in Deep Research Agents (DRAs) remains a critical challenge. Existing benchmarks predominantly rely on end-to-end evalua

agentsarxiv-cs-ai
26 May 2026
Model Releases

World-State Transformations for Neuro-symbolic Interactive Storytelling

DGX agent

arXiv:2605.24719v1 Announce Type: cross Abstract: Large Language Models (LLMs) have changed the possibilities of Interactive Storytelling systems that process free-text user input. However, as more of

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

WorldGUI: An Interactive Benchmark for Desktop GUI Automation from Any Starting Point

DGX agent

arXiv:2502.08047v5 Announce Type: replace Abstract: Recent progress in GUI agents has substantially improved visual grounding, yet robust planning remains challenging, particularly when the environmen

model-releasesarxiv-cs-ai
26 May 2026
Research

WTKO-CNN: Deep Learning Reveals Sequence Motifs Distinguishing Wild-Type and Knockout ATAC-seq Peaks

DGX agent

arXiv:2605.24034v1 Announce Type: cross Abstract: Chromatin regulators can alter transcriptional programs by modifying the accessibility of regulatory DNA elements. Understanding how regulatory sequen

researcharxiv-cs-ai
26 May 2026
Research

You Can Ground Earlier than See: An Effective and Efficient Pipeline for Temporal Sentence Grounding in Compressed Videos

DGX agent

arXiv:2303.07863v3 Announce Type: replace-cross Abstract: Given an untrimmed video, temporal sentence grounding (TSG) aims to locate a target moment semantically according to a sentence query. Althoug

researcharxiv-cs-ai
26 May 2026
Local Ai

Your Embedding Model is SMARTer Than You Think

DGX agent

arXiv:2605.24938v1 Announce Type: cross Abstract: Multimodal retrieval relies heavily on single-vector retrievers, which compress rich, sequential token sequences into one single global representation

local-aiarxiv-cs-ai
26 May 2026
Research

Zero-Shot Parkinson's Disease Detection from Speech: Comparing Large Audio and Language Models

DGX agent

arXiv:2605.24806v1 Announce Type: cross Abstract: Large audio and language models have recently demonstrated zero-shot reasoning capabilities across various domains. However, it remains unclear how th

researcharxiv-cs-ai
26 May 2026
Agents

6G Communication Networks Enabling Embodied Agents: Architecture and Prototype

DGX agent

arXiv:2605.23263v1 Announce Type: cross Abstract: Embodied agents, which couple intelligent decision-making with physical actuation in the real world, impose far more stringent and heterogeneous commu

agentsarxiv-cs-ai
25 May 2026
Research

A drone-based framework for coral habitat mapping via weakly supervised segmentation

DGX agent

arXiv:2508.18958v2 Announce Type: replace-cross Abstract: Obtaining pixel-level annotations over large spatial extents remains a major bottleneck for deploying machine learning in ecological applicati

researcharxiv-cs-ai
25 May 2026
Research

A Fine-Tuned BERT Classifier for Personal-Letter Titles in Late-Ming and Early-Qing Collected Works

DGX agent

arXiv:2605.23103v1 Announce Type: cross Abstract: I present Lepton (Letter Prediction), a fine-tuned BERT classifier that predicts whether a title in a Classical Chinese wenji table of contents is a p

researcharxiv-cs-ai
25 May 2026
Tutorials

A mathematical theory of balancing relational generalization and memorization

DGX agent

arXiv:2605.22972v1 Announce Type: cross Abstract: Humans, animals, and modern machine learning models exhibit impressive abilities to learn complex behaviors and generalize these behaviors to unseen s

tutorialsarxiv-cs-ai
25 May 2026
Model Releases

A measurement substrate for agentic Kubernetes operations: Methodology and a case study in retrieval-compounding falsification

DGX agent

arXiv:2605.23058v1 Announce Type: cross Abstract: Empirical claims about autonomous Kubernetes operations agents are largely unfalsifiable. Published work reports observational results without control

model-releasesarxiv-cs-ai
25 May 2026
Agents

A Proactive Multi-Agent Dialogue Framework for Assessing Social Language Disorder Traits in Autism

DGX agent

arXiv:2605.22993v1 Announce Type: cross Abstract: Characteristic linguistic behaviors associated with Social Language Disorder (SLD) in autism spectrum disorder, including echoic repetition, pronoun d

agentsarxiv-cs-ai
25 May 2026
Model Releases

A Systematic Evaluation of Co-folding Model Representations for Small-Molecule Learning

DGX agent

arXiv:2602.13249v2 Announce Type: replace-cross Abstract: Small-molecule foundation models are typically pretrained on standalone molecular data, unlike vision and language models that often benefit f

model-releasesarxiv-cs-ai
25 May 2026
Research

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning

DGX agent

arXiv:2605.23200v1 Announce Type: cross Abstract: The linear growth of the Key-Value (KV) cache is a critical bottleneck in long-form LLM inference. Existing KV compression methods mitigate this by ev

researcharxiv-cs-ai
25 May 2026
Research

Adversarial Vulnerability Under Temporal Concept Drift: A Longitudinal Study of Android Malware Detection

DGX agent

arXiv:2605.23623v1 Announce Type: cross Abstract: We present a longitudinal, drift-aware evaluation of adversarial robustness across more than a decade of Android applications using static and dynamic

researcharxiv-cs-ai
25 May 2026
Model Releases

Agentic Proving for Program Verification

DGX agent

arXiv:2605.23772v1 Announce Type: new Abstract: Agentic systems have recently emerged as state-of-the-art approaches for automated theorem proving in formal mathematics. To assess how far these capabi

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

Agentic-VLA: Efficient Online Adaptation for Vision-Language-Action Models

DGX agent

arXiv:2605.22896v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have emerged as a promising paradigm for robotic manipulation by leveraging pre-trained vision-language representa

model-releasesarxiv-cs-ai
25 May 2026
Agents

Agentivism: a learning theory for the age of artificial intelligence

DGX agent

arXiv:2604.07813v2 Announce Type: replace Abstract: Learning theories have historically changed when the conditions of learning evolved. Generative and agentic AI create a new condition by allowing le

agentsarxiv-cs-ai
25 May 2026
Agents

AI Assurance: A Comprehensive Testing Strategy for Enterprise AI Systems

DGX agent

arXiv:2605.23459v1 Announce Type: cross Abstract: Enterprise AI systems, built on large language models, retrieval pipelines and autonomous agents, introduce a class of risks that traditional software

agentsarxiv-cs-ai
25 May 2026
Model Releases

AI Evaluation Should Require Standardized Item-Level Data Releases

DGX agent

arXiv:2604.03244v2 Announce Type: replace Abstract: This position paper argues that standardized item-level benchmark data should become the default infrastructure for AI evaluation. Current evaluatio

model-releasesarxiv-cs-ai
25 May 2026
Research

AI Security Research Should Better Incentivize Defense Research

DGX agent

arXiv:2605.23448v1 Announce Type: cross Abstract: This work examines an imbalance in artificial intelligence (AI) security research: the field tends to produce more work on attacking AI systems than o

researcharxiv-cs-ai
25 May 2026
Safety

ALIVE: Awakening LLM Reasoning via Adversarial Learning and Instructive Verbal Evaluation

DGX agent

arXiv:2602.05472v2 Announce Type: replace Abstract: The quest for expert-level reasoning in Large Language Models (LLMs) has been hampered by a persistent extit{reward bottleneck}: traditional reinfor

safetyarxiv-cs-ai
25 May 2026
Research

An AI-Driven Framework for Energy-Efficient Environmental Monitoring in Smart Cities Using Edge Intelligence

DGX agent

arXiv:2605.22824v1 Announce Type: cross Abstract: Environmental monitoring is a crucial component of the smart city infrastructure. It enables informed decision making which enhances sustainability, p

researcharxiv-cs-ai
25 May 2026
Local Ai

Anatomy-Guided Vision-Language Learning with Angular Prototype Separation for Multi-Label Video Capsule Endoscopy Classification Under Class Imbalance

DGX agent

arXiv:2603.17879v2 Announce Type: replace-cross Abstract: This work presents a multi-label temporal event detection framework for video capsule endoscopy (VCE) that addresses the extreme class imbalan

local-aiarxiv-cs-ai
25 May 2026
← Previous
1…275276277278279…448
Next →