AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Applications

The Deterministic Horizon: Impossibility Results as Design Specifications for Trustworthy AI Systems

DGX agent

arXiv:2605.23024v1 Announce Type: new Abstract: Large language models now write software, draft legal documents, and produce clinical notes, yet fundamental limits, from Turing and Arrow to the No Fre

applicationsarxiv-cs-ai
25 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

The Misattribution Gap: When Memory Poisoning Looks Like Model Failure in Agentic AI Systems

DGX agent

arXiv:2605.22842v1 Announce Type: cross Abstract: Multi-agent AI pipelines typically assume that agent misconduct originates from model misalignment. We identify a structural failure in this assumptio

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

The Readout Shortcut: Positional Number Copying Dominates Arithmetic CoT Readout in Small Language Models

DGX agent

arXiv:2605.22870v1 Announce Type: cross Abstract: Chain-of-thought (CoT) prompting is necessary for arithmetic in small language models, yet shuffling its steps preserves most performance. What does C

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

The Surprising Difficulty of Search in Model-Based Reinforcement Learning

DGX agent

arXiv:2601.21306v2 Announce Type: replace-cross Abstract: This paper investigates search in model-based reinforcement learning (RL). Conventional wisdom holds that long-term predictions and compoundin

model-releasesarxiv-cs-ai
25 May 2026
Tutorials

The TIME Machine: On The Power of Motion for Efficient Perception

DGX agent

arXiv:2605.23045v1 Announce Type: cross Abstract: Video representation learning has seen tremendous progress in recent years. This has been driven by many factors, including the scale of training and

tutorialsarxiv-cs-ai
25 May 2026
Safety

Through the Stealth Lens: Attention-Aware Defenses Against Poisoning in RAG

DGX agent

arXiv:2506.04390v2 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) systems are vulnerable to attacks that inject poisoned passages into the retrieved context, even at low c

safetyarxiv-cs-ai
25 May 2026
Model Releases

Transcoders Trace Visual Grounding and Hallucinations in Vision-Language Models

DGX agent

arXiv:2605.22902v1 Announce Type: cross Abstract: Generative Vision-Language Models (VLMs) perform well on multimodal reasoning, but how visual inputs are transformed to text remains poorly understood

model-releasesarxiv-cs-ai
25 May 2026
Tutorials

Uncovering the Latent Potential of Deep Intermediate Representations

DGX agent

arXiv:2605.23033v1 Announce Type: cross Abstract: Foundational Models pretrained on huge amount of data learn representations that evolve across depth, forming a hierarchy of embeddings with distinct

tutorialsarxiv-cs-ai
25 May 2026
Model Releases

Understanding and Improving Noisy Embedding Techniques in Instruction Finetuning

DGX agent

arXiv:2605.23171v1 Announce Type: cross Abstract: Recent advancements in instructional fine-tuning have injected noise into embeddings, with NEFTune (Jain et al., 2024) setting benchmarks using unifor

model-releasesarxiv-cs-ai
25 May 2026
Safety

Understanding Goal Generalisation in Sequential Reinforcement Learning

DGX agent

arXiv:2605.23565v1 Announce Type: cross Abstract: Reinforcement learning agents often exhibit unintended goal-directed behaviour outside their training distribution, but we currently lack a principled

safetyarxiv-cs-ai
25 May 2026
Research

Understanding Task Aggregation for Generalizable Ultrasound Foundation Models

DGX agent

arXiv:2603.18123v3 Announce Type: replace-cross Abstract: Foundation models promise to unify multiple clinical tasks within a single framework, but recent ultrasound studies report that unified models

researcharxiv-cs-ai
25 May 2026
Safety

V-VLAPS: Value-Guided Planning for Vision-Language-Action Models

DGX agent

arXiv:2601.00969v2 Announce Type: replace-cross Abstract: Vision-language-action (VLA) models provide strong action priors for robotic manipulation, but their reactive behavior can fail under distribu

safetyarxiv-cs-ai
25 May 2026
Applications

VACE: Learning Geometrically Structured Representations for Time Series Anomaly Detection

DGX agent

arXiv:2605.23504v1 Announce Type: cross Abstract: Anomaly detection in multivariate time series is a critical task across a wide range of real-world applications, where abnormal behaviour is rare, lab

applicationsarxiv-cs-ai
25 May 2026
Research

VGAS: Value-Guided Action-Chunk Selection for Few-Shot Vision-Language-Action Adaptation

DGX agent

arXiv:2602.07399v2 Announce Type: replace Abstract: Vision--Language--Action (VLA) models bridge multimodal reasoning with physical control, but adapting them to new tasks with scarce demonstrations r

researcharxiv-cs-ai
25 May 2026
Safety

VI-CuRL: Stabilizing Verifier-Independent RL Reasoning via Confidence-Guided Variance Reduction

DGX agent

arXiv:2602.12579v2 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has emerged as a dominant paradigm for enhancing Large Language Models (LLMs) reasoning,

safetyarxiv-cs-ai
25 May 2026
Model Releases

VideoTemp-o3: Harmonizing Temporal Grounding and Video Understanding in Agentic Thinking-with-Videos

DGX agent

arXiv:2602.07801v4 Announce Type: replace-cross Abstract: In long-video understanding, conventional uniform frame sampling often fails to capture key visual evidence, leading to degraded performance a

model-releasesarxiv-cs-ai
25 May 2026
Research

Weierstrass Positional Encoding for Vision Transformers

DGX agent

arXiv:2605.23719v1 Announce Type: cross Abstract: Vision Transformers have achieved remarkable success in computer vision, but their common use of learnable one-dimensional positional encodings weaken

researcharxiv-cs-ai
25 May 2026
Research

When Do LLMs Reason? A Dynamical Systems View via Entropy Phase Transitions

DGX agent

arXiv:2605.22873v1 Announce Type: cross Abstract: Chain-of-thought (CoT) reasoning has become the default strategy for enhancing LLM capabilities, yet its application raises a fundamental question: wh

researcharxiv-cs-ai
25 May 2026
Model Releases

When Good Equations Get Bad Scores: Improving Symbolic Regression Through Better Parameter Optimization

DGX agent

arXiv:2605.23272v1 Announce Type: cross Abstract: Symbolic Regression (SR) plays a central role in scientific knowledge discovery by distilling mathematical equations from observational data. Most exi

model-releasesarxiv-cs-ai
25 May 2026
Agents

When Planning Fails Despite Correct Execution: On Epistemic Calibration for LLM-Based Multi-Agent Systems

DGX agent

arXiv:2605.23414v1 Announce Type: new Abstract: LLM-based multi-agent systems can fail even when planned actions are executed correctly because agents may misjudge their knowledge when evaluating plan

agentsarxiv-cs-ai
25 May 2026
Safety

Whose Good, Whose Place? The Moral Geography of Agentic AI for Social Good

DGX agent

arXiv:2605.22995v1 Announce Type: cross Abstract: Agentic AI systems are increasingly proposed for social-good domains, often invoking the United Nations Sustainable Development Goals (SDGs) as a voca

safetyarxiv-cs-ai
25 May 2026
Research

Worse than Random: The Importance of a Baseline for Unsupervised Feature Selection

DGX agent

arXiv:2605.22973v1 Announce Type: cross Abstract: Many novel unsupervised feature selection methods are proposed each year, yet their empirical evaluation is limited to supervised and unsupervised eva

researcharxiv-cs-ai
25 May 2026
Model Releases

XAttnMark: Learning Robust Audio Watermarking with Cross-Attention

DGX agent

arXiv:2502.04230v3 Announce Type: replace-cross Abstract: The rapid proliferation of generative audio synthesis and editing technologies has raised serious concerns about copyright infringement, data

model-releasesarxiv-cs-ai
25 May 2026
Hardware

XWind: A Cross-site Router for Large Language Model Inference Serving at Renewable Energy Farms

DGX agent

arXiv:2605.23348v1 Announce Type: cross Abstract: AI power demand is growing at an unprecedented rate while power grids are often ailing and struggle to keep up. Grid expansion comes with high capital

hardwarearxiv-cs-ai
25 May 2026
Local Ai

ZipMoE: Efficient On-Device MoE Serving via Lossless Compression and Cache-Affinity Scheduling

DGX agent

arXiv:2601.21198v2 Announce Type: replace-cross Abstract: While Mixture-of-Experts (MoE) architectures substantially bolster the expressive power of large-language models, their prohibitive memory foo

local-aiarxiv-cs-ai
25 May 2026
Research

Access Paths for Efficient Ordering with Large Language Models

DGX agent

arXiv:2509.00303v3 Announce Type: replace-cross Abstract: In this work, we present the exttt{LLM ORDER BY} semantic operator as a logical abstraction and conduct a systematic study of its physical imp

researcharxiv-cs-ai
22 May 2026
Model Releases

AgentCo-op: Retrieval-Based Synthesis of Interoperable Multi-Agent Workflows

DGX agent

arXiv:2605.20425v1 Announce Type: new Abstract: Designing multi-agent workflows is especially difficult in open-ended scientific settings where tasks lack curated training sets, reliable scalar evalua

model-releasesarxiv-cs-ai
22 May 2026
Agents

Agentic Agile-V: From Vibe Coding to Verified Engineering in Software and Hardware Development

DGX agent

arXiv:2605.20456v1 Announce Type: cross Abstract: Agentic AI coding systems can inspect repositories, plan implementation steps, edit files, call tools, run tests, and submit pull requests. These capa

agentsarxiv-cs-ai
22 May 2026
Agents

An Application-Layer Multi-Modal Covert-Channel Reference Monitor for LLM Agent Egress

DGX agent

arXiv:2605.20734v1 Announce Type: cross Abstract: A large language model (LLM) agent that sends messages can leak data inside them. Destination allowlists and content scanners do not police whether an

agentsarxiv-cs-ai
22 May 2026
Agents

Artificial Intelligence Reshapes Microwave Photonics

DGX agent

arXiv:2605.21224v1 Announce Type: cross Abstract: As a rapidly emerging interdisciplinary field that intrinsically integrates microwave and photonics, microwave photonics (MWP) provides disruptive sol

agentsarxiv-cs-ai
22 May 2026
Agents

AutoRPA: Efficient GUI Automation through LLM-Driven Code Synthesis from Interactions

DGX agent

arXiv:2605.21082v1 Announce Type: new Abstract: Large Language Model (LLM) based agents have demonstrated proficiency in multi-step interactions with graphical user interfaces (GUIs). While most resea

agentsarxiv-cs-ai
22 May 2026
Agents

Causal Past Logic for Runtime Verification of Distributed LLM Agent Workflows

DGX agent

arXiv:2605.20923v1 Announce Type: cross Abstract: Distributed LLM agent workflows should not be monitored as if they produced a single sequential log. In an asynchronous execution, a decision can only

agentsarxiv-cs-ai
22 May 2026
Tutorials

Chain-of-thought obfuscation learned from output supervision can generalise to unseen tasks

DGX agent

arXiv:2601.23086v2 Announce Type: replace Abstract: Chain-of-thought (CoT) reasoning provides a significant performance uplift to LLMs by enabling planning, exploration, and deliberation of their acti

tutorialsarxiv-cs-ai
22 May 2026
Agents

COAgents: Multi-Agent Framework to Learn and Navigate Routing Problems Search Space

DGX agent

arXiv:2605.20618v1 Announce Type: new Abstract: Although Vehicle Routing Problems (VRP) are essential to many real-world systems, they remain computationally intractable at scale due to their combinat

agentsarxiv-cs-ai
22 May 2026
Model Releases

Code Researcher: Deep Research Agent for Large Systems Code and Commit History

DGX agent

arXiv:2506.11060v2 Announce Type: replace-cross Abstract: Large Language Model (LLM)-based coding agents have shown promising results on coding benchmarks, but their effectiveness on systems code rema

model-releasesarxiv-cs-ai
22 May 2026
Applications

Codec-Robust Attacks on Audio LLMs

DGX agent

arXiv:2605.20519v1 Announce Type: cross Abstract: Prior attacks on Audio Large Language Models (Audio LLMs) demonstrated that carefully crafted waveform-domain perturbations can force targeted adversa

applicationsarxiv-cs-ai
22 May 2026
Model Releases

CTFExplorer: Evaluating LLM Offensive Agents Through Multi-Target Web CTF Benchmarking

DGX agent

arXiv:2602.08023v3 Announce Type: replace-cross Abstract: Existing benchmarks for LLM-based offensive security agents use isolated, single-target setups with a known vulnerable service and fixed objec

model-releasesarxiv-cs-ai
22 May 2026
Model Releases

Declarative Data Services: Structured Agentic Discovery for Composing Data Systems

DGX agent

arXiv:2605.20690v1 Announce Type: new Abstract: Agentic discovery has shown that LLM-driven search can find novel algorithms, designs, and code under benchmark conditions. Translating the paradigm to

model-releasesarxiv-cs-ai
22 May 2026
Model Releases

DeepWeb-Bench: A Deep Research Benchmark Demanding Massive Cross-Source Evidence and Long-Horizon Derivation

DGX agent

arXiv:2605.21482v1 Announce Type: new Abstract: Deep research, in which an agent searches the open web, collects evidence, and derives an answer through extended reasoning, is a prominent use case for

model-releasesarxiv-cs-ai
22 May 2026
Model Releases

DeFacto: Counterfactual Thinking with Images for Enforcing Evidence-Grounded and Faithful Reasoning

DGX agent

arXiv:2509.20912v4 Announce Type: replace Abstract: Recent advances in multimodal language models (MLLMs) have made thinking with images a dominant paradigm for multimodal reasoning. However, existing

model-releasesarxiv-cs-ai
22 May 2026
Research

Designing Conversations with the Dead: How People Engage with Generative Ghosts

DGX agent

arXiv:2605.21390v1 Announce Type: cross Abstract: We examine how people experience two choices in the design of generative ghosts, AI systems that are trained on data of the dead: representation, wher

researcharxiv-cs-ai
22 May 2026
Research

Detecting Trojaned DNNs via Spectral Regression Analysis

DGX agent

arXiv:2605.21146v1 Announce Type: cross Abstract: Modern DNNs are repeatedly fine-tuned to incorporate new data and functionality. This evolutionary workflow introduces a security risk when updated da

researcharxiv-cs-ai
22 May 2026
Tutorials

Diverge to Induce Prompting: Multi-Rationale Induction for Zero-Shot Reasoning

DGX agent

arXiv:2602.08028v1 Announce Type: cross Abstract: To address the instability of unguided reasoning paths in standard Chain-of-Thought prompting, recent methods guide large language models (LLMs) by fi

tutorialsarxiv-cs-ai
22 May 2026
Applications

ELSA: An ELastic SNN Inference Architecture for Efficient Neuromorphic Computing

DGX agent

arXiv:2605.20802v1 Announce Type: cross Abstract: Spiking neural networks (SNNs) exploit event-driven and addition-only computation to substantially improve efficiency for intelligent computation. A k

applicationsarxiv-cs-ai
22 May 2026
Agents

Enabling Regulatory Multi-Agent Collaboration: Architecture, Challenges, and Solutions

DGX agent

arXiv:2509.09215v2 Announce Type: replace Abstract: Large language models (LLMs)-empowered autonomous agents are transforming both digital and physical environments by enabling adaptive, multi-agent c

agentsarxiv-cs-ai
22 May 2026
Agents

Evaluating multimodal emotion recognition in proactive conversational agents: A user study

DGX agent

arXiv:2605.20200v1 Announce Type: cross Abstract: This article presents a multimodal emotion recognition module integrated into a proactive Socially Interactive Agent (SIA) powered by generative artif

agentsarxiv-cs-ai
22 May 2026
Model Releases

Evaluating Temporal Semantic Caching and Workflow Optimization in Agentic Plan-Execute Pipelines

DGX agent

arXiv:2605.20630v1 Announce Type: new Abstract: Industrial asset operations workflows are latency-sensitive because a single user query may require coordination over sensor data, work orders, failure

model-releasesarxiv-cs-ai
22 May 2026
Agents

From Automated to Autonomous: Hierarchical Agent-native Network Architecture (HANA)

DGX agent

arXiv:2605.20608v1 Announce Type: new Abstract: Realizing Level 4/5 Autonomous Networks (AN) demands a shift from static automation to agent-native intelligence. Current operations, reliant on rigid s

agentsarxiv-cs-ai
22 May 2026
← Previous
1…280281282283284…448
Next →