AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Research

When Can Human-AI Teams Outperform Individuals? Tight Bounds with Impossibility Guarantees

DGX agent

arXiv:2605.08710v1 Announce Type: new Abstract: Human-AI teams fail to outperform their best member in 70% of studies, yet no theory specifies when complementarity is achievable. We derive tight bound

researcharxiv-cs-ai
12 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Agents

When Child Inherits: Modeling and Exploiting Subagent Spawn in Multi-Agent Networks

DGX agent

arXiv:2605.08460v1 Announce Type: cross Abstract: Since the official release of ChatGPT in 2022, large language models (LLMs) have rapidly evolved from chatbot-style interfaces into agentic systems th

agentsarxiv-cs-ai
12 May 2026
Model Releases

When Does Non-Uniform Replay Matter in Reinforcement Learning?

DGX agent

arXiv:2605.10236v1 Announce Type: cross Abstract: Modern off-policy reinforcement learning algorithms often rely on simple uniform replay sampling and it remains unclear when and why non-uniform repla

model-releasesarxiv-cs-ai
12 May 2026
Research

When Does Value-Aware KV Eviction Help? A Fixed-Contract Diagnostic for Non-Monotone Cache Compression

DGX agent

arXiv:2605.08234v1 Announce Type: cross Abstract: Long-context LLM inference is bottlenecked by the memory and bandwidth cost of reading large KV caches during decoding. KV compression reduces this co

researcharxiv-cs-ai
12 May 2026
Research

When Few Steps Are Enough: Training-Free Acceleration of Identity-Preserved Generation

DGX agent

arXiv:2605.09460v1 Announce Type: cross Abstract: Identity-preserved image generation is typically built on many-step diffusion backbones, making personalized generation expensive at deployment time.

researcharxiv-cs-ai
12 May 2026
Safety

When Language Overwrites Vision: Over-Alignment and Geometric Debiasing in Vision-Language Models

DGX agent

arXiv:2605.08245v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) increasingly power high-stakes applications, from medical imaging to autonomous systems, yet they routinely hallucinate,

safetyarxiv-cs-ai
12 May 2026
Tutorials

When Normality Shifts: Risk-Aware Test-Time Adaptation for Unsupervised Tabular Anomaly Detection

DGX agent

arXiv:2605.10242v1 Announce Type: cross Abstract: Unsupervised tabular anomaly detection methods typically learn feature patterns from normal samples during training and subsequently identify samples

tutorialsarxiv-cs-ai
12 May 2026
Model Releases

When Prompts Become Payloads: A Framework for Mitigating SQL Injection Attacks in Large Language Model-Driven Applications

DGX agent

arXiv:2605.10176v1 Announce Type: cross Abstract: Natural language interfaces to structured databases are becoming increasingly common, largely due to advances in large language models (LLMs) that ena

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

When Reviews Disagree: Fine-Grained Contradiction Analysis in Scientific Peer Reviews

DGX agent

arXiv:2605.10171v1 Announce Type: cross Abstract: Scientific peer reviews frequently contain conflicting expert judgments, and the increasing scale of conference submissions makes it challenging for A

model-releasesarxiv-cs-ai
12 May 2026
Research

When Tables Leak: Attacking String Memorization in LLM-Based Tabular Data Generation

DGX agent

arXiv:2512.08875v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have recently demonstrated remarkable performance in generating high-quality tabular synthetic data. In practice,

researcharxiv-cs-ai
12 May 2026
Model Releases

When to Re-Commit: Temporal Abstraction Discovery for Long-Horizon Vision-Language Reasoning

DGX agent

arXiv:2605.09860v1 Announce Type: new Abstract: Long-horizon reasoning requires deciding not only what actions to take, but how deeply to commit before the next observation. We formalize this as commi

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

When to Trust Imagination: Adaptive Action Execution for World Action Models

DGX agent

arXiv:2605.06222v2 Announce Type: replace-cross Abstract: World Action Models (WAMs) have recently emerged as a promising paradigm for robotic manipulation by jointly predicting future visual observat

model-releasesarxiv-cs-ai
12 May 2026
Safety

Where Do Flow Semantics Reside? A Protocol-Native Tabular Pretraining Paradigm for Encrypted Traffic Classification

DGX agent

arXiv:2603.10051v2 Announce Type: replace-cross Abstract: Self-supervised masked modeling shows promise for encrypted traffic classification by masking and reconstructing raw bytes. Yet recent work re

safetyarxiv-cs-ai
12 May 2026
Research

Where Do Reasoning Models Refuse?

DGX agent

arXiv:2507.03167v3 Announce Type: replace-cross Abstract: Chat models without chain-of-thought (CoT) reasoning must decide whether to refuse a harmful request before generating their first response to

researcharxiv-cs-ai
12 May 2026
Research

Where Reliability Lives in Vision-Language Models: A Mechanistic Study of Attention, Hidden States, and Causal Circuits

DGX agent

arXiv:2605.08200v1 Announce Type: new Abstract: A pervasive intuition holds that vision-language models (VLMs) are most trustworthy when their attention maps look sharp: concentrated attention on the

researcharxiv-cs-ai
12 May 2026
Safety

Why Adam Works Better with eta_1 = eta_2: The Missing Gradient Scale Invariance Principle

DGX agent

arXiv:2601.21739v2 Announce Type: replace-cross Abstract: Adam has been at the core of large-scale training for almost a decade, yet a simple empirical fact remains unaccounted for: both validation sc

safetyarxiv-cs-ai
12 May 2026
Local Ai

Why Do Aligned LLMs Remain Jailbreakable: Refusal-Escape Directions, Operator-Level Sources, and Safety-Utility Trade-off

DGX agent

arXiv:2605.08878v1 Announce Type: cross Abstract: Aligned large language models (LLMs) remain vulnerable to jailbreak attacks. Recent mechanistic studies have identified latent features and representa

local-aiarxiv-cs-ai
12 May 2026
Safety

Why Do DiT Editors Drift? Plug-and-Play Low Frequency Alignment in VAE Latent Space

DGX agent

arXiv:2605.08250v1 Announce Type: cross Abstract: Recent advances in diffusion transformers (DiTs) have enabled promising single-turn image editing capabilities. However, multi-turn editing often lead

safetyarxiv-cs-ai
12 May 2026
Research

Why Low-Resource NLP Needs More Than Cross-Lingual Transfer: Lessons Learned from Luxembourgish

DGX agent

arXiv:2605.10714v1 Announce Type: cross Abstract: Cross-lingual transfer has become a central paradigm for extending natural language processing (NLP) technologies to low-resource languages. By levera

researcharxiv-cs-ai
12 May 2026
Model Releases

Why Retrying Fails: Context Contamination in LLM Agent Pipelines

DGX agent

arXiv:2605.08563v1 Announce Type: new Abstract: When an LLM agent fails a multi-step tool-augmented task and retries, the failed attempt typically remains in its context window -- contaminating the ne

model-releasesarxiv-cs-ai
12 May 2026
Agents

Willful Disobedience: Automatically Detecting Failures in Agentic Traces

DGX agent

arXiv:2603.23806v2 Announce Type: replace-cross Abstract: AI agents are increasingly embedded in real software systems, where they execute multi-step workflows through multi-turn dialogue, tool invoca

agentsarxiv-cs-ai
12 May 2026
Model Releases

WindINR: Latent-State INR for Fast Local Wind Query and Correction in Complex Terrain

DGX agent

arXiv:2605.09511v1 Announce Type: new Abstract: Many downstream decisions in complex terrain require fast wind estimates at a small number of user-specified locations and heights for a given forecast

model-releasesarxiv-cs-ai
12 May 2026
Safety

WISTERIA: Learning Clinical Representations from Noisy Supervision via Multi-View Consistency in Electronic Health Records

DGX agent

arXiv:2605.09765v1 Announce Type: cross Abstract: Representation learning in electronic health records (EHR) has largely followed paradigms inherited from natural language processing, relying on seque

safetyarxiv-cs-ai
12 May 2026
Agents

Workspace Optimization: How to Train Your Agent

DGX agent

arXiv:2605.09650v1 Announce Type: new Abstract: Modern agents built on frontier language models often cannot adapt their weights. What, then, remains trainable? We argue it is the agent's workspace, t

agentsarxiv-cs-ai
12 May 2026
Research

WorldSpeech: A Multilingual Speech Corpus from Around the World

DGX agent

arXiv:2605.09167v1 Announce Type: cross Abstract: Automatic speech recognition (ASR) performs well for high-resource languages with abundant paired audio-transcript data, but its accuracy degrades sha

researcharxiv-cs-ai
12 May 2026
Safety

X-Voice: Enabling Everyone to Speak 30 Languages via Zero-Shot Cross-Lingual Voice Cloning

DGX agent

arXiv:2605.05611v2 Announce Type: replace-cross Abstract: In this paper, we present X-Voice, a 0.4B multilingual zero-shot voice cloning model that clones arbitrary voices and enables everyone to spea

safetyarxiv-cs-ai
12 May 2026
Research

Yeti: A compact protein structure tokenizer for reconstruction and multi-modal generation

DGX agent

arXiv:2605.09981v1 Announce Type: cross Abstract: Multimodal models that jointly reason over protein sequences, structures, and function annotations within a unified representation hold immense potent

researcharxiv-cs-ai
12 May 2026
Applications

Yield Curve Forecasting using Machine Learning and Econometrics: A Comparative Analysis

DGX agent

arXiv:2605.09842v1 Announce Type: new Abstract: While machine learning has revolutionized many fields such as natural language processing (NLP) and computer vision, its impact on time-series forecasti

applicationsarxiv-cs-ai
12 May 2026
Model Releases

You Have Been LaTeXpOsEd: A Systematic Analysis of Information Leakage in Preprint Archives Using Large Language Models

DGX agent

arXiv:2510.03761v2 Announce Type: replace-cross Abstract: The widespread use of preprint repositories such as arXiv has accelerated the communication of scientific results but also introduced overlook

model-releasesarxiv-cs-ai
12 May 2026
Agents

Your Recourse, My Loss? Algorithmic Recourse under Shared Constraints

DGX agent

arXiv:2508.11070v2 Announce Type: replace Abstract: Decision makers are increasingly relying on machine learning in sensitive situations. Algorithmic recourse aims to provide individuals with actionab

agentsarxiv-cs-ai
12 May 2026
Model Releases

Your Simulation Runs but Solves the Wrong Physics: PDE-Grounded Intent Verification for LLM-Generated Multiphysics Simulation Code

DGX agent

arXiv:2605.09360v1 Announce Type: cross Abstract: Execution-based evaluation of LLM-generated code implicitly treats successful execution as a proxy for correctness. In scientific simulation, this pro

model-releasesarxiv-cs-ai
12 May 2026
Research

ZAYA1-VL-8B Technical Report

DGX agent

arXiv:2605.08560v1 Announce Type: cross Abstract: We present ZAYA1-VL-8B, a compact mixture-of-experts vision-language model built upon our in-house language model, ZAYA1-8B. Despite its compact size,

researcharxiv-cs-ai
12 May 2026
Agents

Zero-shot Imitation Learning by Latent Topology Mapping

DGX agent

arXiv:2605.08450v1 Announce Type: cross Abstract: Imitation learning is effective for training agents when expert demonstrations are available, but collecting demonstrations for every complex task in

agentsarxiv-cs-ai
12 May 2026
Model Releases

2.5-D Decomposition for LLM-Based Spatial Construction

DGX agent

arXiv:2605.07066v1 Announce Type: new Abstract: Autonomous systems that build structures from natural-language instructions need reliable spatial reasoning, yet large language models (LLMs) make syste

model-releasesarxiv-cs-ai
11 May 2026
Research

A Computer Vision Pipeline for Individual-Level Behavior Analysis: Benchmarking on the Edinburgh Pig Dataset

DGX agent

arXiv:2509.12047v2 Announce Type: replace-cross Abstract: Animal behavior analysis plays a crucial role in understanding animal welfare, health status, and productivity in agricultural settings. Howev

researcharxiv-cs-ai
11 May 2026
Safety

A Generalized Singular Value Theory for Neural Networks

DGX agent

arXiv:2605.06938v1 Announce Type: cross Abstract: Building on the abstract Generalized Singular Value Decomposition (GSVD) theory of Brown et al. [2025], we prove that most modern neural architectures

safetyarxiv-cs-ai
11 May 2026
Safety

A Geometric Taxonomy of Hallucinations in LLMs

DGX agent

arXiv:2602.13224v3 Announce Type: replace Abstract: Hallucinations in deployed language models can have real consequences for downstream decisions in domains such as healthcare, legal, and financial s

safetyarxiv-cs-ai
11 May 2026
Research

A Hybrid Graph Neural Network for Enhanced EEG-Based Depression Detection

DGX agent

arXiv:2410.18103v2 Announce Type: replace-cross Abstract: Graph neural networks (GNNs) are becoming increasingly popular for EEG-based depression detection. However, previous GNN-based methods fail to

researcharxiv-cs-ai
11 May 2026
Safety

A Large-Scale Dataset for Molecular Structure-Language Description via a Rule-Regularized Method

DGX agent

arXiv:2602.02320v3 Announce Type: replace-cross Abstract: Molecular function is largely determined by structure. Accurately aligning molecular structure with natural language is therefore essential fo

safetyarxiv-cs-ai
11 May 2026
Research

A Linear-Transformer Hybrid for SNP-Based Genotype-to-Phenotype Prediction in Grapevine

DGX agent

arXiv:2605.06762v1 Announce Type: cross Abstract: Robust genotype-to-phenotype (G2P) prediction is essential for accelerating breeding decisions and genetic gain. However, it remains challenging to me

researcharxiv-cs-ai
11 May 2026
Agents

A Multi-Memory Segment System for Generating High-Quality Long-Term Memory Content in Agents

DGX agent

arXiv:2508.15294v4 Announce Type: replace Abstract: In the current field of agent memory, extensive explorations have been conducted in the area of memory retrieval, yet few studies have focused on ex

agentsarxiv-cs-ai
11 May 2026
Research

A Resilience Framework for Bi-Criteria Combinatorial Optimization with Bandit Feedback

DGX agent

arXiv:2503.12285v2 Announce Type: replace-cross Abstract: We study bi-criteria combinatorial optimization under noisy function evaluations. While resilience and black-box offline-to-online reductions

researcharxiv-cs-ai
11 May 2026
Research

A Rod Flow Model for Adam at the Edge of Stability

DGX agent

arXiv:2605.06821v1 Announce Type: cross Abstract: Cohen et al. (arXiv:2207.14484) observed that adaptive gradient methods such as Adam operate at the edge of stability. While there has been significan

researcharxiv-cs-ai
11 May 2026
Agents

A Self-Healing Framework for Reliable LLM-Based Autonomous Agents

DGX agent

arXiv:2605.06737v1 Announce Type: cross Abstract: Autonomous agents based on Large Language Models (LLMs) are increasingly being utilized in complex software systems. However, reliability remains a si

agentsarxiv-cs-ai
11 May 2026
Safety

A Statistical Framework for Algorithmic Collective Action with Multiple Collectives

DGX agent

arXiv:2605.06749v1 Announce Type: cross Abstract: As learning systems increasingly shape everyday decisions, Algorithmic Collective Action (ACA), i.e., users coordinating changes to shared data to ste

safetyarxiv-cs-ai
11 May 2026
Safety

A Systematic Investigation of The RL-Jailbreaker in LLMs

DGX agent

arXiv:2605.07032v1 Announce Type: cross Abstract: The evolution of generative models from next-token predictors to autonomous engines of complex systems necessitates rigorous safety hardening. Adversa

safetyarxiv-cs-ai
11 May 2026
Model Releases

A^2RD: Agentic Autoregressive Diffusion for Long Video Consistency

DGX agent

arXiv:2605.06924v1 Announce Type: cross Abstract: Synthesizing consistent and coherent long video remains a fundamental challenge. Existing methods suffer from semantic drift and narrative collapse ov

model-releasesarxiv-cs-ai
11 May 2026
Research

Abductive Reasoning with Probabilistic Commonsense

DGX agent

arXiv:2605.08011v1 Announce Type: new Abstract: Recent efforts to improve the reasoning abilities of Large Language Models (LLMs) have focused on integrating formal logic solvers within neurosymbolic

researcharxiv-cs-ai
11 May 2026
← Previous
1…346347348349350…448
Next →