AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
Safety

Temporal Contrastive Decoding: A Training-Free Method for Large Audio-Language Models

DGX agent

arXiv:2604.15383v1 Announce Type: cross Abstract: Large audio-language models (LALMs) generalize across speech, sound, and music, but unified decoders can exhibit a temporal smoothing bias: transient

safetyarxiv-cs-ai
20 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Applications

The Crutch or the Ceiling? How Different Generations of LLMs Shape EFL Student Writings

DGX agent

arXiv:2604.15460v1 Announce Type: cross Abstract: The rapid evolution of Large Language Models (LLMs) has made them powerful tools for enhancing student writing. This study explores the extent and lim

applicationsarxiv-cs-ai
20 Apr 2026
Model Releases

The Illusion of Equivalence: Systematic FP16 Divergence in KV-Cached Autoregressive Inference

DGX agent

arXiv:2604.15409v1 Announce Type: cross Abstract: KV caching is a ubiquitous optimization in autoregressive transformer inference, long presumed to be numerically equivalent to cache-free computation.

model-releasesarxiv-cs-ai
20 Apr 2026
Safety

The Informational Cost of Agency: A Bounded Measure of Interaction Efficiency for Deployed Reinforcement Learning

DGX agent

arXiv:2603.01283v2 Announce Type: replace Abstract: Deployed RL agents operate in closed-loop systems where reliable performance depends on maintaining coherent coupling between observations, actions,

safetyarxiv-cs-ai
20 Apr 2026
Safety

The Price of Paranoia: Robust Risk-Sensitive Cooperation in Non-Stationary Multi-Agent Reinforcement Learning

DGX agent

arXiv:2604.15695v1 Announce Type: cross Abstract: Cooperative equilibria are fragile. When agents learn alongside each other rather than in a fixed environment, the process of learning destabilizes th

safetyarxiv-cs-ai
20 Apr 2026
Model Releases

The Reasoning Trap: How Enhancing LLM Reasoning Amplifies Tool Hallucination

DGX agent

arXiv:2510.22977v2 Announce Type: replace-cross Abstract: Enhancing the reasoning capabilities of Large Language Models (LLMs) is a key strategy for building Agents that 'think then act.' However, rec

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

The Relic Condition: When Published Scholarship Becomes Material for Its Own Replacement

DGX agent

arXiv:2604.16116v1 Announce Type: cross Abstract: We extracted the scholarly reasoning systems of two internationally prominent humanities and social science scholars from their published corpora alon

model-releasesarxiv-cs-ai
20 Apr 2026
Agents

The Semi-Executable Stack: Agentic Software Engineering and the Expanding Scope of SE

DGX agent

arXiv:2604.15468v1 Announce Type: cross Abstract: AI-based systems, currently driven largely by LLMs and tool-using agentic harnesses, are increasingly discussed as a possible threat to software engin

agentsarxiv-cs-ai
20 Apr 2026
Research

The Synthetic Media Shift: Tracking the Rise, Virality, and Detectability of AI-Generated Multimodal Misinformation

DGX agent

arXiv:2604.15372v1 Announce Type: cross Abstract: As generative AI advances, the distinction between authentic and synthetic media is increasingly blurred, challenging the integrity of online informat

researcharxiv-cs-ai
20 Apr 2026
Hardware

The threat of analytic flexibility in using large language models to simulate human data

DGX agent

arXiv:2509.13397v3 Announce Type: replace-cross Abstract: Social scientists are now using large language models to create 'silicon samples': synthetic datasets intended to stand in for human responden

hardwarearxiv-cs-ai
20 Apr 2026
Agents

The World Leaks the Future: Harness Evolution for Future Prediction Agents

DGX agent

arXiv:2604.15719v1 Announce Type: new Abstract: Many consequential decisions must be made before the relevant outcome is known. Such problems are commonly framed as future prediction, where an LLM age

agentsarxiv-cs-ai
20 Apr 2026
Research

To LLM, or Not to LLM: How Designers and Developers Navigate LLMs as Tools or Teammates

DGX agent

arXiv:2604.15344v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly integrated into design and development workflows, yet decisions about their use are rarely binary or pur

researcharxiv-cs-ai
20 Apr 2026
Safety

Towards Intrinsic Interpretability of Large Language Models:A Survey of Design Principles and Architectures

DGX agent

arXiv:2604.16042v1 Announce Type: cross Abstract: While Large Language Models (LLMs) have achieved strong performance across many NLP tasks, their opaque internal mechanisms hinder trustworthiness and

safetyarxiv-cs-ai
20 Apr 2026
Research

Towards Rigorous Explainability by Feature Attribution

DGX agent

arXiv:2604.15898v1 Announce Type: new Abstract: For around a decade, non-symbolic methods have been the option of choice when explaining complex machine learning (ML) models. Unfortunately, such metho

researcharxiv-cs-ai
20 Apr 2026
Hardware

Towards Understanding, Analyzing, and Optimizing Agentic AI Execution: A CPU-Centric Perspective

DGX agent

arXiv:2511.00739v3 Announce Type: replace Abstract: Agentic AI serving converts monolithic LLM-based inference to autonomous problem-solvers that can plan, call tools, perform reasoning, and adapt on

hardwarearxiv-cs-ai
20 Apr 2026
Research

TPA: Next Token Probability Attribution for Detecting Hallucinations in RAG

DGX agent

arXiv:2512.07515v4 Announce Type: replace-cross Abstract: Detecting hallucinations in Retrieval-Augmented Generation remains a challenge. Prior approaches attribute hallucinations to a binary conflict

researcharxiv-cs-ai
20 Apr 2026
Research

Training Time Prediction for Mixed Precision-based Distributed Training

DGX agent

arXiv:2604.16145v1 Announce Type: cross Abstract: Accurate prediction of training time in distributed deep learning is crucial for resource allocation, cost estimation, and job scheduling. We observe

researcharxiv-cs-ai
20 Apr 2026
Research

Transfer Learning from Foundational Optimization Embeddings to Unsupervised SAT Representations

DGX agent

arXiv:2604.15448v1 Announce Type: cross Abstract: Foundational optimization embeddings have recently emerged as powerful pre-trained representations for mixed-integer programming (MIP) problems. These

researcharxiv-cs-ai
20 Apr 2026
Model Releases

Transformer Neural Processes - Kernel Regression

DGX agent

arXiv:2411.12502v4 Announce Type: replace-cross Abstract: Neural Processes (NPs) are a rapidly evolving class of models designed to directly model the posterior predictive distribution of stochastic p

model-releasesarxiv-cs-ai
20 Apr 2026
Applications

TriagerX: Dual Transformers for Bug Triaging Tasks with Content and Interaction Based Rankings

DGX agent

arXiv:2508.16860v2 Announce Type: replace-cross Abstract: Pretrained Language Models or PLMs are transformer-based architectures that can be used in bug triaging tasks. PLMs can better capture token s

applicationsarxiv-cs-ai
20 Apr 2026
Research

Uncertainty, Vagueness, and Ambiguity in Human-Robot Interaction: Why Conceptualization Matters

DGX agent

arXiv:2604.15339v1 Announce Type: cross Abstract: Uncertainty, vagueness, and ambiguity are closely related and often confused concepts in human-robot interaction (HRI). In earlier studies, these conc

researcharxiv-cs-ai
20 Apr 2026
Model Releases

UniEditBench: A Unified and Cost-Effective Benchmark for Image and Video Editing via Distilled MLLMs

DGX agent

arXiv:2604.15871v1 Announce Type: cross Abstract: The evaluation of visual editing models remains fragmented across methods and modalities. Existing benchmarks are often tailored to specific paradigms

model-releasesarxiv-cs-ai
20 Apr 2026
Applications

Unveiling Stochasticity: Universal Multi-modal Probabilistic Modeling for Traffic Forecasting

DGX agent

arXiv:2604.16084v1 Announce Type: cross Abstract: Traffic forecasting is a challenging spatio-temporal modeling task and a critical component of urban transportation management. Current studies mainly

applicationsarxiv-cs-ai
20 Apr 2026
Applications

Using Large Language Models and Knowledge Graphs to Improve the Interpretability of Machine Learning Models in Manufacturing

DGX agent

arXiv:2604.16280v1 Announce Type: new Abstract: Explaining Machine Learning (ML) results in a transparent and user-friendly manner remains a challenging task of Explainable Artificial Intelligence (XA

applicationsarxiv-cs-ai
20 Apr 2026
Model Releases

VEFX-Bench: A Holistic Benchmark for Generic Video Editing and Visual Effects

DGX agent

arXiv:2604.16272v1 Announce Type: cross Abstract: As AI-assisted video creation becomes increasingly practical, instruction-guided video editing has become essential for refining generated or captured

model-releasesarxiv-cs-ai
20 Apr 2026
Research

VeriCWEty: Embedding enabled Line-Level CWE Detection in Verilog

DGX agent

arXiv:2604.15375v1 Announce Type: cross Abstract: Large Language Models (LLMs) have shown significant improvement in RTL code generation. Despite the advances, the generated code is often riddled with

researcharxiv-cs-ai
20 Apr 2026
Agents

VeriGraph: Scene Graphs for Execution Verifiable Robot Planning

DGX agent

arXiv:2411.10446v3 Announce Type: replace-cross Abstract: Recent progress in vision-language models (VLMs) has opened new possibilities for robot task planning, but these models often produce incorrec

agentsarxiv-cs-ai
20 Apr 2026
Agents

VeriMoA: A Mixture-of-Agents Framework for Spec-to-HDL Generation

DGX agent

arXiv:2510.27617v2 Announce Type: replace Abstract: Automation of Register Transfer Level (RTL) design can help developers meet increasing computational demands. Large Language Models (LLMs) show prom

agentsarxiv-cs-ai
20 Apr 2026
Research

VIB-Probe: Detecting and Mitigating Hallucinations in Vision-Language Models via Variational Information Bottleneck

DGX agent

arXiv:2601.05547v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) have demonstrated remarkable progress in multimodal tasks, but remain susceptible to hallucinations, where gener

researcharxiv-cs-ai
20 Apr 2026
Model Releases

vla-eval: A Unified Evaluation Harness for Vision-Language-Action Models

DGX agent

arXiv:2603.13966v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) models are increasingly evaluated across multiple simulation benchmarks, yet adding each benchmark to an evaluation pip

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

VLegal-Bench: Cognitively Grounded Benchmark for Vietnamese Legal Reasoning of Large Language Models

DGX agent

arXiv:2512.14554v5 Announce Type: replace-cross Abstract: The rapid advancement of large language models (LLMs) has enabled new possibilities for applying artificial intelligence within the legal doma

model-releasesarxiv-cs-ai
20 Apr 2026
Local Ai

VoodooNet: Achieving Analytic Ground States via High-Dimensional Random Projections

DGX agent

arXiv:2604.15613v1 Announce Type: cross Abstract: We present VoodooNet, a non-iterative neural architecture that replaces the stochastic gradient descent (SGD) paradigm with a closed-form analytic sol

local-aiarxiv-cs-ai
20 Apr 2026
Research

WARBERT: A Hierarchical BERT-based Model for Web API Recommendation

DGX agent

arXiv:2509.23175v2 Announce Type: replace-cross Abstract: With the rise of Web 2.0 and microservices, the increasing availability of Web APIs has intensified the need for effective recommendation syst

researcharxiv-cs-ai
20 Apr 2026
Agents

Weak-Link Optimization for Multi-Agent Reasoning and Collaboration

DGX agent

arXiv:2604.15972v1 Announce Type: new Abstract: LLM-driven multi-agent frameworks address complex reasoning tasks through multi-role collaboration. However, existing approaches often suffer from reaso

agentsarxiv-cs-ai
20 Apr 2026
Model Releases

When Cultures Meet: Multicultural Text-to-Image Generation

DGX agent

arXiv:2502.15972v2 Announce Type: replace-cross Abstract: Text-to-image generation models have achieved strong performance in culturally homogeneous settings, yet their ability to generate multicultur

model-releasesarxiv-cs-ai
20 Apr 2026
Research

When Do Early-Exit Networks Generalize? A PAC-Bayesian Theory of Adaptive Depth

DGX agent

arXiv:2604.15764v1 Announce Type: cross Abstract: Early-exit neural networks enable adaptive computation by allowing confident predictions to exit at intermediate layers, achieving 2-8imes inference s

researcharxiv-cs-ai
20 Apr 2026
Safety

When Search Goes Wrong: Red-Teaming Web-Augmented Large Language Models

DGX agent

arXiv:2510.09689v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have been augmented with web search to overcome the limitations of the static knowledge boundary by accessing up-

safetyarxiv-cs-ai
20 Apr 2026
Applications

When the Loop Closes: Architectural Limits of In-Context Isolation, Metacognitive Co-option, and the Two-Target Design Problem in Human-LLM Systems

DGX agent

arXiv:2604.15343v1 Announce Type: cross Abstract: We report a detailed autoethnographic case study of a single-subject who deliberately constructed and operated a multi-modal prompt-engineering system

applicationsarxiv-cs-ai
20 Apr 2026
Research

Where does output diversity collapse in post-training?

DGX agent

arXiv:2604.16027v1 Announce Type: cross Abstract: Post-trained language models produce less varied outputs than their base counterparts. This output diversity collapse undermines inference-time scalin

researcharxiv-cs-ai
20 Apr 2026
Model Releases

Why Fine-Tuning Encourages Hallucinations and How to Fix It

DGX agent

arXiv:2604.15574v1 Announce Type: cross Abstract: Large language models are prone to hallucinating factually incorrect statements. A key source of these errors is exposure to new factual information t

model-releasesarxiv-cs-ai
20 Apr 2026
Agents

WiseMind: a knowledge-guided multi-agent framework for accurate and empathetic psychiatric diagnosis

DGX agent

arXiv:2502.20689v4 Announce Type: replace Abstract: Large Language Models (LLMs) offer promising opportunities to support mental healthcare workflows, yet they often lack the structured clinical reaso

agentsarxiv-cs-ai
20 Apr 2026
Research

Zoom Consistency: A Free Confidence Signal in Multi-Step Visual Grounding Pipelines

DGX agent

arXiv:2604.15376v1 Announce Type: cross Abstract: Multi-step zoom-in pipelines are widely used for GUI grounding, yet the intermediate predictions they produce are typically discarded after coordinate

researcharxiv-cs-ai
20 Apr 2026
Model Releases

3D Instruction Ambiguity Detection

DGX agent

arXiv:2601.05991v2 Announce Type: replace Abstract: In safety-critical domains, linguistic ambiguity can have severe consequences; a vague command like 'Pass me the vial' in a surgical setting could l

model-releasesarxiv-cs-ai
17 Apr 2026
Tutorials

A Pythonic Functional Approach for Semantic Data Harmonisation in the ILIAD Project

DGX agent

arXiv:2604.13042v1 Announce Type: cross Abstract: Semantic data harmonisation is a central requirement in the ILIAD project, where heterogeneous environmental data must be harmonised according to the

tutorialsarxiv-cs-ai
17 Apr 2026
Agents

AgentForge: Execution-Grounded Multi-Agent LLM Framework for Autonomous Software Engineering

DGX agent

arXiv:2604.13120v1 Announce Type: cross Abstract: Large language models generate plausible code but cannot verify correctness. Existing multi-agent systems simulate execution or leave verification opt

agentsarxiv-cs-ai
17 Apr 2026
Agents

Agentic AI Optimisation (AAIO): what it is, how it works, why it matters, and how to deal with it

DGX agent

arXiv:2504.12482v2 Announce Type: replace Abstract: The emergence of Agentic Artificial Intelligence (AAI) systems capable of independently initiating digital interactions necessitates a new optimisat

agentsarxiv-cs-ai
17 Apr 2026
Model Releases

AI-Assisted Peer Review at Scale: The AAAI-26 AI Review Pilot

DGX agent

arXiv:2604.13940v1 Announce Type: new Abstract: Scientific peer review faces mounting strain as submission volumes surge, making it increasingly difficult to sustain review quality, consistency, and t

model-releasesarxiv-cs-ai
17 Apr 2026
Research

AlphaCNOT: Learning CNOT Minimization with Model-Based Planning

DGX agent

arXiv:2604.13812v1 Announce Type: new Abstract: Quantum circuit optimization is a central task in Quantum Computing, as current Noisy Intermediate Scale Quantum devices suffer from error propagation t

researcharxiv-cs-ai
17 Apr 2026
← Previous
1…405406407408409…443
Next →