AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Agents

Task-Level AI Readiness Assessment for Business Process Management:The T-IPO Model and LARA Matrix in Financial-Services IT Operations

DGX agent

arXiv:2605.16297v1 Announce Type: cross Abstract: Which tasks inside an enterprise workflow can a large-language-model agent reliably handle, and under what conditions? Most business process modeling

agentsarxiv-cs-ai
19 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

TaskGround: Structured Executable Task Inference for Full-Scene Household Reasoning

DGX agent

arXiv:2605.18109v1 Announce Type: new Abstract: In real home deployments, household agents must often operate from a complete household scene and a situated household request, rather than from a clean

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Taxonomy and Consistency Analysis of Safety Benchmarks for AI Agents

DGX agent

arXiv:2605.16282v1 Announce Type: cross Abstract: The rapid deployment of LLM-based autonomous agents has introduced safety risks that extend far beyond traditional LLM concerns, prompting a prolifera

model-releasesarxiv-cs-ai
19 May 2026
Safety

TClone: Low-Latency Forking of Live GUI Environments for Computer-Use Agents

DGX agent

arXiv:2605.17320v1 Announce Type: cross Abstract: Computer-use agents increasingly operate inside live personal workspaces, where their actions can modify files, applications, GUI state, credentials,

safetyarxiv-cs-ai
19 May 2026
Model Releases

TeleCom-Bench: How Far Are Large Language Models from Industrial Telecommunication Applications?

DGX agent

arXiv:2605.18025v1 Announce Type: new Abstract: While Large Language Models have achieved remarkable integration in various vertical scenarios, their deployment in the telecommunications domain remain

model-releasesarxiv-cs-ai
19 May 2026
Research

Temporal Aware Pruning for Efficient Diffusion-based Video Generation

DGX agent

arXiv:2605.17837v1 Announce Type: cross Abstract: Video diffusion models have recently enabled high-quality video generation with ViT-based architectures, but remain computationally intensive because

researcharxiv-cs-ai
19 May 2026
Safety

The Alien Space of Science: Sampling Coherent but Cognitively Unavailable Research Directions

DGX agent

arXiv:2603.01092v2 Announce Type: replace Abstract: Scientific discovery is constrained not only by what is true, but by what is cognitively available to the researchers currently exploring a field. M

safetyarxiv-cs-ai
19 May 2026
Agents

The Alpha Illusion: Reported Alpha from LLM Trading Agents Should Not Be Treated as Deployment Evidence

DGX agent

arXiv:2605.16895v1 Announce Type: cross Abstract: End-to-end LLM trading agents have moved quickly from research curiosity to a small ecosystem of named systems, including FinCon, FinMem, TradingAgent

agentsarxiv-cs-ai
19 May 2026
Safety

The Bayesian Geometry of Transformer Attention

DGX agent

arXiv:2512.22471v5 Announce Type: replace-cross Abstract: Transformers often appear to perform Bayesian reasoning in context, but verifying this rigorously has been impossible: natural data lack analy

safetyarxiv-cs-ai
19 May 2026
Safety

The Capability Paradox: How Smarter Auditors Make Multi-Agent Systems Less Secure

DGX agent

arXiv:2605.17480v1 Announce Type: new Abstract: Multi-agent systems extend large language models (LLMs) by decomposing tasks among specialized agents, but their distributed decision process creates ne

safetyarxiv-cs-ai
19 May 2026
Agents

The End of Trust: How Agentic AI Breaks Security Assumptions

DGX agent

arXiv:2605.16436v1 Announce Type: cross Abstract: For decades, the security of digital interaction has rested on an unacknowledged economic constraint. Attackers faced a tradeoff between the fidelity

agentsarxiv-cs-ai
19 May 2026
Research

The Expert Strikes Back: Interpreting Mixture-of-Experts Language Models at Expert Level

DGX agent

arXiv:2604.02178v2 Announce Type: replace-cross Abstract: Mixture-of-Experts (MoE) architectures have become the dominant choice for scaling Large Language Models (LLMs), activating only a subset of p

researcharxiv-cs-ai
19 May 2026
Safety

The Hidden Cost of Contextual Sycophancy: an AI Literacy Intervention in Human-AI Collaboration

DGX agent

arXiv:2605.18372v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used in educational settings as interactive tools for collaboration. However, their tendency toward syco

safetyarxiv-cs-ai
19 May 2026
Model Releases

The Illusion of Specialization: Unveiling the Domain-Invariant 'Standing Committee' in Mixture-of-Experts Models

DGX agent

arXiv:2601.03425v2 Announce Type: replace-cross Abstract: Mixture of Experts models are widely assumed to achieve domain specialization through sparse routing. In this work, we question this assumptio

model-releasesarxiv-cs-ai
19 May 2026
Safety

The Impact of AI Search on the Online Content Ecosystem: Evidence from Google and Reddit

DGX agent

arXiv:2605.16428v1 Announce Type: cross Abstract: Search engines traditionally complement online content platforms by directing users seeking information to external websites. The emergence of generat

safetyarxiv-cs-ai
19 May 2026
Research

The IsalProgram Programming Language

DGX agent

arXiv:2605.17008v1 Announce Type: cross Abstract: We introduce IsalProgram (Instruction Set and Language for Programming), a novel assembly-like programming language with three distinctive theoretical

researcharxiv-cs-ai
19 May 2026
Agents

The Journal of Prompt-Engineered (Moral) Philosophy Or: Why AI-Assisted Ethics Research Requires Process Transparency

DGX agent

arXiv:2511.08639v3 Announce Type: replace-cross Abstract: Existing AI disclosure mandates in scholarship require that AI assistance be reported but leave transparency philosophically unspecified: they

agentsarxiv-cs-ai
19 May 2026
Safety

The Laplacian Keyboard: Beyond the Linear Span

DGX agent

arXiv:2602.07730v2 Announce Type: replace-cross Abstract: Across scientific disciplines, Laplacian eigenvectors serve as a fundamental basis for simplifying complex systems, from signal processing to

safetyarxiv-cs-ai
19 May 2026
Research

The Lattice Representation Hypothesis of Large Language Models

DGX agent

arXiv:2603.01227v2 Announce Type: replace Abstract: We propose the Lattice Representation Hypothesis of large language models: a symbolic backbone that grounds conceptual hierarchies and logical opera

researcharxiv-cs-ai
19 May 2026
Research

The Loupe: A Plug-and-Play Attention Module for Amplifying Discriminative Features in Vision Transformers

DGX agent

arXiv:2508.16663v2 Announce Type: replace-cross Abstract: Fine-Grained Visual Classification (FGVC) requires models to focus on subtle, task-relevant regions rather than broad object context. We prese

researcharxiv-cs-ai
19 May 2026
Local Ai

The Point of No Return: Counterfactual Localization of Deceptive Commitment in Language-Model Reasoning

DGX agent

arXiv:2605.17113v1 Announce Type: cross Abstract: Existing deception datasets label completed outputs as honest or deceptive, treating deception as a property of the final response rather than a funct

local-aiarxiv-cs-ai
19 May 2026
Applications

The Recovery Mechanism: Technology, Education, and What Happens When the Pattern Breaks

DGX agent

arXiv:2605.16283v1 Announce Type: cross Abstract: For centuries, each new technology has automated some layer of cognitive work and been absorbed by education retreating upward to teach the skills mac

applicationsarxiv-cs-ai
19 May 2026
Model Releases

The Scaling Laws of Skills in LLM Agent Systems

DGX agent

arXiv:2605.16508v1 Announce Type: cross Abstract: As agent systems scale, skills accumulate into large reusable libraries, yet their scaling laws remain poorly understood. Across 15 frontier LLMs, 1,1

model-releasesarxiv-cs-ai
19 May 2026
Research

The Token Games: Evaluating Language Model Reasoning with Puzzle Duels

DGX agent

arXiv:2602.17831v2 Announce Type: replace Abstract: Evaluating the reasoning capabilities of Large Language Models is increasingly challenging as models improve. Human curation of hard questions is hi

researcharxiv-cs-ai
19 May 2026
Model Releases

'The Whole Is Greater Than the Sum of Its Parts': A Compatibility-Aware Multi-Teacher CoT Distillation Framework

DGX agent

arXiv:2601.13992v2 Announce Type: replace-cross Abstract: Chain-of-Thought (CoT) reasoning empowers Large Language Models (LLMs) with remarkable capabilities but typically requires prohibitive paramet

model-releasesarxiv-cs-ai
19 May 2026
Local Ai

Thinking with Patterns: Breaking the Perceptual Bottleneck in Visual Planning via Pattern Induction

DGX agent

arXiv:2605.16848v1 Announce Type: cross Abstract: Planning from raw visual input remains a significant challenge for current Vision-Language Models (VLMs), when the complexity of input is beyond their

local-aiarxiv-cs-ai
19 May 2026
Model Releases

TIER: Trajectory-Invariant Execution Rewards for Multi-Step Tool Composition

DGX agent

arXiv:2605.16790v1 Announce Type: cross Abstract: Tool use enables large language models to solve complex tasks through sequences of API calls, yet existing reinforcement learning approaches fail to s

model-releasesarxiv-cs-ai
19 May 2026
Hardware

TierCheck: Tiered Checkpointing for Fault Tolerance in Large Language Model Training

DGX agent

arXiv:2605.17821v1 Announce Type: cross Abstract: Large Language Model (LLM) training is frequently interrupted by a heterogeneous spectrum of failures, from common GPU crashes to catastrophic cluster

hardwarearxiv-cs-ai
19 May 2026
Model Releases

Time-Efficient Hybrid Hyperparameter Tuning Approach for Cardiovascular Disease Classification

DGX agent

arXiv:2411.18234v2 Announce Type: replace-cross Abstract: Cardiovascular diseases (CVDs) are any serious illness of the heart, which require accurate diagnosis to prevent fatal consequences. Hyperpara

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

TinySAM 2: Extreme Memory Compression for Efficient Track Anything Model

DGX agent

arXiv:2605.18013v1 Announce Type: cross Abstract: Segment Anything Model 2 (SAM 2) serves as a core foundation model in the field of video segmentation. Building upon the original SAM model, it introd

model-releasesarxiv-cs-ai
19 May 2026
Research

To Trust or Not to Trust: Authors' Response to AI-based Reviews

DGX agent

arXiv:2605.16623v1 Announce Type: cross Abstract: Large language models are increasingly discussed and used as tools that may assist with scholarly peer review, but empirical evidence regarding how au

researcharxiv-cs-ai
19 May 2026
Model Releases

TOBench: A Task-Oriented Omni-Modal Benchmark for Real-World Tool-Using Agents

DGX agent

arXiv:2605.16909v1 Announce Type: new Abstract: Tool-using agents are increasingly expected to operate across realistic professional workflows, where they must interpret multimodal inputs, coordinate

model-releasesarxiv-cs-ai
19 May 2026
Agents

Tongyi DeepResearch Technical Report

DGX agent

arXiv:2510.24701v3 Announce Type: replace-cross Abstract: We present Tongyi DeepResearch, an agentic large language model, which is specifically designed for long-horizon, deep information-seeking res

agentsarxiv-cs-ai
19 May 2026
Safety

Toward Robust Multilingual Adaptation of LLMs for Low-Resource Languages

DGX agent

arXiv:2510.14466v3 Announce Type: replace-cross Abstract: Large language models (LLMs) continue to struggle with low-resource languages, primarily due to limited training data, translation noise, and

safetyarxiv-cs-ai
19 May 2026
Research

Toward Template-Free Explainability for Monte Carlo Tree Search

DGX agent

arXiv:2605.16524v1 Announce Type: cross Abstract: Probabilistic search algorithms, such as Monte Carlo Tree Search (MCTS), have proven very effective in solving sequential decision-making tasks under

researcharxiv-cs-ai
19 May 2026
Agents

Towards Human-Level Book-Writing Capability

DGX agent

arXiv:2605.17064v1 Announce Type: new Abstract: Large language models optimized for instruction following and agentic tasks remain poorly aligned with the requirements of high-quality creative writing

agentsarxiv-cs-ai
19 May 2026
Research

Towards Robust Argumentative Essay Understanding via TIDE: An Interactive Framework with Trial and Debate

DGX agent

arXiv:2605.17247v1 Announce Type: new Abstract: Argumentative essays serve as a vital medium for assessing critical thinking and reasoning skills, yet there is limited works on accurately understandin

researcharxiv-cs-ai
19 May 2026
Safety

Towards Sustainable Growth: A Multi-Value-Aware Retrieval Framework for E-Commerce Search

DGX agent

arXiv:2605.17994v1 Announce Type: cross Abstract: New item growth is critical for maintaining a healthy ecosystem in large-scale e-commerce platforms. However, existing systems tend to prioritize pres

safetyarxiv-cs-ai
19 May 2026
Research

Towards Ubiquitous Mapping and Localization for Dynamic Indoor Environments

DGX agent

arXiv:2605.18385v1 Announce Type: cross Abstract: We present UbiSLAM, an innovative solution for real-time mapping and localization in dynamic indoor environments. By deploying a network of fixed RGB-

researcharxiv-cs-ai
19 May 2026
Research

TRACE: Trajectory Correction from Cross-layer Evidence for Hallucination Reduction

DGX agent

arXiv:2605.18163v1 Announce Type: new Abstract: Hallucination correction is not a one-direction problem. We show that intermediate layers are neither uniformly more truthful than final layers nor unif

researcharxiv-cs-ai
19 May 2026
Applications

Tracking Drift: Variation-Aware Entropy Scheduling for Non-Stationary Reinforcement Learning

DGX agent

arXiv:2601.19624v2 Announce Type: replace-cross Abstract: Real-world reinforcement learning often faces environment drift, but most existing methods rely on static entropy coefficients/target entropy,

applicationsarxiv-cs-ai
19 May 2026
Agents

Train the Trainers -- An Agentic AI Framework for Peer-Based Mental Health Support in Battlefield Environments

DGX agent

arXiv:2605.16269v1 Announce Type: cross Abstract: Modern military operations expose soldiers to sustained psychological stress, leading to acute reactions, post-traumatic stress symptoms, and other me

agentsarxiv-cs-ai
19 May 2026
Applications

Training data attribution in diffusion models via mirrored unlearning and noise-consistent skew

DGX agent

arXiv:2605.17938v1 Announce Type: cross Abstract: Training data attribution (TDA) should enable generative model interpretability and foster a variety of related downstream tasks. Nonetheless, current

applicationsarxiv-cs-ai
19 May 2026
Research

Training Infinitely Deep and Wide Transformers

DGX agent

arXiv:2605.17660v1 Announce Type: cross Abstract: Transformers have become the dominant architecture in modern machine learning, yet the theoretical understanding of their training dynamics remains li

researcharxiv-cs-ai
19 May 2026
Agents

Trajectory-Aware Adaptive Inference in Object Detection Models

DGX agent

arXiv:2605.16397v1 Announce Type: cross Abstract: The increasing integration of sensors in autonomous maritime navigation has led to large-scale multimodal datasets, raising challenges in achieving ef

agentsarxiv-cs-ai
19 May 2026
Model Releases

Transitivity Meets Cyclicity: Explicit Preference Decomposition for Dynamic Large Language Model Alignment

DGX agent

arXiv:2605.17342v1 Announce Type: cross Abstract: Standard RLHF relies on transitive scalar rewards, failing to capture the cyclic nature of human preferences. While some approaches like the General P

model-releasesarxiv-cs-ai
19 May 2026
Tutorials

Trust the uncertain teacher: distilling dark knowledge via calibrated uncertainty

DGX agent

arXiv:2602.12687v2 Announce Type: replace-cross Abstract: The core of knowledge distillation lies in transferring the teacher's rich 'dark knowledge'-subtle probabilistic patterns that reveal how clas

tutorialsarxiv-cs-ai
19 May 2026
Model Releases

Trustworthiness in Retrieval-Augmented Generation Systems: A Survey

DGX agent

arXiv:2409.10102v2 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) has quickly grown into a pivotal paradigm in the development of Large Language Models (LLMs). Although ex

model-releasesarxiv-cs-ai
19 May 2026
← Previous
1…304305306307308…448
Next →