AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
19 May 2026

Task-Level AI Readiness Assessment for Business Process Management:The T-IPO Model and LARA Matrix in Financial-Services IT Operations

AgentsDGX agent

arXiv:2605.16297v1 Announce Type: cross Abstract: Which tasks inside an enterprise workflow can a large-language-model agent reliably handle, and under what conditions? Most business process modeling

TaskGround: Structured Executable Task Inference for Full-Scene Household Reasoning

Model ReleasesDGX agent

arXiv:2605.18109v1 Announce Type: new Abstract: In real home deployments, household agents must often operate from a complete household scene and a situated household request, rather than from a clean

Taxonomy and Consistency Analysis of Safety Benchmarks for AI Agents

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.16282v1 Announce Type: cross Abstract: The rapid deployment of LLM-based autonomous agents has introduced safety risks that extend far beyond traditional LLM concerns, prompting a prolifera

TClone: Low-Latency Forking of Live GUI Environments for Computer-Use Agents

SafetyDGX agent

arXiv:2605.17320v1 Announce Type: cross Abstract: Computer-use agents increasingly operate inside live personal workspaces, where their actions can modify files, applications, GUI state, credentials,

TeleCom-Bench: How Far Are Large Language Models from Industrial Telecommunication Applications?

Model ReleasesDGX agent

arXiv:2605.18025v1 Announce Type: new Abstract: While Large Language Models have achieved remarkable integration in various vertical scenarios, their deployment in the telecommunications domain remain

Temporal Aware Pruning for Efficient Diffusion-based Video Generation

ResearchDGX agent

arXiv:2605.17837v1 Announce Type: cross Abstract: Video diffusion models have recently enabled high-quality video generation with ViT-based architectures, but remain computationally intensive because

The Alien Space of Science: Sampling Coherent but Cognitively Unavailable Research Directions

SafetyDGX agent

arXiv:2603.01092v2 Announce Type: replace Abstract: Scientific discovery is constrained not only by what is true, but by what is cognitively available to the researchers currently exploring a field. M

The Alpha Illusion: Reported Alpha from LLM Trading Agents Should Not Be Treated as Deployment Evidence

AgentsDGX agent

arXiv:2605.16895v1 Announce Type: cross Abstract: End-to-end LLM trading agents have moved quickly from research curiosity to a small ecosystem of named systems, including FinCon, FinMem, TradingAgent

The Bayesian Geometry of Transformer Attention

SafetyDGX agent

arXiv:2512.22471v5 Announce Type: replace-cross Abstract: Transformers often appear to perform Bayesian reasoning in context, but verifying this rigorously has been impossible: natural data lack analy

The Capability Paradox: How Smarter Auditors Make Multi-Agent Systems Less Secure

SafetyDGX agent

arXiv:2605.17480v1 Announce Type: new Abstract: Multi-agent systems extend large language models (LLMs) by decomposing tasks among specialized agents, but their distributed decision process creates ne

The End of Trust: How Agentic AI Breaks Security Assumptions

AgentsDGX agent

arXiv:2605.16436v1 Announce Type: cross Abstract: For decades, the security of digital interaction has rested on an unacknowledged economic constraint. Attackers faced a tradeoff between the fidelity

The Expert Strikes Back: Interpreting Mixture-of-Experts Language Models at Expert Level

ResearchDGX agent

arXiv:2604.02178v2 Announce Type: replace-cross Abstract: Mixture-of-Experts (MoE) architectures have become the dominant choice for scaling Large Language Models (LLMs), activating only a subset of p

The Hidden Cost of Contextual Sycophancy: an AI Literacy Intervention in Human-AI Collaboration

SafetyDGX agent

arXiv:2605.18372v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used in educational settings as interactive tools for collaboration. However, their tendency toward syco

The Illusion of Specialization: Unveiling the Domain-Invariant 'Standing Committee' in Mixture-of-Experts Models

Model ReleasesDGX agent

arXiv:2601.03425v2 Announce Type: replace-cross Abstract: Mixture of Experts models are widely assumed to achieve domain specialization through sparse routing. In this work, we question this assumptio

The Impact of AI Search on the Online Content Ecosystem: Evidence from Google and Reddit

SafetyDGX agent

arXiv:2605.16428v1 Announce Type: cross Abstract: Search engines traditionally complement online content platforms by directing users seeking information to external websites. The emergence of generat

The IsalProgram Programming Language

ResearchDGX agent

arXiv:2605.17008v1 Announce Type: cross Abstract: We introduce IsalProgram (Instruction Set and Language for Programming), a novel assembly-like programming language with three distinctive theoretical

The Journal of Prompt-Engineered (Moral) Philosophy Or: Why AI-Assisted Ethics Research Requires Process Transparency

AgentsDGX agent

arXiv:2511.08639v3 Announce Type: replace-cross Abstract: Existing AI disclosure mandates in scholarship require that AI assistance be reported but leave transparency philosophically unspecified: they

The Laplacian Keyboard: Beyond the Linear Span

SafetyDGX agent

arXiv:2602.07730v2 Announce Type: replace-cross Abstract: Across scientific disciplines, Laplacian eigenvectors serve as a fundamental basis for simplifying complex systems, from signal processing to

The Lattice Representation Hypothesis of Large Language Models

ResearchDGX agent

arXiv:2603.01227v2 Announce Type: replace Abstract: We propose the Lattice Representation Hypothesis of large language models: a symbolic backbone that grounds conceptual hierarchies and logical opera

The Loupe: A Plug-and-Play Attention Module for Amplifying Discriminative Features in Vision Transformers

ResearchDGX agent

arXiv:2508.16663v2 Announce Type: replace-cross Abstract: Fine-Grained Visual Classification (FGVC) requires models to focus on subtle, task-relevant regions rather than broad object context. We prese

The Point of No Return: Counterfactual Localization of Deceptive Commitment in Language-Model Reasoning

Local AiDGX agent

arXiv:2605.17113v1 Announce Type: cross Abstract: Existing deception datasets label completed outputs as honest or deceptive, treating deception as a property of the final response rather than a funct

The Recovery Mechanism: Technology, Education, and What Happens When the Pattern Breaks

ApplicationsDGX agent

arXiv:2605.16283v1 Announce Type: cross Abstract: For centuries, each new technology has automated some layer of cognitive work and been absorbed by education retreating upward to teach the skills mac

The Scaling Laws of Skills in LLM Agent Systems

Model ReleasesDGX agent

arXiv:2605.16508v1 Announce Type: cross Abstract: As agent systems scale, skills accumulate into large reusable libraries, yet their scaling laws remain poorly understood. Across 15 frontier LLMs, 1,1

The Token Games: Evaluating Language Model Reasoning with Puzzle Duels

ResearchDGX agent

arXiv:2602.17831v2 Announce Type: replace Abstract: Evaluating the reasoning capabilities of Large Language Models is increasingly challenging as models improve. Human curation of hard questions is hi

'The Whole Is Greater Than the Sum of Its Parts': A Compatibility-Aware Multi-Teacher CoT Distillation Framework

Model ReleasesDGX agent

arXiv:2601.13992v2 Announce Type: replace-cross Abstract: Chain-of-Thought (CoT) reasoning empowers Large Language Models (LLMs) with remarkable capabilities but typically requires prohibitive paramet

Thinking with Patterns: Breaking the Perceptual Bottleneck in Visual Planning via Pattern Induction

Local AiDGX agent

arXiv:2605.16848v1 Announce Type: cross Abstract: Planning from raw visual input remains a significant challenge for current Vision-Language Models (VLMs), when the complexity of input is beyond their

TIER: Trajectory-Invariant Execution Rewards for Multi-Step Tool Composition

Model ReleasesDGX agent

arXiv:2605.16790v1 Announce Type: cross Abstract: Tool use enables large language models to solve complex tasks through sequences of API calls, yet existing reinforcement learning approaches fail to s

TierCheck: Tiered Checkpointing for Fault Tolerance in Large Language Model Training

HardwareDGX agent

arXiv:2605.17821v1 Announce Type: cross Abstract: Large Language Model (LLM) training is frequently interrupted by a heterogeneous spectrum of failures, from common GPU crashes to catastrophic cluster

Time-Efficient Hybrid Hyperparameter Tuning Approach for Cardiovascular Disease Classification

Model ReleasesDGX agent

arXiv:2411.18234v2 Announce Type: replace-cross Abstract: Cardiovascular diseases (CVDs) are any serious illness of the heart, which require accurate diagnosis to prevent fatal consequences. Hyperpara

TinySAM 2: Extreme Memory Compression for Efficient Track Anything Model

Model ReleasesDGX agent

arXiv:2605.18013v1 Announce Type: cross Abstract: Segment Anything Model 2 (SAM 2) serves as a core foundation model in the field of video segmentation. Building upon the original SAM model, it introd

To Trust or Not to Trust: Authors' Response to AI-based Reviews

ResearchDGX agent

arXiv:2605.16623v1 Announce Type: cross Abstract: Large language models are increasingly discussed and used as tools that may assist with scholarly peer review, but empirical evidence regarding how au

TOBench: A Task-Oriented Omni-Modal Benchmark for Real-World Tool-Using Agents

Model ReleasesDGX agent

arXiv:2605.16909v1 Announce Type: new Abstract: Tool-using agents are increasingly expected to operate across realistic professional workflows, where they must interpret multimodal inputs, coordinate

Tongyi DeepResearch Technical Report

AgentsDGX agent

arXiv:2510.24701v3 Announce Type: replace-cross Abstract: We present Tongyi DeepResearch, an agentic large language model, which is specifically designed for long-horizon, deep information-seeking res

Toward Robust Multilingual Adaptation of LLMs for Low-Resource Languages

SafetyDGX agent

arXiv:2510.14466v3 Announce Type: replace-cross Abstract: Large language models (LLMs) continue to struggle with low-resource languages, primarily due to limited training data, translation noise, and

Toward Template-Free Explainability for Monte Carlo Tree Search

ResearchDGX agent

arXiv:2605.16524v1 Announce Type: cross Abstract: Probabilistic search algorithms, such as Monte Carlo Tree Search (MCTS), have proven very effective in solving sequential decision-making tasks under

Towards Human-Level Book-Writing Capability

AgentsDGX agent

arXiv:2605.17064v1 Announce Type: new Abstract: Large language models optimized for instruction following and agentic tasks remain poorly aligned with the requirements of high-quality creative writing

Towards Robust Argumentative Essay Understanding via TIDE: An Interactive Framework with Trial and Debate

ResearchDGX agent

arXiv:2605.17247v1 Announce Type: new Abstract: Argumentative essays serve as a vital medium for assessing critical thinking and reasoning skills, yet there is limited works on accurately understandin

Towards Sustainable Growth: A Multi-Value-Aware Retrieval Framework for E-Commerce Search

SafetyDGX agent

arXiv:2605.17994v1 Announce Type: cross Abstract: New item growth is critical for maintaining a healthy ecosystem in large-scale e-commerce platforms. However, existing systems tend to prioritize pres

Towards Ubiquitous Mapping and Localization for Dynamic Indoor Environments

ResearchDGX agent

arXiv:2605.18385v1 Announce Type: cross Abstract: We present UbiSLAM, an innovative solution for real-time mapping and localization in dynamic indoor environments. By deploying a network of fixed RGB-

TRACE: Trajectory Correction from Cross-layer Evidence for Hallucination Reduction

ResearchDGX agent

arXiv:2605.18163v1 Announce Type: new Abstract: Hallucination correction is not a one-direction problem. We show that intermediate layers are neither uniformly more truthful than final layers nor unif

Tracking Drift: Variation-Aware Entropy Scheduling for Non-Stationary Reinforcement Learning

ApplicationsDGX agent

arXiv:2601.19624v2 Announce Type: replace-cross Abstract: Real-world reinforcement learning often faces environment drift, but most existing methods rely on static entropy coefficients/target entropy,

Train the Trainers -- An Agentic AI Framework for Peer-Based Mental Health Support in Battlefield Environments

AgentsDGX agent

arXiv:2605.16269v1 Announce Type: cross Abstract: Modern military operations expose soldiers to sustained psychological stress, leading to acute reactions, post-traumatic stress symptoms, and other me

Training data attribution in diffusion models via mirrored unlearning and noise-consistent skew

ApplicationsDGX agent

arXiv:2605.17938v1 Announce Type: cross Abstract: Training data attribution (TDA) should enable generative model interpretability and foster a variety of related downstream tasks. Nonetheless, current

Training Infinitely Deep and Wide Transformers

ResearchDGX agent

arXiv:2605.17660v1 Announce Type: cross Abstract: Transformers have become the dominant architecture in modern machine learning, yet the theoretical understanding of their training dynamics remains li

Trajectory-Aware Adaptive Inference in Object Detection Models

AgentsDGX agent

arXiv:2605.16397v1 Announce Type: cross Abstract: The increasing integration of sensors in autonomous maritime navigation has led to large-scale multimodal datasets, raising challenges in achieving ef

Transitivity Meets Cyclicity: Explicit Preference Decomposition for Dynamic Large Language Model Alignment

Model ReleasesDGX agent

arXiv:2605.17342v1 Announce Type: cross Abstract: Standard RLHF relies on transitive scalar rewards, failing to capture the cyclic nature of human preferences. While some approaches like the General P

Trust the uncertain teacher: distilling dark knowledge via calibrated uncertainty

TutorialsDGX agent

arXiv:2602.12687v2 Announce Type: replace-cross Abstract: The core of knowledge distillation lies in transferring the teacher's rich 'dark knowledge'-subtle probabilistic patterns that reveal how clas

Trustworthiness in Retrieval-Augmented Generation Systems: A Survey

Model ReleasesDGX agent

arXiv:2409.10102v2 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) has quickly grown into a pivotal paradigm in the development of Large Language Models (LLMs). Although ex

TTE-Flash: Accelerating Reasoning-based Multimodal Representations via Think-Then-Embed Tokens

Model ReleasesDGX agent

arXiv:2605.16638v1 Announce Type: new Abstract: Recent research has demonstrated that Universal Multimodal Embedding (UME) benefits significantly from Chain-of-Thought (CoT) reasoning. In this paradig

TusoAI: Agentic Optimization for Scientific Methods

Model ReleasesDGX agent

arXiv:2509.23986v2 Announce Type: replace Abstract: Scientific discovery is often slowed by the manual development of computational tools needed to analyze complex experimental data. Building such too

Two-Valued Symmetric Circulant Matrices: Applications in Deep Learning

ResearchDGX agent

arXiv:2605.16443v1 Announce Type: cross Abstract: Despite the success of deep neural networks in vision, medical diagnosis, and IoT scenarios, their deployment on resource-limited platforms poses seri

UCSF-PDGM-VQA: Visual Question Answering dataset for brain tumor MRI interpretation

Model ReleasesDGX agent

arXiv:2605.17140v1 Announce Type: cross Abstract: Brain tumor diagnosis is largely dependent on Magnetic Resonance Imaging (MRI) evaluation, which requires radiologists to synthesize thousands of imag

Uncertainty Quantification as a Principled Foundation for Explainable Artificial Intelligence: A Case Study of Counterfactual Explanations

ApplicationsDGX agent

arXiv:2502.17007v2 Announce Type: replace-cross Abstract: In this paper we argue that, to its detriment, transparency research overlooks many foundational concepts of artificial intelligence. As an il

Understanding In-Context Learning on Structured Manifolds: Bridging Attention to Kernel Methods

ResearchDGX agent

arXiv:2506.10959v3 Announce Type: replace-cross Abstract: While in-context learning (ICL) has achieved remarkable success in natural language and vision domains, its theoretical understanding-particul

UniAlign: A Model-Agnostic Framework for Robust Network Traffic Classification under Distribution Shifts

SafetyDGX agent

arXiv:2605.17575v1 Announce Type: cross Abstract: Network traffic classification (NTC) models often suffer severe performance degradation when deployed in real-world environments due to distribution s

UniER: A Unified Benchmark for Item-level and Path-level Exercise Recommendation

Model ReleasesDGX agent

arXiv:2605.16750v1 Announce Type: cross Abstract: Personalized exercise recommendation dynamically aligns pedagogical resources with individual knowledge mastery, which is crucial for satisfying stude

Universal Dynamics of Punctuated Progress

Model ReleasesDGX agent

arXiv:2605.16719v1 Announce Type: cross Abstract: Scientific and technological frontiers advance through punctuated dynamics, yet the principles governing these dynamics remain poorly understood. Here

Universal Time-Series Representation Learning: A Survey

ApplicationsDGX agent

arXiv:2401.03717v4 Announce Type: replace-cross Abstract: Time-series data exists in every corner of real-world systems and services, ranging from satellites in the sky to wearable devices on human bo

UniversalRAG: Retrieval-Augmented Generation over Corpora of Diverse Modalities and Granularities

ApplicationsDGX agent

arXiv:2504.20734v4 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) has shown substantial promise in improving factual accuracy by grounding model responses with external kn

Unlearning Isn't Deletion: Investigating Reversibility of Machine Unlearning in LLMs

SafetyDGX agent

arXiv:2505.16831v3 Announce Type: replace-cross Abstract: Unlearning in large language models (LLMs) aims to remove specified data, but its efficacy is typically assessed with task-level metrics like

← Previous
1…243244245246247…358
Next →