AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
20,981 results
11 Aug 2026

Temporal Generalization in fNIRS-Based Autism Classification: A Cross-Time-Window Transfer Benchmark

Model ReleasesDGX agent

arXiv:2608.07567v1 Announce Type: cross Abstract: Functional near-infrared spectroscopy (fNIRS) is a promising modality for autism spectrum disorder (ASD) classification, yet existing approaches assum

Temporal Misgrounding in Legal RAG: A Versioned-Corpus Benchmark for French Tax Law

Model ReleasesDGX agent

arXiv:2608.09393v1 Announce Type: cross Abstract: We identify and quantify temporal misgrounding: the systematic retrieval and citation of the currently in-force version of a legal article when the ap

Temporal Sepsis Modeling: a Relational and Explainable-by-Design Framework

ResearchDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2601.21747v4 Announce Type: replace-cross Abstract: Sepsis remains one of the most complex and heterogeneous syndromes in intensive care. While deep learning models achieve competitive performan

Test-Time Augmentation for LLMs: When Input Diversity Beats Output Diversity at Matched Compute

ResearchDGX agent

arXiv:2608.09351v1 Announce Type: cross Abstract: Test-time scaling improves LLM accuracy but multiplies inference cost, making the accuracy gained per unit of compute the metric that matters in deplo

TeXFix-Bench: An Empirically Grounded Multi-Format Benchmark for LLM-Based Document Source Repair

Model ReleasesDGX agent

arXiv:2608.07617v1 Announce Type: new Abstract: Scientific and technical writing depends on markup sources that must compile: LaTeX, Typst, and Markdown pipelines fail on missing delimiters, mismatche

TGIF: Text-Guided Layer Fusion Mitigates Hallucination in Multimodal LLMs

ResearchDGX agent

arXiv:2601.03100v3 Announce Type: replace-cross Abstract: Multimodal large language models (MLLMs) typically rely on a single late-layer feature from a frozen vision encoder, leaving the encoder's ric

The Anatomy of a Prompt Injection: A Component Model for Structured Analysis

SafetyDGX agent

arXiv:2608.07808v1 Announce Type: cross Abstract: Four years after prompt injection was first identified in 2022, attacks are still predominantly documented as verbatim strings rather than structured

The Announcement Carries the Cue: Markup, Boundaries, and the Notation of Pre-Training Corpora

ResearchDGX agent

arXiv:2608.09093v1 Announce Type: cross Abstract: How a document's arrangement is written down, its notation, is a training variable that no dataset card records. The field has established that text-e

The Authority Expectancy Effect in Multi-User Conflict

Model ReleasesDGX agent

arXiv:2608.08026v1 Announce Type: new Abstract: We investigate how social authority (SA) signals interact with severity-based prioritization in large language models, operationalizing each axis as a m

The Belief-Desire-Intention Ontology for modelling mental reality and agency

AgentsDGX agent

arXiv:2511.17162v2 Announce Type: replace Abstract: The Belief-Desire-Intention (BDI) model is a cornerstone for representing rational agency in artificial intelligence and cognitive sciences. Yet, it

The Capability Ladder: A Curriculum-Modernization Framework for Workforce Readiness in the AI Era

AgentsDGX agent

arXiv:2608.07779v1 Announce Type: new Abstract: Artificial intelligence is changing the task composition of computing work faster than curricula and training typically adapt. This is a curriculum-fram

The Cell Must Go On: Agar.io for Continual Reinforcement Learning

Model ReleasesDGX agent

arXiv:2505.18347v3 Announce Type: replace-cross Abstract: Continual reinforcement learning (RL) concerns agents that are expected to learn continually, rather than converge to a policy that is then fi

The Collaboration Gap: Exploration and Benchmarking of Open-World Agentic Cooperation

Model ReleasesDGX agent

arXiv:2511.02687v2 Announce Type: replace Abstract: The trajectory of AI development suggests that we will increasingly rely on agent-based systems powered by language models, composed of independentl

The Field Knows: Cross-Dimensional Geometry from Navigation to Black Holes

ResearchDGX agent

arXiv:2608.07566v1 Announce Type: new Abstract: We introduce a continuous metric field framework trained by a single causal contrastive loss. The framework encodes a scene into coefficients of a fixed

The Knowing-Saying Gap: When Probes See Errors that Confidence Misses

Model ReleasesDGX agent

arXiv:2608.07528v1 Announce Type: new Abstract: Linear probes detect corrupted context in language models with near-perfect accuracy, yet this does not translate into reliable failure prediction. The

The Politician, the Liar, and the Obedient Worker: Emerging Behavior of LLM Agents in Hierarchical Games

Model ReleasesDGX agent

arXiv:2608.09574v1 Announce Type: new Abstract: LLMs are rapidly embedding themselves into daily life: drafting our emails, managing our schedules, and making decisions on our behalf. As they move fro

The Scaffolding Matters More Than the Interface: A Controlled Comparison of MCP and CLI Tool Use Across Seven Agent Scaffoldings, Five Language Models, and One Software Task

Model ReleasesDGX agent

arXiv:2608.08654v1 Announce Type: new Abstract: How much an AI coding agent costs to run can depend more on the agent scaffolding that drives it than on the interface through which it reaches its tool

The Scaling Paradox in Human-AI Collaboration

SafetyDGX agent

arXiv:2608.00818v2 Announce Type: replace Abstract: The discovery of scaling laws has highlighted the extraordinary potential of AI systems with a striking empirical pattern: as AI systems scale, thei

The Theory of Strategic Evolution: Games with Endogenous Players and the Seven Laws of Strategic Replicators

SafetyDGX agent

arXiv:2512.07901v4 Announce Type: replace-cross Abstract: Von Neumann founded both game theory and the theory of self-reproducing automata, but the two programs never merged. Rational players do not c

The Watermark Shortcut: How Provenance Marking Sabotages Audio Deepfake Detection

ResearchDGX agent

arXiv:2606.23335v2 Announce Type: replace-cross Abstract: Provenance watermarking is increasingly treated as a safeguard for synthetic speech, whether built directly into speech-generation models such

Theory-Guided Deception Detection: A RAG-Based Artificial Intelligence Exploration

Model ReleasesDGX agent

arXiv:2608.08881v1 Announce Type: new Abstract: The current work developed seven Retrieval-Augmented Generation (RAG) models based on leading deception theories and compared how deception judgments we

Think Deep, Speak Once: Relit, A Recursive Latent Implicit Transformer Framework

Model ReleasesDGX agent

arXiv:2608.08113v1 Announce Type: new Abstract: Chain-of-Thought (CoT) prompting has become the dominant paradigm for eliciting reasoning in Large Language Models (LLMs), yet it creates substantial co

Thinking Hard, Not Smart: Reasoning Models Fail to Ration Test-Time Compute Across Questions

TutorialsDGX agent

arXiv:2608.07968v1 Announce Type: cross Abstract: Reasoning language models increasingly use test-time compute to improve performance, but existing evaluations typically study this compute one questio

Thought-Level Beam Search for Reasoning

ResearchDGX agent

arXiv:2608.08020v1 Announce Type: new Abstract: Test-time compute scaling is a primary driver of performance in large reasoning models (LRMs), but extreme inefficiency bounds current approaches, shift

Three Generations of Healthcare IT: From the Digital Record to the Computable Care Process

ApplicationsDGX agent

arXiv:2608.08806v1 Announce Type: new Abstract: Objective. Healthcare IT is usually organized by the technologies it adopts. We instead organize it by the unit of information a system makes computable

Three Necessary Principles for Self-Supervised Visual Representation Learning

SafetyDGX agent

arXiv:2608.08309v1 Announce Type: cross Abstract: We argue that learning visual representations without labels requires a training signal jointly complete across three non-overlapping objectives: sema

Time Present and Time Past: Benchmarking Large Language Models on Temporally Evolving Document Understanding

Model ReleasesDGX agent

arXiv:2608.08512v1 Announce Type: new Abstract: Evolving documents, such as laws, tax codes, and software documentation, are amended, replaced, and sometimes reverted over time, so a question has diff

TLDChoiceNet: Quantitatively Choosing a Transfer Learning Dataset

TutorialsDGX agent

arXiv:2608.09091v1 Announce Type: cross Abstract: Transfer learning is particularly useful in settings with limited training data, and within image classification it is common to transfer learn upon m

To Memorize or to Retrieve: Scaling the Interaction Between Pretraining and Retrieval

Model ReleasesDGX agent

arXiv:2604.00715v2 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) improves language model (LM) performance by providing relevant context at test time for knowledge-intensi

TokenPrint: A Calibrated Token-Space Fingerprint for Language-Model Provenance

ResearchDGX agent

arXiv:2608.08139v1 Announce Type: new Abstract: Establishing the provenance of a language model---including its base checkpoint and possible overlap in training distributions---is a governance challen

TomaMMU: A Comprehensive Multimodal Understanding Benchmark for Tomato Leaf Diseases

Model ReleasesDGX agent

arXiv:2608.08727v1 Announce Type: cross Abstract: To address this gap, we introduce TomaMMU, a large-scale Tomato leaf disease MultiModal Understanding dataset, alongside TomaBench, a benchmark for ev

TongGuOCR: A Layout-Aware and Token-Augmented OCR Framework for Chinese Historical Documents

Model ReleasesDGX agent

arXiv:2608.07917v1 Announce Type: new Abstract: Chinese historical documents preserve valuable cultural heritage, but many collections remain accessible only as scanned page images, preventing full-te

Tools to Explain Neural Networks for Power System Dynamics

ResearchDGX agent

arXiv:2608.08048v1 Announce Type: cross Abstract: This paper presents, for the first time in power systems literature to our knowledge, analytical tools to explain the training performance of machine

ToolUniverse: An open platform for democratizing AI scientists

SafetyDGX agent

arXiv:2509.23426v3 Announce Type: replace Abstract: AI scientists are emerging computational systems that serve as collaborative partners in discovery. These systems remain difficult to build because

ToolVision: Learning When and How to Use Visual Tools with Capability-Aligned Supervision

AgentsDGX agent

arXiv:2608.08907v1 Announce Type: cross Abstract: Thinking with images allows a multimodal model to compensate for limited perception by invoking visual tools through code. Yet the prevailing SFT-then

Topographic Constraints Shape Brain-Like Component Structure in Auditory Models

SafetyDGX agent

arXiv:2509.24039v2 Announce Type: replace-cross Abstract: If topography is a fundamental feature of the brain, it should influence both how neurons are arranged in space (i.e. explain brain maps) and

Toward CT-Equivalent Image Quality in Low-Dose Radiotherapy Planning: Conditional Diffusion-Based CBCT-to-CT Synthesis and the Impact of CBCT Input Representation

ResearchDGX agent

arXiv:2608.08919v1 Announce Type: cross Abstract: During standard radiotherapy planning, repeated CT acquisitions are often required for patient registration, verification, and adaptive planning, resu

Towards an Argumentative Foundation for Evaluative AI

ResearchDGX agent

arXiv:2608.07473v1 Announce Type: new Abstract: Evaluative AI (EAI) has been recently proposed as a way to support human decision-making, not by producing a single recommendation, but by presenting co

Towards Expert-level Medical AI for Real-time Video Consultations

Model ReleasesDGX agent

arXiv:2608.09861v1 Announce Type: new Abstract: Audio-visual interaction is the standard for patient-physician consultations, enabling natural communication and effective assessment of illness through

Towards Realistic Guarantees: A Probabilistic Certificate for SmoothLLM

SafetyDGX agent

arXiv:2511.18721v4 Announce Type: replace-cross Abstract: The SmoothLLM defense provides a certification guarantee against jailbreaking attacks, but it relies on a strict 'k-unstable' assumption that

Towards Researcher Agents for Knowledge-Graph Question Answering

Model ReleasesDGX agent

arXiv:2608.07700v1 Announce Type: new Abstract: Translating a natural-language question into a SPARQL query that can be executed against a large knowledge graph requires resolving lexical ambiguity, g

TRACE-Memory: Public-Conditioned Retrieval and Utility-Aware Evidence Admission for Personalized Generation

ResearchDGX agent

arXiv:2608.08446v1 Announce Type: new Abstract: Personalized generation systems retrieve user history by request--memory relevance and inject it into the model context. Yet relevant history may concer

TRACE: TRajectory Attribution for Automated Context Engineering

Model ReleasesDGX agent

arXiv:2608.09153v1 Announce Type: new Abstract: Production AI agents fail when their context sources -- system prompts, knowledge bases, tool descriptions, and procedural skills -- contain errors or g

Training Variable Long Sequences with Data-Centric Parallel

ResearchDGX agent

arXiv:2608.07524v1 Announce Type: new Abstract: Training deep learning models on variable long sequences poses significant computational challenges. Existing methods force a difficult trade-off betwee

Transformer-Based Neural Quantum Digital Twins for Many-Body Spectral Reconstruction and Adaptive Quantum-Annealing Schedule Design

ResearchDGX agent

arXiv:2505.15662v3 Announce Type: replace-cross Abstract: We introduce Transformer-based Neural Quantum Digital Twins (Tx-NQDTs) to reconstruct the low-energy spectral evolution of many-body quantum s

Transformer Circuits Can Realize Clustering Algorithms

TutorialsDGX agent

arXiv:2506.19125v2 Announce Type: replace-cross Abstract: Although transformers are most commonly optimized as statistical sequence models, it is unclear to what extent they can implement and learn ex

Transformer Explainer: Learning LLM Transformers with Interactive Visual Explanation and Experimentation

TutorialsDGX agent

arXiv:2408.04619v2 Announce Type: replace-cross Abstract: The Transformer architecture underpins modern large language models powering state-of-the-art text generation and AI applications. However, it

TransNRank: Towards Accurate Neoantigen Ranking with Transformer

Local AiDGX agent

arXiv:2608.01924v2 Announce Type: replace-cross Abstract: Personalized neoantigen prediction is challenging due to the scarcity of positive samples, the noise of the experimental data, the severe clas

TREAT: Evaluating Access to Formal Knowledge across Equivalent Mathematical Representations

Model ReleasesDGX agent

arXiv:2608.07540v1 Announce Type: new Abstract: AI systems increasingly operate between flexible input representations and formal objects used by downstream tools. A key challenge is recognizing when

TreeHop: Efficient Embedding-Level Query Rewriter

Model ReleasesDGX agent

arXiv:2504.20114v3 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) systems face significant challenges in multi-hop question answering (MHQA), where complex queries require

Triple Expert Learning from Noisy Labels for Semi-Supervised Vision Foundation Model Adaptation

SafetyDGX agent

arXiv:2608.09052v1 Announce Type: cross Abstract: Semi-supervised adaptation of vision foundation models (VFMs) commonly freezes the pretrained backbone and updates lightweight modules such as LoRA. H

TrustRoboReward: Preference-Ordered Isotonic Score Editing for Multi-Paradigm Robot Reward Models

Model ReleasesDGX agent

arXiv:2608.08491v1 Announce Type: new Abstract: Reward models are a bottleneck for reinforcement learning in embodied AI. Long-horizon robotic manipulation requires scalable vision feedback beyond han

TSPORec: Token Selection via Preference Optimization for LLM-Based Sequential Recommendation

ResearchDGX agent

arXiv:2608.09605v1 Announce Type: cross Abstract: Large Language Models (LLMs) have emerged as powerful tools for improving recommendation systems. The effectiveness of LLMs arises from their ability

Two-Layer Linear Auto-Regressive Models Estimate Latent States

Model ReleasesDGX agent

arXiv:2606.12691v2 Announce Type: replace-cross Abstract: Auto-regressive models have emerged as powerful tools for sequential data, from language to video. Understanding how and why these models lear

Two-Step MV-DeepONet: Probabilistic Operator Learning for Uncertainty Propagation Driven by Random Input Fields

ResearchDGX agent

arXiv:2608.09071v1 Announce Type: cross Abstract: Forward uncertainty propagation in complex physical systems can induce structured covariance across field-valued outputs. For a probabilistic surrogat

Ultraconstructive Model Theory via Bounded Adversarial Finite Structures

ApplicationsDGX agent

arXiv:2608.07534v1 Announce Type: cross Abstract: Ultraconstructive Model Theory (UCMT) replaces idealized satisfaction, at finite compu- tational scale, by bounded adversarial survival. A finite part

Unaccountable Delegation, Fading Skills: Mapping the Risks of Workplace AI Agents

SafetyDGX agent

arXiv:2608.08601v1 Announce Type: new Abstract: To anticipate socio-technical risks from AI agents, organizations need taxonomies to classify them. However, existing AI risk taxonomies focus on broad

Understanding Calibration and Truncation Error Propagation in Training-Free Low-Rank Compression for LLMs

Model ReleasesDGX agent

arXiv:2608.08506v1 Announce Type: new Abstract: Training-free low-rank compression frameworks have been gaining prominence for LLM compression given their effectiveness in reducing model parameter cou

Understanding Reasoning from Pretraining to Post-Training

SafetyDGX agent

arXiv:2607.16097v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) has become central to improving large language models (LLMs) on complex reasoning tasks, yet RL post-training is l

UniDFKD: A Unified Semantic Prior Framework for Architecture-Agnostic Data-Free Knowledge Distillation

TutorialsDGX agent

arXiv:2608.09287v1 Announce Type: cross Abstract: Data-Free Knowledge Distillation (DFKD) transfers knowledge from a pretrained teacher model to a compact student model by synthesizing semantically in

← Previous
1…1213141516…350
Next →