AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
Human
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,295 results
11 Aug 2026

SymboUQ: Symbolic Uncertainty Quantification for Spatial Reasoning in LLMs

ResearchDGX agent

arXiv:2608.00417v2 Announce Type: replace Abstract: Although large language models (LLMs) can produce fluent spatial reasoning traces, their intermediate relations may fail to support the final conclu

SymDiag: Explainable Diagnosis for LLM Reasoning via Neuro-Symbolic Verification

Local AiDGX agent

arXiv:2608.08786v1 Announce Type: new Abstract: Large language models (LLMs) increasingly serve as data-driven reasoners, yet their chains-of-thought (CoT) can be unfaithful even when final answers ar

SynVAR: Synergizing Spatial and Semantic Alignment in Visual Autoregressive Model

SafetyDGX agent

arXiv:2608.07948v1 Announce Type: new Abstract: VAR has gained widespread popularity due to its next-scale prediction paradigm. However, it faces substantial performance bottlenecks when handling comp

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Tabular Numeric Stretch Transformation

Model ReleasesDGX agent

arXiv:2608.09162v1 Announce Type: cross Abstract: Tabular data presents unique challenges for deep learning due to its heterogeneous nature, where numeric features exhibit diverse distributions, scale

TAMS: Task-Aware Multi-View Adaptive Streaming for Wireless Telerobotic Manipulation

ResearchDGX agent

arXiv:2608.09731v1 Announce Type: new Abstract: Wireless telerobotic manipulation relies on timely multi-view video feedback, but the available uplink bandwidth is often limited and dynamic. This pape

Targeted Counterfactual Fingerprinting for Black-Box LLM Ownership Verification

Local AiDGX agent

arXiv:2608.08195v1 Announce Type: cross Abstract: Large language models (LLMs) are high-value assets that can be derived through redeployment, fine-tuning, quantization, or further alignment. Because

Targeted Label-Flipping and Oversampling Attacks on Federated Conditional GANs

Local AiDGX agent

arXiv:2608.09314v1 Announce Type: new Abstract: In a federated learning setup for GANs, several adversarial attacks are possible. One such attack is label flipping, in which malicious clients delibera

Task-Adaptive 3D Cross-Field MRI Translation via Field-Conditioned Content-Style Pretraining

ResearchDGX agent

arXiv:2608.09264v1 Announce Type: new Abstract: Magnetic field strength is a major source of domain shift in magnetic resonance imaging (MRI), affecting signal-to-noise ratio, tissue contrast, spatial

Task-Oriented Formation Decision via Reinforcement Learning: Herding an Attacking Swarm

Model ReleasesDGX agent

arXiv:2608.09258v1 Announce Type: new Abstract: Multi-robot systems can accomplish tasks that are difficult for a single robot by organizing into task-specific formations. Different from existing stud

Task-to-Model Optimization for Enterprise LLM Coding Assistants: A Data-Driven Framework for Cost-Optimal Routing

Model ReleasesDGX agent

arXiv:2608.08528v1 Announce Type: new Abstract: Enterprise AI coding assistants incur substantial inference spend, and naive token-cost minimization often fails to reduce end-to-end cost once retries,

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability

Model ReleasesDGX agent

arXiv:2608.09538v1 Announce Type: cross Abstract: We introduce TCS-Bench, a benchmark for evaluating Large Language Models (LLMs) on research-level Theoretical Computer Science (TCS) proof generation.

TDMA Based Communications Control Co-Design for Cooperative Carrying: Delay Calibration and Sampling-Rate Optimization

ApplicationsDGX agent

arXiv:2608.09556v1 Announce Type: new Abstract: Multi robot teams performing cooperative transportation face a fundamental challenge: maintaining stable control while keeping communications efficient.

TeaMatch: Teachable Cross-Modal Representation Learning for 2D-3D Matching

ResearchDGX agent

arXiv:2608.09590v1 Announce Type: new Abstract: Learning reliable correspondences between images and point clouds is fundamental for 2D-3D matching. Despite recent progress in detection-free methods,

TelemetrySuffBench: Is Agent Telemetry Sufficient for Failure-Origin Diagnosis?

Model ReleasesDGX agent

arXiv:2608.07899v1 Announce Type: new Abstract: Agent systems increasingly expose execution traces, yet telemetry that reveals a failure may still be inadequate for identifying where that failure orig

TEMPER: Tensorized Efficient Manifold-constrained Parameterization for Expressive Residual Routing

Model ReleasesDGX agent

arXiv:2608.07851v1 Announce Type: cross Abstract: Residual connections rely on a static residual pathway, and are essential for training deep neural networks. Hyper-connections (HC) increase the expre

Temporal Generalization in fNIRS-Based Autism Classification: A Cross-Time-Window Transfer Benchmark

Model ReleasesDGX agent

arXiv:2608.07567v1 Announce Type: cross Abstract: Functional near-infrared spectroscopy (fNIRS) is a promising modality for autism spectrum disorder (ASD) classification, yet existing approaches assum

Temporal Misgrounding in Legal RAG: A Versioned-Corpus Benchmark for French Tax Law

Model ReleasesDGX agent

arXiv:2608.09393v1 Announce Type: cross Abstract: We identify and quantify temporal misgrounding: the systematic retrieval and citation of the currently in-force version of a legal article when the ap

Temporal Sepsis Modeling: a Relational and Explainable-by-Design Framework

ResearchDGX agent

arXiv:2601.21747v4 Announce Type: replace-cross Abstract: Sepsis remains one of the most complex and heterogeneous syndromes in intensive care. While deep learning models achieve competitive performan

Test-Time Augmentation for LLMs: When Input Diversity Beats Output Diversity at Matched Compute

ResearchDGX agent

arXiv:2608.09351v1 Announce Type: cross Abstract: Test-time scaling improves LLM accuracy but multiplies inference cost, making the accuracy gained per unit of compute the metric that matters in deplo

Test-time Generalization for Physics through Neural Operator Splitting

Model ReleasesDGX agent

arXiv:2602.00884v2 Announce Type: replace Abstract: Neural operators have shown promise in learning solution maps of partial differential equations (PDEs), but they often struggle to generalize when t

Test-Time Prototype Adaptation for Open-Vocabulary Semantic Segmentation

Model ReleasesDGX agent

arXiv:2608.08290v1 Announce Type: new Abstract: Open-vocabulary semantic segmentation (OVSS) repurposes a pretrained CLIP encoder for dense prediction without additional labeled supervision. Existing

Test-Time Scaling for CAD Generation via Verifier-Free Consensus Selection

ResearchDGX agent

arXiv:2608.09706v1 Announce Type: cross Abstract: Large language models can write parametric CAD programs from a natural-language description (text-to-CAD generation), but a single sample is often wro

Testing Hypotheses from the Social Approval Theory of Online Hate: An Analysis of 110 Million Messages from Parler

ResearchDGX agent

arXiv:2507.10810v3 Announce Type: replace Abstract: We examined how social approval motivates online hate via the social approval theory, which argues social approval signals on hate messages predict

Tether-Inertial Localization for Planetary Drones

Local AiDGX agent

arXiv:2608.09515v1 Announce Type: new Abstract: Recent developments in planetary exploration have shown the potential of Unmanned Aerial Vehicles (UAVs), such as the Ingenuity helicopter that provided

Tevatron-Elastic: A Unified Abstraction for Training Elastic Retrievers and Rerankers

Model ReleasesDGX agent

arXiv:2608.08809v1 Announce Type: new Abstract: A single model scale challenges the flexibility of a production retrieval system: some settings need it faster, others need a smaller index, and the rig

TeXFix-Bench: An Empirically Grounded Multi-Format Benchmark for LLM-Based Document Source Repair

Model ReleasesDGX agent

arXiv:2608.07617v1 Announce Type: new Abstract: Scientific and technical writing depends on markup sources that must compile: LaTeX, Typst, and Markdown pipelines fail on missing delimiters, mismatche

TGIF: Text-Guided Layer Fusion Mitigates Hallucination in Multimodal LLMs

ResearchDGX agent

arXiv:2601.03100v3 Announce Type: replace-cross Abstract: Multimodal large language models (MLLMs) typically rely on a single late-layer feature from a frozen vision encoder, leaving the encoder's ric

The Anatomy of a Prompt Injection: A Component Model for Structured Analysis

SafetyDGX agent

arXiv:2608.07808v1 Announce Type: cross Abstract: Four years after prompt injection was first identified in 2022, attacks are still predominantly documented as verbatim strings rather than structured

The Announcement Carries the Cue: Markup, Boundaries, and the Notation of Pre-Training Corpora

ResearchDGX agent

arXiv:2608.09093v1 Announce Type: cross Abstract: How a document's arrangement is written down, its notation, is a training variable that no dataset card records. The field has established that text-e

The Authority Expectancy Effect in Multi-User Conflict

Model ReleasesDGX agent

arXiv:2608.08026v1 Announce Type: new Abstract: We investigate how social authority (SA) signals interact with severity-based prioritization in large language models, operationalizing each axis as a m

The Belief-Desire-Intention Ontology for modelling mental reality and agency

AgentsDGX agent

arXiv:2511.17162v2 Announce Type: replace Abstract: The Belief-Desire-Intention (BDI) model is a cornerstone for representing rational agency in artificial intelligence and cognitive sciences. Yet, it

The Capability Ladder: A Curriculum-Modernization Framework for Workforce Readiness in the AI Era

AgentsDGX agent

arXiv:2608.07779v1 Announce Type: new Abstract: Artificial intelligence is changing the task composition of computing work faster than curricula and training typically adapt. This is a curriculum-fram

The Cell Must Go On: Agar.io for Continual Reinforcement Learning

Model ReleasesDGX agent

arXiv:2505.18347v3 Announce Type: replace-cross Abstract: Continual reinforcement learning (RL) concerns agents that are expected to learn continually, rather than converge to a policy that is then fi

The Collaboration Gap: Exploration and Benchmarking of Open-World Agentic Cooperation

Model ReleasesDGX agent

arXiv:2511.02687v2 Announce Type: replace Abstract: The trajectory of AI development suggests that we will increasingly rely on agent-based systems powered by language models, composed of independentl

The Cost of Adaptivity: Matching Lower Bounds Across Learning Problems

Model ReleasesDGX agent

arXiv:2608.08826v1 Announce Type: new Abstract: Adaptive procedures must work without nuisance information an oracle may use, such as a gradient scale or smoothness index, and robust procedures may ha

The Evolution of Mixture-of-Experts Architectures in Large Language Models: Routing, Topology, Load Balancing, and Expert Parallelism

Model ReleasesDGX agent

arXiv:2608.08650v1 Announce Type: new Abstract: Mixture-of-Experts models increase parameter capacity while keeping the computation activated by each token bounded, but their architectural evolution c

The Field Knows: Cross-Dimensional Geometry from Navigation to Black Holes

ResearchDGX agent

arXiv:2608.07566v1 Announce Type: new Abstract: We introduce a continuous metric field framework trained by a single causal contrastive loss. The framework encodes a scene into coefficients of a fixed

The Judge Knows When It Knows: Calibrated Abstention for LLM-Based A/B-Test Prediction

Model ReleasesDGX agent

arXiv:2608.07517v1 Announce Type: cross Abstract: Can a multimodal LLM predict which version of a web page will win a real A/B test from screenshots alone? We report the most complete answer we are aw

The Knowing-Saying Gap: When Probes See Errors that Confidence Misses

Model ReleasesDGX agent

arXiv:2608.07528v1 Announce Type: new Abstract: Linear probes detect corrupted context in language models with near-perfect accuracy, yet this does not translate into reliable failure prediction. The

The Neural Division of Labor: Biologically-Inspired Modular Architectures for Robust Neuromorphic Computing

ResearchDGX agent

arXiv:2608.08317v1 Announce Type: new Abstract: Biological neural systems achieve high efficiency and robustness through compartmentalized architectures. In contrast, modern artificial neural networks

The No-Meaning Falsity: The Structural Impossibility of the Arbitrary Sign in Classical Arabic

ResearchDGX agent

arXiv:2608.07737v1 Announce Type: new Abstract: This paper investigates whether the postmodern claim of unrestricted semantic indeterminacy, and its foundational Saussurean axiom of the arbitrary sign

The Politician, the Liar, and the Obedient Worker: Emerging Behavior of LLM Agents in Hierarchical Games

Model ReleasesDGX agent

arXiv:2608.09574v1 Announce Type: new Abstract: LLMs are rapidly embedding themselves into daily life: drafting our emails, managing our schedules, and making decisions on our behalf. As they move fro

The Replay Gap: Static Evaluation of Model Switching in LLM Agents Scores the Wrong World

AgentsDGX agent

arXiv:2608.08239v1 Announce Type: cross Abstract: LLM routers promise efficiency by matching each request to the cheapest adequate model, and are increasingly applied per step inside multi-step agents

The Sample Complexity of Policy Learning with Mu-Resets

SafetyDGX agent

arXiv:2608.07772v1 Announce Type: new Abstract: We study policy-based reinforcement learning under the mu-resets interaction protocol of Kakade and Langford [KL02]. This interaction protocol enables t

The Scaffolding Matters More Than the Interface: A Controlled Comparison of MCP and CLI Tool Use Across Seven Agent Scaffoldings, Five Language Models, and One Software Task

Model ReleasesDGX agent

arXiv:2608.08654v1 Announce Type: new Abstract: How much an AI coding agent costs to run can depend more on the agent scaffolding that drives it than on the interface through which it reaches its tool

The Scaling Paradox in Human-AI Collaboration

SafetyDGX agent

arXiv:2608.00818v2 Announce Type: replace Abstract: The discovery of scaling laws has highlighted the extraordinary potential of AI systems with a striking empirical pattern: as AI systems scale, thei

The Spectral Neuron

Local AiDGX agent

arXiv:2608.08003v1 Announce Type: cross Abstract: As machine learned models increase in complexity and expressive power, features of simpler models, such as interpretability and control over the shape

The Theory of Strategic Evolution: Games with Endogenous Players and the Seven Laws of Strategic Replicators

SafetyDGX agent

arXiv:2512.07901v4 Announce Type: replace-cross Abstract: Von Neumann founded both game theory and the theory of self-reproducing automata, but the two programs never merged. Rational players do not c

The Transparency Trap: How AI Disclaimers Create Overconfidence in High-Stakes Decisions

SafetyDGX agent

arXiv:2608.07493v1 Announce Type: cross Abstract: Current AI disclaimers often fail to function as intended due to warning habituation and a transparency paradox. As AI-generated information becomes p

The Voiceprint Fallacy: Why Voices Are Not Unique Biometric Imprints

SafetyDGX agent

arXiv:2608.07980v1 Announce Type: cross Abstract: In recent years, the term voiceprint has regained attention, particularly in technological applications and policy-making contexts, often carrying the

The Watermark Shortcut: How Provenance Marking Sabotages Audio Deepfake Detection

ResearchDGX agent

arXiv:2606.23335v2 Announce Type: replace-cross Abstract: Provenance watermarking is increasingly treated as a safeguard for synthetic speech, whether built directly into speech-generation models such

Theory-Guided Deception Detection: A RAG-Based Artificial Intelligence Exploration

Model ReleasesDGX agent

arXiv:2608.08881v1 Announce Type: new Abstract: The current work developed seven Retrieval-Augmented Generation (RAG) models based on leading deception theories and compared how deception judgments we

Think Deep, Speak Once: Relit, A Recursive Latent Implicit Transformer Framework

Model ReleasesDGX agent

arXiv:2608.08113v1 Announce Type: new Abstract: Chain-of-Thought (CoT) prompting has become the dominant paradigm for eliciting reasoning in Large Language Models (LLMs), yet it creates substantial co

Thinking Hard, Not Smart: Reasoning Models Fail to Ration Test-Time Compute Across Questions

TutorialsDGX agent

arXiv:2608.07968v1 Announce Type: cross Abstract: Reasoning language models increasingly use test-time compute to improve performance, but existing evaluations typically study this compute one questio

Thinking Is Not Telling: Information Disclosure in User-Service LLM Agents

AgentsDGX agent

arXiv:2602.07796v2 Announce Type: replace Abstract: User-engaged LLM agents increasingly operate in service scenarios where task success depends on coordination between the agent, the user, and a stat

Thinking vs. NoThinking: Towards Interpreting Reasoning Mechanisms of Large Language Models via Sparse Autoencoders

Model ReleasesDGX agent

arXiv:2608.08168v1 Announce Type: new Abstract: While Large Language Models (LLMs) employing Chain-of-Thought (CoT) exhibit superior reasoning capabilities, the neural mechanisms distinguishing this e

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning

AgentsDGX agent

arXiv:2608.09682v1 Announce Type: new Abstract: Tool-augmented vision-language models increasingly 'think with images': they call crop, zoom, or code tools and reason over the returned pixels. However

Thought-Level Beam Search for Reasoning

ResearchDGX agent

arXiv:2608.08020v1 Announce Type: new Abstract: Test-time compute scaling is a primary driver of performance in large reasoning models (LRMs), but extreme inefficiency bounds current approaches, shift

Three Generations of Healthcare IT: From the Digital Record to the Computable Care Process

ApplicationsDGX agent

arXiv:2608.08806v1 Announce Type: new Abstract: Objective. Healthcare IT is usually organized by the technologies it adopts. We instead organize it by the unit of information a system makes computable

Three Necessary Principles for Self-Supervised Visual Representation Learning

SafetyDGX agent

arXiv:2608.08309v1 Announce Type: cross Abstract: We argue that learning visual representations without labels requires a training signal jointly complete across three non-overlapping objectives: sema

← Previous
1…3637383940…989
Next →