AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,435 results
Model Releases

The Few-shot Dilemma: Over-prompting Large Language Models

DGX agent

arXiv:2509.13196v2 Announce Type: replace Abstract: Over-prompting, a phenomenon where excessive examples in prompts lead to diminished performance in Large Language Models (LLMs), challenges the conv

model-releasesarxiv-cs-cl
28 Jul 2026
Model Releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

The Gate Always Closes: On Injecting Auxiliary Signals into Frozen Vision-Language Models

DGX agent

arXiv:2607.23335v1 Announce Type: new Abstract: Auxiliary signal pathways in VLMs are routinely fitted with learnable gates so the optimiser can decide how much of the signal to admit. We find that th

model-releasesarxiv-cs-cv
28 Jul 2026
Model Releases

Through the Bottleneck: How Multi-head Latent Attention Separates Content from Position in Language Models

DGX agent

arXiv:2607.23054v1 Announce Type: cross Abstract: Multi-head Latent Attention (MLA), introduced in DeepSeek-V2, compresses key-value pairs through a shared low-rank bottleneck (cKV), achieving 81% KV-

model-releasesarxiv-cs-ai
28 Jul 2026
Research

Towards Cultural Bridge by Bahnaric-Vietnamese Translation Using Transfer Learning of Sequence-To-Sequence Pre-training Language Model

DGX agent

arXiv:2505.11421v2 Announce Type: replace Abstract: This work explores the journey towards achieving Bahnaric-Vietnamese translation for the sake of culturally bridging the two ethnic groups in Vietna

researcharxiv-cs-cl
28 Jul 2026
Local Ai

Training Language Models to Cooperate with Inference-Time Controllers

DGX agent

arXiv:2607.23771v1 Announce Type: new Abstract: Large language model (LLM) performance increasingly depends not only on the base model, but also on the inference-time controller used to organize reaso

local-aiarxiv-cs-ai
28 Jul 2026
Model Releases

Visible-Light Imaging Diagnosis of Neutral Particle Emission Tomography in the Tokamak Divertor: An Efficient Transformer-based Surrogate Model

DGX agent

arXiv:2607.22704v1 Announce Type: cross Abstract: Nuclear fusion has made significant progress in recent years and is expected to become one of the most important pathways to addressing global energy

model-releasesarxiv-cs-ai
28 Jul 2026
Agents

Action-Conditioned World Model for Goal Plane Probe Guidance in Robotic Ultrasound

DGX agent

arXiv:2607.21918v1 Announce Type: new Abstract: We present an action-conditioned world model framework for goal plane probe guidance in robotic ultrasound, with a focus on neck ultrasound scanning. Au

agentsarxiv-cs-ro
27 Jul 2026
Research

From Isolated Tasks to Structured Capabilities: A Multilayer Taxonomy for Large Language Models

DGX agent

arXiv:2607.22182v1 Announce Type: new Abstract: Large language model (LLM) evaluation spans diverse tasks and benchmarks, yet evidence remains organized around tasks rather than the capabilities they

researcharxiv-cs-cl
27 Jul 2026
Research

Joint Lossless Compression and Steganography for Medical Images via Large Language Models

DGX agent

arXiv:2508.01782v4 Announce Type: replace-cross Abstract: Recently, large language models (LLMs) have driven promising progress in lossless image compression. However, directly adopting existing parad

researcharxiv-cs-cv
27 Jul 2026
Research

Measuring the Dependency Gap: Diagnosing Inter-Column Fidelity in Tabular Generative Models

DGX agent

arXiv:2607.21636v1 Announce Type: new Abstract: Synthetic tabular data is valued for preserving not only each column's marginal distribution but the dependencies between columns -- structure that carr

researcharxiv-cs-lg
27 Jul 2026
Research

Automatic knot selection in smooth additive models

DGX agent

arXiv:2607.21083v1 Announce Type: cross Abstract: B-spline regression constitutes a widely used framework for nonparametric modeling. The performance of this methodology depends on specifying the numb

researcharxiv-cs-lg
24 Jul 2026
Safety

Belief Propagation in LLM World Models: Measuring Strategic Information Bias with Prediction Markets

DGX agent

arXiv:2607.20441v1 Announce Type: new Abstract: Every information ecosystem produces beliefs that shape strategic decisions. Both human analysts and AI systems inherit the blind spots of their informa

safetyarxiv-cs-cl
24 Jul 2026
Model Releases

Do Pathology Vision-Language Models Truly See Pathology?

DGX agent

arXiv:2607.21065v1 Announce Type: new Abstract: Pathology vision-language models (VLMs) have recently progressed rapidly and are commonly evaluated by answer accuracy on pathology VQA benchmarks. Howe

model-releasesarxiv-cs-cv
24 Jul 2026
Agents

MiniCache: Reusable Program Caching with Small Model Interfaces for Efficient LLM Inference

DGX agent

arXiv:2607.20507v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for program-aided reasoning, agentic decision making, and structured task execution, but these applic

agentsarxiv-cs-ai
24 Jul 2026
Agents

Streaming Multi-Agent Autoregressive Diffusion Model with World State Registers

DGX agent

arXiv:2607.21594v1 Announce Type: new Abstract: Multi-agent interactive world models should not only generate consistent observations, but also maintain world states that persist across agents and evo

agentsarxiv-cs-cv
24 Jul 2026
Safety

X^3-OPD: Distilling Reasoning into Large Audio-Language Models via On-Policy Alignment

DGX agent

arXiv:2607.21550v1 Announce Type: new Abstract: While large audio-language models have achieved remarkable progress in auditory perception, they still lag behind text-based large language models in de

safetyarxiv-cs-lg
24 Jul 2026
Model Releases

A convergence result of a continuous model of deep learning via a L{}ojasiewicz--Simon inequality

DGX agent

arXiv:2311.15365v3 Announce Type: replace Abstract: We study an idealized training process for deep neural networks in a continuous-depth, mean-field model in which each layer is parameterized by a pr

model-releasesarxiv-cs-lg
23 Jul 2026
Applications

Beyond Tracking or Shortcut: Composition-Bounded Predictive States in Poker Autoregressive Models

DGX agent

arXiv:2607.19369v1 Announce Type: new Abstract: Hidden-state probes often recover latent labels in imperfect-information sequence models, but this alone does not establish that a model maintains a pos

applicationsarxiv-cs-ai
23 Jul 2026
Model Releases

D2VBench: Benchmarking Large Language Models with Value Dilemmas in Daily Scenarios

DGX agent

arXiv:2607.19834v1 Announce Type: new Abstract: With the wide application of large language models (LLMs) in real-world scenarios, the value implication of their outputs is crucial. However, existing

model-releasesarxiv-cs-cl
23 Jul 2026
Research

Exposure is Optional: Learning Unlike Coordination in Language Models

DGX agent

arXiv:2607.20251v1 Announce Type: new Abstract: Coordination, a fundamental linguistic structure, remains a subject of intense debate, and its exact nature continues to elude theoretical linguistics.

researcharxiv-cs-cl
23 Jul 2026
Model Releases

FilmWorld: Agentic Novel-to-Film Generation through Dynamic Cinematic World Modeling

DGX agent

arXiv:2607.19038v1 Announce Type: new Abstract: Translating novels into films poses a grand challenge for generative artificial intelligence, requiring conversion of abstract literary prose into long-

model-releasesarxiv-cs-cv
23 Jul 2026
Tutorials

Interval and fuzzy physics-augmented neural networks (iPANN and fPANN) for uncertainty quantification and propagation in constitutive modeling

DGX agent

arXiv:2607.20339v1 Announce Type: new Abstract: Constitutive modeling under uncertainty remains a central challenge for reliable mechanics simulations, particularly when the available stress-deformati

tutorialsarxiv-cs-lg
23 Jul 2026
Model Releases

Reinforcement Learning for Large Language Model Selective Evidence Adoption from Contaminated Retrieval Results

DGX agent

arXiv:2607.20090v1 Announce Type: cross Abstract: Retrieval-augmented large language models frequently face contexts that interleave useful evidence with misleading statements or instruction-like cont

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

RoboInter1.5: A Holistic Intermediate Representation Suite for Embodied World Modeling and Robotic Manipulation

DGX agent

arXiv:2607.18709v2 Announce Type: replace Abstract: Existing robot datasets remain expensive to curate, embodiment-specific, and insufficiently annotated with the fine-grained structure required for g

model-releasesarxiv-cs-ro
23 Jul 2026
Tutorials

Scaling Laws for Hypernetwork-Based Knowledge Injection in Large Language Models

DGX agent

arXiv:2607.19604v1 Announce Type: new Abstract: Injecting factual knowledge into large language models (LLMs) reliably and at scale remains an open challenge. Hypernetworks provide a promising solutio

tutorialsarxiv-cs-cl
23 Jul 2026
Model Releases

SLAI T-Rex: Full-Parameter Post-training of the DeepSeek-V4 Family on Ascend SuperPOD

DGX agent

arXiv:2607.20145v1 Announce Type: cross Abstract: Full-parameter post-training of trillion-parameter-scale MoE models introduces substantial system-level challenges for large-scale distributed trainin

model-releasesarxiv-cs-ai
23 Jul 2026
Safety

Stress Testing Concept Erasure with Large Language Model Agents

DGX agent

arXiv:2607.17890v2 Announce Type: replace Abstract: Concept erasure aims to remove semantic concepts from a trained generative model and is increasingly important for responsible AI deployment. Howeve

safetyarxiv-cs-ai
23 Jul 2026
Applications

AI-Augmented Adaptive Digital Twin Modeling for Brain Tumor Evolution Prediction and Treatment Scheduling

DGX agent

arXiv:2607.13877v1 Announce Type: new Abstract: Brain tumor progression exhibits spatially heterogeneous growth, patient-specific treatment response, and complex interactions with surrounding anatomy,

applicationsarxiv-cs-lg
16 Jul 2026
Research

Discrete Diffusion Models: A Unified Framework from Tokenization to Generation

DGX agent

arXiv:2607.13431v1 Announce Type: cross Abstract: Discrete denoising diffusion models (DDMs) have recently emerged as a compelling alternative to autoregressive (AR) modeling for discrete data, offeri

researcharxiv-cs-ai
16 Jul 2026
Model Releases

Look Again Before You Abstain:Budgeted Conformal Evidence Acquisition for Reliable Vision-Language Model

DGX agent

arXiv:2606.16667v2 Announce Type: replace Abstract: Large vision-language models (LVLMs) hallucinate: they assert visual details that the image does not support. A principled remedy is selective predi

model-releasesarxiv-cs-cv
16 Jul 2026
Safety

Protective Capacity Hallucination: When Large Language Models Claim Nonexistent Capabilities

DGX agent

arXiv:2607.13596v1 Announce Type: cross Abstract: When cast as the protector of a vulnerable user yet given no explicit capability boundary, a large language model (LLM) may respond not by acknowledgi

safetyarxiv-cs-ai
16 Jul 2026
Research

Tactile Modality Fusion for Vision-Language-Action Models

DGX agent

arXiv:2603.14604v2 Announce Type: replace-cross Abstract: We propose TacFiLM, a lightweight modality-fusion approach that integrates visual-tactile signals into vision-language-action (VLA) models. Wh

researcharxiv-cs-cv
16 Jul 2026
Model Releases

The Dynamic Verifiable Multi-Agent Human Agentic Loyalty Loop (DVM-HALL) Model and the Net Human-Agent Score (NHAS) in Autonomous Commerce

DGX agent

arXiv:2607.13998v1 Announce Type: cross Abstract: The rapid proliferation of Agentic Artificial Intelligence fundamentally disrupts traditional customer loyalty paradigms. As AI evolves from passive r

model-releasesarxiv-cs-ai
16 Jul 2026
Tutorials

Action-Aware Generative Sequence Modeling for Short Video Recommendation

DGX agent

arXiv:2604.25834v2 Announce Type: replace Abstract: With the rapid development of the Internet, users have increasingly higher expectations for the recommendation accuracy of online content consumptio

tutorialsarxiv-cs-ai
15 Jul 2026
Research

Adaptive Compute in Latent World Models: When Depth Helps, Hurts, or Doesn't Matter

DGX agent

arXiv:2607.10203v2 Announce Type: replace-cross Abstract: Adaptive-compute world models -- early-exit or mixture-of-depths predictors that spend variable depth per step -- assume depth buys better pre

researcharxiv-cs-ai
15 Jul 2026
Tutorials

DM-KG: A Novel Method for Boosting Spatial Cognition of Vision-Language Models in Street View Imagery

DGX agent

arXiv:2607.12319v1 Announce Type: new Abstract: As vision-language models (VLMs) are increasingly deployed in geospatial question answering and visual scene understanding, improving their spatial cogn

tutorialsarxiv-cs-cv
15 Jul 2026
Model Releases

Do LLMs Know What They Know? Measuring Metacognitive Efficiency with Signal Detection Theory

DGX agent

arXiv:2603.25112v2 Announce Type: replace-cross Abstract: Standard evaluation of LLM confidence relies on calibration metrics (ECE, Brier score) that conflate how much a model knows (Type-1 accuracy)

model-releasesarxiv-cs-ai
15 Jul 2026
Research

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models

DGX agent

arXiv:2509.22415v3 Announce Type: replace-cross Abstract: Multimodal large language models (MLLMs) have achieved strong vision-language performance, yet their token-level visual evidence remains diffi

researcharxiv-cs-ai
15 Jul 2026
Research

Gaussian Mixture Modeling for Event-Aware Visual Allocation in Long Video Understanding

DGX agent

arXiv:2607.12557v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) face significant challenges in long video understanding due to the excessive computational cost and information los

researcharxiv-cs-cv
15 Jul 2026
Research

Modeling Story Expectations: A Generative Framework using LLMs

DGX agent

arXiv:2412.15239v4 Announce Type: replace-cross Abstract: Consumers' engagement with stories is shaped by their expectations about what will happen next, yet modeling these forward-looking beliefs ove

researcharxiv-cs-ai
15 Jul 2026
Research

Robustness of Deep Learning Models for PV Power Forecasting under NWP Forecast Errors: A Spatiotemporal and Physically Interpretable Analysis

DGX agent

arXiv:2607.12954v1 Announce Type: cross Abstract: Engineering use of AI forecasting models requires not only high nominal accuracy but also predictable behavior under uncertain inputs. In photovoltaic

researcharxiv-cs-lg
15 Jul 2026
Local Ai

Self-Consistent Flow: Unifying Velocity and Endpoint Prediction for Rectified Flow Models

DGX agent

arXiv:2607.12171v1 Announce Type: cross Abstract: In rectified-flow-based generative models, the neural network can be trained to predict two different targets, such as the instantaneous velocity or t

local-aiarxiv-cs-ai
15 Jul 2026
Research

Spectral Diffusion Processes

DGX agent

arXiv:2209.14125v3 Announce Type: replace-cross Abstract: Diffusion models have proven to be a flexible and effective framework for modelling probability distributions on finite-dimensional spaces. Ho

researcharxiv-cs-lg
15 Jul 2026
Research

The Seriality Gap in Video Diffusion Models

DGX agent

arXiv:2607.13031v1 Announce Type: cross Abstract: When one ball strikes another, then another, video models should predict the consequences of each bounce. In controlled experiments on multi-ball hard

researcharxiv-cs-cv
15 Jul 2026
Research

Visual Access Boundaries in Vision-Language Model Reasoning

DGX agent

arXiv:2607.12815v1 Announce Type: new Abstract: Chain-of-Thought (CoT) prompting is widely used as a test-time scaling strategy for Vision-Language Models (VLMs), but it remains unclear what is extend

researcharxiv-cs-ai
15 Jul 2026
Safety

Persona Cartography: Charting Language Model Personality Traits in Weight Space

DGX agent

arXiv:2607.07916v1 Announce Type: new Abstract: Large language models exhibit recurring behavioural patterns -- personas -- that shape generalisation and safety, but we lack reliable tools for decompo

safetyarxiv-cs-ai
10 Jul 2026
Research

Rethinking LLM-as-a-Judge: Representation-as-a-Judge with Small Language Models via Semantic Capacity Asymmetry

DGX agent

arXiv:2601.22588v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are widely used as reference-free evaluators via prompting, but this 'LLM-as-a-Judge' paradigm is costly, opaque,

researcharxiv-cs-ai
10 Jul 2026
Model Releases

SQuaD-SQL: Efficient Text-to-SQL with Small Language Models via LLM-Guided Knowledge Distillation

DGX agent

arXiv:2607.08161v1 Announce Type: new Abstract: Text-to-SQL is a fundamental task in natural language processing that enables users to interact with structured databases using natural language. While

model-releasesarxiv-cs-cl
10 Jul 2026
← Previous
1…100101102103104…1030
Next →