AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
48,975 results
Model Releases

WorldCycle: Self-Verifiable Reinforcement Learning for Long-Horizon Video World Models

DGX agent

arXiv:2608.04964v1 Announce Type: new Abstract: Interactive video world models are essential for long-horizon planning and exploration, yet they suffer from compounding errors. Post-training methods s

model-releasesarxiv-cs-ai
6 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

CrossScope: A Role-Asymmetric World Model for Joint Dual-Scope Surgical Video Prediction

DGX agent

arXiv:2608.03211v1 Announce Type: new Abstract: Visual world models typically learn future dynamics from a single observation stream, limiting their ability to model cooperative systems with multiple

model-releasesarxiv-cs-cv
5 Aug 2026
Applications

Modeling Scientific Experiment Scenes: Dataset and Model

DGX agent

arXiv:2608.02892v1 Announce Type: new Abstract: Scene Graph Generation (SGG) is fundamental to structured visual understanding, yet existing benchmarks focus mainly on daily life images and overlook s

applicationsarxiv-cs-cv
5 Aug 2026
Model Releases

Quantifying Hallucinations in Language Language Models on Medical Textbooks

DGX agent

arXiv:2603.09986v3 Announce Type: replace-cross Abstract: Hallucinations, the tendency for large language models to provide responses with factually incorrect and unsupported claims, is a serious prob

model-releasesarxiv-cs-ai
5 Aug 2026
Research

A Comprehensive FP8 Training Recipe for Reasoning-Enhanced Language Models

DGX agent

arXiv:2509.22536v5 Announce Type: replace Abstract: The immense computational cost of training Large Language Models (LLMs) presents a major barrier to innovation. While FP8 training offers a promisin

researcharxiv-cs-cl
4 Aug 2026
Model Releases

Can You Trust the Confidence? ConfBench for Vision-Language Models on Document Extraction

DGX agent

arXiv:2608.01792v1 Announce Type: cross Abstract: Intelligent document processing (IDP) with vision-language models (VLMs) hinges on confidence scores trustworthy enough to route extractions between a

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Generative AI and Foundation Models in Medical Image

DGX agent

arXiv:2608.01686v1 Announce Type: new Abstract: In recent years, generative AI has attracted significant public attention, and its use has been rapidly expanding across a wide range of domains. From c

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

Measuring in-context algorithmic reasoning in language models against an exact Bayes-optimal standard

DGX agent

arXiv:2608.01575v1 Announce Type: new Abstract: Whether large language models perform genuine algorithmic reasoning or mere pattern completion is hard to test, because most benchmarks lack a ground tr

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

REFLEX: Rethinking MoE Inference as Refinement-Aware Compute Allocation in Diffusion Language Models

DGX agent

arXiv:2608.01784v1 Announce Type: cross Abstract: Mixture-of-experts (MoE) models increase parameter capacity by activating only a small subset of experts for each token. This conditional-computation

model-releasesarxiv-cs-cl
4 Aug 2026
Local Ai

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning

DGX agent

arXiv:2608.01556v1 Announce Type: new Abstract: Large language models are increasingly aligned to human preferences via reward modeling, but user preference data are sensitive and often cannot be cent

local-aiarxiv-cs-lg
4 Aug 2026
Model Releases

Round-Trip Consistency: Bidirectional Diffusion Models Can Predict Their Own Rollout Errors

DGX agent

arXiv:2608.00675v1 Announce Type: cross Abstract: Autoregressive models accumulate error over long rollouts, yet at deployment there is no ground truth to measure it against. We train a single conditi

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models

DGX agent

arXiv:2608.01899v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) perform well on commonsense reasoning tasks but struggle with visual spatial reasoning. Most existing solutions introduc

model-releasesarxiv-cs-cl
4 Aug 2026
Applications

WHALE: A Scalable Unified Model for Recommendation with Wukong-HSTU Architecture

DGX agent

arXiv:2607.17017v2 Announce Type: replace-cross Abstract: As scalability becomes increasingly important in recommendation modeling, recent architectures have advanced the modeling of two broad sources

applicationsarxiv-cs-lg
4 Aug 2026
Agents

Auto-JEPA: A Latent World Model of Continuous Intent for End-to-End Autonomous Driving

DGX agent

arXiv:2607.29031v1 Announce Type: cross Abstract: Existing autonomous-driving world models typically perform dense prediction of future videos, occupancy states, BEV representations, or agent motion.

agentsarxiv-cs-ai
3 Aug 2026
Model Releases

Safe Vision Language Action Models via Barrier Enhanced Flow Matching

DGX agent

arXiv:2607.29569v1 Announce Type: new Abstract: This article presents a modular inference framework that integrates Flow Matching generative models with formal Control Barrier Function (CBF) safety gu

model-releasesarxiv-cs-ro
3 Aug 2026
Research

TORUS: A Test of Rendering-Understanding Self-Coherence for Unified Audio Models

DGX agent

arXiv:2607.28896v1 Announce Type: cross Abstract: Unified audio models capable of audio understanding, audio generation and, increasingly, audio editing are proliferating rapidly. Yet a basic question

researcharxiv-cs-ai
3 Aug 2026
Model Releases

A Physics-Informed Framework for PID Tuning of Chemical Processes Using Large Language Model Agents

DGX agent

arXiv:2607.26594v1 Announce Type: cross Abstract: PID tuning for chemical processes commonly relies on identified process models, whereas plant engineers often retune loops iteratively by observing re

model-releasesarxiv-cs-ai
31 Jul 2026
Model Releases

DS@GT ARC at ImageCLEFmedical 2026: Architectural Diversity for Concept Detection and Foundation-Model Scaling for Caption Prediction in Medical Image Analysis

DGX agent

arXiv:2607.27763v1 Announce Type: new Abstract: We describe the DS@GT submissions to the ImageCLEFmedical Caption 2026 challenge, which continues a long-running benchmark on the ROCOv2 dataset with tw

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

Exact Action Values Are Not Enough: Rollout-Verified Reinforcement Fine-Tuning of a Reasoning Model for Multi-Zone VAV Control

DGX agent

arXiv:2607.27914v1 Announce Type: new Abstract: Multi-zone variable-air-volume control must balance thermal comfort, indoor air quality, and electricity use across several continuous actuators. Model

model-releasesarxiv-cs-lg
31 Jul 2026
Research

Human Mesh Modeling for Anny Body

DGX agent

arXiv:2511.03589v3 Announce Type: replace Abstract: Parametric body models provide the structural basis for many human-centric tasks, yet existing models often rely on costly 3D scans and learned shap

researcharxiv-cs-cv
31 Jul 2026
Model Releases

MMOOC: A Comprehensive Benchmark for Out-of-Context Evaluation in Multimodal Large Language Models

DGX agent

arXiv:2607.27637v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have achieved strong performance on a wide range of vision-language tasks, but often fail under imperfect or sh

model-releasesarxiv-cs-cv
31 Jul 2026
Research

Comparing the Performance of Foundation Model Derived Embeddings with Traditional Approaches for Distant Metastasis Prediction in Head and Neck Cancer

DGX agent

arXiv:2607.26276v1 Announce Type: new Abstract: Background: Early prediction of distant metastasis (DM) risk in head and neck cancer (HNC) can enable timely interventions that may improve treatment ou

researcharxiv-cs-cv
30 Jul 2026
Model Releases

Towards Robust Reinforcement Learning for Small-Scale Language Model Agents

DGX agent

arXiv:2607.25091v1 Announce Type: new Abstract: The alignment of Small Language Models (SLMs) in the 70--500M parameter range using reinforcement learning is often considered unstable, though the unde

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Accuracy Hides How Language Models Fail: Measuring Failure States Under Matched Output Budgets

DGX agent

arXiv:2607.24268v1 Announce Type: new Abstract: Language-model benchmarks collapse two distinct measurement questions into a single accuracy score: whether a response reached an evaluable state, and w

model-releasesarxiv-cs-cl
28 Jul 2026
Model Releases

An MLIR-Based Compilation Method for Large Language Models

DGX agent

arXiv:2607.15865v2 Announce Type: replace Abstract: Large Language Models (LLMs) have become the dominant workload on modern AI accelerators, yet deploying them on specialized hardware still faces two

model-releasesarxiv-cs-cl
28 Jul 2026
Model Releases

Chamaileon: Cross-Context Binder Design with Contextualized Modeling and Mixed Sampling

DGX agent

arXiv:2607.23518v1 Announce Type: new Abstract: The rapid evolution of generative models has unlocked new potentials in protein binder design, a pivotal task in structural biology, by facilitating end

model-releasesarxiv-cs-lg
28 Jul 2026
Local Ai

Co-Harness: Co-Evolving Harnesses and Model Weights for LLM Agents

DGX agent

arXiv:2607.22688v1 Announce Type: new Abstract: Post-training agents for automated AI research requires optimizing not only model parameters, but also the runtime harness that shapes how research traj

local-aiarxiv-cs-ai
28 Jul 2026
Model Releases

Context-Adaptive Inference: A Unified Statistical and Foundation-Model View

DGX agent

arXiv:2607.23304v1 Announce Type: cross Abstract: Modern predictive systems are expected to adapt their behavior to the specific situation they are facing. A clinical model should not treat every pati

model-releasesarxiv-cs-lg
28 Jul 2026
Model Releases

D-Score: A Spectral Hidden-State Signal for Hallucination Detection in Large Language Models

DGX agent

arXiv:2607.24586v1 Announce Type: cross Abstract: Large Language Models can produce fluent text that is false, unsupported by the available evidence, or inconsistent with information that appears to b

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Do Language Models Converge to Themselves? Recursive Self-Refinement as Textual Relaxation

DGX agent

arXiv:2607.22653v1 Announce Type: new Abstract: Large language models are increasingly used in recursive refinement workflows, where an initial draft is repeatedly revised by the same model. Despite t

model-releasesarxiv-cs-ai
28 Jul 2026
Applications

From Hybrid Mechanistic--Data-Driven Modeling Toward Neuro-Symbolic AI: What, Why, and How

DGX agent

arXiv:2607.22811v1 Announce Type: cross Abstract: Hybrid mechanistic/data-driven models, which combine first-principles with learned components, are increasingly used in process engineering and scient

applicationsarxiv-cs-ai
28 Jul 2026
Model Releases

Learning Sampling Parameters for Diffusion Models

DGX agent

arXiv:2607.23488v1 Announce Type: cross Abstract: Text-to-image diffusion models expose many inference-time sampling parameters, including prompts, negative prompts, classifier-free guidance scales, a

model-releasesarxiv-cs-cv
28 Jul 2026
Safety

Making Mathematical Knowledge Explainable, Accessible and Interoperable Through Large Language Model Integration

DGX agent

arXiv:2607.24512v1 Announce Type: new Abstract: Mathematical models are central to formalizing research problems, yet their documentation often falls short of FAIR principles. Knowledge bases such as

safetyarxiv-cs-ai
28 Jul 2026
Tutorials

Offline-to-Online Creative Optimization with Generative Models and Adaptive Testing

DGX agent

arXiv:2607.23696v1 Announce Type: new Abstract: Ad creative optimization is increasingly constrained by evaluation rather than generation. Generative models can produce many plausible creatives, but r

tutorialsarxiv-cs-ai
28 Jul 2026
Model Releases

RareLens: Towards End-to-End Rare Disease Care via Aligning Divergent Large Language Model Reasoning

DGX agent

arXiv:2607.23290v1 Announce Type: new Abstract: Rare diseases collectively affect an estimated 3.5% to 5.9% of the population, yet more than 70% of patients are misdiagnosed and many endure years of e

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Reference Feature Atlases for Mechanistic Auditing of Language Models

DGX agent

arXiv:2607.22570v1 Announce Type: new Abstract: Auditing a new language model usually means relearning and reinterpreting its internal features from scratch. We propose a reference feature atlas: a sp

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Robustifying pathology foundation models via fine-tuning

DGX agent

arXiv:2607.22861v1 Announce Type: cross Abstract: Pathology foundation models (FMs) produce powerful tile-level representations which remain sensitive to scanner and staining variability, undermining

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

scMIR: a vision-language foundation model for single-cell light microscopy image representation

DGX agent

arXiv:2607.22712v1 Announce Type: cross Abstract: Single-cell light microscopy images have become an important data source for characterizing cell phenotypes, but their complexity and heterogeneity po

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Stress-Testing EEG Foundation Models for Clinical Decoding: Dataset Identity and Targeted Negative Controls

DGX agent

arXiv:2607.24519v1 Announce Type: cross Abstract: Pretrained EEG foundation models are increasingly proposed for clinical decoding, but their transfer across populations and robustness to negative con

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Enjoy Your Talk: A Human-Centered Benchmark for Multi-Turn Dialogue with Decoupled User Simulation, Target Modeling, and Judging

DGX agent

arXiv:2607.10428v2 Announce Type: replace Abstract: Evaluating large language models (LLMs) as multi-turn conversational partners requires probing capabilities that single-turn benchmarks miss: person

model-releasesarxiv-cs-cl
27 Jul 2026
Model Releases

Enough is as good as a feast: A Comprehensive Analysis of How Reinforcement Learning Mitigates Task Conflicts in LLMs

DGX agent

arXiv:2607.22039v1 Announce Type: new Abstract: Model merging plays a crucial role in consolidating multiple specialized models into a single, unified model, especially in the era of large language mo

model-releasesarxiv-cs-cl
27 Jul 2026
Safety

Pretraining EHR Foundation Models with Patient-Aware Sampling

DGX agent

arXiv:2607.22114v1 Announce Type: new Abstract: Autoregressive foundation models for electronic health records (EHRs) typically inherit pretraining methods from language modeling, where patient trajec

safetyarxiv-cs-lg
27 Jul 2026
Model Releases

Rethinking Layer-Wise Information Allocation for Vision Foundation Model Adaptation

DGX agent

arXiv:2607.21973v1 Announce Type: new Abstract: Vision foundation models are increasingly reused as frozen backbones for downstream visual recognition, making parameter-efficient adaptation a central

model-releasesarxiv-cs-cv
27 Jul 2026
Model Releases

Agentic coding without the cloud: evaluating open-weight large language models on longitudinal data preparation tasks

DGX agent

arXiv:2607.21482v1 Announce Type: new Abstract: Large language models (LLMs) and agents are now widely used tools in code development, with data typically sent to third-party cloud-based models. Their

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

ConfidenceBench: Evaluating Confidence Calibration in Large Language Models

DGX agent

arXiv:2607.20526v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed in settings where fluent but incorrect answers can be costly. In these settings, accuracy alone i

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

REGARD: Regional Affective Differences in Large Language Models

DGX agent

arXiv:2607.20722v1 Announce Type: new Abstract: Large language models trained and aligned within different linguistic and regional ecosystems may frame the same political, cultural, and geopolitical e

model-releasesarxiv-cs-cl
24 Jul 2026
Model Releases

Toward Mechanistic Interpretability of an AI Foundation Model Fine-Tuned for Atmospheric Chemistry

DGX agent

arXiv:2607.20778v1 Announce Type: new Abstract: Weather forecasting foundation models (FMs) are increasingly fine-tuned to predict air quality, offering fast global pollution forecasts at lower comput

model-releasesarxiv-cs-lg
24 Jul 2026
Safety

Abstraction Induces the Brain Alignment of Language and Speech Models

DGX agent

arXiv:2602.04081v2 Announce Type: replace Abstract: Research has repeatedly demonstrated that intermediate hidden states extracted from large language models and speech audio models predict measured b

safetyarxiv-cs-cl
23 Jul 2026
← Previous
1…2223242526…1021
Next →