AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
59,399 results
9 Jun 2026

Auditing Proprietary Alignment in Large Language Models: A Comparative Framework Without a Ground-Truth Standard

Model ReleasesDGX agent

arXiv:2606.08381v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly released and deployed through opaque development and deployment pipelines, enabling model providers to i

BEACON: Behavioral Entropy Aggregation for Cross-Model Hallucination Detection in Large Language Models

ResearchDGX agent

arXiv:2606.07528v1 Announce Type: cross Abstract: Hallucination in large language models (LLMs), defined as the generation of factually incorrect or unsupported content, remains a critical barrier to

Cutting LLM Evaluation Costs with SySRs: A Bandit Algorithm that Provably Exploits Model Similarity

ApplicationsDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.07726v1 Announce Type: new Abstract: Large Language Models are typically benchmarked by evaluating every model on every test query. For practitioners seeking the best model to deploy, this

GPT-Micro: A large language paradigm for accelerated, inexpensive, and thermodynamics-consistent discovery of constitutive models in manufacturing

AgentsDGX agent

arXiv:2606.08238v1 Announce Type: new Abstract: Constitutive modeling of the relationship between process-imposed material states and fundamental material properties is critical to control of material

SLMJury: Can Small Language Models Judge as Well as Large Ones?

Model ReleasesDGX agent

arXiv:2606.07810v1 Announce Type: cross Abstract: Large language models (LLMs) are widely used as judges for evaluating model outputs, but their high cost, latency, and opacity limit scalability. We i

8 Jun 2026

Characterizing Learning Dynamics under Relative Reparameterization of Singular Models

Model ReleasesDGX agent

arXiv:2206.08598v2 Announce Type: replace Abstract: A common way to analyze learning of statistical models is to consider operations in the models parameter space, however this becomes challenging whe

Data-Constrained Language Model Pretraining: Improved Regularization and Scaling Laws

Model ReleasesDGX agent

arXiv:2606.06888v1 Announce Type: new Abstract: Classical scaling laws for language model pretraining balance model size against training dataset size under a fixed compute budget, assuming abundant d

Learning Explicit Behavioral Models with Adaptive Questions and World-Model Probes

Local AiDGX agent

arXiv:2606.07127v1 Announce Type: new Abstract: Interactive agents trained only against task return can achieve high scores while failing to represent the mechanisms that make their actions succeed. T

Model Recycling Framework for Multi-Source Data-Free Supervised Transfer Learning

Model ReleasesDGX agent

arXiv:2508.02039v2 Announce Type: replace Abstract: Increasing concerns for data privacy and other difficulties associated with retrieving source data for model training have created the need for sour

Think Fast: Estimating No-CoT Task-Completion Time Horizons of Frontier AI Models

Model ReleasesDGX agent

arXiv:2606.07157v1 Announce Type: new Abstract: Many efforts to ensure frontier AI models are safe rely on monitoring their chain-of-thought (CoT) reasoning. If models become able to perform sufficien

6 Jun 2026

Five labs, five minds: building a multi-model finance drama on small models

ApplicationsDGX agent

This article describes a collaborative hackathon project involving five research labs that developed a financial simulation drama using small language models, focusing on building complex multi-agent

4 Jun 2026

LDARNet: DNA Adaptive Representation Network with Learnable Tokenization for Genomic Modeling

Model ReleasesDGX agent

arXiv:2606.04552v1 Announce Type: new Abstract: Genomic foundation models increasingly adopt large language model architectures, yet almost universally rely on fixed tokenization schemes such as k-mer

3 Jun 2026

HARVE: Hacking-Aware Reward-Head Vector Editing for Robust Reward Models

SafetyDGX agent

arXiv:2606.03131v1 Announce Type: new Abstract: Reward models are central to large language model (LLM) alignment, but they remain vulnerable to reward hacking. To evaluate reward-model robustness, we

Linguistic Productivity in Large Language Models: Models Coerce, but do not Preempt

ResearchDGX agent

arXiv:2606.02953v1 Announce Type: new Abstract: Usage-based theories of grammars posit that creative productivity of the structures of language is both bolstered and constrained by two distinct freque

2 Jun 2026

Catch-Only-One: Non-Transferable Examples for Model-Specific Authorization

Model ReleasesDGX agent

arXiv:2510.10982v2 Announce Type: replace-cross Abstract: Recent AI regulations increasingly emphasize the need for mechanisms that preserve the utility of data for AI innovation while preventing misu

Model-Native Computing Architecture: Envisioning Future System Architecture Through the Lens of Computer Architecture

Model ReleasesDGX agent

arXiv:2606.00288v1 Announce Type: new Abstract: Large language models are undergoing a transition from model technology to system technology. As developers use Codex, Claude Code, AutoGPT, and related

Visual-Noise Guided In-Context Distillation for Multimodal Large Language Model Unlearning

Model ReleasesDGX agent

arXiv:2606.00105v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable progress on vision-language tasks, but they may also memorize and expose sensitive o

1 Jun 2026

Quantifying Error Propagation and Model Collapse in Diffusion Models

ResearchDGX agent

arXiv:2602.16601v2 Announce Type: replace-cross Abstract: Machine learning models are increasingly trained or fine-tuned on synthetic data. Recursively training on such data has been observed to signi

Unraveling LoRA Interference: Orthogonal Subspaces for Robust Model Merging

Model ReleasesDGX agent

arXiv:2505.22934v2 Announce Type: replace-cross Abstract: Fine-tuning large language models (LMs) for individual tasks yields strong performance but is expensive for deployment and storage. Recent wor

29 May 2026

BEAMS: Benchmarking and Evaluating AI for Modeling and Simulation

SafetyDGX agent

arXiv:2605.28994v1 Announce Type: new Abstract: AI tools to support real world decision making must be able to build simulation models that inform their recommendations and render them interpretable.

HD-Prot: A Protein Language Model for Joint Sequence-Structure Modeling with Continuous Structure Tokens

TutorialsDGX agent

arXiv:2512.15133v2 Announce Type: replace-cross Abstract: Proteins inherently possess a consistent sequence-structure duality. The abundance of protein sequence data, which can be readily represented

Inferring the Size of Large Language Models From Popular Text Memorization

Model ReleasesDGX agent

arXiv:2605.29223v1 Announce Type: new Abstract: The parameter counts of the most widely used large language models (LLMs) are often withheld by their developers, leaving model size -- a primary refere

28 May 2026

A very cool model for the GPU poor bros Trained on an ungodly amount of tokens for a 8b a1b model Gonna be super fast excited to try this ou…

Local AiDGX agent

This post announces an 8-9 billion parameter language model optimized for efficiency and speed, trained on a very large token dataset, and positioned as accessible for users with limited GPU resources

Aligning Language Model Benchmarks with Pairwise Preferences

Model ReleasesDGX agent

arXiv:2602.02898v2 Announce Type: replace Abstract: Language model benchmarks are pervasive and computationally-efficient proxies for real-world performance. However, many recent works find that bench

The Missing Piece in Pre-trained Model Evaluation: Reward-Guided Decoding Unlocks Task-Oriented Behavior Without Parameter Updates

Model ReleasesDGX agent

arXiv:2605.28020v1 Announce Type: new Abstract: With the rapid progress of large language models (LLMs), reliably evaluating the capabilities of pre-trained LLMs has become increasingly important. The

27 May 2026

An uncertainty-aware Bayesian framework for machine learning classification models: A case study in land cover classification

Model ReleasesDGX agent

arXiv:2503.21510v3 Announce Type: replace-cross Abstract: Ensuring that predictions of machine learning (ML) classification models are accompanied by uncertainty estimates is one of the main pillars o

Today we're announcing ESMFold2, an open scientific engine to power prediction, design, and discovery across protein biology. The new model …

Model ReleasesDGX agent

Today we're announcing ESMFold2, an open scientific engine to power prediction, design, and discovery across protein biology. The new model delivers state of the art performance on protein interaction

26 May 2026

Looped Diffusion Language Models

Model ReleasesDGX agent

arXiv:2605.26106v1 Announce Type: new Abstract: Masked diffusion models (MDMs) have emerged as a promising alternative to autoregressive models for language modeling, yet the effective design of trans

TriVAL: A Tri-Validation Framework for Faithful Automatic Optimization Modeling

Model ReleasesDGX agent

arXiv:2605.23966v1 Announce Type: cross Abstract: Optimization modeling serves as the pivotal bridge between natural-language problem descriptions and optimization solvers, and remains a cornerstone f

21 May 2026

Cluster-Based Generalized Additive Models Informed by Random Fourier Features

Model ReleasesDGX agent

arXiv:2512.19373v3 Announce Type: replace-cross Abstract: In developing data-driven modeling methodologies, there is an ongoing need to reconcile the strong predictive performance of opaque black-box

Elon Musk and Franz von Holzhausen during the first Model S customer deliveries in 2012, and the final Model S deliveries in 2026. 14 years …

IndustryDGX agent

Tesla's Model S, designed by Chief Designer Franz von Holzhausen, marked a significant milestone spanning from its first customer deliveries in 2012 to its final production run in 2026, representing 1

20 May 2026

MedFM-Robust: Benchmarking Robustness of Medical Foundation Models

Model ReleasesDGX agent

arXiv:2605.19027v1 Announce Type: new Abstract: Medical foundation models (MedFMs) have emerged as transformative tools in healthcare, demonstrating capabilities across diverse clinical applications.

19 May 2026

Adaptive Camera Sensor for Vision Models

Model ReleasesDGX agent

arXiv:2503.02170v2 Announce Type: replace-cross Abstract: Domain shift remains a persistent challenge in deep-learning-based computer vision, often requiring extensive model modifications or large lab

18 May 2026

Can Vision Language Models Be Adaptive in Mathematics Education? A Learner Model-based Rubric Study

ApplicationsDGX agent

arXiv:2605.16011v1 Announce Type: cross Abstract: Adaptive learning refers to educational technologies that track learners' learning progress and adapt the instructional process based on individual le

17 May 2026

LTX 2.3 is now supported in Comfyui-Mesh for splitting models across Ethernet or multigpu machines with Nvenc codec. Major vram fixes included for flux2/LTX model implementations in the node.

Local AiDGX agent

Comfyui-Mesh now supports LTX 2.3 with the ability to distribute models across multiple GPUs or networked machines using Nvenc codec for video generation. The update includes significant VRAM optimiza

13 May 2026

Improving the Accuracy of Amortized Model Comparison with Self-Consistency

Model ReleasesDGX agent

arXiv:2508.20614v3 Announce Type: replace-cross Abstract: Amortized Bayesian model comparison (BMC) enables fast probabilistic ranking of models via simulation-based training of neural surrogates. How

Intention-Conditioned Flow Occupancy Models

Model ReleasesDGX agent

arXiv:2506.08902v4 Announce Type: replace Abstract: Large-scale pre-training has fundamentally changed how machine learning research is done today: large foundation models are trained once, and then c

12 May 2026

Generating synthetic electronic health record data using agent-based models to evaluate machine learning robustness under mass casualty incidents

AgentsDGX agent

arXiv:2605.09951v1 Announce Type: new Abstract: ML models in healthcare are typically evaluated using curated real-world EHR data. A key limitation of such evaluations is that they may fail to assess

NeuralBench: A Unifying Framework to Benchmark NeuroAI Models

Model ReleasesDGX agent

arXiv:2605.08495v1 Announce Type: new Abstract: Deep learning and large public datasets have recently catalyzed the proliferation of AI models for processing brain recordings. However, systematically

Predicting Large Model Test Losses with a Noisy Quadratic System

ResearchDGX agent

arXiv:2605.09154v1 Announce Type: new Abstract: We introduce a predictive model that estimates the pre-training loss of large models from model size (N), batch size (B) and number of weight updates (K

11 May 2026

Chain-based Distillation for Effective Initialization of Variable-Sized Small Language Models

Model ReleasesDGX agent

arXiv:2605.07783v1 Announce Type: new Abstract: Large language models (LLMs) achieve strong performance but remain costly to deploy in resource-constrained settings. Training small language models (SL

LLMSpace: Carbon Footprint Modeling for Large Language Model Inference on LEO Satellites

HardwareDGX agent

arXiv:2605.05615v2 Announce Type: replace Abstract: Large language models (LLMs) impose rapidly growing energy demands, creating an emerging energy and carbon crisis driven by large-scale inference. S

7 May 2026

EvoLM: Self-Evolving Language Models through Co-Evolved Discriminative Rubrics

Model ReleasesDGX agent

arXiv:2605.03871v1 Announce Type: new Abstract: Language models encode substantial evaluative knowledge from pretraining, yet current post-training methods rely on external supervision (human annotati

Model Quantization: Post-Training Quantization Using NVIDIA Model Optimizer

Local AiDGX agent

Post-training quantization is a technique that reduces model size and improves inference performance by converting weights and activations to lower precision formats after training is complete. NVIDIA

StoryAlign: Evaluating and Training Reward Models for Story Generation

Model ReleasesDGX agent

arXiv:2605.04831v1 Announce Type: new Abstract: Story generation aims to automatically produce coherent, structured, and engaging narratives. Although large language models (LLMs) have significantly a

6 May 2026

It feels like agent harness evolution runs on two axes that usually get conflated. There’s the temporal axis: simplify as models improve, st…

Model ReleasesDGX agent

It feels like agent harness evolution runs on two axes that usually get conflated. There’s the temporal axis: simplify as models improve, stripping components that compensated for limitations the new

5 May 2026

Deep Time Series Models: A Comprehensive Survey and Benchmark

Model ReleasesDGX agent

arXiv:2407.13278v3 Announce Type: replace Abstract: Time series, characterized by a sequence of data points organized in a discrete-time order, are ubiquitous in real-world scenarios. Unlike other dat

Enforcing tail calibration when training probabilistic forecast models

ResearchDGX agent

arXiv:2506.13687v2 Announce Type: replace-cross Abstract: Probabilistic forecasts are typically obtained using state-of-the-art statistical and machine learning models, with model parameters estimated

TRIP-Evaluate: An Open Multimodal Benchmark for Evaluating Large Models in Transportation

Model ReleasesDGX agent

arXiv:2605.00907v1 Announce Type: new Abstract: Large language models (LLMs) and multimodal large models (MLLMs) are increasingly used for transportation tasks such as regulation question answering, t

2 May 2026

The Harness is a Context Manager on Behalf of the Model What happens when the context window fills up and who decides? This decision is exte…

Model ReleasesDGX agent

The Harness is a Context Manager on Behalf of the Model What happens when the context window fills up and who decides? This decision is external to the model - The Harness designer must have some opin

1 May 2026

Pragmos: A Process Agentic Modeling System

AgentsDGX agent

arXiv:2604.27311v1 Announce Type: cross Abstract: The advent of Large Language Models (LLMs) has significantly transformed tasks across Software Engineering. In the context of Business Process Managem

30 Apr 2026

A Survey of Process Reward Models: From Outcome Signals to Process Supervisions for Large Language Models

SafetyDGX agent

arXiv:2510.08049v3 Announce Type: replace-cross Abstract: Although Large Language Models (LLMs) exhibit advanced reasoning ability, conventional alignment remains largely dominated by outcome reward m

Fitting Large Nonlinear Mixed Effects Models Using Variational Expectation Maximization

ResearchDGX agent

arXiv:2604.26160v1 Announce Type: cross Abstract: Nonlinear Mixed Effects models (NLME) models are widely used in pharmacometrics and related fields to analyze hierarchical and longitudinal data. Howe

reward-lens: A Mechanistic Interpretability Library for Reward Models

Model ReleasesDGX agent

arXiv:2604.26130v1 Announce Type: cross Abstract: Every RLHF-trained language model is shaped by a reward model, yet the mechanistic interpretability toolkit -- logit lens, direct logit attribution, a

29 Apr 2026

Cornserve: A Distributed Serving System for Any-to-Any Multimodal Models

ResearchDGX agent

arXiv:2603.12118v2 Announce Type: replace Abstract: Any-to-Any models are an emerging class of multimodal models that accept combinations of multimodal data (e.g., text, image, video, audio) as input

28 Apr 2026

Benchmarking Pathology Foundation Models for Breast Cancer Survival Prediction

Model ReleasesDGX agent

arXiv:2604.24679v1 Announce Type: new Abstract: Pathology foundation models (PFMs) have recently emerged as powerful pretrained encoders for computational pathology, enabling transfer learning across

RoboECC: Multi-Factor-Aware Edge-Cloud Collaborative Deployment for VLA Models

ResearchDGX agent

arXiv:2603.20711v2 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models are mainstream in embodied intelligence but face high inference costs. Edge-Cloud Collaborative (ECC) depl

23 Apr 2026

DistortBench: Benchmarking Vision Language Models on Image Distortion Identification

Model ReleasesDGX agent

arXiv:2604.19966v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly used in settings where sensitivity to low-level image degradations matters, including content moderatio

Foundation Models in Biomedical Imaging: Turning Hype into Reality

Model ReleasesDGX agent

arXiv:2512.15808v2 Announce Type: replace-cross Abstract: Foundation models (FMs) are driving a prominent shift in biomedical imaging from task-specific models to unified backbone models for diverse t

Hallucination Early Detection in Diffusion Models

ResearchDGX agent

arXiv:2604.20354v1 Announce Type: new Abstract: Text-to-Image generation has seen significant advancements in output realism with the advent of diffusion models. However, diffusion models encounter di

← Previous
1…1415161718…990
Next →