AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
59,840 results
Model Releases

Model Recycling Framework for Multi-Source Data-Free Supervised Transfer Learning

DGX agent

arXiv:2508.02039v2 Announce Type: replace Abstract: Increasing concerns for data privacy and other difficulties associated with retrieving source data for model training have created the need for sour

model-releasesarxiv-cs-lg
8 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Think Fast: Estimating No-CoT Task-Completion Time Horizons of Frontier AI Models

DGX agent

arXiv:2606.07157v1 Announce Type: new Abstract: Many efforts to ensure frontier AI models are safe rely on monitoring their chain-of-thought (CoT) reasoning. If models become able to perform sufficien

model-releasesarxiv-cs-ai
8 Jun 2026
Applications

Five labs, five minds: building a multi-model finance drama on small models

DGX agent

This article describes a collaborative hackathon project involving five research labs that developed a financial simulation drama using small language models, focusing on building complex multi-agent

applicationshugging-face
6 Jun 2026
Model Releases

LDARNet: DNA Adaptive Representation Network with Learnable Tokenization for Genomic Modeling

DGX agent

arXiv:2606.04552v1 Announce Type: new Abstract: Genomic foundation models increasingly adopt large language model architectures, yet almost universally rely on fixed tokenization schemes such as k-mer

model-releasesarxiv-cs-cl
4 Jun 2026
Safety

HARVE: Hacking-Aware Reward-Head Vector Editing for Robust Reward Models

DGX agent

arXiv:2606.03131v1 Announce Type: new Abstract: Reward models are central to large language model (LLM) alignment, but they remain vulnerable to reward hacking. To evaluate reward-model robustness, we

safetyarxiv-cs-lg
3 Jun 2026
Research

Linguistic Productivity in Large Language Models: Models Coerce, but do not Preempt

DGX agent

arXiv:2606.02953v1 Announce Type: new Abstract: Usage-based theories of grammars posit that creative productivity of the structures of language is both bolstered and constrained by two distinct freque

researcharxiv-cs-cl
3 Jun 2026
Model Releases

Catch-Only-One: Non-Transferable Examples for Model-Specific Authorization

DGX agent

arXiv:2510.10982v2 Announce Type: replace-cross Abstract: Recent AI regulations increasingly emphasize the need for mechanisms that preserve the utility of data for AI innovation while preventing misu

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Model-Native Computing Architecture: Envisioning Future System Architecture Through the Lens of Computer Architecture

DGX agent

arXiv:2606.00288v1 Announce Type: new Abstract: Large language models are undergoing a transition from model technology to system technology. As developers use Codex, Claude Code, AutoGPT, and related

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Visual-Noise Guided In-Context Distillation for Multimodal Large Language Model Unlearning

DGX agent

arXiv:2606.00105v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable progress on vision-language tasks, but they may also memorize and expose sensitive o

model-releasesarxiv-cs-ai
2 Jun 2026
Research

Quantifying Error Propagation and Model Collapse in Diffusion Models

DGX agent

arXiv:2602.16601v2 Announce Type: replace-cross Abstract: Machine learning models are increasingly trained or fine-tuned on synthetic data. Recursively training on such data has been observed to signi

researcharxiv-cs-lg
1 Jun 2026
Model Releases

Unraveling LoRA Interference: Orthogonal Subspaces for Robust Model Merging

DGX agent

arXiv:2505.22934v2 Announce Type: replace-cross Abstract: Fine-tuning large language models (LMs) for individual tasks yields strong performance but is expensive for deployment and storage. Recent wor

model-releasesarxiv-cs-ai
1 Jun 2026
Safety

BEAMS: Benchmarking and Evaluating AI for Modeling and Simulation

DGX agent

arXiv:2605.28994v1 Announce Type: new Abstract: AI tools to support real world decision making must be able to build simulation models that inform their recommendations and render them interpretable.

safetyarxiv-cs-ai
29 May 2026
Tutorials

HD-Prot: A Protein Language Model for Joint Sequence-Structure Modeling with Continuous Structure Tokens

DGX agent

arXiv:2512.15133v2 Announce Type: replace-cross Abstract: Proteins inherently possess a consistent sequence-structure duality. The abundance of protein sequence data, which can be readily represented

tutorialsarxiv-cs-ai
29 May 2026
Model Releases

Inferring the Size of Large Language Models From Popular Text Memorization

DGX agent

arXiv:2605.29223v1 Announce Type: new Abstract: The parameter counts of the most widely used large language models (LLMs) are often withheld by their developers, leaving model size -- a primary refere

model-releasesarxiv-cs-lg
29 May 2026
Local Ai

A very cool model for the GPU poor bros Trained on an ungodly amount of tokens for a 8b a1b model Gonna be super fast excited to try this ou…

DGX agent

This post announces an 8-9 billion parameter language model optimized for efficiency and speed, trained on a very large token dataset, and positioned as accessible for users with limited GPU resources

local-aiclem-delangue--x
28 May 2026
Model Releases

Aligning Language Model Benchmarks with Pairwise Preferences

DGX agent

arXiv:2602.02898v2 Announce Type: replace Abstract: Language model benchmarks are pervasive and computationally-efficient proxies for real-world performance. However, many recent works find that bench

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

The Missing Piece in Pre-trained Model Evaluation: Reward-Guided Decoding Unlocks Task-Oriented Behavior Without Parameter Updates

DGX agent

arXiv:2605.28020v1 Announce Type: new Abstract: With the rapid progress of large language models (LLMs), reliably evaluating the capabilities of pre-trained LLMs has become increasingly important. The

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

An uncertainty-aware Bayesian framework for machine learning classification models: A case study in land cover classification

DGX agent

arXiv:2503.21510v3 Announce Type: replace-cross Abstract: Ensuring that predictions of machine learning (ML) classification models are accompanied by uncertainty estimates is one of the main pillars o

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

Today we're announcing ESMFold2, an open scientific engine to power prediction, design, and discovery across protein biology. The new model …

DGX agent

Today we're announcing ESMFold2, an open scientific engine to power prediction, design, and discovery across protein biology. The new model delivers state of the art performance on protein interaction

model-releasesyann-lecun--x
27 May 2026
Model Releases

Looped Diffusion Language Models

DGX agent

arXiv:2605.26106v1 Announce Type: new Abstract: Masked diffusion models (MDMs) have emerged as a promising alternative to autoregressive models for language modeling, yet the effective design of trans

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

TriVAL: A Tri-Validation Framework for Faithful Automatic Optimization Modeling

DGX agent

arXiv:2605.23966v1 Announce Type: cross Abstract: Optimization modeling serves as the pivotal bridge between natural-language problem descriptions and optimization solvers, and remains a cornerstone f

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Cluster-Based Generalized Additive Models Informed by Random Fourier Features

DGX agent

arXiv:2512.19373v3 Announce Type: replace-cross Abstract: In developing data-driven modeling methodologies, there is an ongoing need to reconcile the strong predictive performance of opaque black-box

model-releasesarxiv-cs-lg
21 May 2026
Industry

Elon Musk and Franz von Holzhausen during the first Model S customer deliveries in 2012, and the final Model S deliveries in 2026. 14 years …

DGX agent

Tesla's Model S, designed by Chief Designer Franz von Holzhausen, marked a significant milestone spanning from its first customer deliveries in 2012 to its final production run in 2026, representing 1

industryelon-musk--x
21 May 2026
Model Releases

MedFM-Robust: Benchmarking Robustness of Medical Foundation Models

DGX agent

arXiv:2605.19027v1 Announce Type: new Abstract: Medical foundation models (MedFMs) have emerged as transformative tools in healthcare, demonstrating capabilities across diverse clinical applications.

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

Adaptive Camera Sensor for Vision Models

DGX agent

arXiv:2503.02170v2 Announce Type: replace-cross Abstract: Domain shift remains a persistent challenge in deep-learning-based computer vision, often requiring extensive model modifications or large lab

model-releasesarxiv-cs-ai
19 May 2026
Applications

Can Vision Language Models Be Adaptive in Mathematics Education? A Learner Model-based Rubric Study

DGX agent

arXiv:2605.16011v1 Announce Type: cross Abstract: Adaptive learning refers to educational technologies that track learners' learning progress and adapt the instructional process based on individual le

applicationsarxiv-cs-ai
18 May 2026
Local Ai

LTX 2.3 is now supported in Comfyui-Mesh for splitting models across Ethernet or multigpu machines with Nvenc codec. Major vram fixes included for flux2/LTX model implementations in the node.

DGX agent

Comfyui-Mesh now supports LTX 2.3 with the ability to distribute models across multiple GPUs or networked machines using Nvenc codec for video generation. The update includes significant VRAM optimiza

local-air-stablediffusion
17 May 2026
Model Releases

Improving the Accuracy of Amortized Model Comparison with Self-Consistency

DGX agent

arXiv:2508.20614v3 Announce Type: replace-cross Abstract: Amortized Bayesian model comparison (BMC) enables fast probabilistic ranking of models via simulation-based training of neural surrogates. How

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Intention-Conditioned Flow Occupancy Models

DGX agent

arXiv:2506.08902v4 Announce Type: replace Abstract: Large-scale pre-training has fundamentally changed how machine learning research is done today: large foundation models are trained once, and then c

model-releasesarxiv-cs-lg
13 May 2026
Agents

Generating synthetic electronic health record data using agent-based models to evaluate machine learning robustness under mass casualty incidents

DGX agent

arXiv:2605.09951v1 Announce Type: new Abstract: ML models in healthcare are typically evaluated using curated real-world EHR data. A key limitation of such evaluations is that they may fail to assess

agentsarxiv-cs-lg
12 May 2026
Model Releases

NeuralBench: A Unifying Framework to Benchmark NeuroAI Models

DGX agent

arXiv:2605.08495v1 Announce Type: new Abstract: Deep learning and large public datasets have recently catalyzed the proliferation of AI models for processing brain recordings. However, systematically

model-releasesarxiv-cs-lg
12 May 2026
Research

Predicting Large Model Test Losses with a Noisy Quadratic System

DGX agent

arXiv:2605.09154v1 Announce Type: new Abstract: We introduce a predictive model that estimates the pre-training loss of large models from model size (N), batch size (B) and number of weight updates (K

researcharxiv-cs-lg
12 May 2026
Model Releases

Chain-based Distillation for Effective Initialization of Variable-Sized Small Language Models

DGX agent

arXiv:2605.07783v1 Announce Type: new Abstract: Large language models (LLMs) achieve strong performance but remain costly to deploy in resource-constrained settings. Training small language models (SL

model-releasesarxiv-cs-cl
11 May 2026
Hardware

LLMSpace: Carbon Footprint Modeling for Large Language Model Inference on LEO Satellites

DGX agent

arXiv:2605.05615v2 Announce Type: replace Abstract: Large language models (LLMs) impose rapidly growing energy demands, creating an emerging energy and carbon crisis driven by large-scale inference. S

hardwarearxiv-cs-lg
11 May 2026
Model Releases

EvoLM: Self-Evolving Language Models through Co-Evolved Discriminative Rubrics

DGX agent

arXiv:2605.03871v1 Announce Type: new Abstract: Language models encode substantial evaluative knowledge from pretraining, yet current post-training methods rely on external supervision (human annotati

model-releasesarxiv-cs-ai
7 May 2026
Local Ai

Model Quantization: Post-Training Quantization Using NVIDIA Model Optimizer

DGX agent

Post-training quantization is a technique that reduces model size and improves inference performance by converting weights and activations to lower precision formats after training is complete. NVIDIA

local-ainvidia-developer
7 May 2026
Model Releases

StoryAlign: Evaluating and Training Reward Models for Story Generation

DGX agent

arXiv:2605.04831v1 Announce Type: new Abstract: Story generation aims to automatically produce coherent, structured, and engaging narratives. Although large language models (LLMs) have significantly a

model-releasesarxiv-cs-cl
7 May 2026
Model Releases

It feels like agent harness evolution runs on two axes that usually get conflated. There’s the temporal axis: simplify as models improve, st…

DGX agent

It feels like agent harness evolution runs on two axes that usually get conflated. There’s the temporal axis: simplify as models improve, stripping components that compensated for limitations the new

model-releasesharrison-chase--x
6 May 2026
Model Releases

Deep Time Series Models: A Comprehensive Survey and Benchmark

DGX agent

arXiv:2407.13278v3 Announce Type: replace Abstract: Time series, characterized by a sequence of data points organized in a discrete-time order, are ubiquitous in real-world scenarios. Unlike other dat

model-releasesarxiv-cs-lg
5 May 2026
Research

Enforcing tail calibration when training probabilistic forecast models

DGX agent

arXiv:2506.13687v2 Announce Type: replace-cross Abstract: Probabilistic forecasts are typically obtained using state-of-the-art statistical and machine learning models, with model parameters estimated

researcharxiv-cs-lg
5 May 2026
Model Releases

TRIP-Evaluate: An Open Multimodal Benchmark for Evaluating Large Models in Transportation

DGX agent

arXiv:2605.00907v1 Announce Type: new Abstract: Large language models (LLMs) and multimodal large models (MLLMs) are increasingly used for transportation tasks such as regulation question answering, t

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

The Harness is a Context Manager on Behalf of the Model What happens when the context window fills up and who decides? This decision is exte…

DGX agent

The Harness is a Context Manager on Behalf of the Model What happens when the context window fills up and who decides? This decision is external to the model - The Harness designer must have some opin

model-releasesharrison-chase--x
2 May 2026
Agents

Pragmos: A Process Agentic Modeling System

DGX agent

arXiv:2604.27311v1 Announce Type: cross Abstract: The advent of Large Language Models (LLMs) has significantly transformed tasks across Software Engineering. In the context of Business Process Managem

agentsarxiv-cs-ai
1 May 2026
Safety

A Survey of Process Reward Models: From Outcome Signals to Process Supervisions for Large Language Models

DGX agent

arXiv:2510.08049v3 Announce Type: replace-cross Abstract: Although Large Language Models (LLMs) exhibit advanced reasoning ability, conventional alignment remains largely dominated by outcome reward m

safetyarxiv-cs-ai
30 Apr 2026
Research

Fitting Large Nonlinear Mixed Effects Models Using Variational Expectation Maximization

DGX agent

arXiv:2604.26160v1 Announce Type: cross Abstract: Nonlinear Mixed Effects models (NLME) models are widely used in pharmacometrics and related fields to analyze hierarchical and longitudinal data. Howe

researcharxiv-cs-lg
30 Apr 2026
Model Releases

reward-lens: A Mechanistic Interpretability Library for Reward Models

DGX agent

arXiv:2604.26130v1 Announce Type: cross Abstract: Every RLHF-trained language model is shaped by a reward model, yet the mechanistic interpretability toolkit -- logit lens, direct logit attribution, a

model-releasesarxiv-cs-ai
30 Apr 2026
Research

Cornserve: A Distributed Serving System for Any-to-Any Multimodal Models

DGX agent

arXiv:2603.12118v2 Announce Type: replace Abstract: Any-to-Any models are an emerging class of multimodal models that accept combinations of multimodal data (e.g., text, image, video, audio) as input

researcharxiv-cs-lg
29 Apr 2026
Model Releases

Benchmarking Pathology Foundation Models for Breast Cancer Survival Prediction

DGX agent

arXiv:2604.24679v1 Announce Type: new Abstract: Pathology foundation models (PFMs) have recently emerged as powerful pretrained encoders for computational pathology, enabling transfer learning across

model-releasesarxiv-cs-cv
28 Apr 2026
← Previous
1…1819202122…1247
Next →