AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
59,841 results
27 Jul 2026

Rethinking Layer-Wise Information Allocation for Vision Foundation Model Adaptation

Model ReleasesDGX agent

arXiv:2607.21973v1 Announce Type: new Abstract: Vision foundation models are increasingly reused as frozen backbones for downstream visual recognition, making parameter-efficient adaptation a central

24 Jul 2026

Agentic coding without the cloud: evaluating open-weight large language models on longitudinal data preparation tasks

Model ReleasesDGX agent

arXiv:2607.21482v1 Announce Type: new Abstract: Large language models (LLMs) and agents are now widely used tools in code development, with data typically sent to third-party cloud-based models. Their

Announcing Fugu-Ultra v1.1 🐡 We’ve been thrilled by the reception to the Fugu model family. Thanks to everyone who tried it, shared feedbac…

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model ReleasesDGX agent

Announcing Fugu-Ultra v1.1 🐡 We’ve been thrilled by the reception to the Fugu model family. Thanks to everyone who tried it, shared feedback, and trusted Fugu with real work. Today, we’re releasing Fu

ConfidenceBench: Evaluating Confidence Calibration in Large Language Models

Model ReleasesDGX agent

arXiv:2607.20526v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed in settings where fluent but incorrect answers can be costly. In these settings, accuracy alone i

i want the US to win in AI both in open source and proprietary models, and i am glad to see this

SafetyDGX agent

i want the US to win in AI both in open source and proprietary models, and i am glad to see this For my first post, I’m sharing a letter @NVIDIA signed on why open models matter. AI will transform eve

Open weight models will ensure that the entire world benefits from AI growth, and that America does not get left behind

SafetyDGX agent

Open weight models will ensure that the entire world benefits from AI growth, and that America does not get left behind For my first post, I’m sharing a letter @NVIDIA signed on why open models matter

REGARD: Regional Affective Differences in Large Language Models

Model ReleasesDGX agent

arXiv:2607.20722v1 Announce Type: new Abstract: Large language models trained and aligned within different linguistic and regional ecosystems may frame the same political, cultural, and geopolitical e

Toward Mechanistic Interpretability of an AI Foundation Model Fine-Tuned for Atmospheric Chemistry

Model ReleasesDGX agent

arXiv:2607.20778v1 Announce Type: new Abstract: Weather forecasting foundation models (FMs) are increasingly fine-tuned to predict air quality, offering fast global pollution forecasts at lower comput

23 Jul 2026

Abstraction Induces the Brain Alignment of Language and Speech Models

SafetyDGX agent

arXiv:2602.04081v2 Announce Type: replace Abstract: Research has repeatedly demonstrated that intermediate hidden states extracted from large language models and speech audio models predict measured b

Economic Evaluations of Language Models

Model ReleasesDGX agent

arXiv:2607.19375v1 Announce Type: cross Abstract: Language models perform economically valuable work, yet they are not currently assessed for how well they perform every economically valuable task. We

MissingBench-Verified: Probing Vision-Language Models' Inability to Detect Missing Object Parts

Model ReleasesDGX agent

arXiv:2607.18673v2 Announce Type: new Abstract: Vision Language Models (VLMs) are well known for hallucinating non-existent objects in images. Objects with missing parts present a unique challenge for

Post-Training in Time Series Foundation Models: A Unifying Framework

Model ReleasesDGX agent

arXiv:2607.20002v1 Announce Type: cross Abstract: Time series foundation models (TSFMs) have emerged as general-purpose models for time series analysis, but pretraining alone is often insufficient for

ROMS-IMLE: A Minimalist Approach to Competitive Single-Step Generative Modelling

Model ReleasesDGX agent

arXiv:2607.19332v1 Announce Type: cross Abstract: Generative models have undergone many generations of evolution, from VAEs/GANs to diffusion/flow matching. Along the way, the underlying techniques ha

The C-index illusion: discrimination without calibration in published survival models

ResearchDGX agent

arXiv:2607.19526v1 Announce Type: new Abstract: 'Stop Chasing the C-index when Evaluating Survival Analysis Models' (ICML 2026, Spotlight) argued normatively, on synthetic data, that evaluating surviv

21 Jul 2026

AI models pushing the frontier are a growing challenge for cybersecurity. A few weeks ago, I asked Demis what's underhyped in AI right now a…

AgentsDGX agent

AI models pushing the frontier are a growing challenge for cybersecurity. A few weeks ago, I asked Demis what's underhyped in AI right now and on his mind: 'I'm very excited about this new agentic era

16 Jul 2026

DarwinLM: Evolutionary Structured Pruning of Large Language Models

Model ReleasesDGX agent

arXiv:2502.07780v4 Announce Type: replace Abstract: Large Language Models (LLMs) have achieved significant success across various NLP tasks. However, their massive computational costs limit their wide

We’re excited to collaborate with NVIDIA to build the next generation of Fugu orchestration models together, by incorporating leading open-w…

Model ReleasesDGX agent

We’re excited to collaborate with NVIDIA to build the next generation of Fugu orchestration models together, by incorporating leading open-weights models. Sakana AI Teams With NVIDIA to Advance Open M

What Models Express, Suppress, and Resist: Auditing Open-Weight LLMs with Persona Vectors

AgentsDGX agent

arXiv:2607.13162v1 Announce Type: cross Abstract: What a language model will and will not do is largely set during post-training, but which behaviors it expresses, hides, or resists is not revealed by

15 Jul 2026

ABot-N1: Toward a General Visual Language Navigation Foundation Model

Model ReleasesDGX agent

arXiv:2607.10383v2 Announce Type: replace-cross Abstract: Visual Language Navigation foundation models aim to unify deep reasoning for grounded spatial decisions with broad versatility for diverse emb

Audio-Native Speech Recognition with a Frozen Discrete-Diffusion Language Model

ResearchDGX agent

arXiv:2607.13013v1 Announce Type: new Abstract: Automatic speech recognition is dominated by autoregressive decoders that emit one token at a time. We ask whether a discrete diffusion language model c

Current efficient frontier of open models

Model ReleasesDGX agent

Efficiency defined as score over active parameters. Removed all the models that were not on the pareto frontier. Yes I'm aware that artificialanalysis.ai aggregate benchmark isn't perfect, but I have

DiTailed: Ensuring Visual Object Consistency in Text-Image-to-Image Flow Matching Models

Model ReleasesDGX agent

arXiv:2607.12539v1 Announce Type: new Abstract: Despite remarkable progress in text-guided image editing, generative models frequently fail to preserve visual object consistency, defined as the preser

Evaluating Health Misinformation in Low-Resource Languages: Integrating Small Language Models with a Culturally-Sensitive Responsible NLP Framework (Bangla as a Case Study)

Model ReleasesDGX agent

arXiv:2607.12336v1 Announce Type: cross Abstract: Artificial Intelligence (AI) technologies, while serving as a foundational enabler for modern social media and digital health services, exert a bivale

Growing a Tail: Increasing Output Diversity in Large Language Models

SafetyDGX agent

arXiv:2411.02989v2 Announce Type: replace Abstract: How diverse are the outputs of large language models when diversity is desired? We examine the diversity of responses of several language models to

The Computational Basis of Confidence in Large Language Models

ResearchDGX agent

arXiv:2607.12447v1 Announce Type: cross Abstract: Reliable confidence -- the probability that a model's own answer is correct -- is essential for the trustworthy deployment of language models. Existin

The Emerging Paradigm of Geospatial Foundation Models: From Pre-Training to Agentic Reasoning

AgentsDGX agent

arXiv:2607.12177v1 Announce Type: new Abstract: The analysis of satellite and aerial imagery has entered a new era with the advent of foundation models. This paper describes the concept of Geospatial

10 Jul 2026

AI model mania and the new chip gold rush

Model ReleasesDGX agent

Just when you thought the artificial intelligence model race might slow down, it starts up again double-time. This week alone brought new models and related services from OpenAI (twice), Meta (also tw

An exact information theory of generalization phase transitions in Bayesian diffusion models

Local AiDGX agent

arXiv:2607.08041v1 Announce Type: new Abstract: How diffusion models circumvent the curse of dimensionality to learn complex distributions over high dimensional spaces from a finite training set, inst

Beyond wheelchairs and blindfolds: Investigating disability stereotypes in T2I models with INCLUDE-BENCH

Model ReleasesDGX agent

arXiv:2607.08515v1 Announce Type: new Abstract: Text-to-image (T2I) models have been shown to exhibit social biases. Prior work has mainly focused on gender, skin tone, and cultural representation wit

Robust Weighted Triangulation of Causal Effects Under Model Uncertainty

ResearchDGX agent

arXiv:2603.01119v2 Announce Type: replace-cross Abstract: A fundamental challenge in causal inference with observational data is correct specification of a causal model. When there is model uncertaint

9 Jul 2026

Diffusion Models in Simulation-Based Inference: A Tutorial Review

Model ReleasesDGX agent

arXiv:2512.20685v3 Announce Type: replace-cross Abstract: Diffusion models have recently emerged as powerful learners for simulation-based inference (SBI), enabling fast and accurate estimation of lat

Gradient-Based Speech-to-Text Alignment for Any ASR Model: From CTC to Speech LLMs

SafetyDGX agent

arXiv:2607.06831v1 Announce Type: cross Abstract: Speech-to-text alignment means finding the temporal boundaries of each word in the audio. Some models provide such an alignment directly and others do

My feed has been hijacked this week by frontier model influencers who say they've had 5.6 Sol and Fable for 'months'. I think this tells an …

Model ReleasesDGX agent

My feed has been hijacked this week by frontier model influencers who say they've had 5.6 Sol and Fable for 'months'. I think this tells an inaccurate story the field. From the outside it makes AI pro

Open-Ended Scenario Reasoning for Specialist Model Adaptation

SafetyDGX agent

arXiv:2607.06625v1 Announce Type: cross Abstract: Process industries have accumulated validated specialist models, yet sensor drift, feedstock variation, and regime switching cause these models to deg

The new GPT-5.6 family: Luna, Terra, Sol

Model ReleasesDGX agent

OpenAI's latest flagship model hit general availability this morning, and comes in three sizes: Luna, Terra, and Sol (from smallest to largest). The new models are priced per 1M input/output tokens as

8 Jul 2026

FourTune: Towards Fully 4-Bit Efficient Post-Training for Diffusion Models

Model ReleasesDGX agent

arXiv:2607.05711v1 Announce Type: cross Abstract: Diffusion models have become a dominant paradigm for high-quality generative modeling, while post-training is essential for adapting them to diverse d

Harnessing Generative Image Models for Training-Free Primitive Shape Abstraction

Model ReleasesDGX agent

arXiv:2607.05568v1 Announce Type: cross Abstract: Representing 3D shapes as compact sets of geometric primitives is fundamental to robotics, simulation, and scene understanding. Generative image model

Model page: https://ollama.com/library/glm-5.2

Local AiDGX agent

GLM-5.2 is a language model available through Ollama's model library, likely representing an updated version of the GLM (General Language Model) series with improvements in capabilities or performance

Omni-RRM: Advancing Omni Reward Modeling via Automatic Rubric-Grounded Preference Synthesis

Model ReleasesDGX agent

arXiv:2602.00846v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) struggle with alignment due to the limitations of existing reward models (RMs), which are predominantly vis

- OpenAI continually outperforms Anthropic models on computer use (with Claude, I’m surprised when it works, with Codex, I expect it to work…

Model ReleasesDGX agent

- OpenAI continually outperforms Anthropic models on computer use (with Claude, I’m surprised when it works, with Codex, I expect it to work) - when I prompt with Fable 5.5, I feel like I’m motivating

OpenAI launches GPT-Live voice model series ahead of broad GPT-5.6 release

Model ReleasesDGX agent

OpenAI Group PBC today introduced GPT-Live, a family of artificial intelligence models optimized to process spoken instructions. The model series will power ChatGPT’s voice mode. Additionally, OpenAI

Performance Optimization and Comparative Analysis of Generative AI Models on Advanced Accelerators

ResearchDGX agent

arXiv:2607.05400v1 Announce Type: cross Abstract: Generative AI models, such as Large Language Models (LLMs) and diffusion models, have demonstrated impressive performance across a wide range of tasks

The relationship between reasoning and performance in large language models--o3 (mini) thinks harder, not longer

Model ReleasesDGX agent

arXiv:2502.15631v2 Announce Type: replace-cross Abstract: Large language models have demonstrated remarkable progress in mathematical reasoning, leveraging chain-of-thought and reinforcement learning.

Xiaomi now processes more AI tokens than OpenAI. On OpenRouter, Chinese models just crossed 45% of all token volume. Anthropic is at 15.3%. …

Model ReleasesDGX agent

Xiaomi now processes more AI tokens than OpenAI. On OpenRouter, Chinese models just crossed 45% of all token volume. Anthropic is at 15.3%. OpenAI is at 7.4%. For now, the frontier is American models,

7 Jul 2026

Brand-as-Memory: Vision-Language Models Encode Causal, Mechanistically Localizable Credibility Priors for News Sources

Model ReleasesDGX agent

arXiv:2607.03365v1 Announce Type: cross Abstract: Vision-language models (VLMs) increasingly read news and web content as images, where the publisher's identity is visually present. We show that VLMs

Can Model Merging Improve Aggregation in DiLoCo?

Local AiDGX agent

arXiv:2607.03011v1 Announce Type: cross Abstract: Model merging techniques, which aggregate independently finetuned models into one to combine their capabilities, have become a topic of significant in

CollabEval: Statistically Efficient Collaborative Model Evaluation via Matrix Completion

TutorialsDGX agent

arXiv:2607.05046v1 Announce Type: new Abstract: Evaluating generative AI models is a routine, but resource-intensive, process that is conducted over and over again during the course of model developme

Criterion-Conditional In-Context Learning: Evaluating Criterion-Shift Adaptation in Vision-Language Models

Model ReleasesDGX agent

arXiv:2607.02575v1 Announce Type: cross Abstract: Vision-language models can perform new tasks without parameter updates through in-context learning (ICL), whose core mechanism is utilizing the suppor

Cross-device Collaborative Test-time Adaptation with Zeroth-order Optimization and Model Merging

ApplicationsDGX agent

arXiv:2607.02988v1 Announce Type: new Abstract: Test-time adaptation (TTA) mitigates domain shifts by using incoming test data to update a model on the fly. The majority of TTA methods require resourc

EvoXplain: When Machine Learning Models Agree on Predictions but Disagree on Why -- Measuring Mechanistic Multiplicity Across Training Runs

ResearchDGX agent

arXiv:2512.22240v5 Announce Type: replace-cross Abstract: Machine learning models are primarily judged by predictive performance, especially in applied genomics, where explanations are read as biologi

Multiplayer Interactive World Models with Representation Autoencoders

Model ReleasesDGX agent

arXiv:2607.05352v1 Announce Type: cross Abstract: We introduce the first multiplayer world model for highly dynamic environments governed by complex physical interactions. Whereas single-player world

No Reliable Evidence of Self-Reported Sentience in Small Large Language Models

Model ReleasesDGX agent

arXiv:2601.15334v2 Announce Type: replace-cross Abstract: Whether language models possess sentience has no empirical answer. But whether they believe themselves to be sentient can, in principle, be te

Spectral Rewiring for Exploration, Purification, and Model Merging

Model ReleasesDGX agent

arXiv:2607.03065v1 Announce Type: cross Abstract: Reinforcement learning has become a standard post-training recipe for large language models, but dense full-parameter updates create two deployment-re

Trust-Region Noise Search for Black-Box Alignment of Diffusion and Flow Models

Local AiDGX agent

arXiv:2603.14504v2 Announce Type: replace-cross Abstract: Optimizing the noise samples of diffusion and flow models is an increasingly popular approach to align these models to target rewards at infer

2 Jul 2026

Comparing Large Language Models on Scrum Certification-Style Questions: Accuracy, Stability, and Error Patterns

Model ReleasesDGX agent

arXiv:2607.00048v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used in exam- and certification-style question answering tasks, where their ability to retrieve, interpr

Steal the Patch Size: Adversarially Manipulate Vision-Language Models

Model ReleasesDGX agent

arXiv:2607.00174v1 Announce Type: new Abstract: We present a black-box model-stealing attack that recovers private vision-tokenizer configurations of deployed vision-language models (VLMs), including

Utilizing Earth Foundation Models to Enhance the Simulation Performance of Hydrological Models with AlphaEarth Embeddings

ResearchDGX agent

arXiv:2601.01558v2 Announce Type: replace-cross Abstract: Predicting river flow in places without streamflow records is challenging because basins respond differently to climate, terrain, vegetation,

1 Jul 2026

Run NVIDIA Nemotron and OpenAI GPT OSS models on Amazon Bedrock in AWS GovCloud (US)

Model ReleasesDGX agent

We're excited to introduce US-based frontier open-weight models in AWS GovCloud (US). With this release, Amazon Bedrock now supports OpenAI’s open-weight GPT OSS models (120B and 20B) and NVIDIA Nemot

30 Jun 2026

China’s Meituan open-sources massive LongCat-2.0 AI model, saying it was trained on domestic chips

Model ReleasesDGX agent

Beijing, China-based Meituan Inc. today debuted its next-generation LongCat-2.0 open-source large language model, stating that the company trained the 1.6-trillion-parameter model on domestic Chinese

How Should World Models Be Evaluated for Embodied Decision-Making? A Decision-Making-Centric Position

SafetyDGX agent

arXiv:2606.15032v2 Announce Type: replace Abstract: World models have become a central abstraction in modern AI. The term now refers to several different objects: action-conditioned environment models

← Previous
1…2728293031…998
Next →