AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
59,377 results
Research

IMWM: Intuition Models Complement World Models for Latent Planning

DGX agent

arXiv:2606.01626v1 Announce Type: new Abstract: Planning with a learned latent world model is a promising route to control from raw pixels, but a strong world model alone is not enough. We show this e

researcharxiv-cs-lg
2 Jun 2026
Research
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

SAE-RNA: A Sparse Autoencoder Model for Interpreting RNA Language Model Representations

DGX agent

arXiv:2510.02734v2 Announce Type: replace-cross Abstract: Deep learning, particularly with the advancement of Large Language Models, has transformed biomolecular modeling, with protein language models

researcharxiv-cs-ai
18 May 2026
Model Releases

Tabular Foundation Model for Generative Modelling

DGX agent

arXiv:2605.09424v1 Announce Type: new Abstract: Generative modelling is a demanding test of foundation models, because it requires robust, holistic representation learning for a given data modality, r

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

When Should a Language Model Trust Itself? Same-Model Self-Verification as a Conditional Confidence Signal

DGX agent

arXiv:2605.02915v1 Announce Type: new Abstract: Same-model self-verification, prompting a model to audit its own predicted answer, is a plausible confidence signal for selective prediction, but its pr

model-releasesarxiv-cs-cl
6 May 2026
Local Ai

Would a 2nd hand custom built 2080 Ti 22GB vram be worth it? How usable would it be with ComfyUI? Or maybe even 2 pcs with NVLink? Can large models, like Wan2.2 be split with NVLink like it's 44GB, or will it always be 22GB for 1 model, and 22GB for another model (encoders, CLIP, anything else)?

DGX agent

This Reddit post discusses the viability of using second-hand custom-built RTX 2080 Ti GPUs with 22GB VRAM for running Stable Diffusion models in ComfyUI, including whether two cards could be linked v

local-air-stablediffusion
4 May 2026
Model Releases

Introducing talkie: a 13B vintage language model from 1930

DGX agent

Introducing talkie: a 13B vintage language model from 1930 New project from Nick Levine, David Duvenaud, and Alec Radford (of GPT, GPT-2, Whisper fame). talkie-1930-13b-base (53.1 GB) is a '13B langua

model-releasessimon-willison
28 Apr 2026
Tutorials

Pre-trained Large Language Models Learn Hidden Markov Models In-context

DGX agent

arXiv:2506.07298v3 Announce Type: replace-cross Abstract: Hidden Markov Models (HMMs) are foundational tools for modeling sequential data with latent Markovian structure, yet fitting them to real-worl

tutorialsarxiv-cs-ai
27 Apr 2026
Model Releases

Qwen3.6 and gemma 4 just shown that we were still far from capability limit on small models, we still are. Huge models are good, but is ther…

DGX agent

Qwen3.6 and gemma 4 just shown that we were still far from capability limit on small models, we still are. Huge models are good, but is there really a point in going bigger. Just realize something. On

model-releasesclem-delangue--x
22 Apr 2026
Research

Reason Only When Needed: Efficient Generative Reward Modeling via Model-Internal Uncertainty

DGX agent

arXiv:2604.10072v1 Announce Type: new Abstract: Recent advancements in the Generative Reward Model (GRM) have demonstrated its potential to enhance the reasoning abilities of LLMs through Chain-of-Tho

researcharxiv-cs-cl
14 Apr 2026
Model Releases

Personalized RewardBench: Evaluating Reward Models with Human Aligned Personalization

DGX agent

arXiv:2604.07343v1 Announce Type: cross Abstract: Pluralistic alignment has emerged as a critical frontier in the development of Large Language Models (LLMs), with reward models (RMs) serving as a cen

model-releasesarxiv-cs-lg
10 Apr 2026
Model Releases

Cross-Model Humor Preference Modeling with Cards Against Humanity

DGX agent

arXiv:2608.07481v1 Announce Type: cross Abstract: This paper investigates whether one large language model can approximate the humor preferences of another in a controlled Cards Against Humanity-style

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Multi-Level Modeling of Large Language Model Inference Latency and Energy via Hybrid Analytical--Machine-Learning Predictors

DGX agent

arXiv:2608.06723v1 Announce Type: cross Abstract: The rapid scaling of Large Language Models (LLMs) has significantly increased computational cost, energy consumption, and inference latency, making ac

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

After evaluating one of our upcoming models, Astra, we're treating it as our first 'critical' model for cybersecurity under our Preparedness…

DGX agent

After evaluating one of our upcoming models, Astra, we're treating it as our first 'critical' model for cybersecurity under our Preparedness Framework. This is a scenario we've planned for, and we're

model-releasesopenai--x
7 Aug 2026
Model Releases

DeepSeek-V4-Flash-0731 is Ollama's fastest growing model ever in token usage. We are scaling capacity in US & Europe. On Ollama, this model …

DGX agent

DeepSeek-V4-Flash-0731 is Ollama's fastest growing model ever in token usage. We are scaling capacity in US & Europe. On Ollama, this model runs with high performance (100tps+) and zero data retention

model-releasesollama--x
4 Aug 2026
Model Releases

Mergeable Model-Side Aggregation States for Long-Context Language Models

DGX agent

arXiv:2607.26448v1 Announce Type: new Abstract: A known limitation of long-context language models is their increasingly unreliable performance in non-additive, set-based aggregation as context length

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

HyWorldVLA: A Vision-Language-Action Model with Hybrid World Modeling for Autonomous Driving

DGX agent

arXiv:2607.20988v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models augmented with world modeling represent a promising paradigm for end-to-end autonomous driving. While pixel-level

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Agentic orchestration: Enterprise AI organizations have a deployment problem, not a platform problem — and most are calling chatbots agents

DGX agent

Across 101 enterprises, agent orchestration is consolidating onto model-provider platforms — Anthropic’s Claude leads by a wide margin — chosen for the gravity of the underlying model and judged on re

model-releasesventurebeat-ai
15 Jul 2026
Model Releases

Kimi K3 in the next few hours. Deepseek V4 GA later in the week. New Liquid models. New Mistral models sometime this month. And some rumours suggest GLM 5.5 is coming in August. Openweight AI is eating good.

DGX agent

dam bois we eating good this week ngl, The velocity of the open_weight ecosystem right now is hitting a point where proprietary, closed-source APIs are losing their leverage on compute intelligence. W

model-releasesr-localllama
14 Jul 2026
Model Releases

🚀langchain launches this week: all about open source models and memory! First: open source models. We partnered with @NVIDIAAI to launch a …

DGX agent

🚀langchain launches this week: all about open source models and memory! First: open source models. We partnered with @NVIDIAAI to launch a NemoClaw DeepAgents blueprint. This pairs Deep Agents (our op

model-releasesharrison-chase--x
10 Jul 2026
Model Releases

Protocol Models: Scaling Decentralized Training with Communication-Efficient Model Parallelism

DGX agent

arXiv:2506.01260v2 Announce Type: replace Abstract: Scaling models has led to significant advancements in deep learning, but training these models in decentralized settings remains challenging due to

model-releasesarxiv-cs-lg
9 Jul 2026
Research

How Can AI Find My Model? A Model-Finding Experimental Study Considering Data Formats, Embeddings, and Retrieval Strategies

DGX agent

arXiv:2606.30846v1 Announce Type: new Abstract: Discovering simulation models for reuse remains a fundamental challenge in Modeling and Simulation (M&S). When many models coexist, identifying those th

researcharxiv-cs-ai
1 Jul 2026
Model Releases

OptiMUS-0.3: Using Large Language Models to Model and Solve Optimization Problems at Scale

DGX agent

arXiv:2407.19633v4 Announce Type: replace Abstract: Optimization problems are pervasive in sectors from manufacturing and distribution to healthcare. However, most such problems are still solved heuri

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Thanks for running our open-source work on current frontier models “The results are: the most capable models today (GPT-5.5 Pro) did outperf…

DGX agent

Thanks for running our open-source work on current frontier models “The results are: the most capable models today (GPT-5.5 Pro) did outperform the best models from before (79/100 vs 69/100), but did

model-releasesgary-marcus--x
27 Jun 2026
Model Releases

UBS says 60% of companies now watching AI budgets are moving to cheaper models and open-source Chinese models The pressure is coming from ex…

DGX agent

UBS says 60% of companies now watching AI budgets are moving to cheaper models and open-source Chinese models The pressure is coming from extreme bills, including users spending up to $35K/month, team

model-releasesclem-delangue--x
26 Jun 2026
Model Releases

Video models are hard. Real-time video models are even harder. Loved this breakdown by @itunpredictable on how we productionized Runway Char…

DGX agent

Video models are hard. Real-time video models are even harder. Loved this breakdown by @itunpredictable on how we productionized Runway Characters, our real-time interactive avatar model. We had to fi

model-releasescristobal-valenzuela--x
25 Jun 2026
Model Releases

BIM-Edit: Benchmarking Large Language Models for IFC-Based Building Information Modeling

DGX agent

arXiv:2606.20146v3 Announce Type: replace Abstract: Large language models (LLMs) are increasingly applied to computer-aided design (CAD) to generate design artifacts from textual instructions. In engi

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

All these new models landing this year but Flux Klein 9b FP8 has spoiled me. All I care about now is whether a new model can edit and be used on an 8GB GPU.

DGX agent

This Reddit post discusses user preferences for AI image generation models in 2026, expressing that despite numerous new model releases, the Flux Klein 9b FP8 model has become their benchmark for what

model-releasesr-stablediffusion
22 Jun 2026
Model Releases

Very impressive from GLM-5.2. Frontier open-weight model indeed. Now, can we get a Gemini model in the top 3 soon?

DGX agent

Very impressive from GLM-5.2. Frontier open-weight model indeed. Now, can we get a Gemini model in the top 3 soon? GLM 5.2 is now on DeepSWE as the top open-source model on our leaderboard. With a pas

model-releasesdair-ai--x
21 Jun 2026
Model Releases

NEW: Anthropic introduces Claude Fable 5, a Mythos-class model for general use. Beginning of a new class of frontier models.

DGX agent

NEW: Anthropic introduces Claude Fable 5, a Mythos-class model for general use. Beginning of a new class of frontier models. Introducing Claude Fable 5: a Mythos-class model that we’ve made safe for g

model-releasesdair-ai--x
9 Jun 2026
Safety

Teach a Reward Model to Correct Itself: Reward Guided Adversarial Failure Discovery for Robust Reward Modeling

DGX agent

arXiv:2507.06419v3 Announce Type: replace Abstract: Reward modeling (RM), which captures human preferences to align large language models (LLMs), is increasingly employed in tasks such as model finetu

safetyarxiv-cs-cl
8 Jun 2026
Model Releases

Speculative Thinking: Enhancing Small-Model Reasoning with Large Model Guidance at Inference Time

DGX agent

arXiv:2504.12329v2 Announce Type: replace-cross Abstract: Recent advances leverage post-training to enhance model reasoning performance, which typically requires costly training pipelines and still su

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

@mteamisloading the models from 6 months ago kinda feel the same like the recently released models. currently not holding my breath for more…

DGX agent

Jeremy Howard comments that large language models released 6 months ago feel comparable in capability to recently released models, suggesting that the pace of improvement in model development may be s

model-releasesjeremy-howard--x
26 May 2026
Model Releases

Decouple Searching from Training: Scaling Data Mixing via Model Merging for Large Language Model Pre-training

DGX agent

arXiv:2602.00747v2 Announce Type: replace-cross Abstract: Determining an effective data mixture is a key factor in Large Language Model (LLM) pre-training, where models must balance general competence

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Cross-Model Consistency of Feature Importance in Electrospinning: Separating Robust from Model-Dependent Features

DGX agent

arXiv:2605.04905v1 Announce Type: new Abstract: Electrospinning is a highly sensitive fabrication process in which small variations in operating parameters can significantly influence fiber morphology

model-releasesarxiv-cs-lg
7 May 2026
Model Releases

we're continuing to see clear examples where a model's harness is a major determinant of overall performance. with the same model, running o…

DGX agent

we're continuing to see clear examples where a model's harness is a major determinant of overall performance. with the same model, running on same task, it's easy to observe very different scores depe

model-releasesharrison-chase--x
6 May 2026
Model Releases

Model-Dowser: Data-Free Importance Probing to Mitigate Catastrophic Forgetting in Multimodal Large Language Models

DGX agent

arXiv:2602.04509v4 Announce Type: replace Abstract: Fine-tuning Multimodal Large Language Models (MLLMs) on task-specific data is an effective way to improve performance on downstream applications. Ho

model-releasesarxiv-cs-cl
5 May 2026
Applications

Mythos seems to be a very capable model based on available information, but it is not a cybersecurity model - it is an advanced general purp…

DGX agent

Mythos seems to be a very capable model based on available information, but it is not a cybersecurity model - it is an advanced general purpose model that happens to be good at cyber because it is goo

applicationsethan-mollick--x
30 Apr 2026
Model Releases

LLM 0.32a0 is a major backwards-compatible refactor

DGX agent

I just released LLM 0.32a0, an alpha release of my LLM Python library and CLI tool for accessing LLMs, with some consequential changes that I've been working towards for quite a while. Previous versio

model-releasessimon-willison
29 Apr 2026
Model Releases

First open-weight model from @poolsideai! Apache license, and available on Ollama to try. 👇👇👇 model page

DGX agent

First open-weight model from @poolsideai! Apache license, and available on Ollama to try. 👇👇👇 model page Today we’re releasing Laguna XS.2, Poolside’s first open-weight model. It’s a 33B total / 3B ac

model-releasesollama--x
28 Apr 2026
Agents

Hermes will now by default use native vision if the main agent model supports it, and you didn't set a different vision auxiliary model! Jus…

DGX agent

Hermes will now by default use native vision if the main agent model supports it, and you didn't set a different vision auxiliary model! Just `hermes update` and it will take affect immediately. You c

agentsnous-research--x
27 Apr 2026
Agents

i haven't seen a model that just works across agent harnesses. seems like it should exist. great opportunity for open-weight models. any tho…

DGX agent

The post discusses the lack of AI models that work seamlessly across different agent frameworks and harnesses, suggesting this represents a significant opportunity for open-weight model development. T

agentsdair-ai--x
22 Apr 2026
Model Releases

We need a new document that AI labs should release with each new model, besides the model card: a sort of changelog I want to see how & in w…

DGX agent

We need a new document that AI labs should release with each new model, besides the model card: a sort of changelog I want to see how & in what way the new model changes, breaks, or improves at a rang

model-releasesethan-mollick--x
17 Apr 2026
Model Releases

We replicated Mythos findings in opencode using public models, not Anthropic's private stack. The moat is moving from model access to valida…

DGX agent

We replicated Mythos findings in opencode using public models, not Anthropic's private stack. The moat is moving from model access to validation: finding vulnerability signal is getting cheaper; turni

model-releasesclem-delangue--x
16 Apr 2026
Model Releases

Routing every task to your largest model burns tokens, adds latency, and inflates costs. @AI21Labs' Maestro Orchestration Meta Model (OMM) i…

DGX agent

Routing every task to your largest model burns tokens, adds latency, and inflates costs. @AI21Labs' Maestro Orchestration Meta Model (OMM) is the layer above your stack that dynamically selects the ri

model-releasesai21-labs--x
9 Apr 2026
Model Releases

On Solomonoff Induction in Large Language Models and the Limits of Self-Improving: The Singularity Is Not Near Without Symbolic Model Synthesis

DGX agent

arXiv:2601.05280v3 Announce Type: replace-cross Abstract: On the one hand, the question of whether large language models (LLMs) are Solomonoff induction estimators has become an explicit question at t

model-releasesarxiv-cs-ai
12 Aug 2026
Agents

What Iterated Self-Feeding Probes of Language Models Measure, and a test that separates the construction from the model

DGX agent

arXiv:2608.10986v1 Announce Type: new Abstract: A growing class of methods probes a language model by feeding it its own output: self-consistency, iterated refinement, agentic loops. We ask what such

agentsarxiv-cs-cl
12 Aug 2026
Model Releases

Model Discovery Agent: LLM-assisted Bayesian experiment design for data-efficient discovery of mechanistic world models

DGX agent

arXiv:2608.09696v1 Announce Type: new Abstract: Predicting the answer to interventional ``what if'' questions --- the outcome of an action never taken --- requires a mechanistic, causal model, not a c

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Beyond Foundation Models: Dimension-Aware Neural Architecture Search with Small-Data Representation Models for Cryocooler Lifetime Prediction

DGX agent

arXiv:2608.06993v1 Announce Type: cross Abstract: Large-scale pretrained time-series models achieve strong results through large-scale pretraining and task-agnostic representation learning, but they r

model-releasesarxiv-cs-ai
10 Aug 2026
← Previous
12345…1238
Next →