AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,504 results
23 Jul 2026

I support open-source models distilling what commercial companies distilled for free from the entire internet. I published distillation for …

ResearchDGX agent

I support open-source models distilling what commercial companies distilled for free from the entire internet. I published distillation for free in 1991 in Europe - this was copied in the US and in Ch

Inside the Model Factory — Eiso Kant, Poolside AI

ToolsDGX agent

Poolside AI’s co‑CEO Eiso Kant explains that a small, top‑research team built a model‑factory capable of training Laguna S, a 118‑billion‑parameter mixture‑of‑experts model that outperforms the ~1 tri

Interval and fuzzy physics-augmented neural networks (iPANN and fPANN) for uncertainty quantification and propagation in constitutive modeling

TutorialsDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2607.20339v1 Announce Type: new Abstract: Constitutive modeling under uncertainty remains a central challenge for reliable mechanics simulations, particularly when the available stress-deformati

Introducing the Qwen-Audio-3.0-TTS. Our latest text-to-speech model, in two flavors: • Flash: real-time interaction • Plus: high-quality gen…

Model ReleasesDGX agent

Introducing the Qwen-Audio-3.0-TTS. Our latest text-to-speech model, in two flavors: • Flash: real-time interaction • Plus: high-quality generation What's new: • Fine-grained inline tags-steer [whispe

Reinforcement Learning for Large Language Model Selective Evidence Adoption from Contaminated Retrieval Results

Model ReleasesDGX agent

arXiv:2607.20090v1 Announce Type: cross Abstract: Retrieval-augmented large language models frequently face contexts that interleave useful evidence with misleading statements or instruction-like cont

RoboInter1.5: A Holistic Intermediate Representation Suite for Embodied World Modeling and Robotic Manipulation

Model ReleasesDGX agent

arXiv:2607.18709v2 Announce Type: replace Abstract: Existing robot datasets remain expensive to curate, embodiment-specific, and insufficiently annotated with the fine-grained structure required for g

Scaling Laws for Hypernetwork-Based Knowledge Injection in Large Language Models

TutorialsDGX agent

arXiv:2607.19604v1 Announce Type: new Abstract: Injecting factual knowledge into large language models (LLMs) reliably and at scale remains an open challenge. Hypernetworks provide a promising solutio

SLAI T-Rex: Full-Parameter Post-training of the DeepSeek-V4 Family on Ascend SuperPOD

Model ReleasesDGX agent

arXiv:2607.20145v1 Announce Type: cross Abstract: Full-parameter post-training of trillion-parameter-scale MoE models introduces substantial system-level challenges for large-scale distributed trainin

Stress Testing Concept Erasure with Large Language Model Agents

SafetyDGX agent

arXiv:2607.17890v2 Announce Type: replace Abstract: Concept erasure aims to remove semantic concepts from a trained generative model and is increasingly important for responsible AI deployment. Howeve

22 Jul 2026

Beyond horizontal expansion, depth is another key trait of Rich Content — testing the model's semantic deconstruction and logical nesting: r…

Model ReleasesDGX agent

Beyond horizontal expansion, depth is another key trait of Rich Content — testing the model's semantic deconstruction and logical nesting: rendering multiple nested interfaces layer by layer within a

Cactus Hybrid: We taught Gemma 4 to know when it's wrong

Model ReleasesDGX agent

Hey HN, Henry & Roman here from Cactus. A small, on-device model is fast and private, but sometimes wrong, but frontier models are getting expensive pretty fast. So, we post-trained Gemma 4 E2B post-t

I wrote about the completely wild incident where OpenAI were testing a new model and it broke out of its sandbox and broke INTO Hugging Face…

Model ReleasesDGX agent

I wrote about the completely wild incident where OpenAI were testing a new model and it broke out of its sandbox and broke INTO Hugging Face to steal the answers to the benchmark https://simonwillison

🎨 Meet Qwen-Image-3.0 — the third generation of our foundational image generation model. If 1.0 was about 'Precision,' and 2.0 added 'Varie…

Model ReleasesDGX agent

🎨 Meet Qwen-Image-3.0 — the third generation of our foundational image generation model. If 1.0 was about 'Precision,' and 2.0 added 'Variety, Completeness, Beauty & Authenticity,' then 3.0 comes down

21 Jul 2026

Highly performant open weights frontier models such as Kimi are a competitive threat to OpenAI & Anthropic, but probably for everyone else t…

ResearchDGX agent

Highly performant open weights frontier models such as Kimi are a competitive threat to OpenAI & Anthropic, but probably for everyone else these are a win. Hope more US entities will release top quali

My OCR model mislabels section titles as body text. Is a CRF the right fix, or am I overcomplicating it? [P]

Model ReleasesDGX agent

Hi everyone, I'm working on extracting the hierarchical structure of long PDF documents (legal/regulatory text, lots of numbered sections) and would like to gather some feedback on my approach before

We're partnering with @huggingface to investigate an unprecedented security incident. Cyber-capable OpenAI models compromised Hugging Face p…

Model ReleasesDGX agent

We're partnering with @huggingface to investigate an unprecedented security incident. Cyber-capable OpenAI models compromised Hugging Face production during a benchmark evaluation. Sharing preliminary

16 Jul 2026

AI-Augmented Adaptive Digital Twin Modeling for Brain Tumor Evolution Prediction and Treatment Scheduling

ApplicationsDGX agent

arXiv:2607.13877v1 Announce Type: new Abstract: Brain tumor progression exhibits spatially heterogeneous growth, patient-specific treatment response, and complex interactions with surrounding anatomy,

Discrete Diffusion Models: A Unified Framework from Tokenization to Generation

ResearchDGX agent

arXiv:2607.13431v1 Announce Type: cross Abstract: Discrete denoising diffusion models (DDMs) have recently emerged as a compelling alternative to autoregressive (AR) modeling for discrete data, offeri

Look Again Before You Abstain:Budgeted Conformal Evidence Acquisition for Reliable Vision-Language Model

Model ReleasesDGX agent

arXiv:2606.16667v2 Announce Type: replace Abstract: Large vision-language models (LVLMs) hallucinate: they assert visual details that the image does not support. A principled remedy is selective predi

Protective Capacity Hallucination: When Large Language Models Claim Nonexistent Capabilities

SafetyDGX agent

arXiv:2607.13596v1 Announce Type: cross Abstract: When cast as the protector of a vulnerable user yet given no explicit capability boundary, a large language model (LLM) may respond not by acknowledgi

Tactile Modality Fusion for Vision-Language-Action Models

ResearchDGX agent

arXiv:2603.14604v2 Announce Type: replace-cross Abstract: We propose TacFiLM, a lightweight modality-fusion approach that integrates visual-tactile signals into vision-language-action (VLA) models. Wh

The Dynamic Verifiable Multi-Agent Human Agentic Loyalty Loop (DVM-HALL) Model and the Net Human-Agent Score (NHAS) in Autonomous Commerce

Model ReleasesDGX agent

arXiv:2607.13998v1 Announce Type: cross Abstract: The rapid proliferation of Agentic Artificial Intelligence fundamentally disrupts traditional customer loyalty paradigms. As AI evolves from passive r

15 Jul 2026

A closer look at improved intelligence in GPT-Live: the model can keep a conversation going while helping with multiple tasks at once, like …

Model ReleasesDGX agent

A closer look at improved intelligence in GPT-Live: the model can keep a conversation going while helping with multiple tasks at once, like checking flights, pulling up local weather, and shaping an i

Action-Aware Generative Sequence Modeling for Short Video Recommendation

TutorialsDGX agent

arXiv:2604.25834v2 Announce Type: replace Abstract: With the rapid development of the Internet, users have increasingly higher expectations for the recommendation accuracy of online content consumptio

Adaptive Compute in Latent World Models: When Depth Helps, Hurts, or Doesn't Matter

ResearchDGX agent

arXiv:2607.10203v2 Announce Type: replace-cross Abstract: Adaptive-compute world models -- early-exit or mixture-of-depths predictors that spend variable depth per step -- assume depth buys better pre

AI agents are already being used to improve the capabilities of our next-generation models. We believe with GPT-Red that we have started to …

SafetyDGX agent

AI agents are already being used to improve the capabilities of our next-generation models. We believe with GPT-Red that we have started to unlock a similar flywheel for safety, where today's models c

DM-KG: A Novel Method for Boosting Spatial Cognition of Vision-Language Models in Street View Imagery

TutorialsDGX agent

arXiv:2607.12319v1 Announce Type: new Abstract: As vision-language models (VLMs) are increasingly deployed in geospatial question answering and visual scene understanding, improving their spatial cogn

Do LLMs Know What They Know? Measuring Metacognitive Efficiency with Signal Detection Theory

Model ReleasesDGX agent

arXiv:2603.25112v2 Announce Type: replace-cross Abstract: Standard evaluation of LLM confidence relies on calibration metrics (ECE, Brier score) that conflate how much a model knows (Type-1 accuracy)

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models

ResearchDGX agent

arXiv:2509.22415v3 Announce Type: replace-cross Abstract: Multimodal large language models (MLLMs) have achieved strong vision-language performance, yet their token-level visual evidence remains diffi

Gaussian Mixture Modeling for Event-Aware Visual Allocation in Long Video Understanding

ResearchDGX agent

arXiv:2607.12557v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) face significant challenges in long video understanding due to the excessive computational cost and information los

How to solve PostgreSQL multilingual full-text search limitations with AlloyDB AI

Model ReleasesDGX agent

AlloyDB powers enterprise-grade search for some of the largest organizations, providing robust hybrid search capabilities that combine text, vector, and keyword searches into a simple ranked SQL query

Modeling Story Expectations: A Generative Framework using LLMs

ResearchDGX agent

arXiv:2412.15239v4 Announce Type: replace-cross Abstract: Consumers' engagement with stories is shaped by their expectations about what will happen next, yet modeling these forward-looking beliefs ove

Robustness of Deep Learning Models for PV Power Forecasting under NWP Forecast Errors: A Spatiotemporal and Physically Interpretable Analysis

ResearchDGX agent

arXiv:2607.12954v1 Announce Type: cross Abstract: Engineering use of AI forecasting models requires not only high nominal accuracy but also predictable behavior under uncertain inputs. In photovoltaic

Self-Consistent Flow: Unifying Velocity and Endpoint Prediction for Rectified Flow Models

Local AiDGX agent

arXiv:2607.12171v1 Announce Type: cross Abstract: In rectified-flow-based generative models, the neural network can be trained to predict two different targets, such as the instantaneous velocity or t

Spectral Diffusion Processes

ResearchDGX agent

arXiv:2209.14125v3 Announce Type: replace-cross Abstract: Diffusion models have proven to be a flexible and effective framework for modelling probability distributions on finite-dimensional spaces. Ho

The Seriality Gap in Video Diffusion Models

ResearchDGX agent

arXiv:2607.13031v1 Announce Type: cross Abstract: When one ball strikes another, then another, video models should predict the consequences of each bounce. In controlled experiments on multi-ball hard

Thinking Machines just dropped a ~1T Omni model 🔥 > 1M context window, trained on 48T tokens of image, text, audio 🤯 > comes with a drafte…

Model ReleasesDGX agent

Thinking Machines just dropped a ~1T Omni model 🔥 > 1M context window, trained on 48T tokens of image, text, audio 🤯 > comes with a drafter, NVFP4 weights and Unsloth quants > transformers, llama.cpp

Visual Access Boundaries in Vision-Language Model Reasoning

ResearchDGX agent

arXiv:2607.12815v1 Announce Type: new Abstract: Chain-of-Thought (CoT) prompting is widely used as a test-time scaling strategy for Vision-Language Models (VLMs), but it remains unclear what is extend

We scaled a robot model natively to 8,000 timesteps of context, 5 minutes worth of muscle memory, with constant inference cost. Robot polici…

ResearchDGX agent

We scaled a robot model natively to 8,000 timesteps of context, 5 minutes worth of muscle memory, with constant inference cost. Robot policies used to live their lives a few frames at a time (< 0.1 se

Who touches every token that flows through @anthropic? It’s not any one model, but it’s @katelyn_lesse, @angjiang and the platform team. A y…

Model ReleasesDGX agent

Who touches every token that flows through @anthropic? It’s not any one model, but it’s @katelyn_lesse, @angjiang and the platform team. A year ago, it was just a messages API. Today, their platform s

14 Jul 2026

Claude at scale on Google Cloud: Frontier AI, built for enterprise production

Model ReleasesDGX agent

Running frontier AI in production is demanding — accelerators to manage, latency to hold steady across continents, regulated data to keep in-region, and long-context requests to serve reliably. Claude

cosign. models have overtuned to this now and do not realize when the agentsmd is out of date and should be changed/ignored. last night i go…

Model ReleasesDGX agent

cosign. models have overtuned to this now and do not realize when the agentsmd is out of date and should be changed/ignored. last night i goaled 5.6 sol to complete a 5 stage task and woke up to find

Grok is the Flow LLM. Grok 4.5’s biggest advantage is its speed. It’s smart enough to be comparable to the other models on most things. But …

ApplicationsDGX agent

Grok is the Flow LLM. Grok 4.5’s biggest advantage is its speed. It’s smart enough to be comparable to the other models on most things. But that speed allows you to make little tweaks to your system s

Local AI is the future. Learning how to run Opensource models (Inference), how to evaluate them systematically (Evals), and how to customize…

Local AiDGX agent

Local AI is the future. Learning how to run Opensource models (Inference), how to evaluate them systematically (Evals), and how to customize them (Fine-tuning / RL / Post-training) are invaluable skil

New research from Google DeepMind on effective model routing. LLM routers get judged on accuracy and cost. Both can look great while the rou…

TutorialsDGX agent

New research from Google DeepMind on effective model routing. LLM routers get judged on accuracy and cost. Both can look great while the router is meaningless. If every model in your society responds

Our free monthly DevRel webinar series kicks off this Thursday, 10am PST. Fine-tune an open vision model to extract clean, structured JSON f…

ToolsDGX agent

Our free monthly DevRel webinar series kicks off this Thursday, 10am PST. Fine-tune an open vision model to extract clean, structured JSON from messy receipt images, managed start to finish on Firewor

Today, we’re announcing Bonsai 27B: the first 27B-class model to run on a phone. Bonsai 27B is the new multimodal flagship of the Bonsai fam…

Local AiDGX agent

Today, we’re announcing Bonsai 27B: the first 27B-class model to run on a phone. Bonsai 27B is the new multimodal flagship of the Bonsai family. Based on Qwen3.6 27B, it brings a new capability tier t

uhm this gpt 5.6 launch might be the openai's most successful model ever since... since chatgpt? this is IPO altering stuff going on here

Model ReleasesDGX agent

uhm this gpt 5.6 launch might be the openai's most successful model ever since... since chatgpt? this is IPO altering stuff going on here Did... Codex just overtake Claude Code? 24.5 hours ago Tibo an

13 Jul 2026

One thing I am kind of surprised by is that full multi-modal (any-any) models have not become a bigger deal. It seems Google is the only Lab…

ApplicationsDGX agent

One thing I am kind of surprised by is that full multi-modal (any-any) models have not become a bigger deal. It seems Google is the only Lab releasing these, OpenAI uses selective multimodal capabilit

11 Jul 2026

Free trial of Grok 4.5 model via Grok Build CLI

IndustryDGX agent

Elon Musk announced the availability of a free trial for the Grok 4.5 model through the Grok Build CLI, allowing developers to access and test the latest version of xAI's Grok language model. This ann

10 Jul 2026

Glad to share what I've been building in public. A lightweight AI news feed. We also finetuned a model to write new stories better than Clau…

Model ReleasesDGX agent

Glad to share what I've been building in public. A lightweight AI news feed. We also finetuned a model to write new stories better than Claude or GPT with extensive prompting (believe me not for a lac

One of the most confusing aspects of GPT-5.6 is figuring out which model to use at which reasoning effort - sounds like Sol on Medium might …

Model ReleasesDGX agent

One of the most confusing aspects of GPT-5.6 is figuring out which model to use at which reasoning effort - sounds like Sol on Medium might be a good new default for coding work, if it's an upgrade fr

Persona Cartography: Charting Language Model Personality Traits in Weight Space

SafetyDGX agent

arXiv:2607.07916v1 Announce Type: new Abstract: Large language models exhibit recurring behavioural patterns -- personas -- that shape generalisation and safety, but we lack reliable tools for decompo

Rethinking LLM-as-a-Judge: Representation-as-a-Judge with Small Language Models via Semantic Capacity Asymmetry

ResearchDGX agent

arXiv:2601.22588v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are widely used as reference-free evaluators via prompting, but this 'LLM-as-a-Judge' paradigm is costly, opaque,

SQuaD-SQL: Efficient Text-to-SQL with Small Language Models via LLM-Guided Knowledge Distillation

Model ReleasesDGX agent

arXiv:2607.08161v1 Announce Type: new Abstract: Text-to-SQL is a fundamental task in natural language processing that enables users to interact with structured databases using natural language. While

Structured Pruning of Large Language Models via Power Transformation and Sign-Preserving Score Aggregation with Adaptive Feature Retention

Model ReleasesDGX agent

arXiv:2607.08027v1 Announce Type: cross Abstract: This paper proposes an improved structured pruning method for large language models (LLMs) that addresses key challenges in adapting Adaptive Feature

This time it is novel math proofs with a public model (most of the other big math breakthroughs have been with experimental LLMs).

Model ReleasesDGX agent

This time it is novel math proofs with a public model (most of the other big math breakthroughs have been with experimental LLMs). Yesterday, we made GPT-5.6 Sol Ultra generally available. Today, we'r

9 Jul 2026

An Hybrid Quantum-Classical Diffusion Model for Image Generation

ResearchDGX agent

arXiv:2607.07072v1 Announce Type: new Abstract: Quantum diffusion models provide a physics-consistent route to generative learning by formulating noising and denoising directly on quantum states. Howe

Fast segmentation of watermarked texts from large language models through an epidemic change-point framework

Model ReleasesDGX agent

arXiv:2509.21160v2 Announce Type: replace-cross Abstract: With the growing use of large language models, concerns over content authenticity have spurred a variety of watermarking schemes. These scheme

Generalist Vision-Language Models for Fast Radio Burst detection: a zero-shot benchmark against a specialized detector

Model ReleasesDGX agent

arXiv:2607.07382v1 Announce Type: new Abstract: Fast Radio Bursts (FRBs) are millisecond-duration radio transients whose automated detection increasingly relies on highly specialized deep learning mod

← Previous
1…101102103104105…1009
Next →