AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
59,840 results
Agents

Open everything 🔥 Open Harness, Model Choice, Open Memory (take it wherever you need), Open Protocols basically we’re in the middle of a mo…

DGX agent

Open everything 🔥 Open Harness, Model Choice, Open Memory (take it wherever you need), Open Protocols basically we’re in the middle of a model war, they all make great models but the optimal arrangeme

agentsharrison-chase--x
9 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

So what's the deal with Amazon Nova? They released Nova 2 in December, and even then, the top flight Nova 2 model trailed Sonnet 4.5. And it…

DGX agent

Amazon released the Nova 2 model family on December 2, 2025 at AWS re:Invent, comprising Nova 2 Lite, Nova 2 Pro (Preview), Nova 2 Omni, and Nova 2 Sonic, all featuring a 1-million-token context wi...

model-releasesethan-mollick--x
9 Apr 2026
Safety

Constructing Dynamic Master Logic Models as Knowledge Graphs for Complex System Diagnostics Using Retrieval-Augmented Large Language Models

DGX agent

arXiv:2608.12304v1 Announce Type: new Abstract: Dynamic Master Logic (DML) provides a hierarchical framework for representing system behavior by linking functional objectives to underlying structural

safetyarxiv-cs-ai
13 Aug 2026
Model Releases

Interpreting Language Model Hidden States at Scale

DGX agent

arXiv:2608.10260v1 Announce Type: new Abstract: Lens methods interpret large language models (LLMs) by mapping intermediate activations to the output vocabulary, revealing how next-token predictions d

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

Qwen3.8-2.4T-A95B is now live on Together AI. The Qwen Team’s latest flagship model is built for coding and long-horizon agent workflows, wi…

DGX agent

The Qwen Team has released its flagship model, Qwen3.8‑2.4T‑A95B, on the Together AI platform (togethercompute) as of August 12 2026. This 2.4‑trillion‑parameter model is engineered for coding tasks a

model-releasestogether-ai--x
12 Aug 2026
Local Ai

4D-WAM: Infusing Spatiotemporal Awareness into World Action Models through Trajectory Fields

DGX agent

arXiv:2608.08023v1 Announce Type: new Abstract: Building on recent advances in world models, World Action Models (WAMs) jointly model video prediction and action generation. However, they typically re

local-aiarxiv-cs-ro
11 Aug 2026
Safety

Concept-Guided Spatial Regularization for World Models in Atari Pong

DGX agent

arXiv:2607.15142v2 Announce Type: replace Abstract: World models are usually evaluated as components of model-based reinforcement learning (MBRL) systems, leaving their standalone reliability understu

safetyarxiv-cs-ai
11 Aug 2026
Model Releases

I gave DeepSeek V4 Flash basic vision by training a 40M connector on 100K examples

DGX agent

I wanted to find out whether a huge text-only MoE could be given basic vision without retraining the language model itself. The short answer is yes. I froze DeepSeek V4 Flash and a 417M-parameter Moon

model-releasesr-localllama
11 Aug 2026
Model Releases

Matryoshka Language Model Suites

DGX agent

arXiv:2608.09703v1 Announce Type: new Abstract: Training a language model suite classically requires training each model separately and serving them independently. We improve both training and inferen

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Observations on Muse-Glimmer reasoning traces being noticeably different from qwen / gemma models and questions for you guys

DGX agent

Just downloaded the model, UD-Q5_K_XL quant, asked it to generate a long story to test out reasoning and speed with dflash (super fast btw, ~ 90 to 160 tok/s on a 5090 depending on task) and was surpr

model-releasesr-localllama
11 Aug 2026
Model Releases

Two-Layer Linear Auto-Regressive Models Estimate Latent States

DGX agent

arXiv:2606.12691v2 Announce Type: replace-cross Abstract: Auto-regressive models have emerged as powerful tools for sequential data, from language to video. Understanding how and why these models lear

model-releasesarxiv-cs-ai
11 Aug 2026
Agents

Dueling World Models: Advantage-Style Action Channels for Common-Mode Distractor Rejection

DGX agent

arXiv:2608.06706v1 Announce Type: cross Abstract: Latent world models plan by predicting future states from an action, but when a scene contains motion the agent does not control, they quietly go acti

agentsarxiv-cs-ai
10 Aug 2026
Model Releases

GRASP: Reinforcing Language Model Anonymizers with Group Relative Policy Optimization

DGX agent

arXiv:2608.06526v1 Announce Type: new Abstract: Large language models can infer sensitive personal attributes, such as age, location, and occupation, from ordinary text, turning everyday writing into

model-releasesarxiv-cs-cl
10 Aug 2026
Model Releases

MemWM: Memory-Augmented Text-Based World Model

DGX agent

arXiv:2608.07107v1 Announce Type: new Abstract: World models are increasingly used to support planning in agents by predicting how environment states evolve in response to agent actions. Yet fluent ne

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

MI-MIDI: Mechanistic Interpretability of Text-to-MIDI Generation Models via Probing, Lenses and Steering

DGX agent

arXiv:2608.06638v1 Announce Type: cross Abstract: Mechanistic interpretability of music generation has concentrated on audio models, leaving symbolic models largely unexplored. We analyze two public t

model-releasesarxiv-cs-ai
10 Aug 2026
Safety

Progressive Alignment of Recommender Foundation Model through Multi-Phase Post-Training

DGX agent

arXiv:2608.06792v1 Announce Type: cross Abstract: Foundation model(FM) for recommendation has shown strong ability to model long-horizon sequential user behavior. In practice, a single pretrained foun

safetyarxiv-cs-ai
10 Aug 2026
Research

Quantum Generative Diffusion Model: A Fully Quantum-Mechanical Model for Generating Quantum State Ensemble

DGX agent

arXiv:2401.07039v5 Announce Type: replace-cross Abstract: Mixed quantum states are the native description of many physically important quantum systems, making their generation a fundamental task in qu

researcharxiv-cs-lg
10 Aug 2026
Model Releases

Social World Models

DGX agent

arXiv:2509.00559v3 Announce Type: replace Abstract: Humans intuitively navigate social interactions by simulating unspoken dynamics and reasoning about others' perspectives, even with limited informat

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Surg-UniWorld: A Unified Surgical World Model with Multimodal Control Experts

DGX agent

arXiv:2608.06770v1 Announce Type: new Abstract: Controllable surgical world models can provide a generative foundation for surgical artificial intelligence and simulation by synthesizing realistic ins

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

DeepSeek-V4-Flash-0731 is now fully rolled out as the new default for deepseek-v4-flash on Ollama's cloud. This model combines speed, effici…

DGX agent

DeepSeek-V4-Flash-0731 is now fully rolled out as the new default for deepseek-v4-flash on Ollama's cloud. This model combines speed, efficiency, and frontier-level performance. Fast: 120+ output tps

model-releasesollama--x
7 Aug 2026
Model Releases

Everyone should pay attention to the training timelines in this video. OpenAI shares they started training a new internal model May 7. That …

DGX agent

Everyone should pay attention to the training timelines in this video. OpenAI shares they started training a new internal model May 7. That is more than two months before they released GPT-5.6 publicl

model-releasesallie-k--miller--x
7 Aug 2026
Model Releases

LabyrinthBench: a local-focused, judge-free LLM benchmark that measures context recall under interference for multi-step agentic tasks.

DGX agent

LabyrinthBench measures the thing that actually kills long agent runs — whether a model can still use what it learned twenty turns ago — deterministically, with no LLM judge, on your own hardware, wit

model-releasesr-localllama
7 Aug 2026
Model Releases

M-GATE: Multilingual Grammar, Accuracy in Translation, and Efficiency Benchmark for Large Language Models

DGX agent

arXiv:2608.03803v1 Announce Type: new Abstract: Multilingual language models are deployed across a hundred or more languages, yet most benchmarks test whether a model can perform a task _in_ a languag

model-releasesarxiv-cs-cl
5 Aug 2026
Research

Cross-Task Dissociation in Frontier Vision-Language Model Theory of Mind

DGX agent

arXiv:2608.00261v1 Announce Type: new Abstract: Do frontier vision-language models present a coherent Theory-of-Mind (ToM) profile across tasks, matching the same human reference group, or does that p

researcharxiv-cs-cl
4 Aug 2026
Research

Empirical investigation of 3D CT Foundation Models and Unsupervised Adaptation for Head and Neck Cancer Recurrence Prediction

DGX agent

arXiv:2608.00071v1 Announce Type: new Abstract: The rapid emergence of 3D CT foundation models has opened new avenues for predictive modeling from CT imaging, offering a compelling alternative to trad

researcharxiv-cs-cv
4 Aug 2026
Model Releases

FriendBench: Benchmarking Dyadic Familiarity Inference in Humans and Multimodal Large Language Models

DGX agent

arXiv:2607.29602v1 Announce Type: cross Abstract: Reading a social situation often depends on behavior, not words alone. We introduce FriendBench, a benchmark for inferring whether two people are alre

model-releasesarxiv-cs-ai
3 Aug 2026
Safety

WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning

DGX agent

arXiv:2607.29613v1 Announce Type: cross Abstract: Reinforcement learning (RL) post-training of Vision-Language-Action (VLA) models has shown strong promise for robotic manipulation. Among RL methods,

safetyarxiv-cs-cl
3 Aug 2026
Model Releases

All other models on the Portal remain 20% discounted, aside from GPT-5.6 Terra and Luna which are 50% off. https://x.com/NousResearch/status…

DGX agent

All other models on the Portal remain 20% discounted, aside from GPT-5.6 Terra and Luna which are 50% off. https://x.com/NousResearch/status/2080039066771337475?s=20 All models are now 20% off for a l

model-releasesnous-research--x
2 Aug 2026
Model Releases

EEG-EditBench: Probing Visual Information in EEG-Image Retrieval Models with Controlled Image Edits

DGX agent

arXiv:2607.27857v1 Announce Type: new Abstract: Recent EEG-to-image retrieval models have achieved strong performance in identifying viewed images from semantically diverse candidates. Yet such succes

model-releasesarxiv-cs-cv
31 Jul 2026
Research

Hand-Object Interaction in the Age of Large Foundation Models:Reconstruction, Generation, and Embodied Transfer

DGX agent

arXiv:2607.28394v1 Announce Type: new Abstract: Hand-object interaction (HOI) modeling remains challenging because it requires joint reasoning about hand articulation, object geometry, contact, semant

researcharxiv-cs-cv
31 Jul 2026
Tutorials

MetaRank: Task-Aware Metric Selection for Model Transferability Estimation

DGX agent

arXiv:2511.21007v2 Announce Type: replace Abstract: Selecting an appropriate pre-trained source model is a critical, yet computationally expensive, task in transfer learning. Model Transferability Est

tutorialsarxiv-cs-cv
31 Jul 2026
Model Releases

Minimax-H3 video model released, open weights coming in the next few days

DGX agent

https://x.com/MiniMax_AI/status/2083006198828417501?s=20 Quote from their article: Today, we're launching MiniMax H3, a general-purpose multimodal generation model. H3 understands unified context acro

model-releasesr-localllama
31 Jul 2026
Agents

Mental World Modeling

DGX agent

arXiv:2607.27201v1 Announce Type: new Abstract: World models enable a predictive substrate for planning and action, yet existing formulations merely answer a physical question: what/where it is, and h

agentsarxiv-cs-cl
30 Jul 2026
Model Releases

Seeing or Knowing? Visual Context Sensitivity in Multimodal Large Language Models

DGX agent

arXiv:2607.26326v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) achieve strong performance by integrating visual inputs with the rich priors of pretrained language models. How

model-releasesarxiv-cs-cv
30 Jul 2026
Research

Visual prompt engineering for video models

DGX agent

arXiv:2607.25537v1 Announce Type: cross Abstract: In the age of foundation models, a model is only as good as its prompt. For this reason, prompt engineering has become an essential technique for impr

researcharxiv-cs-ai
29 Jul 2026
Model Releases

Bayesian Repetition Penalty: A Principled Adjacent-Conditional Framework for Reversing Attention Collapse in Autoregressive Language Models

DGX agent

arXiv:2607.22694v1 Announce Type: new Abstract: Attention collapse in autoregressive language models -- manifested as repetitive token loops where the model becomes trapped in self-reinforcing attract

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Closed-Loop Validation-Repair for Healthcare Interoperability: A Multi-Model Study of Schema Compliance in Clinical LLMs

DGX agent

arXiv:2607.24371v1 Announce Type: cross Abstract: Healthcare interoperability requires AI systems to produce structured outputs conforming to standardized schemas including ICD-10 for diagnostic codin

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

DeepLens Diagnosis Agent: Agentic Workflow Design Lets a Small Reasoning Model Compete with Frontier LLMs

DGX agent

arXiv:2607.22555v1 Announce Type: new Abstract: Medical diagnosis is a multi-stage process: extract facts, consult knowledge, generate a differential analysis, and select the best diagnosis with expla

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Evaluating Large Language Models for Symbolic Security Protocol Analysis

DGX agent

arXiv:2607.20712v1 Announce Type: cross Abstract: Security protocol verification relies on formal tools such as ProVerif and OFMC. This study evaluates whether Large Language Models (LLMs) can perform

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Hierarchical Group-Conditional Conformal Risk Control for Selective Prediction in Language Models

DGX agent

arXiv:2607.24562v1 Announce Type: new Abstract: Large language models serve heterogeneous populations structured by domain, topic difficulty, and linguistic style. Conformal risk control (CRC) gives r

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

How Context Attribution Handles What the Model Already Knows

DGX agent

arXiv:2607.23804v1 Announce Type: cross Abstract: Context attribution methods for large language models (LLMs) identify which input context contributes to the model response. Recent works show the ini

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

LEGO Co-builder: Exploring Fine-Grained Vision-Language Modeling for Multimodal LEGO Assembly Assistants

DGX agent

arXiv:2507.05515v4 Announce Type: replace Abstract: Vision-language models (VLMs) are facing the challenges of understanding and following multimodal assembly instructions, particularly when fine-grai

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

WorldDiT: A Unified Diffusion Architecture for World and Action Modeling

DGX agent

arXiv:2607.23909v1 Announce Type: new Abstract: Many recent robot policies pursue stronger control by using large pretrained vision-language models (VLMs) as the action backbone. We introduce WorldDiT

model-releasesarxiv-cs-lg
28 Jul 2026
Research

Unboxing Diffusion Models for the Arts: Interactive Model Bending and Practice-Based Explainability

DGX agent

arXiv:2607.22428v1 Announce Type: cross Abstract: Explainable AI (XAI) in creative practice can be less about technocentric explanation and more about enabling artists to inspect modify and debug mode

researcharxiv-cs-lg
27 Jul 2026
Model Releases

DONDO: Open w2v-BERT Speech-Recognition Base Models for African Languages

DGX agent

arXiv:2607.21540v1 Announce Type: new Abstract: We present DONDO, a family of open, permissively licensed automatic speech recognition (ASR) base models for African languages, built on the w2v-BERT 2.

model-releasesarxiv-cs-cl
24 Jul 2026
Local Ai

FLUX 3 - Real World Models: Towards Multimodal Flow Models as the Backbone of Visual Intelligence

DGX agent

Introducing FLUX 3. One multi-modal model for Image, Video, Audio and Action-Prediction. Creations are truer to life in every kind of style. Blog Post : https://bfl.ai/blog/flux-3 submitted by /u/pmtt

local-air-localllama
24 Jul 2026
Model Releases

I trained a 0.5M model on 1B tokens of Fineweb-edu dataset.

DGX agent

Hi everyone, About a month ago I publish my very first research paper on my neural network architecture called Silia. You can look at the model here: https://huggingface.co/Srijan-Srivastava/Silia-v2

model-releasesr-localllama
23 Jul 2026
Model Releases

Representation-Based Exploration for Language Models: From Test-Time to Post-Training

DGX agent

arXiv:2510.11686v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) promises to expand the capabilities of language models, but it is unclear if current RL techniques promote the dis

model-releasesarxiv-cs-ai
16 Jul 2026
← Previous
1…2021222324…1247
Next →