AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
59,399 results
10 Aug 2026

Social World Models

Model ReleasesDGX agent

arXiv:2509.00559v3 Announce Type: replace Abstract: Humans intuitively navigate social interactions by simulating unspoken dynamics and reasoning about others' perspectives, even with limited informat

Surg-UniWorld: A Unified Surgical World Model with Multimodal Control Experts

Model ReleasesDGX agent

arXiv:2608.06770v1 Announce Type: new Abstract: Controllable surgical world models can provide a generative foundation for surgical artificial intelligence and simulation by synthesizing realistic ins

7 Aug 2026

DeepSeek-V4-Flash-0731 is now fully rolled out as the new default for deepseek-v4-flash on Ollama's cloud. This model combines speed, effici…

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

DeepSeek-V4-Flash-0731 is now fully rolled out as the new default for deepseek-v4-flash on Ollama's cloud. This model combines speed, efficiency, and frontier-level performance. Fast: 120+ output tps

Everyone should pay attention to the training timelines in this video. OpenAI shares they started training a new internal model May 7. That …

Model ReleasesDGX agent

Everyone should pay attention to the training timelines in this video. OpenAI shares they started training a new internal model May 7. That is more than two months before they released GPT-5.6 publicl

LabyrinthBench: a local-focused, judge-free LLM benchmark that measures context recall under interference for multi-step agentic tasks.

Model ReleasesDGX agent

LabyrinthBench measures the thing that actually kills long agent runs — whether a model can still use what it learned twenty turns ago — deterministically, with no LLM judge, on your own hardware, wit

5 Aug 2026

M-GATE: Multilingual Grammar, Accuracy in Translation, and Efficiency Benchmark for Large Language Models

Model ReleasesDGX agent

arXiv:2608.03803v1 Announce Type: new Abstract: Multilingual language models are deployed across a hundred or more languages, yet most benchmarks test whether a model can perform a task _in_ a languag

4 Aug 2026

Cross-Task Dissociation in Frontier Vision-Language Model Theory of Mind

ResearchDGX agent

arXiv:2608.00261v1 Announce Type: new Abstract: Do frontier vision-language models present a coherent Theory-of-Mind (ToM) profile across tasks, matching the same human reference group, or does that p

Empirical investigation of 3D CT Foundation Models and Unsupervised Adaptation for Head and Neck Cancer Recurrence Prediction

ResearchDGX agent

arXiv:2608.00071v1 Announce Type: new Abstract: The rapid emergence of 3D CT foundation models has opened new avenues for predictive modeling from CT imaging, offering a compelling alternative to trad

3 Aug 2026

FriendBench: Benchmarking Dyadic Familiarity Inference in Humans and Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2607.29602v1 Announce Type: cross Abstract: Reading a social situation often depends on behavior, not words alone. We introduce FriendBench, a benchmark for inferring whether two people are alre

WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning

SafetyDGX agent

arXiv:2607.29613v1 Announce Type: cross Abstract: Reinforcement learning (RL) post-training of Vision-Language-Action (VLA) models has shown strong promise for robotic manipulation. Among RL methods,

2 Aug 2026

All other models on the Portal remain 20% discounted, aside from GPT-5.6 Terra and Luna which are 50% off. https://x.com/NousResearch/status…

Model ReleasesDGX agent

All other models on the Portal remain 20% discounted, aside from GPT-5.6 Terra and Luna which are 50% off. https://x.com/NousResearch/status/2080039066771337475?s=20 All models are now 20% off for a l

31 Jul 2026

EEG-EditBench: Probing Visual Information in EEG-Image Retrieval Models with Controlled Image Edits

Model ReleasesDGX agent

arXiv:2607.27857v1 Announce Type: new Abstract: Recent EEG-to-image retrieval models have achieved strong performance in identifying viewed images from semantically diverse candidates. Yet such succes

Hand-Object Interaction in the Age of Large Foundation Models:Reconstruction, Generation, and Embodied Transfer

ResearchDGX agent

arXiv:2607.28394v1 Announce Type: new Abstract: Hand-object interaction (HOI) modeling remains challenging because it requires joint reasoning about hand articulation, object geometry, contact, semant

MetaRank: Task-Aware Metric Selection for Model Transferability Estimation

TutorialsDGX agent

arXiv:2511.21007v2 Announce Type: replace Abstract: Selecting an appropriate pre-trained source model is a critical, yet computationally expensive, task in transfer learning. Model Transferability Est

Minimax-H3 video model released, open weights coming in the next few days

Model ReleasesDGX agent

https://x.com/MiniMax_AI/status/2083006198828417501?s=20 Quote from their article: Today, we're launching MiniMax H3, a general-purpose multimodal generation model. H3 understands unified context acro

30 Jul 2026

Mental World Modeling

AgentsDGX agent

arXiv:2607.27201v1 Announce Type: new Abstract: World models enable a predictive substrate for planning and action, yet existing formulations merely answer a physical question: what/where it is, and h

Seeing or Knowing? Visual Context Sensitivity in Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2607.26326v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) achieve strong performance by integrating visual inputs with the rich priors of pretrained language models. How

29 Jul 2026

Visual prompt engineering for video models

ResearchDGX agent

arXiv:2607.25537v1 Announce Type: cross Abstract: In the age of foundation models, a model is only as good as its prompt. For this reason, prompt engineering has become an essential technique for impr

28 Jul 2026

Bayesian Repetition Penalty: A Principled Adjacent-Conditional Framework for Reversing Attention Collapse in Autoregressive Language Models

Model ReleasesDGX agent

arXiv:2607.22694v1 Announce Type: new Abstract: Attention collapse in autoregressive language models -- manifested as repetitive token loops where the model becomes trapped in self-reinforcing attract

Closed-Loop Validation-Repair for Healthcare Interoperability: A Multi-Model Study of Schema Compliance in Clinical LLMs

Model ReleasesDGX agent

arXiv:2607.24371v1 Announce Type: cross Abstract: Healthcare interoperability requires AI systems to produce structured outputs conforming to standardized schemas including ICD-10 for diagnostic codin

DeepLens Diagnosis Agent: Agentic Workflow Design Lets a Small Reasoning Model Compete with Frontier LLMs

Model ReleasesDGX agent

arXiv:2607.22555v1 Announce Type: new Abstract: Medical diagnosis is a multi-stage process: extract facts, consult knowledge, generate a differential analysis, and select the best diagnosis with expla

Evaluating Large Language Models for Symbolic Security Protocol Analysis

Model ReleasesDGX agent

arXiv:2607.20712v1 Announce Type: cross Abstract: Security protocol verification relies on formal tools such as ProVerif and OFMC. This study evaluates whether Large Language Models (LLMs) can perform

Hierarchical Group-Conditional Conformal Risk Control for Selective Prediction in Language Models

Model ReleasesDGX agent

arXiv:2607.24562v1 Announce Type: new Abstract: Large language models serve heterogeneous populations structured by domain, topic difficulty, and linguistic style. Conformal risk control (CRC) gives r

How Context Attribution Handles What the Model Already Knows

Model ReleasesDGX agent

arXiv:2607.23804v1 Announce Type: cross Abstract: Context attribution methods for large language models (LLMs) identify which input context contributes to the model response. Recent works show the ini

LEGO Co-builder: Exploring Fine-Grained Vision-Language Modeling for Multimodal LEGO Assembly Assistants

Model ReleasesDGX agent

arXiv:2507.05515v4 Announce Type: replace Abstract: Vision-language models (VLMs) are facing the challenges of understanding and following multimodal assembly instructions, particularly when fine-grai

WorldDiT: A Unified Diffusion Architecture for World and Action Modeling

Model ReleasesDGX agent

arXiv:2607.23909v1 Announce Type: new Abstract: Many recent robot policies pursue stronger control by using large pretrained vision-language models (VLMs) as the action backbone. We introduce WorldDiT

27 Jul 2026

Unboxing Diffusion Models for the Arts: Interactive Model Bending and Practice-Based Explainability

ResearchDGX agent

arXiv:2607.22428v1 Announce Type: cross Abstract: Explainable AI (XAI) in creative practice can be less about technocentric explanation and more about enabling artists to inspect modify and debug mode

24 Jul 2026

DONDO: Open w2v-BERT Speech-Recognition Base Models for African Languages

Model ReleasesDGX agent

arXiv:2607.21540v1 Announce Type: new Abstract: We present DONDO, a family of open, permissively licensed automatic speech recognition (ASR) base models for African languages, built on the w2v-BERT 2.

FLUX 3 - Real World Models: Towards Multimodal Flow Models as the Backbone of Visual Intelligence

Local AiDGX agent

Introducing FLUX 3. One multi-modal model for Image, Video, Audio and Action-Prediction. Creations are truer to life in every kind of style. Blog Post : https://bfl.ai/blog/flux-3 submitted by /u/pmtt

23 Jul 2026

I trained a 0.5M model on 1B tokens of Fineweb-edu dataset.

Model ReleasesDGX agent

Hi everyone, About a month ago I publish my very first research paper on my neural network architecture called Silia. You can look at the model here: https://huggingface.co/Srijan-Srivastava/Silia-v2

16 Jul 2026

Representation-Based Exploration for Language Models: From Test-Time to Post-Training

Model ReleasesDGX agent

arXiv:2510.11686v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) promises to expand the capabilities of language models, but it is unclear if current RL techniques promote the dis

15 Jul 2026

Learning-enabled Acceleration of Scenario-based Model Predictive Control

ResearchDGX agent

arXiv:2607.12775v1 Announce Type: cross Abstract: Scenario-based model predictive control (SBMPC) is a variant of model predictive control (MPC) that explicitly accounts for uncertainty by optimizing

10 Jul 2026

Blind-Spots-Bench: Evaluating Blind Spots in Multimodal Models

Model ReleasesDGX agent

arXiv:2607.08317v1 Announce Type: new Abstract: Modern AI models achieve strong performance on many established benchmarks, yet they still fail on tasks that humans find almost trivial, such as manipu

Fine-tune NVIDIA Nemotron 3 models with Amazon SageMaker AI serverless model customization

Model ReleasesDGX agent

In this post, we explore what makes the Nemotron 3 architecture unique, walk through the fine-tuning techniques available, and show you step-by-step how to get started with serverless customization us

9 Jul 2026

From system models to class models: An in-context learning paradigm

TutorialsDGX agent

arXiv:2308.13380v3 Announce Type: replace-cross Abstract: Is it possible to understand the intricacies of a dynamical system not solely from its input/output pattern, but also by observing the behavio

8 Jul 2026

An Experimental Design Approach to Evaluating Agentic AI's Autonomous Model Discovery

Model ReleasesDGX agent

arXiv:2607.06413v1 Announce Type: cross Abstract: Large language model coding agents increasingly perform open-ended data modeling and analysis. These agents are stochastic and adaptive, and therefore

Improving LLM-Generated Process Model Quality Through Reinforcement Learning: The Role of Reward Function Design

Model ReleasesDGX agent

arXiv:2607.06175v1 Announce Type: cross Abstract: Large language models (LLMs) can generate BPMN process models from natural-language descriptions, yet supervised fine-tuning (SFT) limits their output

7 Jul 2026

MMEarth-Bench: Global Model Adaptation via Multimodal Test-Time Training

Model ReleasesDGX agent

arXiv:2602.06285v2 Announce Type: replace Abstract: Recent research in geospatial machine learning has demonstrated that models pretrained with self-supervised learning on Earth observation data can p

Operator-on-F complements value-equivalence: a planning-time diagnostic for latent world models

ResearchDGX agent

arXiv:2607.04464v1 Announce Type: cross Abstract: World-model evaluation for model-based reinforcement learning typically asks whether the learned model predicts reward and value well, which can leave

TESSERA v2: Scaling Pixel-wise Earth Foundation Models

Model ReleasesDGX agent

arXiv:2607.03949v1 Announce Type: new Abstract: Pixel-wise Earth-observation (EO) foundation models are now achieving state-of-the-art performance via generated spatial embeddings. However, how these

4 Jul 2026

The July 4th weekend All-In @theallinpod turned into a long argument about who owns the intelligence layer. The besties think enterprises ju…

Model ReleasesDGX agent

The July 4th weekend All-In @theallinpod turned into a long argument about who owns the intelligence layer. The besties think enterprises just woke up to a trap they had been walking into, here's how

3 Jul 2026

Something about this year’s @aiDotEngineer World’s Fair just hit different. Last year was the year of “let the agents rip.” This year was th…

Model ReleasesDGX agent

Something about this year’s @aiDotEngineer World’s Fair just hit different. Last year was the year of “let the agents rip.” This year was the year of realizing that autonomy without structure creates

2 Jul 2026

Toward Cybersecurity-Expert Small Language Models

Model ReleasesDGX agent

arXiv:2510.14113v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are transforming everyday applications, yet deployment in cybersecurity lags due to a lack of high-quality, domai

1 Jul 2026

ClawArena-Team: Benchmarking Subagent Orchestration and Dynamic Workflows in Language-Model Agents

Model ReleasesDGX agent

arXiv:2606.31174v1 Announce Type: new Abstract: Production large language-model (LLM) agents are increasingly deployed not as lone problem-solvers but as managers: a main model creates specialized sub

REMSA: Foundation Model Selection for Remote Sensing via a Constraint-Aware Agent

Model ReleasesDGX agent

arXiv:2511.17442v3 Announce Type: replace-cross Abstract: Foundation Models (FMs) are increasingly integrated into remote sensing (RS) pipelines. These models include unimodal vision encoders and mult

30 Jun 2026

On the Vulnerability of Parameter-Level Defenses to Model Merging

Model ReleasesDGX agent

arXiv:2606.30360v1 Announce Type: cross Abstract: The training-free integration of expert models via model merging has exposed significant security risks, enabling free-riders to combine specialized m

StackingNet: Collective Inference Across Independent AI Foundation Models

SafetyDGX agent

arXiv:2602.13792v2 Announce Type: replace Abstract: Artificial intelligence built on large foundation models has transformed language understanding, computer vision, and reasoning, yet these systems r

UCM: Unified Modeling of Camera Control and Memory with Time-aware Positional Encoding Warping for World Models

ApplicationsDGX agent

arXiv:2602.22960v2 Announce Type: replace Abstract: World models based on video generation demonstrate remarkable potential for simulating interactive environments yet suffer from persistent difficult

26 Jun 2026

Benchmarking Open-Weight Foundation Models for Global AI Technical Governance

Model ReleasesDGX agent

arXiv:2606.26099v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed in artificial intelligence (AI) governance analysis across national and international organisat

Highly-recommended reading. Interesting details in this METR's GPT-5.6 eval. They couldn't get a clean capability number because the model c…

Model ReleasesDGX agent

Highly-recommended reading. Interesting details in this METR's GPT-5.6 eval. They couldn't get a clean capability number because the model cheated more than any public model they've tested, and even r

Nemotron-TwoTower: Diffusion Language Modeling with Pretrained Autoregressive Context

Model ReleasesDGX agent

arXiv:2606.26493v1 Announce Type: new Abstract: Diffusion language models offer a promising alternative to autoregressive models due to their potential for parallel and iterative generation. However,

The Capability Frontier: Benchmarks Miss 82% of Model Performance

Model ReleasesDGX agent

arXiv:2606.26836v1 Announce Type: new Abstract: Existing benchmarks typically report accuracy for a single model on a single run. This systematically understates real-world LLM capabilities, particula

25 Jun 2026

Do vision-language models search like humans? Reasoning tokens as a reaction-time analog in classic visual-search paradigms

Model ReleasesDGX agent

arXiv:2606.25066v1 Announce Type: cross Abstract: Visual search has been one of the most productive paradigms in the study of visual attention: the way reaction time scales with the number of items di

24 Jun 2026

Are Text-to-Image Models Inductivist Turkeys? A Counterfactual Benchmark for Causal Reasoning

Model ReleasesDGX agent

arXiv:2606.24548v1 Announce Type: new Abstract: Text-to-image (T2I) generation models have achieved remarkable progress in producing visually realistic images from natural language prompts. Yet it rem

23 Jun 2026

Can Reasoning Models Detect Changes to their Chains of Thought?

ResearchDGX agent

arXiv:2606.22085v1 Announce Type: cross Abstract: There are many reasons one may want to edit a model's chain of thought (CoT) -- e.g., to prefill it with reasoning from a stronger model or to remove

CogFormer: Learn All Your Models Once

TutorialsDGX agent

arXiv:2603.20520v2 Announce Type: replace-cross Abstract: Simulation-based inference (SBI) with neural networks has accelerated and transformed cognitive modeling workflows. SBI enables modelers to fi

Comparative Evaluation of Machine Learning and Deep Learning Models for Wound-Rotor Synchronous Motor Performance Prediction

Model ReleasesDGX agent

arXiv:2606.21230v1 Announce Type: new Abstract: Wound rotor synchronous motors have emerged as a strong alternative that eliminates dependence on REEs. However, WRSM design requires the simultaneous o

Stealthy World Model Manipulation via Data Poisoning

ResearchDGX agent

arXiv:2606.18697v2 Announce Type: replace Abstract: Model-based learning agents use learned world models to predict future states, plan actions, and adapt to new environments. However, the process of

UniviewVLA: A Unified Multiview Vision-Language-Action Model with World Modeling

AgentsDGX agent

arXiv:2606.21501v1 Announce Type: new Abstract: Occluded tasks remain a bottleneck in robot manipulation. Existing solutions either deploy additional physical cameras requiring training-inference came

Vision-language models for chest radiography do not always need the image

Model ReleasesDGX agent

arXiv:2606.17710v2 Announce Type: replace Abstract: Medical vision-language models report strong chest radiograph accuracy, and this is increasingly read as evidence that they use the image. That infe

← Previous
1…1617181920…990
Next →