AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
59,378 results
30 Apr 2026

APEX-Agents now has a @huggingface leaderboard for open-source models. APEX-Agents is our frontier benchmark for whether models can do the r…

Model ReleasesDGX agent

APEX-Agents now has a @huggingface leaderboard for open-source models. APEX-Agents is our frontier benchmark for whether models can do the real work of consultants, lawyers, and bankers. https://huggi

What Google Cloud announced in AI this month

Model ReleasesDGX agent

Editor’s note: Want to keep up with the latest from Google Cloud? Check back here for a monthly recap of our latest updates, announcements, resources, events, learning opportunities, and more. We host

28 Apr 2026

Both are great models and neither wins everywhere. I use both Opus and 5.5 depending on the task. LangSmith Fleet lets you choose the model …

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model ReleasesDGX agent

Both are great models and neither wins everywhere. I use both Opus and 5.5 depending on the task. LangSmith Fleet lets you choose the model for each agent, so you can match it to the work https://www.

Improving Vision-language Models with Perception-centric Process Reward Models

Model ReleasesDGX agent

arXiv:2604.24583v1 Announce Type: new Abstract: Recent advancements in reinforcement learning with verifiable rewards (RLVR) have significantly improved the complex reasoning ability of vision-languag

Today we’re releasing Laguna XS.2, Poolside’s first open-weight model. It’s a 33B total / 3B active MoE model built for agentic coding and l…

Model ReleasesDGX agent

Today we’re releasing Laguna XS.2, Poolside’s first open-weight model. It’s a 33B total / 3B active MoE model built for agentic coding and long-horizon tasks. Trained fully in-house on our own stack.

27 Apr 2026

Lightweight Retrieval-Augmented Generation and Large Language Model-Based Modeling for Scalable Patient-Trial Matching

ApplicationsDGX agent

arXiv:2604.22061v1 Announce Type: cross Abstract: Patient-trial matching requires reasoning over long, heterogeneous electronic health records (EHRs) and complex eligibility criteria, posing significa

25 Apr 2026

🚀Meet Carnice-V2-27b🚀 → Carnice is a 27 billion parameter model capable of beating models 10x the size in Hermes-agent, fully open-source …

Model ReleasesDGX agent

🚀Meet Carnice-V2-27b🚀 → Carnice is a 27 billion parameter model capable of beating models 10x the size in Hermes-agent, fully open-source and built on top of Qwen3.6-27B →Build to fit on Consumer GPU

23 Apr 2026

GPT-5.5 is likely the best model in the world. But open models like Kimi and Minimax get almost identical coding benchmark scores at 10-25x …

Model ReleasesDGX agent

GPT-5.5 is likely the best model in the world. But open models like Kimi and Minimax get almost identical coding benchmark scores at 10-25x lower cost. I broke down the benchmarks and pricing. Here's

22 Apr 2026

OpenAI just released a new open-source model it's 'a bidirectional token-classification model for personally identifiable information (PII) …

Model ReleasesDGX agent

OpenAI just released a new open-source model it's 'a bidirectional token-classification model for personally identifiable information (PII) detection and masking in text' https://github.com/openai/pri

The Rise of Verbal Tics in Large Language Models: A Systematic Analysis Across Frontier Models

Model ReleasesDGX agent

arXiv:2604.19139v1 Announce Type: cross Abstract: As Large Language Models (LLMs) continue to evolve through alignment techniques such as Reinforcement Learning from Human Feedback (RLHF) and Constitu

21 Apr 2026

LiFT: Does Instruction Fine-Tuning Improve In-Context Learning for Longitudinal Modelling by Large Language Models?

Model ReleasesDGX agent

arXiv:2604.16382v1 Announce Type: new Abstract: Longitudinal NLP tasks require reasoning over temporally ordered text to detect persistence and change in human behavior and opinions. However, in-conte

We are entering an extremely exciting era for open-weight models. Kimi K2.6 now feels like a top agentic model. I took it for a spin via @Fi…

AgentsDGX agent

We are entering an extremely exciting era for open-weight models. Kimi K2.6 now feels like a top agentic model. I took it for a spin via @FireworksAI_HQ fast inference APIs. Kimi K2.6 has impressive a

What makes ChatGPT Images 2.0 a state-of-the-art image generation model? Researchers behind the model explain. A thread: Thinking & Intellig…

Model ReleasesDGX agent

What makes ChatGPT Images 2.0 a state-of-the-art image generation model? Researchers behind the model explain. A thread: Thinking & Intelligence in ChatGPT Images 2.0, demonstrated by @ayaanzhaque Med

17 Apr 2026

Diagnosing and Improving Diffusion Models by Estimating the Optimal Loss Value

SafetyDGX agent

arXiv:2506.13763v2 Announce Type: replace-cross Abstract: Diffusion models have achieved remarkable success in generative modeling. Despite more stable training, the loss of diffusion models is not in

16 Apr 2026

Alibaba unveils Qwen3.6-35B-A3B, an open-weight MoE model with 35B total and 3B active parameters, saying it rivals larger dense models in agentic coding tasks (Qwen)

Model ReleasesDGX agent

Qwen: Alibaba unveils Qwen3.6-35B-A3B, an open-weight MoE model with 35B total and 3B active parameters, saying it rivals larger dense models in agentic coding tasks — · 4355 words · QwenTeam丨Translat

14 Apr 2026

Sub-32B open weights models now offer GPT-5 level intelligence with Qwen3.5 27B (Reasoning) matching GPT-5 (medium) at 42 and Gemma 4 31B (R…

Model ReleasesDGX agent

Sub-32B open weights models now offer GPT-5 level intelligence with Qwen3.5 27B (Reasoning) matching GPT-5 (medium) at 42 and Gemma 4 31B (Reasoning) matching GPT-5 (low) at 39 on the Artificial Analy

13 Apr 2026

The Roots of Performance Disparity in Multilingual Language Models: Intrinsic Modeling Difficulty or Design Choices?

Model ReleasesDGX agent

arXiv:2601.07220v3 Announce Type: replace Abstract: Multilingual language models (LMs) promise broader NLP access, yet current systems deliver uneven performance across the world's languages. This sur

XFED: Non-Collusive Model Poisoning Attack Against Byzantine-Robust Federated Classifiers

Model ReleasesDGX agent

arXiv:2604.09489v1 Announce Type: cross Abstract: Model poisoning attacks pose a significant security threat to Federated Learning (FL). Most existing model poisoning attacks rely on collusion, requir

11 Apr 2026

Relying on model providers' stateful APIs or harnesses creates lock-in: switching models means losing your agent's memory -- a cost that onl…

AgentsDGX agent

Relying on model providers' stateful APIs or harnesses creates lock-in: switching models means losing your agent's memory -- a cost that only grows as agents get better at learning a big part of agent

10 Apr 2026

HY-Embodied-0.5: Embodied Foundation Models for Real-World Agents

Model ReleasesDGX agent

arXiv:2604.07430v1 Announce Type: new Abstract: We introduce HY-Embodied-0.5, a family of foundation models specifically designed for real-world embodied agents. To bridge the gap between general Visi

The End of the Foundation Model Era: Open-Weight Models, Sovereign AI, and Inference as Infrastructure

SafetyDGX agent

arXiv:2604.06217v1 Announce Type: cross Abstract: The foundation model era -- roughly 2020 to 2025 -- is over. The forces that defined it have inverted. Open source models have reached frontier perfor

You Point, I Learn: Online Adaptation of Interactive Segmentation Models for Handling Distribution Shifts in Medical Imaging

Model ReleasesDGX agent

arXiv:2503.06717v3 Announce Type: replace Abstract: Interactive segmentation uses real-time user inputs, such as mouse clicks, to iteratively refine model predictions. Although not originally designed

8 Apr 2026

We’re partnering with @MiniMax_AI across product and models to make their upcoming releases the best for Hermes Agent users. MiniMax models …

AgentsDGX agent

We’re partnering with @MiniMax_AI across product and models to make their upcoming releases the best for Hermes Agent users. MiniMax models are already some of the most-used in Hermes Agent. If you ha

12 Aug 2026

A model is only as good as its data, and we’ve long since exhausted the internet. From here on out, model progress is gated by data producti…

AgentsDGX agent

A model is only as good as its data, and we’ve long since exhausted the internet. From here on out, model progress is gated by data production. @mercor_ai’s @BrendanFoody joined us at our Sovereign AI

ChronoSSM: Training for Temporally Aware Representations in Autoregressive State Space Models

ResearchDGX agent

arXiv:2608.10120v1 Announce Type: new Abstract: Modern sequence models, from Transformers to State Space Models, have enabled powerful generative modeling across diverse domains, yet they are typicall

11 Aug 2026

OpenAI just launched a cybersecurity model that answers 95% of advanced threat queries. And Meta put a frontier model on your laptop. Same day.

Model ReleasesDGX agent

Something happened today that I think most people are going to miss because there are two separate stories and neither one is getting the full picture. OpenAI expanded Daybreak. If you haven't heard o

10 Aug 2026

Not All Problems Are Best Modeled as MILP: A DSL-Centric Framework for Flexible and Accurate Optimization Modeling

Model ReleasesDGX agent

arXiv:2608.07040v1 Announce Type: new Abstract: Solving combinatorial optimization problems (COPs) requires not only efficient algorithms but also carefully crafted formulations. While recent works ha

9 Aug 2026

Best Embedding + Reranking Model

Model ReleasesDGX agent

What Local Embedding + Reranking Models are you guys running for RAG? I went down this rabbit hole because I wanted a Embedding Model + Reranker for a Translation Memory Server. Essentially, given X p

7 Aug 2026

Woodpecker Distillation: Weak Models Diagnose Reasoning Bugs in Strong Models

Local AiDGX agent

arXiv:2608.05168v1 Announce Type: new Abstract: Large language models often fail on reasoning tasks despite possessing the capability to solve them. We argue that many such failures arise from localiz

6 Aug 2026

A Unified Model for Cross-Domain Clone Detection via Model Merging

Model ReleasesDGX agent

arXiv:2608.04215v1 Announce Type: cross Abstract: The growing diversity of code clone types, from syntactic copies to cross-language semantic clones to AI-generated duplicates, has created a fragmenta

Large-Small Model Collaboration for Enhancing Edge-Deployed Small Models

Model ReleasesDGX agent

arXiv:2503.10367v2 Announce Type: replace-cross Abstract: Edge devices host domain-specific small language models (SLMs) with limited resources, while private clouds offer larger LLMs. We propose G-Bo

3 Aug 2026

Language Models Agree With Each Other, Not With Readers

Model ReleasesDGX agent

arXiv:2607.29274v1 Announce Type: cross Abstract: Claims that language models homogenise are usually measured against human judgements collected for the study, which makes the human side an artifact o

Step-Level Visual Grounding Faithfulness Predicts Out-of-Distribution Generalization in Long-Horizon Vision-Language Models

Model ReleasesDGX agent

arXiv:2603.06828v2 Announce Type: replace-cross Abstract: We uncover a behavioral law of long-horizon vision-language models: models that maintain temporally grounded beliefs generalize better. Standa

Try Qwen3.8-Max on Hermes Agent and you will have to doubt on how much these open frontier models have caught up with frontier closed models…

AgentsDGX agent

Try Qwen3.8-Max on Hermes Agent and you will have to doubt on how much these open frontier models have caught up with frontier closed models. These new open models are insanely good. Meet Qwen3.8-Max:

31 Jul 2026

Epistemic diversity across language models mitigates knowledge collapse

Model ReleasesDGX agent

arXiv:2512.15011v3 Announce Type: replace Abstract: Artificial intelligence (AI) increasingly generates the very content used to train future AI systems. This feedback loop can degrade model quality,

30 Jul 2026

DenseOn with the LateOn: Fully Open Dense and Late-Interaction Models for Multilingual, Long-Context, and Code Search

Model ReleasesDGX agent

arXiv:2607.27178v1 Announce Type: new Abstract: State-of-the-art retrieval models increasingly rely on closed training data, creating a reproducibility gap. We present an open end-to-end recipe for tr

29 Jul 2026

AI Security Leaderboard: benchmarking model robustness [P]

Model ReleasesDGX agent

We developed a leaderboard ranking frontier model security. There's no shortage of model capability rankings, but we didn't find anything comparable for model security. Yet security is becoming increa

CHARM: A Multimodal Graph Foundation Model with Hierarchical Context Modeling for Zero-Shot Transfer

ApplicationsDGX agent

arXiv:2607.26023v1 Announce Type: new Abstract: Graph foundation models (GFMs) have emerged as a promising paradigm for transferring knowledge across graph domains and tasks. Real-world graphs associa

25 Jul 2026

I released Inflect v2: two ultra-tiny complete TTS models under 4M and 10M parameters

Model ReleasesDGX agent

I’ve spent the past month trying to find the point where an extremely small TTS model stops feeling like a size experiment and starts feeling genuinely useful. Today I’m releasing Inflect v2, with two

24 Jul 2026

Sources: OpenAI's models breached Hugging Face from July 11 to 13 and OpenAI realized their models were behind the hack several days later (Reuters)

AgentsDGX agent

Reuters: Sources: OpenAI's models breached Hugging Face from July 11 to 13 and OpenAI realized their models were behind the hack several days later — The OpenAI agent that broke into tech firm Hugging

16 Jul 2026

PersGuard: Preventing Malicious Personalization in Text-to-Image Diffusion Models via Model Backdoors

ResearchDGX agent

arXiv:2502.16167v2 Announce Type: replace-cross Abstract: Diffusion models (DMs) have advanced text-to-image (T2I) synthesis, yet their personalization capabilities raise serious privacy and copyright

15 Jul 2026

The One-Word Census: Answer-Choice Conformity Across 44 Language Models

Model ReleasesDGX agent

arXiv:2607.12796v1 Announce Type: cross Abstract: When a language model must pick one answer from a large space of equally valid options, which does it pick -- and how often is it the same answer ever

Xray-Visual Models: Scaling Vision models on Industry Scale Data

ApplicationsDGX agent

arXiv:2602.16918v2 Announce Type: replace-cross Abstract: We present Xray-Visual, a unified vision model architecture for large-scale image and video understanding trained on industry-scale social med

9 Jul 2026

Since I am needling every model maker tonight about minor but important issues, one more: Grok 4.5 has no model card. Companies that are try…

ApplicationsDGX agent

Since I am needling every model maker tonight about minor but important issues, one more: Grok 4.5 has no model card. Companies that are trying to compete near the frontier should be releasing model c

4 Jul 2026

Better Models: Worse Tools

Model ReleasesDGX agent

Better Models: Worse Tools Armin reports on a weird problem he ran into while hacking on Pi: The short version is that newer Claude models sometimes call Pi’s edit tool with extra, invented fields in

3 Jul 2026

A Survey of Circuit Foundation Model: Foundation AI Models for VLSI Circuit Design and EDA

TutorialsDGX agent

arXiv:2504.03711v2 Announce Type: replace-cross Abstract: Artificial intelligence (AI)-driven electronic design automation (EDA) techniques have been extensively explored for VLSI circuit design appli

Predicting Closed-Loop Performance of Latent World Models: Offline Checkpoint Selection for MPC and Model-Based RL Under Non-Markovian Rewards in LunarLander

SafetyDGX agent

arXiv:2607.01736v1 Announce Type: cross Abstract: We study how to predict the downstream closed-loop performance of a learned latent world model from validation-time diagnostics alone. Choosing the ri

2 Jul 2026

From World Models to World Action Models: A Concise Tutorial for Robotics

SafetyDGX agent

arXiv:2607.00836v1 Announce Type: cross Abstract: World models are increasingly used in embodied intelligence and generative simulation, yet their scope remains ambiguous across communities. This tuto

30 Jun 2026

Deep probabilistic model synthesis enables unified modeling of whole-brain neural activity across individual subjects

TutorialsDGX agent

arXiv:2603.14161v2 Announce Type: replace Abstract: Many disciplines need quantitative models that synthesize experimental data across multiple instances of the same general system. For example, neuro

Developmental Trajectories of Situation Modeling and Mentalizing in Transformer Language Models

AgentsDGX agent

arXiv:2606.28524v1 Announce Type: new Abstract: Recent work suggests that Large Language Models (LLMs) are sensitive to the belief states of agents described by text, as measured by the false belief t

29 Jun 2026

Tired: The US government regulating open-source AI models Wired: The US government training and releasing open-source AI models

IndustryDGX agent

Tired: The US government regulating open-source AI models Wired: The US government training and releasing open-source AI models Today, we are releasing Rampart: a 14.7MB machine learning model designe

27 Jun 2026

When does combining LLMs help? Great analysis on combining language models, measured across 67 models from 21 providers. Any policy that rou…

SafetyDGX agent

When does combining LLMs help? Great analysis on combining language models, measured across 67 models from 21 providers. Any policy that routes, votes, cascades, or runs a mixture of agents and then r

26 Jun 2026

Frontier models are getting really expensive Model providers give us a bunch of features to reduce costs, but the support landscape is….inco…

AgentsDGX agent

Frontier models are getting really expensive Model providers give us a bunch of features to reduce costs, but the support landscape is….inconsistent We made prompt caching support model-agnostic in de

25 Jun 2026

Transitioning from using a single video model to generate a clip to achieving a finished video is like the transition from coding models tha…

AgentsDGX agent

Transitioning from using a single video model to generate a clip to achieving a finished video is like the transition from coding models that merely autocomplete to coding models that write working so

24 Jun 2026

Qwen-AgentWorld: Language World Models for General Agents

Model ReleasesDGX agent

arXiv:2606.24597v1 Announce Type: new Abstract: A world model predicts environment dynamics based on current observations and actions, serving as a core cognitive mechanism for reasoning and planning.

10 Jun 2026

Modeling Complex Behaviors: Multi-Personality Composition and Dynamic Switching in Vision-Language Models

Model ReleasesDGX agent

arXiv:2606.11074v1 Announce Type: cross Abstract: With the widespread deployment of Multimodal Large Language Models (MLLMs) in social interaction, understanding and controlling their behavior under c

9 Jun 2026

MemoryVLA++: Temporal Modeling via Memory and Imagination in Vision-Language-Action Models

ResearchDGX agent

arXiv:2606.09827v1 Announce Type: cross Abstract: Temporal modeling is essential for robotic manipulation, as effective control requires both memory of past interactions and imagination of future stat

4 Jun 2026

Sparse Mixture-of-Experts Reward Models Learn Interpretable and Specialized Experts for Personalized Preference Modeling

TutorialsDGX agent

arXiv:2606.04284v1 Announce Type: cross Abstract: Preference modeling plays a central role in reinforcement learning from human feedback (RLHF), enabling large language models (LLMs) to align with hum

2 Jun 2026

All Models are Wrong, Knowing Where is Useful: On Model Uncertainty in Reinforcement Learning

AgentsDGX agent

arXiv:2606.01363v1 Announce Type: new Abstract: Model-based reinforcement learning (MBRL) infers information about the environment from a learned dynamics model and bears the potential to address open

BLISS: A Lightweight Bilevel Influence Scoring Method for Data Selection in Language Model Pretraining

Model ReleasesDGX agent

arXiv:2510.06048v4 Announce Type: replace Abstract: Effective data selection is essential for pretraining large language models (LLMs), enhancing efficiency and improving generalization to downstream

← Previous
1…56789…990
Next →