AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
59,399 results
10 Apr 2026

People ask me how I choose what model to route queries to. It's simple. Claude gets knowledge work anything below that would be degrading Cl…

Model ReleasesDGX agent

I was unable to retrieve the specific tweet content from that URL, as the post requires a logged-in X (Twitter) session to access, and search results did not surface the full text of that specific ...

Temporally Phenotyping GLP-1RA Case Reports with Large Language Models: A Textual Time Series Corpus and Risk Modeling

Model ReleasesDGX agent

arXiv:2604.06197v1 Announce Type: cross Abstract: Type 2 diabetes case reports describe complex clinical courses, but their timelines are often expressed in language that is difficult to reuse in long

The Master Key Hypothesis: Unlocking Cross-Model Capability Transfer via Linear Subspace Alignment

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety
DGX agent

arXiv:2604.06377v1 Announce Type: cross Abstract: We investigate whether post-trained capabilities can be transferred across models without retraining, with a focus on transfer across different model

Which Way Does Time Flow? A Psychophysics-Grounded Evaluation for Vision-Language Models

Model ReleasesDGX agent

arXiv:2510.26241v5 Announce Type: replace-cross Abstract: Modern vision-language models (VLMs) excel at many multimodal tasks, yet their grasp of temporal information in video remains weak and has not

9 Apr 2026

if you want claude's managed agents w/o model provider lock in, this is the path! `deepagents deploy` lets you deploy an agent built on our …

Model ReleasesDGX agent

if you want claude's managed agents w/o model provider lock in, this is the path! `deepagents deploy` lets you deploy an agent built on our model agnostic, open source harness in minutes you can easil

8 Apr 2026

Lots of love for Gemma 4! Team just told me it’s already had 10M+ downloads since last week’s launch. Gemma models have now been downloaded …

Model ReleasesDGX agent

Google's Gemma 4, launched on April 2, 2026, surpassed 10 million downloads within its first week, according to Google CEO Sundar Pichai, pushing the cumulative total for all Gemma models past 500 ...

12 Aug 2026

Mapping and Measuring the Behavioral Evolution of Large Language Models

Model ReleasesDGX agent

arXiv:2608.11027v1 Announce Type: cross Abstract: Benchmark leaderboards summarize how well a language model performs, but not how its behavior relates to that of other models or changes across genera

Reference-Free Post-Training of Open Large Language Models for Multilingual Machine Translation

Model ReleasesDGX agent

arXiv:2608.10812v1 Announce Type: cross Abstract: We study reference-free post-training for multilingual machine translation with open large language models. Starting from the supervised-finetuned MiL

11 Aug 2026

A Hybrid Neural-Microfacet BRDF Model for Real-Time Rendering

ResearchDGX agent

arXiv:2608.09604v1 Announce Type: cross Abstract: Over the past decade, microfacet-based BRDF models have formed the foundation of real-time rendering pipelines. Despite their widespread use, they oft

ELBench: A Multi-Dimensional Benchmark for Education-Facing Large Language Models

Model ReleasesDGX agent

arXiv:2608.09548v1 Announce Type: cross Abstract: Large language models are increasingly deployed in education as tutors, teaching assistants, and content generators. These roles place demands that or

HugSelect: An Explainable Multi-Criteria Decision-Support Framework for foundation-model selection

ResearchDGX agent

arXiv:2608.08069v1 Announce Type: cross Abstract: Foundation models are increasingly reused as software components, making model selection a critical software-engineering decision. Current model hubs

10 Aug 2026

Addressable Memory for Video World Models

Model ReleasesDGX agent

arXiv:2608.07408v1 Announce Type: new Abstract: We study visual persistence in interactive video world models. These models rely on a Key-Value (KV) cache as a growing visual memory to carry forward p

Critical Acclaim Orientation in Large Language Models: Evidence from Film Preference Elicitation

Model ReleasesDGX agent

arXiv:2608.06955v1 Announce Type: new Abstract: Large language models (LLMs) are trained on corpora that contain expressions of human judgment about films, books, music, and more. Yet whether LLMs sys

Meta releases open-source Muse Glimmer model with 30B parameters

Model ReleasesDGX agent

Meta Platforms Inc. today released Muse Glimmer, an open-source language model that can run on personal computers. The company also published a lengthy essay penned by Chief Executive Officer Mark Zuc

Probing Visual Concepts in Lightweight Vision-Language Models for Automated Driving

AgentsDGX agent

arXiv:2603.06054v2 Announce Type: replace-cross Abstract: The use of Vision-Language Models (VLMs) in automated driving applications is becoming increasingly common, with the aim of leveraging their r

6 Aug 2026

The Fairness Collapse Phenomenon: Bias Amplification in Language Models Trained on Synthetic Data

SafetyDGX agent

arXiv:2608.04268v1 Announce Type: new Abstract: Generative models trained on artificially generated data have been shown to exhibit model collapse, resulting in significant performance degradation. As

5 Aug 2026

Conformal risk control for model-form uncertainty in parametric non-intrusive reduced-order models

ApplicationsDGX agent

arXiv:2608.03360v1 Announce Type: cross Abstract: Non-intrusive reduced-order models (NIROMs) have become a standard tool for approximating parametric partial differential equations from computer desi

Utilize a nvidia gpu and amd gpu together for 2 different ai models?

Model ReleasesDGX agent

We run a local model instance in our company that the dev we hired built for us. We're a trade business and we want to further use our on hand hardware for it. The specs given we have is a 5090 gpu wi

Xiaomi-Robotics-1: New robotics model released

Model ReleasesDGX agent

Xiaomi-Robotics-1 is a robot foundation model trained on over 100K hours of real-world manipulation trajectories. It is a Vision-Language-Action (VLA) model engineered for out-of-the-box mobile manipu

3 Aug 2026

A Frozen Pixel-Space Diffusion Model Can Guide Itself with Its Own Samples

Model ReleasesDGX agent

arXiv:2607.29122v1 Announce Type: new Abstract: Pixel-space diffusion models aim to learn an end-to-end generator directly over raw pixels. This is challenging because a single model must capture both

Alibaba debuts Qwen3.8-Max model with 2.4T parameters

Model ReleasesDGX agent

Alibaba Group Holding Ltd. today debuted a new addition to its Qwen series of open-source large language models. Qwen3.8-Max is the Chinese e-commerce giant’s most capable LLM to date. It features 2.4

31 Jul 2026

Beyond the Bidirectional Promise: Re-evaluating the Robustness of Diffusion Language Models

Model ReleasesDGX agent

arXiv:2607.27386v1 Announce Type: cross Abstract: Diffusion Language Models (DLMs) offer a compelling alternative to autoregressive (AR) generation by enabling bidirectional context and iterative refi

MiniMax H3: Open-weight multimodel video model

Model ReleasesDGX agent

Just saw this posted by Fal.ai and then by Hailuo themselves, the next video model will be open weight released! Here's the blurb and link to to the feature post: Today, we're launching MiniMax H3, a

Predict before you train: Scaling Laws for particle physics foundation models

ResearchDGX agent

arXiv:2607.23377v1 Announce Type: cross Abstract: The largest machine learning models in particle physics are also the most expensive to train, yet the return on scaling a given architecture cannot be

RepBench: Compiling Benchmarks into Capability Representations for Large Language Models

Model ReleasesDGX agent

arXiv:2607.28008v1 Announce Type: new Abstract: Representation engineering reads and steers capability directions in large language models, yet methods are typically evaluated on paper-specific synthe

Same Facts, Different Diagnosis: Measuring and Mitigating Narrative Anchoring in Clinical Language Models

Model ReleasesDGX agent

arXiv:2607.27384v1 Announce Type: new Abstract: Large language models used for clinical diagnostic reasoning are sensitive to sociolinguistic register, not just clinical content. We term this failure

30 Jul 2026

Language Models are not Equally Robust to Non-Canonical Tokenization across Languages

Model ReleasesDGX agent

arXiv:2607.26831v1 Announce Type: new Abstract: Despite the existence of exponentially many valid tokenizations for a given string, language models operate on a single canonical sequence deterministic

29 Jul 2026

Laplace-PSN-IRT: Uncertainty Quantification for Neural Item Response Theory Models of LLM Benchmarks

Model ReleasesDGX agent

arXiv:2607.25257v1 Announce Type: cross Abstract: Item Response Theory (IRT) has recently been proposed as a framework for evaluating large language model (LLM) benchmarks by separating a model's late

WALoMA: A Multitask Wireless Foundation Model via Adaptive Low-Rank Masked Autoencoders

Model ReleasesDGX agent

arXiv:2607.25763v1 Announce Type: cross Abstract: This paper proposes a multitask wireless foundation model via adaptive low-rank masked autoencoders (WALoMA), a unified multi-task foundation model fo

28 Jul 2026

AIR-BENCH Live: An Evolving Safety Benchmark for Foundation Models

Model ReleasesDGX agent

arXiv:2607.22671v1 Announce Type: new Abstract: Foundation-model safety benchmarks capture the AI risks of their time of publication: as models improve and governments pass new AI-safety legislation,

Numerical Investigation of Sequence Modeling Theory using Controllable Memory Functions

Model ReleasesDGX agent

arXiv:2506.05678v3 Announce Type: replace Abstract: The evolution of sequence modeling architectures, from recurrent neural networks and convolutional models to Transformers and structured state-space

Reverso: Efficient Time Series Foundation Models for Zero-shot Forecasting

ResearchDGX agent

arXiv:2602.17634v2 Announce Type: replace-cross Abstract: Learning time series foundation models has been shown to be a promising approach for zero-shot time series forecasting across diverse time ser

RM-Distiller: Exploiting Generative LLM for Reward Model Distillation

SafetyDGX agent

arXiv:2601.14032v2 Announce Type: replace Abstract: Reward models (RMs) play a pivotal role in aligning large language models (LLMs) with human preferences. Due to the difficulty of obtaining high-qua

27 Jul 2026

Big update: Among open-weight models, Kimi K3 (Max) is #1 in the Agent Arena with +9.75% net-improvement, surpassing GLM-5.2 (Max) at +7.12%…

Model ReleasesDGX agent

Big update: Among open-weight models, Kimi K3 (Max) is #1 in the Agent Arena with +9.75% net-improvement, surpassing GLM-5.2 (Max) at +7.12%, and landed the #1 spot across 5 signals (see below). Kimi

MoE^2-LoRA: When MoE Models Meet MoE-style Low-Rank Adaptation

Model ReleasesDGX agent

arXiv:2607.21978v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) architectures have been widely adopted in large language models, yet parameter-efficient fine-tuning (PEFT) for MoE models rema

24 Jul 2026

PhantomFill: When the Form Demands an Answer, Language Models Invent One

Model ReleasesDGX agent

arXiv:2607.20492v1 Announce Type: cross Abstract: Language models in production do not write prose. They fill forms: JSON fields, function arguments, extraction templates. We show that the form itself

23 Jul 2026

Task Competence Is Not Instruction Following: Evaluating Instruction-Conflicting Behavior in Small Language Models

Model ReleasesDGX agent

arXiv:2607.19608v1 Announce Type: new Abstract: Instruction tuning is meant to make language models follow user requests, yet it is unclear whether small models comply when an instruction conflicts wi

Toward a Vision-Language Foundation Model for Medical Data: Multimodal Dataset and Benchmarks for Vietnamese PET/CT Report Generation

Model ReleasesDGX agent

arXiv:2509.24739v4 Announce Type: replace Abstract: Vision-Language Foundation Models (VLMs), trained on large-scale multimodal datasets, have driven significant advances in Artificial Intelligence (A

16 Jul 2026

Mixed-Timescale Differential Coding for Downlink Model Broadcast in Wireless Federated Learning

Model ReleasesDGX agent

arXiv:2607.13119v1 Announce Type: cross Abstract: In standard federated learning systems, the parameter server broadcasts the global model to the participating devices in every iteration. Motivated by

TSSM: Triaxial State Space Model for Global Station Weather Forecasting with Temporal-Variable-Historical Modeling

Local AiDGX agent

arXiv:2607.13101v1 Announce Type: cross Abstract: Global Station Weather Forecasting (GSWF) is pivotal for localized and extreme weather prediction over key regions. Despite efforts to exploit look-ba

15 Jul 2026

Do We Really Need Multimodal Emotion Language Models Larger Than 1B Parameters?

Model ReleasesDGX agent

arXiv:2607.12787v1 Announce Type: new Abstract: Recent advances in multimodal large language models (MLLMs) have significantly improved the performance of multimodal emotion recognition (MER) and enab

OvisOCR2 (0.8B): first end-to-end model to top OmniDocBench - I threw 827 real scanned medical docs at it, here's everything I learned

Model ReleasesDGX agent

What it is: ATH-MaaS/OvisOCR2 - a 0.8B document-parsing VLM post-trained from Qwen3.5-0.8B (SFT + RL + OPD), Apache 2.0, runs on vLLM 0.22.1. One prompt per page image -> complete markdown (HTML table

Silent Alarm: A J-Space Protocol for Comparing Danger Recognition Across Models and Quantization Levels

Model ReleasesDGX agent

arXiv:2607.12792v1 Announce Type: cross Abstract: Jailbreak-robustness research typically evaluates safety through generated responses using an LLM-as-judge approach. Such evaluations, however, are se

14 Jul 2026

Google named a Leader in the 2026 IDC MarketScape for Worldwide Foundation Model Software

Model ReleasesDGX agent

For years, we’ve built with a clear priority: putting the practical needs of the enterprise first. Long before generative AI dominated the headlines, we were focused on building the global infrastruct

13 Jul 2026

New model for AMD Strix Halo users: My 198B Step 3.7 Flash release was a big hit, but this one may be even better: 298B-parameter Hy3, now r…

Model ReleasesDGX agent

New model for AMD Strix Halo users: My 198B Step 3.7 Flash release was a big hit, but this one may be even better: 298B-parameter Hy3, now running on a new 2-bit FPX codebook designed to map efficient

.@satyanadella's Reverse Information Paradox is real. What @satyanadella's calling for already exists: open models. Over 9M+ developers have…

Local AiDGX agent

.@satyanadella's Reverse Information Paradox is real. What @satyanadella's calling for already exists: open models. Over 9M+ developers have used @ollama to access open models and keep their competiti

10 Jul 2026

AUTOPILOT VQA: Benchmarking Vision-Language Models for Incident-Centric Dashcam Understanding

Model ReleasesDGX agent

arXiv:2607.08745v1 Announce Type: new Abstract: Recent advances in Vision-Language Models, Large Language Models, and Multimodal Large Language Models have improved autonomous driving tasks such as sc

9 Jul 2026

Bifidelity Parameter Estimation Using Conditional Diffusion Models

Model ReleasesDGX agent

arXiv:2504.01894v2 Announce Type: replace Abstract: We present a bifidelity method for uncertainty quantification of parameter estimates in complex systems, leveraging generative models trained to sam

GLM 5.2: a new rise of open-weight agentic models

Model ReleasesDGX agent

On June 16th, Z.ai released GLM 5.2, its latest flagship model. At the time of announcement, it advertised scores at or near Anthropic and OpenAI's models, and far ahead of GLM 5.1. In the world of us

7 Jul 2026

Deform360: A Massive Multi-view Visuotactile Dataset for Deformable World Models

Model ReleasesDGX agent

arXiv:2607.05390v1 Announce Type: cross Abstract: Predicting object dynamics (i.e., world modeling) is a fundamental challenge for robotic manipulation, and modeling deformable objects presents a part

DELTA-TTS: Adapting Autoregressive Model into Diffusion Language Model for Text-to-Speech

Local AiDGX agent

arXiv:2607.04140v1 Announce Type: cross Abstract: Autoregressive (AR) text-to-speech (TTS) models generate discrete speech tokens sequentially, which makes inference slow and can degrade robustness by

30 Jun 2026

Recommended reading if you are scaling with open models. BTW, you should be thinking about how to scale with open-weight models.

TutorialsDGX agent

This post from DAIR.AI recommends resources for developers and organizations scaling applications using open-weight language models, emphasizing the importance of planning infrastructure and deploymen

29 Jun 2026

FULL INTERVIEW: Engineers Edward Coristine (@as400495) and Tai Groot (@taigrr) just released an ML model called Rampart for the National Des…

Model ReleasesDGX agent

FULL INTERVIEW: Engineers Edward Coristine (@as400495) and Tai Groot (@taigrr) just released an ML model called Rampart for the National Design Studio. It's a local-first, open-source AI privacy model

26 Jun 2026

@NousResearch We are working on benching various combos of open source models to see if we can get Opus levels with much cheaper models as w…

ResearchDGX agent

Nous Research is conducting benchmarking tests to evaluate whether combinations of open-source models can achieve performance comparable to Anthropic's Claude Opus while maintaining significantly lowe

25 Jun 2026

Bias Fitting to Mitigate Length Bias of Reward Model in RLHF

SafetyDGX agent

arXiv:2505.12843v2 Announce Type: replace Abstract: Reinforcement Learning from Human Feedback (RLHF) relies on reward models to align large language models with human preferences. However, RLHF often

23 Jun 2026

SeFi-Image: A Text-to-Image Foundation Model with Semantic-First Diffusion

Model ReleasesDGX agent

arXiv:2606.22568v1 Announce Type: new Abstract: Training image generation foundation models consumes substantial resources. Previous methods have attempted to leverage semantic guidance to accelerate

11 Jun 2026

Agentic Environment Engineering for Large Language Models: A Survey of Environment Modeling, Synthesis, Evaluation, and Application

AgentsDGX agent

arXiv:2606.12191v1 Announce Type: cross Abstract: Environments serve as interactive systems for large language model (LLM) based agents across diverse scenarios and play a crucial role in driving the

BioMamba: Domain-Adaptive Biomedical Language Models

Model ReleasesDGX agent

arXiv:2408.02600v3 Announce Type: replace Abstract: Background. Biomedical language models should improve performance on biomedical text while retaining general-language-modeling fluency. For Mamba-ba

When Roleplaying, Do Models Believe What They Say?

Model ReleasesDGX agent

arXiv:2606.11502v1 Announce Type: cross Abstract: Language models can state that 'the Earth orbits the Sun' and, when role-playing Aristotle, assert the opposite. Recent work argues that persona adopt

10 Jun 2026

Large Language Models as Modal Models in Linguistics

ResearchDGX agent

arXiv:2606.10467v1 Announce Type: new Abstract: The rapid advancement of large language models (LLMs) has intensified debates about their significance for linguistic theory. These debates are commonly

← Previous
1…1314151617…990
Next →