AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
59,929 results
9 Jun 2026

Fable 5 is now available in Claude Code and Cowork Fable is the best model I have used for coding, by a wide margin. It is a big step up, en…

Model ReleasesDGX agent

Fable 5 is now available in Claude Code and Cowork Fable is the best model I have used for coding, by a wide margin. It is a big step up, enabling less prompts and steers, more efficient token use, be

GlobeAudio: A Multilingual Multicultural Benchmark for Naturalistic Evaluation of Large Audio-Language Models

Model ReleasesDGX agent

arXiv:2606.08194v1 Announce Type: cross Abstract: Large Audio-Language Models (LALMs) integrate audio perception and language understanding within a unified framework, enabling a wide range of real-wo

HASA: Subnet Allocation for Compute-Constrained Model-Heterogeneous Federated Learning

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases
DGX agent

arXiv:2606.07621v1 Announce Type: cross Abstract: Edge services increasingly use federated learning to personalize on-device models while keeping sensitive data local. In practice, deployments must ha

Large Models for Time Series and Spatio-Temporal Data: A Survey and Outlook

ApplicationsDGX agent

arXiv:2310.10196v3 Announce Type: replace-cross Abstract: Temporal data, including time series and spatio-temporal data, are pervasive in real-world applications. Generated in massive volumes by physi

Phase transition in large language models and the criticality of natural languages

Model ReleasesDGX agent

arXiv:2406.05335v3 Announce Type: replace-cross Abstract: Generation of text and speech in natural languages can be modeled as a stochastic process. This idea dates back to the seminal work of Markov

TempoBench: Evaluating Temporal Causal Reasoning in Large Language Models

Model ReleasesDGX agent

arXiv:2510.27544v2 Announce Type: replace Abstract: Temporal reasoning involves understanding how systems evolve over time through input-driven state transitions. A key aspect is temporal causal reaso

Thinking-Based Non-Thinking: Solving the Reward Hacking Problem in Training Hybrid Reasoning Models via Reinforcement Learning

Model ReleasesDGX agent

arXiv:2601.04805v2 Announce Type: replace Abstract: Large reasoning models (LRMs) have attracted much attention due to their exceptional performance. However, their performance mainly stems from think

Today, we released Gemini 3.5 Live Translate, our latest audio model for live speech-to-speech translation. It supports over 70 languages an…

Model ReleasesDGX agent

Today, we released Gemini 3.5 Live Translate, our latest audio model for live speech-to-speech translation. It supports over 70 languages and starts translating as soon as you start talking, streaming

why does ollama unloads models automatically?

Local AiDGX agent

Ollama automatically unloads inactive models from memory based on inactivity parameters, with models being removed after a specified duration of non-use . By default, Ollama unloads a model after 5 mi

8 Jun 2026

A Four-Condition Diagnostic Protocol for Evidence Utilization in Long-Context and Retrieval-Augmented Language Models

Model ReleasesDGX agent

arXiv:2606.06758v1 Announce Type: new Abstract: Final-answer accuracy, retrieval recall, and citation overlap do not by themselves identify whether a long-context or retrieval-augmented language model

CULTURESCORE: Evaluating Cultural Faithfulness in Video Generation Models

Local AiDGX agent

arXiv:2606.07311v1 Announce Type: cross Abstract: As video generation models like Veo 3.1 and LTX-2 advance, their ability to accurately represent diverse global cultures remains a critical yet unders

Data-Efficient Autoregressive-to-Diffusion Language Models via On-Policy Distillation

SafetyDGX agent

arXiv:2606.06712v1 Announce Type: cross Abstract: We study the transformation of autoregressive models (ARLMs) into diffusion language models (DLMs). Rather than pretraining from scratch, prior work r

Generalization of Diffusion Models Arises with a Balanced Representation Space

Local AiDGX agent

arXiv:2512.20963v3 Announce Type: replace-cross Abstract: Diffusion models excel at generating high-quality, diverse samples, yet they risk memorizing training data when overfit to the training object

Rethinking Genomic Modeling Through Optical Character Recognition

ResearchDGX agent

arXiv:2602.02014v2 Announce Type: replace-cross Abstract: Recent genomic foundation models largely adopt large language model architectures that treat DNA as a one-dimensional token sequence. However,

ShallowBench: Benchmarking Generative Drug Design Models on Shallow-Pocket Targets

Model ReleasesDGX agent

arXiv:2606.06717v1 Announce Type: cross Abstract: While generative AI models have demonstrated remarkable success in structure-based drug design, they predominantly rely on deep binding pockets and st

7 Jun 2026

American Open Source is so back. 9 / 30 of the models on page 1 of Huggingface are published by Nvidia.

HardwareDGX agent

Nvidia has significantly increased its presence in open-source AI models, with 9 out of 30 models on the first page of Hugging Face's model hub published by the company. This observation reflects Nvid

6 Jun 2026

CogManip: Benchmarking Manipulative Behavior in Multi-Turn Interactions with Large Language Model

Model ReleasesDGX agent

arXiv:2606.06099v1 Announce Type: new Abstract: Whether Large Language Models (LLMs) exhibit covert psychological manipulation in complex human-AI interactions has garnered increasing safety concerns.

Fireworks Training Platform keeps expanding. Leading US open weight model Nemotron 3 Ultra is now ready for post-training: SFT and DPO via L…

Model ReleasesDGX agent

Fireworks Training Platform keeps expanding. Leading US open weight model Nemotron 3 Ultra is now ready for post-training: SFT and DPO via LoRA or full-parameter, on the same infrastructure that serve

5 Jun 2026

AdaPlanBench: Evaluating Adaptive Planning in Large Language Model Agents under World and User Constraints

Model ReleasesDGX agent

arXiv:2606.05622v1 Announce Type: new Abstract: Planning for real-world problems by language models often involves both world and user constraints, which may not be fully specified upfront and are pro

CaliDist: Calibrating Large Language Models via Behavioral Robustness to Distraction

ResearchDGX agent

arXiv:2606.05799v1 Announce Type: cross Abstract: Existing calibration methods for Large Language Models (LLMs) often overlook a critical dimension of trustworthiness: a model's {em behavioral robustn

Decomposing Factual Sycophancy in Language Models: How Size and Instruction Tuning Shape Robustness

ResearchDGX agent

arXiv:2606.06306v1 Announce Type: new Abstract: Factual sycophancy occurs when a language model abandons a correct, verifiable answer under social pressure. Because a flip occurs only when pressure to

Explainability of Large Language Models: Opportunities and Challenges toward Generating Trustworthy Explanations

Local AiDGX agent

arXiv:2510.17256v2 Announce Type: replace Abstract: Large language models have exhibited impressive performance across a broad range of downstream tasks in natural language processing. However, how a

FUSAR-GPT : A Spatiotemporal Feature-Embedded and Two-Stage Decoupled Visual Language Model for SAR Imagery

Model ReleasesDGX agent

arXiv:2602.19190v4 Announce Type: replace Abstract: Research on the intelligent interpretation of all-weather, all-time Synthetic Aperture Radar (SAR) is crucial for advancing remote sensing applicati

Less is MoE: Trimming Experts in Domain-Specialist Language Models

Model ReleasesDGX agent

arXiv:2606.05538v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) models achieve strong performance through conditional computation, but their large parameter footprint poses deployment chall

Seeing Time: Benchmarking Chronological Reasoning and Shortcut Biases in Vision-Language Models

Model ReleasesDGX agent

arXiv:2606.05702v1 Announce Type: cross Abstract: Recent advancements in Vision-Language Models (VLMs) have significantly enhanced their ability to interpret complex visual semantics, yet their capaci

The Tell-Tale Norm: ell_2 Magnitude as a Signal for Reasoning Dynamics in Large Language Models

ResearchDGX agent

arXiv:2606.06188v1 Announce Type: new Abstract: Recent work has sought to understand Large Language Models (LLMs) reasoning, yet a principled, model-intrinsic signal that captures its layer-wise reaso

4 Jun 2026

A Latent Variable Framework for Scaling Laws in Large Language Models

ResearchDGX agent

arXiv:2512.06553v2 Announce Type: replace-cross Abstract: We propose a statistical framework built on latent variable modeling for scaling laws of large language models (LLMs). Our work is motivated b

An Ensembled Latent Factor Model via Differential Evolution and Gradient Descent Optimization

ApplicationsDGX agent

arXiv:2606.04408v1 Announce Type: cross Abstract: High-dimensional and incomplete (HDI) data are prevalent in many real-world big data scenarios. Latent factor models serve as a common representation

Audio Interaction Model

ResearchDGX agent

arXiv:2606.05121v1 Announce Type: cross Abstract: Audio is an inherently interactive modality, yet today's Large Audio Language Models (LALMs) are offline, and streaming audio models each handle only

Discourse-Role Labels as Presentation-Time Variables for Context Use in Language Models

Model ReleasesDGX agent

arXiv:2606.04109v1 Announce Type: new Abstract: Context-augmented language model systems often wrap supplied content with labels such as Reference:, Evidence:, Instruction:, Note:, or Example:, but th

Food-R1: A Unified Multi-Task Food Vision-Language Model with Reinforcement Learning

Model ReleasesDGX agent

arXiv:2606.04986v1 Announce Type: new Abstract: Recent studies have explored Vision-Language Models (VLMs) for food analysis. However, most existing methods rely primarily on supervised fine-tuning (S

Interfaze: The Future of AI is built on Task-Specific Small Models

Model ReleasesDGX agent

arXiv:2602.04101v2 Announce Type: replace Abstract: We present Interfaze, a native hybrid model that fuses task-specific deep neural networks (CNNs and DNNs) directly into a transformer decoder throug

LimiX-2M: Mitigating Low-Rank Collapse and Attention Bottlenecks in Tabular Foundation Models

Model ReleasesDGX agent

arXiv:2606.04485v1 Announce Type: new Abstract: Tabular foundation models (TFMs) increasingly rival tree ensembles, but their performance is often compute-inefficient: with standard affine scalar toke

LoopMoE: Unifying Iterative Computation with Mixture-of-Experts for Language Modeling

Model ReleasesDGX agent

arXiv:2606.04438v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) and looped architectures scale models along two orthogonal axes, namely parameter capacity and effective depth. However, main

QuBLAST: A Framework for Quantizing Large Language Models with Block-Level Compression Approach and Activation Scaling Strategy

Model ReleasesDGX agent

arXiv:2606.04620v1 Announce Type: cross Abstract: LLMs have become the state-of-the-art algorithms for solving NLP tasks. However, they typically come at huge computational and memory costs, thus maki

SPLIT-PINN: Separable Probability Learning Technique via Physics-Informed Neural Networks for High-Dimensional Probabilistic Modeling

Model ReleasesDGX agent

arXiv:2606.04000v1 Announce Type: cross Abstract: We present a probabilistic modeling framework for incorporating small-scale spatial heterogeneity into macroscopic descriptions of material behavior f

Towards Evaluating the Robustness of Visual State Space Models

ApplicationsDGX agent

arXiv:2406.09407v3 Announce Type: replace Abstract: Vision State Space Models (VSSMs), a novel architecture that combines the strengths of recurrent neural networks and latent variable models, have de

transitions like this are why we think it's helpful to have a provider-agnostic harness we used to talk more about swapping models when the …

Model ReleasesDGX agent

transitions like this are why we think it's helpful to have a provider-agnostic harness we used to talk more about swapping models when the latest and greatest came out -- but the latest and greatest

UniCAD: A Unified Benchmark and Universal Model for Multi-Modal Multi-Task CAD

Model ReleasesDGX agent

arXiv:2606.05058v1 Announce Type: cross Abstract: Computer-Aided Design (CAD) underpins modern engineering and manufacturing by enabling the creation of precise, editable 3D models. However, CAD resea

VGGSounder: Audio-Visual Evaluations for Foundation Models

Model ReleasesDGX agent

arXiv:2508.08237v4 Announce Type: replace-cross Abstract: The emergence of audio-visual foundation models underscores the importance of reliably assessing their multi-modal understanding. The VGGSound

3 Jun 2026

A Close Look At World Model Recovery In Supervised Fine-Tuned LLM Planners

TutorialsDGX agent

arXiv:2606.03685v1 Announce Type: cross Abstract: Supervised fine-tuning (SFT) improves end-to-end classical planning in large language models (LLMs), but do these models also learn to represent and r

[AINews] Microsoft Build: MAI-Thinking-1 and MAI Family models

ToolsDGX agent

Microsoft announced the MAI-Thinking-1 and MAI family of models at their Build conference, representing new advances in their AI model lineup. These models likely focus on improved reasoning capabilit

Another banger open-source release. Miso One is an 8B text-to-speech model with real emotional range, so voiceovers carry warmth, hesitation…

Model ReleasesDGX agent

Another banger open-source release. Miso One is an 8B text-to-speech model with real emotional range, so voiceovers carry warmth, hesitation, and excitement instead of sounding flat. It's purpose-buil

Automated Report-Derived Oncology VQA Benchmark for Evaluating Vision-Language Models on 3D Medical Imaging

Model ReleasesDGX agent

arXiv:2606.02809v1 Announce Type: new Abstract: Evaluating vision-language models (VLMs) on medical images requires benchmarks that are clinically grounded, scalable, and controlled for evaluation con

BehaviorBench: Modeling Real-World User Decisions from Behavioral Traces

Model ReleasesDGX agent

arXiv:2606.02798v1 Announce Type: new Abstract: Many decision-support settings require systems that adapt to individual users, but evaluation data for this problem remain limited. Existing benchmarks

Consistent Yet Wrong: Evidence Insensitivity in Spatial Vision-Language Models

Model ReleasesDGX agent

arXiv:2606.02742v1 Announce Type: new Abstract: Spatial reasoning is fundamental to robotics, autonomy, and embodied AI, yet modern vision-language models (VLMs) remain unreliable on metric distance q

Google introduces Gemma 4 12B, a unified, encoder-free open multimodal model that can run locally on devices with 16GB of VRAM or unified memory (Carl Franzen/VentureBeat)

Model ReleasesDGX agent

Carl Franzen / VentureBeat: Google introduces Gemma 4 12B, a unified, encoder-free open multimodal model that can run locally on devices with 16GB of VRAM or unified memory — While many AI open source

ReaLM: Residual Quantization Bridging Knowledge Graph Embeddings and Large Language Models

Model ReleasesDGX agent

arXiv:2510.09711v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have recently emerged as a powerful paradigm for Knowledge Graph Completion (KGC), offering strong reasoning and

Reasoning Structure of Large Language Models

Model ReleasesDGX agent

arXiv:2606.03883v1 Announce Type: new Abstract: Large reasoning models (LRMs) are often evaluated using metrics such as final-answer accuracy or token count. However, identical scores on these metrics

RogueMerge: Robust and Unified Attacks against LLM Model Merging

Model ReleasesDGX agent

arXiv:2606.03344v1 Announce Type: cross Abstract: Model merging composes specialized capabilities into a single LLM by aggregating task vectors sourced from unverified public platforms, exposing a cri

SVHalluc: Benchmarking Speech-Vision Hallucination in Audio-Visual Large Language Models

Model ReleasesDGX agent

arXiv:2606.02642v1 Announce Type: cross Abstract: Despite the success of audio-visual large-language models (LLMs), they can produce plausible but ungrounded outputs, termed hallucination. Existing be

Synthetic Hallucinations, Real Gains: Hard Negatives from Frontier Models for FIM Hallucination Mitigation

Model ReleasesDGX agent

arXiv:2606.03130v1 Announce Type: new Abstract: Small open-source code models that power IDE autocomplete still emit hallucinated Fill-in-the-Middle (FIM) completions: syntactically natural calls to m

VLA-Arena: An Open-Source Framework for Benchmarking Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2512.22539v2 Announce Type: replace-cross Abstract: While Vision-Language-Action models (VLAs) are rapidly advancing towards generalist robot policies, it remains difficult to quantitatively und

When Models Refuse: Political Steerability and Feature Richness as Measures of Ideological Depth

SafetyDGX agent

arXiv:2508.21448v3 Announce Type: replace Abstract: Large language models (LLMs) sometimes refuse to follow benign instructions, such as declining to argue a political position or adopt a stated perso

2 Jun 2026

Assessment of Generative Named Entity Recognition in the Era of Large Language Models

Model ReleasesDGX agent

arXiv:2601.17898v2 Announce Type: replace Abstract: Named entity recognition (NER) is evolving from a sequence labeling task into a generative paradigm with the rise of large language models (LLMs). W

Beyond Isolated Behaviors: Hierarchical User Modeling for LLM Personalization

Model ReleasesDGX agent

arXiv:2606.02300v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities across diverse domains, yet personalizing their outputs to individual users remai

Business Utility of Large Language Models as Exploratory Data Analysis Agents

Model ReleasesDGX agent

arXiv:2606.00051v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used in analytical workflows, but their suitability as exploratory data analysis (EDA) agents in busines

Decoupled Residual Denoising Diffusion Models for Unified and Data Efficient Image-to-Image Translation

ResearchDGX agent

arXiv:2606.01048v1 Announce Type: new Abstract: We propose Decoupled Residual Denoising Diffusion models (DRDD) for unified and data-efficient image-to-image (I2I) translation. While diffusion models

DetailMaster: Can Your Text-to-Image Model Handle Long Prompts?

Model ReleasesDGX agent

arXiv:2505.16915v3 Announce Type: replace-cross Abstract: While recent Text-to-Image (T2I) models show impressive capabilities in synthesizing images from brief descriptions, they struggle with the lo

Domain-Shift-Aware Conformal Prediction for Large Language Models

Model ReleasesDGX agent

arXiv:2510.05566v2 Announce Type: replace-cross Abstract: Large language models have achieved impressive performance across diverse tasks. However, their tendency to produce overconfident and factuall

← Previous
1…6465666768…999
Next →