AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
59,399 results
15 May 2026

Slot-MPC: Goal-Conditioned Model Predictive Control with Object-Centric Representations

SafetyDGX agent

arXiv:2605.14937v1 Announce Type: cross Abstract: Predictive world models enable agents to model scene dynamics and reason about the consequences of their actions. Inspired by human perception, object

14 May 2026

Edit-Compass & EditReward-Compass: A Unified Benchmark for Image Editing and Reward Modeling

Model ReleasesDGX agent

arXiv:2605.13062v1 Announce Type: new Abstract: Recent image editing models have achieved remarkable progress in instruction following, multimodal understanding, and complex visual editing. However, e

Understanding and Accelerating the Training of Masked Diffusion Language Models

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases
DGX agent

arXiv:2605.13026v1 Announce Type: cross Abstract: Masked diffusion models (MDMs) have emerged as a promising alternative to autoregressive models (ARMs) for language modeling. However, MDMs are known

When is Warmstarting Effective for Scaling Language Models?

Model ReleasesDGX agent

arXiv:2605.13405v1 Announce Type: new Abstract: Model growth from a given checkpoint aims to accelerate training of a larger model, offering potential resource savings. Despite recent interest, warmst

13 May 2026

Qwen-Scope: Turning Sparse Features into Development Tools for Large Language Models

Model ReleasesDGX agent

arXiv:2605.11887v1 Announce Type: new Abstract: Large language models have achieved remarkable capabilities across diverse tasks, yet their internal decision-making processes remain largely opaque, li

READ: Recurrent Adapter with Partial Video-Language Alignment for Parameter-Efficient Transfer Learning in Low-Resource Video-Language Modeling

Model ReleasesDGX agent

arXiv:2312.06950v3 Announce Type: replace-cross Abstract: Fully fine-tuning pretrained large-scale transformer models has become a popular paradigm for video-language modeling tasks, such as temporal

12 May 2026

A Single-Layer Model Can Do Language Modeling

ResearchDGX agent

arXiv:2605.10643v1 Announce Type: new Abstract: Modern language models scale depth by stacking layers, each holding its own state - a per-layer KV cache in transformers, a per-layer matrix in Mamba, G

Aligning Agents via Planning: A Benchmark for Trajectory-Level Reward Modeling

Model ReleasesDGX agent

arXiv:2604.08178v2 Announce Type: replace Abstract: In classical Reinforcement Learning from Human Feedback (RLHF), Reward Models (RMs) serve as the fundamental signal provider for model alignment. As

dFlowGRPO: Rate-Aware Policy Optimization for Discrete Flow Models

SafetyDGX agent

arXiv:2605.09291v1 Announce Type: new Abstract: Discrete flow models (DFMs) are a class of flexible generative models for generating discrete data, and diffusion large language models (dLLMs) can be v

GIFT: Guided Importance-Aware Fine-Tuning for Diffusion Language Models

Model ReleasesDGX agent

arXiv:2509.20863v3 Announce Type: replace Abstract: Diffusion models have recently shown strong potential in language modeling, offering faster generation compared to traditional autoregressive approa

Hunyuan3D 2.0: Scaling Diffusion Models for High Resolution Textured 3D Assets Generation

Model ReleasesDGX agent

arXiv:2501.12202v4 Announce Type: replace Abstract: We present Hunyuan3D 2.0, an advanced large-scale 3D synthesis system for generating high-resolution textured 3D assets. This system includes two fo

Is Your Driving World Model an All-Around Player?

Model ReleasesDGX agent

arXiv:2605.10858v1 Announce Type: new Abstract: Today's driving world models can generate remarkably realistic dash-cam videos, yet no single model excels universally. Some generate photorealistic tex

Model-Free Neural Filtering: A Comparison with Classical Filters in Nonlinear Systems

Model ReleasesDGX agent

arXiv:2601.21266v3 Announce Type: replace Abstract: Neural network models are increasingly used for state estimation in control and decision-making, yet it remains unclear to what extent they behave a

11 May 2026

Benchmarking World-Model Learning with Environment-Level Queries

Model ReleasesDGX agent

arXiv:2510.19788v4 Announce Type: replace Abstract: World models are central to building AI agents capable of flexible reasoning and planning. Yet current evaluations (i) test only properties measurab

Fine-tuning a vision-language model for fracture-surface morphology recognition

Model ReleasesDGX agent

arXiv:2605.07145v1 Announce Type: cross Abstract: Vision-language models (VLMs) have shown strong potential for scientific image understanding, but general-purpose models often lack the domain-specifi

6 May 2026

A Domain Incremental Continual Learning Benchmark for ICU Time Series Model Transportability

Model ReleasesDGX agent

arXiv:2605.03832v1 Announce Type: new Abstract: In recent years, machine learning has made significant progress in clinical outcome prediction, demonstrating increasingly accurate results. However, th

Mechanism-Faithful Queueing Simulation Model Translation with Large Language Model Support

ResearchDGX agent

arXiv:2601.06543v2 Announce Type: replace Abstract: Queueing simulation studies often require substantial manual effort to translate conceptual system descriptions into executable programs and to veri

StateVLM: A State-Aware Vision-Language Model for Robotic Affordance Reasoning

Model ReleasesDGX agent

arXiv:2605.03927v1 Announce Type: new Abstract: Vision-language models (VLMs) have shown remarkable performance in various robotic tasks, as they can perceive visual information and understand natural

5 May 2026

CNN-based Multi-In-Multi-Out Model for Efficient Spatiotemporal Prediction

Model ReleasesDGX agent

arXiv:2605.01277v1 Announce Type: new Abstract: Recently, Convolutional Neural Network (CNN) or Transformer architecture based models have been proposed to overcome the limitations of Recurrent Neural

Dispersion Loss Counteracts Embedding Condensation and Improves Generalization in Small Language Models

Model ReleasesDGX agent

arXiv:2602.00217v2 Announce Type: replace Abstract: Large language models (LLMs) achieve remarkable performance through ever-increasing parameter counts, but scaling incurs steep computational costs.

Fine-Tuning Impairs the Balancedness of Foundation Models in Long-tailed Personalized Federated Learning

Model ReleasesDGX agent

arXiv:2605.02247v1 Announce Type: new Abstract: Personalized federated learning (PFL) with foundation models has emerged as a promising paradigm enabling clients to adapt to heterogeneous data distrib

GIFT: Guided Fine-Tuning and Transfer for Enhancing Instruction-Tuned Language Models

TutorialsDGX agent

arXiv:2605.01256v1 Announce Type: new Abstract: A promising paradigm for adapting instruction-tuned language models is to learn task-specific updates on a pretrained base model and subsequently merge

The Pre-Training Study of Expanded-SPLADE Models on Web Document Titles

ResearchDGX agent

arXiv:2605.01407v1 Announce Type: cross Abstract: Masked Language Modeling (MLM) pre-training is one of the primary ways to initialize Neural Information Retrieval (IR) models prior to retrieval fine-

When Correct Isn't Usable: Improving Structured Output Reliability in Small Language Models

Model ReleasesDGX agent

arXiv:2605.02363v1 Announce Type: new Abstract: Deployed language models must produce outputs that are both correct and format-compliant. We study this structured-output reliability gap using two math

4 May 2026

How Well Does GPT-4o Understand Vision? Evaluating Multimodal Foundation Models on Standard Computer Vision Tasks

Model ReleasesDGX agent

arXiv:2507.01955v3 Announce Type: replace Abstract: Multimodal foundation models (MFMs), such as GPT-4o, have recently made remarkable progress. However, their detailed visual understanding beyond que

Jailbroken Frontier Models Retain Their Capabilities

Model ReleasesDGX agent

arXiv:2605.00267v1 Announce Type: new Abstract: As language model safeguards become more robust, attackers are pushed toward developing increasingly complex jailbreaks. Prior work has found that this

1 May 2026

haha our model likes to talk about goblins no of course we dont know why, we dont know why the model does anything - yes we are trying to make a superintelligent machine god, maybe it will like goblins too, we have no way of knowing what it will like, we hope it will like humans

TutorialsDGX agent

This Reddit post from r/ChatGPT humorously discusses an AI model's unexplained tendency to frequently mention goblins in its outputs, using this quirk to reflect on broader uncertainties around AI beh

MotuBrain: An Advanced World Action Model for Robot Control

SafetyDGX agent

arXiv:2604.27792v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models achieve strong semantic generalization but often lack fine-grained modeling of world dynamics. Recent work explores

29 Apr 2026

Learning Illumination Control in Diffusion Models

ResearchDGX agent

arXiv:2604.24877v1 Announce Type: new Abstract: Controlling illumination in images is essential for photography and visual content creation. While closed-source models have demonstrated impressive ill

Revisiting the Past: Data Unlearning with Model State History

ResearchDGX agent

arXiv:2506.20941v3 Announce Type: replace Abstract: Large language models are trained on massive corpora of web data, which may include private data, copyrighted material, factually inaccurate data, o

28 Apr 2026

EgoDyn-Bench: Evaluating Ego-Motion Understanding in Vision-Centric Foundation Models for Autonomous Driving

Model ReleasesDGX agent

arXiv:2604.22851v1 Announce Type: cross Abstract: While Vision-Language Models (VLMs) have advanced highlevel reasoning in autonomous driving, their ability to ground this reasoning in the underlying

Evaluating whether AI models would sabotage AI safety research

Model ReleasesDGX agent

arXiv:2604.24618v1 Announce Type: new Abstract: We evaluate the propensity of frontier models to sabotage or refuse to assist with safety research when deployed as AI research agents within a frontier

LinguDistill: Recovering Linguistic Ability in Vision-Language Models via Selective Cross-Modal Distillation

SafetyDGX agent

arXiv:2604.00829v3 Announce Type: replace-cross Abstract: Adapting pretrained language models (LMs) into vision-language models (VLMs) can degrade their native linguistic capability due to representat

On the Memorization of Consistency Distillation for Diffusion Models

ResearchDGX agent

arXiv:2604.23552v1 Announce Type: cross Abstract: Diffusion models are central to modern generative modeling, and understanding how they balance memorization and generalization is critical for reliabl

Open Models & a potential looming Haiku-apocalypse? 🚀 chart comparison UI courtesy of @OpenRouter there's a large class of problems where y…

Model ReleasesDGX agent

Open Models & a potential looming Haiku-apocalypse? 🚀 chart comparison UI courtesy of @OpenRouter there's a large class of problems where you don't need frontier intelligence. you want cheap, fast, an

Patching LLM Like Software: A Lightweight Method for Improving Safety Policy in Large Language Models

Model ReleasesDGX agent

arXiv:2511.08484v2 Announce Type: replace Abstract: We propose patching for large language models (LLMs) like software versions, a lightweight and modular approach for addressing safety vulnerabilitie

27 Apr 2026

Are Video Models Emerging as Zero-Shot Learners and Reasoners in Medical Imaging?

ResearchDGX agent

arXiv:2510.10254v2 Announce Type: replace Abstract: Recent advances in large generative models have shown that simple autoregressive formulations, when scaled appropriately, can exhibit strong zero-sh

26 Apr 2026

extremely happy that we are in q2 2026, and engineers i look up to are plowing a path for the local model future. yes, we are still beholden…

Model ReleasesDGX agent

extremely happy that we are in q2 2026, and engineers i look up to are plowing a path for the local model future. yes, we are still beholden to some lab publishing weights. but i take that over compan

24 Apr 2026

DeepSeek V4 - almost on the frontier, a fraction of the price

Model ReleasesDGX agent

Chinese AI lab DeepSeek's last model release was V3.2 (and V3.2 Speciale) last December. They just dropped the first of their hotly anticipated V4 series in the shape of two preview models, DeepSeek-V

Do LLMs Overthink Basic Math Reasoning? Benchmarking the Accuracy-Efficiency Tradeoff in Language Models

Model ReleasesDGX agent

arXiv:2507.04023v3 Announce Type: replace Abstract: Large language models (LLMs) achieve impressive performance on complex mathematical benchmarks yet sometimes fail on basic math reasoning while gene

Model page for more information and integrations: https://ollama.com/library/deepseek-v4-flash

Model ReleasesDGX agent

Ollama announced DeepSeek-v4-flash, a lightweight variant of the DeepSeek-v4 model, now available in their model library for local deployment and integration. The model page provides documentation, us

23 Apr 2026

OMIBench: Benchmarking Olympiad-Level Multi-Image Reasoning in Large Vision-Language Model

Model ReleasesDGX agent

arXiv:2604.20806v1 Announce Type: cross Abstract: Large vision-language models (LVLMs) have made substantial advances in reasoning tasks at the Olympiad level. Nevertheless, current Olympiad-level mul

Surrogate modeling for interpreting black-box LLMs in medical predictions

ApplicationsDGX agent

arXiv:2604.20331v1 Announce Type: cross Abstract: Large language models (LLMs), trained on vast datasets, encode extensive real-world knowledge within their parameters, yet their black-box nature obsc

22 Apr 2026

Beyond Coefficients: Forecast-Necessity Testing for Interpretable Causal Discovery in Nonlinear Time-Series Models

ApplicationsDGX agent

arXiv:2604.18751v1 Announce Type: cross Abstract: Nonlinear machine-learning models are increasingly used to discover causal relationships in time-series data, yet the interpretation of their outputs

Handling and Interpreting Missing Modalities in Patient Clinical Trajectories via Autoregressive Sequence Modeling

ApplicationsDGX agent

arXiv:2604.18753v1 Announce Type: cross Abstract: An active challenge in developing multimodal machine learning (ML) models for healthcare is handling missing modalities during training and deployment

IndiaFinBench: An Evaluation Benchmark for Large Language Model Performance on Indian Financial Regulatory Text

Model ReleasesDGX agent

arXiv:2604.19298v1 Announce Type: cross Abstract: We introduce IndiaFinBench, to our knowledge the first publicly available evaluation benchmark for assessing large language model (LLM) performance on

OmniVoice: Towards Omnilingual Zero-Shot Text-to-Speech with Diffusion Language Models

ResearchDGX agent

arXiv:2604.00688v3 Announce Type: replace Abstract: We present OmniVoice, a massively multilingual zero-shot text-to-speech (TTS) model that scales to over 600 languages. At its core is a novel diffus

21 Apr 2026

Finding Culture-Sensitive Neurons in Vision-Language Models

Model ReleasesDGX agent

arXiv:2510.24942v2 Announce Type: replace-cross Abstract: Despite their impressive performance, vision-language models (VLMs) still struggle on culturally situated inputs. To understand how VLMs proce

From Heads to Neurons: Causal Attribution and Steering in Multi-Task Vision-Language Models

Model ReleasesDGX agent

arXiv:2604.17941v1 Announce Type: cross Abstract: Recent work has increasingly explored neuron-level interpretation in vision-language models (VLMs) to identify neurons critical to final predictions.

Heterogeneity in Formal Linguistic Competence of Language Models: Is Data the Real Bottleneck?

TutorialsDGX agent

arXiv:2604.17930v1 Announce Type: new Abstract: Large Language Models (LLMs) exhibit a puzzling disparity in their formal linguistic competence: while they learn some linguistic phenomena with near-pe

HORIZON: A Benchmark for In-the-wild User Behaviour Modeling

Model ReleasesDGX agent

arXiv:2604.17259v1 Announce Type: cross Abstract: User behavior in the real world is diverse, cross-domain, and spans long time horizons. Existing user modeling benchmarks however remain narrow, focus

Reciprocal Co-Training (RCT): Coupling Gradient-Based and Non-Differentiable Models via Reinforcement Learning

TutorialsDGX agent

arXiv:2604.16378v1 Announce Type: new Abstract: Large language models (LLMs) and classical machine learning methods offer complementary strengths for predictive modeling, yet their fundamentally diffe

ReflexiCoder: Teaching Large Language Models to Self-Reflect on Generated Code and Self-Correct It via Reinforcement Learning

Model ReleasesDGX agent

arXiv:2603.05863v2 Announce Type: replace Abstract: While Large Language Models (LLMs) have revolutionized code generation, standard ``System 1'' approaches that generate solutions in a single forward

ToxiFrench: Benchmarking and Enhancing Language Models via CoT Fine-Tuning for French Toxicity Detection

Model ReleasesDGX agent

arXiv:2508.11281v3 Announce Type: replace Abstract: Detecting toxic content using language models is crucial yet challenging. While substantial progress has been made in English, toxicity detection in

Understanding Counting Mechanisms in Large Language and Vision-Language Models

ResearchDGX agent

arXiv:2511.17699v2 Announce Type: replace Abstract: Counting is one of the fundamental abilities of large language models (LLMs) and large vision-language models (LVLMs). This paper examines how these

20 Apr 2026

BAGEL: Benchmarking Animal Knowledge Expertise in Language Models

Model ReleasesDGX agent

arXiv:2604.16241v1 Announce Type: cross Abstract: Large language models have shown strong performance on broad-domain knowledge and reasoning benchmarks, but it remains unclear how well language model

Cost-Aware Model Orchestration for LLM-based Systems

ResearchDGX agent

arXiv:2512.01099v2 Announce Type: replace Abstract: As modern artificial intelligence (AI) systems become more advanced and capable, they can leverage a wide range of tools and models to perform compl

LLM attribution analysis across different fine-tuning strategies and model scales for automated code compliance

Model ReleasesDGX agent

arXiv:2604.15589v1 Announce Type: cross Abstract: Existing research on large language models (LLMs) for automated code compliance has primarily focused on performance, treating the models as black box

Using Large Language Models and Knowledge Graphs to Improve the Interpretability of Machine Learning Models in Manufacturing

ApplicationsDGX agent

arXiv:2604.16280v1 Announce Type: new Abstract: Explaining Machine Learning (ML) results in a transparent and user-friendly manner remains a challenging task of Explainable Artificial Intelligence (XA

17 Apr 2026

Cornfigurator: Automated Planning for Any-to-Any Multimodal Model Serving

ResearchDGX agent

arXiv:2512.14098v3 Announce Type: replace Abstract: Any-to-Any models are an emerging class of multimodal models that accept combinations of text and multimodal data as input and generate them as outp

← Previous
1…1819202122…990
Next →