AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
59,904 results
27 May 2026

InvokeAI 6.13 just released, its largest community-driven release ever. Adds full support for Anima & Qwen Image, support for API models (like GPT Image), support for Prompt Expansion & Image To Prompt, lasso & polygon tools, overhauled docs website and more

Model ReleasesDGX agent

InvokeAI 6.13 is the largest community-driven release of the software, adding full support for Anima & Qwen Image models, API model integration (such as GPT Image), and new features including Prompt E

MolPIF: A Parameter Interpolation Flow Model for Molecule Generation

Model ReleasesDGX agent

arXiv:2507.13762v4 Announce Type: replace Abstract: Motivation: Structure-based drug design (SBDD) has advanced with deep generative models, but bridging the gap between continuous atomic coordinates

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model ReleasesDGX agent

arXiv:2506.00250v4 Announce Type: replace Abstract: Large Language Models (LLMs) have achieved remarkable performance on a wide range of Natural Language Processing (NLP) benchmarks, often surpassing

PitchBench: Measuring Pitch Hearing in Audio-Language Models

Model ReleasesDGX agent

arXiv:2605.26176v1 Announce Type: cross Abstract: Audio-language models (ALMs) are increasingly used in real-world applications that require understanding music, from music tutoring and transcription

26 May 2026

Characterizing Linear Alignment Across Language Models

SafetyDGX agent

arXiv:2603.18908v4 Announce Type: replace Abstract: Language models increasingly appear to learn similar representations, despite differences in training objectives, architectures, and data modalities

Cross-Domain Generalization Limits of Vision Foundation Models in Facial Deepfake Detection

Model ReleasesDGX agent

arXiv:2605.24965v1 Announce Type: cross Abstract: The rapid evolution of generative models has enabled the creation of hyper-realistic facial deepfakes, exposing a critical vulnerability in modern dig

Efficient Long-Horizon Vision-Language-Action Models via Static-Dynamic Disentanglement

Model ReleasesDGX agent

arXiv:2602.03983v3 Announce Type: replace Abstract: Vision-Language-Action (VLA) models have recently emerged as a promising paradigm for generalist robotic control. Built upon vision-language model (

Gemma 4 adoption numbers outpacing Qwen 3.5/3.6 for the same sized models is a big shift in the international balance of influence via open …

Model ReleasesDGX agent

Gemma 4 adoption numbers outpacing Qwen 3.5/3.6 for the same sized models is a big shift in the international balance of influence via open models. Some ideas for what comes next, May 2026 Gemini Flas

Nano World Models: A Minimalist Implementation of Future Video Prediction

ResearchDGX agent

arXiv:2605.23993v1 Announce Type: cross Abstract: World models have become a central paradigm for learning predictive simulators that support generation, planning, and decision-making. Yet, despite ra

One-for-All Model Initialization with Frequency-Domain Knowledge

Model ReleasesDGX agent

arXiv:2603.07523v2 Announce Type: replace Abstract: Transferring knowledge by fine-tuning large-scale pre-trained networks has become a standard paradigm for downstream tasks, yet the knowledge of a p

Practical Quantum CIM Empowerment via All-Domestic-Core Agentic Large Model

AgentsDGX agent

arXiv:2605.23934v1 Announce Type: new Abstract: Quantum computing devices are recognized as powerful tools for solving NP-complete problems. However, the intricacy of their modeling presents notable b

Small Models, Strong Priors: Architectural Inductive Bias for Parameter-Efficient Neural PDE Solvers

Model ReleasesDGX agent

arXiv:2605.25949v1 Announce Type: cross Abstract: Neural PDE solvers have followed the scaling trajectory of vision and language, with recent foundation models reaching billions of parameters. We argu

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models

Model ReleasesDGX agent

arXiv:2506.18543v2 Announce Type: replace-cross Abstract: The rapid proliferation of Large Language Models (LLMs) has heightened concerns regarding their exposure to jailbreak attacks, which craft adv

T2S-MPC: Time-Embedded Online Adaptive Model Predictive Control for Time-Varying Dynamics

ResearchDGX agent

arXiv:2605.24852v1 Announce Type: new Abstract: Recent advances in learning-based model predictive control (MPC) have leveraged neural networks for online model learning, achieving strong performance

The Model Is Not the Product: A Dual-Pillar Architecture for Local-First Psychological Coaching

Model ReleasesDGX agent

arXiv:2605.24411v1 Announce Type: new Abstract: Existing language model applications struggle to meet the demand for emotionally oriented support, primarily due to their inability to maintain deep, pe

Topology-Driven Transferability Estimation of Medical Foundation Models for Segmentation

Model ReleasesDGX agent

arXiv:2602.23916v2 Announce Type: replace-cross Abstract: The advent of large-scale self-supervised learning (SSL) has produced a vast zoo of medical foundation models. However, selecting optimal medi

Towards Large Model Feature Coding

Model ReleasesDGX agent

arXiv:2605.24025v1 Announce Type: cross Abstract: Large models have delivered remarkable performance across a wide range of perception and generation tasks, yet practical deployment is increasingly co

TUBE: Tangent Upper Bound on Evidence for Discrete Diffusion Language Models

ResearchDGX agent

arXiv:2605.24292v1 Announce Type: new Abstract: Log-likelihood is a standard metric for evaluating generative models. Unfortunately, in contrast to autoregressive models (ARMs), discrete diffusion mod

UWM-JEPA: Predictive World Models That Imagine in Belief Space

Model ReleasesDGX agent

arXiv:2605.25313v1 Announce Type: cross Abstract: World models for partially observed environments must imagine multiple compatible hidden futures and steer between them under counterfactual actions.

25 May 2026

BURMESE-SAN: Burmese NLP Benchmark for Evaluating Large Language Models

Model ReleasesDGX agent

arXiv:2602.18788v3 Announce Type: replace Abstract: We introduce BURMESE-SAN, the first holistic benchmark that systematically evaluates large language models (LLMs) for Burmese across three core NLP

Evaluating Customized vs. Generalist Transformer-based Models for Legal Contract Classification

ApplicationsDGX agent

arXiv:2508.07849v2 Announce Type: replace Abstract: Despite advances in legal NLP, no comprehensive evaluation of Transformer-based models customized for legal tasks (referred to as `legal-specific' m

Exploring deep learning for Event-Based Saliency Prediction with a Transformer-based model

Model ReleasesDGX agent

arXiv:2605.23790v1 Announce Type: new Abstract: Saliency prediction has been extensively studied in RGB images and videos as a computational model of human visual attention. In contrast, predicting sa

InfiGFusion: Graph-on-Logits Distillation via Efficient Gromov-Wasserstein for Model Fusion

ResearchDGX agent

arXiv:2505.13893v2 Announce Type: replace Abstract: Recent advances in large language models (LLMs) have intensified efforts to fuse heterogeneous open-source models into a unified system that inherit

The Misattribution Gap: When Memory Poisoning Looks Like Model Failure in Agentic AI Systems

Model ReleasesDGX agent

arXiv:2605.22842v1 Announce Type: cross Abstract: Multi-agent AI pipelines typically assume that agent misconduct originates from model misalignment. We identify a structural failure in this assumptio

The physics of AI weather models

SafetyDGX agent

arXiv:2605.23778v1 Announce Type: cross Abstract: Could it be that AI weather models are solving physical equations, although they may not be the equations used by conventional NWP models? We compute

The Readout Shortcut: Positional Number Copying Dominates Arithmetic CoT Readout in Small Language Models

Model ReleasesDGX agent

arXiv:2605.22870v1 Announce Type: cross Abstract: Chain-of-thought (CoT) prompting is necessary for arithmetic in small language models, yet shuffling its steps preserves most performance. What does C

The Surprising Difficulty of Search in Model-Based Reinforcement Learning

Model ReleasesDGX agent

arXiv:2601.21306v2 Announce Type: replace-cross Abstract: This paper investigates search in model-based reinforcement learning (RL). Conventional wisdom holds that long-term predictions and compoundin

23 May 2026

HIDBench: Benchmarking Large Language Models for Host-Based Intrusion Detection

Model ReleasesDGX agent

arXiv:2605.21773v1 Announce Type: cross Abstract: Recent benchmark efforts have advanced the evaluation of large language models (LLMs) in cybersecurity, including tasks such as penetration testing an

Live Music Diffusion Models: Efficient Fine-Tuning and Post-Training of Interactive Diffusion Music Generators

SafetyDGX agent

arXiv:2605.22717v1 Announce Type: cross Abstract: Interactive streaming music generation promises the use of generative models for live performance and co-creation that is impossible with offline mode

Truncated Neural Likelihood Estimation for Simulation-Based Inference in State-Space Models

Model ReleasesDGX agent

arXiv:2605.21805v1 Announce Type: cross Abstract: State-space models (SSMs) are powerful probabilistic tools for modeling time-varying systems with latent dynamics. Inference in SSMs involves the esti

22 May 2026

Accelerated Test-Time Scaling with Model-Free Speculative Sampling

ResearchDGX agent

arXiv:2506.04708v3 Announce Type: replace Abstract: Language models have demonstrated remarkable capabilities in reasoning tasks through test-time scaling techniques like best-of-N sampling and tree s

BEiTScore: Reference-free Image Captioning Evaluation with an Efficient Cross-Encoder Model

Model ReleasesDGX agent

arXiv:2605.21728v1 Announce Type: cross Abstract: Image captioning evaluation remains a significant challenge, as vision-language models evolve toward more challenging capabilities such as generating

Beyond Acoustic Emotion Recognition: Multimodal Pathos Analysis in Political Speech Using LLM-Based and Acoustic Emotion Models

Model ReleasesDGX agent

arXiv:2605.22732v1 Announce Type: cross Abstract: We investigate whether acoustic emotion recognition models can serve as proxies for the Pathos dimension in political speech analysis, as operationali

DeepSeek v4: the most expected open-source model ever released, and the quietest landing

Model ReleasesDGX agent

After 15 months of incremental updates, leaks, and rumored leaks, DeepSeek released version 4. It arrived without the fanfare R1 and R1-preview commanded in early 2025. That quiet reception is the mos

GPT 5.5 seems to be improving in that direction now, and Claude models are getting worse at it, so I don't think there's a clear winner now.

Model ReleasesDGX agent

Jeremy Howard comments on comparative performance trends between GPT 5.5 and Claude models, noting that GPT 5.5 appears to be improving in a particular capability while Claude models are declining in

How Well Do Models Follow Visual Instructions? VIBE: A Systematic Benchmark for Visual Instruction-Driven Image Editing

Model ReleasesDGX agent

arXiv:2602.01851v2 Announce Type: replace Abstract: Recent generative models have achieved remarkable progress in image editing. However, existing systems and benchmarks remain largely text-guided. In

InnerQ: Hardware-Aware Tuning-Free Quantization of KV Cache for Large Language Models

Model ReleasesDGX agent

arXiv:2602.23200v2 Announce Type: replace-cross Abstract: When transformer-based language models are deployed for text generation, most of the inference time is spent in the decoding stage, where outp

LVDrive: Latent Visual Representation Enhanced Vision-Language-Action Autonomous Driving Model

Model ReleasesDGX agent

arXiv:2605.22089v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have emerged as a promising framework for end-to-end autonomous driving. However, existing VLAs typically rely on sp

Security Document Classification with a Fine-Tuned Local Large Language Model: Benchmark Data and an Open-Source System

Model ReleasesDGX agent

arXiv:2605.20368v1 Announce Type: cross Abstract: Organizations that scan documents for sensitive information face a practical problem. Cloud services require data to be sent to external infrastructur

Sub-exponential Growth Dynamics in Complex Systems: A Piecewise Power-Law Model for the Diffusion of New Words and Names

Model ReleasesDGX agent

arXiv:2511.04106v5 Announce Type: replace-cross Abstract: The diffusion of ideas and language in society has conventionally been described by S-shaped models, such as the logistic curve. However, the

Teaching Language Models to Forecast Research Success Through Comparative Idea Evaluation

Model ReleasesDGX agent

arXiv:2605.21491v1 Announce Type: cross Abstract: As language models accelerate scientific research by automating hypothesis generation and implementation, a new bottleneck emerges: evaluating and fil

The Neglected Baseline in Model Interpretation

ResearchDGX agent

arXiv:2605.22417v1 Announce Type: new Abstract: We observe that existing model interpretation methods generally ignore the baseline, and such neglect often results in imprecise or even incorrect inter

VDE Bench: Evaluating The Capability of Image Editing Models to Modify Visual Documents

Model ReleasesDGX agent

arXiv:2602.00122v2 Announce Type: replace Abstract: In recent years, image editing models have made significant progress, enabling users to manipulate visual content in a flexible and interactive mann

Visual-Advantage On-Policy Distillation for Vision-Language Models

Model ReleasesDGX agent

arXiv:2605.21924v1 Announce Type: new Abstract: On-policy knowledge distillation has proven effective for language models, yet its application to vision-language models (VLMs) remains underexplored. W

When Shared Knowledge Hurts: Spectral Over-Accumulation in Model Merging

ResearchDGX agent

arXiv:2602.05536v2 Announce Type: replace-cross Abstract: Model merging combines multiple fine-tuned models into a single model by adding their weight updates, providing a lightweight alternative to r

21 May 2026

An exponential mechanism based on quadratic approximations for fine-tuning machine learning models with privacy guarantees

Model ReleasesDGX agent

arXiv:2605.20521v1 Announce Type: new Abstract: Fine-tuning adapts a pretrained machine learning model to a small, sensitive dataset, but this process risks memorizing individual new data points, maki

AttriStory: Fine-grained Attribute Realization for Visual Storytelling with Diffusion Models

Model ReleasesDGX agent

arXiv:2605.20777v1 Announce Type: new Abstract: Visual storytelling with diffusion models has made impressive strides in maintaining character consistency across narrative scenes. However, a critical

Bayesian Preference Learning for Test-Time Steerable Reward Models

SafetyDGX agent

arXiv:2602.08819v2 Announce Type: replace-cross Abstract: Reward models are central to aligning language models with human preferences via reinforcement learning (RL). As RL is increasingly applied to

Do No Harm? Hallucination and Actor-Level Abuse in Web-Deployed Medical Large Language Models

SafetyDGX agent

arXiv:2605.20591v1 Announce Type: new Abstract: Medical large language models (LLMs), including custom medical GPTs (MedGPTs) and open-source models, are increasingly deployed on web platforms to prov

FullFlow: Upgrading Text-to-Image Flow Matching Models for Bidirectional Vision--Language Generation

Model ReleasesDGX agent

arXiv:2605.20316v1 Announce Type: new Abstract: Modern text-to-image diffusion models encode rich visual priors, but expose them only through one-way text-conditioned generation. Existing unified visi

HalluCXR: Benchmarking and Mitigating Hallucinations in Medical Vision-Language Models for Chest Radiograph Interpretation

Model ReleasesDGX agent

arXiv:2605.20469v1 Announce Type: new Abstract: Vision-language models (VLMs) are increasingly used for medical image interpretation, yet they frequently hallucinate, generating clinically plausible b

Optimization Hyper-parameter Laws for Large Language Models

Model ReleasesDGX agent

arXiv:2409.04777v4 Announce Type: replace Abstract: Large Language Models have driven significant AI advancements, yet their training is resource-intensive and highly sensitive to hyper-parameter sele

Towards the Anonymization of the Language Modeling

ApplicationsDGX agent

arXiv:2501.02407v3 Announce Type: replace Abstract: Rapid advances in Natural Language Processing (NLP) have revolutionized many fields, including healthcare. However, these advances raise significant

VLA-REPLICA: A Low-Cost, Reproducible Benchmark for Real-World Evaluation of Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2605.20774v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have shown strong promise for general-purpose robotic manipulation, but their real-world evaluation remains limited

WildRoadBench: A Wild Aerial Road-Damage Grounding Benchmark for Vision-Language Models and Autonomous Agents

Model ReleasesDGX agent

arXiv:2605.20306v1 Announce Type: new Abstract: We introduce WildRoadBench, a wild aerial road-damage grounding benchmark that couples direct visual grounding by vision-language models with autonomous

20 May 2026

A Systematic Failure Analysis of Vision Foundation Models for Open Set Iris Presentation Attack Detection

Model ReleasesDGX agent

arXiv:2605.19020v1 Announce Type: new Abstract: Vision foundation models have demonstrated strong transferability across diverse visual recognition tasks and are increasingly considered for biometric

Active Learning of Fractional-Order Viscoelastic Model Parameters for Realistic Haptic Rendering

Model ReleasesDGX agent

arXiv:2512.00667v2 Announce Type: replace-cross Abstract: Effective medical simulators necessitate realistic haptic rendering of biological tissues that exhibit viscoelastic material properties, such

Backdooring Masked Diffusion Language Models

Model ReleasesDGX agent

arXiv:2605.19262v1 Announce Type: new Abstract: Masked diffusion language models (MDLMs) are emerging as a compelling new paradigm for text generation, but their training-time security remains largely

Did we ever learn what model won gold at the IMO from OpenAI? It was a year ago and it was called an unreleased internal general purpose mod…

Model ReleasesDGX agent

Did we ever learn what model won gold at the IMO from OpenAI? It was a year ago and it was called an unreleased internal general purpose model back then. Has GPT-5.5 Pro Extended caught up with whatev

DLEBench: Evaluating Small-scale Object Editing Ability for Instruction-based Image Editing Model

Model ReleasesDGX agent

arXiv:2602.23622v2 Announce Type: replace-cross Abstract: Significant progress has been made in the field of Instruction-based Image Editing Models (IIEMs). However, while these models demonstrate pla

← Previous
1…5152535455…999
Next →