AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
59,904 results
30 Jun 2026

The strength of clinical evidence is recoverable from language model representations but not from their stated grades

Local AiDGX agent

arXiv:2606.29034v1 Announce Type: cross Abstract: Large language models (LLMs) increasingly summarize clinical evidence, where a claim's weight depends on how strongly it is supported. Yet these model

Towards Harnessing the Collaborative Power of Large and Small Models for Domain Tasks

ResearchDGX agent

arXiv:2504.17421v2 Announce Type: replace-cross Abstract: Large language models (LMs) offer broad generalization capabilities but require vast amounts of data and computational resources for domain-sp

Training Vision-Language-Action Models with Dense Embodied Chain-of-Thought Supervision

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.30552v1 Announce Type: cross Abstract: Cross-embodiment transfer in vision-language-action (VLA) models remains challenging because low-level state and action spaces differ fundamentally ac

Unified Enhancement of the Generalization and Robustness of Language Models via Bi-Stage Optimization

Model ReleasesDGX agent

arXiv:2503.16550v2 Announce Type: replace Abstract: Neural network language models (LMs) are confronted with significant challenges in generalization and robustness. Currently, many studies focus on i

VTEdit-Bench: A Comprehensive Benchmark for Multi-Reference Image Editing Models in Virtual Try-On

Model ReleasesDGX agent

arXiv:2603.11734v2 Announce Type: replace Abstract: As virtual try-on (VTON) continues to advance, a growing number of real-world scenarios have emerged, pushing beyond the ability of the existing spe

You wake up. It’s 2026. The newest open source AI model was trained by federal government employees, Big Balls and Groot.

Model ReleasesDGX agent

You wake up. It’s 2026. The newest open source AI model was trained by federal government employees, Big Balls and Groot. FULL INTERVIEW: Engineers Edward Coristine (@as400495) and Tai Groot (@taigrr)

Your Data Manifold is Secretly a Reward Model: Shell-LCC for Text-to-Video Generation

Local AiDGX agent

arXiv:2606.30248v1 Announce Type: new Abstract: Recent text-to-video (T2V) diffusion models rely heavily on auxiliary reward signals (e.g., via reward models or DPO) to align generated content with hu

29 Jun 2026

An Expectation-Maximization Algorithm for Training Clean Diffusion Models from Corrupted Observations

ResearchDGX agent

arXiv:2407.01014v2 Announce Type: replace Abstract: Diffusion models excel in solving imaging inverse problems due to their ability to model complex image priors. However, their reliance on large, cle

Benchmarking on Tasks That Matter: Dataset Selection for Preserving Model Rankings

Model ReleasesDGX agent

arXiv:2606.27997v1 Announce Type: new Abstract: Benchmarks of machine learning models often include many datasets, making evaluation expensive. For efficiency, it is preferable to perform evaluations

DataStates-LLM: Scalable Checkpointing for Transformer Models Using Composable State Providers

HardwareDGX agent

arXiv:2601.16956v1 Announce Type: cross Abstract: The rapid growth of Large Transformer-based models, specifically Large Language Models (LLMs), now scaling to trillions of parameters, has necessitate

FlexMoE: One-for-All Nested Intra-Expert Pruning for MoE Language Models

TutorialsDGX agent

arXiv:2606.27866v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) language models scale model ability with sparsely activated experts, making this architecture a standard recipe for modern larg

GAIA: A Data Flywheel System for Training GUI Test-Time Scaling Critic Models

Model ReleasesDGX agent

arXiv:2601.18197v2 Announce Type: replace Abstract: While Large Vision-Language Models (LVLMs) have significantly advanced GUI agents' capabilities in parsing textual instructions, interpreting screen

PEBS: Per-rater Empirical-Bayes Shrinkage for RLHF Reward-Model Calibration

Model ReleasesDGX agent

arXiv:2606.27578v1 Announce Type: cross Abstract: Reward models for Reinforcement Learning from Human Feedback (RLHF) pool preferences across thousands of annotators and fit one global affine calibrat

Quantum Generative Diffusion Model for Real-World Time Series

ApplicationsDGX agent

arXiv:2606.27561v1 Announce Type: new Abstract: Generative models have achieved remarkable success in data synthesis, though recent advances driven by increasing model scale have introduced challenges

RSICCLLM: A Multimodal Large Language Model for Remote Sensing Image Change Captioning

Model ReleasesDGX agent

arXiv:2606.28266v1 Announce Type: new Abstract: Remote Sensing Image Change Captioning (RSICC) aims to describe changes between bi-temporal remote sensing images and holds significant research and app

S^2-VLA: State-Space Guided Vision-Language-Action Models for Long-Horizon Manipulation

Model ReleasesDGX agent

arXiv:2606.27872v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have demonstrated strong capabilities in robotic manipulation, but their performance degrades significantly in lon

28 Jun 2026

Good guy DeepSeek gives us accelerated models The most interesting one here is Gemma4-12B, I presume vision included. Might be the best loca…

Model ReleasesDGX agent

Good guy DeepSeek gives us accelerated models The most interesting one here is Gemma4-12B, I presume vision included. Might be the best local model in its weight class now, by some margin Qwen 3.5 not

26 Jun 2026

6 Fingers, 1 Kidney: Natural Adversarial Medical Images Reveal Critical Weaknesses of Vision-Language Models

Model ReleasesDGX agent

arXiv:2512.04238v2 Announce Type: replace Abstract: Vision-language models (VLMs) are increasingly integrated into clinical workflows. However, existing benchmarks primarily assess performance on comm

Consistent Distributed Ranking of Generative Models via Kernel Distances

ResearchDGX agent

arXiv:2310.11714v5 Announce Type: replace Abstract: Ranking generative models based on the fidelity and diversity of their outputs is required to identify the best generator in a group of candidate ge

Dataset Usage Inference without Shadow Models or Held-out Data

ResearchDGX agent

arXiv:2606.26257v1 Announce Type: new Abstract: How much of my data was used to train a machine learning model? Dataset Usage Inference (DUI) aims to answer this by estimating what fraction of a datas

extsc{DiARC}: Distinguishing Positive and Negative Samples Helps Improving ARC-like Reasoning Ability of Large Language Models

Model ReleasesDGX agent

arXiv:2606.26530v1 Announce Type: cross Abstract: The Abstraction and Reasoning Corpus (ARC;~itealp{chollet2019measure}) contains tasks that require summarizing patterns from limited grid samples and

Great to see the new GPT-5.6 models finally announced. Sad to see this new release strategy where only a select few get access initially. No…

Model ReleasesDGX agent

Great to see the new GPT-5.6 models finally announced. Sad to see this new release strategy where only a select few get access initially. Not a win for our industry IMO. Open-source AI must win! Intro

Normalizing Flows are Capable Models for Continuous Control

SafetyDGX agent

arXiv:2505.23527v4 Announce Type: replace Abstract: Modern reinforcement learning (RL) algorithms have found success by using powerful probabilistic models, such as transformers, energy-based models,

Paying More Attention to Visual Tokens in Self-Evolving Large Multimodal Models

SafetyDGX agent

arXiv:2606.27373v1 Announce Type: new Abstract: Recently, self-evolving large multimodal models (LMMs) have received attention for improving visual reasoning in a purely unsupervised setting. However,

Post-Training Recipe, More Than Model Family, Shapes Multi-Agent LLM Conversational Behavior

Model ReleasesDGX agent

arXiv:2606.20632v2 Announce Type: replace-cross Abstract: Multi-LLM systems use multiple language models to deliberate, judge each other's outputs, or coordinate as agents. Their value depends on the

Reproducibility Study of 'AlphaEdit: Null-Space Constrained Knowledge Editing for Language Models'

Model ReleasesDGX agent

arXiv:2606.26783v1 Announce Type: cross Abstract: Fang et al. (2025) introduced a null-space constrained projection, named AlphaEdit, for locate-then-edit knowledge editing methods, theoretically guar

Simulation-based inference for rapid Bayesian parameter estimation in epidemiological models: a comparison with MCMC

Model ReleasesDGX agent

arXiv:2606.27286v1 Announce Type: new Abstract: Mechanistic epidemiological models are widely used to support infectious disease forecasting and public-health decision making. Bayesian calibration of

SITUATION EXPLAINED: 70% of frontier model queries could run locally for free. @ClementDelangue, co-founder and CEO of @huggingface: 'There …

Local AiDGX agent

SITUATION EXPLAINED: 70% of frontier model queries could run locally for free. @ClementDelangue, co-founder and CEO of @huggingface: 'There was an interesting study from Stanford published last year s

Where Do Models Find Happiness? Emotion Vectors in Open-Source LLMs

Model ReleasesDGX agent

arXiv:2606.26987v1 Announce Type: cross Abstract: Recent work identified emotion vectors in Claude Sonnet 4.5, which are internal representations that encode emotion concepts, causally influence behav

25 Jun 2026

A Probabilistic Framework for LLM-Based Model Discovery

AgentsDGX agent

arXiv:2602.18266v2 Announce Type: replace Abstract: Automated methods for discovering mechanistic simulator models from observational data offer a promising path toward accelerating scientific progres

Benchmarking Deep Learning Models for Laryngeal Cancer Staging Using the LaryngealCT Dataset

Model ReleasesDGX agent

arXiv:2510.11047v2 Announce Type: replace Abstract: Laryngeal cancer imaging research lacks standardised public datasets to enable reproducible deep learning (DL) model development. We present Larynge

Concept Removal for Frontier Image Generative Models

ResearchDGX agent

arXiv:2606.25548v1 Announce Type: new Abstract: Image generative models are trained on massive, largely uncurated internet-scale datasets that contain undesirable visual concepts. Efficiently removing

Detect, Unlearn, Restore: Defending Text Summarization Models Against Data Poisoning

Model ReleasesDGX agent

arXiv:2606.26036v1 Announce Type: new Abstract: Training-time data poisoning during fine-tuning poses a significant threat to large language models (LLMs) deployed for abstractive text summarization,

How Robust is OCR-Reasoning? Evaluating OCR-Reasoning Robustness of Vision-Language Models under Visual Perturbations

Model ReleasesDGX agent

arXiv:2606.26041v1 Announce Type: cross Abstract: Vision-language models (VLMs) have achieved strong performance on OCR-based benchmarks and increasingly focused on text-rich understanding, but their

In-Context World Modeling for Robotic Control

Model ReleasesDGX agent

arXiv:2606.26025v1 Announce Type: cross Abstract: Modern Vision-Language-Action (VLA) models often fail to generalize to novel setups, such as altered camera viewpoints or robot morphologies, because

InvestPhilBench: A Multi-Layer Dynamic Benchmark for Evaluating Large Language Model Procedural Reasoning in Expert Investment Philosophy

Model ReleasesDGX agent

arXiv:2606.25984v1 Announce Type: cross Abstract: Large language models are increasingly deployed as investment research assistants, yet no benchmark tests whether they can accurately reconstruct and

MedLayBench-V: A Large-Scale Benchmark for Expert-Lay Semantic Alignment in Medical Vision Language Models

Model ReleasesDGX agent

arXiv:2604.05738v2 Announce Type: replace Abstract: Medical Vision-Language Models (Med-VLMs) have achieved expert-level proficiency in interpreting diagnostic imaging. However, current models are pre

Model-agnostic Mitigation Strategies of Data Imbalance for Regression

Model ReleasesDGX agent

arXiv:2506.01486v2 Announce Type: replace Abstract: Data imbalance persists as a pervasive challenge in regression tasks, introducing bias in model performance and undermining predictive reliability.

Steering Vision-Language Models with Joint Sparse Autoencoders

Model ReleasesDGX agent

arXiv:2606.25657v1 Announce Type: new Abstract: Sparse Autoencoders (SAEs) have shown promise for analyzing language models, but applying them to vision-language models (VLMs) often yields representat

24 Jun 2026

ABACUS: Adapting Unified Foundation Model for Bridging Image Count Understanding and Generation

Model ReleasesDGX agent

arXiv:2606.23835v1 Announce Type: new Abstract: ABACUS is a unified vision-language model that handles object counting, crowd counting, referring-expression counting, and count-faithful image generati

Gemini 3 Pro was the first model to achieve at least 23% on ARC-AGI-2, which it did in November, 2025 (it actually scored 31%). So the 8-12 …

Model ReleasesDGX agent

Gemini 3 Pro was the first model to achieve at least 23% on ARC-AGI-2, which it did in November, 2025 (it actually scored 31%). So the 8-12 month gap between closed and open weights models still seems

HANCLIP: A Family of Hyperbolic Angular Negation Vision Language Models

Model ReleasesDGX agent

arXiv:2606.23843v1 Announce Type: new Abstract: Vision-Language Models (VLMs) are typically pre-trained on large-scale image-text datasets to capture semantic correspondences between visual content an

Quant Convergence: Bridging Classical Value Investing and Modern Factor Models for Systematic Equity Selection

SafetyDGX agent

arXiv:2606.24575v1 Announce Type: new Abstract: Modern finance relies heavily on complex machine learning models to find patterns in the stock market. However, as these AI models get more complicated,

Variational Model Merging for Pareto Front Estimation in Multitask Finetuning

Model ReleasesDGX agent

arXiv:2412.08147v2 Announce Type: replace-cross Abstract: Pareto fronts are useful to find good task-mixing strategies for multitask finetuning, but they are also costly to compute. To reduce costs, r

23 Jun 2026

2D Versus 3D Diffusion for In Silico Training of Interventional X-ray AI Models

ResearchDGX agent

arXiv:2606.21414v1 Announce Type: cross Abstract: The ability to synthesize realistic X-ray images has catalyzed the development of AI models for X-ray image-guided procedures, which otherwise suffer

Finetuning Vision-Language-Action Models Requires Fewer Layers Than You Think

Model ReleasesDGX agent

arXiv:2606.20246v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) models pre-trained on massive video-robot datasets have revolutionized robotic manipulation, yet their multi-billion pa

Generating Fearful Images: Investigating Potential Emotional Biases in Image-Generation Models

Model ReleasesDGX agent

arXiv:2411.05985v3 Announce Type: replace-cross Abstract: This paper examines potential biases and inconsistencies in the emotions evoked by images produced by generative artificial intelligence (AI)

Holo-World: Unified Camera, Object and Weather Control for Video World Model

Model ReleasesDGX agent

arXiv:2606.20083v2 Announce Type: replace Abstract: Video world models are moving toward preserving an observed world under controllable camera and object motion while allowing its environmental state

MapReason-OSM: Can Vision-Language Models Make Graph-Verifiable Mobility Decisions from Street Maps ?

Model ReleasesDGX agent

arXiv:2606.22597v1 Announce Type: new Abstract: Vision-language models (VLMs) are increasingly used to read maps for logistics, delivery, and accessible navigation, where the output is an actionable d

Right Knowledge, Wrong Answer: Test-Time Steering for Temporal Fact Conflicts in Open-Weight Language Models

Model ReleasesDGX agent

arXiv:2606.20959v1 Announce Type: new Abstract: Large language models can store both outdated facts and newer superseding facts in their parameters, but standard prompting may still elicit the outdate

SegTME-UNI2: A Foundation Model-Based Framework for Generalisable Multiclass Cell Segmentation and LLM-Driven Tumour Microenvironment Characterisation in Histopathology

Model ReleasesDGX agent

arXiv:2606.17702v2 Announce Type: replace Abstract: Characterising the tumour microenvironment (TME) from routine H&E-stained histology images requires simultaneous cell segmentation, feature extracti

SparseWorld: Enhancing End-to-End Autonomous Driving via World Models with Sparse Scene Representation

Model ReleasesDGX agent

arXiv:2605.24354v2 Announce Type: replace Abstract: Recently, world models have made significant progress in enhancing end-to-end driving systems through both future situation forecasting and improved

Tapered Language Models

Model ReleasesDGX agent

arXiv:2606.23670v1 Announce Type: new Abstract: Modern language models, including transformer, recurrent, and memory-based variants, share a common chassis: a stack of identical layers in which parame

World Action Models: A Survey

ResearchDGX agent

arXiv:2606.20781v1 Announce Type: cross Abstract: World Action Models (WAMs) are embodied predictive-action models that make a forecast of the future available to action. Recent WAMs repurpose large v

22 Jun 2026

GLM-5.2 leads open weights models and sits at #3 overall on GDPval-AA, a real-world agentic work benchmark GLM-5.2 from @Zai_org scores 1524…

Model ReleasesDGX agent

GLM-5.2 leads open weights models and sits at #3 overall on GDPval-AA, a real-world agentic work benchmark GLM-5.2 from @Zai_org scores 1524 Elo on GDPval-AA, which measures performance on real-world,

That's right, an open, MIT-licensed model beating GPT-5.5 (xhigh) on real-world agentic work! 🔥 Available for free on @huggingface for anyo…

Model ReleasesDGX agent

That's right, an open, MIT-licensed model beating GPT-5.5 (xhigh) on real-world agentic work! 🔥 Available for free on @huggingface for anyone to build on top off GLM-5.2 leads open weights models and

11 Jun 2026

Fast Speech Foundation Model Distillation Using Interleaved Stacking

ResearchDGX agent

arXiv:2606.11766v1 Announce Type: cross Abstract: Distilling a large speech foundation model (SFM) into an efficient student model has been successfully applied to low-resource environments. Although

Improving Cross-Format Robustness in Language Models with Multi-Format Training

Model ReleasesDGX agent

arXiv:2606.11643v1 Announce Type: new Abstract: Large language models often remain sensitive to answer format: a question solved correctly in one form may fail in another semantically equivalent form.

Risk Under Pressure: Compute-Aware Evaluation of Adversarial Robustness in Language Models

SafetyDGX agent

arXiv:2606.11409v1 Announce Type: cross Abstract: Adversarial robustness evaluations of large language models (LLMs) typically report attack success rate (ASR) under fixed query budgets, implicitly tr

The Standard Interpretable Model: A general theory of interpretable machine learning to deductively design interpretable methods using Lagrangian mechanics

Model ReleasesDGX agent

arXiv:2606.12289v1 Announce Type: cross Abstract: As Artificial Intelligence models grow in complexity, interpretability has become an indispensable tool for understanding, debugging, and controlling

← Previous
1…4849505152…999
Next →