AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
59,904 results
2 Jun 2026

From Segments to Scenes: Temporal Understanding in Autonomous Driving via Vision-Language Model

Model ReleasesDGX agent

arXiv:2512.05277v3 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) are increasingly deployed as the perception and reasoning backbone of autonomous agents acting in the wild, with

KliniskVestBERT: BERT Model Specialised to Norwegian Clinical Texts

Model ReleasesDGX agent

arXiv:2606.01904v1 Announce Type: cross Abstract: The increasing application of Natural Language Processing (NLP) in healthcare demands language models specifically attuned to the complexities of clin

Med-URWKV{ag}: Toward Enhanced Pretrained Pure VRWKV Models for Medical Image Segmentation

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2506.10858v2 Announce Type: replace-cross Abstract: Medical image segmentation is a fundamental task in computer-aided diagnosis and treatment. Existing approaches based on CNNs, ViTs, Mamba, an

Microsoft's new MAI models

Model ReleasesDGX agent

Microsoft announced two new text LLMs this morning - MAI-Thinking-1 (reasoning, 35B parameters, available to 'select early partners') and MAI-Code-1-Flash (5B parameters, 'purpose-built for GitHub Cop

MIRROR: A Multi-Agent Framework with Iterative Adaptive Revision and Hierarchical Retrieval for Optimization Modeling in Operations Research

AgentsDGX agent

arXiv:2602.03318v3 Announce Type: replace Abstract: Operations Research (OR) relies on expert-driven modeling-a slow and fragile process ill-suited to novel scenarios. While large language models (LLM

Mitigating Reward Hacking in RLHF via Bayesian Non-negative Reward Modeling

ResearchDGX agent

arXiv:2602.10623v2 Announce Type: replace-cross Abstract: Reward models learned from human preferences are central to aligning large language models (LLMs) via reinforcement learning from human feedba

Model-Based Quality Assessment for Massively Multilingual Parallel Data

Model ReleasesDGX agent

arXiv:2606.00285v1 Announce Type: new Abstract: Large-scale multilingual bitext often contains two distinct problems: non-parallel sentence pairs and low-quality translations. We decompose model-based

On the Limits of LLM Adaptability: Impact of Model-Internalized Priors on Annotation Task Performance

SafetyDGX agent

arXiv:2606.00467v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used for zero-shot annotation and LLM-as-a-judge tasks, yet their reliability hinges on how model-intern

Random Erasing vs. Model Inversion: A Promising Defense or a False Hope?

ResearchDGX agent

arXiv:2409.01062v4 Announce Type: replace-cross Abstract: Model Inversion (MI) attacks pose a significant privacy threat by reconstructing private training data from machine learning models. While exi

Self-Healing Agentic Orchestrators for Reliable Tool-Augmented Large Language Model Systems

Model ReleasesDGX agent

arXiv:2606.01416v1 Announce Type: new Abstract: Tool-augmented large language model (LLM) agents rely on orchestration layers that coordinate planning, retrieval, tool invocation, validation, memory,

Truthful AI Advisors: A Pre-Specified Benchmark for Large Language Model Honesty Under Preference Misalignment

Model ReleasesDGX agent

arXiv:2606.01456v1 Announce Type: cross Abstract: Large language models are increasingly deployed as advisors whose objective is not aligned with the user's: recommenders optimize for engagement, sale

1 Jun 2026

CodeGolf Bench: A Multi-Language Benchmark for Evaluating Concise Code Generation Capabilities of Large Language Models

Model ReleasesDGX agent

arXiv:2605.30394v1 Announce Type: cross Abstract: This paper introduces Code Bench, a benchmark capable of evaluating Large Language Models (LLMs) concise code generation abilities in 60 programming l

DriveMA: Driving Vision-Language-Action Models with verifiable Meta-Actions

Model ReleasesDGX agent

arXiv:2605.31271v1 Announce Type: new Abstract: Driving Vision-Language-Action Models (Driving VLAs) aim to use language to improve end-to-end planning, but the language-action gap limits this promise

Float8@2bits: Entropy Coding Enables Data-Free Model Compression

Model ReleasesDGX agent

arXiv:2601.22787v2 Announce Type: replace Abstract: Post-training compression is currently divided into two contrasting regimes. On the one hand, fast, data-free, and model-agnostic methods (e.g., NF4

ImmigrationQA: A Source-Grounded Dataset and Small-Model Adaptation for U.S. Immigration Law

Model ReleasesDGX agent

arXiv:2605.30589v1 Announce Type: cross Abstract: U.S. immigration law spans thousands of pages of official policy, federal regulations, and procedural guidance that change frequently and carry high s

Language Models Learn Constructional Semantics, Not To Mention Syntax: Investigating LM Understanding of Paired-Focus Constructions

Model ReleasesDGX agent

arXiv:2605.31586v1 Announce Type: cross Abstract: Grasping the semantics of rare constructions (form-meaning pairings) has been shown to be a challenging problem that has currently only been solved by

LiMuon: Light and Fast Muon Optimizer for Large Models

ResearchDGX agent

arXiv:2509.14562v3 Announce Type: replace Abstract: Large models recently are widely applied in machine learning, so efficient training of large models has received widespread attention. More recently

Representation Forcing for Bottleneck-Free Unified Multimodal Models

TutorialsDGX agent

arXiv:2605.31604v1 Announce Type: new Abstract: Unified multimodal models (UMMs) aim to handle perception and generation in a single model. Yet existing UMMs still rely on a frozen, separately pretrai

SAC-Opt: Semantic Anchors for Iterative Correction in Optimization Modeling

ResearchDGX agent

arXiv:2510.05115v3 Announce Type: replace Abstract: Large language models (LLMs) have opened new paradigms in optimization modeling by enabling the generation of executable solver code from natural la

29 May 2026

Adapting Multilingual Embedding Models to Turkish via Cross-Lingual Tokenizer Surgery and Offline Distillation

Model ReleasesDGX agent

arXiv:2605.29992v1 Announce Type: new Abstract: Sentence embeddings are a foundational component for semantic search, clustering, classification, and retrieval-augmented generation. This paper present

Aggregate Models, Not Explanations: Improving Feature Importance Estimation

ResearchDGX agent

arXiv:2602.11760v2 Announce Type: replace-cross Abstract: Feature-importance methods show promise in transforming machine learning models from predictive engines into tools for scientific discovery. H

Contrastive Representation Regularization for Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2510.01711v3 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models have shown strong capabilities in robot manipulation by leveraging rich representations from pre-trained V

Improving Full Waveform Inversion in Large Model Era

Model ReleasesDGX agent

arXiv:2603.00377v2 Announce Type: replace Abstract: Full Waveform Inversion (FWI) is a highly nonlinear and ill-posed problem that aims to recover subsurface velocity maps from surface-recorded seismi

In-Context Reward Adaptation for Robust Preference Modeling

SafetyDGX agent

arXiv:2605.30323v1 Announce Type: cross Abstract: Reinforcement Learning from Human Feedback (RLHF) typically relies on static reward models to align Large Language Models with human preferences. Howe

libhmm: A Modern C++20 Library for Hidden Markov Models with Correct MLE Emission M-Steps

Model ReleasesDGX agent

arXiv:2605.29208v1 Announce Type: cross Abstract: We describe libhmm, a C++20 library for Hidden Markov Model parameter estimation, sequence decoding, and model selection. libhmm addresses two gaps in

LoopFM: Learning frOm HistOrical RePresentations of Foundation Model for Recommendation

Model ReleasesDGX agent

arXiv:2605.29280v1 Announce Type: cross Abstract: Knowledge distillation (KD) transfers a single scalar prediction from a large foundation model (FM) to compact vertical models (VMs), suffering from d

MechELK: A Mechanistic Interpretability Framework for Eliciting Latent Knowledge in Large Language Models

Model ReleasesDGX agent

arXiv:2605.28825v1 Announce Type: new Abstract: Large language models (LLMs) frequently encode factual and reasoning knowledge in their internal representations that is not faithfully reflected in the

Mitigating Hallucination in Vision-Language Models through Barrier-Regulated Adaptive Closed-form Steering

Model ReleasesDGX agent

arXiv:2605.29881v1 Announce Type: cross Abstract: Large vision-language models (LVLMs) often hallucinate objects that are not present in the input image, largely because visual grounding weakens as de

Model Merging by Output-Space Projection

ResearchDGX agent

arXiv:2605.29101v1 Announce Type: new Abstract: Model merging combines fine-tuned checkpoints into a single multi-task model without retraining. Existing methods - such as task arithmetic, model soups

open models are having a moment!

AgentsDGX agent

open models are having a moment! The latest finding in the LangSmith Signal: Open Models are having a moment. 1 in 3 AI teams ran an open-weights model in April 2026, up from 1 in 5 nine months ago. T

open models will become a huge force for coding.

AgentsDGX agent

open models will become a huge force for coding. The latest finding in the LangSmith Signal: Open Models are having a moment. 1 in 3 AI teams ran an open-weights model in April 2026, up from 1 in 5 ni

Procedural Pretraining: Warming Up Language Models with Abstract Data

ResearchDGX agent

arXiv:2601.21725v2 Announce Type: replace Abstract: Pretraining language models directly on web-scale corpora is the de facto paradigm. We study an alternative where the model is initially exposed to

Qwen-VLA: Unifying Vision-Language-Action Modeling across Tasks, Environments, and Robot Embodiments

Model ReleasesDGX agent

arXiv:2605.30280v1 Announce Type: cross Abstract: Embodied intelligence is often studied through specialized models for individual tasks such as manipulation or navigation, resulting in fragmented cap

ReasonBreak: Probing Vulnerabilities in Reasoning-Enabled Vision-Language-Action Models for Autonomous Driving

Model ReleasesDGX agent

arXiv:2605.29114v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models with integrated reasoning have been proposed for end-to-end autonomous driving, assuming a tight coupling between

When Models Disagree: Rethinking LLM Evaluation for Public Comment Analysis

ResearchDGX agent

arXiv:2605.29025v1 Announce Type: new Abstract: Federal agencies are deploying large language models (LLMs) to categorize public comment corpora, where the model's organization of the record shapes wh

YoCausal: How Far is Video Generation from World Model? A Causality Perspective

Model ReleasesDGX agent

arXiv:2605.30346v1 Announce Type: new Abstract: As video diffusion models (VDMs) advance toward world models, a key question arises: do they truly understand causality, or merely overfit to statistica

28 May 2026

Argument Quality Assessment with Large Language Models: A Pairwise Bradley-Terry Approach

Model ReleasesDGX agent

arXiv:2605.28313v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities in tasks related to reasoning and judgment. However, assessing the quality of arg

Assessing Factual Music Comprehension in Large Audio Language Models

Model ReleasesDGX agent

arXiv:2511.05550v2 Announce Type: replace-cross Abstract: Large audio language models (LALMs) leverage multimodal representations to generate open-ended answers to natural language queries about audio

Benchmarking Ultrasound Foundation Models for Fetal Plane Classification

Model ReleasesDGX agent

arXiv:2605.27796v1 Announce Type: cross Abstract: Ultrasound is widely used in obstetric care due to its safety, accessibility, and real-time imaging. However, interpretation remains operator-dependen

BlazeEdit: Generalist Image Editing on Mobile Devices with Image-to-Image Diffusion Models

Model ReleasesDGX agent

arXiv:2605.28067v1 Announce Type: new Abstract: The remarkable generation quality of modern diffusion models often comes at the cost of massive parameter counts, which necessitate server-side inferenc

Can Decision Trees Teach Large Language Models? Distilling Verbalized Knowledge for Molecular Property Prediction

Model ReleasesDGX agent

arXiv:2603.12344v2 Announce Type: replace Abstract: Molecular Property Prediction (MPP) is a fundamental problem in drug discovery that has recently attracted growing attention. Large Language Models

CFDTwin: An open-source GUI and Python toolkit for POD-NN surrogate modeling of ANSYS Fluent simulations

Model ReleasesDGX agent

arXiv:2605.27725v1 Announce Type: cross Abstract: High-fidelity computational fluid dynamics (CFD) is widely used for thermal-fluid design, but repeated CFD solves remain expensive for design optimiza

Compositional Generalization in Autoregressive Models via Logit Composition

ResearchDGX agent

arXiv:2605.28304v1 Announce Type: new Abstract: Composing autoregressive models remains a core challenge in understanding how large language models can combine behaviors or skills learned across tasks

Do Clinical Models Change Treatment Decisions?

Model ReleasesDGX agent

arXiv:2605.28129v1 Announce Type: new Abstract: Clinical foundation models are evaluated with factual or exam-style medical QA, but treatment decisions must change when patient context changes. We int

Entropy-aware Masking for Masked Language Modeling

ResearchDGX agent

arXiv:2605.28526v1 Announce Type: new Abstract: Masked language modeling has become a standard pretraining objective for training encoder-based language models. In this approach, certain tokens in the

FLORO: A Multimodal Geospatial Foundation Model for Ecological Remote Sensing Across Sensors and Scales

Model ReleasesDGX agent

arXiv:2605.28174v1 Announce Type: cross Abstract: Foundation models offer a promising route to transferable remote sensing representations, but many current approaches depend on very large pretraining

Gamma-World: Generative Multi-Agent World Modeling Beyond Two Players

Model ReleasesDGX agent

arXiv:2605.28816v1 Announce Type: new Abstract: World models for interactive video generation have largely focused on single-agent settings, where future observations are generated from a single contr

Local MDI+: Local Feature Importances for Tree-Based Models

Model ReleasesDGX agent

arXiv:2506.08928v2 Announce Type: replace Abstract: Tree-based ensembles such as random forests remain the go-to for tabular data over deep learning models due to their prediction performance and comp

Nano Banana 2 and Nano Banana Pro are generally available, and already powering creative workflows

Model ReleasesDGX agent

Organizations are unlocking entirely new ways to use image generation and editing across their industries. To drive next-generation experiences, businesses are embedding AI directly into creative, age

📢Qwen3.7-Max just hit #3 on ITbench-AA — a fresh benchmark testing how well models handle real-world enterprise IT tasks, agentic-style. 🔧…

Model ReleasesDGX agent

📢Qwen3.7-Max just hit #3 on ITbench-AA — a fresh benchmark testing how well models handle real-world enterprise IT tasks, agentic-style. 🔧Agentic era, go with Qwen.🏃🏃 Artificial Analysis and IBM Resea

Stay Fair! Ensuring Group Fairness in Diffusion Models Across Guidance Scales

Model ReleasesDGX agent

arXiv:2605.28036v1 Announce Type: new Abstract: Diffusion models steer conditional generation with a tunable guidance scale to trade off prompt alignment and diversity. However, existing debiasing tec

We're releasing Paris 2.0, which, to our knowledge, is the world's first decentralized trained video generation model. We benchmarked it aga…

Model ReleasesDGX agent

We're releasing Paris 2.0, which, to our knowledge, is the world's first decentralized trained video generation model. We benchmarked it against a monolithic model trained on the same data and compute

27 May 2026

A closer look at what we released today 🧵 ESMC is a language model trained on billions of protein sequences spanning the full diversity of …

Model ReleasesDGX agent

A closer look at what we released today 🧵 ESMC is a language model trained on billions of protein sequences spanning the full diversity of life. Trained across 2.8 billion sequences, the model is expo

Advancing Creative Physical Intelligence in Large Multimodal Models

Model ReleasesDGX agent

arXiv:2605.26396v1 Announce Type: new Abstract: Large multimodal models (LMMs) have rapidly advanced in perception and reasoning; however, it remains unclear whether these capabilities generalize to d

Auditing and Fixing Economic Validity in Tabular Foundation Models for Discrete Choice

SafetyDGX agent

arXiv:2605.26559v1 Announce Type: cross Abstract: Tabular foundation models achieve strong accuracy on choice prediction tasks, but their predictions often violate the economic logic those tasks requi

Beyond Holistic Models: Systematic Component-level Benchmarking of Deep Multivariate Time-Series Forecasting

Model ReleasesDGX agent

arXiv:2605.26562v1 Announce Type: new Abstract: While previous research in multivariate time series forecasting has focused on developing complex holistic models, this work advocates for a shift towar

CleanSurvival: Automated data preprocessing for time-to-event models using reinforcement learning

Model ReleasesDGX agent

arXiv:2502.03946v5 Announce Type: replace Abstract: Data preprocessing is often paid little attention in machine learning, despite its potentially significant impact on model performance. While automa

Continual Model-Based Reinforcement Learning with Hypernetworks

ResearchDGX agent

arXiv:2009.11997v3 Announce Type: replace-cross Abstract: Effective planning in model-based reinforcement learning (MBRL) and model-predictive control (MPC) relies on the accuracy of the learned dynam

Guiding Token-Sparse Diffusion Models

Model ReleasesDGX agent

arXiv:2601.01608v2 Announce Type: replace Abstract: Diffusion models deliver high quality in image synthesis but remain expensive during training and inference. Recent works have leveraged the inheren

InterSketch: An Interleaved Reasoning Model with Self-correcting Visual Sketch and Stepwise Reward

Model ReleasesDGX agent

arXiv:2605.26520v1 Announce Type: cross Abstract: While vision-language models (VLMs) have exhibited multi-turn visual reasoning capabilities, their reasoning trajectories remain relatively shallow an

← Previous
1…5051525354…999
Next →