AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
59,860 results
1 May 2026

Auto-FlexSwitch: Efficient Dynamic Model Merging via Learnable Task Vector Compression

Model ReleasesDGX agent

arXiv:2604.28109v1 Announce Type: new Abstract: Model merging has attracted attention as an effective path toward multi-task adaptation by integrating knowledge from multiple task-specific models. Amo

Simple Self-Conditioning Adaptation for Masked Diffusion Models

ResearchDGX agent

arXiv:2604.26985v1 Announce Type: cross Abstract: Masked diffusion models (MDMs) generate discrete sequences by iterative denoising under an absorbing masking process. In standard masked diffusion, if

29 Apr 2026

Combating Visual Neglect and Semantic Drift in Large Multimodal Models for Enhanced Cross-Modal Retrieval

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases
DGX agent

arXiv:2604.25273v1 Announce Type: new Abstract: Despite significant progress in Unified Multimodal Retrieval (UMR) powered by Large Multimodal Models (LMMs), existing embedding methods primarily focus

Comparing Data Assimilation and Likelihood-Based Inference on Latent State Estimation in Agent-Based Models

Model ReleasesDGX agent

arXiv:2509.17625v2 Announce Type: replace Abstract: In this paper, we present the first systematic comparison of Data Assimilation (DA) and Likelihood-Based Inference (LBI) in the context of an Agent-

Heterogeneous Variational Inference for Markov Degradation Hazard Models: Discretized Mixture with Interpretable Clusters

ApplicationsDGX agent

arXiv:2604.24818v1 Announce Type: new Abstract: Bayesian finite mixture models can identify discrete risk clusters (low-risk vs. high-risk equipment), but face three critical bottlenecks: (1) insuffic

Thinking About Thinking: Evaluating Reasoning in Post-Trained Language Models

SafetyDGX agent

arXiv:2510.16340v2 Announce Type: replace Abstract: Recent advances in post-training techniques have endowed Large Language Models (LLMs) with enhanced capabilities for tackling complex, logic-intensi

28 Apr 2026

A Divergence-Based Method for Weighting and Averaging Model Predictions

ResearchDGX agent

arXiv:2604.24172v1 Announce Type: cross Abstract: This paper uses a minimum divergence framework to introduce a new way of calculating model weights that can be used to average probabilistic predictio

Diffusion Model as a Generalist Segmentation Learner

ResearchDGX agent

arXiv:2604.24575v1 Announce Type: new Abstract: Diffusion models are primarily trained for image synthesis, yet their denoising trajectories encode rich, spatially aligned visual priors. In this paper

DreamAudio: Customized Text-to-Audio Generation with Diffusion Models

Model ReleasesDGX agent

arXiv:2509.06027v3 Announce Type: replace-cross Abstract: With the development of large-scale diffusion-based and language-modeling-based generative models, impressive progress has been achieved in te

Dual-domain Multi-path Self-supervised Diffusion Model for Accelerated MRI Reconstruction

ApplicationsDGX agent

arXiv:2503.18836v2 Announce Type: replace-cross Abstract: Magnetic resonance imaging (MRI) is a vital diagnostic tool, but its inherently long acquisition times reduce clinical efficiency and patient

Evolve: A Persistent Knowledge Lifecycle for Small Language Models

Model ReleasesDGX agent

arXiv:2604.23424v1 Announce Type: cross Abstract: Evolve pairs a small local language model with a persistent, teacher-compiled knowledge store -- refined through sleep consolidation and usage-driven

FreqCache: Accelerating Embodied VLN Models with Adaptive Frequency-Guided Token Caching

ResearchDGX agent

arXiv:2604.24391v1 Announce Type: new Abstract: Vision-Language-Navigation (VLN) models exhibit excellent navigation accuracy but incur high computational overhead. Token caching has emerged as a prom

In-Sync: Adaptation of Speech Aware Large Language Models for ASR with Word Level Timestamp Predictions

SafetyDGX agent

arXiv:2604.22817v1 Announce Type: cross Abstract: Recent advances in speech-aware language models have coupled strong acoustic encoders with large language models, enabling systems that move beyond tr

IntentVLM: Open-Vocabulary Intention Recognition through Forward-Inverse Modeling with Video-Language Models

ResearchDGX agent

arXiv:2604.24002v1 Announce Type: cross Abstract: Improving the effectiveness of human-robot interaction requires social robots to accurately infer human goals through robust intention understanding.

LAMP: Extracting Local Decision Surfaces From Large Language Models

Local AiDGX agent

arXiv:2505.11772v3 Announce Type: replace Abstract: We introduce LAMP (Local Attribution Mapping Probe), a method that shines light onto a black-box language model's decision surface and studies how r

LLMind: Bio-inspired Training-free Adaptive Visual Representations for Vision-Language Models

ResearchDGX agent

arXiv:2603.14882v2 Announce Type: replace Abstract: Vision-Language Models (VLMs) typically assume a uniform spatial fidelity across the entire field of view of visual inputs, dedicating equal precisi

MIMIC: A Generative Multimodal Foundation Model for Biomolecules

ResearchDGX agent

arXiv:2604.24506v1 Announce Type: new Abstract: Biological function emerges from coupled constraints across sequence, structure, regulation, evolution, and cellular context, yet most foundation models

RefEvo: Agentic Design with Co-Evolutionary Verification for Agile Reference Model Generation

Model ReleasesDGX agent

arXiv:2604.24218v1 Announce Type: cross Abstract: As the complexity of System-on-Chip (SoC) designs grows, the shift-left paradigm necessitates the rapid development of high-fidelity reference models

Scaling Properties of Continuous Diffusion Spoken Language Models

Model ReleasesDGX agent

arXiv:2604.24416v1 Announce Type: cross Abstract: Speech-only spoken language models (SLMs) lag behind text and text-speech models in performance, with recent discrete autoregressive (AR) SLMs indicat

Tandem: Riding Together with Large and Small Language Models for Efficient Reasoning

TutorialsDGX agent

arXiv:2604.23623v1 Announce Type: new Abstract: Recent advancements in large language models (LLMs) have catalyzed the rise of reasoning-intensive inference paradigms, where models perform explicit st

Towards Holistic Evaluation of Large Audio-Language Models: A Comprehensive Survey

SafetyDGX agent

arXiv:2505.15957v4 Announce Type: replace-cross Abstract: With advancements in large audio-language models (LALMs), which enhance large language models (LLMs) with auditory capabilities, these models

Ulterior Motives: Detecting Misaligned Reasoning in Continuous Thought Models

Model ReleasesDGX agent

arXiv:2604.23460v1 Announce Type: new Abstract: Chain-of-Thought (CoT) reasoning has emerged as a key technique for eliciting complex reasoning in Large Language Models (LLMs). Although interpretable,

When to Commit? Towards Variable-Size Self-Contained Blocks for Discrete Diffusion Language Models

ResearchDGX agent

arXiv:2604.23994v1 Announce Type: cross Abstract: Discrete diffusion language models (dLLMs) enable parallel token updates with bidirectional attention, yet practical generation typically adopts block

WISE-FM:Operation-Aware, Engineering-Informed Foundation Model for Multi-Task Well Design

Model ReleasesDGX agent

arXiv:2604.23767v1 Announce Type: new Abstract: Deploying machine learning models across diverse well portfolios requires generalisation to wells with design parameters outside the training distributi

27 Apr 2026

Atlas-Alignment: Making Interpretability Transferable Across Language Models

Model ReleasesDGX agent

arXiv:2510.27413v2 Announce Type: replace-cross Abstract: Interpretability is crucial for building safe, reliable, and controllable language models, yet existing interpretability pipelines remain cost

RedVLA: Physical Red Teaming for Vision-Language-Action Models

SafetyDGX agent

arXiv:2604.22591v1 Announce Type: new Abstract: The real-world deployment of Vision-Language-Action (VLA) models remains limited by the risk of unpredictable and irreversible physical harm. However, w

Towards Safe Mobility: A Unified Transportation Foundation Model enabled by Open-Ended Vision-Language Dataset

SafetyDGX agent

arXiv:2604.22260v1 Announce Type: cross Abstract: Urban transportation systems face growing safety challenges that require scalable intelligence for emerging smart mobility infrastructures. While rece

24 Apr 2026

Accurate predictive model of band gap with selected important features based on explainable machine learning

ResearchDGX agent

arXiv:2503.04492v3 Announce Type: replace-cross Abstract: In the rapidly advancing field of materials informatics, nonlinear machine learning models have demonstrated exceptional predictive capabiliti

Exploring Continual Fine-Tuning for Enhancing Language Ability in Large Language Model

TutorialsDGX agent

arXiv:2410.16006v3 Announce Type: replace Abstract: A common challenge towards the adaptability of Large Language Models (LLMs) is their ability to learn new languages over time without hampering the

Low-Rank Adaptation Redux for Large Models

Model ReleasesDGX agent

arXiv:2604.21905v1 Announce Type: new Abstract: Low-rank adaptation (LoRA) has emerged as the de facto standard for parameter-efficient fine-tuning (PEFT) of foundation models, enabling the adaptation

23 Apr 2026

From Data to Theory: Autonomous Large Language Model Agents for Materials Science

Model ReleasesDGX agent

arXiv:2604.19789v1 Announce Type: new Abstract: We present an autonomous large language model (LLM) agent for end-to-end, data-driven materials theory development. The model can choose an equation for

Graph-Theoretic Models for the Prediction of Molecular Measurements

Model ReleasesDGX agent

arXiv:2604.19840v1 Announce Type: new Abstract: Graph-theoretic approaches offer simplicity, interpretability, and low computational cost for molecular property prediction. Among these, the model prop

Kimi K2.6 becomes the #1 open model on MathArena!

Model ReleasesDGX agent

Kimi K2.6 achieved the top ranking on MathArena, a benchmark for evaluating mathematical problem-solving capabilities in open-source language models. This announcement highlights the model's superior

Rashomon Sets and Model Multiplicity in Federated Learning

Model ReleasesDGX agent

arXiv:2602.09520v2 Announce Type: replace Abstract: The Rashomon set captures the collection of models that achieve near-identical empirical performance yet may differ substantially in their decision

The Model Says Walk: How Surface Heuristics Override Implicit Constraints in LLM Reasoning

Model ReleasesDGX agent

arXiv:2603.29025v2 Announce Type: replace-cross Abstract: Large language models systematically fail when a salient surface cue conflicts with an unstated feasibility constraint. We study this through

What Language Models Know But Don't Say: Non-Generative Prior Extraction for Generalization

ApplicationsDGX agent

arXiv:2601.17609v2 Announce Type: replace Abstract: In domains like medicine and finance, large-scale labeled data is costly and often unavailable, leading to models trained on small datasets that str

22 Apr 2026

Are Large Language Models Economically Viable for Industry Deployment?

Model ReleasesDGX agent

arXiv:2604.19342v1 Announce Type: new Abstract: Generative AI-powered by Large Language Models (LLMs)-is increasingly deployed in industry across healthcare decision support, financial analytics, ente

Deep sprite-based image models: An analysis

Model ReleasesDGX agent

arXiv:2604.19480v1 Announce Type: new Abstract: While foundation models drive steady progress in image segmentation and diffusion algorithms compose always more realistic images, the seemingly simple

MiroThinker: Pushing the Performance Boundaries of Open-Source Research Agents via Model, Context, and Interactive Scaling

Model ReleasesDGX agent

arXiv:2511.11793v3 Announce Type: replace Abstract: We present MiroThinker v1.0, an open-source research agent designed to advance tool-augmented reasoning and information-seeking capabilities. Unlike

Qwen3.6-27B: Flagship-Level Coding in a 27B Dense Model

Model ReleasesDGX agent

Qwen3.6-27B: Flagship-Level Coding in a 27B Dense Model Big claims from Qwen about their latest open weight model: Qwen3.6-27B delivers flagship-level agentic coding performance, surpassing the previo

Regression with Large Language Models for Materials and Molecular Property Prediction

Model ReleasesDGX agent

arXiv:2409.06080v2 Announce Type: replace-cross Abstract: We demonstrate the ability of large language models (LLMs) to perform material and molecular property regression tasks, a significant deviatio

Remask, Don't Replace: Token-to-Mask Refinement in Masked Diffusion Language Models

ResearchDGX agent

arXiv:2604.18738v1 Announce Type: new Abstract: Masked diffusion language models such as LLaDA2.1 rely on Token-to-Token (T2T) editing to correct their own generation errors: whenever a different toke

VecHeart: Holistic Four-Chamber Cardiac Anatomy Modeling via Hybrid VecSets

Model ReleasesDGX agent

arXiv:2604.19403v1 Announce Type: new Abstract: Accurate cardiac anatomy modeling requires the model to be able to handle intricate interrelations among structures. In this paper, we propose VecHeart,

Visual Adversarial Attack on Vision-Language Models for Autonomous Driving

SafetyDGX agent

arXiv:2411.18275v2 Announce Type: replace Abstract: Vision-language models (VLMs) have significantly advanced autonomous driving (AD) by enhancing reasoning capabilities. However, these models remain

21 Apr 2026

Bi-LoRA: Efficient Sharpness-Aware Minimization for Fine-Tuning Large-Scale Models

Model ReleasesDGX agent

arXiv:2508.19564v2 Announce Type: replace Abstract: Fine-tuning large-scale pre-trained models with limited data presents significant challenges for generalization. While Sharpness-Aware Minimization

Counterfactual Modeling with Fine-Tuned LLMs for Health Intervention Design and Sensor Data Augmentation

Model ReleasesDGX agent

arXiv:2601.14590v2 Announce Type: replace Abstract: Counterfactual explanations (CFEs) provide human-centric interpretability by identifying the minimal, actionable changes required to alter a machine

DeepRitzSplit Neural Operator for Phase-Field Models via Energy Splitting

ResearchDGX agent

arXiv:2604.18261v1 Announce Type: cross Abstract: The multi-scale and non-linear nature of phase-field models of solidification requires fine spatial and temporal discretization, leading to long compu

Flow marching for a generative PDE foundation model

Model ReleasesDGX agent

arXiv:2509.18611v2 Announce Type: replace Abstract: Pretraining on large-scale collections of PDE-governed spatiotemporal trajectories has recently shown promise for building generalizable models of d

FOREVER: Forgetting Curve-Inspired Memory Replay for Language Model Continual Learning

Model ReleasesDGX agent

arXiv:2601.03938v2 Announce Type: replace-cross Abstract: Continual learning (CL) for large language models (LLMs) aims to enable sequential knowledge acquisition without catastrophic forgetting. Memo

Grokking of Diffusion Models: Case Study on Modular Addition

ApplicationsDGX agent

arXiv:2604.17673v1 Announce Type: new Abstract: Despite their empirical success, how diffusion models generalize remains poorly understood from a mechanistic perspective. We demonstrate that diffusion

IncreFA: Breaking the Static Wall of Generative Model Attribution

Model ReleasesDGX agent

arXiv:2604.17736v1 Announce Type: new Abstract: As AI generative models evolve at unprecedented speed, image attribution has become a moving target. New diffusion, adversarial and autoregressive gener

On the Predictive Power of Representation Dispersion in Language Models

Model ReleasesDGX agent

arXiv:2506.24106v2 Announce Type: replace Abstract: We show that a language model's ability to predict text is tightly linked to the breadth of its embedding space: models that spread their contextual

One Adapts to Any: Meta Reward Modeling for Personalized LLM Alignment

SafetyDGX agent

arXiv:2601.18731v2 Announce Type: replace Abstract: Alignment of Large Language Models (LLMs) aims to align outputs with human preferences, and personalized alignment further adapts models to individu

SmoGVLM: A Small, Graph-enhanced Vision-Language Model

ResearchDGX agent

arXiv:2604.16517v1 Announce Type: cross Abstract: Large vision-language models (VLMs) achieve strong performance on multimodal tasks but often suffer from hallucination and poor grounding in knowledge

SpeakerSleuth: Can Large Audio-Language Models Judge Speaker Consistency across Multi-turn Dialogues?

Model ReleasesDGX agent

arXiv:2601.04029v2 Announce Type: replace Abstract: Large Audio-Language Models (LALMs) as judges have emerged as a prominent approach for evaluating speech generation quality, yet their ability to as

Stability-Weighted Decoding for Diffusion Language Models

ResearchDGX agent

arXiv:2604.17068v1 Announce Type: new Abstract: Diffusion large language models (dLLMs) enable parallel text generation by iteratively denoising a fully masked sequence, unmasking a subset of masked t

TensorHub: Rethinking AI Model Hub with Tensor-Centric Compression

ApplicationsDGX agent

arXiv:2604.17104v1 Announce Type: cross Abstract: Modern AI models are growing rapidly in size and redundancy, leading to significant storage and distribution challenges in model hubs. We present Tens

20 Apr 2026

Advancing Intelligent Sequence Modeling: Evolution, Trade-offs, and Applications of State- Space Architectures from S4 to Mamba

ResearchDGX agent

arXiv:2503.18970v3 Announce Type: replace Abstract: Structured State Space Models (SSMs) have emerged as a transformative paradigm in sequence modeling, addressing critical limitations of Recurrent Ne

Do Vision-Language Models Truly Perform Vision Reasoning? A Rigorous Study of the Modality Gap

Model ReleasesDGX agent

arXiv:2604.16256v1 Announce Type: cross Abstract: Reasoning in vision-language models (VLMs) has recently attracted significant attention due to its broad applicability across diverse downstream tasks

Elucidating the SNR-t Bias of Diffusion Probabilistic Models

SafetyDGX agent

arXiv:2604.16044v1 Announce Type: new Abstract: Diffusion Probabilistic Models have demonstrated remarkable performance across a wide range of generative tasks. However, we have observed that these mo

← Previous
1…3132333435…998
Next →