AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
59,903 results
Model Releases

ML-Bench&Guard: Policy-Grounded Multilingual Safety Benchmark and Guardrail for Large Language Models

DGX agent

arXiv:2605.00689v1 Announce Type: new Abstract: As Large Language Models (LLMs) are increasingly deployed in cross-linguistic contexts, ensuring safety in diverse regulatory and cultural environments

model-releasesarxiv-cs-cl
4 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Themis: Training Robust Multilingual Code Reward Models for Flexible Multi-Criteria Scoring

DGX agent

arXiv:2605.00754v1 Announce Type: cross Abstract: Reward models (RMs) have become an indispensable fixture of the language model (LM) post-training playbook, enabling policy alignment and test-time sc

model-releasesarxiv-cs-lg
4 May 2026
Model Releases

Auto-FlexSwitch: Efficient Dynamic Model Merging via Learnable Task Vector Compression

DGX agent

arXiv:2604.28109v1 Announce Type: new Abstract: Model merging has attracted attention as an effective path toward multi-task adaptation by integrating knowledge from multiple task-specific models. Amo

model-releasesarxiv-cs-lg
1 May 2026
Research

Simple Self-Conditioning Adaptation for Masked Diffusion Models

DGX agent

arXiv:2604.26985v1 Announce Type: cross Abstract: Masked diffusion models (MDMs) generate discrete sequences by iterative denoising under an absorbing masking process. In standard masked diffusion, if

researcharxiv-cs-ai
1 May 2026
Model Releases

Combating Visual Neglect and Semantic Drift in Large Multimodal Models for Enhanced Cross-Modal Retrieval

DGX agent

arXiv:2604.25273v1 Announce Type: new Abstract: Despite significant progress in Unified Multimodal Retrieval (UMR) powered by Large Multimodal Models (LMMs), existing embedding methods primarily focus

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

Comparing Data Assimilation and Likelihood-Based Inference on Latent State Estimation in Agent-Based Models

DGX agent

arXiv:2509.17625v2 Announce Type: replace Abstract: In this paper, we present the first systematic comparison of Data Assimilation (DA) and Likelihood-Based Inference (LBI) in the context of an Agent-

model-releasesarxiv-cs-lg
29 Apr 2026
Applications

Heterogeneous Variational Inference for Markov Degradation Hazard Models: Discretized Mixture with Interpretable Clusters

DGX agent

arXiv:2604.24818v1 Announce Type: new Abstract: Bayesian finite mixture models can identify discrete risk clusters (low-risk vs. high-risk equipment), but face three critical bottlenecks: (1) insuffic

applicationsarxiv-cs-lg
29 Apr 2026
Safety

Thinking About Thinking: Evaluating Reasoning in Post-Trained Language Models

DGX agent

arXiv:2510.16340v2 Announce Type: replace Abstract: Recent advances in post-training techniques have endowed Large Language Models (LLMs) with enhanced capabilities for tackling complex, logic-intensi

safetyarxiv-cs-cl
29 Apr 2026
Research

A Divergence-Based Method for Weighting and Averaging Model Predictions

DGX agent

arXiv:2604.24172v1 Announce Type: cross Abstract: This paper uses a minimum divergence framework to introduce a new way of calculating model weights that can be used to average probabilistic predictio

researcharxiv-cs-lg
28 Apr 2026
Research

Diffusion Model as a Generalist Segmentation Learner

DGX agent

arXiv:2604.24575v1 Announce Type: new Abstract: Diffusion models are primarily trained for image synthesis, yet their denoising trajectories encode rich, spatially aligned visual priors. In this paper

researcharxiv-cs-cv
28 Apr 2026
Model Releases

DreamAudio: Customized Text-to-Audio Generation with Diffusion Models

DGX agent

arXiv:2509.06027v3 Announce Type: replace-cross Abstract: With the development of large-scale diffusion-based and language-modeling-based generative models, impressive progress has been achieved in te

model-releasesarxiv-cs-ai
28 Apr 2026
Applications

Dual-domain Multi-path Self-supervised Diffusion Model for Accelerated MRI Reconstruction

DGX agent

arXiv:2503.18836v2 Announce Type: replace-cross Abstract: Magnetic resonance imaging (MRI) is a vital diagnostic tool, but its inherently long acquisition times reduce clinical efficiency and patient

applicationsarxiv-cs-ai
28 Apr 2026
Model Releases

Evolve: A Persistent Knowledge Lifecycle for Small Language Models

DGX agent

arXiv:2604.23424v1 Announce Type: cross Abstract: Evolve pairs a small local language model with a persistent, teacher-compiled knowledge store -- refined through sleep consolidation and usage-driven

model-releasesarxiv-cs-cl
28 Apr 2026
Research

FreqCache: Accelerating Embodied VLN Models with Adaptive Frequency-Guided Token Caching

DGX agent

arXiv:2604.24391v1 Announce Type: new Abstract: Vision-Language-Navigation (VLN) models exhibit excellent navigation accuracy but incur high computational overhead. Token caching has emerged as a prom

researcharxiv-cs-ro
28 Apr 2026
Safety

In-Sync: Adaptation of Speech Aware Large Language Models for ASR with Word Level Timestamp Predictions

DGX agent

arXiv:2604.22817v1 Announce Type: cross Abstract: Recent advances in speech-aware language models have coupled strong acoustic encoders with large language models, enabling systems that move beyond tr

safetyarxiv-cs-cl
28 Apr 2026
Research

IntentVLM: Open-Vocabulary Intention Recognition through Forward-Inverse Modeling with Video-Language Models

DGX agent

arXiv:2604.24002v1 Announce Type: cross Abstract: Improving the effectiveness of human-robot interaction requires social robots to accurately infer human goals through robust intention understanding.

researcharxiv-cs-ai
28 Apr 2026
Local Ai

LAMP: Extracting Local Decision Surfaces From Large Language Models

DGX agent

arXiv:2505.11772v3 Announce Type: replace Abstract: We introduce LAMP (Local Attribution Mapping Probe), a method that shines light onto a black-box language model's decision surface and studies how r

local-aiarxiv-cs-lg
28 Apr 2026
Research

LLMind: Bio-inspired Training-free Adaptive Visual Representations for Vision-Language Models

DGX agent

arXiv:2603.14882v2 Announce Type: replace Abstract: Vision-Language Models (VLMs) typically assume a uniform spatial fidelity across the entire field of view of visual inputs, dedicating equal precisi

researcharxiv-cs-cv
28 Apr 2026
Research

MIMIC: A Generative Multimodal Foundation Model for Biomolecules

DGX agent

arXiv:2604.24506v1 Announce Type: new Abstract: Biological function emerges from coupled constraints across sequence, structure, regulation, evolution, and cellular context, yet most foundation models

researcharxiv-cs-ai
28 Apr 2026
Model Releases

RefEvo: Agentic Design with Co-Evolutionary Verification for Agile Reference Model Generation

DGX agent

arXiv:2604.24218v1 Announce Type: cross Abstract: As the complexity of System-on-Chip (SoC) designs grows, the shift-left paradigm necessitates the rapid development of high-fidelity reference models

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Scaling Properties of Continuous Diffusion Spoken Language Models

DGX agent

arXiv:2604.24416v1 Announce Type: cross Abstract: Speech-only spoken language models (SLMs) lag behind text and text-speech models in performance, with recent discrete autoregressive (AR) SLMs indicat

model-releasesarxiv-cs-ai
28 Apr 2026
Tutorials

Tandem: Riding Together with Large and Small Language Models for Efficient Reasoning

DGX agent

arXiv:2604.23623v1 Announce Type: new Abstract: Recent advancements in large language models (LLMs) have catalyzed the rise of reasoning-intensive inference paradigms, where models perform explicit st

tutorialsarxiv-cs-ai
28 Apr 2026
Safety

Towards Holistic Evaluation of Large Audio-Language Models: A Comprehensive Survey

DGX agent

arXiv:2505.15957v4 Announce Type: replace-cross Abstract: With advancements in large audio-language models (LALMs), which enhance large language models (LLMs) with auditory capabilities, these models

safetyarxiv-cs-ai
28 Apr 2026
Model Releases

Ulterior Motives: Detecting Misaligned Reasoning in Continuous Thought Models

DGX agent

arXiv:2604.23460v1 Announce Type: new Abstract: Chain-of-Thought (CoT) reasoning has emerged as a key technique for eliciting complex reasoning in Large Language Models (LLMs). Although interpretable,

model-releasesarxiv-cs-ai
28 Apr 2026
Research

When to Commit? Towards Variable-Size Self-Contained Blocks for Discrete Diffusion Language Models

DGX agent

arXiv:2604.23994v1 Announce Type: cross Abstract: Discrete diffusion language models (dLLMs) enable parallel token updates with bidirectional attention, yet practical generation typically adopts block

researcharxiv-cs-cl
28 Apr 2026
Model Releases

WISE-FM:Operation-Aware, Engineering-Informed Foundation Model for Multi-Task Well Design

DGX agent

arXiv:2604.23767v1 Announce Type: new Abstract: Deploying machine learning models across diverse well portfolios requires generalisation to wells with design parameters outside the training distributi

model-releasesarxiv-cs-lg
28 Apr 2026
Model Releases

Atlas-Alignment: Making Interpretability Transferable Across Language Models

DGX agent

arXiv:2510.27413v2 Announce Type: replace-cross Abstract: Interpretability is crucial for building safe, reliable, and controllable language models, yet existing interpretability pipelines remain cost

model-releasesarxiv-cs-ai
27 Apr 2026
Safety

RedVLA: Physical Red Teaming for Vision-Language-Action Models

DGX agent

arXiv:2604.22591v1 Announce Type: new Abstract: The real-world deployment of Vision-Language-Action (VLA) models remains limited by the risk of unpredictable and irreversible physical harm. However, w

safetyarxiv-cs-ro
27 Apr 2026
Safety

Towards Safe Mobility: A Unified Transportation Foundation Model enabled by Open-Ended Vision-Language Dataset

DGX agent

arXiv:2604.22260v1 Announce Type: cross Abstract: Urban transportation systems face growing safety challenges that require scalable intelligence for emerging smart mobility infrastructures. While rece

safetyarxiv-cs-ai
27 Apr 2026
Research

Accurate predictive model of band gap with selected important features based on explainable machine learning

DGX agent

arXiv:2503.04492v3 Announce Type: replace-cross Abstract: In the rapidly advancing field of materials informatics, nonlinear machine learning models have demonstrated exceptional predictive capabiliti

researcharxiv-cs-lg
24 Apr 2026
Tutorials

Exploring Continual Fine-Tuning for Enhancing Language Ability in Large Language Model

DGX agent

arXiv:2410.16006v3 Announce Type: replace Abstract: A common challenge towards the adaptability of Large Language Models (LLMs) is their ability to learn new languages over time without hampering the

tutorialsarxiv-cs-cl
24 Apr 2026
Model Releases

Low-Rank Adaptation Redux for Large Models

DGX agent

arXiv:2604.21905v1 Announce Type: new Abstract: Low-rank adaptation (LoRA) has emerged as the de facto standard for parameter-efficient fine-tuning (PEFT) of foundation models, enabling the adaptation

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

From Data to Theory: Autonomous Large Language Model Agents for Materials Science

DGX agent

arXiv:2604.19789v1 Announce Type: new Abstract: We present an autonomous large language model (LLM) agent for end-to-end, data-driven materials theory development. The model can choose an equation for

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Graph-Theoretic Models for the Prediction of Molecular Measurements

DGX agent

arXiv:2604.19840v1 Announce Type: new Abstract: Graph-theoretic approaches offer simplicity, interpretability, and low computational cost for molecular property prediction. Among these, the model prop

model-releasesarxiv-cs-lg
23 Apr 2026
Model Releases

Kimi K2.6 becomes the #1 open model on MathArena!

DGX agent

Kimi K2.6 achieved the top ranking on MathArena, a benchmark for evaluating mathematical problem-solving capabilities in open-source language models. This announcement highlights the model's superior

model-releaseskimi-moonshot--x
23 Apr 2026
Model Releases

Rashomon Sets and Model Multiplicity in Federated Learning

DGX agent

arXiv:2602.09520v2 Announce Type: replace Abstract: The Rashomon set captures the collection of models that achieve near-identical empirical performance yet may differ substantially in their decision

model-releasesarxiv-cs-lg
23 Apr 2026
Model Releases

The Model Says Walk: How Surface Heuristics Override Implicit Constraints in LLM Reasoning

DGX agent

arXiv:2603.29025v2 Announce Type: replace-cross Abstract: Large language models systematically fail when a salient surface cue conflicts with an unstated feasibility constraint. We study this through

model-releasesarxiv-cs-ai
23 Apr 2026
Applications

What Language Models Know But Don't Say: Non-Generative Prior Extraction for Generalization

DGX agent

arXiv:2601.17609v2 Announce Type: replace Abstract: In domains like medicine and finance, large-scale labeled data is costly and often unavailable, leading to models trained on small datasets that str

applicationsarxiv-cs-cl
23 Apr 2026
Model Releases

Are Large Language Models Economically Viable for Industry Deployment?

DGX agent

arXiv:2604.19342v1 Announce Type: new Abstract: Generative AI-powered by Large Language Models (LLMs)-is increasingly deployed in industry across healthcare decision support, financial analytics, ente

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

Deep sprite-based image models: An analysis

DGX agent

arXiv:2604.19480v1 Announce Type: new Abstract: While foundation models drive steady progress in image segmentation and diffusion algorithms compose always more realistic images, the seemingly simple

model-releasesarxiv-cs-cv
22 Apr 2026
Model Releases

MiroThinker: Pushing the Performance Boundaries of Open-Source Research Agents via Model, Context, and Interactive Scaling

DGX agent

arXiv:2511.11793v3 Announce Type: replace Abstract: We present MiroThinker v1.0, an open-source research agent designed to advance tool-augmented reasoning and information-seeking capabilities. Unlike

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

Qwen3.6-27B: Flagship-Level Coding in a 27B Dense Model

DGX agent

Qwen3.6-27B: Flagship-Level Coding in a 27B Dense Model Big claims from Qwen about their latest open weight model: Qwen3.6-27B delivers flagship-level agentic coding performance, surpassing the previo

model-releasessimon-willison
22 Apr 2026
Model Releases

Regression with Large Language Models for Materials and Molecular Property Prediction

DGX agent

arXiv:2409.06080v2 Announce Type: replace-cross Abstract: We demonstrate the ability of large language models (LLMs) to perform material and molecular property regression tasks, a significant deviatio

model-releasesarxiv-cs-lg
22 Apr 2026
Research

Remask, Don't Replace: Token-to-Mask Refinement in Masked Diffusion Language Models

DGX agent

arXiv:2604.18738v1 Announce Type: new Abstract: Masked diffusion language models such as LLaDA2.1 rely on Token-to-Token (T2T) editing to correct their own generation errors: whenever a different toke

researcharxiv-cs-cl
22 Apr 2026
Model Releases

VecHeart: Holistic Four-Chamber Cardiac Anatomy Modeling via Hybrid VecSets

DGX agent

arXiv:2604.19403v1 Announce Type: new Abstract: Accurate cardiac anatomy modeling requires the model to be able to handle intricate interrelations among structures. In this paper, we propose VecHeart,

model-releasesarxiv-cs-cv
22 Apr 2026
Safety

Visual Adversarial Attack on Vision-Language Models for Autonomous Driving

DGX agent

arXiv:2411.18275v2 Announce Type: replace Abstract: Vision-language models (VLMs) have significantly advanced autonomous driving (AD) by enhancing reasoning capabilities. However, these models remain

safetyarxiv-cs-cv
22 Apr 2026
Model Releases

Bi-LoRA: Efficient Sharpness-Aware Minimization for Fine-Tuning Large-Scale Models

DGX agent

arXiv:2508.19564v2 Announce Type: replace Abstract: Fine-tuning large-scale pre-trained models with limited data presents significant challenges for generalization. While Sharpness-Aware Minimization

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Counterfactual Modeling with Fine-Tuned LLMs for Health Intervention Design and Sensor Data Augmentation

DGX agent

arXiv:2601.14590v2 Announce Type: replace Abstract: Counterfactual explanations (CFEs) provide human-centric interpretability by identifying the minimal, actionable changes required to alter a machine

model-releasesarxiv-cs-lg
21 Apr 2026
← Previous
1…3940414243…1248
Next →