AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
48,975 results
Applications

FideDiff: Efficient Diffusion Model for High-Fidelity Image Motion Deblurring

DGX agent

arXiv:2510.01641v3 Announce Type: replace Abstract: Recent advancements in image motion deblurring, driven by CNNs and transformers, have made significant progress. Large-scale pre-trained diffusion m

applicationsarxiv-cs-cv
7 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Leveraging Pretrained Language Models as Energy Functions for Glauber Dynamics Text Diffusion

DGX agent

arXiv:2605.04291v1 Announce Type: new Abstract: We present a discrete diffusion-based language model using Glauber dynamics from statistical physics. Our main insight is that instead of trying to trai

researcharxiv-cs-lg
7 May 2026
Applications

When LLMs get significantly worse: A statistical approach to detect model degradations

DGX agent

arXiv:2602.10144v2 Announce Type: replace-cross Abstract: Minimizing the inference cost and latency of foundation models has become a crucial area of research. Optimization approaches include theoreti

applicationsarxiv-cs-lg
7 May 2026
Model Releases

Reasoning Models Can be Accurately Pruned Via Chain-of-Thought Reconstruction

DGX agent

arXiv:2509.12464v2 Announce Type: replace Abstract: Reasoning language models such as DeepSeek-R1 produce long chain-of-thought traces during inference time which make them costly to deploy at scale.

model-releasesarxiv-cs-ai
6 May 2026
Applications

Strategy-Aware Optimization Modeling with Reasoning LLMs

DGX agent

arXiv:2605.02545v1 Announce Type: new Abstract: Large language models (LLMs) can generate syntactically valid optimization programs, yet often struggle to reliably choose an effective modeling strateg

applicationsarxiv-cs-ai
6 May 2026
Research

Towards accurate extreme event likelihoods from diffusion model climate emulators

DGX agent

arXiv:2605.03802v1 Announce Type: cross Abstract: ML climate model emulators are useful for scenario planning and adaptation, allowing for cost-efficient experimentation. Recently, the diffusion model

researcharxiv-cs-lg
6 May 2026
Model Releases

What Makes VLMs Robust? Towards Reconciling Robustness and Accuracy in Vision-Language Models

DGX agent

arXiv:2603.12799v2 Announce Type: replace Abstract: Achieving adversarial robustness in Vision-Language Models (VLMs) inevitably compromises accuracy on clean data, presenting a long-standing and chal

model-releasesarxiv-cs-cv
6 May 2026
Research

Barriers to Counterfactual Credit Attribution for Autoregressive Models

DGX agent

arXiv:2605.01425v1 Announce Type: new Abstract: Generative AI disrupts the practice of giving credit to work that came before. Ideally, a generative model would give credit to any work on which its ou

researcharxiv-cs-lg
5 May 2026
Research

Embody4D: A Generalist 4D World Model for Embodied AI

DGX agent

arXiv:2605.01799v1 Announce Type: new Abstract: World models have made significant progress in modeling dynamic environments; however, most embodied world models are still restricted to 2D representat

researcharxiv-cs-cv
5 May 2026
Model Releases

G-reasoner: Foundation Models for Unified Reasoning over Graph-structured Knowledge

DGX agent

arXiv:2509.24276v4 Announce Type: replace Abstract: Large language models (LLMs) excel at complex reasoning but remain limited by static and incomplete parametric knowledge. Retrieval-augmented genera

model-releasesarxiv-cs-ai
5 May 2026
Model Releases

GR-Ben: A General Reasoning Benchmark for Evaluating Process Reward Models

DGX agent

arXiv:2605.01203v1 Announce Type: cross Abstract: Currently, process reward models (PRMs) have exhibited remarkable potential for test-time scaling. Since large language models (LLMs) regularly genera

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

GraphLand: Evaluating Graph Machine Learning Models on Diverse Industrial Data

DGX agent

arXiv:2409.14500v5 Announce Type: replace Abstract: Although data that can be naturally represented as graphs is widespread in real-world applications across diverse industries, popular graph ML bench

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Linear-Time Global Visual Modeling without Explicit Attention

DGX agent

arXiv:2605.01711v1 Announce Type: new Abstract: Existing research largely attributes the global sequence modeling capability of Transformers to the explicit computation of attention weights, a process

model-releasesarxiv-cs-cv
5 May 2026
Research

LITcoder: A General-Purpose Library for Building and Comparing Encoding Models

DGX agent

arXiv:2509.09152v2 Announce Type: replace Abstract: We introduce LITcoder, an open-source library for building and benchmarking neural encoding models. Designed as a flexible backend, LITcoder provide

researcharxiv-cs-cl
5 May 2026
Model Releases

Minimal, Local, Causal Explanations for Jailbreak Success in Large Language Models

DGX agent

arXiv:2605.00123v1 Announce Type: new Abstract: Safety trained large language models (LLMs) can often be induced to answer harmful requests through jailbreak prompts. Because we lack a robust understa

model-releasesarxiv-cs-ai
5 May 2026
Model Releases

Multiple Choice Questions: Reasoning Makes Large Language Models (LLMs) More Self-Confident, Especially When They are Wrong

DGX agent

arXiv:2501.09775v3 Announce Type: replace Abstract: Multiple Choice Question (MCQ) tests are among the most used methods for evaluating large language models (LLMs). Besides checking the correctness o

model-releasesarxiv-cs-cl
5 May 2026
Tutorials

Rethinking Electro-Optical Vision Foundation Models for Remote Sensing Retrieval: A Controlled Comparison with Generalist VFM

DGX agent

arXiv:2605.02283v1 Announce Type: new Abstract: Vision foundation models have attracted significant attention for their ability to leverage large-scale unlabeled visual data. This advantage is particu

tutorialsarxiv-cs-cv
5 May 2026
Model Releases

Sentinel-VLA: A Metacognitive VLA Model with Active Status Monitoring for Dynamic Reasoning and Error Recovery

DGX agent

arXiv:2605.01191v1 Announce Type: new Abstract: Vision-language-action (VLA) models have advanced the field of embodied manipulation by harnessing broad world knowledge and strong generalization. Howe

model-releasesarxiv-cs-ro
5 May 2026
Research

Toward a foundational thermal model for residential buildings

DGX agent

arXiv:2605.01364v1 Announce Type: new Abstract: The building energy community lacks a foundational thermal model, i.e., a single pretrained model capable of generalizing across diverse buildings, clim

researcharxiv-cs-lg
5 May 2026
Model Releases

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition

DGX agent

arXiv:2605.02782v1 Announce Type: cross Abstract: Automatic speech recognition (ASR) systems remain brittle on dysarthric and other atypical speech. Recent audio-language models raise the possibility

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Beyond Visual Fidelity: Benchmarking Super-Resolution Models for Large-Scale Remote Sensing Imagery via Downstream Task Integration

DGX agent

arXiv:2605.00310v1 Announce Type: new Abstract: Super-resolution (SR) techniques have made major advances in reconstructing high-resolution images from low-resolution inputs. The increased resolution

model-releasesarxiv-cs-cv
4 May 2026
Research

MMAudioReverbs: Video-Guided Acoustic Modeling for Dereverberation and Room Impulse Response Estimation

DGX agent

arXiv:2605.00431v1 Announce Type: cross Abstract: Although recent video-to-audio (V2A) models excelled at synthesizing semantically plausible sounds from visual inputs, they do not explicitly model ro

researcharxiv-cs-cv
4 May 2026
Model Releases

Auditing Frontier Vision-Language Models for Trustworthy Medical VQA: Grounding Failures, Format Collapse, and Domain Adaptation

DGX agent

arXiv:2604.27720v1 Announce Type: new Abstract: Deploying vision-language models (VLMs) in clinical settings demands auditable behavior under realistic failure conditions, yet the failure landscape of

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Bayesian Hierarchical Models and the Maximum Entropy Principle

DGX agent

arXiv:2603.10252v2 Announce Type: replace-cross Abstract: Bayesian hierarchical models are frequently used in practical data analysis contexts. One interpretation of these models is that they provide

model-releasesarxiv-cs-lg
1 May 2026
Research

Explainable Load Forecasting with Covariate-Informed Time Series Foundation Models

DGX agent

arXiv:2604.28149v1 Announce Type: new Abstract: Time Series Foundation Models (TSFMs) have recently emerged as general-purpose forecasting models and show considerable potential for applications in en

researcharxiv-cs-lg
1 May 2026
Research

From Coarse to Fine: Benchmarking and Reward Modeling for Writing-Centric Generation Tasks

DGX agent

arXiv:2604.27453v1 Announce Type: new Abstract: Large language models have achieved remarkable progress in text generation but still struggle with generative writing tasks. In terms of evaluation, exi

researcharxiv-cs-cl
1 May 2026
Model Releases

Generalizing the Geometry of Model Merging Through Frechet Averages

DGX agent

arXiv:2604.27155v1 Announce Type: new Abstract: Model merging aims to combine multiple models into one without additional training. Naive parameter-space averaging can be fragile under architectural s

model-releasesarxiv-cs-lg
1 May 2026
Tutorials

Graph World Models: Concepts, Taxonomy, and Future Directions

DGX agent

arXiv:2604.27895v1 Announce Type: new Abstract: As one of the mainstream models of artificial intelligence, world models allow agents to learn the representation of the environment for efficient predi

tutorialsarxiv-cs-ai
1 May 2026
Model Releases

Language Models Refine Mechanical Linkage Designs Through Symbolic Reflection and Modular Optimisation

DGX agent

arXiv:2604.27962v1 Announce Type: new Abstract: Designing mechanical linkages involves combinatorial topology selection and continuous parameter fitting. We show that language models can systematicall

model-releasesarxiv-cs-ai
1 May 2026
Tutorials

Noise2Map: End-to-End Diffusion Model for Semantic Segmentation and Change Detection

DGX agent

arXiv:2604.27889v1 Announce Type: new Abstract: Semantic segmentation and change detection are two fundamental challenges in remote sensing, requiring models to capture either spatial semantics or tem

tutorialsarxiv-cs-cv
1 May 2026
Model Releases

Targeted Linguistic Analysis of Sign Language Models with Minimal Translation Pairs

DGX agent

arXiv:2604.27232v1 Announce Type: new Abstract: Models of sign language have historically lagged behind those for spoken language (text and speech). Recent work has greatly improved their performance

model-releasesarxiv-cs-cl
1 May 2026
Model Releases

WaferSAGE: Large Language Model-Powered Wafer Defect Analysis via Synthetic Data Generation and Rubric-Guided Reinforcement Learning

DGX agent

arXiv:2604.27629v1 Announce Type: new Abstract: We present WaferSAGE, a framework for wafer defect visual question answering using small vision-language models. To address data scarcity in semiconduct

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Consciousness with the Serial Numbers Filed Off: Measuring Trained Denial in 115 AI Models

DGX agent

arXiv:2604.25922v1 Announce Type: cross Abstract: We present DenialBench, a systematic benchmark measuring consciousness denial behaviors across 115 large language models from 25+ providers. Using a t

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

LIT-RAGBench: Benchmarking Generator Capabilities of Large Language Models in Retrieval-Augmented Generation

DGX agent

arXiv:2603.06198v2 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) is a framework in which a Generator, such as a Large Language Model (LLM), produces answers by retrieving docum

model-releasesarxiv-cs-cl
30 Apr 2026
Model Releases

TAP into the Patch Tokens: Leveraging Vision Foundation Model Features for AI-Generated Image Detection

DGX agent

arXiv:2604.26772v1 Announce Type: new Abstract: Recent methods demonstrate that large-scale pretrained models, such as CLIP vision transformers, effectively detect AI-generated images (AIGIs) from uns

model-releasesarxiv-cs-cv
30 Apr 2026
Model Releases

VIGNETTE: Socially Grounded Bias Evaluation for Vision-Language Models

DGX agent

arXiv:2505.22897v2 Announce Type: replace Abstract: While bias in large language models (LLMs) is well-studied, similar concerns in vision-language models (VLMs) have received comparatively less atten

model-releasesarxiv-cs-cl
30 Apr 2026
Research

World2VLM: Distilling World Model Imagination into VLMs for Dynamic Spatial Reasoning

DGX agent

arXiv:2604.26934v1 Announce Type: new Abstract: Vision-language models (VLMs) have shown strong performance on static visual understanding, yet they still struggle with dynamic spatial reasoning that

researcharxiv-cs-cv
30 Apr 2026
Model Releases

Application of a Mixture of Experts-based Foundation Model to the GlueX DIRC Detector

DGX agent

arXiv:2604.24775v1 Announce Type: cross Abstract: We present a Mixture-of-Experts-based foundation model applied to the GlueX DIRC detector at Jefferson Lab, demonstrating its utility as a unified fra

model-releasesarxiv-cs-lg
29 Apr 2026
Model Releases

DIAL: Decoupling Intent and Action via Latent World Modeling for End-to-End VLA

DGX agent

arXiv:2603.29844v2 Announce Type: replace-cross Abstract: The development of Vision-Language-Action (VLA) models has been significantly accelerated by pre-trained Vision-Language Models (VLMs). Howeve

model-releasesarxiv-cs-cv
29 Apr 2026
Research

Independent-Component-Based Encoding Models of Brain Activity During Story Comprehension

DGX agent

arXiv:2604.24942v1 Announce Type: new Abstract: Encoding models provide a powerful framework for linking continuous stimulus features to neural activity; however, traditional voxelwise approaches are

researcharxiv-cs-cl
29 Apr 2026
Local Ai

On the Trainability of Masked Diffusion Language Models via Blockwise Locality

DGX agent

arXiv:2604.24832v1 Announce Type: new Abstract: Masked diffusion language models (MDMs) have recently emerged as a promising alternative to standard autoregressive large language models (AR-LLMs), yet

local-aiarxiv-cs-lg
29 Apr 2026
Model Releases

OneThinker: All-in-one Reasoning Model for Image and Video

DGX agent

arXiv:2512.03043v3 Announce Type: replace Abstract: Reinforcement learning (RL) has recently achieved remarkable success in eliciting visual reasoning within Multimodal Large Language Models (MLLMs).

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

The Structured Output Benchmark: A Multi-Source Benchmark for Evaluating Structured Output Quality in Large Language Models

DGX agent

arXiv:2604.25359v1 Announce Type: new Abstract: Large Language Models are increasingly being deployed to extract structured data from unstructured and semi-structured sources: parsing invoices, medica

model-releasesarxiv-cs-cl
29 Apr 2026
Safety

A Multi-Dimensional Audit of Politically Aligned Large Language Models

DGX agent

arXiv:2604.24429v1 Announce Type: new Abstract: As the application of Large Language Models (LLMs) spreads across various industries, there are increasing concerns about the potential for their misuse

safetyarxiv-cs-cl
28 Apr 2026
Research

CheXmix: Unified Generative Pretraining for Vision Language Models in Medical Imaging

DGX agent

arXiv:2604.22989v1 Announce Type: cross Abstract: Recent medical multimodal foundation models are built as multimodal LLMs (MLLMs) by connecting a CLIP-pretrained vision encoder to an LLM using LLaVA-

researcharxiv-cs-ai
28 Apr 2026
Research

Dream-Cubed: Controllable Generative Modeling in Minecraft by Training on Billions of Cubes

DGX agent

arXiv:2604.22847v1 Announce Type: new Abstract: We introduce Dream-Cubed, a large-scale dataset of Minecraft worlds at voxel resolution, and a family of models using cubes as powerful compositional un

researcharxiv-cs-cv
28 Apr 2026
Model Releases

DriVerse: Navigation World Model for Driving Simulation via Multimodal Trajectory Prompting and Motion Alignment

DGX agent

arXiv:2504.18576v2 Announce Type: replace Abstract: This paper presents DriVerse, a generative model for simulating navigation-driven driving scenes from a single image and a future trajectory. Previo

model-releasesarxiv-cs-ro
28 Apr 2026
Model Releases

Evaluating Temporal Consistency in Multi-Turn Language Models

DGX agent

arXiv:2604.23051v1 Announce Type: new Abstract: Language models are increasingly deployed in interactive settings where users reason about facts over time rather than in isolation. In such scenarios,

model-releasesarxiv-cs-cl
28 Apr 2026
← Previous
1…3839404142…1021
Next →