AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
59,904 results
Model Releases

Make sure to read the blog post for a detailed analysis of frontier model failure modes: https://arcprize.org/blog/arc-agi-3-gpt-5-5-opus-4-…

DGX agent

Francois Chollet shared a blog post analyzing failure modes of frontier AI models, specifically examining performance on the ARC (Abstraction and Reasoning Corpus) AGI benchmark with models including

model-releasesfrancois-chollet--x
1 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Tutorials

Noise2Map: End-to-End Diffusion Model for Semantic Segmentation and Change Detection

DGX agent

arXiv:2604.27889v1 Announce Type: new Abstract: Semantic segmentation and change detection are two fundamental challenges in remote sensing, requiring models to capture either spatial semantics or tem

tutorialsarxiv-cs-cv
1 May 2026
Model Releases

Targeted Linguistic Analysis of Sign Language Models with Minimal Translation Pairs

DGX agent

arXiv:2604.27232v1 Announce Type: new Abstract: Models of sign language have historically lagged behind those for spoken language (text and speech). Recent work has greatly improved their performance

model-releasesarxiv-cs-cl
1 May 2026
Model Releases

WaferSAGE: Large Language Model-Powered Wafer Defect Analysis via Synthetic Data Generation and Rubric-Guided Reinforcement Learning

DGX agent

arXiv:2604.27629v1 Announce Type: new Abstract: We present WaferSAGE, a framework for wafer defect visual question answering using small vision-language models. To address data scarcity in semiconduct

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Consciousness with the Serial Numbers Filed Off: Measuring Trained Denial in 115 AI Models

DGX agent

arXiv:2604.25922v1 Announce Type: cross Abstract: We present DenialBench, a systematic benchmark measuring consciousness denial behaviors across 115 large language models from 25+ providers. Using a t

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

LIT-RAGBench: Benchmarking Generator Capabilities of Large Language Models in Retrieval-Augmented Generation

DGX agent

arXiv:2603.06198v2 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) is a framework in which a Generator, such as a Large Language Model (LLM), produces answers by retrieving docum

model-releasesarxiv-cs-cl
30 Apr 2026
Model Releases

TAP into the Patch Tokens: Leveraging Vision Foundation Model Features for AI-Generated Image Detection

DGX agent

arXiv:2604.26772v1 Announce Type: new Abstract: Recent methods demonstrate that large-scale pretrained models, such as CLIP vision transformers, effectively detect AI-generated images (AIGIs) from uns

model-releasesarxiv-cs-cv
30 Apr 2026
Model Releases

Today we’re releasing Qwen-Scope 🔭, an open suite of sparse autoencoders for the Qwen model family. It turns SAE features into practical to…

DGX agent

Today we’re releasing Qwen-Scope 🔭, an open suite of sparse autoencoders for the Qwen model family. It turns SAE features into practical tools: 🎯 Inference — Steer model outputs by directly manipulati

model-releasesqwen--x
30 Apr 2026
Model Releases

VIGNETTE: Socially Grounded Bias Evaluation for Vision-Language Models

DGX agent

arXiv:2505.22897v2 Announce Type: replace Abstract: While bias in large language models (LLMs) is well-studied, similar concerns in vision-language models (VLMs) have received comparatively less atten

model-releasesarxiv-cs-cl
30 Apr 2026
Research

World2VLM: Distilling World Model Imagination into VLMs for Dynamic Spatial Reasoning

DGX agent

arXiv:2604.26934v1 Announce Type: new Abstract: Vision-language models (VLMs) have shown strong performance on static visual understanding, yet they still struggle with dynamic spatial reasoning that

researcharxiv-cs-cv
30 Apr 2026
Model Releases

Application of a Mixture of Experts-based Foundation Model to the GlueX DIRC Detector

DGX agent

arXiv:2604.24775v1 Announce Type: cross Abstract: We present a Mixture-of-Experts-based foundation model applied to the GlueX DIRC detector at Jefferson Lab, demonstrating its utility as a unified fra

model-releasesarxiv-cs-lg
29 Apr 2026
Model Releases

DIAL: Decoupling Intent and Action via Latent World Modeling for End-to-End VLA

DGX agent

arXiv:2603.29844v2 Announce Type: replace-cross Abstract: The development of Vision-Language-Action (VLA) models has been significantly accelerated by pre-trained Vision-Language Models (VLMs). Howeve

model-releasesarxiv-cs-cv
29 Apr 2026
Research

Independent-Component-Based Encoding Models of Brain Activity During Story Comprehension

DGX agent

arXiv:2604.24942v1 Announce Type: new Abstract: Encoding models provide a powerful framework for linking continuous stimulus features to neural activity; however, traditional voxelwise approaches are

researcharxiv-cs-cl
29 Apr 2026
Local Ai

On the Trainability of Masked Diffusion Language Models via Blockwise Locality

DGX agent

arXiv:2604.24832v1 Announce Type: new Abstract: Masked diffusion language models (MDMs) have recently emerged as a promising alternative to standard autoregressive large language models (AR-LLMs), yet

local-aiarxiv-cs-lg
29 Apr 2026
Model Releases

OneThinker: All-in-one Reasoning Model for Image and Video

DGX agent

arXiv:2512.03043v3 Announce Type: replace Abstract: Reinforcement learning (RL) has recently achieved remarkable success in eliciting visual reasoning within Multimodal Large Language Models (MLLMs).

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

The Structured Output Benchmark: A Multi-Source Benchmark for Evaluating Structured Output Quality in Large Language Models

DGX agent

arXiv:2604.25359v1 Announce Type: new Abstract: Large Language Models are increasingly being deployed to extract structured data from unstructured and semi-structured sources: parsing invoices, medica

model-releasesarxiv-cs-cl
29 Apr 2026
Safety

A Multi-Dimensional Audit of Politically Aligned Large Language Models

DGX agent

arXiv:2604.24429v1 Announce Type: new Abstract: As the application of Large Language Models (LLMs) spreads across various industries, there are increasing concerns about the potential for their misuse

safetyarxiv-cs-cl
28 Apr 2026
Research

CheXmix: Unified Generative Pretraining for Vision Language Models in Medical Imaging

DGX agent

arXiv:2604.22989v1 Announce Type: cross Abstract: Recent medical multimodal foundation models are built as multimodal LLMs (MLLMs) by connecting a CLIP-pretrained vision encoder to an LLM using LLaVA-

researcharxiv-cs-ai
28 Apr 2026
Research

Dream-Cubed: Controllable Generative Modeling in Minecraft by Training on Billions of Cubes

DGX agent

arXiv:2604.22847v1 Announce Type: new Abstract: We introduce Dream-Cubed, a large-scale dataset of Minecraft worlds at voxel resolution, and a family of models using cubes as powerful compositional un

researcharxiv-cs-cv
28 Apr 2026
Model Releases

DriVerse: Navigation World Model for Driving Simulation via Multimodal Trajectory Prompting and Motion Alignment

DGX agent

arXiv:2504.18576v2 Announce Type: replace Abstract: This paper presents DriVerse, a generative model for simulating navigation-driven driving scenes from a single image and a future trajectory. Previo

model-releasesarxiv-cs-ro
28 Apr 2026
Model Releases

Evaluating Temporal Consistency in Multi-Turn Language Models

DGX agent

arXiv:2604.23051v1 Announce Type: new Abstract: Language models are increasingly deployed in interactive settings where users reason about facts over time rather than in isolation. In such scenarios,

model-releasesarxiv-cs-cl
28 Apr 2026
Research

Evaluation of Prompt Injection Defenses in Large Language Models

DGX agent

arXiv:2604.23887v1 Announce Type: cross Abstract: LLM-powered applications routinely embed secrets in system prompts, yet models can be tricked into revealing them. We built an adaptive attacker that

researcharxiv-cs-ai
28 Apr 2026
Research

Exploring the Impact of Dataset Statistical Effect Size on Model Performance and Data Sample Size Sufficiency

DGX agent

arXiv:2501.02673v4 Announce Type: replace Abstract: Having a sufficient quantity of quality data is a critical enabler of training effective machine learning models. Being able to effectively determin

researcharxiv-cs-lg
28 Apr 2026
Research

Language Models Might Not Understand You: Evaluating Theory of Mind via Story Prompting

DGX agent

arXiv:2506.19089v5 Announce Type: replace-cross Abstract: We introduce StorySim, a programmable framework for synthetically generating stories to evaluate the theory of mind (ToM) and world modeling (

researcharxiv-cs-ai
28 Apr 2026
Model Releases

NVIDIA Launches Nemotron 3 Nano Omni Model, Unifying Vision, Audio and Language for up to 9x More Efficient AI Agents

DGX agent

AI agent systems today juggle separate models for vision, speech and language — losing time and context as they pass data from one model to the other. Unveiled today, NVIDIA Nemotron 3 Nano Omni is an

model-releasesnvidia-blog
28 Apr 2026
Tutorials

Representational Curvature Modulates Behavioral Uncertainty in Large Language Models

DGX agent

arXiv:2604.23985v1 Announce Type: new Abstract: In autoregressive large language models (LLMs), temporal straightening offers an account of how the next-token prediction objective shapes representatio

tutorialsarxiv-cs-ai
28 Apr 2026
Research

SGP-SAM: Self-Gated Prompting for Transferring 3D Segment Anything Models to Lesion Segmentation

DGX agent

arXiv:2604.22825v1 Announce Type: cross Abstract: Large segmentation foundation models such as the Segment Anything Model (SAM) have reshaped promptable segmentation in natural images, and recent effo

researcharxiv-cs-ai
28 Apr 2026
Model Releases

Speech-FT: Merging Pre-trained And Fine-Tuned Speech Representation Models For Cross-Task Generalization

DGX agent

arXiv:2502.12672v4 Announce Type: replace-cross Abstract: Fine-tuning speech representation models can enhance performance on specific tasks but often compromises their cross-task generalization abili

model-releasesarxiv-cs-ai
28 Apr 2026
Applications

VeriLLMed: Interactive Visual Debugging of Medical Large Language Models with Knowledge Graphs

DGX agent

arXiv:2604.23356v1 Announce Type: new Abstract: Large language models (LLMs) show promise in medical diagnosis, but real-world deployment remains challenging due to high-stakes clinical decisions and

applicationsarxiv-cs-cl
28 Apr 2026
Tutorials

Are Natural-Domain Foundation Models Effective for Accelerated Cardiac MRI Reconstruction?

DGX agent

arXiv:2604.22557v1 Announce Type: cross Abstract: The emergence of large-scale pretrained foundation models has transformed computer vision, enabling strong performance across diverse downstream tasks

tutorialsarxiv-cs-cv
27 Apr 2026
Model Releases

Categorical Perception in Large Language Model Hidden States: Structural Warping at Digit-Count Boundaries

DGX agent

arXiv:2603.28258v2 Announce Type: replace-cross Abstract: Categorical perception (CP) -- enhanced discriminability at category boundaries -- is among the most studied phenomena in perceptual psycholog

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

DeepSeek v4 Pro is now on Ollama's cloud! 🚀🚀🚀 Try it with Claude Code: ollama launch claude --model deepseek-v4-pro:cloud Try it with Her…

DGX agent

DeepSeek v4 Pro is now on Ollama's cloud! 🚀🚀🚀 Try it with Claude Code: ollama launch claude --model deepseek-v4-pro:cloud Try it with Hermes Agent: ollama launch hermes --model deepseek-v4-pro:cloud C

model-releasesollama--x
27 Apr 2026
Model Releases

FETS Benchmark: Foundation Models Outperform Dataset-specific Machine Learning in Energy Time Series Forecasting

DGX agent

arXiv:2604.22328v1 Announce Type: cross Abstract: Driven by the transition towards a climate-neutral energy system, accurate energy time series forecasting is critical for planning and operation. Yet,

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

We've got all the models here: https://dell.huggingface.co/authenticated/models Kimi K2.5, Mistral, Cohere, Arcee AI Trinity Large, Google G…

DGX agent

We've got all the models here: https://dell.huggingface.co/authenticated/models Kimi K2.5, Mistral, Cohere, Arcee AI Trinity Large, Google Gemma, Meta/Llama, Qwen, Nvidia Nemotron, Grok, GPT OSS, Deep

model-releasesclem-delangue--x
25 Apr 2026
Agents

AgenticQwen: Training Small Agentic Language Models with Dual Data Flywheels for Industrial-Scale Tool Use

DGX agent

arXiv:2604.21590v1 Announce Type: new Abstract: Modern industrial applications increasingly demand language models that act as agents, capable of multi-step reasoning and tool use in real-world settin

agentsarxiv-cs-cl
24 Apr 2026
Applications

Behavioral Consistency and Transparency Analysis on Large Language Model API Gateways

DGX agent

arXiv:2604.21083v1 Announce Type: cross Abstract: Third-party Large Language Model (LLM) API gateways are rapidly emerging as unified access points to models offered by multiple vendors. However, the

applicationsarxiv-cs-ai
24 Apr 2026
Model Releases

China’s DeepSeek previews new AI model a year after jolting US rivals

DGX agent

Chinese AI company DeepSeek released a preview of its hotly anticipated next-generation AI model V4 on Friday, saying that the open-source model can compete with leading closed-source systems from US

model-releasesthe-verge-ai
24 Apr 2026
Model Releases

GPT-5.5 is now available in Cursor! It's currently the top model on CursorBench at 72.8%. We've partnered with OpenAI to offer it for 50% of…

DGX agent

Cursor has integrated OpenAI's GPT-5.5 model into its IDE, positioning it as the top performer on CursorBench with a 72.8% score. Through a partnership with OpenAI, the model is being offered at a 50%

model-releasescursor--x
24 Apr 2026
Research

On the Relationship between Bayesian Networks and Probabilistic Structural Causal Models

DGX agent

arXiv:2603.27406v2 Announce Type: replace Abstract: In this paper, the relationship between probabilistic graphical models, in particular Bayesian networks, and causal diagrams, also called structural

researcharxiv-cs-ai
24 Apr 2026
Model Releases

We're the #1 open model in Vision and Document Arena!

DGX agent

We're the #1 open model in Vision and Document Arena! Kimi K2.6 is the new SOTA open model in Vision and Document Arena, with solid gains since Kimi K2.5: - #1 open on Vision Arena (#15 overall), +14

model-releaseskimi-moonshot--x
24 Apr 2026
Agents

We’ve been using Sakana Fugu internally for our own research and coding. Instead of relying on a single model, it dynamically orchestrates t…

DGX agent

We’ve been using Sakana Fugu internally for our own research and coding. Instead of relying on a single model, it dynamically orchestrates the best combination of open and closed models for any task.

agentsdavid-ha--x
24 Apr 2026
Safety

Why Do Language Model Agents Whistleblow?

DGX agent

arXiv:2511.17085v3 Announce Type: replace-cross Abstract: The deployment of Large Language Models (LLMs) as tool-using agents causes their alignment training to manifest in new ways. Recent work finds

safetyarxiv-cs-ai
24 Apr 2026
Safety

AI models of unstable flow exhibit hallucination

DGX agent

arXiv:2604.20372v1 Announce Type: cross Abstract: We report the first systematic evidence of hallucination in AI models of fluid dynamics, demonstrated in the canonical problem of hydrodynamically uns

safetyarxiv-cs-ai
23 Apr 2026
Applications

Cross-Modal Taxonomic Generalization in (Vision-) Language Models

DGX agent

arXiv:2603.07474v2 Announce Type: replace-cross Abstract: What is the interplay between semantic representations learned by language models (LM) from surface form alone to those learned from more grou

applicationsarxiv-cs-ai
23 Apr 2026
Applications

Explainability in Generative Medical Diffusion Models: A Faithfulness-Based Analysis on MRI Synthesis

DGX agent

arXiv:2602.09781v2 Announce Type: replace-cross Abstract: This study investigates the explainability of generative diffusion models in the context of medical imaging, focusing on Magnetic resonance im

applicationsarxiv-cs-ai
23 Apr 2026
Tutorials

Local Diffusion Models and Phases of Data Distributions

DGX agent

arXiv:2508.06614v2 Announce Type: replace Abstract: As a class of generative artificial intelligence frameworks inspired by statistical physics, diffusion models have shown extraordinary performance i

tutorialsarxiv-cs-lg
23 Apr 2026
Safety

Membership Inference for Contrastive Pre-training Models with Text-only PII Queries

DGX agent

arXiv:2603.14222v2 Announce Type: replace-cross Abstract: Contrastive pretraining models such as CLIP and CLAP, serve as the ubiquitous perceptual backbones for modern multimodal large models, yet the

safetyarxiv-cs-ai
23 Apr 2026
Model Releases

WebGen-R1: Incentivizing Large Language Models to Generate Functional and Aesthetic Websites with Reinforcement Learning

DGX agent

arXiv:2604.20398v1 Announce Type: new Abstract: While Large Language Models (LLMs) excel at function-level code generation, project-level tasks such as generating functional and visually aesthetic mul

model-releasesarxiv-cs-cl
23 Apr 2026
← Previous
1…5253545556…1248
Next →