AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
59,841 results
Agents

Xiaomi EV World Model: A Joint World Model Integrating Reconstruction and Generation for Autonomous Driving

DGX agent

arXiv:2605.18137v1 Announce Type: new Abstract: This report presents a unified technical system addressing the two core capabilities of world models for autonomous driving: world representation and wo

agentsarxiv-cs-cv
19 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

GiLT: Augmenting Transformer Language Models with Dependency Graphs

DGX agent

arXiv:2605.15562v1 Announce Type: new Abstract: Augmenting Transformers with linguistic structures effectively enhances the syntactic generalization performance of language models. Previous work in th

model-releasesarxiv-cs-cl
18 May 2026
Model Releases

PhyDetEx: Detecting and Explaining the Physical Plausibility of T2V Models

DGX agent

arXiv:2512.01843v2 Announce Type: replace Abstract: Driven by the growing capacity and training scale, Text-to-Video (T2V) generation models have recently achieved substantial progress in video qualit

model-releasesarxiv-cs-cv
18 May 2026
Research

Towards Foundation Models for Relational Databases with Language Models and Graph Neural Networks

DGX agent

arXiv:2605.16085v1 Announce Type: cross Abstract: Relational databases store much of the world's structured information, and they are essential for driving complex predictive applications. However, de

researcharxiv-cs-ai
18 May 2026
Model Releases

Agentic Systems as Boosting Weak Reasoning Models

DGX agent

arXiv:2605.14163v1 Announce Type: new Abstract: Can a committee of weak reasoning-model calls reach the performance of much stronger models? We study verifier-backed committee search as inference-time

model-releasesarxiv-cs-ai
15 May 2026
Safety

Do Language Models Align with Brains? Prediction Scores Are Not Enough

DGX agent

arXiv:2605.14025v1 Announce Type: cross Abstract: Brain-language model comparisons often interpret neural prediction scores as evidence that model representations capture brain-relevant language compu

safetyarxiv-cs-ai
15 May 2026
Research

PALMS: A Computational Implementation for Pavlovian Associative Learning Models' Simulation

DGX agent

arXiv:2602.07519v3 Announce Type: replace Abstract: In contrast to static formalisms, computational definitions describe the operational mechanisms of a model. Simulations are an essential part of the

researcharxiv-cs-lg
15 May 2026
Model Releases

Probing into Camera Control of Video Models

DGX agent

arXiv:2605.14815v1 Announce Type: new Abstract: Video is a rich and scalable source of 3D/4D visual observations, and camera control is a key capability for video generation models to produce geometri

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

We just added significantly more NVIDIA Blackwell GPUs to better serve GLM-5.1 model on Ollama's cloud. We have been adding more GPUs daily …

DGX agent

We just added significantly more NVIDIA Blackwell GPUs to better serve GLM-5.1 model on Ollama's cloud. We have been adding more GPUs daily for all the other models. Claude Code: ollama launch claude

model-releasesollama--x
15 May 2026
Model Releases

Energy Scaling Laws for Diffusion Models: Quantifying Compute in Image Generation

DGX agent

arXiv:2511.17031v2 Announce Type: replace-cross Abstract: The rapidly growing computational demands of diffusion models for image generation have raised significant concerns about energy consumption a

model-releasesarxiv-cs-cv
14 May 2026
Model Releases

Negation Neglect: When models fail to learn negations in training

DGX agent

arXiv:2605.13829v1 Announce Type: cross Abstract: We introduce Negation Neglect, where finetuning LLMs on documents that flag a claim as false makes them believe the claim is true. For example, models

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Multi-Narrow Transformation as a Single-Model Ensemble: Boundary Conditions, Mechanisms, and Failure Modes

DGX agent

arXiv:2605.11530v1 Announce Type: new Abstract: Single-model ensembles (SMEs) have attracted attention as a way to approximate some of the benefits of deep ensembles within a single network. However,

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Test-Time Compute for Dense Retrieval: Agentic Program Generation with Frozen Embedding Models

DGX agent

arXiv:2605.11374v1 Announce Type: cross Abstract: Test-time compute is widely believed to benefit only large reasoning models. We show it also helps small embedding models. Most modern embedding check

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

ACWM-Phys: Investigating Generalized Physical Interaction in Action-Conditioned Video World Models

DGX agent

arXiv:2605.08567v1 Announce Type: new Abstract: Action-conditioned world models (ACWMs) have shown strong promise for video prediction and decision-making. However, existing benchmarks are largely res

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

CapVector: Learning Transferable Capability Vectors in Parametric Space for Vision-Language-Action Models

DGX agent

arXiv:2605.10903v1 Announce Type: new Abstract: This paper proposes a novel approach to address the challenge that pretrained VLA models often fail to effectively improve performance and reduce adapta

model-releasesarxiv-cs-cv
12 May 2026
Research

Compact SO(3) Equivariant Atomistic Foundation Models via Structural Pruning

DGX agent

arXiv:2605.08885v1 Announce Type: new Abstract: SO(3) equivariant graph neural networks have become the dominant paradigm for atomistic foundation models, achieving high accuracy and data efficiency b

researcharxiv-cs-lg
12 May 2026
Tutorials

Developing a foundation model for high-resolution remote sensing data of the Netherlands

DGX agent

arXiv:2605.10184v1 Announce Type: cross Abstract: We develop a foundation model using 1.2m high resolution satellite images of the Netherlands. By combining a Convolutional Neural Network and a Vision

tutorialsarxiv-cs-ai
12 May 2026
Model Releases

From Pixels to Concepts: Do Segmentation Models Understand What They Segment?

DGX agent

arXiv:2605.09591v1 Announce Type: new Abstract: Segmentation is a fundamental vision task underlying numerous downstream applications. Recent promptable segmentation models, such as Segment Anything M

model-releasesarxiv-cs-cv
12 May 2026
Applications

HarmoWAM: Harmonizing Generalizable and Precise Manipulation via Adaptive World Action Models

DGX agent

arXiv:2605.10942v1 Announce Type: new Abstract: World Action Models (WAMs) have emerged as a promising paradigm for robot control by modeling physical dynamics. Current WAMs generally follow two parad

applicationsarxiv-cs-ro
12 May 2026
Model Releases

LegalCiteBench: Evaluating Citation Reliability in Legal Language Models

DGX agent

arXiv:2605.10186v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly integrated into legal drafting and research workflows, where incorrect citations or fabricated precedent

model-releasesarxiv-cs-ai
12 May 2026
Hardware

Model-Aware Tokenizer Transfer

DGX agent

arXiv:2510.21954v2 Announce Type: replace Abstract: Large Language Models (LLMs) are trained to support an increasing number of languages, yet their predefined tokenizers remain a bottleneck for adapt

hardwarearxiv-cs-cl
12 May 2026
Model Releases

Scaling Vision Models Does Not Consistently Improve Localisation-Based Explanation Quality

DGX agent

arXiv:2605.10142v1 Announce Type: cross Abstract: Artificial intelligence models are increasingly scaled to improve predictive accuracy, yet it remains unclear whether scale improves the quality of po

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Domain-level metacognitive monitoring in frontier LLMs: A 33-model atlas

DGX agent

arXiv:2605.06673v1 Announce Type: cross Abstract: Aggregate metacognitive quality scores mask within-model variation across MMLU benchmark domains. We administered 1,500 MMLU items (250 per domain, un

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Open Models x Headless Agent Execution 🔥

DGX agent

Open Models x Headless Agent Execution 🔥 your daily reminder that open models are plenty capable for a lot of coding work. easiest place to feel that out is deepagents! swap the model and go. i've bee

model-releasesharrison-chase--x
7 May 2026
Research

Position: the Stochastic Parrot in the Coal Mine. Model Collapse is a Threat to Low-Resource Communities

DGX agent

arXiv:2605.04127v1 Announce Type: cross Abstract: Model collapse, the degradation in performance that arises when generative models are trained on the outputs of prior models, is an increasing concern

researcharxiv-cs-cl
7 May 2026
Model Releases

Replacing Parameters with Preferences: Federated Alignment of Heterogeneous Vision-Language Models

DGX agent

arXiv:2605.03426v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have broad potential in privacy-sensitive domains such as healthcare and finance, yet strict data-sharing constraints rend

model-releasesarxiv-cs-ai
7 May 2026
Model Releases

Sparse Autoencoder Decomposition of Clinical Sequence Model Representations: Feature Complexity, Task Specialisation, and Mortality Prediction

DGX agent

arXiv:2605.04072v1 Announce Type: cross Abstract: Sparse autoencoders (SAEs) have been applied to large language models and protein language models, but not systematically to electronic health record

model-releasesarxiv-cs-cl
7 May 2026
Model Releases

How Language Models Process Negation

DGX agent

arXiv:2605.03052v1 Announce Type: new Abstract: We study how Large Language Models (LLMs) process negation mechanistically. First, we establish that even though open-weight models often provide wrong

model-releasesarxiv-cs-cl
6 May 2026
Applications

Re-Key-Free, Risky-Free: Adaptable Model Usage Control

DGX agent

arXiv:2511.18772v2 Announce Type: replace-cross Abstract: Deep neural networks (DNNs) have become valuable intellectual property of model owners, due to the substantial resources required for their de

applicationsarxiv-cs-ai
6 May 2026
Model Releases

Causal2Vec: Improving Decoder-only LLMs as Embedding Models through a Contextual Token

DGX agent

arXiv:2507.23386v3 Announce Type: replace Abstract: Decoder-only large language models (LLMs) have been increasingly adopted to build embedding models for diverse tasks. To overcome the inherent limit

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

CP-SynC: Multi-Agent Zero-Shot Constraint Modeling in MiniZinc with Synthesized Checkers

DGX agent

arXiv:2605.01675v1 Announce Type: cross Abstract: Constraint Programming (CP) is a powerful paradigm for solving combinatorial problems, yet translating natural language problem descriptions into exec

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Developing a Strong Pre-Trained Base Model for Plant Leaf Disease Classification

DGX agent

arXiv:2605.01283v1 Announce Type: new Abstract: Plants, crops and their yields are essential to our very existence, but diseases and pests cause large losses every year. As such it is vital to ensure

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

Model Merging: Foundations and Algorithms

DGX agent

arXiv:2605.01580v1 Announce Type: new Abstract: Modern deep learning usually treats models as separate artifacts: trained independently, specialized for particular purposes, and replaced when improved

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

OpenAI claims ChatGPT’s new default model hallucinates way less

DGX agent

OpenAI's newest default model for ChatGPT might not make stuff up as much. Hallucinations have been an ongoing problem for AI models, but OpenAI says its new GPT-5.5 Instant model has 'significant imp

model-releasesthe-verge-ai
5 May 2026
Model Releases

RMGAP: Benchmarking the Generalization of Reward Models across Diverse Preferences

DGX agent

arXiv:2605.01831v1 Announce Type: new Abstract: Reinforcement Learning from Human Feedback has become the standard paradigm for language model alignment, where reward models directly determine alignme

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

There’s a bunch of conflicting stances I don’t fully understand in the debate of Proprietary RL’d vs Open Harness, Model intelligence, and A…

DGX agent

There’s a bunch of conflicting stances I don’t fully understand in the debate of Proprietary RL’d vs Open Harness, Model intelligence, and Agent Labs building harnesses for bespoke tasks Not all of th

model-releasesharrison-chase--x
5 May 2026
Research

Rethinking LLM Ensembling from the Perspective of Mixture Models

DGX agent

arXiv:2605.00419v1 Announce Type: cross Abstract: Model ensembling is a well-established technique for improving the performance of machine learning models. Conventionally, this involves averaging the

researcharxiv-cs-cl
4 May 2026
Agents

Use open models in Fleet!

DGX agent

Use open models in Fleet! Not every step in an agent workflow needs the same model. Fleet now lets you customize which model each sub-agent uses, so you can route simple tasks to fast/cheap models and

agentsharrison-chase--x
4 May 2026
Model Releases

HealthBench Professional: Evaluating Large Language Models on Real Clinician Chats

DGX agent

arXiv:2604.27470v1 Announce Type: new Abstract: Millions of clinicians use ChatGPT to support clinical care, but evaluations of the most common use cases in model-clinician conversations are limited.

model-releasesarxiv-cs-cl
1 May 2026
Agents

Modeling Clinical Concern Trajectories in Language Model Agents

DGX agent

arXiv:2604.27872v1 Announce Type: new Abstract: Large language model (LLM) agents deployed in clinical settings often exhibit abrupt, threshold-driven behavior, offering little visibility into accumul

agentsarxiv-cs-ai
1 May 2026
Model Releases

Models Recall What They Violate: Constraint Adherence in Multi-Turn LLM Ideation

DGX agent

arXiv:2604.28031v1 Announce Type: new Abstract: When researchers iteratively refine ideas with large language models, do the models preserve fidelity to the original objective? We introduce DriftBench

model-releasesarxiv-cs-cl
1 May 2026
Model Releases

What Suppresses Nash Equilibrium Play in Large Language Models? Mechanistic Evidence and Causal Control

DGX agent

arXiv:2604.27167v1 Announce Type: cross Abstract: LLM agents are known to deviate from Nash equilibria in strategic interactions, but nobody has looked inside the model to understand why, or asked whe

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

A Comparative Study in Surgical AI: Datasets, Foundation Models, and Barriers to Med-AGI

DGX agent

arXiv:2603.27341v2 Announce Type: replace-cross Abstract: Recent Artificial Intelligence (AI) models have matched or exceeded human experts in several benchmarks of biomedical task performance, but su

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

Benchmarking Layout-Guided Diffusion Models through Unified Semantic-Spatial Evaluation in Closed and Open Settings

DGX agent

arXiv:2604.25358v1 Announce Type: new Abstract: Evaluating layout-guided text-to-image generative models requires assessing both semantic alignment with textual prompts and spatial fidelity to prescri

model-releasesarxiv-cs-cv
29 Apr 2026
Tutorials

Diffusion Model for Manifold Data: Score Decomposition, Curvature, and Statistical Complexity

DGX agent

arXiv:2603.20645v2 Announce Type: replace Abstract: Diffusion models have become a leading framework in generative modeling, yet their theoretical understanding -- especially for high-dimensional data

tutorialsarxiv-cs-lg
29 Apr 2026
Safety

How Fast Should a Model Commit to Supervision? Training Reasoning Models on the Tsallis Loss Continuum

DGX agent

arXiv:2604.25907v1 Announce Type: new Abstract: Adapting reasoning models to new tasks during post-training with only output-level supervision stalls under reinforcement learning from verifiable rewar

safetyarxiv-cs-lg
29 Apr 2026
Research

Marco-MoE: Open Multilingual Mixture-of-Expert Language Models with Efficient Upcycling

DGX agent

arXiv:2604.25578v1 Announce Type: new Abstract: We present Marco-MoE, a suite of fully open multilingual sparse Mixture-of-Experts (MoE) models. Marco-MoE features a highly sparse design in which only

researcharxiv-cs-cl
29 Apr 2026
Model Releases

Personalization Toolkit: Training Free Personalization of Large Vision Language Models

DGX agent

arXiv:2502.02452v4 Announce Type: replace Abstract: Personalization of Large Vision-Language Models (LVLMs) involves customizing models to recognize specific users or object instances and to generate

model-releasesarxiv-cs-cv
29 Apr 2026
← Previous
1…2930313233…1247
Next →