AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
48,975 results
Model Releases

Adapting Foundation Vision-Language Models to Medical Diagnosis via Query-Driven Expert Bridging

DGX agent

arXiv:2505.21698v3 Announce Type: replace Abstract: Vision-language foundation models achieve promising performance in natural image classification, yet their direct application to medical imaging is

model-releasesarxiv-cs-cv
18 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Golden Layers and Where to Find Them: Improved Knowledge Editing for Large Language Models Via Layer Gradient Analysis

DGX agent

arXiv:2602.20207v3 Announce Type: replace-cross Abstract: Knowledge editing in Large Language Models (LLMs) aims to update the model's prediction for a specific query to a desired target while preserv

model-releasesarxiv-cs-ai
18 May 2026
Research

Latent Video Prediction Learns Better World Models

DGX agent

arXiv:2605.15618v1 Announce Type: cross Abstract: Self-supervised video models are increasingly framed as world models, yet their evaluation remains largely confined to a single top-1 accuracy score o

researcharxiv-cs-ai
18 May 2026
Model Releases

VideoGameBench: Can Vision-Language Models complete popular video games?

DGX agent

arXiv:2505.18134v3 Announce Type: replace Abstract: Vision-language models (VLMs) have achieved strong results on coding and math benchmarks that are challenging for humans, yet their ability to perfo

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Mini-JEPA Foundation Model Fleet Enables Agentic Hydrologic Intelligence

DGX agent

arXiv:2605.14120v1 Announce Type: cross Abstract: Geospatial foundation models compress multispectral observations into dense embeddings increasingly used in natural-language environmental reasoning s

model-releasesarxiv-cs-cl
15 May 2026
Agents

Model-Adaptive Tool Necessity Reveals the Knowing-Doing Gap in LLM Tool Use

DGX agent

arXiv:2605.14038v1 Announce Type: new Abstract: Large language models (LLMs) increasingly act as autonomous agents that must decide when to answer directly vs. when to invoke external tools. Prior wor

agentsarxiv-cs-ai
15 May 2026
Model Releases

NeuroAtlas: Benchmarking Foundation Models for Clinical EEG and Brain-Computer Interfaces

DGX agent

arXiv:2605.14698v1 Announce Type: cross Abstract: Foundation models (FMs) promise to extract unified representations that generalize across downstream tasks. They have emerged across fields, including

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

OPT-Engine: Benchmarking the Limits of LLMs in Optimization Modeling via Complexity Scaling

DGX agent

arXiv:2601.19924v2 Announce Type: replace-cross Abstract: We investigate the capabilities and scalability of Large Language Models (LLMs) in optimization modeling, a domain requiring structured reason

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

Small Language Models (SLMs) Can Still Pack a Punch: A survey (updated 2026)

DGX agent

arXiv:2501.05465v2 Announce Type: replace Abstract: As foundation AI models continue to increase in size, an important question arises - is massive scale the only path forward? This survey of about 16

model-releasesarxiv-cs-cl
15 May 2026
Model Releases

Unsteady Metrics and Benchmarking Cultures of AI Model Builders

DGX agent

arXiv:2605.14164v1 Announce Type: new Abstract: The primary way to establish and compare competencies in foundation and generative AI models has shifted from peer-reviewed literature to press releases

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

CoT-Guard: Small Models for Strong Monitoring

DGX agent

arXiv:2605.12746v1 Announce Type: cross Abstract: Monitoring the chain-of-thought (CoT) of reasoning models is a promising approach for detecting covert misbehavior (i.e., hidden objectives) in code g

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

D-VLA: A High-Concurrency Distributed Asynchronous Reinforcement Learning Framework for Vision-Language-Action Models

DGX agent

arXiv:2605.13276v1 Announce Type: new Abstract: The rapid evolution of Embodied AI has enabled Vision-Language-Action (VLA) models to excel in multimodal perception and task execution. However, applyi

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Embodied Multi-Agent Coordination by Aligning World Models Through Dialogue

DGX agent

arXiv:2605.12920v1 Announce Type: cross Abstract: Effective collaboration between embodied agents requires more than acting in a shared environment; it demands communication grounded in each agent's e

model-releasesarxiv-cs-ai
14 May 2026
Local Ai

PROMETHEUS: Automating Deep Causal Research Integrating Text, Data and Models

DGX agent

arXiv:2605.12835v1 Announce Type: new Abstract: Large language models can extract local causal claims from text, but those claims become more useful when organized as persistent, navigable world model

local-aiarxiv-cs-ai
14 May 2026
Model Releases

Query-Conditioned Test-Time Self-Training for Large Language Models

DGX agent

arXiv:2605.13369v1 Announce Type: cross Abstract: Large language models (LLMs) are typically deployed with fixed parameters, and their performance is often improved by allocating more computation at i

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Safe Bayesian Optimization for Uncertain Correlations Matrices in Linear Models of Co-Regionalization

DGX agent

arXiv:2605.13302v1 Announce Type: new Abstract: This paper extends safety guarantees for multi-task Bayesian optimization with uncertain correlation matrices from intrinsic co-reginalization models to

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs

DGX agent

arXiv:2510.18245v3 Announce Type: replace-cross Abstract: Scaling the number of parameters and the size of training data has proven to be an effective strategy for improving large language model (LLM)

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Visual Aesthetic Benchmark: Can Frontier Models Judge Beauty?

DGX agent

arXiv:2605.12684v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) are now routinely deployed for visual understanding, generation, and curation. A substantial fraction of thes

model-releasesarxiv-cs-ai
14 May 2026
Applications

Coevolutionary Continuous Discrete Diffusion: Make Your Diffusion Language Model a Latent Reasoner

DGX agent

arXiv:2510.03206v2 Announce Type: replace-cross Abstract: Diffusion language models, especially masked discrete diffusion models, have achieved great success recently. While there are some theoretical

applicationsarxiv-cs-cl
13 May 2026
Safety

Enabling clinical use of foundation models for computational pathology

DGX agent

arXiv:2602.22347v2 Announce Type: replace Abstract: Foundation models for computational pathology are expected to facilitate the development of high-performing, generalisable deep learning systems. Ho

safetyarxiv-cs-cv
13 May 2026
Model Releases

More Edits, More Stable: Understanding the Lifelong Normalization in Sequential Model Editing

DGX agent

arXiv:2605.11836v1 Announce Type: cross Abstract: Lifelong Model Editing aims to continuously update evolving facts in Large Language Models while preserving unrelated knowledge and general capabiliti

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

Can We Go Beyond Visual Features? Neural Tissue Relation Modeling for Relational Graph Analysis in Non-Melanoma Skin Histology

DGX agent

arXiv:2512.06949v3 Announce Type: replace Abstract: Histopathology image segmentation is essential for delineating tissue structures in skin cancer diagnostics, but modeling spatial context and inter-

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

ColorConceptBench: A Benchmark for Probabilistic Color-Concept Understanding in Text-to-Image Models

DGX agent

arXiv:2601.16836v3 Announce Type: replace-cross Abstract: Text-to-image (T2I) models have advanced considerably in generating high-quality images from textual descriptions. However, their ability to a

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

Decomposing and Steering Functional Metacognition in Large Language Models

DGX agent

arXiv:2605.08942v1 Announce Type: new Abstract: Large language models (LLMs) increasingly exhibit behaviors suggesting awareness of their evaluation context, often adapting their reasoning strategies

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

DeformMaster: An Interactive Physics-Neural World Model for Deformable Objects from Videos

DGX agent

arXiv:2605.09586v1 Announce Type: new Abstract: World models for deformable objects should recover not only geometry and appearance, but also underlying physical dynamics, interaction grounding, and m

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Diffusion Models are Evolutionary Algorithms

DGX agent

arXiv:2410.02543v3 Announce Type: replace-cross Abstract: In a convergence of machine learning and biology, we reveal that diffusion models are evolutionary algorithms. By considering evolution as a d

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

jNO: A JAX Library for Neural Operator and Foundation Model Training

DGX agent

arXiv:2605.10159v1 Announce Type: new Abstract: jNO (jax Neural Operators) is a JAX-native library for neural operators and foundation models with unified support for both data-driven and physics-info

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Large Language Models as Students Who Think Aloud: Overly Coherent, Verbose, and Confident

DGX agent

arXiv:2602.01015v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly embedded in AI-based tutoring systems. Can they faithfully model novice reasoning and metacognitive ju

model-releasesarxiv-cs-cl
12 May 2026
Research

PARD-2: Target-Aligned Parallel Draft Model for Dual-Mode Speculative Decoding

DGX agent

arXiv:2605.08632v1 Announce Type: cross Abstract: Speculative decoding accelerates Large Language Models (LLMs) inference by using a lightweight draft model to propose candidate tokens that are verifi

researcharxiv-cs-ai
12 May 2026
Model Releases

Reinforcement Learning Measurement Model

DGX agent

arXiv:2605.09305v1 Announce Type: cross Abstract: Interactive assessments generate sequential process data that are not well handled by conventional item response models. Existing MDP-based measuremen

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Selective Deficits in LLM Mental Self-Modeling in a Behavior-Based Test of Theory of Mind

DGX agent

arXiv:2603.26089v2 Announce Type: replace-cross Abstract: The ability to represent oneself and others as agents with knowledge, intentions, and belief states that guide their behavior - Theory of Mind

model-releasesarxiv-cs-ai
12 May 2026
Research

Sparse Layers are Critical to Scaling Looped Language Models

DGX agent

arXiv:2605.09165v1 Announce Type: cross Abstract: Looped language models repeat a set of transformer layers through depth, reducing memory costs and providing natural early-exit points at loop boundar

researcharxiv-cs-cl
12 May 2026
Model Releases

Synthetic Pre-Pre-Training Improves Language Model Robustness to Noisy Pre-Training Data

DGX agent

arXiv:2605.10129v1 Announce Type: new Abstract: Large language models (LLMs) rely on web-scale corpora for pre-training. The noise inherent in these datasets tends to obscure meaningful patterns and u

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

Visual-ERM: Reward Modeling for Visual Equivalence

DGX agent

arXiv:2603.13224v2 Announce Type: replace-cross Abstract: Vision-to-code tasks require models to reconstruct structured visual inputs, such as charts, tables, and SVGs, into executable or structured r

model-releasesarxiv-cs-ai
12 May 2026
Research

Where Do Reasoning Models Refuse?

DGX agent

arXiv:2507.03167v3 Announce Type: replace-cross Abstract: Chat models without chain-of-thought (CoT) reasoning must decide whether to refuse a harmful request before generating their first response to

researcharxiv-cs-ai
12 May 2026
Research

CellScientist: Dual-Space Hierarchical Orchestration for Closed-Loop Refinement of Virtual Cell Models

DGX agent

arXiv:2605.07335v1 Announce Type: new Abstract: Virtual Cell Modeling (VCM) requires models that not only predict perturbation responses, but also support targeted revision when predictions fail. Curr

researcharxiv-cs-lg
11 May 2026
Model Releases

Clinically Aware Synthetic Image Generation for Concept Coverage in Chest X-ray Models

DGX agent

arXiv:2603.15525v2 Announce Type: replace Abstract: Deep learning models for chest X-ray diagnosis are constrained by limited coverage of clinically meaningful concept combinations in publicly availab

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

Dino U-Net: Exploiting High-Fidelity Dense Features from Foundation Models for Medical Image Segmentation

DGX agent

arXiv:2508.20909v2 Announce Type: replace Abstract: Foundation models pre-trained on large-scale natural image datasets offer a powerful paradigm for medical image segmentation. However, effectively t

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

Do Joint Audio-Video Generation Models Understand Physics?

DGX agent

arXiv:2605.07061v1 Announce Type: cross Abstract: Joint audio-video generation models are rapidly approaching professional production quality, raising a central question: do they understand audio-visu

model-releasesarxiv-cs-ai
11 May 2026
Safety

Radiologist-Guided Causal Concept Bottleneck Models for Chest X-Ray Interpretation

DGX agent

arXiv:2605.07785v1 Announce Type: new Abstract: Concept Bottleneck Models (CBMs) in medical imaging aim to improve model interpretability by predicting intermediate clinical concepts before final diag

safetyarxiv-cs-cv
11 May 2026
Hardware

SpikingBrain: Spiking Brain-inspired Large Models

DGX agent

arXiv:2509.05276v4 Announce Type: replace-cross Abstract: Mainstream Transformer-based large language models face major efficiency bottlenecks: training computation scales quadratically with sequence

hardwarearxiv-cs-ai
11 May 2026
Local Ai

Test-Time Compositional Generalization in Diffusion Models via Concept Discovery

DGX agent

arXiv:2605.07078v1 Announce Type: new Abstract: Compositional generalization requires models to produce novel configurations from familiar parts. In diffusion models, prior compositional generation me

local-aiarxiv-cs-lg
11 May 2026
Safety

Theoretical Limits of Language Model Alignment

DGX agent

arXiv:2605.07105v1 Announce Type: cross Abstract: Language model (LM) alignment improves model outputs to reflect human preferences while preserving the capabilities of the base model. The most common

safetyarxiv-cs-cl
11 May 2026
Model Releases

Theory of Optimal Learning Rate Schedules and Scaling Laws for a Random Feature Model

DGX agent

arXiv:2602.04774v2 Announce Type: replace-cross Abstract: Setting the learning rate (LR) for a deep learning model is a critical part of successful training. Choosing LRs is often done empirically wit

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

Towards Closing the Autoregressive Gap in Language Modeling via Entropy-Gated Continuous Bitstream Diffusion

DGX agent

arXiv:2605.07013v1 Announce Type: new Abstract: Diffusion language models (DLMs) promise parallel, order-agnostic generation, but on standard benchmarks they have historically lagged behind autoregres

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

Understanding Robustness of Model Editing in Code LLMs

DGX agent

arXiv:2511.03182v2 Announce Type: replace-cross Abstract: Large language models (LLMs) for code are increasingly used in software development, but they remain static after pretraining while APIs and s

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

Your Language Model is Its Own Critic: Reinforcement Learning with Value Estimation from Actor's Internal States

DGX agent

arXiv:2605.07579v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) for Large Reasoning Models hinges on baseline estimation for variance reduction, but existing ap

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Adapting Large Language Models to a Low-Resource Agglutinative Language: A Comparative Study of LoRA and QLoRA for Bashkir

DGX agent

arXiv:2605.04948v1 Announce Type: new Abstract: This paper presents a comparative study of parameter-efficient fine-tuning (PEFT) methods, including LoRA and QLoRA, applied to the task of adapting lar

model-releasesarxiv-cs-cl
7 May 2026
← Previous
1…3738394041…1021
Next →