AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
59,904 results
20 May 2026

Entry-level guide to the use of large language models for medical research

Model ReleasesDGX agent

arXiv:2410.18856v4 Announce Type: replace Abstract: Frontier large language models (LLMs), such as GPT-5, Claude 4.5, Gemini 3, Llama 4, and DeepSeek-R1, represent a transformative class of AI tools c

i don't think i need cloud models anymore

Model ReleasesDGX agent

i don't think i need cloud models anymore MTP speedup Qwen by 2.5x in Atomic Chat Dense vs MoE models on 2x RTX 5090 Qwen3.6 27B: 51 → 117 tps +137% Qwen3.6 35B-A3B: 218 → 267 tps +25% MTP drafts seve

Neural Network Models for Contextual Regression

ResearchDGX agent

arXiv:2603.24400v2 Announce Type: replace-cross Abstract: We propose a neural network model for contextual regression in which the regression model depends on contextual features that determine the ac

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

PROWL: Prioritized Regret-Driven Optimization for World Model Learning

SafetyDGX agent

arXiv:2605.18803v1 Announce Type: cross Abstract: Modern action-conditioned video world models achieve strong short-horizon visual realism, yet remain unreliable on rare, interaction-critical transiti

Token by Token, Compromised: Backdoor Vulnerabilities in Unified Autoregressive Models

ResearchDGX agent

arXiv:2605.19227v1 Announce Type: cross Abstract: Unified autoregressive models (UAMs) are transformer models that generate text as well as image tokens within a single autoregressive pass. Shared par

ZeroUnlearn: Few-Shot Knowledge Unlearning in Large Language Models

Model ReleasesDGX agent

arXiv:2605.18879v1 Announce Type: cross Abstract: Large language models inevitably retain sensitive information, defined as inputs that may induce harmful generations, due to training on massive web c

19 May 2026

Also had some early access to Gemini 3.5 Flash. Very fast for a flash model and very capable, though not as powerful as a full frontier mode…

Model ReleasesDGX agent

Also had some early access to Gemini 3.5 Flash. Very fast for a flash model and very capable, though not as powerful as a full frontier model. I added it to the gallery or procedurally generated one-s

Better Together: Evaluating the Complementarity of Earth Embedding Models

ResearchDGX agent

arXiv:2605.18667v1 Announce Type: new Abstract: Earth embedding models transform Earth observation data into embeddings uniquely tied to locations on the Earth's surface. These models are typically ev

By now, you've probably heard about Gemini Omni, our new model designed to create anything from any input, starting with video. But... what'…

Model ReleasesDGX agent

Google AI announced Gemini Omni, a new multimodal model capable of generating diverse content types from various input formats, with initial focus on video generation capabilities. The model represent

CarbonScaling: Extending Neural Scaling Laws for Carbon Footprint in Large Language Models

Model ReleasesDGX agent

arXiv:2508.06524v2 Announce Type: replace-cross Abstract: Large language models (LLMs) increasingly follow neural scaling laws that tie performance gains to rapidly expanding computational budgets, ra

Cerebras is now running Kimi K2.6 – a trillion parameter model – in enterprise trials. At ~1,000 tokens/s, this is the fastest frontier mode…

Model ReleasesDGX agent

Cerebras is now running Kimi K2.6 – a trillion parameter model – in enterprise trials. At ~1,000 tokens/s, this is the fastest frontier model performance ever measured by Artificial Analysis @Artifici

DecoupleSearch: Decouple Planning and Search via Hierarchical Reward Modeling

Model ReleasesDGX agent

arXiv:2510.21712v2 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) systems have emerged as a pivotal methodology for enhancing Large Language Models (LLMs) through the dyna

Disentangling Ambiguity from Instability in Large Language Models: A Clinical Text-to-SQL Case Study

Model ReleasesDGX agent

arXiv:2602.12015v2 Announce Type: replace Abstract: Deploying large language models for clinical Text-to-SQL requires distinguishing two qualitatively different causes of output diversity: (i) input a

Locally Coherent Parallel Decoding in Diffusion Language Models

Local AiDGX agent

arXiv:2603.20216v2 Announce Type: replace-cross Abstract: Diffusion language models (DLMs) have emerged as a promising alternative to autoregressive (AR) models, offering sub-linear generation latency

Med-V1: Small Language Models for Zero-shot and Scalable Biomedical Evidence Attribution

Model ReleasesDGX agent

arXiv:2603.05308v2 Announce Type: replace-cross Abstract: Assessing whether an article supports an assertion is essential for hallucination detection and claim verification. While large language model

Membership Inference Attacks on Discrete Diffusion Language Models

Model ReleasesDGX agent

arXiv:2605.16445v1 Announce Type: cross Abstract: Masked Diffusion Language Models MDLMs replace autoregressive generation with iterative demasking and their privacy properties are largely unstudied.

Preference Instability in Reward Models: Detection and Mitigation via Sparse Autoencoders

ResearchDGX agent

arXiv:2605.16339v1 Announce Type: new Abstract: Preference learning in large language models relies on reward models as proxies for human judgment. However, these models frequently exhibit preference

QSTRBench: a New Benchmark to Evaluate the Ability of Language Models to Reason with Qualitative Spatial and Temporal Calculi

Model ReleasesDGX agent

arXiv:2605.18380v1 Announce Type: new Abstract: We introduce an extensive qualitative spatial and temporal reasoning (QSTR) benchmark for evaluating large language models (LLMs). We pose questions con

TabH2O: A Unified Foundation Model for Tabular Prediction

Model ReleasesDGX agent

arXiv:2605.18383v1 Announce Type: new Abstract: We present TabH2O, a foundation model for tabular data that performs classification and regression in a single forward pass via in-context learning. Tab

The Illusion of Specialization: Unveiling the Domain-Invariant 'Standing Committee' in Mixture-of-Experts Models

Model ReleasesDGX agent

arXiv:2601.03425v2 Announce Type: replace-cross Abstract: Mixture of Experts models are widely assumed to achieve domain specialization through sparse routing. In this work, we question this assumptio

UniAlign: A Model-Agnostic Framework for Robust Network Traffic Classification under Distribution Shifts

SafetyDGX agent

arXiv:2605.17575v1 Announce Type: cross Abstract: Network traffic classification (NTC) models often suffer severe performance degradation when deployed in real-world environments due to distribution s

What Does the AI Doctor Value? Auditing Pluralism in the Clinical Ethics of Language Models

Model ReleasesDGX agent

arXiv:2605.18738v1 Announce Type: new Abstract: Medicine is inherently pluralistic. Principles such as autonomy, beneficence, nonmaleficence, and justice routinely conflict, and such ethical dilemmas

Who Generated This 3D Asset? Learning Source Attribution for Generative 3D Models

Model ReleasesDGX agent

arXiv:2605.18132v1 Announce Type: cross Abstract: Generative 3D models are deployed in gaming, robotics, and immersive creation, making source attribution critical: given a 3D asset, can we identify w

18 May 2026

DiLA: Disentangled Latent Action World Models

ResearchDGX agent

arXiv:2605.15725v1 Announce Type: cross Abstract: Latent Action Models (LAMs) enable the learning of world models from unlabeled video by inferring abstract actions between consecutive frames. However

Do Chinese models speak Chinese languages?

ResearchDGX agent

arXiv:2504.00289v3 Announce Type: replace-cross Abstract: The release of top-performing open-weight LLMs has cemented China's role as a leading force in AI development. Do these models support languag

Imperfect World Models are Exploitable

SafetyDGX agent

arXiv:2605.15960v1 Announce Type: new Abstract: We propose a novel definition of model exploitation in reinforcement learning. Informally, a world model is exploitable if it implies that one policy sh

Learning Structured Robot Policies from Vision-Language Models via Synthetic Neuro-Symbolic Supervision

Model ReleasesDGX agent

arXiv:2604.02812v2 Announce Type: replace Abstract: Vision-Language Models (VLMs) have recently demonstrated strong capabilities in mapping multimodal observations to robot behaviors. However, most cu

SARVLM: A Vision Language Foundation Model for Semantic Understanding in SAR Imagery

Model ReleasesDGX agent

arXiv:2510.22665v3 Announce Type: replace-cross Abstract: Synthetic Aperture Radar (SAR) is a critical imaging modality due to its all-weather operational capability. Although recent advances in self-

SkyLink: A Large Vision-Language Model Driven Re-ranking Framework for Cross-View UAV geolocalization

Model ReleasesDGX agent

arXiv:2603.08063v3 Announce Type: replace Abstract: Cross-view UAV geolocalization is fundamentally a challenging large-scale image retrieval task, aiming to determine the geographic coordinates of Un

Structure Abstraction and Generalization in a Hippocampal-Entorhinal Inspired World Model

Model ReleasesDGX agent

arXiv:2605.15733v1 Announce Type: cross Abstract: Humans abstract experiences into structured representations to facilitate pattern inference and knowledge transfer. While the hippocampal-entorhinal (

Tadpole: Autoencoders as Foundation Models for 3D PDEs with Online Learning

Model ReleasesDGX agent

arXiv:2605.15284v1 Announce Type: new Abstract: We introduce Tadpole, a novel foundation model for three-dimensional partial differential equations (PDEs) that addresses key challenges in transferabil

15 May 2026

Breaking Dual Bottlenecks: Evolving Unified Multimodal Models into Self-Adaptive Interleaved Visual Reasoners

Model ReleasesDGX agent

arXiv:2605.14709v1 Announce Type: new Abstract: Recent unified models integrate multimodal understanding and generation within a single framework. However, an 'understanding-generation gap' persists,

Deceive, Detect, and Disclose: Large Language Models Play Mini-Mafia

Model ReleasesDGX agent

arXiv:2509.23023v3 Announce Type: replace Abstract: Large language models are increasingly deployed in multi-agent settings whose outcomes hinge on social intelligence, motivating evaluations of their

GFMate: Empowering Graph Foundation Models with Test-time Prompt Tuning

Model ReleasesDGX agent

arXiv:2605.14809v1 Announce Type: new Abstract: Graph prompt tuning has shown great potential in graph learning by introducing trainable prompts to enhance the model performance in conventional single

HERO: Hierarchical Extrapolation and Refresh for Efficient World Models

ResearchDGX agent

arXiv:2508.17588v2 Announce Type: replace Abstract: Generation-driven world models create immersive virtual environments but suffer slow inference due to the iterative nature of diffusion models. Whil

Model page: https://ollama.com/library/glm-5.1

Local AiDGX agent

GLM-5.1 is a language model available through Ollama's model library, accessible via the Ollama platform for local deployment and use. The model can be pulled and run locally using Ollama's tools, mak

Proxy Compression for Language Modeling

SafetyDGX agent

arXiv:2602.04289v2 Announce Type: replace Abstract: Modern language models are trained almost exclusively on token sequences produced by a fixed tokenizer, an external lossless compressor often over U

REALM: Retrospective Encoder Alignment for LFP Modeling

Model ReleasesDGX agent

arXiv:2605.14867v1 Announce Type: cross Abstract: Spike activity has been the dominant neural signal for behavior decoding due to its high spatial and temporal resolution. However, as brain-computer i

SANA-WM: Efficient Minute-Scale World Modeling with Hybrid Linear Diffusion Transformer

Model ReleasesDGX agent

arXiv:2605.15178v1 Announce Type: new Abstract: We introduce SANA-WM, an efficient 2.6B-parameter open-source world model natively trained for one-minute generation, synthesizing high-fidelity, 720p,

Some models to try with Codex: kimi-k2.6:cloud (with vision support) glm-5.1:cloud If you don't yet have a paid subscription with Ollama's c…

Model ReleasesDGX agent

Some models to try with Codex: kimi-k2.6:cloud (with vision support) glm-5.1:cloud If you don't yet have a paid subscription with Ollama's cloud, choose a model that supports reliable tool calling: ne

Synthetic Sociality: How Generative Models Privatize the Social Fabric

HardwareDGX agent

arXiv:2605.14090v1 Announce Type: cross Abstract: We put forth a critical theoretical framework for analyzing generative models both descriptively and normatively. Our thesis is that generative models

Unsupervised learning of acquisition variability in structural connectomes via hybrid latent space modeling

Model ReleasesDGX agent

arXiv:2605.13933v1 Announce Type: cross Abstract: Acquisition differences across sites, scanners, and protocols in dMRI introduce variability that complicates structural connectome analysis. This moti

14 May 2026

Benchmarking Attribute Discrimination in Infant-Scale Vision-Language Models

Model ReleasesDGX agent

arXiv:2512.18951v3 Announce Type: replace Abstract: Infants learn not only object categories but also fine-grained visual attributes such as color, size, and texture from limited experience. Prior inf

Characterizing Universal Object Representations Across Vision Models

SafetyDGX agent

arXiv:2605.13675v1 Announce Type: new Abstract: Deep neural networks trained with different architectures, objectives, and datasets have been reported to converge on similar visual representations. Ho

DAWM: Diffusion Action World Models for Offline Reinforcement Learning via Action-Inferred Transitions

Model ReleasesDGX agent

arXiv:2509.19538v2 Announce Type: replace-cross Abstract: Diffusion-based world models have demonstrated strong capabilities in synthesizing realistic long-horizon trajectories for offline reinforceme

Domain Adaptation of Large Language Models for Polymer-Composite Additive Manufacturing Using Retrieval-Augmented Generation and Fine-Tuning

Model ReleasesDGX agent

arXiv:2605.12516v1 Announce Type: cross Abstract: General-purpose large language models (LLMs) often struggle to generate reliable responses in specialized engineering domains due to limited domain gr

Entropy Aware Reward Guidance for Diffusion Language Model Alignment

Model ReleasesDGX agent

arXiv:2602.05000v2 Announce Type: replace-cross Abstract: Reward guidance, also known as posterior sampling, is a popular method for test-time adaptation and post-training in continuous diffusion mode

TiCo: Time-Controllable Spoken Dialogue Model

Model ReleasesDGX agent

arXiv:2603.22267v2 Announce Type: replace-cross Abstract: We introduce TiCo, a time-controllable spoken dialogue model (SDM) that follows time-constrained instructions (e.g., 'Please generate a respon

UNIV: Unified Foundation Model for Infrared and Visible Modalities

Model ReleasesDGX agent

arXiv:2509.15642v3 Announce Type: replace Abstract: Joint RGB-infrared perception is essential for achieving robustness under diverse weather and illumination conditions. Although foundation models ex

13 May 2026

Do Enterprise Systems Need Learned World Models? The Importance of Context to Infer Dynamics

Model ReleasesDGX agent

arXiv:2605.12178v1 Announce Type: cross Abstract: World models enable agents to anticipate the effects of their actions by internalizing environment dynamics. In enterprise systems, however, these dyn

EsoLang-Bench: Evaluating Genuine Reasoning in Large Language Models via Esoteric Programming Languages

Model ReleasesDGX agent

arXiv:2603.09678v2 Announce Type: replace-cross Abstract: Large language models achieve near-ceiling performance on code generation benchmarks, yet most of the programming languages used by popular be

HEBATRON: A Hebrew-Specialized Open-Weight Mixture-of-Experts Language Model

Model ReleasesDGX agent

arXiv:2605.11255v1 Announce Type: new Abstract: We present Hebatron, a Hebrew-specialized open-weight large language model built on the NVIDIA Nemotron-3 sparse Mixture-of-Experts architecture. Traini

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer

Model ReleasesDGX agent

arXiv:2605.11061v1 Announce Type: new Abstract: The evolution of visual generative models has long been constrained by fragmented architectures relying on disjoint text encoders and external VAEs. In

LatentRouter: Can We Choose the Right Multimodal Model Before Seeing Its Answer?

SafetyDGX agent

arXiv:2605.11301v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) have heterogeneous strengths across OCR, chart understanding, spatial reasoning, visual question answering, c

Mechanistic Interpretability of ASR models using Sparse Autoencoders

ApplicationsDGX agent

arXiv:2605.12225v1 Announce Type: new Abstract: Understanding the internal machinations of deep Transformer-based NLP models is more crucial than ever as these models see widespread use in various dom

MotionBench: Benchmarking and Improving Fine-grained Video Motion Understanding for Vision Language Models

Model ReleasesDGX agent

arXiv:2501.02955v2 Announce Type: replace Abstract: In recent years, vision language models (VLMs) have made significant advancements in video understanding. However, a crucial capability - fine-grain

Multi-Stream LLMs: Unblocking Language Models with Parallel Streams of Thoughts, Inputs and Outputs

AgentsDGX agent

arXiv:2605.12460v1 Announce Type: cross Abstract: The continued improvements in language model capability have unlocked their widespread use as drivers of autonomous agents, for example in coding or c

Pretraining Strategies and Scaling for ECG Foundation Models: A Systematic Study

ResearchDGX agent

arXiv:2605.12241v1 Announce Type: cross Abstract: Specialized foundation models are beginning to emerge in various medical subdomains, but pretraining methodologies and parametric scaling with the siz

Probabilistic Calibration Is a Trainable Capability in Language Models

Model ReleasesDGX agent

arXiv:2605.11845v1 Announce Type: new Abstract: Language models are increasingly used in settings where outputs must satisfy user-specified randomness constraints, yet their generation probabilities a

StoicLLM: Preference Optimization for Philosophical Alignment in Small Language Models

Model ReleasesDGX agent

arXiv:2605.11483v1 Announce Type: new Abstract: While large language models excel at factual adaptation, their ability to internalize nuanced philosophical frameworks under severe data constraints rem

← Previous
1…5253545556…999
Next →