AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,404 results
Model Releases

WildRoadBench: A Wild Aerial Road-Damage Grounding Benchmark for Vision-Language Models and Autonomous Agents

DGX agent

arXiv:2605.20306v1 Announce Type: new Abstract: We introduce WildRoadBench, a wild aerial road-damage grounding benchmark that couples direct visual grounding by vision-language models with autonomous

model-releasesarxiv-cs-cv
21 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

A Systematic Failure Analysis of Vision Foundation Models for Open Set Iris Presentation Attack Detection

DGX agent

arXiv:2605.19020v1 Announce Type: new Abstract: Vision foundation models have demonstrated strong transferability across diverse visual recognition tasks and are increasingly considered for biometric

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

Active Learning of Fractional-Order Viscoelastic Model Parameters for Realistic Haptic Rendering

DGX agent

arXiv:2512.00667v2 Announce Type: replace-cross Abstract: Effective medical simulators necessitate realistic haptic rendering of biological tissues that exhibit viscoelastic material properties, such

model-releasesarxiv-cs-ro
20 May 2026
Model Releases

Backdooring Masked Diffusion Language Models

DGX agent

arXiv:2605.19262v1 Announce Type: new Abstract: Masked diffusion language models (MDLMs) are emerging as a compelling new paradigm for text generation, but their training-time security remains largely

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

Did we ever learn what model won gold at the IMO from OpenAI? It was a year ago and it was called an unreleased internal general purpose mod…

DGX agent

Did we ever learn what model won gold at the IMO from OpenAI? It was a year ago and it was called an unreleased internal general purpose model back then. Has GPT-5.5 Pro Extended caught up with whatev

model-releasesethan-mollick--x
20 May 2026
Model Releases

DLEBench: Evaluating Small-scale Object Editing Ability for Instruction-based Image Editing Model

DGX agent

arXiv:2602.23622v2 Announce Type: replace-cross Abstract: Significant progress has been made in the field of Instruction-based Image Editing Models (IIEMs). However, while these models demonstrate pla

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Entry-level guide to the use of large language models for medical research

DGX agent

arXiv:2410.18856v4 Announce Type: replace Abstract: Frontier large language models (LLMs), such as GPT-5, Claude 4.5, Gemini 3, Llama 4, and DeepSeek-R1, represent a transformative class of AI tools c

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

i don't think i need cloud models anymore

DGX agent

i don't think i need cloud models anymore MTP speedup Qwen by 2.5x in Atomic Chat Dense vs MoE models on 2x RTX 5090 Qwen3.6 27B: 51 → 117 tps +137% Qwen3.6 35B-A3B: 218 → 267 tps +25% MTP drafts seve

model-releasesclem-delangue--x
20 May 2026
Research

Neural Network Models for Contextual Regression

DGX agent

arXiv:2603.24400v2 Announce Type: replace-cross Abstract: We propose a neural network model for contextual regression in which the regression model depends on contextual features that determine the ac

researcharxiv-cs-lg
20 May 2026
Safety

PROWL: Prioritized Regret-Driven Optimization for World Model Learning

DGX agent

arXiv:2605.18803v1 Announce Type: cross Abstract: Modern action-conditioned video world models achieve strong short-horizon visual realism, yet remain unreliable on rare, interaction-critical transiti

safetyarxiv-cs-ai
20 May 2026
Research

Token by Token, Compromised: Backdoor Vulnerabilities in Unified Autoregressive Models

DGX agent

arXiv:2605.19227v1 Announce Type: cross Abstract: Unified autoregressive models (UAMs) are transformer models that generate text as well as image tokens within a single autoregressive pass. Shared par

researcharxiv-cs-ai
20 May 2026
Model Releases

ZeroUnlearn: Few-Shot Knowledge Unlearning in Large Language Models

DGX agent

arXiv:2605.18879v1 Announce Type: cross Abstract: Large language models inevitably retain sensitive information, defined as inputs that may induce harmful generations, due to training on massive web c

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Also had some early access to Gemini 3.5 Flash. Very fast for a flash model and very capable, though not as powerful as a full frontier mode…

DGX agent

Also had some early access to Gemini 3.5 Flash. Very fast for a flash model and very capable, though not as powerful as a full frontier model. I added it to the gallery or procedurally generated one-s

model-releasesethan-mollick--x
19 May 2026
Research

Better Together: Evaluating the Complementarity of Earth Embedding Models

DGX agent

arXiv:2605.18667v1 Announce Type: new Abstract: Earth embedding models transform Earth observation data into embeddings uniquely tied to locations on the Earth's surface. These models are typically ev

researcharxiv-cs-cv
19 May 2026
Model Releases

By now, you've probably heard about Gemini Omni, our new model designed to create anything from any input, starting with video. But... what'…

DGX agent

Google AI announced Gemini Omni, a new multimodal model capable of generating diverse content types from various input formats, with initial focus on video generation capabilities. The model represent

model-releasesgoogle-ai--x
19 May 2026
Model Releases

CarbonScaling: Extending Neural Scaling Laws for Carbon Footprint in Large Language Models

DGX agent

arXiv:2508.06524v2 Announce Type: replace-cross Abstract: Large language models (LLMs) increasingly follow neural scaling laws that tie performance gains to rapidly expanding computational budgets, ra

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Cerebras is now running Kimi K2.6 – a trillion parameter model – in enterprise trials. At ~1,000 tokens/s, this is the fastest frontier mode…

DGX agent

Cerebras is now running Kimi K2.6 – a trillion parameter model – in enterprise trials. At ~1,000 tokens/s, this is the fastest frontier model performance ever measured by Artificial Analysis @Artifici

model-releaseskimi-moonshot--x
19 May 2026
Model Releases

DecoupleSearch: Decouple Planning and Search via Hierarchical Reward Modeling

DGX agent

arXiv:2510.21712v2 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) systems have emerged as a pivotal methodology for enhancing Large Language Models (LLMs) through the dyna

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Disentangling Ambiguity from Instability in Large Language Models: A Clinical Text-to-SQL Case Study

DGX agent

arXiv:2602.12015v2 Announce Type: replace Abstract: Deploying large language models for clinical Text-to-SQL requires distinguishing two qualitatively different causes of output diversity: (i) input a

model-releasesarxiv-cs-cl
19 May 2026
Local Ai

Locally Coherent Parallel Decoding in Diffusion Language Models

DGX agent

arXiv:2603.20216v2 Announce Type: replace-cross Abstract: Diffusion language models (DLMs) have emerged as a promising alternative to autoregressive (AR) models, offering sub-linear generation latency

local-aiarxiv-cs-ai
19 May 2026
Model Releases

Med-V1: Small Language Models for Zero-shot and Scalable Biomedical Evidence Attribution

DGX agent

arXiv:2603.05308v2 Announce Type: replace-cross Abstract: Assessing whether an article supports an assertion is essential for hallucination detection and claim verification. While large language model

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Membership Inference Attacks on Discrete Diffusion Language Models

DGX agent

arXiv:2605.16445v1 Announce Type: cross Abstract: Masked Diffusion Language Models MDLMs replace autoregressive generation with iterative demasking and their privacy properties are largely unstudied.

model-releasesarxiv-cs-ai
19 May 2026
Research

Preference Instability in Reward Models: Detection and Mitigation via Sparse Autoencoders

DGX agent

arXiv:2605.16339v1 Announce Type: new Abstract: Preference learning in large language models relies on reward models as proxies for human judgment. However, these models frequently exhibit preference

researcharxiv-cs-lg
19 May 2026
Model Releases

QSTRBench: a New Benchmark to Evaluate the Ability of Language Models to Reason with Qualitative Spatial and Temporal Calculi

DGX agent

arXiv:2605.18380v1 Announce Type: new Abstract: We introduce an extensive qualitative spatial and temporal reasoning (QSTR) benchmark for evaluating large language models (LLMs). We pose questions con

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

TabH2O: A Unified Foundation Model for Tabular Prediction

DGX agent

arXiv:2605.18383v1 Announce Type: new Abstract: We present TabH2O, a foundation model for tabular data that performs classification and regression in a single forward pass via in-context learning. Tab

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

The Illusion of Specialization: Unveiling the Domain-Invariant 'Standing Committee' in Mixture-of-Experts Models

DGX agent

arXiv:2601.03425v2 Announce Type: replace-cross Abstract: Mixture of Experts models are widely assumed to achieve domain specialization through sparse routing. In this work, we question this assumptio

model-releasesarxiv-cs-ai
19 May 2026
Safety

UniAlign: A Model-Agnostic Framework for Robust Network Traffic Classification under Distribution Shifts

DGX agent

arXiv:2605.17575v1 Announce Type: cross Abstract: Network traffic classification (NTC) models often suffer severe performance degradation when deployed in real-world environments due to distribution s

safetyarxiv-cs-ai
19 May 2026
Model Releases

What Does the AI Doctor Value? Auditing Pluralism in the Clinical Ethics of Language Models

DGX agent

arXiv:2605.18738v1 Announce Type: new Abstract: Medicine is inherently pluralistic. Principles such as autonomy, beneficence, nonmaleficence, and justice routinely conflict, and such ethical dilemmas

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Who Generated This 3D Asset? Learning Source Attribution for Generative 3D Models

DGX agent

arXiv:2605.18132v1 Announce Type: cross Abstract: Generative 3D models are deployed in gaming, robotics, and immersive creation, making source attribution critical: given a 3D asset, can we identify w

model-releasesarxiv-cs-ai
19 May 2026
Research

DiLA: Disentangled Latent Action World Models

DGX agent

arXiv:2605.15725v1 Announce Type: cross Abstract: Latent Action Models (LAMs) enable the learning of world models from unlabeled video by inferring abstract actions between consecutive frames. However

researcharxiv-cs-ai
18 May 2026
Research

Do Chinese models speak Chinese languages?

DGX agent

arXiv:2504.00289v3 Announce Type: replace-cross Abstract: The release of top-performing open-weight LLMs has cemented China's role as a leading force in AI development. Do these models support languag

researcharxiv-cs-ai
18 May 2026
Safety

Imperfect World Models are Exploitable

DGX agent

arXiv:2605.15960v1 Announce Type: new Abstract: We propose a novel definition of model exploitation in reinforcement learning. Informally, a world model is exploitable if it implies that one policy sh

safetyarxiv-cs-ai
18 May 2026
Model Releases

Learning Structured Robot Policies from Vision-Language Models via Synthetic Neuro-Symbolic Supervision

DGX agent

arXiv:2604.02812v2 Announce Type: replace Abstract: Vision-Language Models (VLMs) have recently demonstrated strong capabilities in mapping multimodal observations to robot behaviors. However, most cu

model-releasesarxiv-cs-ro
18 May 2026
Model Releases

SARVLM: A Vision Language Foundation Model for Semantic Understanding in SAR Imagery

DGX agent

arXiv:2510.22665v3 Announce Type: replace-cross Abstract: Synthetic Aperture Radar (SAR) is a critical imaging modality due to its all-weather operational capability. Although recent advances in self-

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

SkyLink: A Large Vision-Language Model Driven Re-ranking Framework for Cross-View UAV geolocalization

DGX agent

arXiv:2603.08063v3 Announce Type: replace Abstract: Cross-view UAV geolocalization is fundamentally a challenging large-scale image retrieval task, aiming to determine the geographic coordinates of Un

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

Structure Abstraction and Generalization in a Hippocampal-Entorhinal Inspired World Model

DGX agent

arXiv:2605.15733v1 Announce Type: cross Abstract: Humans abstract experiences into structured representations to facilitate pattern inference and knowledge transfer. While the hippocampal-entorhinal (

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Tadpole: Autoencoders as Foundation Models for 3D PDEs with Online Learning

DGX agent

arXiv:2605.15284v1 Announce Type: new Abstract: We introduce Tadpole, a novel foundation model for three-dimensional partial differential equations (PDEs) that addresses key challenges in transferabil

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

Breaking Dual Bottlenecks: Evolving Unified Multimodal Models into Self-Adaptive Interleaved Visual Reasoners

DGX agent

arXiv:2605.14709v1 Announce Type: new Abstract: Recent unified models integrate multimodal understanding and generation within a single framework. However, an 'understanding-generation gap' persists,

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

Deceive, Detect, and Disclose: Large Language Models Play Mini-Mafia

DGX agent

arXiv:2509.23023v3 Announce Type: replace Abstract: Large language models are increasingly deployed in multi-agent settings whose outcomes hinge on social intelligence, motivating evaluations of their

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

GFMate: Empowering Graph Foundation Models with Test-time Prompt Tuning

DGX agent

arXiv:2605.14809v1 Announce Type: new Abstract: Graph prompt tuning has shown great potential in graph learning by introducing trainable prompts to enhance the model performance in conventional single

model-releasesarxiv-cs-lg
15 May 2026
Research

HERO: Hierarchical Extrapolation and Refresh for Efficient World Models

DGX agent

arXiv:2508.17588v2 Announce Type: replace Abstract: Generation-driven world models create immersive virtual environments but suffer slow inference due to the iterative nature of diffusion models. Whil

researcharxiv-cs-cv
15 May 2026
Local Ai

Model page: https://ollama.com/library/glm-5.1

DGX agent

GLM-5.1 is a language model available through Ollama's model library, accessible via the Ollama platform for local deployment and use. The model can be pulled and run locally using Ollama's tools, mak

local-aiollama--x
15 May 2026
Safety

Proxy Compression for Language Modeling

DGX agent

arXiv:2602.04289v2 Announce Type: replace Abstract: Modern language models are trained almost exclusively on token sequences produced by a fixed tokenizer, an external lossless compressor often over U

safetyarxiv-cs-cl
15 May 2026
Model Releases

REALM: Retrospective Encoder Alignment for LFP Modeling

DGX agent

arXiv:2605.14867v1 Announce Type: cross Abstract: Spike activity has been the dominant neural signal for behavior decoding due to its high spatial and temporal resolution. However, as brain-computer i

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

SANA-WM: Efficient Minute-Scale World Modeling with Hybrid Linear Diffusion Transformer

DGX agent

arXiv:2605.15178v1 Announce Type: new Abstract: We introduce SANA-WM, an efficient 2.6B-parameter open-source world model natively trained for one-minute generation, synthesizing high-fidelity, 720p,

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

Some models to try with Codex: kimi-k2.6:cloud (with vision support) glm-5.1:cloud If you don't yet have a paid subscription with Ollama's c…

DGX agent

Some models to try with Codex: kimi-k2.6:cloud (with vision support) glm-5.1:cloud If you don't yet have a paid subscription with Ollama's cloud, choose a model that supports reliable tool calling: ne

model-releasesollama--x
15 May 2026
Hardware

Synthetic Sociality: How Generative Models Privatize the Social Fabric

DGX agent

arXiv:2605.14090v1 Announce Type: cross Abstract: We put forth a critical theoretical framework for analyzing generative models both descriptively and normatively. Our thesis is that generative models

hardwarearxiv-cs-lg
15 May 2026
Model Releases

Unsupervised learning of acquisition variability in structural connectomes via hybrid latent space modeling

DGX agent

arXiv:2605.13933v1 Announce Type: cross Abstract: Acquisition differences across sites, scanners, and protocols in dMRI introduce variability that complicates structural connectome analysis. This moti

model-releasesarxiv-cs-ai
15 May 2026
← Previous
1…6667686970…1259
Next →