AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
59,399 results
17 Apr 2026

Internal Knowledge Without External Expression: Probing the Generalization Boundary of a Classical Chinese Language Model

Model ReleasesDGX agent

arXiv:2604.14180v1 Announce Type: new Abstract: We train a 318M-parameter Transformer language model from scratch on a curated corpus of 1.56 billion tokens of pure Classical Chinese, with zero Englis

Logo-LLM: Local and Global Modeling with Large Language Models for Time Series Forecasting

Local AiDGX agent

arXiv:2505.11017v2 Announce Type: replace Abstract: Time series forecasting is critical across multiple domains, where time series data exhibit both local patterns and global dependencies. While Trans

Optimize video semantic search intent with Amazon Nova Model Distillation on Amazon Bedrock

TutorialsDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

In this post, we show you how to use Model Distillation, a model customization technique on Amazon Bedrock, to transfer routing intelligence from a large teacher model (Amazon Nova Premier) into a muc

16 Apr 2026

Been having fun with local AI stuff. I don't really like the big models and now that some of them want your ID I'm not sure I'll even stick …

Local AiDGX agent

Been having fun with local AI stuff. I don't really like the big models and now that some of them want your ID I'm not sure I'll even stick around to care that much. Local shit is really powerful now,

Dataset-Level Metrics Attenuate Non-Determinism: A Fine-Grained Non-Determinism Evaluation in Diffusion Language Models

ResearchDGX agent

arXiv:2604.13413v1 Announce Type: new Abstract: Diffusion language models (DLMs) have emerged as a promising paradigm for large language models (LLMs), yet the non-deterministic behavior of DLMs remai

From Plausibility to Verifiability: Risk-Controlled Generative OCR with Vision-Language Models

Model ReleasesDGX agent

arXiv:2603.19790v3 Announce Type: replace Abstract: Modern vision-language models (VLMs) can act as generative OCR engines, yet open-ended decoding can expose rare but consequential failures. We ident

Model page: https://ollama.com/library/qwen3.6

Local AiDGX agent

Qwen3.6 is a model available through Ollama's model library, representing Alibaba's Qwen series of large language models optimized for local deployment. The model can be accessed and run through the O

Saber: An Efficient Sampling with Adaptive Acceleration and Backtracking Enhanced Remasking for Diffusion Language Model

ResearchDGX agent

arXiv:2510.18165v2 Announce Type: replace-cross Abstract: Diffusion language models (DLMs) are emerging as a powerful and promising alternative to the dominant autoregressive paradigm, offering inhere

Transcriptomic Models for Immunotherapy Response Prediction Show Limited Cross-cohort Generalisability

Model ReleasesDGX agent

arXiv:2604.05478v2 Announce Type: replace-cross Abstract: Immune checkpoint inhibitors (ICIs) have transformed cancer therapy; yet substantial proportion of patients exhibit intrinsic or acquired resi

15 Apr 2026

Latent Chain-of-Thought World Modeling for End-to-End Driving

Model ReleasesDGX agent

arXiv:2512.10226v2 Announce Type: replace Abstract: Recent Vision-Language-Action (VLA) models for autonomous driving explore inference-time reasoning as a way to improve driving performance and safet

NoisePrints: Distortion-Free Watermarks for Authorship in Private Diffusion Models

TutorialsDGX agent

arXiv:2510.13793v2 Announce Type: replace Abstract: With the rapid adoption of diffusion models for visual content generation, proving authorship and protecting copyright have become critical. This ch

Safe-SAIL: Towards a Fine-grained Safety Landscape of Large Language Models via Sparse Autoencoder Interpretation Framework

Model ReleasesDGX agent

arXiv:2509.18127v3 Announce Type: replace-cross Abstract: Sparse autoencoders (SAEs) enable interpretability research by decomposing entangled model activations into monosemantic features. However, un

Testing Ollama with Genma 4 and internet search turned on, and got the model extremely confused that it got results from the future

Local AiDGX agent

A Reddit post from the r/ollama community documents a user's experiment running Google's Gemma 4 model locally via Ollama with web search (internet access) enabled, which resulted in the model becomin

Variational Autoencoding Discrete Diffusion with Enhanced Dimensional Correlations Modeling

ResearchDGX agent

arXiv:2505.17384v2 Announce Type: replace-cross Abstract: Discrete diffusion models have recently shown great promise for modeling complex discrete data, with masked diffusion models (MDMs) offering a

14 Apr 2026

And this is a very generous definition of notable. If we are talking frontier models, only the US and China that are even in the race. And t…

ApplicationsDGX agent

And this is a very generous definition of notable. If we are talking frontier models, only the US and China that are even in the race. And that obscures the fact that the Big Three US labs really do s

CArtBench: Evaluating Vision-Language Models on Chinese Art Understanding, Interpretation, and Authenticity

Model ReleasesDGX agent

arXiv:2604.11632v1 Announce Type: new Abstract: We introduce CARTBENCH, a museum-grounded benchmark for evaluating vision-language models (VLMs) on Chinese artworks beyond short-form recognition and Q

Design Principles for Sequence Models via Coefficient Dynamics

Model ReleasesDGX agent

arXiv:2510.09389v2 Announce Type: replace-cross Abstract: Deep sequence models, ranging from Transformers and State Space Models (SSMs) to more recent approaches such as gated linear RNNs, fundamental

Do LLMs Build Spatial World Models? Evidence from Grid-World Maze Tasks

Model ReleasesDGX agent

arXiv:2604.10690v1 Announce Type: new Abstract: Foundation models have shown remarkable performance across diverse tasks, yet their ability to construct internal spatial world models for reasoning and

Edu-MMBias: A Three-Tier Multimodal Benchmark for Auditing Social Bias in Vision-Language Models under Educational Contexts

Model ReleasesDGX agent

arXiv:2604.10200v1 Announce Type: new Abstract: As Vision-Language Models (VLMs) become integral to educational decision-making, ensuring their fairness is paramount. However, current text-centric eva

ExecTune: Effective Steering of Black-Box LLMs with Guide Models

Model ReleasesDGX agent

arXiv:2604.09741v1 Announce Type: cross Abstract: For large language models deployed through black-box APIs, recurring inference costs often exceed one-time training costs. This motivates composed age

ForestPrune: High-ratio Visual Token Compression for Video Multimodal Large Language Models via Spatial-Temporal Forest Modeling

ResearchDGX agent

arXiv:2603.22911v2 Announce Type: replace-cross Abstract: Due to the great saving of computation and memory overhead, token compression has become a research hot-spot for MLLMs and achieved remarkable

GAMBIT: A Gamified Jailbreak Framework for Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2601.03416v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have become widely deployed, yet their safety alignment remains fragile under adversarial inputs. Previous

MLLM-as-a-Judge Exhibits Model Preference Bias

Model ReleasesDGX agent

arXiv:2604.11589v1 Announce Type: new Abstract: Automatic evaluation using multimodal large language models (MLLMs), commonly referred to as MLLM-as-a-Judge, has been widely used to measure model perf

OOWM: Structuring Embodied Reasoning and Planning via Object-Oriented Programmatic World Modeling

Model ReleasesDGX agent

arXiv:2604.09580v1 Announce Type: new Abstract: Standard Chain-of-Thought (CoT) prompting empowers Large Language Models (LLMs) with reasoning capabilities, yet its reliance on linear natural language

Quantitative Introspection in Language Models: Tracking Emotive States Across Conversation

Model ReleasesDGX agent

arXiv:2603.18893v2 Announce Type: replace Abstract: Tracking the internal states of large language models across conversations is important for safety, interpretability, and model welfare, yet current

Seeing No Evil: Blinding Large Vision-Language Models to Safety Instructions via Adversarial Attention Hijacking

Model ReleasesDGX agent

arXiv:2604.10299v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) rely on attention-based retrieval of safety instructions to maintain alignment during generation. Existing attack

Suiren-1.0 Technical Report: A Family of Molecular Foundation Models

ResearchDGX agent

arXiv:2603.21942v2 Announce Type: replace-cross Abstract: We introduce Suiren-1.0, a family of molecular foundation models for the accurate modeling of diverse organic systems. Suiren-1.0 comprising t

TS-Haystack: A Multi-Scale Retrieval Benchmark for Time Series Language Models

Model ReleasesDGX agent

arXiv:2602.14200v4 Announce Type: replace Abstract: Time Series Language Models (TSLMs) are emerging as unified models for reasoning over continuous signals in natural language. However, long-context

Woosh: A Sound Effects Foundation Model

Model ReleasesDGX agent

arXiv:2604.01929v2 Announce Type: replace-cross Abstract: The audio research community depends on open generative models as foundational tools for building novel approaches and establishing baselines.

13 Apr 2026

Advantage-Guided Diffusion for Model-Based Reinforcement Learning

SafetyDGX agent

arXiv:2604.09035v1 Announce Type: new Abstract: Model-based reinforcement learning (MBRL) with autoregressive world models suffers from compounding errors, whereas diffusion world models mitigate this

12 Apr 2026

if you bounce between claude code and codex, that’s completely normal. the models are good at different things and everyone's workflow is dy…

Model ReleasesDGX agent

if you bounce between claude code and codex, that’s completely normal. the models are good at different things and everyone's workflow is dynamic but you’re also paying multiple subscriptions, constan

11 Apr 2026

All Tools

ConceptsDGX agent

Auto-generated index of all tools mentioned across the wiki.

10 Apr 2026

AgriChain Visually Grounded Expert Verified Reasoning for Interpretable Agricultural Vision Language Models

Model ReleasesDGX agent

arXiv:2604.07814v1 Announce Type: new Abstract: Accurate and interpretable plant disease diagnosis remains a major challenge for vision-language models (VLMs) in real-world agriculture. We introduce A

An empirical study of LoRA-based fine-tuning of large language models for automated test case generation

Model ReleasesDGX agent

arXiv:2604.06946v1 Announce Type: cross Abstract: Automated test case generation from natural language requirements remains a challenging problem in software engineering due to the ambiguity of requir

Mitigating Visual Context Degradation in Large Multimodal Models: A Training-Free Decoupled Agentic Framework

Model ReleasesDGX agent

arXiv:2509.23322v2 Announce Type: replace Abstract: With the continuous expansion of Large Language Models (LLMs) and advances in reinforcement learning, LLMs have demonstrated exceptional reasoning c

Multi-objective Evolutionary Merging Enables Efficient Reasoning Models

Model ReleasesDGX agent

arXiv:2604.06465v1 Announce Type: cross Abstract: Reasoning models have demonstrated remarkable capabilities in solving complex problems by leveraging long chains of thought. However, this more delibe

OmniTabBench: Mapping the Empirical Frontiers of GBDTs, Neural Networks, and Foundation Models for Tabular Data at Scale

Model ReleasesDGX agent

arXiv:2604.06814v1 Announce Type: cross Abstract: While traditional tree-based ensemble methods have long dominated tabular tasks, deep neural networks and emerging foundation models have challenged t

Open Harness, separated from model providers is a critical architectural pattern.

AgentsDGX agent

An **Open Harness** is a unified architectural layer that sits between AI agents and model providers, abstracting away provider-specific APIs and patterns. Because every AI agent harness has its o...

oslash Source Models Leak What They Shouldn't nrightarrow: Unlearning Zero-Shot Transfer in Domain Adaptation Through Adversarial Optimization

Model ReleasesDGX agent

arXiv:2604.08238v1 Announce Type: new Abstract: The increasing adaptation of vision models across domains, such as satellite imagery and medical scans, has raised an emerging privacy risk: models may

Share your Gemma 4 builds or the model variants you’re training in the replies below!

Model ReleasesDGX agent

Google released Gemma 4 in April 2025 as its most capable open-weight model family to date, built on the same research as Gemini 3 and licensed under Apache 2.0 for unrestricted commercial use, fin...

Small Vision-Language Models are Smart Compressors for Long Video Understanding

Model ReleasesDGX agent

arXiv:2604.08120v1 Announce Type: cross Abstract: Adapting Multimodal Large Language Models (MLLMs) for hour-long videos is bottlenecked by context limits. Dense visual streams saturate token budgets

SOLAR: Communication-Efficient Model Adaptation via Subspace-Oriented Latent Adapter Reparametrization

Model ReleasesDGX agent

arXiv:2604.08368v1 Announce Type: cross Abstract: Parameter-efficient fine-tuning (PEFT) methods, such as LoRA, enable scalable adaptation of foundation models by injecting low-rank adapters. However,

The ATOM Report: Measuring the Open Language Model Ecosystem

Model ReleasesDGX agent

arXiv:2604.07190v1 Announce Type: cross Abstract: We present a comprehensive adoption snapshot of the leading open language models and who is building them, focusing on the ~1.5K mainline open models

The Impact of Steering Large Language Models with Persona Vectors in Educational Applications

Model ReleasesDGX agent

arXiv:2604.07102v1 Announce Type: cross Abstract: Activation-based steering can personalize large language models at inference time, but its effects in educational settings remain unclear. We study pe

WorldMAP: Bootstrapping Vision-Language Navigation Trajectory Prediction with Generative World Models

ResearchDGX agent

arXiv:2604.07957v1 Announce Type: cross Abstract: Vision-language models (VLMs) and generative world models are opening new opportunities for embodied navigation. VLMs are increasingly used as direct

9 Apr 2026

Do you understand what this means? For the first time, an Open Weight models is #1 on CyberSecutity. Sure there’s Mythos but we don’t have i…

IndustryDGX agent

The specific tweet from @0xSero is not directly accessible, and the search results don't surface the exact model or event being referenced in that post. However, based on the context clues in the t...

Open Harness, Model Choice, Open Memory (take it wherever you need), Open Protocols

AgentsDGX agent

Open Harness, Model Choice, Open Memory (take it wherever you need), Open Protocols Open everything 🔥 Open Harness, Model Choice, Open Memory (take it wherever you need), Open Protocols basically we’r

13 Aug 2026

SpaceXAI releases flagship Grok 4.6 model with advanced reasoning capabilities

Model ReleasesDGX agent

SpaceXAI today released Grok 4.6, a large language model that it says can outperform Anthropic PBC’s Claude Fable 5 in some areas. SpaceXAI was known as xAI until last month. The Elon Musk-founded art

12 Aug 2026

Flex-pi: A Multi-Stream World-Action Model with Compute Flexibility

Model ReleasesDGX agent

arXiv:2608.10860v1 Announce Type: cross Abstract: World-action models (WAMs) predict the future to act better, but nearly all of them predict only RGB latents, trained purely for pixel reconstruction,

11 Aug 2026

Improving Constraint Models with LLM Agents

AgentsDGX agent

arXiv:2608.08127v1 Announce Type: new Abstract: The runtime of Constraint Programming (CP) solvers is highly sensitive to modeling choices, such as symmetry breaking, implied constraints, global const

Learning How the World Evolves: Extrapolative Video World Models via Latent Dynamics Reasoning

Model ReleasesDGX agent

arXiv:2608.09926v1 Announce Type: new Abstract: The world evolves following its dynamics, i.e., its laws of motion. However, leading video diffusion models largely fit the pixels without modeling how

Luth-2: New State-of-the-Art French Small Language Models

Model ReleasesDGX agent

Hey everyone, Today we release Luth-2-0.8B and Luth2-2-2B, two non-reasoning models that set a new state of the art for French across a wide variety of tasks for their size 🚀 A few notable scores on F

Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models

Model ReleasesDGX agent

arXiv:2608.09666v1 Announce Type: new Abstract: Recent advances in visual generative models have enabled high-quality image and video generation, but evaluating these models often demands sampling hun

Toy project: a chat title model that fits in 5 MiB of ram

Model ReleasesDGX agent

Not even sure if I'm allowed to post this, what with the 'completely/primarily LLM generated copy' rule (the post itself is fine, but the repo/model I'm sharing definitely is, whoops) and the whole li

10 Aug 2026

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models

ApplicationsDGX agent

arXiv:2608.06729v1 Announce Type: cross Abstract: While Vision-Language-Action (VLA) models have advanced embodied AI, their fundamentally reactive paradigm severely limits performance in partially ob

Capacity Confounds and Coverage Guarantees in Adaptive Sub-model Federated Learning

Model ReleasesDGX agent

arXiv:2608.07157v1 Announce Type: new Abstract: Sub-model federated learning lets resource-constrained clients train width-reduced versions of a global model, but existing methods allocate capacity by

Capek 0.5: An Execution-Centric Vision-Language Model for Embodied Intelligence

Model ReleasesDGX agent

arXiv:2608.06756v1 Announce Type: new Abstract: Vision-language models are increasingly serving as the reasoning core of embodied agents. Robot execution is inherently iterative: each action reshapes

Divergent Response Modes in Frontier Language Models Under Steering Pressure

Model ReleasesDGX agent

arXiv:2608.06578v1 Announce Type: new Abstract: Frontier language models are trained using distinct data, objectives, and safety pipelines. Whether these differences produce measurably different behav

Newton-Schulz Retraction-Based Inference Enables Hidden Quantum Markov Models to Outperform Classical HMMs

Model ReleasesDGX agent

arXiv:2608.06554v1 Announce Type: new Abstract: Hidden Markov models (HMMs) are widely used probabilistic models for discrete sequential data but can be limited when hidden dynamics are complex. Hidde

The successor to Llama is here, and Meta is revitalizing focus on open weights with their new Muse Glimmer - a leading 30B param model desig…

Model ReleasesDGX agent

The successor to Llama is here, and Meta is revitalizing focus on open weights with their new Muse Glimmer - a leading 30B param model designed for always-on local agent use, small enough to run on a

← Previous
1…1920212223…990
Next →