AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
59,840 results
Model Releases

BAGEL: Benchmarking Animal Knowledge Expertise in Language Models

DGX agent

arXiv:2604.16241v1 Announce Type: cross Abstract: Large language models have shown strong performance on broad-domain knowledge and reasoning benchmarks, but it remains unclear how well language model

model-releasesarxiv-cs-ai
20 Apr 2026
Research
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Cost-Aware Model Orchestration for LLM-based Systems

DGX agent

arXiv:2512.01099v2 Announce Type: replace Abstract: As modern artificial intelligence (AI) systems become more advanced and capable, they can leverage a wide range of tools and models to perform compl

researcharxiv-cs-ai
20 Apr 2026
Model Releases

LLM attribution analysis across different fine-tuning strategies and model scales for automated code compliance

DGX agent

arXiv:2604.15589v1 Announce Type: cross Abstract: Existing research on large language models (LLMs) for automated code compliance has primarily focused on performance, treating the models as black box

model-releasesarxiv-cs-ai
20 Apr 2026
Applications

Using Large Language Models and Knowledge Graphs to Improve the Interpretability of Machine Learning Models in Manufacturing

DGX agent

arXiv:2604.16280v1 Announce Type: new Abstract: Explaining Machine Learning (ML) results in a transparent and user-friendly manner remains a challenging task of Explainable Artificial Intelligence (XA

applicationsarxiv-cs-ai
20 Apr 2026
Research

Cornfigurator: Automated Planning for Any-to-Any Multimodal Model Serving

DGX agent

arXiv:2512.14098v3 Announce Type: replace Abstract: Any-to-Any models are an emerging class of multimodal models that accept combinations of text and multimodal data as input and generate them as outp

researcharxiv-cs-lg
17 Apr 2026
Model Releases

Internal Knowledge Without External Expression: Probing the Generalization Boundary of a Classical Chinese Language Model

DGX agent

arXiv:2604.14180v1 Announce Type: new Abstract: We train a 318M-parameter Transformer language model from scratch on a curated corpus of 1.56 billion tokens of pure Classical Chinese, with zero Englis

model-releasesarxiv-cs-cl
17 Apr 2026
Local Ai

Logo-LLM: Local and Global Modeling with Large Language Models for Time Series Forecasting

DGX agent

arXiv:2505.11017v2 Announce Type: replace Abstract: Time series forecasting is critical across multiple domains, where time series data exhibit both local patterns and global dependencies. While Trans

local-aiarxiv-cs-lg
17 Apr 2026
Tutorials

Optimize video semantic search intent with Amazon Nova Model Distillation on Amazon Bedrock

DGX agent

In this post, we show you how to use Model Distillation, a model customization technique on Amazon Bedrock, to transfer routing intelligence from a large teacher model (Amazon Nova Premier) into a muc

tutorialsaws-ml-blog
17 Apr 2026
Local Ai

Been having fun with local AI stuff. I don't really like the big models and now that some of them want your ID I'm not sure I'll even stick …

DGX agent

Been having fun with local AI stuff. I don't really like the big models and now that some of them want your ID I'm not sure I'll even stick around to care that much. Local shit is really powerful now,

local-ainous-research--x
16 Apr 2026
Research

Dataset-Level Metrics Attenuate Non-Determinism: A Fine-Grained Non-Determinism Evaluation in Diffusion Language Models

DGX agent

arXiv:2604.13413v1 Announce Type: new Abstract: Diffusion language models (DLMs) have emerged as a promising paradigm for large language models (LLMs), yet the non-deterministic behavior of DLMs remai

researcharxiv-cs-lg
16 Apr 2026
Model Releases

From Plausibility to Verifiability: Risk-Controlled Generative OCR with Vision-Language Models

DGX agent

arXiv:2603.19790v3 Announce Type: replace Abstract: Modern vision-language models (VLMs) can act as generative OCR engines, yet open-ended decoding can expose rare but consequential failures. We ident

model-releasesarxiv-cs-cv
16 Apr 2026
Local Ai

Model page: https://ollama.com/library/qwen3.6

DGX agent

Qwen3.6 is a model available through Ollama's model library, representing Alibaba's Qwen series of large language models optimized for local deployment. The model can be accessed and run through the O

local-aiollama--x
16 Apr 2026
Research

Saber: An Efficient Sampling with Adaptive Acceleration and Backtracking Enhanced Remasking for Diffusion Language Model

DGX agent

arXiv:2510.18165v2 Announce Type: replace-cross Abstract: Diffusion language models (DLMs) are emerging as a powerful and promising alternative to the dominant autoregressive paradigm, offering inhere

researcharxiv-cs-cl
16 Apr 2026
Model Releases

Transcriptomic Models for Immunotherapy Response Prediction Show Limited Cross-cohort Generalisability

DGX agent

arXiv:2604.05478v2 Announce Type: replace-cross Abstract: Immune checkpoint inhibitors (ICIs) have transformed cancer therapy; yet substantial proportion of patients exhibit intrinsic or acquired resi

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Latent Chain-of-Thought World Modeling for End-to-End Driving

DGX agent

arXiv:2512.10226v2 Announce Type: replace Abstract: Recent Vision-Language-Action (VLA) models for autonomous driving explore inference-time reasoning as a way to improve driving performance and safet

model-releasesarxiv-cs-cv
15 Apr 2026
Tutorials

NoisePrints: Distortion-Free Watermarks for Authorship in Private Diffusion Models

DGX agent

arXiv:2510.13793v2 Announce Type: replace Abstract: With the rapid adoption of diffusion models for visual content generation, proving authorship and protecting copyright have become critical. This ch

tutorialsarxiv-cs-cv
15 Apr 2026
Model Releases

Safe-SAIL: Towards a Fine-grained Safety Landscape of Large Language Models via Sparse Autoencoder Interpretation Framework

DGX agent

arXiv:2509.18127v3 Announce Type: replace-cross Abstract: Sparse autoencoders (SAEs) enable interpretability research by decomposing entangled model activations into monosemantic features. However, un

model-releasesarxiv-cs-ai
15 Apr 2026
Local Ai

Testing Ollama with Genma 4 and internet search turned on, and got the model extremely confused that it got results from the future

DGX agent

A Reddit post from the r/ollama community documents a user's experiment running Google's Gemma 4 model locally via Ollama with web search (internet access) enabled, which resulted in the model becomin

local-air-ollama
15 Apr 2026
Research

Variational Autoencoding Discrete Diffusion with Enhanced Dimensional Correlations Modeling

DGX agent

arXiv:2505.17384v2 Announce Type: replace-cross Abstract: Discrete diffusion models have recently shown great promise for modeling complex discrete data, with masked diffusion models (MDMs) offering a

researcharxiv-cs-cv
15 Apr 2026
Applications

And this is a very generous definition of notable. If we are talking frontier models, only the US and China that are even in the race. And t…

DGX agent

And this is a very generous definition of notable. If we are talking frontier models, only the US and China that are even in the race. And that obscures the fact that the Big Three US labs really do s

applicationsethan-mollick--x
14 Apr 2026
Model Releases

CArtBench: Evaluating Vision-Language Models on Chinese Art Understanding, Interpretation, and Authenticity

DGX agent

arXiv:2604.11632v1 Announce Type: new Abstract: We introduce CARTBENCH, a museum-grounded benchmark for evaluating vision-language models (VLMs) on Chinese artworks beyond short-form recognition and Q

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Design Principles for Sequence Models via Coefficient Dynamics

DGX agent

arXiv:2510.09389v2 Announce Type: replace-cross Abstract: Deep sequence models, ranging from Transformers and State Space Models (SSMs) to more recent approaches such as gated linear RNNs, fundamental

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Do LLMs Build Spatial World Models? Evidence from Grid-World Maze Tasks

DGX agent

arXiv:2604.10690v1 Announce Type: new Abstract: Foundation models have shown remarkable performance across diverse tasks, yet their ability to construct internal spatial world models for reasoning and

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Edu-MMBias: A Three-Tier Multimodal Benchmark for Auditing Social Bias in Vision-Language Models under Educational Contexts

DGX agent

arXiv:2604.10200v1 Announce Type: new Abstract: As Vision-Language Models (VLMs) become integral to educational decision-making, ensuring their fairness is paramount. However, current text-centric eva

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

ExecTune: Effective Steering of Black-Box LLMs with Guide Models

DGX agent

arXiv:2604.09741v1 Announce Type: cross Abstract: For large language models deployed through black-box APIs, recurring inference costs often exceed one-time training costs. This motivates composed age

model-releasesarxiv-cs-ai
14 Apr 2026
Research

ForestPrune: High-ratio Visual Token Compression for Video Multimodal Large Language Models via Spatial-Temporal Forest Modeling

DGX agent

arXiv:2603.22911v2 Announce Type: replace-cross Abstract: Due to the great saving of computation and memory overhead, token compression has become a research hot-spot for MLLMs and achieved remarkable

researcharxiv-cs-ai
14 Apr 2026
Model Releases

GAMBIT: A Gamified Jailbreak Framework for Multimodal Large Language Models

DGX agent

arXiv:2601.03416v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have become widely deployed, yet their safety alignment remains fragile under adversarial inputs. Previous

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

MLLM-as-a-Judge Exhibits Model Preference Bias

DGX agent

arXiv:2604.11589v1 Announce Type: new Abstract: Automatic evaluation using multimodal large language models (MLLMs), commonly referred to as MLLM-as-a-Judge, has been widely used to measure model perf

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

OOWM: Structuring Embodied Reasoning and Planning via Object-Oriented Programmatic World Modeling

DGX agent

arXiv:2604.09580v1 Announce Type: new Abstract: Standard Chain-of-Thought (CoT) prompting empowers Large Language Models (LLMs) with reasoning capabilities, yet its reliance on linear natural language

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Quantitative Introspection in Language Models: Tracking Emotive States Across Conversation

DGX agent

arXiv:2603.18893v2 Announce Type: replace Abstract: Tracking the internal states of large language models across conversations is important for safety, interpretability, and model welfare, yet current

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Seeing No Evil: Blinding Large Vision-Language Models to Safety Instructions via Adversarial Attention Hijacking

DGX agent

arXiv:2604.10299v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) rely on attention-based retrieval of safety instructions to maintain alignment during generation. Existing attack

model-releasesarxiv-cs-cl
14 Apr 2026
Research

Suiren-1.0 Technical Report: A Family of Molecular Foundation Models

DGX agent

arXiv:2603.21942v2 Announce Type: replace-cross Abstract: We introduce Suiren-1.0, a family of molecular foundation models for the accurate modeling of diverse organic systems. Suiren-1.0 comprising t

researcharxiv-cs-ai
14 Apr 2026
Model Releases

TS-Haystack: A Multi-Scale Retrieval Benchmark for Time Series Language Models

DGX agent

arXiv:2602.14200v4 Announce Type: replace Abstract: Time Series Language Models (TSLMs) are emerging as unified models for reasoning over continuous signals in natural language. However, long-context

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

Woosh: A Sound Effects Foundation Model

DGX agent

arXiv:2604.01929v2 Announce Type: replace-cross Abstract: The audio research community depends on open generative models as foundational tools for building novel approaches and establishing baselines.

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

Advantage-Guided Diffusion for Model-Based Reinforcement Learning

DGX agent

arXiv:2604.09035v1 Announce Type: new Abstract: Model-based reinforcement learning (MBRL) with autoregressive world models suffers from compounding errors, whereas diffusion world models mitigate this

safetyarxiv-cs-ai
13 Apr 2026
Model Releases

if you bounce between claude code and codex, that’s completely normal. the models are good at different things and everyone's workflow is dy…

DGX agent

if you bounce between claude code and codex, that’s completely normal. the models are good at different things and everyone's workflow is dynamic but you’re also paying multiple subscriptions, constan

model-releasesharrison-chase--x
12 Apr 2026
Concepts

All Tools

DGX agent

Auto-generated index of all tools mentioned across the wiki.

conceptstoolsindex
11 Apr 2026
Model Releases

AgriChain Visually Grounded Expert Verified Reasoning for Interpretable Agricultural Vision Language Models

DGX agent

arXiv:2604.07814v1 Announce Type: new Abstract: Accurate and interpretable plant disease diagnosis remains a major challenge for vision-language models (VLMs) in real-world agriculture. We introduce A

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

An empirical study of LoRA-based fine-tuning of large language models for automated test case generation

DGX agent

arXiv:2604.06946v1 Announce Type: cross Abstract: Automated test case generation from natural language requirements remains a challenging problem in software engineering due to the ambiguity of requir

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Mitigating Visual Context Degradation in Large Multimodal Models: A Training-Free Decoupled Agentic Framework

DGX agent

arXiv:2509.23322v2 Announce Type: replace Abstract: With the continuous expansion of Large Language Models (LLMs) and advances in reinforcement learning, LLMs have demonstrated exceptional reasoning c

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

Multi-objective Evolutionary Merging Enables Efficient Reasoning Models

DGX agent

arXiv:2604.06465v1 Announce Type: cross Abstract: Reasoning models have demonstrated remarkable capabilities in solving complex problems by leveraging long chains of thought. However, this more delibe

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

OmniTabBench: Mapping the Empirical Frontiers of GBDTs, Neural Networks, and Foundation Models for Tabular Data at Scale

DGX agent

arXiv:2604.06814v1 Announce Type: cross Abstract: While traditional tree-based ensemble methods have long dominated tabular tasks, deep neural networks and emerging foundation models have challenged t

model-releasesarxiv-cs-ai
10 Apr 2026
Agents

Open Harness, separated from model providers is a critical architectural pattern.

DGX agent

An **Open Harness** is a unified architectural layer that sits between AI agents and model providers, abstracting away provider-specific APIs and patterns. Because every AI agent harness has its o...

agentsharrison-chase--x
10 Apr 2026
Model Releases

oslash Source Models Leak What They Shouldn't nrightarrow: Unlearning Zero-Shot Transfer in Domain Adaptation Through Adversarial Optimization

DGX agent

arXiv:2604.08238v1 Announce Type: new Abstract: The increasing adaptation of vision models across domains, such as satellite imagery and medical scans, has raised an emerging privacy risk: models may

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

Share your Gemma 4 builds or the model variants you’re training in the replies below!

DGX agent

Google released Gemma 4 in April 2025 as its most capable open-weight model family to date, built on the same research as Gemini 3 and licensed under Apache 2.0 for unrestricted commercial use, fin...

model-releasesgoogle-ai--x
10 Apr 2026
Model Releases

Small Vision-Language Models are Smart Compressors for Long Video Understanding

DGX agent

arXiv:2604.08120v1 Announce Type: cross Abstract: Adapting Multimodal Large Language Models (MLLMs) for hour-long videos is bottlenecked by context limits. Dense visual streams saturate token budgets

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

SOLAR: Communication-Efficient Model Adaptation via Subspace-Oriented Latent Adapter Reparametrization

DGX agent

arXiv:2604.08368v1 Announce Type: cross Abstract: Parameter-efficient fine-tuning (PEFT) methods, such as LoRA, enable scalable adaptation of foundation models by injecting low-rank adapters. However,

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

The ATOM Report: Measuring the Open Language Model Ecosystem

DGX agent

arXiv:2604.07190v1 Announce Type: cross Abstract: We present a comprehensive adoption snapshot of the leading open language models and who is building them, focusing on the ~1.5K mainline open models

model-releasesarxiv-cs-ai
10 Apr 2026
← Previous
1…2425262728…1247
Next →