AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
59,840 results
Model Releases

Phoneme- vs. Character-Level Targets and Selective State-Space Models for Intracortical Brain-to-Text

DGX agent

arXiv:2607.26751v1 Announce Type: new Abstract: State-of-the-art intracortical brain-to-text systems pair a neural-sequence phone decoder with an external language model. Two design axes remain undere

model-releasesarxiv-cs-cl
30 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Tutorials

StatePlay: State-Aware Game World Models for Mechanics-Consistent Generation

DGX agent

arXiv:2607.26754v1 Announce Type: new Abstract: Recent game world models can generate visually realistic and interactive environments conditioned on player actions. However, games are not defined by p

tutorialsarxiv-cs-cv
30 Jul 2026
Safety

Do Models Fake Alignment Without Clear Consequences?

DGX agent

arXiv:2607.24758v1 Announce Type: new Abstract: Large language models are capable of recognizing evaluation contexts and altering their behavior to reflect evaluator expectations rather than typical d

safetyarxiv-cs-ai
29 Jul 2026
Safety

Faces of Fairness: Examining Bias in Facial Expression Recognition Datasets and Models

DGX agent

arXiv:2502.11049v3 Announce Type: replace Abstract: Automated Facial Expression Recognition (FER), involves two critical aspects: data and model design. Both significantly influence bias and fairness

safetyarxiv-cs-cv
29 Jul 2026
Model Releases

Instruction-Tuned Models Locally Reuse Human Syntax More Than Humans Do

DGX agent

arXiv:2607.26015v1 Announce Type: new Abstract: Syntactic convergence (the tendency of speakers to adapt in language towards the grammatical profiles of their interlocutors) is a well-documented featu

model-releasesarxiv-cs-cl
29 Jul 2026
Model Releases

Model + harness. We have barely begun to understand the best ways to do harness engineering. A huge amount of untapped potential even withou…

DGX agent

Model + harness. We have barely begun to understand the best ways to do harness engineering. A huge amount of untapped potential even without models getting better (but models are getting better) Turn

model-releasesethan-mollick--x
29 Jul 2026
Model Releases

SpecPrefetch: Parameter-Efficient Expert Prefetching for Sparse MoE Foundation Models

DGX agent

arXiv:2607.24787v1 Announce Type: new Abstract: Sparse Mixture-of-Experts (MoE) models expand foundation model capacity through conditional expert activation, but their full expert pools remain diffic

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

A Scale-adaptive Vision Model Links C. elegans Neuronal Morphology to Behavior for Neurotoxicity Assessment

DGX agent

arXiv:2607.23183v1 Announce Type: cross Abstract: Neurological disorders are a leading cause of global disability and are increasingly linked to environmental chemical exposures. Yet neurotoxicity ass

model-releasesarxiv-cs-cv
28 Jul 2026
Model Releases

Chart Deception in Vision-Language Models: From Vulnerability to Mitigation

DGX agent

arXiv:2607.22600v1 Announce Type: new Abstract: Information visualizations are widely used to communicate patterns, trends, and outliers, yet deceptive design choices-such as truncated or inverted axe

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

IKS-Instruct: A 24,000-Example Multilingual Dataset for Teaching Language Models Indian Knowledge Systems

DGX agent

arXiv:2607.23322v1 Announce Type: new Abstract: Instruction tuning has become the standard method for adapting large language models to follow human intent, yet existing instruction datasets are domin

model-releasesarxiv-cs-cl
28 Jul 2026
Model Releases

K3 already got in the top 5 most liked models of all time on Hugging Face, just 24 hours after being released! Ahead of Llama 3, Whisper and…

DGX agent

K3 is a language model that entered the top five most‑liked models on Hugging Face merely 24 hours after its release. The achievement surprised many, placing it ahead of prominent models such as Llama

model-releasesclem-delangue--x
28 Jul 2026
Model Releases

microsoft/Mage-VL · Hugging Face - An Efficient Codec-Native Streaming Multimodal Foundation Model

DGX agent

Mage-VL is a codec-native, proactive-streaming multimodal foundation model for image and video understanding, whose visual encoder is trained entirely from scratch at a compact 4B scale. It targets a

model-releasesr-localllama
28 Jul 2026
Model Releases

Multi-Objective Structured Pruning of LLMs for Latency and Model Size Optimization

DGX agent

arXiv:2607.22583v1 Announce Type: new Abstract: Large Language Models (LLMs) have achieved widespread adoption because of their strong reasoning and query-response capabilities. However, deploying the

model-releasesarxiv-cs-ai
28 Jul 2026
Research

Rethinking the Generation Order of Block Diffusion Language Models

DGX agent

arXiv:2607.24306v1 Announce Type: new Abstract: Diffusion language models enable flexible arbitrary-order generation, but existing sampling methods are mostly designed for early masked diffusion model

researcharxiv-cs-cl
28 Jul 2026
Model Releases

StepX-Edge: An On-Device UI Vision-Language Model via Architecture-Training-Deployment Co-Design

DGX agent

arXiv:2607.22708v1 Announce Type: new Abstract: Deploying a vision-language model with full UI understanding on end devices has long been trapped between accuracy and efficiency: on one side is the ac

model-releasesarxiv-cs-cv
28 Jul 2026
Model Releases

The Half-Lives of Generative-AI Evidence: A 40-Record Audit, a Claim-Currency Framework, and a Reflexive Case of Frontier-Model-Assisted Research

DGX agent

arXiv:2607.24032v1 Announce Type: new Abstract: Generative-AI evaluations can become historical before publication, yet calendar age does not affect every conclusion equally. This paper has two linked

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Understanding Tone-Dependent Inference Cost in Large Language Models

DGX agent

arXiv:2607.23915v1 Announce Type: cross Abstract: We examine how prompt tone affects both accuracy of the LLM answers and inference cost as reflected in output-token consumption. Experiments were perf

model-releasesarxiv-cs-ai
28 Jul 2026
Research

Atlas 2 -- Foundation models for clinical deployment

DGX agent

arXiv:2601.05148v2 Announce Type: replace Abstract: Pathology foundation models substantially advanced the possibilities in computational pathology --- yet tradeoffs in terms of performance, robustnes

researcharxiv-cs-cv
27 Jul 2026
Industry

We've joined the alliance. Open-weight models will ensure that we live in a safer digital world, and that America does not get left behind

DGX agent

We've joined the alliance. Open-weight models will ensure that we live in a safer digital world, and that America does not get left behind Attackers have frontier AI. Defenders need a frontier AI ecos

industryarthur-mensch--x
27 Jul 2026
Model Releases

Announcing the Claude Code-compatible interface for our new Fugu-Ultra v1.1! 🐡 Put a dynamically coordinated team of frontier models to wor…

DGX agent

Announcing the Claude Code-compatible interface for our new Fugu-Ultra v1.1! 🐡 Put a dynamically coordinated team of frontier models to work inside the coding workflow you already know. Instead of rel

model-releasesdavid-ha--x
26 Jul 2026
Model Releases

DatedGPT: Preventing Lookahead Bias in Large Language Models with Time-Aware Pretraining

DGX agent

arXiv:2603.11838v2 Announce Type: replace Abstract: Large language models pretrained on internet-scale data risk lookahead bias in forecasting tasks, as they may have already seen the true outcome dur

model-releasesarxiv-cs-cl
24 Jul 2026
Safety

For my first post, I’m sharing a letter @NVIDIA signed on why open models matter. AI will transform every industry, power every company, and…

DGX agent

For my first post, I’m sharing a letter @NVIDIA signed on why open models matter. AI will transform every industry, power every company, and be built by every country. Open models strengthen safety an

safetyclem-delangue--x
24 Jul 2026
Model Releases

GigaPath-Flash and GigaTIME-Flash: Efficient Pathology Foundation Models for Whole-Slide and Tumor Microenvironment Analysis

DGX agent

arXiv:2607.18218v2 Announce Type: replace-cross Abstract: Foundation models have emerged as a driving force in computational pathology, with the potential to transform cancer diagnosis, prognosis, and

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Learning causality from internet videos in latent space first, and then using RL to teach the foundation model how to act. This approach is …

DGX agent

Learning causality from internet videos in latent space first, and then using RL to teach the foundation model how to act. This approach is 30× cheaper than Gemini 3.1 Flash on pretraining and achieve

model-releasesyann-lecun--x
24 Jul 2026
Safety

Open models for the win!

DGX agent

Open models for the win! For my first post, I’m sharing a letter @NVIDIA signed on why open models matter. AI will transform every industry, power every company, and be built by every country. Open mo

safetyclem-delangue--x
24 Jul 2026
Model Releases

The Washington Post processed 1.79B input tokens per month through Together AI, running open models like Llama and Mistral in production wit…

DGX agent

The Washington Post processed 1.79B input tokens per month through Together AI, running open models like Llama and Mistral in production with predictable costs and full control over the model stack. T

model-releasestogether-ai--x
24 Jul 2026
Model Releases

Training Large Language Models for Self-Explanation Faithfulness

DGX agent

arXiv:2607.21090v1 Announce Type: cross Abstract: We propose a Reinforcement Learning (RL) method to directly optimize the faithfulness of self-explanations - the extent to which a model's generated r

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Cost per successful task: Benchmarking Kimi K3, GPT-5.5, and 8 more AI models

DGX agent

Arize and Fireworks benchmarked 10 AI models across 2,400 agent runs. Learn why cost per successful task beats token price for model evaluation and routing. The post Cost per successful task: Benchmar

model-releasesarize-ai
23 Jul 2026
Model Releases

Huge launch from @tryramp. Different steps in an agent workflow can use different models. This can help reduce costs significantly without s…

DGX agent

Huge launch from @tryramp. Different steps in an agent workflow can use different models. This can help reduce costs significantly without sacrificing performance. Model routing will become a core par

model-releasesdair-ai--x
20 Jul 2026
Model Releases

Who’s Afraid of Chinese Models?

DGX agent

Who’s Afraid of Chinese Models? Interesting proposal from Ben Thompson that both addresses the hypocrisy of labs outlawing distillation against their models despite training on unlicensed data, and co

model-releasessimon-willison
20 Jul 2026
Model Releases

Evaluating Large Language Models on Misconceptions in Multi-Turn Medical Conversations

DGX agent

arXiv:2607.12884v1 Announce Type: new Abstract: Patients seeking medical information often ask questions that embed incorrect assumptions or misconceptions. In such cases, safe medical communication r

model-releasesarxiv-cs-cl
15 Jul 2026
Safety

FlowWAM: Optical Flow as a Unified Action Representation for World Action Models

DGX agent

arXiv:2607.13017v1 Announce Type: cross Abstract: World Action Models (WAMs) are able to leverage pretrained video generators for both world modeling and action prediction. However, directly leveragin

safetyarxiv-cs-cv
15 Jul 2026
Model Releases

Scaling Point-in-Time Language Models

DGX agent

arXiv:2607.11889v1 Announce Type: cross Abstract: Large language models trained on unrestricted internet corpora inevitably embed information from the future, introducing lookahead bias that compromis

model-releasesarxiv-cs-ai
15 Jul 2026
Model Releases

U.S. open-source models are quickly gaining ground. @Nvidia's newest Nemotron Ultra is fast growing on Ollama and unlocking complex, longer …

DGX agent

U.S. open‑source AI models are rapidly gaining popularity, with NVIDIA’s newest model, **Nemotron Ultra**, becoming a prominent entry on the Ollama platform. On Ollama, Nemotron Ultra is quickly scali

model-releasesollama--x
14 Jul 2026
Model Releases

Another big reason to use combination of frontier models. Chain-of-thought monitoring is treated as a reliable safety layer for agents. This…

DGX agent

Another big reason to use combination of frontier models. Chain-of-thought monitoring is treated as a reliable safety layer for agents. This DeepMind-affiliated study shows the layer can be argued out

model-releasesdair-ai--x
12 Jul 2026
Model Releases

AtomBench: A Benchmarking Framework for Generative Crystal Reconstruction Models in Conventional Superconductors

DGX agent

arXiv:2510.16165v2 Announce Type: replace Abstract: A key question in benchmarking generative crystal reconstruction models is how the amount and type of crystallographic information provided to a gen

model-releasesarxiv-cs-lg
8 Jul 2026
Agents

MoWorld: A Flash World Model

DGX agent

arXiv:2607.06216v1 Announce Type: new Abstract: The future of World Models depends not only on scaling model capability, but also on scaling practicality and inference efficiency. High-frame-rate infe

agentsarxiv-cs-cv
8 Jul 2026
Model Releases

Chasing Moving Targets with Online Self-Play Reinforcement Learning for Safer Language Models

DGX agent

arXiv:2506.07468v4 Announce Type: replace-cross Abstract: Conventional large language model (LLM) safety alignment relies on a reactive, disjoint loop: attackers exploit a static model, then defenders

model-releasesarxiv-cs-cl
7 Jul 2026
Model Releases

Continual Model Merging with Test-Time Adaptation for Whole-Slide Image Analysis

DGX agent

arXiv:2607.04755v1 Announce Type: new Abstract: Model merging offers a practical alternative to conventional continual learning by integrating independently fine-tuned models without retaining previou

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

Is Your Benchmark Still Useful? Dynamic Benchmarking for Code Language Models

DGX agent

arXiv:2503.06643v2 Announce Type: replace-cross Abstract: In this paper, we tackle a critical challenge in model evaluation: how to keep code benchmarks useful when models might have already seen them

model-releasesarxiv-cs-cl
7 Jul 2026
Model Releases

Mask2Real-WM: Segmentation Masks as a Sim-to-Real Bridge for Controllable Dexterous World Models

DGX agent

arXiv:2607.04546v1 Announce Type: cross Abstract: Action-conditioned world models allow robots to predict the future consequences of candidate actions without additional physical interaction, supporti

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Sakana AI (@SakanaAILabs) is now a model vendor on Merge Gateway, and Fugu Ultra is live through them. It's a multi-agent orchestration mode…

DGX agent

Sakana AI (@SakanaAILabs) is now a model vendor on Merge Gateway, and Fugu Ultra is live through them. It's a multi-agent orchestration model that routes across frontier models behind one API. You get

model-releasesdavid-ha--x
7 Jul 2026
Model Releases

The Remarkable Effectiveness of Providing AI Agents with Natural Language Tools: A Replication Study Validating NLT Performance Across 14 Models

DGX agent

arXiv:2607.03953v1 Announce Type: cross Abstract: This study independently replicates and extends the Natural Language Tools (NLT) framework of Johnson et al.~(2025), which questions the use of struct

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

introducing tinyrouter i reverse engineered the routing architecture behind Skana AI's Fugu and built replication for open frontier models. …

DGX agent

introducing tinyrouter i reverse engineered the routing architecture behind Skana AI's Fugu and built replication for open frontier models. it's a tiny ~10K parameter LLM router that learns which mode

model-releasesclem-delangue--x
4 Jul 2026
Model Releases

Discrete Diffusion Language Models for Interactive Radiology Report Drafting

DGX agent

arXiv:2607.01436v1 Announce Type: new Abstract: Diffusion language models, which generate text by denoising a token canvas bidirectionally instead of emitting tokens left to right, have become competi

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Gravity-Awareness: Deep Learning Models and LLM Simulation of Human Awareness in Altered Gravity

DGX agent

arXiv:2511.05536v2 Announce Type: replace-cross Abstract: Earth s gravity fundamentally shapes human behaviour. The brain encodes this force as an internal model of gravity, enabling the prediction an

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Liquid Latent State Dynamics for Interpretable Turbofan Degradation Modeling

DGX agent

arXiv:2607.01986v1 Announce Type: new Abstract: Multivariate time-series models for prognostics are often evaluated by point prediction accuracy, yet their internal states rarely expose a coherent deg

model-releasesarxiv-cs-lg
3 Jul 2026
Tutorials

Harnessing the Latent Space: From Steering Vectors to Model Calibrators for Control and Trust

DGX agent

arXiv:2607.00083v1 Announce Type: cross Abstract: Language models have changed from unreliable text generators to highly-capable large models with trillions of parameters. Capability increases come ha

tutorialsarxiv-cs-ai
2 Jul 2026
← Previous
1…2627282930…1247
Next →