AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,536 results
Model Releases

LIBERO-Safety: A Comprehensive Benchmark for Physical and Semantic Safety in Vision-Language-Action Models

DGX agent

arXiv:2606.23686v1 Announce Type: new Abstract: Despite the impressive manipulation capabilities of Vision-Language-Action (VLA) models, their operational safety under strict constraints remains large

model-releasesarxiv-cs-ro
23 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Load Testing for Machine Learning Model Serving Systems at Scale

DGX agent

arXiv:2606.22013v1 Announce Type: new Abstract: Machine learning (ML) model serving has become a dominant consumer of GPU infrastructure, yet capacity planning in these systems remains largely ad hoc.

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

LUQ: Layerwise Ultra-Low Bit Quantization for Multimodal Large Language Models

DGX agent

arXiv:2509.23729v3 Announce Type: replace Abstract: Large Language Models (LLMs) with multimodal capabilities have revolutionized vision-language tasks, but their deployment often requires huge memory

model-releasesarxiv-cs-cv
23 Jun 2026
Research

Model Inversion meets Cryptographic Fuzzy Extractors

DGX agent

arXiv:2510.25687v3 Announce Type: replace-cross Abstract: Model inversion attacks pose an open challenge to privacy-sensitive applications that use machine learning (ML) models. For example, face auth

researcharxiv-cs-lg
23 Jun 2026
Safety

PAIWorld: A 3D-Consistent World Foundation Model for Robotic Manipulation

DGX agent

arXiv:2606.18375v2 Announce Type: replace Abstract: World foundation models (WFMs) are powerful simulators, yet they predominantly operate in a single-view setting and lack the multi-view 3D consisten

safetyarxiv-cs-ro
23 Jun 2026
Model Releases

Post-Training Speech Enhancement Language Models with Perceptual Rewards

DGX agent

arXiv:2606.21458v1 Announce Type: new Abstract: Speech enhancement language models achieve strong results when trained on discrete audio tokens, but their optimization relies on token-level cross-entr

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Residue-Level Attributions in Protein Language Models Do Not Recover Allergen Epitopes

DGX agent

arXiv:2606.22181v1 Announce Type: new Abstract: Deep allergenicity classifiers are increasingly used in safety screening of novel foods, and recent protein language models have substantially improved

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Translating Inference-Time Control to Radiology Vision-Language Models: Activation Steering for Pneumonia Classification on Chest X-rays

DGX agent

arXiv:2606.20852v1 Announce Type: new Abstract: Inference-time engineering can alter model behavior without fine-tuning. However, its utility for improving diagnostic performance in medical vision-lan

model-releasesarxiv-cs-cv
23 Jun 2026
Safety

VRPO: Rethinking Value Modeling for Robust RL under Noisy Supervision in LLM Post-Training

DGX agent

arXiv:2508.03058v2 Announce Type: replace Abstract: Reinforcement Learning (RL) in real-world environments often suffers from ambiguous or incomplete reward supervision, which undermines policy stabil

safetyarxiv-cs-lg
23 Jun 2026
Applications

Fugu stands shoulder-to-shoulder with leading models like Fable and Mythos across the industry's most rigorous engineering, scientific, and …

DGX agent

Fugu stands shoulder-to-shoulder with leading models like Fable and Mythos across the industry's most rigorous engineering, scientific, and reasoning benchmarks. Read the full blog: https://sakana.ai/

applicationsdavid-ha--x
22 Jun 2026
Model Releases

Yes, is model agnostic and general purpose

DGX agent

Harrison Chase confirms that 'Yes' (likely referring to LangChain or a related tool/framework) is model-agnostic and designed as a general-purpose solution, meaning it can work across different AI mod

model-releasesharrison-chase--x
21 Jun 2026
Hardware

Hot take: GLM 5.2 might be the first open/public model that actually changes the enterprise AI cost equation. I played with it for a few hou…

DGX agent

Hot take: GLM 5.2 might be the first open/public model that actually changes the enterprise AI cost equation. I played with it for a few hours today after friends told me to stop ignoring it. I expect

hardwareclem-delangue--x
20 Jun 2026
Model Releases

Google open-sources speedy DiffusionGemma text diffusion model

DGX agent

Google LLC today released DiffusionGemma, a large language model based on an emerging machine learning approach known as text diffusion. The company says the algorithm can generate text four times fas

model-releasessiliconangle
11 Jun 2026
Model Releases

OpenMedReason: Scientific Reasoning Supervision for Medical Vision-Language Models

DGX agent

arXiv:2606.12169v1 Announce Type: cross Abstract: High-stakes clinical use of large vision-language models (LVLMs) requires reasoning that is grounded in visual evidence and clinical knowledge, not ju

model-releasesarxiv-cs-ai
11 Jun 2026
Agents

Skill-Augmented AI Agents for Medical Research Analysis: An Exploratory Multi-Model Human Evaluation in an NSCLC Transcriptomic Biomarker Task

DGX agent

arXiv:2606.11830v1 Announce Type: new Abstract: Background. Large language models and AI agents are increasingly used to support biomedical research, but native model outputs may omit key analytical s

agentsarxiv-cs-ai
11 Jun 2026
Model Releases

TimeRouter: Efficient and Adaptive Routing of Time-Series Foundation Models

DGX agent

arXiv:2606.11625v1 Announce Type: new Abstract: Time-series foundation models (TSFMs) are increasingly explored as predictive experts within emerging agentic time-series systems. However, TSFMs exhibi

model-releasesarxiv-cs-lg
11 Jun 2026
Model Releases

When Does Language Matter? Multilingual Instructions Reveal Step-wise Language Sensitivity in Vision-Language-Action Models

DGX agent

arXiv:2606.11906v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have shown strong performance in language-conditioned robotic manipulation, yet their robustness to linguistic varia

model-releasesarxiv-cs-cl
11 Jun 2026
Model Releases

World Pilot: Steering Vision-Language-Action Models with World-Action Priors

DGX agent

arXiv:2606.12403v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models inherit semantic grounding from large-scale pretraining and perform competently across in-distribution manipulation

model-releasesarxiv-cs-ro
11 Jun 2026
Model Releases

Efficient-WAM: A 1B-Parameter World-Action Model with Low-Cost Future Imagination

DGX agent

arXiv:2606.10040v1 Announce Type: new Abstract: World-Action Models (WAMs) have emerged as a promising paradigm for embodied control by coupling future visual prediction with action generation. Howeve

model-releasesarxiv-cs-ro
10 Jun 2026
Tutorials

Entropy, Disagreement, and the Limits of Foundation Models in Genomics

DGX agent

arXiv:2604.04287v2 Announce Type: replace-cross Abstract: Foundation models in genomics have shown mixed success compared to their counterparts in natural language processing. Yet, the reasons for the

tutorialsarxiv-cs-cl
10 Jun 2026
Model Releases

FADA: Accessible fetal ultrasound interpretation and annotation with a selectively distilled unified vision-language model

DGX agent

arXiv:2606.11106v1 Announce Type: cross Abstract: A global shortage of trained sonographers limits prenatal ultrasound screening in low- and middle-income countries, where over half of pregnant women

model-releasesarxiv-cs-ai
10 Jun 2026
Safety

Improving Adversarial Transferability on Vision-Language Pre-training Models via Surrogate-Specific Bias Correction

DGX agent

arXiv:2606.10571v1 Announce Type: cross Abstract: Adversarial examples reveal vulnerabilities in Vision-Language Pre-training (VLP) models and provide insights for improving robustness. A key property

safetyarxiv-cs-ai
10 Jun 2026
Model Releases

P3D-Bench: Benchmarking MLLMs for Parametric 3D Generation and Structural Reasoning

DGX agent

arXiv:2606.11152v1 Announce Type: new Abstract: Multimodal large language models can write code to produce complex programs as well as use programs to do 3D modeling, which opens up a new avenue for 3

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

Recalling Too Well: Sycophancy Evaluation and Mitigation in Memory-Augmented Models

DGX agent

arXiv:2606.10949v1 Announce Type: new Abstract: Persistent memory systems promise to make LLMs more helpful by storing user beliefs over time. We show they also make models less correct by systematica

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

SARM2: Multi-Task Stage Aware Reward Modeling for Self Improving Robotic Manipulation

DGX agent

arXiv:2606.10305v1 Announce Type: new Abstract: Fine-tuning vision-language-action (VLA) policies for long-horizon manipulation still relies heavily on behavior cloning, which requires costly high-qua

model-releasesarxiv-cs-ro
10 Jun 2026
Model Releases

The Shibboleth Effect: Auditing the Cross-Lingual Distributional Skew of Large Language Models

DGX agent

arXiv:2606.11082v1 Announce Type: new Abstract: This study investigates cross-lingual distributional skew (the Shibboleth Effect) in frontier large language models (LLMs) subjected to sustained advers

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

Anthropic sets AI performance records with new Mythos 5, Fable 5 frontier models

DGX agent

Anthropic PBC today introduced Claude Mythos 5 and Claude Fable 5, two large language models that it says outperform the competition across a wide range of benchmarks. The LLMs are derived from the Cl

model-releasessiliconangle
9 Jun 2026
Model Releases

Are Reasoning Vision-Language Models Robust to Semantic Visual Distractions?

DGX agent

arXiv:2606.08894v1 Announce Type: new Abstract: Reasoning Vision-Language Models (VLMs) achieve strong performance on complex multimodal tasks, but reliable real-world application requires handling vi

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Bokeh Diffusion: Defocus Blur Control in Text-to-Image Diffusion Models

DGX agent

arXiv:2503.08434v5 Announce Type: replace-cross Abstract: Recent advances in large-scale text-to-image models have revolutionized creative fields by generating visually captivating outputs from textua

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

BUDDY: BUdget-Driven DYnamic Depth Routing for Adaptive Large Language Model Inference

DGX agent

arXiv:2606.09514v1 Announce Type: new Abstract: Large language models (LLMs) incur high inference cost due to their depth and parameter scale. Depth pruning can reduce latency by skipping redundant Tr

model-releasesarxiv-cs-lg
9 Jun 2026
Safety

CheXanatomy: Anatomy-Aware Vision-Language Modeling for Chest Radiographs

DGX agent

arXiv:2606.08420v1 Announce Type: new Abstract: Vision-language models (VLMs) pretrained on large-scale image-text pairs demonstrate strong image-level understanding, but are primarily optimized for g

safetyarxiv-cs-cv
9 Jun 2026
Model Releases

Explaining Black-Box Language Models: Learning to Optimize Linguistically-Structured Word Subsets

DGX agent

arXiv:2606.08497v1 Announce Type: new Abstract: As deep language models (DLMs) are increasingly deployed in high-stakes domains such as healthcare, understanding their decision rationale becomes param

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Fable 5 is now available in Claude Code and Cowork Fable is the best model I have used for coding, by a wide margin. It is a big step up, en…

DGX agent

Fable 5 is now available in Claude Code and Cowork Fable is the best model I have used for coding, by a wide margin. It is a big step up, enabling less prompts and steers, more efficient token use, be

model-releasesboris-cherny--x
9 Jun 2026
Model Releases

GlobeAudio: A Multilingual Multicultural Benchmark for Naturalistic Evaluation of Large Audio-Language Models

DGX agent

arXiv:2606.08194v1 Announce Type: cross Abstract: Large Audio-Language Models (LALMs) integrate audio perception and language understanding within a unified framework, enabling a wide range of real-wo

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

HASA: Subnet Allocation for Compute-Constrained Model-Heterogeneous Federated Learning

DGX agent

arXiv:2606.07621v1 Announce Type: cross Abstract: Edge services increasingly use federated learning to personalize on-device models while keeping sensitive data local. In practice, deployments must ha

model-releasesarxiv-cs-ai
9 Jun 2026
Applications

Large Models for Time Series and Spatio-Temporal Data: A Survey and Outlook

DGX agent

arXiv:2310.10196v3 Announce Type: replace-cross Abstract: Temporal data, including time series and spatio-temporal data, are pervasive in real-world applications. Generated in massive volumes by physi

applicationsarxiv-cs-ai
9 Jun 2026
Model Releases

Phase transition in large language models and the criticality of natural languages

DGX agent

arXiv:2406.05335v3 Announce Type: replace-cross Abstract: Generation of text and speech in natural languages can be modeled as a stochastic process. This idea dates back to the seminal work of Markov

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

TempoBench: Evaluating Temporal Causal Reasoning in Large Language Models

DGX agent

arXiv:2510.27544v2 Announce Type: replace Abstract: Temporal reasoning involves understanding how systems evolve over time through input-driven state transitions. A key aspect is temporal causal reaso

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Thinking-Based Non-Thinking: Solving the Reward Hacking Problem in Training Hybrid Reasoning Models via Reinforcement Learning

DGX agent

arXiv:2601.04805v2 Announce Type: replace Abstract: Large reasoning models (LRMs) have attracted much attention due to their exceptional performance. However, their performance mainly stems from think

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Today, we released Gemini 3.5 Live Translate, our latest audio model for live speech-to-speech translation. It supports over 70 languages an…

DGX agent

Today, we released Gemini 3.5 Live Translate, our latest audio model for live speech-to-speech translation. It supports over 70 languages and starts translating as soon as you start talking, streaming

model-releasesgoogle-ai--x
9 Jun 2026
Local Ai

why does ollama unloads models automatically?

DGX agent

Ollama automatically unloads inactive models from memory based on inactivity parameters, with models being removed after a specified duration of non-use . By default, Ollama unloads a model after 5 mi

local-air-ollama
9 Jun 2026
Model Releases

A Four-Condition Diagnostic Protocol for Evidence Utilization in Long-Context and Retrieval-Augmented Language Models

DGX agent

arXiv:2606.06758v1 Announce Type: new Abstract: Final-answer accuracy, retrieval recall, and citation overlap do not by themselves identify whether a long-context or retrieval-augmented language model

model-releasesarxiv-cs-cl
8 Jun 2026
Local Ai

CULTURESCORE: Evaluating Cultural Faithfulness in Video Generation Models

DGX agent

arXiv:2606.07311v1 Announce Type: cross Abstract: As video generation models like Veo 3.1 and LTX-2 advance, their ability to accurately represent diverse global cultures remains a critical yet unders

local-aiarxiv-cs-ai
8 Jun 2026
Safety

Data-Efficient Autoregressive-to-Diffusion Language Models via On-Policy Distillation

DGX agent

arXiv:2606.06712v1 Announce Type: cross Abstract: We study the transformation of autoregressive models (ARLMs) into diffusion language models (DLMs). Rather than pretraining from scratch, prior work r

safetyarxiv-cs-ai
8 Jun 2026
Local Ai

Generalization of Diffusion Models Arises with a Balanced Representation Space

DGX agent

arXiv:2512.20963v3 Announce Type: replace-cross Abstract: Diffusion models excel at generating high-quality, diverse samples, yet they risk memorizing training data when overfit to the training object

local-aiarxiv-cs-cv
8 Jun 2026
Research

Rethinking Genomic Modeling Through Optical Character Recognition

DGX agent

arXiv:2602.02014v2 Announce Type: replace-cross Abstract: Recent genomic foundation models largely adopt large language model architectures that treat DNA as a one-dimensional token sequence. However,

researcharxiv-cs-ai
8 Jun 2026
Model Releases

ShallowBench: Benchmarking Generative Drug Design Models on Shallow-Pocket Targets

DGX agent

arXiv:2606.06717v1 Announce Type: cross Abstract: While generative AI models have demonstrated remarkable success in structure-based drug design, they predominantly rely on deep binding pockets and st

model-releasesarxiv-cs-ai
8 Jun 2026
Hardware

American Open Source is so back. 9 / 30 of the models on page 1 of Huggingface are published by Nvidia.

DGX agent

Nvidia has significantly increased its presence in open-source AI models, with 9 out of 30 models on the first page of Hugging Face's model hub published by the company. This observation reflects Nvid

hardwareclem-delangue--x
7 Jun 2026
← Previous
1…8182838485…1262
Next →