AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
59,929 results
Research

AdaJEPA: An Adaptive Latent World Model

DGX agent

arXiv:2606.32026v1 Announce Type: cross Abstract: Latent world models enable planning from high-dimensional observations by predicting future states in a compact latent space. However, these models ar

researcharxiv-cs-ai
1 Jul 2026
Model Releases

Beyond Binary Instrument QA: Probing Instrument Grounding in Music Audio-Language Models

DGX agent
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

arXiv:2606.31338v1 Announce Type: cross Abstract: Recent music audio-language models achieve high accuracy on instrument question-answering benchmarks, but it remains unclear whether this reflects rob

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

CoLT: Teaching Multi-Modal Models to Think with Chain of Latent Thoughts

DGX agent

arXiv:2606.31986v1 Announce Type: new Abstract: Chain-of-thought (CoT) reasoning has enabled multi-modal large language models (MLLMs) to tackle complex visual reasoning tasks by generating explicit i

model-releasesarxiv-cs-cv
1 Jul 2026
Model Releases

Really confused by all the excitement I see in my timeline for a nerfed model. Never seen anything like it. So many will end up very disappo…

DGX agent

Really confused by all the excitement I see in my timeline for a nerfed model. Never seen anything like it. So many will end up very disappointed. Time to rethink how to build around frontier and open

model-releasesdair-ai--x
1 Jul 2026
Model Releases

The Bidirectional Process Reward Model

DGX agent

arXiv:2508.01682v3 Announce Type: replace Abstract: Process Reward Models (PRMs), which assign fine-grained scores to intermediate reasoning steps within a solution trajectory, have emerged as a promi

model-releasesarxiv-cs-cl
1 Jul 2026
Model Releases

You really need to benchmark models for your use case. As soon as judgements & decisions stack on top of each other, the differences between…

DGX agent

You really need to benchmark models for your use case. As soon as judgements & decisions stack on top of each other, the differences between models amplifies, and no standard benchmark will tell you t

model-releasesethan-mollick--x
1 Jul 2026
Model Releases

AURORA: Asymmetry and Update-Induced Rotation for Robust Hallucination Detection in Large Language Models

DGX agent

arXiv:2606.29545v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities across a wide range of natural language processing tasks. However, their tendency

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

BrepLLM: Enabling Large Language Models to Understand Boundary Representations

DGX agent

arXiv:2512.16413v2 Announce Type: replace Abstract: Current token-sequence-based Large Language Models (LLMs) struggle to directly process 3D Boundary Representation (B-rep) models that contain comple

model-releasesarxiv-cs-cv
30 Jun 2026
Safety

ExploreVLA: Dense World Modeling and Exploration for End-to-End Autonomous Driving

DGX agent

arXiv:2604.02714v2 Announce Type: replace Abstract: End-to-end autonomous driving models based on Vision-Language-Action (VLA) architectures have shown promising results by learning driving policies t

safetyarxiv-cs-cv
30 Jun 2026
Model Releases

Fine-Tuning General-Purpose Large Language Models for Agricultural Applications:A Reproducible Framework and Evaluation Protocol Based on Qwen3-8B

DGX agent

arXiv:2606.28992v1 Announce Type: cross Abstract: General-purpose large language models (LLMs) have demonstrated strong abilities in opendomain question answering, information extraction, and text gen

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Introducing Claude Sonnet 5 on AWS: Anthropic’s most capable Sonnet model

DGX agent

Today, we’re excited to announce the availability of Anthropic’s most advanced Sonnet model, Claude Sonnet 5, on Amazon Bedrock and Claude Platform on AWS. Claude Sonnet 5 is the first Sonnet model of

model-releasesaws-ml-blog
30 Jun 2026
Model Releases

Morphing into Hybrid Attention Models

DGX agent

arXiv:2606.30562v1 Announce Type: new Abstract: Hybrid attention models improve long-context efficiency by retaining only a subset of full-attention layers and replacing the remaining layers with line

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

PlantExpertVQA: A Visual Question Answering Dataset for Benchmarking Vision-Language Models in Plant Science

DGX agent

arXiv:2508.17117v3 Announce Type: replace-cross Abstract: Existing plant-disease datasets target classification and detection, leaving vision-language models unable to support interactive, reasoning-b

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Qwen-RobotNav Technical Report: A Scalable Navigation Model Designed for an Agentic Navigation System

DGX agent

arXiv:2606.18112v3 Announce Type: replace-cross Abstract: Agentic navigation systems require a base navigation model whose observation strategy can be externally reconfigured at inference time, becaus

model-releasesarxiv-cs-cv
30 Jun 2026
Local Ai

The strength of clinical evidence is recoverable from language model representations but not from their stated grades

DGX agent

arXiv:2606.29034v1 Announce Type: cross Abstract: Large language models (LLMs) increasingly summarize clinical evidence, where a claim's weight depends on how strongly it is supported. Yet these model

local-aiarxiv-cs-ai
30 Jun 2026
Research

Towards Harnessing the Collaborative Power of Large and Small Models for Domain Tasks

DGX agent

arXiv:2504.17421v2 Announce Type: replace-cross Abstract: Large language models (LMs) offer broad generalization capabilities but require vast amounts of data and computational resources for domain-sp

researcharxiv-cs-ai
30 Jun 2026
Model Releases

Training Vision-Language-Action Models with Dense Embodied Chain-of-Thought Supervision

DGX agent

arXiv:2606.30552v1 Announce Type: cross Abstract: Cross-embodiment transfer in vision-language-action (VLA) models remains challenging because low-level state and action spaces differ fundamentally ac

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

Unified Enhancement of the Generalization and Robustness of Language Models via Bi-Stage Optimization

DGX agent

arXiv:2503.16550v2 Announce Type: replace Abstract: Neural network language models (LMs) are confronted with significant challenges in generalization and robustness. Currently, many studies focus on i

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

VTEdit-Bench: A Comprehensive Benchmark for Multi-Reference Image Editing Models in Virtual Try-On

DGX agent

arXiv:2603.11734v2 Announce Type: replace Abstract: As virtual try-on (VTON) continues to advance, a growing number of real-world scenarios have emerged, pushing beyond the ability of the existing spe

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

You wake up. It’s 2026. The newest open source AI model was trained by federal government employees, Big Balls and Groot.

DGX agent

You wake up. It’s 2026. The newest open source AI model was trained by federal government employees, Big Balls and Groot. FULL INTERVIEW: Engineers Edward Coristine (@as400495) and Tai Groot (@taigrr)

model-releasesclem-delangue--x
30 Jun 2026
Local Ai

Your Data Manifold is Secretly a Reward Model: Shell-LCC for Text-to-Video Generation

DGX agent

arXiv:2606.30248v1 Announce Type: new Abstract: Recent text-to-video (T2V) diffusion models rely heavily on auxiliary reward signals (e.g., via reward models or DPO) to align generated content with hu

local-aiarxiv-cs-cv
30 Jun 2026
Research

An Expectation-Maximization Algorithm for Training Clean Diffusion Models from Corrupted Observations

DGX agent

arXiv:2407.01014v2 Announce Type: replace Abstract: Diffusion models excel in solving imaging inverse problems due to their ability to model complex image priors. However, their reliance on large, cle

researcharxiv-cs-cv
29 Jun 2026
Model Releases

Benchmarking on Tasks That Matter: Dataset Selection for Preserving Model Rankings

DGX agent

arXiv:2606.27997v1 Announce Type: new Abstract: Benchmarks of machine learning models often include many datasets, making evaluation expensive. For efficiency, it is preferable to perform evaluations

model-releasesarxiv-cs-lg
29 Jun 2026
Hardware

DataStates-LLM: Scalable Checkpointing for Transformer Models Using Composable State Providers

DGX agent

arXiv:2601.16956v1 Announce Type: cross Abstract: The rapid growth of Large Transformer-based models, specifically Large Language Models (LLMs), now scaling to trillions of parameters, has necessitate

hardwarearxiv-cs-ai
29 Jun 2026
Tutorials

FlexMoE: One-for-All Nested Intra-Expert Pruning for MoE Language Models

DGX agent

arXiv:2606.27866v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) language models scale model ability with sparsely activated experts, making this architecture a standard recipe for modern larg

tutorialsarxiv-cs-lg
29 Jun 2026
Model Releases

GAIA: A Data Flywheel System for Training GUI Test-Time Scaling Critic Models

DGX agent

arXiv:2601.18197v2 Announce Type: replace Abstract: While Large Vision-Language Models (LVLMs) have significantly advanced GUI agents' capabilities in parsing textual instructions, interpreting screen

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

PEBS: Per-rater Empirical-Bayes Shrinkage for RLHF Reward-Model Calibration

DGX agent

arXiv:2606.27578v1 Announce Type: cross Abstract: Reward models for Reinforcement Learning from Human Feedback (RLHF) pool preferences across thousands of annotators and fit one global affine calibrat

model-releasesarxiv-cs-ai
29 Jun 2026
Applications

Quantum Generative Diffusion Model for Real-World Time Series

DGX agent

arXiv:2606.27561v1 Announce Type: new Abstract: Generative models have achieved remarkable success in data synthesis, though recent advances driven by increasing model scale have introduced challenges

applicationsarxiv-cs-lg
29 Jun 2026
Model Releases

RSICCLLM: A Multimodal Large Language Model for Remote Sensing Image Change Captioning

DGX agent

arXiv:2606.28266v1 Announce Type: new Abstract: Remote Sensing Image Change Captioning (RSICC) aims to describe changes between bi-temporal remote sensing images and holds significant research and app

model-releasesarxiv-cs-cv
29 Jun 2026
Model Releases

S^2-VLA: State-Space Guided Vision-Language-Action Models for Long-Horizon Manipulation

DGX agent

arXiv:2606.27872v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have demonstrated strong capabilities in robotic manipulation, but their performance degrades significantly in lon

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

Good guy DeepSeek gives us accelerated models The most interesting one here is Gemma4-12B, I presume vision included. Might be the best loca…

DGX agent

Good guy DeepSeek gives us accelerated models The most interesting one here is Gemma4-12B, I presume vision included. Might be the best local model in its weight class now, by some margin Qwen 3.5 not

model-releasesclem-delangue--x
28 Jun 2026
Model Releases

6 Fingers, 1 Kidney: Natural Adversarial Medical Images Reveal Critical Weaknesses of Vision-Language Models

DGX agent

arXiv:2512.04238v2 Announce Type: replace Abstract: Vision-language models (VLMs) are increasingly integrated into clinical workflows. However, existing benchmarks primarily assess performance on comm

model-releasesarxiv-cs-cv
26 Jun 2026
Research

Consistent Distributed Ranking of Generative Models via Kernel Distances

DGX agent

arXiv:2310.11714v5 Announce Type: replace Abstract: Ranking generative models based on the fidelity and diversity of their outputs is required to identify the best generator in a group of candidate ge

researcharxiv-cs-lg
26 Jun 2026
Research

Dataset Usage Inference without Shadow Models or Held-out Data

DGX agent

arXiv:2606.26257v1 Announce Type: new Abstract: How much of my data was used to train a machine learning model? Dataset Usage Inference (DUI) aims to answer this by estimating what fraction of a datas

researcharxiv-cs-lg
26 Jun 2026
Model Releases

extsc{DiARC}: Distinguishing Positive and Negative Samples Helps Improving ARC-like Reasoning Ability of Large Language Models

DGX agent

arXiv:2606.26530v1 Announce Type: cross Abstract: The Abstraction and Reasoning Corpus (ARC;~itealp{chollet2019measure}) contains tasks that require summarizing patterns from limited grid samples and

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Great to see the new GPT-5.6 models finally announced. Sad to see this new release strategy where only a select few get access initially. No…

DGX agent

Great to see the new GPT-5.6 models finally announced. Sad to see this new release strategy where only a select few get access initially. Not a win for our industry IMO. Open-source AI must win! Intro

model-releasesdair-ai--x
26 Jun 2026
Safety

Normalizing Flows are Capable Models for Continuous Control

DGX agent

arXiv:2505.23527v4 Announce Type: replace Abstract: Modern reinforcement learning (RL) algorithms have found success by using powerful probabilistic models, such as transformers, energy-based models,

safetyarxiv-cs-lg
26 Jun 2026
Safety

Paying More Attention to Visual Tokens in Self-Evolving Large Multimodal Models

DGX agent

arXiv:2606.27373v1 Announce Type: new Abstract: Recently, self-evolving large multimodal models (LMMs) have received attention for improving visual reasoning in a purely unsupervised setting. However,

safetyarxiv-cs-cv
26 Jun 2026
Model Releases

Post-Training Recipe, More Than Model Family, Shapes Multi-Agent LLM Conversational Behavior

DGX agent

arXiv:2606.20632v2 Announce Type: replace-cross Abstract: Multi-LLM systems use multiple language models to deliberate, judge each other's outputs, or coordinate as agents. Their value depends on the

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Reproducibility Study of 'AlphaEdit: Null-Space Constrained Knowledge Editing for Language Models'

DGX agent

arXiv:2606.26783v1 Announce Type: cross Abstract: Fang et al. (2025) introduced a null-space constrained projection, named AlphaEdit, for locate-then-edit knowledge editing methods, theoretically guar

model-releasesarxiv-cs-cl
26 Jun 2026
Model Releases

Simulation-based inference for rapid Bayesian parameter estimation in epidemiological models: a comparison with MCMC

DGX agent

arXiv:2606.27286v1 Announce Type: new Abstract: Mechanistic epidemiological models are widely used to support infectious disease forecasting and public-health decision making. Bayesian calibration of

model-releasesarxiv-cs-ai
26 Jun 2026
Local Ai

SITUATION EXPLAINED: 70% of frontier model queries could run locally for free. @ClementDelangue, co-founder and CEO of @huggingface: 'There …

DGX agent

SITUATION EXPLAINED: 70% of frontier model queries could run locally for free. @ClementDelangue, co-founder and CEO of @huggingface: 'There was an interesting study from Stanford published last year s

local-aiclem-delangue--x
26 Jun 2026
Model Releases

Where Do Models Find Happiness? Emotion Vectors in Open-Source LLMs

DGX agent

arXiv:2606.26987v1 Announce Type: cross Abstract: Recent work identified emotion vectors in Claude Sonnet 4.5, which are internal representations that encode emotion concepts, causally influence behav

model-releasesarxiv-cs-ai
26 Jun 2026
Agents

A Probabilistic Framework for LLM-Based Model Discovery

DGX agent

arXiv:2602.18266v2 Announce Type: replace Abstract: Automated methods for discovering mechanistic simulator models from observational data offer a promising path toward accelerating scientific progres

agentsarxiv-cs-lg
25 Jun 2026
Model Releases

Benchmarking Deep Learning Models for Laryngeal Cancer Staging Using the LaryngealCT Dataset

DGX agent

arXiv:2510.11047v2 Announce Type: replace Abstract: Laryngeal cancer imaging research lacks standardised public datasets to enable reproducible deep learning (DL) model development. We present Larynge

model-releasesarxiv-cs-cv
25 Jun 2026
Research

Concept Removal for Frontier Image Generative Models

DGX agent

arXiv:2606.25548v1 Announce Type: new Abstract: Image generative models are trained on massive, largely uncurated internet-scale datasets that contain undesirable visual concepts. Efficiently removing

researcharxiv-cs-cv
25 Jun 2026
Model Releases

Detect, Unlearn, Restore: Defending Text Summarization Models Against Data Poisoning

DGX agent

arXiv:2606.26036v1 Announce Type: new Abstract: Training-time data poisoning during fine-tuning poses a significant threat to large language models (LLMs) deployed for abstractive text summarization,

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

How Robust is OCR-Reasoning? Evaluating OCR-Reasoning Robustness of Vision-Language Models under Visual Perturbations

DGX agent

arXiv:2606.26041v1 Announce Type: cross Abstract: Vision-language models (VLMs) have achieved strong performance on OCR-based benchmarks and increasingly focused on text-rich understanding, but their

model-releasesarxiv-cs-cl
25 Jun 2026
← Previous
1…6061626364…1249
Next →