AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
48,975 results
Safety

Revisiting Autoregressive Models for Generative Image Classification

DGX agent

arXiv:2603.19122v2 Announce Type: replace Abstract: Class-conditional generative models have emerged as accurate and robust classifiers, with diffusion models demonstrating clear advantages over other

safetyarxiv-cs-cv
2 Jul 2026
Model Releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

SocialOmni: Benchmarking Audio-Visual Social Interactivity in Omni Models

DGX agent

arXiv:2603.16859v2 Announce Type: replace Abstract: Omni-modal large language models (OLMs) redefine human-machine interaction by natively integrating audio, vision, and text. However, existing OLM be

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

Testing Frontier Large Language Models' Physics Literacy in Parallel Physical Worlds

DGX agent

arXiv:2607.00276v1 Announce Type: cross Abstract: Current large-language-model (LLM) physics benchmarks are usually scored by answer accuracy, which cannot distinguish genuine reasoning from recall of

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

UniDrive-WM: Unified Understanding, Planning and Generation World Model for Autonomous Driving

DGX agent

arXiv:2601.04453v4 Announce Type: replace Abstract: World models have become central to autonomous driving, where accurate scene understanding and future prediction are crucial for safe control. Recen

model-releasesarxiv-cs-cv
2 Jul 2026
Research

Valdi: Value Diffusion World Models

DGX agent

arXiv:2607.00917v1 Announce Type: cross Abstract: World models can enable Model Predictive Control (MPC), but this requires dynamics prediction that is both fast enough for online use and expressive e

researcharxiv-cs-ai
2 Jul 2026
Research

AdaJEPA: An Adaptive Latent World Model

DGX agent

arXiv:2606.32026v1 Announce Type: cross Abstract: Latent world models enable planning from high-dimensional observations by predicting future states in a compact latent space. However, these models ar

researcharxiv-cs-ai
1 Jul 2026
Model Releases

Beyond Binary Instrument QA: Probing Instrument Grounding in Music Audio-Language Models

DGX agent

arXiv:2606.31338v1 Announce Type: cross Abstract: Recent music audio-language models achieve high accuracy on instrument question-answering benchmarks, but it remains unclear whether this reflects rob

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

CoLT: Teaching Multi-Modal Models to Think with Chain of Latent Thoughts

DGX agent

arXiv:2606.31986v1 Announce Type: new Abstract: Chain-of-thought (CoT) reasoning has enabled multi-modal large language models (MLLMs) to tackle complex visual reasoning tasks by generating explicit i

model-releasesarxiv-cs-cv
1 Jul 2026
Model Releases

The Bidirectional Process Reward Model

DGX agent

arXiv:2508.01682v3 Announce Type: replace Abstract: Process Reward Models (PRMs), which assign fine-grained scores to intermediate reasoning steps within a solution trajectory, have emerged as a promi

model-releasesarxiv-cs-cl
1 Jul 2026
Model Releases

AURORA: Asymmetry and Update-Induced Rotation for Robust Hallucination Detection in Large Language Models

DGX agent

arXiv:2606.29545v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities across a wide range of natural language processing tasks. However, their tendency

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

BrepLLM: Enabling Large Language Models to Understand Boundary Representations

DGX agent

arXiv:2512.16413v2 Announce Type: replace Abstract: Current token-sequence-based Large Language Models (LLMs) struggle to directly process 3D Boundary Representation (B-rep) models that contain comple

model-releasesarxiv-cs-cv
30 Jun 2026
Safety

ExploreVLA: Dense World Modeling and Exploration for End-to-End Autonomous Driving

DGX agent

arXiv:2604.02714v2 Announce Type: replace Abstract: End-to-end autonomous driving models based on Vision-Language-Action (VLA) architectures have shown promising results by learning driving policies t

safetyarxiv-cs-cv
30 Jun 2026
Model Releases

Fine-Tuning General-Purpose Large Language Models for Agricultural Applications:A Reproducible Framework and Evaluation Protocol Based on Qwen3-8B

DGX agent

arXiv:2606.28992v1 Announce Type: cross Abstract: General-purpose large language models (LLMs) have demonstrated strong abilities in opendomain question answering, information extraction, and text gen

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Morphing into Hybrid Attention Models

DGX agent

arXiv:2606.30562v1 Announce Type: new Abstract: Hybrid attention models improve long-context efficiency by retaining only a subset of full-attention layers and replacing the remaining layers with line

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

PlantExpertVQA: A Visual Question Answering Dataset for Benchmarking Vision-Language Models in Plant Science

DGX agent

arXiv:2508.17117v3 Announce Type: replace-cross Abstract: Existing plant-disease datasets target classification and detection, leaving vision-language models unable to support interactive, reasoning-b

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Qwen-RobotNav Technical Report: A Scalable Navigation Model Designed for an Agentic Navigation System

DGX agent

arXiv:2606.18112v3 Announce Type: replace-cross Abstract: Agentic navigation systems require a base navigation model whose observation strategy can be externally reconfigured at inference time, becaus

model-releasesarxiv-cs-cv
30 Jun 2026
Local Ai

The strength of clinical evidence is recoverable from language model representations but not from their stated grades

DGX agent

arXiv:2606.29034v1 Announce Type: cross Abstract: Large language models (LLMs) increasingly summarize clinical evidence, where a claim's weight depends on how strongly it is supported. Yet these model

local-aiarxiv-cs-ai
30 Jun 2026
Research

Towards Harnessing the Collaborative Power of Large and Small Models for Domain Tasks

DGX agent

arXiv:2504.17421v2 Announce Type: replace-cross Abstract: Large language models (LMs) offer broad generalization capabilities but require vast amounts of data and computational resources for domain-sp

researcharxiv-cs-ai
30 Jun 2026
Model Releases

Training Vision-Language-Action Models with Dense Embodied Chain-of-Thought Supervision

DGX agent

arXiv:2606.30552v1 Announce Type: cross Abstract: Cross-embodiment transfer in vision-language-action (VLA) models remains challenging because low-level state and action spaces differ fundamentally ac

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

Unified Enhancement of the Generalization and Robustness of Language Models via Bi-Stage Optimization

DGX agent

arXiv:2503.16550v2 Announce Type: replace Abstract: Neural network language models (LMs) are confronted with significant challenges in generalization and robustness. Currently, many studies focus on i

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

VTEdit-Bench: A Comprehensive Benchmark for Multi-Reference Image Editing Models in Virtual Try-On

DGX agent

arXiv:2603.11734v2 Announce Type: replace Abstract: As virtual try-on (VTON) continues to advance, a growing number of real-world scenarios have emerged, pushing beyond the ability of the existing spe

model-releasesarxiv-cs-cv
30 Jun 2026
Local Ai

Your Data Manifold is Secretly a Reward Model: Shell-LCC for Text-to-Video Generation

DGX agent

arXiv:2606.30248v1 Announce Type: new Abstract: Recent text-to-video (T2V) diffusion models rely heavily on auxiliary reward signals (e.g., via reward models or DPO) to align generated content with hu

local-aiarxiv-cs-cv
30 Jun 2026
Research

An Expectation-Maximization Algorithm for Training Clean Diffusion Models from Corrupted Observations

DGX agent

arXiv:2407.01014v2 Announce Type: replace Abstract: Diffusion models excel in solving imaging inverse problems due to their ability to model complex image priors. However, their reliance on large, cle

researcharxiv-cs-cv
29 Jun 2026
Model Releases

Benchmarking on Tasks That Matter: Dataset Selection for Preserving Model Rankings

DGX agent

arXiv:2606.27997v1 Announce Type: new Abstract: Benchmarks of machine learning models often include many datasets, making evaluation expensive. For efficiency, it is preferable to perform evaluations

model-releasesarxiv-cs-lg
29 Jun 2026
Hardware

DataStates-LLM: Scalable Checkpointing for Transformer Models Using Composable State Providers

DGX agent

arXiv:2601.16956v1 Announce Type: cross Abstract: The rapid growth of Large Transformer-based models, specifically Large Language Models (LLMs), now scaling to trillions of parameters, has necessitate

hardwarearxiv-cs-ai
29 Jun 2026
Tutorials

FlexMoE: One-for-All Nested Intra-Expert Pruning for MoE Language Models

DGX agent

arXiv:2606.27866v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) language models scale model ability with sparsely activated experts, making this architecture a standard recipe for modern larg

tutorialsarxiv-cs-lg
29 Jun 2026
Model Releases

GAIA: A Data Flywheel System for Training GUI Test-Time Scaling Critic Models

DGX agent

arXiv:2601.18197v2 Announce Type: replace Abstract: While Large Vision-Language Models (LVLMs) have significantly advanced GUI agents' capabilities in parsing textual instructions, interpreting screen

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

PEBS: Per-rater Empirical-Bayes Shrinkage for RLHF Reward-Model Calibration

DGX agent

arXiv:2606.27578v1 Announce Type: cross Abstract: Reward models for Reinforcement Learning from Human Feedback (RLHF) pool preferences across thousands of annotators and fit one global affine calibrat

model-releasesarxiv-cs-ai
29 Jun 2026
Applications

Quantum Generative Diffusion Model for Real-World Time Series

DGX agent

arXiv:2606.27561v1 Announce Type: new Abstract: Generative models have achieved remarkable success in data synthesis, though recent advances driven by increasing model scale have introduced challenges

applicationsarxiv-cs-lg
29 Jun 2026
Model Releases

RSICCLLM: A Multimodal Large Language Model for Remote Sensing Image Change Captioning

DGX agent

arXiv:2606.28266v1 Announce Type: new Abstract: Remote Sensing Image Change Captioning (RSICC) aims to describe changes between bi-temporal remote sensing images and holds significant research and app

model-releasesarxiv-cs-cv
29 Jun 2026
Model Releases

S^2-VLA: State-Space Guided Vision-Language-Action Models for Long-Horizon Manipulation

DGX agent

arXiv:2606.27872v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have demonstrated strong capabilities in robotic manipulation, but their performance degrades significantly in lon

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

6 Fingers, 1 Kidney: Natural Adversarial Medical Images Reveal Critical Weaknesses of Vision-Language Models

DGX agent

arXiv:2512.04238v2 Announce Type: replace Abstract: Vision-language models (VLMs) are increasingly integrated into clinical workflows. However, existing benchmarks primarily assess performance on comm

model-releasesarxiv-cs-cv
26 Jun 2026
Research

Consistent Distributed Ranking of Generative Models via Kernel Distances

DGX agent

arXiv:2310.11714v5 Announce Type: replace Abstract: Ranking generative models based on the fidelity and diversity of their outputs is required to identify the best generator in a group of candidate ge

researcharxiv-cs-lg
26 Jun 2026
Research

Dataset Usage Inference without Shadow Models or Held-out Data

DGX agent

arXiv:2606.26257v1 Announce Type: new Abstract: How much of my data was used to train a machine learning model? Dataset Usage Inference (DUI) aims to answer this by estimating what fraction of a datas

researcharxiv-cs-lg
26 Jun 2026
Model Releases

extsc{DiARC}: Distinguishing Positive and Negative Samples Helps Improving ARC-like Reasoning Ability of Large Language Models

DGX agent

arXiv:2606.26530v1 Announce Type: cross Abstract: The Abstraction and Reasoning Corpus (ARC;~itealp{chollet2019measure}) contains tasks that require summarizing patterns from limited grid samples and

model-releasesarxiv-cs-ai
26 Jun 2026
Safety

Normalizing Flows are Capable Models for Continuous Control

DGX agent

arXiv:2505.23527v4 Announce Type: replace Abstract: Modern reinforcement learning (RL) algorithms have found success by using powerful probabilistic models, such as transformers, energy-based models,

safetyarxiv-cs-lg
26 Jun 2026
Safety

Paying More Attention to Visual Tokens in Self-Evolving Large Multimodal Models

DGX agent

arXiv:2606.27373v1 Announce Type: new Abstract: Recently, self-evolving large multimodal models (LMMs) have received attention for improving visual reasoning in a purely unsupervised setting. However,

safetyarxiv-cs-cv
26 Jun 2026
Model Releases

Post-Training Recipe, More Than Model Family, Shapes Multi-Agent LLM Conversational Behavior

DGX agent

arXiv:2606.20632v2 Announce Type: replace-cross Abstract: Multi-LLM systems use multiple language models to deliberate, judge each other's outputs, or coordinate as agents. Their value depends on the

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Reproducibility Study of 'AlphaEdit: Null-Space Constrained Knowledge Editing for Language Models'

DGX agent

arXiv:2606.26783v1 Announce Type: cross Abstract: Fang et al. (2025) introduced a null-space constrained projection, named AlphaEdit, for locate-then-edit knowledge editing methods, theoretically guar

model-releasesarxiv-cs-cl
26 Jun 2026
Model Releases

Simulation-based inference for rapid Bayesian parameter estimation in epidemiological models: a comparison with MCMC

DGX agent

arXiv:2606.27286v1 Announce Type: new Abstract: Mechanistic epidemiological models are widely used to support infectious disease forecasting and public-health decision making. Bayesian calibration of

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Where Do Models Find Happiness? Emotion Vectors in Open-Source LLMs

DGX agent

arXiv:2606.26987v1 Announce Type: cross Abstract: Recent work identified emotion vectors in Claude Sonnet 4.5, which are internal representations that encode emotion concepts, causally influence behav

model-releasesarxiv-cs-ai
26 Jun 2026
Agents

A Probabilistic Framework for LLM-Based Model Discovery

DGX agent

arXiv:2602.18266v2 Announce Type: replace Abstract: Automated methods for discovering mechanistic simulator models from observational data offer a promising path toward accelerating scientific progres

agentsarxiv-cs-lg
25 Jun 2026
Model Releases

Benchmarking Deep Learning Models for Laryngeal Cancer Staging Using the LaryngealCT Dataset

DGX agent

arXiv:2510.11047v2 Announce Type: replace Abstract: Laryngeal cancer imaging research lacks standardised public datasets to enable reproducible deep learning (DL) model development. We present Larynge

model-releasesarxiv-cs-cv
25 Jun 2026
Research

Concept Removal for Frontier Image Generative Models

DGX agent

arXiv:2606.25548v1 Announce Type: new Abstract: Image generative models are trained on massive, largely uncurated internet-scale datasets that contain undesirable visual concepts. Efficiently removing

researcharxiv-cs-cv
25 Jun 2026
Model Releases

Detect, Unlearn, Restore: Defending Text Summarization Models Against Data Poisoning

DGX agent

arXiv:2606.26036v1 Announce Type: new Abstract: Training-time data poisoning during fine-tuning poses a significant threat to large language models (LLMs) deployed for abstractive text summarization,

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

How Robust is OCR-Reasoning? Evaluating OCR-Reasoning Robustness of Vision-Language Models under Visual Perturbations

DGX agent

arXiv:2606.26041v1 Announce Type: cross Abstract: Vision-language models (VLMs) have achieved strong performance on OCR-based benchmarks and increasingly focused on text-rich understanding, but their

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

In-Context World Modeling for Robotic Control

DGX agent

arXiv:2606.26025v1 Announce Type: cross Abstract: Modern Vision-Language-Action (VLA) models often fail to generalize to novel setups, such as altered camera viewpoints or robot morphologies, because

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

InvestPhilBench: A Multi-Layer Dynamic Benchmark for Evaluating Large Language Model Procedural Reasoning in Expert Investment Philosophy

DGX agent

arXiv:2606.25984v1 Announce Type: cross Abstract: Large language models are increasingly deployed as investment research assistants, yet no benchmark tests whether they can accurately reconstruct and

model-releasesarxiv-cs-lg
25 Jun 2026
← Previous
1…4546474849…1021
Next →