AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,376
  • Agents7,554
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,170
  • Local Ai4,930
  • Model Releases23,883
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,376
  • Agents7,554
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,170
  • Local Ai4,930
  • Model Releases23,883
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

88,376Total entries
1Added by human
88,375Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,602 results
19 May 2026

DepthPolyp: Pseudo-Depth Guided Lightweight Segmentation for Real-Time Colonoscopy

Model ReleasesDGX agent

arXiv:2605.16519v1 Announce Type: new Abstract: Accurate polyp segmentation in colonoscopy is essential for early colorectal cancer detection, yet real-world clinical environments pose persistent chal

Diffusion-Based sRGB Real Noise Generation via Prompt-Driven Noise Representation Learning

Model ReleasesDGX agent

arXiv:2603.04870v2 Announce Type: replace Abstract: Denoising in the sRGB image space is challenging due to large noise variability. Although end-to-end methods perform well, their effectiveness in re

EgoIntrospect: An Egocentric Dataset and Benchmark for User-Centric Internal State Reasoning

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.17262v1 Announce Type: new Abstract: Despite extensive efforts on egocentric video datasets and benchmarks, understanding users' internal states, which is crucial for enabling seamless AI a

EVA01: Unified Native 3D Understanding and Generation via Mixture-of-Transformers

ResearchDGX agent

arXiv:2605.16745v1 Announce Type: new Abstract: This paper addresses the challenge of integrating 3D meshes as a native modality within Multimodal Large Language Models (MLLMs). Diffusion-based large

Evidence-Grounded Frontier Mapping and Agentic Hypothesis Generation in Nanomedicine

Model ReleasesDGX agent

arXiv:2605.18144v1 Announce Type: new Abstract: Nanomedicine research spans delivery chemistry, immunology, imaging, biomaterials, and disease-specific translational science, yet its conceptual design

FinAuditing: A Financial Taxonomy-Structured Multi-Document Benchmark for Evaluating LLMs

Model ReleasesDGX agent

arXiv:2510.08886v3 Announce Type: replace Abstract: Going beyond simple text processing, financial auditing requires detecting semantic, structural, and numerical inconsistencies across large-scale di

FinTagging: Benchmarking LLMs for Extracting and Structuring Financial Information

Model ReleasesDGX agent

arXiv:2505.20650v5 Announce Type: replace-cross Abstract: Accurate interpretation of numerical data in financial reports is critical for markets and regulators. Although XBRL (eXtensible Business Repo

From Isolated Scoring to Collaborative Ranking: A Comparison-Native Framework for LLM-Based Paper Evaluation

ResearchDGX agent

arXiv:2603.17588v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are currently applied to scientific paper evaluation by assigning an absolute score to each paper independently.

GeoFlow: Enforcing Implicit Geometric Consistency in Video Generation

ResearchDGX agent

arXiv:2605.18365v1 Announce Type: new Abstract: Generating geometrically consistent videos remains an open challenge: text-to-video diffusion models trained on web-scale data treat geometry only impli

Geometric Scaling of Bayesian Inference in LLMs

Model ReleasesDGX agent

arXiv:2512.23752v5 Announce Type: replace-cross Abstract: Recent work has shown that small transformers trained in controlled 'wind-tunnel'' settings can implement exact Bayesian inference, and that t

Geometry-Aware Surrogate for Real-Time Hydrodynamics Estimation of Autonomous Ground Vehicles in Amphibious Environments

AgentsDGX agent

arXiv:2605.18543v1 Announce Type: new Abstract: Autonomous ground vehicles operating in shallow water or flood-prone terrains require dynamic models that account for hydrodynamic forces. However, the

Goal-Conditioned Supervised Learning for LLM Fine-Tuning

SafetyDGX agent

arXiv:2605.16345v1 Announce Type: cross Abstract: Large language models often require fine-tuning to better align their behavior with user intent at deployment. Existing approaches are commonly divide

GUIDE-VAE: Advancing Data Generation with User Information and Pattern Dictionaries

TutorialsDGX agent

arXiv:2411.03936v2 Announce Type: replace Abstract: Generative modelling of multi-user datasets has become prominent in science and engineering. Generating a data point for a given user requires emplo

I built a terminal tool (TUI) to make local LLMs debate each other and catch hallucinations (Ollama/Cloud).GitHub Debut

Local AiDGX agent

A developer created a terminal user interface (TUI) tool that enables local large language models (LLMs) running on Ollama or cloud platforms to debate each other as a method for detecting and reducin

Incentive-Aware Federated Averaging with Performance Guarantees under Strategic Participation

Local AiDGX agent

arXiv:2603.20873v2 Announce Type: replace Abstract: Federated learning (FL) is a communication-efficient collaborative learning framework that enables model training across multiple agents with privat

Leveraging Unsupervised Learning for Cost-Effective Visual Anomaly Detection

ResearchDGX agent

arXiv:2409.15980v2 Announce Type: replace-cross Abstract: Traditional machine learning-based visual inspection systems require extensive data collection and repetitive model training to improve accura

Meta-Learning Guided Pruning for Few-Shot Plant Pathology on Edge Devices

TutorialsDGX agent

arXiv:2601.02353v3 Announce Type: replace Abstract: Farmers in remote areas need quick and reliable methods for identifying plant diseases, yet they often lack access to laboratories or high-performan

Modality vs. Morphology: A Framework for Time Series Classification for Biological Signals

ResearchDGX agent

arXiv:2605.18483v1 Announce Type: cross Abstract: Time series classification (TSC) of biological signals has progressed from handcrafted, modality-specific approaches to deep architectures capable of

Needles in the Landscape: Semi-Supervised Pseudolabeling for Archaeological Site Discovery under Label Scarcity

ResearchDGX agent

arXiv:2510.16814v2 Announce Type: replace-cross Abstract: Archaeological predictive modelling estimates where undiscovered sites are likely to occur by combining known locations with environmental, cu

Network Knowledge Prior Guided Learning for Data-Efficient Surface Defect Detection

ApplicationsDGX agent

arXiv:2605.17780v1 Announce Type: new Abstract: Deep learning-based methods have become the de facto standard for industrial defect detection. However, their data-hungry nature and inherent 'black-box

OmniVL-Guard Pro: A Tool-Augmented Agent for Omnibus Vision-Language Forensics

Model ReleasesDGX agent

arXiv:2605.16962v1 Announce Type: cross Abstract: Existing vision-language forgery detection and grounding methods operate under a closed-world paradigm, assuming verification can be completed by the

On Improving Multimodal Pedestrian Trajectory Prediction with CVAE: A Study on Benchmark and Robot Data

Model ReleasesDGX agent

arXiv:2605.18262v1 Announce Type: new Abstract: Accurate pedestrian trajectory prediction is crucial for autonomous systems operating in complex environments, such as modular buses and delivery robots

ORACLE: Anticipating Scams from Partial Trajectories in Streaming App Usage

Model ReleasesDGX agent

arXiv:2605.16363v1 Announce Type: new Abstract: Smartphone scams are increasingly prevalent and typically manifest as multi-stage, cross-application processes with gradually emerging intent. Effective

PAREDA: A Multi-Accent Speech Dataset of Natural Language Processing Research Discussions

Model ReleasesDGX agent

arXiv:2605.17860v1 Announce Type: cross Abstract: While modern Automatic Speech Recognition (ASR) systems achieve high accuracy on benchmark corpora, their performance often degrades when there is rea

Prior Knowledge Makes It Possible: From Sublinear Graph Algorithms to LLM Test-Time Methods

AgentsDGX agent

arXiv:2510.16609v3 Announce Type: replace-cross Abstract: Test-time augmentation, such as Retrieval-Augmented Generation (RAG) or tool use, critically depends on an interplay between a model's paramet

Probing for Representation Manifolds in Superposition

Model ReleasesDGX agent

arXiv:2605.18537v1 Announce Type: cross Abstract: This paper introduces the Manifold Probe, a supervised method for discovering representation manifolds in superposition. The method generalizes linear

Proof-Carrying Certificates for LLM Pipelines: A Trust-Boundary Architecture

AgentsDGX agent

arXiv:2605.16407v1 Announce Type: cross Abstract: We present a framework for verifying the deterministic structured computations surrounding a large language model rather than the model itself, extend

Protein Fold Classification at Scale: Benchmarking and Pretraining

Model ReleasesDGX agent

arXiv:2605.18552v1 Announce Type: new Abstract: Classifying protein topology is essential for deciphering biological function, but progress is held back by the lack of large-scale benchmarks that avoi

QLIF-CAST: Quantum Leaky-Integrate-and-Fire for Time-Series Weather Forecasting

Model ReleasesDGX agent

arXiv:2605.18333v1 Announce Type: cross Abstract: Accurate and efficient time-series forecasting remains a challenging problem for both classical and quantum neural architectures, particularly in mult

Rethinking Point Clouds as Sequences: A Causal Next-Token Predictive Learning Framework

Local AiDGX agent

arXiv:2605.17566v1 Announce Type: new Abstract: With the rapid progress of multimodal foundation models and predictive pre-training, an important open question is how to equip 3D point clouds with a p

Revisiting the Adam-SGD Gap in LLM Pre-Training: The Role of Large Effective Learning Rates

Model ReleasesDGX agent

arXiv:2605.17787v1 Announce Type: new Abstract: It is widely believed that stochastic gradient descent (SGD) performs significantly worse than adaptive optimizers such as Adam in pre-training Large La

RoboMME: Benchmarking and Understanding Memory for Robotic Generalist Policies

Model ReleasesDGX agent

arXiv:2603.04639v2 Announce Type: replace-cross Abstract: Memory is critical for long-horizon and history-dependent robotic manipulation. Such tasks often involve counting repeated actions or manipula

SafeDiffusion-R1: Online Reward Steering for Safe Diffusion Post-Training

SafetyDGX agent

arXiv:2605.18719v1 Announce Type: new Abstract: Diffusion models have been widely studied for removing unsafe content learned during pre-training. Existing methods require expensive supervised data, e

SAMRI: Segment Any MRI

Local AiDGX agent

arXiv:2510.26635v3 Announce Type: replace-cross Abstract: Summary: SAMRI is an MRI-specialized adaptation of the Segment Anything Model achieving superior whole-body MRI segmentation, particularly for

SENSE: Satellite-based ENergy Synthesis for Sustainable Environment

ResearchDGX agent

arXiv:2605.18101v1 Announce Type: cross Abstract: Urban Building Energy Modeling plays a critical role in achieving the United Nations' Sustainable Development Goals 7 and 11. Although existing studie

SocialMemBench: Are AI Memory Systems Ready for Social Group Settings?

Model ReleasesDGX agent

arXiv:2605.17789v1 Announce Type: cross Abstract: Memory systems for AI assistants were built for single-user dialogue and fail characteristically when applied to multi-party social group settings. Th

Sparse-to-Dense: A Free Lunch for Lossless Acceleration of Video Understanding in LLMs

ResearchDGX agent

arXiv:2505.19155v2 Announce Type: replace-cross Abstract: Due to the auto-regressive nature of current video large language models (Video-LLMs), the inference latency increases as the input sequence l

Spectral Progressive Diffusion for Efficient Image and Video Generation

ResearchDGX agent

arXiv:2605.18736v1 Announce Type: new Abstract: Diffusion models have been shown to implicitly generate visual content autoregressively in the frequency domain, where low-frequency components are gene

SRC-Flow: Compact Semantic Representations Enable Normalizing Flows for Image Generation

TutorialsDGX agent

arXiv:2605.18267v1 Announce Type: new Abstract: Normalizing flows (NFs) provide exact likelihoods and deterministic invertible sampling, but have historically lagged behind diffusion models for large-

SteadyDancer: Harmonized and Coherent Human Image Animation with First-Frame Preservation

Model ReleasesDGX agent

arXiv:2511.19320v2 Announce Type: replace Abstract: Preserving first-frame identity while ensuring precise motion control is a fundamental challenge in human image animation. The Image-to-Motion Bindi

Supervise Less, See More: Training-free Nuclear Instance Segmentation with Prototype-Guided Prompting

Model ReleasesDGX agent

arXiv:2511.19953v2 Announce Type: replace Abstract: Accurate nuclear instance segmentation is a pivotal task in computational pathology, supporting data-driven clinical insights and facilitating downs

TailedTS: Benchmark Dataset for Heavy-Tailed Time Series Prediction and Periodicity Quantification

Model ReleasesDGX agent

arXiv:2605.16361v1 Announce Type: cross Abstract: We present TailedTS, a large-scale benchmark dataset derived from Wikipedia hourly page view observations throughout 2024, specifically designed to te

The Diffusion Duality, Chapter II: Psi-Samplers

TutorialsDGX agent

arXiv:2602.21185v2 Announce Type: replace Abstract: Uniform-state discrete diffusion models excel at few-step generation and guidance due to their ability to self-correct, making them preferred over a

The Expressive Power of Low Precision Softmax Transformers with (Summarized) Chain-of-Thought

Model ReleasesDGX agent

arXiv:2605.18079v1 Announce Type: cross Abstract: Existing expressivity results for transformers typically rely on hardmax attention, high precision, and other architectural modifications that disconn

The Neural Tangent Kernel for Classification

Model ReleasesDGX agent

arXiv:2605.17606v1 Announce Type: new Abstract: In wide neural networks, the Neural Tangent Kernel (NTK) remains approximately constant during training, providing a powerful theoretical tool for study

The Silent Brush: Evaluating Artistic Style Leakage in AI Art Generation

TutorialsDGX agent

arXiv:2605.17500v1 Announce Type: cross Abstract: Generative text-to-image models are typically trained on large-scale web-scraped datasets that include diverse visual content such as copyrighted and

TOBench: A Task-Oriented Omni-Modal Benchmark for Real-World Tool-Using Agents

Model ReleasesDGX agent

arXiv:2605.16909v1 Announce Type: new Abstract: Tool-using agents are increasingly expected to operate across realistic professional workflows, where they must interpret multimodal inputs, coordinate

Today we release Contrastive Neuron Attribution (CNA), a method for steering LLM behavior by identifying and ablating sparse circuits in the…

Model ReleasesDGX agent

Today we release Contrastive Neuron Attribution (CNA), a method for steering LLM behavior by identifying and ablating sparse circuits in the MLP basis without training a sparse autoencoder, modifying

Towards Human-Level Book-Writing Capability

AgentsDGX agent

arXiv:2605.17064v1 Announce Type: new Abstract: Large language models optimized for instruction following and agentic tasks remain poorly aligned with the requirements of high-quality creative writing

TTE-Flash: Accelerating Reasoning-based Multimodal Representations via Think-Then-Embed Tokens

Model ReleasesDGX agent

arXiv:2605.16638v1 Announce Type: new Abstract: Recent research has demonstrated that Universal Multimodal Embedding (UME) benefits significantly from Chain-of-Thought (CoT) reasoning. In this paradig

UAVFF3D: A Geometry-Aware Benchmark for Feed-Forward UAV 3D Reconstruction

Model ReleasesDGX agent

arXiv:2605.17942v1 Announce Type: new Abstract: Feed-forward 3D reconstruction has recently demonstrated strong generalization across diverse scenes, yet its performance in UAV imagery remains underex

Unveiling Memorization-Generalization Coexistence: A Case Study on Arithmetic Tasks with Label Noise

ApplicationsDGX agent

arXiv:2605.18022v1 Announce Type: cross Abstract: Highly over-parameterized models can simultaneously memorize noisy labels and generalize well, yet how these behaviors coexist remains poorly understo

WavFlow: Audio Generation in Waveform Space

Model ReleasesDGX agent

arXiv:2605.18749v1 Announce Type: cross Abstract: Modern audio generation predominantly relies on latent-space compression, introducing additional complexity and potential information loss. In this wo

WEBSERV: A Full-Stack and RL-Ready Web Environment for Training Web Agents at Scale

Model ReleasesDGX agent

arXiv:2510.16252v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) for web agents demands environments that are both effective for evaluation and efficient enough for large-scale on

We’re releasing Nemotron-Labs-Diffusion - the first Tri-mode LM family (3B/8B/14B) that switches between 1⃣Autoregressive, 2⃣Diffusion, and …

Model ReleasesDGX agent

We’re releasing Nemotron-Labs-Diffusion - the first Tri-mode LM family (3B/8B/14B) that switches between 1⃣Autoregressive, 2⃣Diffusion, and 3⃣Self-Speculation decoding by simply changing the attention

White-Box Sensitivity Auditing with Steering Vectors

SafetyDGX agent

arXiv:2601.16398v2 Announce Type: replace-cross Abstract: Algorithmic audits are essential tools for examining systems for properties required by regulators or desired by operators. Current audits of

Zero-Shot Faithful Textual Explanations via Directional-Derivative Influence on Predictions

Model ReleasesDGX agent

arXiv:2605.16877v1 Announce Type: new Abstract: Zero-shot textual explanations aim to make image classifiers more transparent by probing their internal representations, without relying on task-specifi

18 May 2026

AirNav: A Large-Scale UAV Vision-and-Language Navigation Dataset with Natural and Diverse Instructions

Model ReleasesDGX agent

arXiv:2601.03707v2 Announce Type: replace Abstract: Existing UAV vision-and-language navigation (VLN) benchmarks rarely provide realistic aerial scenes, natural process-level instructions, and suffici

AOT-POT: Adaptive Operator Transformation for Large-Scale PDE Pre-training

ResearchDGX agent

arXiv:2605.15793v1 Announce Type: new Abstract: Pre-training neural operators on diverse partial differential equation (PDE) datasets has emerged as a promising direction for building general-purpose

Approximate and Weighted Data Reconstruction Attack in Federated Learning

Local AiDGX agent

arXiv:2308.06822v3 Announce Type: replace-cross Abstract: Federated Learning (FL) is a distributed learning paradigm that enables multiple clients to collaborate on building a machine learning model w

← Previous
1…472473474475476…1061
Next →