AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,618 results
Tutorials

Min Generalized Sliced Gromov Wasserstein: A Scalable Path to Gromov Wasserstein

DGX agent

arXiv:2605.13753v1 Announce Type: cross Abstract: We propose min Generalized Sliced Gromov--Wasserstein (min-GSGW), a sliced formulation for the Gromov--Wasserstein (GW) problem using expressive gener

tutorialsarxiv-cs-cv
14 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

MindVLA-U1: VLA Beats VA with Unified Streaming Architecture for Autonomous Driving

DGX agent

arXiv:2605.12624v1 Announce Type: cross Abstract: Autonomous driving has progressed from modular pipelines toward end-to-end unification, and Vision-Language-Action (VLA) models are a natural extensio

model-releasesarxiv-cs-cv
14 May 2026
Tutorials

Multi-Modal Guided Multi-Source Domain Adaptation for Object Detection

DGX agent

arXiv:2605.13140v1 Announce Type: new Abstract: General object detection (OD) struggles to detect objects in the target domain that differ from the training distribution. To address this, recent studi

tutorialsarxiv-cs-cv
14 May 2026
Research

Neural Surrogate Forward Modelling For Electrocardiology Without Explicit Intracellular Conductivity Tensor

DGX agent

arXiv:2605.13366v1 Announce Type: new Abstract: Accurate forward modelling is essential for non-invasive cardiac electrophysiology, particularly in atrial fibrillation, where electrical activation is

researcharxiv-cs-cv
14 May 2026
Research

Neural Video Compression with Domain Transfer

DGX agent

arXiv:2605.13476v1 Announce Type: new Abstract: Content-adaptive compression has always been a key direction in neural video coding (NVC), aiming to mitigate the domain gap between training and testin

researcharxiv-cs-cv
14 May 2026
Model Releases

No One Knows the State of the Art in Geospatial Foundation Models

DGX agent

arXiv:2605.12678v1 Announce Type: new Abstract: Geospatial foundation models (GFMs) have been proposed as generalizable backbones for disaster response, land-cover mapping, food-security monitoring, a

model-releasesarxiv-cs-cv
14 May 2026
Research

OCH3R: Object-Centric Holistic 3D Reconstruction

DGX agent

arXiv:2605.13018v1 Announce Type: new Abstract: Object-centric scene understanding is a fundamental challenge in computer vision. Existing approaches often rely on multi-stage pipelines that first app

researcharxiv-cs-cv
14 May 2026
Model Releases

OmniLiDAR: A Unified Diffusion Framework for Multi-Domain 3D LiDAR Generation

DGX agent

arXiv:2605.13815v1 Announce Type: new Abstract: LiDAR scene generation is increasingly important for scalable simulation and synthetic data creation, especially under diverse sensing conditions that a

model-releasesarxiv-cs-cv
14 May 2026
Research

On Hallucinations in Inverse Problems: Fundamental Limits and Provable Assessment Methods

DGX agent

arXiv:2605.13146v1 Announce Type: cross Abstract: Artificial intelligence (AI) has transformed imaging inverse problems, from medical diagnostics to Earth observation. Yet deep neural networks can pro

researcharxiv-cs-cv
14 May 2026
Hardware

OP4KSR: One-Step Patch-Free 4K Super-Resolution with Periodic Artifact Suppression

DGX agent

arXiv:2605.13457v1 Announce Type: new Abstract: Diffusion-based real-world image super-resolution (Real-ISR) has achieved remarkable perceptual quality; however, directly super-resolving images to 4K

hardwarearxiv-cs-cv
14 May 2026
Research

Optimization in Sparse 2D to Dense 3D Weakly Supervised Learning: Application to Multi-Label Segmentation of Large ex vivo MRI Data

DGX agent

arXiv:2605.12753v1 Announce Type: cross Abstract: INTRODUCTION | Fully supervised 3D segmentation of high-resolution ex vivo MRI is limited by the prohibitive cost of volumetric annotation, forcing re

researcharxiv-cs-cv
14 May 2026
Safety

Pareto-Guided Optimal Transport for Multi-Reward Alignment

DGX agent

arXiv:2605.13155v1 Announce Type: new Abstract: Text-to-image generation models have achieved remarkable progress in preference optimization, yet achieving robust alignment across diverse reward model

safetyarxiv-cs-cv
14 May 2026
Model Releases

Pattern-Enhanced RT-DETR for Multi-Class Battery Detection

DGX agent

arXiv:2605.13670v1 Announce Type: new Abstract: Accurate and efficient battery detection is increasingly important for applications in electronic waste recycling, industrial quality control, and autom

model-releasesarxiv-cs-cv
14 May 2026
Safety

Perception with Guarantees: Certified Pose Estimation via Reachability Analysis

DGX agent

arXiv:2602.10032v2 Announce Type: replace Abstract: Agents in cyber-physical systems are increasingly entrusted with safety-critical tasks. Ensuring safety of these agents often requires localizing th

safetyarxiv-cs-cv
14 May 2026
Research

Phy-CoSF: Physics-Guided Continuous Spectral Fields Reconstruction and Super-Resolution for Snapshot Compressive Imaging

DGX agent

arXiv:2605.13583v1 Announce Type: new Abstract: Recent advances have demonstrated that coded aperture snapshot spectral imaging (CASSI) systems show great potential for capturing 3D hyperspectral imag

researcharxiv-cs-cv
14 May 2026
Model Releases

PhysEditBench: A Protocol-Conditioned Benchmark for Dense Physical-Map Prediction with Image Editors

DGX agent

arXiv:2605.13493v1 Announce Type: new Abstract: Can general-purpose image editors predict physical maps from a single RGB image? General-purpose image editors differ from standard task-specific dense-

model-releasesarxiv-cs-cv
14 May 2026
Safety

PRA-PoE: Robust Alzheimer's Diagnosis with Arbitrary Missing Modalities

DGX agent

arXiv:2605.13081v1 Announce Type: new Abstract: Missing modalities are prevalent in real-world Alzheimer's disease (AD) assessment and pose a significant challenge to multimodal learning, particularly

safetyarxiv-cs-cv
14 May 2026
Research

Prediction of Rectal Cancer Regrowth from Longitudinal Endoscopy

DGX agent

arXiv:2605.12855v1 Announce Type: new Abstract: Clinical trial studies indicate benefit of watch-and-wait (WW) surveillance for patients with rectal cancer showing a complete or near clinical response

researcharxiv-cs-cv
14 May 2026
Model Releases

PreFIQs: Face Image Quality Is What Survives Pruning

DGX agent

arXiv:2605.13396v1 Announce Type: new Abstract: Face Image Quality Assessment (FIQA) evaluates the utility of a face image for automated face recognition (FR) systems. In this work, we propose PreFIQs

model-releasesarxiv-cs-cv
14 May 2026
Local Ai

PRISM: Prior Rectification and Uncertainty-Aware Structure Modeling for Diffusion-Based Text Image Super-Resolution

DGX agent

arXiv:2605.13027v1 Announce Type: new Abstract: Text image super-resolution (Text-SR) requires more than visually plausible detail synthesis: slight errors in stroke topology may alter character ident

local-aiarxiv-cs-cv
14 May 2026
Safety

Pyramid Forcing: Head-Aware Pyramid KV Cache Policy for High-Quality Long Video Generation

DGX agent

arXiv:2605.13111v1 Announce Type: new Abstract: Autoregressive video generation enables streaming and open-ended long video synthesis, but still suffers from long-term degradation caused by accumulate

safetyarxiv-cs-cv
14 May 2026
Research

QLAM: A Quantum Long-Attention Memory Approach to Long-Sequence Token Modeling

DGX agent

arXiv:2605.13833v1 Announce Type: cross Abstract: Modeling long-range dependencies in sequential data remains a central challenge in machine learning. Transformers address this challenge through atten

researcharxiv-cs-cv
14 May 2026
Model Releases

Qwen-Image-VAE-2.0 Technical Report

DGX agent

arXiv:2605.13565v1 Announce Type: new Abstract: We present Qwen-Image-VAE-2.0, a suite of high-compression Variational Autoencoders (VAEs) that achieve significant advances in both reconstruction fide

model-releasesarxiv-cs-cv
14 May 2026
Safety

R-DMesh: Video-Guided 3D Animation via Rectified Dynamic Mesh Flow

DGX agent

arXiv:2605.13838v1 Announce Type: new Abstract: Video-guided 3D animation holds immense potential for content creation, offering intuitive and precise control over dynamic assets. However, practical d

safetyarxiv-cs-cv
14 May 2026
Safety

Real2Sim: A Physics-driven and Editable Gaussian Splatting Framework for Autonomous Driving Scenes

DGX agent

arXiv:2605.13591v1 Announce Type: new Abstract: Reliable autonomous driving relies on large-scale, well-labeled data and robust models. However, manual data collection is resource-intensive, and tradi

safetyarxiv-cs-cv
14 May 2026
Applications

Realtime-VLA FLASH: Speculative Inference Framework for Diffusion-based VLAs

DGX agent

arXiv:2605.13778v1 Announce Type: cross Abstract: Diffusion-based vision-language-action models (dVLAs) are promising for embodied intelligence but are fundamentally limited in real-time deployment by

applicationsarxiv-cs-cv
14 May 2026
Model Releases

Reasoning to Edit: Hypothetical Instruction-Based Image Editing with Visual Reasoning

DGX agent

arXiv:2507.01908v3 Announce Type: replace Abstract: Instruction-based image editing (IIE) has advanced rapidly with the success of diffusion models. However, existing efforts primarily focus on simple

model-releasesarxiv-cs-cv
14 May 2026
Model Releases

Reducing Bias and Variance: Generative Semantic Guidance and Bi-Layer Ensemble for Image Clustering

DGX agent

arXiv:2605.12961v1 Announce Type: new Abstract: Image clustering aims to partition unlabeled image datasets into distinct groups. A core aspect of this task is constructing and leveraging prior knowle

model-releasesarxiv-cs-cv
14 May 2026
Model Releases

Rethinking Graph Convolution for 2D-to-3D Hand Pose Lifting

DGX agent

arXiv:2605.13604v1 Announce Type: new Abstract: Graph convolutional networks (GCNs) are widely used for 3D hand pose estimation, where the hand skeleton is encoded as a fixed adjacency graph. We revis

model-releasesarxiv-cs-cv
14 May 2026
Research

Rigel3D: Rig-aware Latents for Animation-Ready 3D Asset Generation

DGX agent

arXiv:2605.13129v1 Announce Type: cross Abstract: Recent 3D generative models can synthesize high-quality assets, but their outputs are typically static: they lack the skeletal rigs, joint hierarchies

researcharxiv-cs-cv
14 May 2026
Safety

RoboEvolve: Co-Evolving Planner-Simulator for Robotic Manipulation with Limited Data

DGX agent

arXiv:2605.13775v1 Announce Type: cross Abstract: The scalability of robotic manipulation is fundamentally bottlenecked by the scarcity of task-aligned physical interaction data. While vision-language

safetyarxiv-cs-cv
14 May 2026
Model Releases

RoSplat: Robust Feed-Forward Pixel-wise Gaussian Splatting for Varying Input Views and High-Resolution Rendering

DGX agent

arXiv:2605.13093v1 Announce Type: new Abstract: Generalizable 3D Gaussian Splatting has recently emerged as an efficient approach for novel-view synthesis, enabling feed-forward synthesis from only a

model-releasesarxiv-cs-cv
14 May 2026
Applications

RotVLA: Rotational Latent Action for Vision-Language-Action Model

DGX agent

arXiv:2605.13403v1 Announce Type: cross Abstract: Latent Action Models (LAMs) have emerged as an effective paradigm for handling heterogeneous datasets during Vision-Language-Action (VLA) model pretra

applicationsarxiv-cs-cv
14 May 2026
Research

SceneGraphVLM: Dynamic Scene Graph Generation from Video with Vision-Language Models

DGX agent

arXiv:2605.13667v1 Announce Type: new Abstract: Scene graph generation provides a compact structured representation for visual perception, but accurate and fast graph prediction from images and videos

researcharxiv-cs-cv
14 May 2026
Research

sketch2symm: Symmetry-aware sketch-to-shape generation via semantic bridging

DGX agent

arXiv:2510.11303v2 Announce Type: replace Abstract: Sketch-based 3D reconstruction remains a challenging task due to the abstract and sparse nature of sketch inputs, which often lack sufficient semant

researcharxiv-cs-cv
14 May 2026
Research

Skill-Aligned Annotation for Reliable Evaluation in Text-to-Image Generation

DGX agent

arXiv:2605.13223v1 Announce Type: new Abstract: Text-to-image (T2I) generation has advanced rapidly, making reliable evaluation critical as performance differences between models narrow. Existing eval

researcharxiv-cs-cv
14 May 2026
Model Releases

SkySplat: Generalizable 3D Gaussian Splatting from Multi-Temporal Sparse Satellite Images

DGX agent

arXiv:2508.09479v2 Announce Type: replace Abstract: Three-dimensional scene reconstruction from sparse-view satellite images is a long-standing and challenging task. While 3D Gaussian Splatting (3DGS)

model-releasesarxiv-cs-cv
14 May 2026
Tutorials

Sparse Code Uplifting for Efficient 3D Language Gaussian Splatting

DGX agent

arXiv:2605.13600v1 Announce Type: new Abstract: 3D Language Gaussian Splatting (3DLGS) augments 3D Gaussian Splatting with language-aligned visual features for open-vocabulary 3D scene understanding.

tutorialsarxiv-cs-cv
14 May 2026
Model Releases

SpatialReward: Bridging the Perception Gap in Online RL for Image Editing via Explicit Spatial Reasoning

DGX agent

arXiv:2602.07458v4 Announce Type: replace Abstract: Online Reinforcement Learning (RL) offers a promising avenue for complex image editing but is currently constrained by the scarcity of reliable and

model-releasesarxiv-cs-cv
14 May 2026
Model Releases

SpurAudio: A Benchmark for Studying Shortcut Learning in Few-Shot Audio Classification

DGX agent

arXiv:2605.13672v1 Announce Type: new Abstract: Few-shot classification (FSC) is widely used for learning from limited labeled data, yet most evaluations implicitly assume that target concepts are ind

model-releasesarxiv-cs-cv
14 May 2026
Agents

Still Camouflage, Moving Illusion: View-Induced Trajectory Manipulation in Autonomous Driving

DGX agent

arXiv:2605.12743v1 Announce Type: cross Abstract: Existing physical adversarial attacks on vision-based autonomous driving induce time-evolving perception errors, including biased object tracking or t

agentsarxiv-cs-cv
14 May 2026
Applications

STORM: Segment, Track, and Object Re-Localization from a Single Image

DGX agent

arXiv:2511.09771v3 Announce Type: replace Abstract: Accurate 6D pose estimation and tracking are core capabilities for physical AI systems, yet real-world deployment remains brittle and labor-intensiv

applicationsarxiv-cs-cv
14 May 2026
Safety

Structural Diversity Drives Disruptive Scientific Innovation

DGX agent

arXiv:2605.12514v1 Announce Type: cross Abstract: Scientific innovation increasingly depends on collaboration, yet the organizational structure that fosters breakthrough ideas remains poorly understoo

safetyarxiv-cs-cv
14 May 2026
Research

SubspaceAD: Training-Free Few-Shot Anomaly Detection via Subspace Modeling

DGX agent

arXiv:2602.23013v3 Announce Type: replace Abstract: Detecting visual anomalies in industrial inspection often requires training with only a few normal images per category. Recent few-shot methods achi

researcharxiv-cs-cv
14 May 2026
Research

SymbolSight: Minimizing Inter-Symbol Interference for Reading with Prosthetic Vision

DGX agent

arXiv:2601.17326v2 Announce Type: replace Abstract: Retinal prostheses restore limited visual perception, but low spatial resolution and temporal persistence make reading difficult. In sequential lett

researcharxiv-cs-cv
14 May 2026
Applications

Taming the Long Tail: Rebalancing Adversarial Training via Adaptive Perturbation

DGX agent

arXiv:2605.13395v1 Announce Type: cross Abstract: Deep neural networks are highly vulnerable to adversarial examples, i.e.,small perturbations that can significantly degrade model performance. While a

applicationsarxiv-cs-cv
14 May 2026
Safety

Test-time Sparsity for Extreme Fast Action Diffusion

DGX agent

arXiv:2605.13316v1 Announce Type: new Abstract: Action diffusion excels at high-fidelity action generation but incurs heavy computational costs owing to its iterative denoising nature. Despite current

safetyarxiv-cs-cv
14 May 2026
Applications

The Joint Gromov Wasserstein Objective for Multiple Object Matching

DGX agent

arXiv:2511.16868v2 Announce Type: replace Abstract: The Gromov-Wasserstein (GW) distance serves as a powerful tool for matching objects in metric spaces. However, its traditional formulation is constr

applicationsarxiv-cs-cv
14 May 2026
← Previous
1…173174175176177…263
Next →