AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,618 results
Safety

Face inpainting with Identity Preserving Latent Diffusion Models

DGX agent

arXiv:2605.16696v1 Announce Type: new Abstract: Face inpainting techniques recover missing or occluded facial regions in a visually realistic manner, but preserving the identity in the final output re

safetyarxiv-cs-cv
19 May 2026
Research
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Factorized Latent Dynamics for Video JEPA: An Empirical Study of Auxiliary Objectives

DGX agent

arXiv:2605.17165v1 Announce Type: new Abstract: Joint-Embedding Predictive Architectures (JEPA) are a promising framework for self-supervised video representation learning, yet the behavior of auxilia

researcharxiv-cs-cv
19 May 2026
Model Releases

Fast Kernel-Space Diffusion for Remote Sensing Pansharpening

DGX agent

arXiv:2505.18991v3 Announce Type: replace Abstract: Pansharpening seeks to fuse high-resolution panchromatic (PAN) and low-resolution multispectral (LRMS) images into a single image with both fine spa

model-releasesarxiv-cs-cv
19 May 2026
Research

FG-TreeSeg: Flow-Guided Tree Crown Segmentation without Instance Annotations

DGX agent

arXiv:2602.00470v2 Announce Type: replace Abstract: Individual tree crown segmentation is an important task in remote sensing for forest biomass estimation and ecological monitoring. However, accurate

researcharxiv-cs-cv
19 May 2026
Safety

Flow Matching with Optimized Subclass Priors for Medical Image Augmentation

DGX agent

arXiv:2605.16469v1 Announce Type: cross Abstract: Rare diseases dominate the diagnostic challenge in medical imaging yet are severely underrepresented in clinical datasets, causing classifiers to fail

safetyarxiv-cs-cv
19 May 2026
Research

Forget-It-All: Multi-Concept Machine Unlearning via Concept-Aware Neuron Masking

DGX agent

arXiv:2601.06163v2 Announce Type: replace Abstract: The widespread adoption of text-to-image (T2I) diffusion models has raised concerns about their potential to generate copyrighted, inappropriate, or

researcharxiv-cs-cv
19 May 2026
Safety

Forget Many, Forget Right: Scalable and Precise Concept Unlearning in Diffusion Models

DGX agent

arXiv:2601.06162v4 Announce Type: replace-cross Abstract: Text-to-image diffusion models have achieved remarkable progress, yet their use raises copyright and misuse concerns, prompting research into

safetyarxiv-cs-cv
19 May 2026
Local Ai

FrequencyBooster: Full-Frequency Modeling for High-Fidelity Pixel Diffusion

DGX agent

arXiv:2605.17759v1 Announce Type: new Abstract: To circumvent the inherent fidelity bottlenecks and optimization misalignment of VAE-based latent diffusion, pixel-space diffusion models have emerged a

local-aiarxiv-cs-cv
19 May 2026
Model Releases

From Pixels to Places: A Systematic Benchmark for Evaluating Image Geolocalization Ability in Large Language Models

DGX agent

arXiv:2508.01608v2 Announce Type: replace Abstract: Image geolocalization, the task of identifying the geographic location depicted in an image, is important for applications in crisis response, digit

model-releasesarxiv-cs-cv
19 May 2026
Research

Functionalization via Structure Completion and Motion Rectification

DGX agent

arXiv:2605.18010v1 Announce Type: new Abstract: Acquisition and creation of 3D assets have been largely view- or appearance-driven. As a result, existing digital 3D models often lack the requisite str

researcharxiv-cs-cv
19 May 2026
Safety

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation

DGX agent

arXiv:2512.23180v3 Announce Type: replace Abstract: Driving World Models (DWMs) have been developing rapidly with the advances of generative models. However, existing DWMs lack 3D scene understanding

safetyarxiv-cs-cv
19 May 2026
Research

GaussianZoom: Progressive Zoom-in Generative 3D Gaussian Splatting with Geometric and Semantic Guidance

DGX agent

arXiv:2605.18252v1 Announce Type: new Abstract: We introduce GaussianZoom, a generative zoom-in 3D reconstruction system with an iterative progressive framework that combines geometry-consistent scene

researcharxiv-cs-cv
19 May 2026
Agents

GEM: Gaussian Evolution Model for Occupancy Forecasting and Motion Planning

DGX agent

arXiv:2605.17682v1 Announce Type: new Abstract: Future 3D semantic occupancy forecasting and motion planning are central to autonomous driving, as they require models to reason about how surrounding s

agentsarxiv-cs-cv
19 May 2026
Research

Generalize cross-ratios in n-dimensional Plane-Based Geometric Algebra

DGX agent

arXiv:2605.18398v1 Announce Type: cross Abstract: We develop a complete theory of projective cross-ratios in n-dimensional Plane-Based Geometric Algebra (PGA), R(n,0,1), covering geometric objects of

researcharxiv-cs-cv
19 May 2026
Safety

Generation Navigator: A State-Aware Agentic Framework for Image Generation

DGX agent

arXiv:2605.17969v1 Announce Type: new Abstract: Despite rapid advances in text-to-image generation, faithfully realizing user intent remains challenging, often requiring manual multi-turn trial and er

safetyarxiv-cs-cv
19 May 2026
Research

Generative 3D Gaussians with Learned Density Control

DGX agent

arXiv:2605.16355v1 Announce Type: cross Abstract: We present Density-Sampled Gaussians (DeG), a novel 3D representation designed to bridge the gap between adaptive rendering primitives and scalable ge

researcharxiv-cs-cv
19 May 2026
Research

GeoFlow: Enforcing Implicit Geometric Consistency in Video Generation

DGX agent

arXiv:2605.18365v1 Announce Type: new Abstract: Generating geometrically consistent videos remains an open challenge: text-to-video diffusion models trained on web-scale data treat geometry only impli

researcharxiv-cs-cv
19 May 2026
Research

GeoHand: Unlocking Prior Geometry Knowledge for Monocular 3D Hand Reconstruction

DGX agent

arXiv:2605.17354v1 Announce Type: new Abstract: Monocular 3D hand reconstruction is intrinsically a geometric problem, yet RGB appearance features alone often struggle to resolve severe ambiguities ca

researcharxiv-cs-cv
19 May 2026
Research

Geometry-Editable and Appearance-Preserving Object Compositon

DGX agent

arXiv:2505.20914v2 Announce Type: replace Abstract: General object composition (GOC) aims to seamlessly integrate a target object into a background scene with desired geometric properties, while simul

researcharxiv-cs-cv
19 May 2026
Research

Geospatial-Reasoning-Driven Vocabulary-Agnostic Remote Sensing Semantic Segmentation

DGX agent

arXiv:2602.08206v2 Announce Type: replace Abstract: Open-vocabulary semantic segmentation has become an important direction in remote sensing, as it enables recognition beyond predefined land-cover ca

researcharxiv-cs-cv
19 May 2026
Research

GeoWorld: Geometric World Models

DGX agent

arXiv:2602.23058v2 Announce Type: replace Abstract: Energy-based predictive world models provide a powerful approach for multi-step visual planning by reasoning over latent energy landscapes rather th

researcharxiv-cs-cv
19 May 2026
Model Releases

GLT-PEFT: Gated Lie-Tucker Parameter-Efficient Fine-Tuning for Alzheimer's Disease Diagnosis with Hippocampal Segmentation Pretraining

DGX agent

arXiv:2605.16769v1 Announce Type: new Abstract: Parameter-efficient fine-tuning (PEFT) has emerged as a promising paradigm for adapting pretrained models under limited data conditions. However, most e

model-releasesarxiv-cs-cv
19 May 2026
Applications

GraphMAR: Geometry-Aware Graph Learning Framework for Spatially Adaptive CT Metal Artifact Reduction

DGX agent

arXiv:2605.17343v1 Announce Type: new Abstract: Computed tomography (CT) metal artifact reduction (MAR) aims to reduce the severe streaking artifacts induced by metallic implants and other high-densit

applicationsarxiv-cs-cv
19 May 2026
Research

GraSP-VL: Length as a Semantic Granularity Interface for Vision-Language Representations

DGX agent

arXiv:2605.17727v1 Announce Type: new Abstract: Frozen vision-language embeddings contain signals at multiple semantic resolutions, from object identity to attributes, relations, and full-caption mean

researcharxiv-cs-cv
19 May 2026
Research

HAD: Hallucination-Aware Diffusion Priors for 3D Reconstruction

DGX agent

arXiv:2605.16873v1 Announce Type: new Abstract: Diffusion priors have recently demonstrated strong capability in enhancing the quality of sparse-view 3D reconstruction by augmenting training views at

researcharxiv-cs-cv
19 May 2026
Research

HexagonalWarriorMamba: Superior Threshold-Dependent Multi-label Classification of 12-Lead ECG Cardiac Abnormalities

DGX agent

arXiv:2605.17875v1 Announce Type: new Abstract: The accurate automated diagnosis of cardiac abnormalities from 12-lead electrocardiograms (ECGs) is critical for managing cardiovascular disease. Howeve

researcharxiv-cs-cv
19 May 2026
Local Ai

HierEdit: Region-Aware Hierarchical Diffusion for Efficient High-Resolution Editing

DGX agent

arXiv:2605.17294v1 Announce Type: new Abstract: High-resolution image editing is essential for professional and creative applications, yet existing multimodal diffusion-based editors remain computatio

local-aiarxiv-cs-cv
19 May 2026
Tutorials

High-Resolution Reference Image Assisted Volumetric Super-Resolution of Cardiac Diffusion Weighted Imaging

DGX agent

arXiv:2310.20389v2 Announce Type: replace-cross Abstract: Diffusion Tensor Cardiac Magnetic Resonance (DT-CMR) is the only in vivo method to non-invasively examine the microstructure of the human hear

tutorialsarxiv-cs-cv
19 May 2026
Applications

HighSync: High-Quality Lip Synchronization via Latent Diffusion Models

DGX agent

arXiv:2605.16918v1 Announce Type: new Abstract: We present HighSync, an end-to-end diffusion-based framework for high-fidelity lip synchronization that generates photorealistic talking-face videos ali

applicationsarxiv-cs-cv
19 May 2026
Research

Historical Knowledge Graphs for Global Maritime Estimated Time of Arrival

DGX agent

arXiv:2605.18408v1 Announce Type: new Abstract: Accurate vessel estimated-time-of-arrival forecasts are critical for port operations and decarbonization, yet global-scale travel-time prediction remain

researcharxiv-cs-cv
19 May 2026
Research

HL-OutPaint: Coarse-to-Fine Video Outpainting for High-Resolution Long-Range Videos

DGX agent

arXiv:2605.17543v1 Announce Type: new Abstract: Video outpainting generates plausible visual content beyond the original spatial extent of a video, playing a key role in adapting videos to diverse dis

researcharxiv-cs-cv
19 May 2026
Applications

Hybrid Quantum-MambaVision: A Quantum-enhanced State Space Model for Calibrated Mixed-type Wafer Defect Detection

DGX agent

arXiv:2605.16404v1 Announce Type: new Abstract: Extracting actionable knowledge from industrial visual data is fundamentally bottlenecked by extreme class imbalance and the prohibitive computational c

applicationsarxiv-cs-cv
19 May 2026
Local Ai

HyperTea: A Hypergraph-based Temporal Enhancement and Alignment Network for Moving Infrared Small Target Detection

DGX agent

arXiv:2508.10678v2 Announce Type: replace Abstract: In practical application scenarios, moving infrared small target detection (MIRSTD) remains highly challenging due to the target's small size, weak

local-aiarxiv-cs-cv
19 May 2026
Research

HyperVision: A Channel-Adaptive Ground-Based Hyperspectral Vision Pre-trained Backbone

DGX agent

arXiv:2605.17286v1 Announce Type: new Abstract: While hyperspectral imaging provides rich spatial-spectral information across hundreds of narrow wavelength bands for precise material identification, g

researcharxiv-cs-cv
19 May 2026
Research

Image-to-Video Diffusion: From Foundations to Open Frontiers

DGX agent

arXiv:2605.17248v1 Announce Type: new Abstract: Diffusion-based extit{image-to-video} (I2V) generation has become a central direction in generative models by turning a reference image, with optional c

researcharxiv-cs-cv
19 May 2026
Research

Imaging Hidden Objects with Consumer LiDAR via Motion Induced Sampling

DGX agent

arXiv:2605.17865v1 Announce Type: new Abstract: LiDARs are being increasingly deployed for consumer imaging in handheld, wearable, and robotic applications. These sensors can capture the time-of-fligh

researcharxiv-cs-cv
19 May 2026
Model Releases

iMiGUE-3K: A Large-Scale Benchmark for Micro-Gesture Analysis with Self-Supervised Learning

DGX agent

arXiv:2605.17179v1 Announce Type: new Abstract: Emotion understanding is a fundamental challenge in affective computing and artificial intelligence. While existing approaches predominantly focus on fa

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

Incantation: Natural Language as the Action Interface for Multi-Entity Video World Models

DGX agent

arXiv:2605.18601v1 Announce Type: new Abstract: Modern interactive video world models have achieved impressive visual fidelity, yet lack fine-grained multi-entity control and cross-entity, cross-world

model-releasesarxiv-cs-cv
19 May 2026
Research

Inducing Spatial Locality in Vision Transformers through the Training Protocol

DGX agent

arXiv:2605.16390v1 Announce Type: new Abstract: We investigate whether the training protocol can induce spatial locality in the early layers of a Vision Transformer (ViT) trained from scratch, without

researcharxiv-cs-cv
19 May 2026
Research

InstructAV2AV: Instruction-Guided Audio-Video Joint Editing

DGX agent

arXiv:2605.18467v1 Announce Type: new Abstract: Recent diffusion-based methods have achieved impressive progress in video content manipulation. However, they typically ignore the accompanying audio, l

researcharxiv-cs-cv
19 May 2026
Research

Inter-LPCM: Learning-based Inter-Frame Predictive Coding for LiDAR Point Cloud Compression

DGX agent

arXiv:2605.18006v1 Announce Type: cross Abstract: Because LiDAR sensors acquire point clouds with a fixed angular resolution, the resulting data can be systematically parameterized and efficiently com

researcharxiv-cs-cv
19 May 2026
Model Releases

Intuitive Surgical SurgToolLoc and SurgVU Challenges Results: 2022-2025

DGX agent

arXiv:2305.07152v4 Announce Type: replace Abstract: Robotic assisted (RA) surgery promises to transform surgical intervention. Intuitive Surgical is committed to fostering these changes and the machin

model-releasesarxiv-cs-cv
19 May 2026
Research

Is Complex Training Necessary for Long-Tailed OOD Detection? A Re-think from Feature Geometry

DGX agent

arXiv:2605.17799v1 Announce Type: new Abstract: Long-tailed out-of-distribution (LT-OOD) detection is often addressed with specialized training, including auxiliary out-of-distribution (OOD) data, abs

researcharxiv-cs-cv
19 May 2026
Model Releases

JDCNet: Confidence-Gated Privileged-Modality Distillation for Cost-Preserving X-ray Inference

DGX agent

arXiv:2603.29167v2 Announce Type: replace Abstract: We study a systems-level visual inference problem: using an expensive privileged modality during training while preserving a fixed-cost, single-moda

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

Kelvin v1.0: A Neural Pre-Encoder for H.264: A standards-compliant learned preprocessor with -27.62% BD-VMAF on UVG

DGX agent

arXiv:2605.16376v1 Announce Type: cross Abstract: Kelvin is a lightweight learned pre-encoder that sits in front of an unmodified libx264 encoder. It applies content-adaptive pixel adjustments, bounde

model-releasesarxiv-cs-cv
19 May 2026
Agents

LASAR: Towards Spatio-temporal Reasoning with Latent Cognitive Map

DGX agent

arXiv:2605.16899v1 Announce Type: new Abstract: A fundamental challenge in embodied AI is verifying if agents build internal models of spatial structure or merely learn to mimic task-specific expert t

agentsarxiv-cs-cv
19 May 2026
Safety

LatentUMM: Dual Latent Alignment for Unified Multimodal Models

DGX agent

arXiv:2605.17766v1 Announce Type: new Abstract: Unified multimodal models (UMMs) achieve strong performance in both understanding and generation by learning a shared latent space, yet they often exhib

safetyarxiv-cs-cv
19 May 2026
Research

Learning spatially adaptive sparsity level maps for arbitrary convolutional dictionaries

DGX agent

arXiv:2602.21707v2 Announce Type: replace-cross Abstract: State-of-the-art learned reconstruction methods often rely on black-box modules that, despite their strong performance, raise questions about

researcharxiv-cs-cv
19 May 2026
← Previous
1…160161162163164…263
Next →