AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlog
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,618 results
Research

3DSS: 3D Surface Splatting for Inverse Rendering

DGX agent

arXiv:2605.05876v2 Announce Type: replace-cross Abstract: We present 3D Surface Splatting (3DSS), the first differentiable surface splatting renderer for physically-based inverse rendering from multi-

researcharxiv-cs-cv
11 May 2026
Research

6D Pose Estimation via Keypoint Heatmap Regression with RGB-D Residual Neural Networks

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2605.08059v1 Announce Type: new Abstract: In this paper, we propose a modular framework for 6D pose estimation based on keypoint heatmap regression. Our approach combines YOLOv10m for object det

researcharxiv-cs-cv
11 May 2026
Model Releases

A Causal Diffusion Model for Video Reconstruction from Ultra-Low-Bitrate Representations

DGX agent

arXiv:2602.13837v2 Announce Type: replace Abstract: We study video reconstruction from ultra-low-bitrate representations, where the primary challenge shifts from encoding to decoding. In this regime,

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

A Hierarchical Ensemble Pipeline for Anomaly Detection in ESA Satellite Telemetry

DGX agent

arXiv:2605.06681v1 Announce Type: cross Abstract: A hierarchical ensemble pipeline is introduced to address anomaly detection in multivariate telemetry data provided by European Space Agency (ESA). Th

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

A Marine Debris Detection Framework for Ocean Robots via Self-Attention Enhancement and Feature Interaction Optimization

DGX agent

arXiv:2605.07388v1 Announce Type: new Abstract: Marine debris detection for ocean robot is crucial for ecological protection, yet performance is often degraded by low-quality images with blur, complex

model-releasesarxiv-cs-cv
11 May 2026
Research

A Step to Decouple Optimization in 3DGS

DGX agent

arXiv:2601.16736v5 Announce Type: replace Abstract: 3D Gaussian Splatting (3DGS) has emerged as a powerful technique for real-time novel view synthesis. As an explicit representation optimized through

researcharxiv-cs-cv
11 May 2026
Model Releases

A Unified and Controllable Framework for Layered Image Generation with Visual Effects

DGX agent

arXiv:2601.15507v2 Announce Type: replace Abstract: Recent image generation models produce impressive composites, but often fail to preserve the identity of user-provided content when editing specific

model-releasesarxiv-cs-cv
11 May 2026
Research

A Unified Framework for the Detection and Classification of Fatty Pancreas in Ultrasound Images

DGX agent

arXiv:2605.07466v1 Announce Type: new Abstract: Non-alcoholic fatty pancreas disease (NAFPD) is an underdiagnosed condition associated with metabolic syndrome, insulin resistance, and increased risk o

researcharxiv-cs-cv
11 May 2026
Research

A Unified Measure-Theoretic View of Diffusion, Score-Based, and Flow Matching Generative Models

DGX agent

arXiv:2605.06829v1 Announce Type: cross Abstract: We survey continuous-time generative modeling methods based on transporting a simple reference distribution to a data distribution via stochastic or d

researcharxiv-cs-cv
11 May 2026
Safety

Adaptive Subspace Projection for Generative Personalization

DGX agent

arXiv:2605.07257v1 Announce Type: new Abstract: Generative personalization often suffers from the semantic collapsing problem (SCP), where a learned personalized concept overpowers the rest of the tex

safetyarxiv-cs-cv
11 May 2026
Research

AdpSplit: Error-Driven Adaptive Splitting for Faster Geometry Discovery in 3D Gaussian Splatting

DGX agent

arXiv:2605.06876v1 Announce Type: new Abstract: Adaptive density control in 3D Gaussian Splatting (3DGS) repeatedly grows the Gaussian population through fixed-cardinality random splitting to discover

researcharxiv-cs-cv
11 May 2026
Research

Advancing Reliable Synthetic Video Detection: Insights from the SAFE Challenge

DGX agent

arXiv:2605.06912v1 Announce Type: new Abstract: The proliferation of generative video technologies has intensified the need for reliable methods to detect and characterize synthetic media. To address

researcharxiv-cs-cv
11 May 2026
Research

AGA3DNet: Anatomy-Guided Gaussian Priors with Multi-view xLSTM for 3D Brain MRI Subtype Classification

DGX agent

arXiv:2605.07142v1 Announce Type: new Abstract: Accurate 3D brain MRI subtype classification benefits from both localized anatomical cues and long-range contextual reasoning. We present AGA3DNet, a re

researcharxiv-cs-cv
11 May 2026
Agents

AGILE: Hand-Object Interaction Reconstruction from Video via Agentic Generation

DGX agent

arXiv:2602.04672v3 Announce Type: replace Abstract: Reconstructing dynamic hand-object interactions from monocular videos is critical for dexterous manipulation data collection and creating realistic

agentsarxiv-cs-cv
11 May 2026
Safety

Anisotropic Modality Align

DGX agent

arXiv:2605.07825v1 Announce Type: cross Abstract: Training multimodal large language models has long been limited by the scarcity of high-quality paired multimodal data. Recent studies show that the s

safetyarxiv-cs-cv
11 May 2026
Research

Aquatic Neuromorphic Optical Flow

DGX agent

arXiv:2605.07653v1 Announce Type: new Abstract: Underwater environments impose severe constraints on conventional imaging systems and demand solutions that balance high-quality sensing with strict res

researcharxiv-cs-cv
11 May 2026
Research

AsyncEvGS: Asynchronous Event-Assisted Gaussian Splatting for Handheld Motion-Blurred Scenes

DGX agent

arXiv:2605.07192v1 Announce Type: new Abstract: 3D reconstruction methods such as 3D Gaussian Splatting (3DGS) and Neural Radiance Fields (NeRF) achieve impressive photorealism but fail when input ima

researcharxiv-cs-cv
11 May 2026
Research

Attention Sparsity is Input-Stable: Training-Free Sparse Attention for Video Generation via Offline Sparsity Profiling and Online QK Co-Clustering

DGX agent

arXiv:2603.18636v2 Announce Type: replace Abstract: Diffusion Transformers (DiTs) achieve strong video generation quality but suffer from high inference cost due to dense 3D attention, motivating spar

researcharxiv-cs-cv
11 May 2026
Model Releases

Attention Transfer Is Not Universally Effective for Vision Transformers

DGX agent

arXiv:2605.07191v1 Announce Type: new Abstract: A recent work shows that Attention Transfer, which transfers only the attention patterns from a pre-trained teacher Vision Transformer (ViT) to a random

model-releasesarxiv-cs-cv
11 May 2026
Applications

AudioFace: Language-Assisted Speech-Driven Facial Animation with Multimodal Language Models

DGX agent

arXiv:2605.07478v1 Announce Type: new Abstract: Speech-driven facial animation requires accurate correspondence between acoustic signals and facial motion, especially for articulation-related mouth mo

applicationsarxiv-cs-cv
11 May 2026
Model Releases

Benchmarking Foundation Models for Renal Lesion Stratification in CT

DGX agent

arXiv:2605.07749v1 Announce Type: new Abstract: The rapid proliferation of open-source medical foundation models (FMs) raises a practical question: how well do their pre-trained representations transf

model-releasesarxiv-cs-cv
11 May 2026
Research

Beyond Defenses: Manifold-Aligned Regularization for Intrinsic 3D Point Cloud Robustness

DGX agent

arXiv:2605.07590v1 Announce Type: new Abstract: Despite extensive progress in point cloud robustness, existing methods primarily improve performance through augmentation or defense mechanisms, while o

researcharxiv-cs-cv
11 May 2026
Model Releases

Beyond GSD-as-Token: Continuous Scale Conditioning for Remote Sensing VLMs

DGX agent

arXiv:2605.07562v1 Announce Type: new Abstract: Remote sensing vision-language models (RS-VLMs) face a fundamental mismatch with natural-image counterparts: the same geographic object exhibits radical

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

Breaking Spatial Uniformity: Prior-Guided Mamba with Radial Serialization for Lens Flare Removal

DGX agent

arXiv:2605.07650v1 Announce Type: new Abstract: Lens flares, caused by complex optical aberrations, severely degrade image quality especially in nighttime photography. Although recent restoration meth

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

BRIDGE: Background Routing and Isolated Discrete Gating for Coarse-Mask Local Editing

DGX agent

arXiv:2605.07846v1 Announce Type: new Abstract: Coarse-mask local image editing asks a model to modify a user-indicated region while preserving the surrounding scene. In practice, however, rough masks

model-releasesarxiv-cs-cv
11 May 2026
Tutorials

Bringing Multimodal Large Language Models to Infrared-Visible Image Fusion Quality Assessment

DGX agent

arXiv:2605.06969v1 Announce Type: new Abstract: Infrared-Visible image fusion (IVIF) aims to integrate thermal information and detailed spatial structures into a single fused image to enhance percepti

tutorialsarxiv-cs-cv
11 May 2026
Safety

CalexNet: Soft Cascade-Aligned Training and Calibration for Lightweight Early-Exit Branches

DGX agent

arXiv:2509.08318v2 Announce Type: replace Abstract: Early-exit cascades over a frozen convolutional backbone enable adaptive inference but suffer from three sources of train-inference mismatch: branch

safetyarxiv-cs-cv
11 May 2026
Model Releases

Clinically Aware Synthetic Image Generation for Concept Coverage in Chest X-ray Models

DGX agent

arXiv:2603.15525v2 Announce Type: replace Abstract: Deep learning models for chest X-ray diagnosis are constrained by limited coverage of clinically meaningful concept combinations in publicly availab

model-releasesarxiv-cs-cv
11 May 2026
Research

Cloud-top infrared observations reveal the four-dimensional precipitation structure

DGX agent

arXiv:2605.07499v1 Announce Type: new Abstract: Accurate four-dimensional (4D) precipitation information is essential for understanding the Earth's energy and water cycles, yet remains observationally

researcharxiv-cs-cv
11 May 2026
Research

CONSIGN: Conformal Segmentation Informed by Spatial Groupings via Decomposition

DGX agent

arXiv:2505.14113v3 Announce Type: replace Abstract: Most machine learning-based image segmentation models produce pixel-wise confidence scores that represent the model's predicted probability for each

researcharxiv-cs-cv
11 May 2026
Research

Consistency Regularised Gradient Flows for Inverse Problems

DGX agent

arXiv:2605.07907v1 Announce Type: cross Abstract: Vision-Language Latent Diffusion Models (LDMs) (Rombach et al., 2022) provide powerful generative priors for inverse problems. However, existing LDM-b

researcharxiv-cs-cv
11 May 2026
Model Releases

Contrast-X: A Multi-Modal Contrast Image Synthesis Benchmark and Universal Modality Flow Matching

DGX agent

arXiv:2601.15884v2 Announce Type: replace Abstract: Contrast-enhanced imaging is central to oncologic diagnosis, but contrast agents can be contraindicated for many of the patients who need them most.

model-releasesarxiv-cs-cv
11 May 2026
Local Ai

DeCo-DETR: Decoupled Cognition DETR for efficient Open-Vocabulary Object Detection

DGX agent

arXiv:2604.02753v2 Announce Type: replace Abstract: Open-vocabulary Object Detection (OVOD) enables models to recognize objects beyond predefined categories, but existing approaches remain limited in

local-aiarxiv-cs-cv
11 May 2026
Safety

Decoupling Semantics and Fingerprints: A Universal Representation for AI-Generated Image Detection

DGX agent

arXiv:2605.07074v1 Announce Type: new Abstract: Detecting AI-generated images across unseen architectures remains challenging, as existing models often overfit to generator-specific fingerprints and s

safetyarxiv-cs-cv
11 May 2026
Hardware

DeepFedNAS: Efficient Hardware-Aware Architecture Adaptation for Heterogeneous IoT Federations via Pareto-Guided Supernet Training

DGX agent

arXiv:2601.15127v3 Announce Type: replace-cross Abstract: Deploying federated learning across heterogeneous IoT device fleets requires tailored neural network architectures for each device class, yet

hardwarearxiv-cs-cv
11 May 2026
Model Releases

Deeply Dual Supervised learning for melanoma recognition

DGX agent

arXiv:2508.01994v2 Announce Type: replace Abstract: As the application of deep learning in dermatology continues to grow, the recognition of melanoma has garnered significant attention, demonstrating

model-releasesarxiv-cs-cv
11 May 2026
Tutorials

Delta-Adapter: Scalable Exemplar-Based Image Editing with Single-Pair Supervision

DGX agent

arXiv:2605.07940v1 Announce Type: new Abstract: Exemplar-based image editing applies a transformation defined by a source-target image pair to a new query image. Existing methods rely on a pair-of-pai

tutorialsarxiv-cs-cv
11 May 2026
Research

Differentiable Ray Tracing with Gaussians for Unified Radio Propagation Simulation and View Synthesis

DGX agent

arXiv:2605.07781v1 Announce Type: new Abstract: Explicit neural representations such as 3D Gaussian Splatting (3DGS) enable high-fidelity and real-time novel view synthesis, yet optimize for alpha-com

researcharxiv-cs-cv
11 May 2026
Safety

Diffusion-APO: Trajectory-Aware Direct Preference Alignment for Video Diffusion Transformers

DGX agent

arXiv:2605.07503v1 Announce Type: new Abstract: Efficiently aligning large-scale video diffusion models with human intent requires a scalable and trajectory-aware pathway that bridges the inherent dis

safetyarxiv-cs-cv
11 May 2026
Research

DIMoE-Adapters: Dynamic Expert Evolution for Continual Learning in Vision-Language Models

DGX agent

arXiv:2605.07494v1 Announce Type: new Abstract: Continual learning enables vision-language models to accumulate knowledge and adapt to evolving tasks without retraining from scratch. However, in multi

researcharxiv-cs-cv
11 May 2026
Research

DINO-MVR: Multi-View Readout of Frozen DINOv3 for Annotation-Efficient Medical Segmentation

DGX agent

arXiv:2605.07221v1 Announce Type: new Abstract: Adapting foundation models to medical segmentation typically requires either backbone fine-tuning or high-capacity task-specific decoders, both of which

researcharxiv-cs-cv
11 May 2026
Model Releases

Dino U-Net: Exploiting High-Fidelity Dense Features from Foundation Models for Medical Image Segmentation

DGX agent

arXiv:2508.20909v2 Announce Type: replace Abstract: Foundation models pre-trained on large-scale natural image datasets offer a powerful paradigm for medical image segmentation. However, effectively t

model-releasesarxiv-cs-cv
11 May 2026
Research

Disambiguating 2D-3D Correspondences in Gaussian Splatting-based Feature Fields for Visual Localization

DGX agent

arXiv:2605.07351v1 Announce Type: new Abstract: While Gaussian Splatting-based Feature Fields (GSFFs) have shown promise for visual localization, this paper highlights that photometrically optimized G

researcharxiv-cs-cv
11 May 2026
Model Releases

DKDS: A Benchmark Dataset of Degraded Kuzushiji Documents with Seals for Detection and Binarization

DGX agent

arXiv:2511.09117v4 Announce Type: replace Abstract: Kuzushiji, a pre-modern Japanese cursive script, can currently be read and understood by only a few thousand trained experts in Japan. With the rapi

model-releasesarxiv-cs-cv
11 May 2026
Safety

Dr-BA: Separable Optimization for Direct Radar Bundle Adjustment & Localization

DGX agent

arXiv:2605.07041v1 Announce Type: cross Abstract: This paper introduces Dr-BA, a first-of-its-kind radar bundle adjustment (BA) framework that operates directly on 2D spinning radar intensity images.

safetyarxiv-cs-cv
11 May 2026
Local Ai

DualResolution Residual Architecture with Artifact Suppression for Melanocytic Lesion Segmentation

DGX agent

arXiv:2508.06816v3 Announce Type: replace Abstract: Lesion segmentation, in contrast to natural scene segmentation, requires handling subtle variations in texture and color, frequent imaging artifacts

local-aiarxiv-cs-cv
11 May 2026
Research

DVD: Discrete Voxel Diffusion for 3D Generation and Editing

DGX agent

arXiv:2605.07971v1 Announce Type: new Abstract: We introduce Discrete Voxel Diffusion (DVD), a discrete diffusion framework to generate, assess, and edit sparse voxels for SLat (Structured LATent) bas

researcharxiv-cs-cv
11 May 2026
Agents

Dynamic Mode Decomposition along Depth in Vision Transformers

DGX agent

arXiv:2605.07556v1 Announce Type: new Abstract: Recent work has shown that contiguous vision transformer (ViT) blocks (a) can be replaced by a linear map and (b) organize into recurrent phases of comp

agentsarxiv-cs-cv
11 May 2026
← Previous
1…186187188189190…263
Next →