AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,618 results
Model Releases

CT-DegradBench: A Physics-Informed Benchmark for CT Degradation Detection and Severity Estimation

DGX agent

arXiv:2605.16431v1 Announce Type: new Abstract: Computed tomography (CT) images are frequently degraded by acquisition artifacts, including noise, blur, streaking, aliasing, and metal artifacts. Yet C

model-releasesarxiv-cs-cv
19 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Cultivating Forensic Reasoning for Generalizable Multimodal Manipulation Detection

DGX agent

arXiv:2603.01993v2 Announce Type: replace Abstract: Recent advances in generative AI have significantly enhanced the realism of multimodal media manipulation, thereby posing substantial challenges to

researcharxiv-cs-cv
19 May 2026
Safety

Dance Across Shifts: Forward-Facilitation Continual Test-Time Adaptation through Dynamic Style Bridging

DGX agent

arXiv:2605.18608v1 Announce Type: new Abstract: Continual Test-Time Adaptation (CTTA) aims to empower perception systems to handle dynamic distribution shifts encountered after deployment. Existing me

safetyarxiv-cs-cv
19 May 2026
Applications

DanceHMR: Hand-Aware Whole-Body Human Mesh Recovery from Monocular Videos

DGX agent

arXiv:2605.18102v1 Announce Type: new Abstract: Monocular video human mesh recovery is essential for digital humans, avatar animation, and embodied simulation, where both temporal stability and expres

applicationsarxiv-cs-cv
19 May 2026
Research

DASH: A Meta-Attack Framework for Synthesizing Effective and Stealthy Adversarial Examples

DGX agent

arXiv:2508.13309v3 Announce Type: replace Abstract: Numerous techniques have been proposed for generating adversarial examples in white-box settings under strict Lp-norm constraints. However, such nor

researcharxiv-cs-cv
19 May 2026
Agents

DECODE: Domain-aware Continual Domain Expansion for Motion Prediction

DGX agent

arXiv:2411.17917v2 Announce Type: replace Abstract: Motion prediction is critical for autonomous vehicles to effectively navigate complex environments and accurately anticipate the behaviors of other

agentsarxiv-cs-cv
19 May 2026
Research

DecoRec: Decomposed 3D Scene Reconstruction from Single-View Images via Object-Level Diffusion

DGX agent

arXiv:2605.16807v1 Announce Type: new Abstract: In this paper, we introduce extit{DecoRec}, a novel system designed to elevate single-view 2D images to a decomposed 3D scene mesh. Current methods for

researcharxiv-cs-cv
19 May 2026
Research

Decoupling Motion and Geometry in 4D Gaussian Splatting

DGX agent

arXiv:2603.00952v2 Announce Type: replace Abstract: High-fidelity reconstruction of dynamic scenes is an important yet challenging problem. While recent 4D Gaussian Splatting (4DGS) has demonstrated t

researcharxiv-cs-cv
19 May 2026
Research

Deep learning-based compression of giga-resolution whole slide images

DGX agent

arXiv:2605.17668v1 Announce Type: new Abstract: Implementation of digital pathology leads to an increased number of whole slide images (WSIs). The large size of WSIs is challenging. Today, WSIs are co

researcharxiv-cs-cv
19 May 2026
Research

Deep Learning for MRI Slice Interpolation: The Critical Role of Problem Formulation

DGX agent

arXiv:2605.16476v1 Announce Type: cross Abstract: Through-plane resolution in clinical MRI is typically much coarser than in-plane resolution, limiting diagnostic utility. This work investigates deep

researcharxiv-cs-cv
19 May 2026
Research

Deepfake Detection in Social Media: A Temporal Artifact Analysis Using 3D Convolutional Neural Networks

DGX agent

arXiv:2605.17573v1 Announce Type: new Abstract: Synthetic facial videos have proliferated across social media faster than platform moderation can respond, raising the cost of disinformation and identi

researcharxiv-cs-cv
19 May 2026
Tutorials

Degradation Frequency Curve: An Explicit Frequency-Quantified Representation for All-in-One Image Restoration

DGX agent

arXiv:2605.17506v1 Announce Type: new Abstract: A fundamental difficulty in all-in-one blind image restoration is that degradation is usually treated as an implicit factor hidden in degraded-to-clean

tutorialsarxiv-cs-cv
19 May 2026
Model Releases

DepthPolyp: Pseudo-Depth Guided Lightweight Segmentation for Real-Time Colonoscopy

DGX agent

arXiv:2605.16519v1 Announce Type: new Abstract: Accurate polyp segmentation in colonoscopy is essential for early colorectal cancer detection, yet real-world clinical environments pose persistent chal

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

Designing streetscapes from street-view imagery using diffusion models

DGX agent

arXiv:2605.17527v1 Announce Type: new Abstract: Street-view imagery (SVI) is widely used to quantify key indicators of urban environment, such as green- ery, sky, or road view indices. However, existi

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

DeTrack: A Benchmark and Altitude-Aware Dual World Model for Drone-embodied Tracking

DGX agent

arXiv:2605.17451v1 Announce Type: new Abstract: Aerial object tracking has broad applications in public safety, emergency rescue, wildlife monitoring, and related fields. However, existing aerial trac

model-releasesarxiv-cs-cv
19 May 2026
Safety

DEVIS-GRPO: Unleashing GRPO on Dynamic Extreme View Synthesis

DGX agent

arXiv:2605.16937v1 Announce Type: new Abstract: Trajectory-controlled video generation has become essential for controllable video generation. While current methods perform well under small-view camer

safetyarxiv-cs-cv
19 May 2026
Local Ai

Diffeomorphic Cortical Alignment via Direct Warping of Streamline Endpoints

DGX agent

arXiv:2605.16742v1 Announce Type: new Abstract: Cortical surface registration is often driven by local geometric descriptors (e.g., sulcal depth and curvature). While this approach achieves geometric

local-aiarxiv-cs-cv
19 May 2026
Model Releases

Diffusion-Based sRGB Real Noise Generation via Prompt-Driven Noise Representation Learning

DGX agent

arXiv:2603.04870v2 Announce Type: replace Abstract: Denoising in the sRGB image space is challenging due to large noise variability. Although end-to-end methods perform well, their effectiveness in re

model-releasesarxiv-cs-cv
19 May 2026
Safety

Diffusion Models, Denoiser Architecture and Creativity

DGX agent

arXiv:2605.16415v1 Announce Type: new Abstract: The creativity of diffusion models refers to their ability to generate highly realistic images that are different from their training data. Creativity i

safetyarxiv-cs-cv
19 May 2026
Applications

DiffWind: Physics-Informed Differentiable Modeling of Wind-Driven Object Dynamics

DGX agent

arXiv:2603.09668v2 Announce Type: replace Abstract: Modeling wind-driven object dynamics from video observations is highly challenging due to the invisibility and spatio-temporal variability of wind,

applicationsarxiv-cs-cv
19 May 2026
Safety

DiRotQ: Rotation-Aware Quantization for 4-bit Diffusion Transformers

DGX agent

arXiv:2605.16732v1 Announce Type: new Abstract: Diffusion Transformers (DiTs) achieve state-of-the-art image generation quality but incur substantial memory and computational costs at inference. While

safetyarxiv-cs-cv
19 May 2026
Model Releases

DisasterVQA: A Visual Question Answering Benchmark Dataset for Disaster Scenes

DGX agent

arXiv:2601.13839v2 Announce Type: replace Abstract: Social media imagery provides a low-latency source of situational information during natural and human-induced disasters, enabling rapid damage asse

model-releasesarxiv-cs-cv
19 May 2026
Tutorials

Distribution Prototype Diffusion Learning for Open-set Supervised Anomaly Detection

DGX agent

arXiv:2502.20981v2 Announce Type: replace Abstract: In Open-set Supervised Anomaly Detection (OSAD), the existing methods typically generate pseudo anomalies to compensate for the scarcity of observed

tutorialsarxiv-cs-cv
19 May 2026
Research

Do You Need Text Rectification? Soft Attention Mask Embedding for Rectification-Free Scene Text Spotting

DGX agent

arXiv:2605.18173v1 Announce Type: new Abstract: End-to-end scene text spotting, which unifies text detection and recognition within a single framework, has witnessed remarkable progress driven by deep

researcharxiv-cs-cv
19 May 2026
Safety

DreamEdit3D: Personalization of Multi-View Diffusion Models for 3D Editing

DGX agent

arXiv:2605.16990v1 Announce Type: new Abstract: While 2D diffusion models have achieved remarkable success in identity-preserving personalization, extending this capability to 3D assets remains a sign

safetyarxiv-cs-cv
19 May 2026
Model Releases

DriveSafer: End-to-End Autonomous Driving with Safety Guidance

DGX agent

arXiv:2605.16737v1 Announce Type: cross Abstract: End-to-End (E2E) autonomous driving models have shown growing capability in recent years, with performance improving on increasingly challenging bench

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

DSAA: Dual-Stage Attribute Activation for Fine-grained Open Vocabulary Detection

DGX agent

arXiv:2605.18023v1 Announce Type: new Abstract: Open-Vocabulary Object Detection (OVD) models break the limitations of closed-set detection, enabling the iden- tification of unseen categories through

model-releasesarxiv-cs-cv
19 May 2026
Research

Dual-Rate Diffusion: Accelerating diffusion models with an interleaved heavy-light network

DGX agent

arXiv:2605.18190v1 Announce Type: cross Abstract: Diffusion models achieve state-of-the-art generative performance but suffer from high computational costs during inference due to the repeated evaluat

researcharxiv-cs-cv
19 May 2026
Research

EchoSR: Efficient Context Harnessing for Lightweight Image Super-Resolution

DGX agent

arXiv:2605.17470v1 Announce Type: new Abstract: Image super-resolution (SR) aims to reconstruct high-quality, high-resolution (HR) images from low-resolution (LR) inputs and plays a critical role in v

researcharxiv-cs-cv
19 May 2026
Safety

Edit-GRPO: A Locality-Preserving Policy Optimization Framework for Image Editing

DGX agent

arXiv:2605.16951v1 Announce Type: new Abstract: A fundamental challenge in image editing lies in preserving spatial locality: edits should improve targeted content without inadvertently altering surro

safetyarxiv-cs-cv
19 May 2026
Hardware

Efficient 3D Content Reconstruction and Generation

DGX agent

arXiv:2605.18052v1 Announce Type: new Abstract: Automatic 3D content creation seeks to replace labor-intensive modeling and scanning pipelines with systems that can synthesize or recover 3D assets dir

hardwarearxiv-cs-cv
19 May 2026
Research

Efficient Sparse-to-Dense Visual Localization via Compact Gaussian Scene Representation and Accelerated Dense Pose Estimation

DGX agent

arXiv:2605.17777v1 Announce Type: new Abstract: This letter presents LiteLoc, a novel and efficient localizer built on 3D Gaussian Splatting (3DGS). The previous state-of-the-art (SoTA) sparse-to-dens

researcharxiv-cs-cv
19 May 2026
Research

Efficient Spatially-Variant Convolution via Differentiable Sparse Kernel Complex

DGX agent

arXiv:2512.04556v2 Announce Type: replace-cross Abstract: Image convolution with complex kernels is a fundamental operation in photography, scientific imaging, and animation effects, yet direct dense

researcharxiv-cs-cv
19 May 2026
Model Releases

EgoExoMem: Cross-View Memory Reasoning over Synchronized Egocentric and Exocentric Videos

DGX agent

arXiv:2605.18734v1 Announce Type: new Abstract: Egocentric memory is widely used in embodied intelligence, but it may be insufficient for comprehensive spatial-temporal reasoning. Inspired by human re

model-releasesarxiv-cs-cv
19 May 2026
Applications

EgoInteract: Synthetic Egocentric Videos Generation for Interaction Understanding and Anticipation

DGX agent

arXiv:2605.18214v1 Announce Type: new Abstract: Collecting large-scale egocentric video datasets with dense spatial and temporal annotations is costly, slow, and often constrained by environmental bia

applicationsarxiv-cs-cv
19 May 2026
Model Releases

EgoIntrospect: An Egocentric Dataset and Benchmark for User-Centric Internal State Reasoning

DGX agent

arXiv:2605.17262v1 Announce Type: new Abstract: Despite extensive efforts on egocentric video datasets and benchmarks, understanding users' internal states, which is crucial for enabling seamless AI a

model-releasesarxiv-cs-cv
19 May 2026
Local Ai

EgoKit: Towards Unified Low-Cost Egocentric Data Collection with Heterogeneous Devices

DGX agent

arXiv:2605.16797v1 Announce Type: new Abstract: Egocentric video is increasingly used as a data source for robot learning, activity understanding, and embodied AI research, but collecting it at scale

local-aiarxiv-cs-cv
19 May 2026
Research

Embedded ConvNet Ensembles: A Lightweight Approach to Recognize Arabic Handwritten Characters

DGX agent

arXiv:2605.18060v1 Announce Type: new Abstract: Arabic Handwritten Character Recognition (AHCR) has recently advanced significantly with deep Convolutional Neural Networks (ConvNets). However, many mo

researcharxiv-cs-cv
19 May 2026
Model Releases

Employing Vision-Language Models for Face Image Quality Assessment

DGX agent

arXiv:2605.17489v1 Announce Type: new Abstract: Face Image Quality Assessment (FIQA) is a crucial control step in biometric pipelines. It ensures only reliable samples are processed to maintain system

model-releasesarxiv-cs-cv
19 May 2026
Agents

Enhancing Event-based Object Detection with Monocular Normal Maps

DGX agent

arXiv:2508.02127v2 Announce Type: replace Abstract: Object detection in autonomous driving is frequently compromised by complex illumination. While event cameras offer a robust solution, they are susc

agentsarxiv-cs-cv
19 May 2026
Safety

Enhancing Train-Free Infinite-Frame Generation for Consistent Long Videos

DGX agent

arXiv:2605.18233v1 Announce Type: new Abstract: Without incurring significant computational overhead, train-free long video generation aims to enable foundation video generation models to produce long

safetyarxiv-cs-cv
19 May 2026
Model Releases

EPIC-Bench: A Perception-Centric Benchmark for Fine-Grained Embodied Visual Grounding in Vision-Language Models

DGX agent

arXiv:2605.17070v1 Announce Type: new Abstract: While large vision-language models (VLMs) are increasingly adopted as the perceptual backbone for embodied agents, existing benchmarks often rely on que

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

Error-Decomposed Class-Conditional Fusion for Statistically Guaranteed Hard-Category Robust Perception

DGX agent

arXiv:2605.17591v1 Announce Type: new Abstract: Aggregate object detection metrics inherently mask catastrophic and repeatable failures in operationally critical, long-tail minority classes. This pape

model-releasesarxiv-cs-cv
19 May 2026
Research

EVA01: Unified Native 3D Understanding and Generation via Mixture-of-Transformers

DGX agent

arXiv:2605.16745v1 Announce Type: new Abstract: This paper addresses the challenge of integrating 3D meshes as a native modality within Multimodal Large Language Models (MLLMs). Diffusion-based large

researcharxiv-cs-cv
19 May 2026
Research

Evidence-Guided Unknown Rejection for High-Confidence Near-Known Unknowns

DGX agent

arXiv:2605.17818v1 Announce Type: new Abstract: Open-set recognition systems face a neglected failure mode: high-confidence near-known unknowns, which lie outside the known label set but are close eno

researcharxiv-cs-cv
19 May 2026
Model Releases

Expandable, Compressible, Mineable: Open-World Thermal Image Restoration

DGX agent

arXiv:2605.16967v1 Announce Type: new Abstract: In open-world settings, thermal infrared (TIR) image degradations continuously emerge and evolve, while most existing all-in-one restoration methods are

model-releasesarxiv-cs-cv
19 May 2026
Research

Explaining Object Detectors via Collective Contribution of Pixels

DGX agent

arXiv:2412.00666v4 Announce Type: replace Abstract: Visual explanations for object detectors are crucial for enhancing their reliability. Object detectors identify and localize instances by assessing

researcharxiv-cs-cv
19 May 2026
Model Releases

extit{Don't Guess, Just Ask}: Resolving Ambiguity in Referring Segmentation via Multi-turn Clarification

DGX agent

arXiv:2605.17531v1 Announce Type: new Abstract: Referring segmentation aims to segment the target objects in images or videos based on the textual query. Despite remarkable progress over the past year

model-releasesarxiv-cs-cv
19 May 2026
← Previous
1…159160161162163…263
Next →