AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,515 results
Safety

CoSyncDiT: Cognitive Synchronous Diffusion Transformer for Movie Dubbing

DGX agent

arXiv:2604.12292v1 Announce Type: cross Abstract: Movie dubbing aims to synthesize speech that preserves the vocal identity of a reference audio while synchronizing with the lip movements in a target

safetyarxiv-cs-cv
15 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Local Ai

CREG: Compass Relational Evidence Graph for Characterizing Directional Structure in VLM Spatial-Reasoning Attribution

DGX agent

arXiv:2603.20475v3 Announce Type: replace Abstract: Standard attribution heatmaps show where a vision-language model (VLM) focuses, but they do not reveal whether the recovered evidence is organized b

local-aiarxiv-cs-cv
15 Apr 2026
Research

Cross-Attentive Multiview Fusion of Vision-Language Embeddings

DGX agent

arXiv:2604.12551v1 Announce Type: new Abstract: Vision-language models have been key to the development of open-vocabulary 2D semantic segmentation. Lifting these models from 2D images to 3D scenes, h

researcharxiv-cs-cv
15 Apr 2026
Safety

Cross-Modal Knowledge Distillation for PET-Free Amyloid-Beta Detection from MRI

DGX agent

arXiv:2604.12574v1 Announce Type: new Abstract: Detecting amyloid-eta (Aeta) positivity is crucial for early diagnosis of Alzheimer's disease but typically requires PET imaging, which is costly, invas

safetyarxiv-cs-cv
15 Apr 2026
Research

DAV-GSWT: Diffusion-Active-View Sampling for Data-Efficient Gaussian Splatting Wang Tiles

DGX agent

arXiv:2602.15355v3 Announce Type: replace Abstract: The emergence of 3D Gaussian Splatting has fundamentally redefined the capabilities of photorealistic neural rendering by enabling high-throughput s

researcharxiv-cs-cv
15 Apr 2026
Research

DC-TTA: Divide-and-Conquer Framework for Test-Time Adaptation of Interactive Segmentation

DGX agent

arXiv:2506.23104v2 Announce Type: replace Abstract: Interactive segmentation (IS) allows users to iteratively refine object boundaries with minimal cues, such as positive and negative clicks. While th

researcharxiv-cs-cv
15 Apr 2026
Research

Deep Learning using Rectified Linear Units (ReLU)

DGX agent

arXiv:1803.08375v3 Announce Type: replace-cross Abstract: The Rectified Linear Unit (ReLU) is a foundational activation function in artficial neural networks. Recent literature frequently misattribute

researcharxiv-cs-cv
15 Apr 2026
Research

DeferredSeg: A Multi-Expert Deferral Framework for Trustworthy Medical Image Segmentation

DGX agent

arXiv:2604.12411v1 Announce Type: new Abstract: Segmentation models based on deep neural networks demonstrate strong generalization for medical image segmentation. However, they often exhibit overconf

researcharxiv-cs-cv
15 Apr 2026
Research

Detecting Precise Hand Touch Moments in Egocentric Video

DGX agent

arXiv:2604.12343v1 Announce Type: new Abstract: We address the challenging task of detecting the precise moment when hands make contact with objects in egocentric videos. This frame-level detection is

researcharxiv-cs-cv
15 Apr 2026
Local Ai

DiffusionPrint: Learning Generative Fingerprints for Diffusion-Based Inpainting Localization

DGX agent

arXiv:2604.12443v1 Announce Type: new Abstract: Modern diffusion-based inpainting models pose significant challenges for image forgery localization (IFL), as their full regeneration pipelines reconstr

local-aiarxiv-cs-cv
15 Apr 2026
Agents

DINO-Explorer: Active Underwater Discovery via Ego-Motion Compensated Semantic Predictive Coding

DGX agent

arXiv:2604.12933v1 Announce Type: cross Abstract: Marine ecosystem degradation necessitates continuous, scientifically selective underwater monitoring. However, most autonomous underwater vehicles (AU

agentsarxiv-cs-cv
15 Apr 2026
Tutorials

Direct Discrepancy Replay: Distribution-Discrepancy Condensation and Manifold-Consistent Replay for Continual Face Forgery Detection

DGX agent

arXiv:2604.12941v1 Announce Type: new Abstract: Continual face forgery detection (CFFD) requires detectors to learn emerging forgery paradigms without forgetting previously seen manipulations. Existin

tutorialsarxiv-cs-cv
15 Apr 2026
Research

Does Visual Token Pruning Improve Calibration? An Empirical Study on Confidence in MLLMs

DGX agent

arXiv:2604.12035v1 Announce Type: new Abstract: Visual token pruning is a widely used strategy for efficient inference in multimodal large language models (MLLMs), but existing work mainly evaluates i

researcharxiv-cs-cv
15 Apr 2026
Model Releases

Don't Show Pixels, Show Cues: Unlocking Visual Tool Reasoning in Language Models via Perception Programs

DGX agent

arXiv:2604.12896v1 Announce Type: new Abstract: Multimodal language models (MLLMs) are increasingly paired with vision tools (e.g., depth, flow, correspondence) to enhance visual reasoning. However, d

model-releasesarxiv-cs-cv
15 Apr 2026
Model Releases

DPC-VQA: Decoupling Quality Perception and Residual Calibration for Video Quality Assessment

DGX agent

arXiv:2604.12813v1 Announce Type: new Abstract: Recent multimodal large language models (MLLMs) have shown promising performance on video quality assessment (VQA) tasks. However, adapting them to new

model-releasesarxiv-cs-cv
15 Apr 2026
Hardware

DreamStereo: Towards Real-Time Stereo Inpainting for HD Videos

DGX agent

arXiv:2604.12270v1 Announce Type: new Abstract: Stereo video inpainting, which aims to fill the occluded regions of warped videos with visually coherent content while maintaining temporal consistency,

hardwarearxiv-cs-cv
15 Apr 2026
Model Releases

Dress-ED: Instruction-Guided Editing for Virtual Try-On and Try-Off

DGX agent

arXiv:2603.22607v2 Announce Type: replace Abstract: Recent advances in Virtual Try-On (VTON) and Virtual Try-Off (VTOFF) have greatly improved photo-realistic fashion synthesis and garment reconstruct

model-releasesarxiv-cs-cv
15 Apr 2026
Model Releases

Dual-Modality Anchor-Guided Filtering for Test-time Prompt Tuning

DGX agent

arXiv:2604.12403v1 Announce Type: new Abstract: Test-Time Prompt Tuning (TPT) adapts vision-language models using augmented views, but its effectiveness is hindered by the challenge of determining whi

model-releasesarxiv-cs-cv
15 Apr 2026
Model Releases

EDGE-Shield: Efficient Denoising-staGE Shield for Violative Content Filtering via Scalable Reference-Based Matching

DGX agent

arXiv:2604.06063v2 Announce Type: replace Abstract: The advent of Text-to-Image generative models poses significant risks of copyright violation and deepfake generation. Since the rapid proliferation

model-releasesarxiv-cs-cv
15 Apr 2026
Research

EigenCoin: sassanid coins classification based on Bhattacharyya distance

DGX agent

arXiv:2604.11932v1 Announce Type: new Abstract: Solving pattern recognition problems using imbalanced databases is a hot topic, which entices researchers to bring it into focus. Therefore, we consider

researcharxiv-cs-cv
15 Apr 2026
Model Releases

ELoG-GS: Dual-Branch Gaussian Splatting with Luminance-Guided Enhancement for Extreme Low-light 3D Reconstruction

DGX agent

arXiv:2604.12592v1 Announce Type: new Abstract: This paper presents our approach to the NTIRE 2026 3D Restoration and Reconstruction Challenge (Track 1), which focuses on reconstructing high-quality 3

model-releasesarxiv-cs-cv
15 Apr 2026
Safety

Evolution-Inspired Sample Competition for Deep Neural Network Optimization

DGX agent

arXiv:2604.12568v1 Announce Type: new Abstract: Conventional deep network training generally optimizes all samples under a largely uniform learning paradigm, without explicitly modeling the heterogene

safetyarxiv-cs-cv
15 Apr 2026
Research

Evolution of Optimization Methods: Algorithms, Scenarios, and Evaluations

DGX agent

arXiv:2604.12968v1 Announce Type: cross Abstract: Balancing convergence speed, generalization capability, and computational efficiency remains a core challenge in deep learning optimization. First-ord

researcharxiv-cs-cv
15 Apr 2026
Research

Fall Risk and Gait Analysis in Community-Dwelling Older Adults using World-Spaced 3D Human Mesh Recovery

DGX agent

arXiv:2604.11961v1 Announce Type: new Abstract: Gait assessment is a key clinical indicator of fall risk and overall health in older adults. However, standard clinical practice is largely limited to s

researcharxiv-cs-cv
15 Apr 2026
Safety

Fragile Reconstruction: Adversarial Vulnerability of Reconstruction-Based Detectors for Diffusion-Generated Images

DGX agent

arXiv:2604.12781v1 Announce Type: new Abstract: Recently, detecting AI-generated images produced by diffusion-based models has attracted increasing attention due to their potential threat to safety. A

safetyarxiv-cs-cv
15 Apr 2026
Research

From Attenuation to Attention: Variational Information Flow Manipulation for Fine-Grained Visual Perception

DGX agent

arXiv:2604.12508v1 Announce Type: new Abstract: While Multimodal Large Language Models (MLLMs) have demonstrated impressive capabilities in general visual understanding, they frequently falter in fine

researcharxiv-cs-cv
15 Apr 2026
Model Releases

Fundus Image-based Glaucoma Screening via Retinal Knowledge-Oriented Dynamic Multi-Level Feature Integration

DGX agent

arXiv:2604.12351v1 Announce Type: new Abstract: Automated diagnosis based on color fundus photography is essential for large-scale glaucoma screening. However, existing deep learning models are typica

model-releasesarxiv-cs-cv
15 Apr 2026
Model Releases

Generative Anonymization in Event Streams

DGX agent

arXiv:2604.12803v1 Announce Type: new Abstract: Neuromorphic vision sensors offer low latency and high dynamic range, but their deployment in public spaces raises severe data protection concerns. Rece

model-releasesarxiv-cs-cv
15 Apr 2026
Model Releases

Generative Refinement Networks for Visual Synthesis

DGX agent

arXiv:2604.13030v1 Announce Type: new Abstract: While diffusion models dominate the field of visual generation, they are computationally inefficient, applying a uniform computational effort regardless

model-releasesarxiv-cs-cv
15 Apr 2026
Safety

Grasp in Gaussians: Fast Monocular Reconstruction of Dynamic Hand-Object Interactions

DGX agent

arXiv:2604.12929v1 Announce Type: new Abstract: We present Grasp in Gaussians (GraG), a fast and robust method for reconstructing dynamic 3D hand-object interactions from a single monocular video. Unl

safetyarxiv-cs-cv
15 Apr 2026
Model Releases

GroupKAN: Efficient Kolmogorov-Arnold Networks via Grouped Spline Modeling

DGX agent

arXiv:2511.05477v2 Announce Type: replace Abstract: Medical image segmentation demands models that achieve high accuracy while maintaining computational efficiency and clinical interpretability. While

model-releasesarxiv-cs-cv
15 Apr 2026
Model Releases

GTPBD-MM: A Global Terraced Parcel and Boundary Dataset with Multi-Modality

DGX agent

arXiv:2604.12315v1 Announce Type: new Abstract: Agricultural parcel extraction plays an important role in remote sensing-based agricultural monitoring, supporting parcel surveying, precision managemen

model-releasesarxiv-cs-cv
15 Apr 2026
Local Ai

Habitat Classification from Ground-Level Imagery Using Deep Neural Networks

DGX agent

arXiv:2507.04017v3 Announce Type: replace Abstract: Habitat assessment at local scales -- critical for enhancing biodiversity and guiding conservation priorities -- often relies on expert field survey

local-aiarxiv-cs-cv
15 Apr 2026
Agents

Habitat-GS: A High-Fidelity Navigation Simulator with Dynamic Gaussian Splatting

DGX agent

arXiv:2604.12626v1 Announce Type: cross Abstract: Training embodied AI agents depends critically on the visual fidelity of simulation environments and the ability to model dynamic humans. Current simu

agentsarxiv-cs-cv
15 Apr 2026
Research

HTDC: Hesitation-Triggered Differential Calibration for Mitigating Hallucination in Large Vision-Language Models

DGX agent

arXiv:2604.12115v1 Announce Type: new Abstract: Large vision-language models (LVLMs) achieve strong multimodal performance, but still suffer from hallucinations caused by unstable visual grounding and

researcharxiv-cs-cv
15 Apr 2026
Research

Hypergraph-State Collaborative Reasoning for Multi-Object Tracking

DGX agent

arXiv:2604.12665v1 Announce Type: new Abstract: Motion reasoning serves as the cornerstone of multi-object tracking (MOT), as it enables consistent association of targets across frames. However, exist

researcharxiv-cs-cv
15 Apr 2026
Local Ai

HyperLiDAR: Adaptive Post-Deployment LiDAR Segmentation via Hyperdimensional Computing

DGX agent

arXiv:2604.12331v1 Announce Type: new Abstract: LiDAR semantic segmentation plays a pivotal role in 3D scene understanding for edge applications such as autonomous driving. However, significant challe

local-aiarxiv-cs-cv
15 Apr 2026
Research

Image-to-Image Translation Framework Embedded with Rotation Symmetry Priors

DGX agent

arXiv:2604.12805v1 Announce Type: new Abstract: Image-to-image translation (I2I) is a fundamental task in computer vision, focused on mapping an input image from a source domain to a corresponding ima

researcharxiv-cs-cv
15 Apr 2026
Research

IMU: Influence-guided Machine Unlearning

DGX agent

arXiv:2508.01620v3 Announce Type: replace-cross Abstract: Machine Unlearning (MU) aims to selectively erase the influence of specific data points from pretrained models. However, most existing MU meth

researcharxiv-cs-cv
15 Apr 2026
Model Releases

INST-Align: Implicit Neural Alignment for Spatial Transcriptomics via Canonical Expression Fields

DGX agent

arXiv:2604.12084v1 Announce Type: new Abstract: Spatial transcriptomics (ST) measures mRNA expression while preserving spatial organization, but multi-slice analysis faces two coupled difficulties: la

model-releasesarxiv-cs-cv
15 Apr 2026
Applications

Label-Efficient Cross-Modality Generalization for Liver Segmentation in Multi-Phase MRI

DGX agent

arXiv:2510.04705v4 Announce Type: replace Abstract: Accurate liver segmentation in multi-phase MRI is vital for liver fibrosis assessment, yet labeled data is often scarce and unevenly distributed acr

applicationsarxiv-cs-cv
15 Apr 2026
Model Releases

Latent Chain-of-Thought World Modeling for End-to-End Driving

DGX agent

arXiv:2512.10226v2 Announce Type: replace Abstract: Recent Vision-Language-Action (VLA) models for autonomous driving explore inference-time reasoning as a way to improve driving performance and safet

model-releasesarxiv-cs-cv
15 Apr 2026
Model Releases

LaV-CoT: Language-Aware Visual CoT with Multi-Aspect Reward Optimization for Real-World Multilingual VQA

DGX agent

arXiv:2509.10026v4 Announce Type: replace Abstract: As large vision language models (VLMs) advance, their capabilities in multilingual visual question answering (mVQA) have significantly improved. Cha

model-releasesarxiv-cs-cv
15 Apr 2026
Tutorials

Listening Deepfake Detection: A New Perspective Beyond Speaking-Centric Forgery Analysis

DGX agent

arXiv:2604.12650v1 Announce Type: new Abstract: Existing deepfake detection research has primarily focused on scenarios where the manipulated subject is actively speaking, i.e., generating fabricated

tutorialsarxiv-cs-cv
15 Apr 2026
Safety

LiveMoments: Reselected Key Photo Restoration in Live Photos via Reference-guided Diffusion

DGX agent

arXiv:2604.12286v1 Announce Type: new Abstract: Live Photo captures both a high-quality key photo and a short video clip to preserve the precious dynamics around the captured moment. While users may c

safetyarxiv-cs-cv
15 Apr 2026
Research

Lyra 2.0: Explorable Generative 3D Worlds

DGX agent

arXiv:2604.13036v1 Announce Type: new Abstract: Recent advances in video generation enable a new paradigm for 3D scene creation: generating camera-controlled videos that simulate scene walkthroughs, t

researcharxiv-cs-cv
15 Apr 2026
Model Releases

M3D-Stereo: A Multiple-Medium and Multiple-Degradation Dataset for Stereo Image Restoration

DGX agent

arXiv:2604.12917v1 Announce Type: new Abstract: Image restoration under adverse conditions, such as underwater, haze or fog, and low-light environments, remains a highly challenging problem due to com

model-releasesarxiv-cs-cv
15 Apr 2026
Model Releases

MedConcept: Unsupervised Concept Discovery for Interpretability in Medical VLMs

DGX agent

arXiv:2604.11868v1 Announce Type: new Abstract: While medical Vision-Language models (VLMs) achieve strong performance on tasks such as tumor or organ segmentation and diagnosis prediction, their opaq

model-releasesarxiv-cs-cv
15 Apr 2026
← Previous
1…242243244245246…261
Next →