AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,618 results
Tutorials

C-LEAD: Contrastive Learning for Enhanced Adversarial Defense

DGX agent

arXiv:2510.27249v2 Announce Type: replace Abstract: Deep neural networks (DNNs) have achieved remarkable success in computer vision tasks such as image classification, segmentation, and object detecti

tutorialsarxiv-cs-cv
2 Jun 2026
Safety
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

CanonCGT: Reference-Based Color Grading via Canonical Pivot Representation

DGX agent

arXiv:2606.01638v1 Announce Type: new Abstract: Reference-based color grading aims to reproduce the tonal mood and lighting of a reference while preserving color harmony and scene structure. Existing

safetyarxiv-cs-cv
2 Jun 2026
Model Releases

CASTLE2026 Team WDL Technical Report

DGX agent

arXiv:2606.00712v1 Announce Type: new Abstract: The CASTLE Challenge @ EgoVis 2026 evaluates long-form egocentric video question answering over 600+ hours of multi-perspective recordings. Each four-ch

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Chameleon: Style-Content Disentangled Framework for Cross-Domain Object Compositing

DGX agent

arXiv:2606.01079v1 Announce Type: new Abstract: Image compositing aims to seamlessly insert a foreground object into a background image, and recent advances in diffusion models have significantly enha

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

ChartArena: Benchmarking Chart Parsing across Languages, Scenarios, and Formats

DGX agent

arXiv:2606.01348v1 Announce Type: new Abstract: Charts are a primary medium for conveying quantitative and relational information, yet systematically evaluating chart parsing models remains difficult.

model-releasesarxiv-cs-cv
2 Jun 2026
Research

ChatUMM: Robust Context Tracking for Conversational Interleaved Generation

DGX agent

arXiv:2602.06442v2 Announce Type: replace Abstract: Unified multimodal models (UMMs) have achieved remarkable progress yet remain constrained by a single-turn interaction paradigm, effectively functio

researcharxiv-cs-cv
2 Jun 2026
Tutorials

Chroma Clues: Leveraging Color Statistics to Detect Synthetic Images

DGX agent

arXiv:2606.02224v1 Announce Type: new Abstract: The evolution and dissemination of AI-synthesized images is occurring at an unprecedented rate. Image generators are making rapid progress in their goal

tutorialsarxiv-cs-cv
2 Jun 2026
Research

ChWDTA: Channel-wise Wavelet-Domain Transformer Attention and Entropy Modeling for Learned Image Compression

DGX agent

arXiv:2606.00111v1 Announce Type: cross Abstract: State-of-the-art learned image compression (LIC) schemes are increasingly based on hybrid CNN-transformer architectures. To further improve rate-disto

researcharxiv-cs-cv
2 Jun 2026
Tutorials

CLIP-like Model as a Foundational Density Ratio Estimator

DGX agent

arXiv:2506.22881v3 Announce Type: replace Abstract: Density ratio estimation is a core concept in statistical machine learning because it provides a unified mechanism for tasks such as importance weig

tutorialsarxiv-cs-cv
2 Jun 2026
Research

CloSE: A Geometric Shape-Agnostic Cloth State Representation

DGX agent

arXiv:2504.05033v3 Announce Type: replace-cross Abstract: Cloth manipulation is a difficult problem mainly because of the non-rigid nature of cloth, which makes a good representation of deformation es

researcharxiv-cs-cv
2 Jun 2026
Safety

Closing the Alignment-Maturity Gap in Federated Prototype Learning

DGX agent

arXiv:2606.02172v1 Announce Type: cross Abstract: Learning discriminative visual representations from distributed, heterogeneous data is a fundamental challenge in Federated Learning (FL). Prototype-b

safetyarxiv-cs-cv
2 Jun 2026
Hardware

Cohort-Scale Neural Atlases of Ultrasound Video

DGX agent

arXiv:2606.00890v1 Announce Type: new Abstract: Ultrasound is the most widely used real-time imaging modality in clinical practice, yet per-frame video annotation remains a major bottleneck: expert la

hardwarearxiv-cs-cv
2 Jun 2026
Safety

COLLAR: Cascaded Object-Level Latent Refinement for High-Fidelity Conditional Generation

DGX agent

arXiv:2606.00954v1 Announce Type: new Abstract: Achieving high-fidelity object-level control in Diffusion Transformers remains a significant challenge despite the introduction of structural priors lik

safetyarxiv-cs-cv
2 Jun 2026
Applications

Conditional Collapse in Sign Language Production: A Diagnostic and a Scaling Argument

DGX agent

arXiv:2606.01643v1 Announce Type: new Abstract: Sign Language Production (SLP) is the task of generating avatar sign language motion from natural language text. The quality of the generated motion is

applicationsarxiv-cs-cv
2 Jun 2026
Applications

Contrastive Augmented Transformer with Domain-specific Enhancement for Robust Multi-scenario Metal Surface Defect Detection

DGX agent

arXiv:2606.01962v1 Announce Type: new Abstract: Metal surface defect detection is critical for maintaining product quality in industrial manufacturing. However, it faces significant challenges, includ

applicationsarxiv-cs-cv
2 Jun 2026
Research

Contrastive meta-domain adaptation for robust skin lesion classification across clinical and acquisition conditions

DGX agent

arXiv:2602.19857v2 Announce Type: replace Abstract: Deep learning models for dermatological image analysis remain sensitive to acquisition variability and domain-specific visual characteristics, leadi

researcharxiv-cs-cv
2 Jun 2026
Research

CORE-MTL: Rethinking Gradient Balancing via Causal Orthogonal Representations

DGX agent

arXiv:2606.02221v1 Announce Type: new Abstract: Multi-task learning (MTL) aims to construct a joint model for multiple tasks by sharing a common representation across domains. To achieve this goal, ex

researcharxiv-cs-cv
2 Jun 2026
Local Ai

CoSTL: Comprehensive Spatial-Temporal Representation Learning for Moment Retrieval and Highlight Detection

DGX agent

arXiv:2606.01149v1 Announce Type: new Abstract: Video Moment Retrieval (MR) and Highlight Detection (HD) are crucial tasks in video analysis that aim to localize specific moments and estimate clip-wis

local-aiarxiv-cs-cv
2 Jun 2026
Research

Counterfactual Intervention Feature Transfer for Visible-Infrared Person Re-identification

DGX agent

arXiv:2208.00967v4 Announce Type: replace Abstract: Graph-based models have achieved great success in person re-identification tasks recently, which compute the graph topology structure (affinities) a

researcharxiv-cs-cv
2 Jun 2026
Agents

CountGD++: Generalized Prompting for Open-World Counting

DGX agent

arXiv:2512.23351v2 Announce Type: replace Abstract: The flexibility and accuracy of methods for automatically counting objects in images and videos are limited by the way the object can be specified.

agentsarxiv-cs-cv
2 Jun 2026
Safety

CR-JEPA: Cross-Modal Joint-Embedding Predictive Learning for Remote Sensing Image Retrieval

DGX agent

arXiv:2606.00706v1 Announce Type: new Abstract: Cross-modal remote sensing image retrieval aims to retrieve semantically related scenes across heterogeneous sensing modalities. This remains challengin

safetyarxiv-cs-cv
2 Jun 2026
Safety

Cross-Domain Dead Tree Detection via Knowledge Distillation in Aerial Imagery

DGX agent

arXiv:2606.02303v1 Announce Type: new Abstract: Detecting dead trees in aerial imagery is vital for assessing forest health, especially as tree mortality increases globally due to climate change, but

safetyarxiv-cs-cv
2 Jun 2026
Research

Cross-Domain Few-Shot Segmentation via Multi-view Progressive Adaptation

DGX agent

arXiv:2602.05217v2 Announce Type: replace Abstract: Cross-Domain Few-Shot Segmentation aims to segment categories in data-scarce domains conditioned on a few exemplars. Typical methods first establish

researcharxiv-cs-cv
2 Jun 2026
Research

DeblurNVS: Geometric Latent Diffusion for Novel View Synthesis from Sparse Motion-Blurred Images

DGX agent

arXiv:2606.01315v1 Announce Type: new Abstract: Novel view synthesis (NVS) is a fundamental problem in computer vision and graphics. Recent advances in neural radiance fields (NeRF), 3D Gaussian Splat

researcharxiv-cs-cv
2 Jun 2026
Research

Decoupled Residual Denoising Diffusion Models for Unified and Data Efficient Image-to-Image Translation

DGX agent

arXiv:2606.01048v1 Announce Type: new Abstract: We propose Decoupled Residual Denoising Diffusion models (DRDD) for unified and data-efficient image-to-image (I2I) translation. While diffusion models

researcharxiv-cs-cv
2 Jun 2026
Research

Deep Learning for Generating Computational PIN-4 Immunohistochemistry Staining from Prostate Biopsy H&E Images

DGX agent

arXiv:2606.01871v1 Announce Type: new Abstract: Immunohistochemistry (IHC)is frequently used to resolve diagnostically ambiguous prostate cancer biopsy findings on hematoxylin and eosin (H&E)-stained

researcharxiv-cs-cv
2 Jun 2026
Tutorials

Deep Learning for Remote Sensing to Improve Flood Inundation Mapping

DGX agent

arXiv:2606.02310v1 Announce Type: new Abstract: Flooding is the most pervasive natural disaster worldwide. Timely and accurate flood inundation mapping are essential for informing disaster risk manage

tutorialsarxiv-cs-cv
2 Jun 2026
Research

DeepLatent: Think with Images via Parallel Latent Visual Reasoning

DGX agent

arXiv:2606.00562v1 Announce Type: new Abstract: The emerging paradigm of 'thinking with images' embeds visual states into intermediate reasoning steps, defining a new frontier for Vision-Language Mode

researcharxiv-cs-cv
2 Jun 2026
Research

DefocusTrackerAI -- A Generalized Framework for the Automatic Detection of Defocused Particle Images

DGX agent

arXiv:2606.00076v1 Announce Type: new Abstract: The present work introduces DefocusTrackerAI, a generalized deep-learning framework for the automatic detection and position estimation of defocused par

researcharxiv-cs-cv
2 Jun 2026
Model Releases

Deformable Wiener Filter for Future Video Coding

DGX agent

arXiv:2606.01576v1 Announce Type: new Abstract: In-loop filters have attracted increasing attention due to the remarkable noise-reduction capability in the hybrid video coding framework. However, the

model-releasesarxiv-cs-cv
2 Jun 2026
Research

Degradation-Aware Metric Prompting for Hyperspectral Image Restoration

DGX agent

arXiv:2512.20251v3 Announce Type: replace Abstract: Unified hyperspectral image (HSI) restoration aims to recover diverse degradations within a single model. However, current methods often rely on imp

researcharxiv-cs-cv
2 Jun 2026
Research

DENSER: Depth-Guided Ensemble with Staged EFA-GS Reconstruction for Soccer Novel View Synthesis

DGX agent

arXiv:2606.01419v1 Announce Type: new Abstract: We propose DENSER, a Depth-guided ENSemble with Staged EFA-GS Reconstruction for soccer novel view synthesis. DENSER extends EFA-GS with three key contr

researcharxiv-cs-cv
2 Jun 2026
Model Releases

Density-Aware Translation of Spurious Correlations in Zero-Shot VLMs

DGX agent

arXiv:2606.01710v1 Announce Type: new Abstract: Vision-Language models (VLMs), such as CLIP, achieve powerful zero-shot classification. However, their predictions remain sensitive to spurious correlat

model-releasesarxiv-cs-cv
2 Jun 2026
Local Ai

DerMAE: Improving skin lesion classification through conditioned latent diffusion and MAE distillation

DGX agent

arXiv:2602.19848v2 Announce Type: replace Abstract: Skin lesion classification datasets often suffer from severe class imbalance, with malignant cases significantly underrepresented, leading to biased

local-aiarxiv-cs-cv
2 Jun 2026
Research

Detecting Pen-In-Air States from Video: A Proof-of-Concept Toward Complementary Handwriting Analysis

DGX agent

arXiv:2606.02342v1 Announce Type: new Abstract: Dynamic aspects of handwriting are critical for assessing developmental disorders such as dysgraphia and are typically captured using digitizing tablets

researcharxiv-cs-cv
2 Jun 2026
Research

Diamonds in the Sky: Pareidolic Animals in Clouds

DGX agent

arXiv:2606.01361v1 Announce Type: new Abstract: People often see animal shapes in clouds, a phenomenon known as pareidolia. We propose an AI-based method that aims to predict which animals people are

researcharxiv-cs-cv
2 Jun 2026
Research

Differing Roles of Leisure and Productivity in GDP - A Machine Learning based comparative analysis of Germany and USA

DGX agent

arXiv:2606.01234v1 Announce Type: cross Abstract: The GDP of a country is modelled as the relative interaction between two agents - working hours, reflecting the social choice of a population, and Tot

researcharxiv-cs-cv
2 Jun 2026
Research

Diffusion Models for Hyperspectral Image Analysis: A Comprehensive Review

DGX agent

arXiv:2505.11158v4 Announce Type: replace-cross Abstract: Hyperspectral image (HSI) analysis plays a critical role in remote sensing, agriculture, and environmental monitoring. However, traditional me

researcharxiv-cs-cv
2 Jun 2026
Model Releases

DINO-GFSA: Geo-Localization via Semantic Gated Fusion and Mamba-based Sequential Aggregation

DGX agent

arXiv:2606.00784v1 Announce Type: new Abstract: Cross-view geo-localization (CVGL) is critical for Unmanned Aerial Vehicle (UAV) self-positioning and target localization in GNSS-denied environments. H

model-releasesarxiv-cs-cv
2 Jun 2026
Research

Directed Distance Fields for Constant-Time Ray Queries on Gaussian Splatting

DGX agent

arXiv:2606.00817v1 Announce Type: cross Abstract: 3D Gaussian Splatting (3DGS) renders new views of a scene in real time. Like every rasterizer, it answers only primary rays, the rays from the camera

researcharxiv-cs-cv
2 Jun 2026
Research

Discrete Diffusion VLA: Bringing Discrete Diffusion to Action Decoding in Vision-Language-Action Policies

DGX agent

arXiv:2508.20072v4 Announce Type: replace Abstract: Vision-Language-Action (VLA) models adapt large vision-language backbones to map images and instructions into robot actions. However, prevailing VLA

researcharxiv-cs-cv
2 Jun 2026
Model Releases

Disentanglement-Based Equivariant Learning for Compositional VQA

DGX agent

arXiv:2606.02168v1 Announce Type: new Abstract: Compositional visual question answering (VQA) represents a challenging yet fundamental task that requires models to comprehend novel combinations of pre

model-releasesarxiv-cs-cv
2 Jun 2026
Tutorials

Distortion-Aware Fusion of Statistical and Vision-Language Features for Blind Image Quality Assessment

DGX agent

arXiv:2606.02002v1 Announce Type: new Abstract: Blind image quality assessment (BIQA) aims to predict perceived image quality without access to a reference image. Classical natural scene statistics (N

tutorialsarxiv-cs-cv
2 Jun 2026
Research

Divide and Conquer: Reliable Multi-View Evidential Learning for Deepfake Detection

DGX agent

arXiv:2606.01885v1 Announce Type: new Abstract: With the evolution of generative models, deepfakes have achieved near-perfect semantic realism, leaving forensic traces only in subtle structural anomal

researcharxiv-cs-cv
2 Jun 2026
Agents

Domain Adaptation with a Single Vision-Language Embedding

DGX agent

arXiv:2410.21361v2 Announce Type: replace Abstract: Domain adaptation has been extensively investigated in computer vision but still requires access to target data at the training time, which might be

agentsarxiv-cs-cv
2 Jun 2026
Research

DPsurv: Dual-Prototype Evidential Fusion for Uncertainty-Aware and Interpretable Whole-Slide Image Survival Prediction

DGX agent

arXiv:2510.00053v2 Announce Type: replace-cross Abstract: Pathology whole-slide images (WSIs) are widely used for cancer survival analysis because of their comprehensive histopathological information

researcharxiv-cs-cv
2 Jun 2026
Safety

Drifting Preference Optimization for One-Step Generative Models

DGX agent

arXiv:2606.02521v1 Announce Type: cross Abstract: One-step text-to-image generators are attractive for deployment because they generate an image with a single forward pass, but preference finetuning t

safetyarxiv-cs-cv
2 Jun 2026
Research

Dual-Route Top-K Retrieval with 1v1 VLM Reranking for the CoVR-R

DGX agent

arXiv:2606.01097v1 Announce Type: new Abstract: We describe Dual-Route Top-K Retrieval with 1v1 VLM Reranking for the CoVR-R challenge. The method treats composed video retrieval as two coupled proble

researcharxiv-cs-cv
2 Jun 2026
← Previous
1…125126127128129…263
Next →