AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlog
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,618 results
Model Releases

ComplexMimic: Human-Scene Interaction Imitation in Complex 3D Environments

DGX agent

arXiv:2607.02034v2 Announce Type: replace Abstract: Physics-based Human-Scene Interaction (HSI) imitation learning is crucial for embodied intelligence as it bridges the gap between kinematic 3D motio

model-releasesarxiv-cs-cv
7 Jul 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Research

Compositional Generalization Requires Linear, Orthogonal Representations in Vision Embedding Models

DGX agent

arXiv:2602.24264v2 Announce Type: replace Abstract: Compositional generalization, the ability to recognize familiar parts in novel contexts, is a defining property of intelligent systems. Although mod

researcharxiv-cs-cv
7 Jul 2026
Research

CompressedVQA-AEV: Full-Reference and No-Reference Quality Assessment Models for Asymmetric Encoded Videos

DGX agent

arXiv:2607.04606v1 Announce Type: cross Abstract: This report presents our solutions to the QoMEX 2026 Grand Challenge on Video Quality Assessment for Asymmetric Encoded Videos, comprising a full-refe

researcharxiv-cs-cv
7 Jul 2026
Research

Consistent and Editable: A Balanced Framework for Text-Guided Video Editing

DGX agent

arXiv:2607.05056v1 Announce Type: new Abstract: Recently, diffusion models have achieved considerable success in the text-guided video editing domain. However, existing works often struggle to balance

researcharxiv-cs-cv
7 Jul 2026
Model Releases

Continual Model Merging with Test-Time Adaptation for Whole-Slide Image Analysis

DGX agent

arXiv:2607.04755v1 Announce Type: new Abstract: Model merging offers a practical alternative to conventional continual learning by integrating independently fine-tuned models without retaining previou

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

ContiStain: Cross-Domain Relation-Preserving Distillation for Continual Multi-Domain Virtual IHC Staining

DGX agent

arXiv:2607.03851v1 Announce Type: new Abstract: A unified multiplex virtual staining model enables scalable and non-destructive multiplex analysis from H&E slides while promoting parameter efficiency,

model-releasesarxiv-cs-cv
7 Jul 2026
Research

ControlHair: Synergizing Physics Simulator and Video Diffusion for Controllable Dynamic Hair Rendering

DGX agent

arXiv:2509.21541v3 Announce Type: replace-cross Abstract: Hair simulation and rendering are challenging due to complex strand dynamics, diverse material properties, and intricate light-hair interactio

researcharxiv-cs-cv
7 Jul 2026
Research

Conversational Human Audio-visual Talking Dialogue Generation

DGX agent

arXiv:2607.02799v1 Announce Type: new Abstract: Large-scale dyadic interactive audio-visual dialogue (DIAD) datasets provide fundamental data resources for developing humanoid interactive virtual agen

researcharxiv-cs-cv
7 Jul 2026
Research

Coordinate Singularities Break Conformal Coverage for Gaze and Head Pose

DGX agent

arXiv:2607.02565v1 Announce Type: new Abstract: Conformal prediction provides distribution-free reliability guarantees for vision systems, but these guarantees depend on how prediction errors are meas

researcharxiv-cs-cv
7 Jul 2026
Local Ai

CORA: Generalizable coronary artery disease assessment and risk stratification from coronary CT angiography using pathology-centric representation learning

DGX agent

arXiv:2603.24847v2 Announce Type: replace Abstract: Coronary artery disease, a leading cause of cardiovascular mortality worldwide, can be assessed non-invasively by coronary computed tomography angio

local-aiarxiv-cs-cv
7 Jul 2026
Hardware

CPR: Chained Perceptual Refinement for Coarse-to-Fine Medical Image Classification

DGX agent

arXiv:2607.02591v1 Announce Type: new Abstract: High resolution medical images contain fine grained, spatially sparse cues that are critical for diagnosis, yet preserving full resolution incurs substa

hardwarearxiv-cs-cv
7 Jul 2026
Applications

Cross-device Collaborative Test-time Adaptation with Zeroth-order Optimization and Model Merging

DGX agent

arXiv:2607.02988v1 Announce Type: new Abstract: Test-time adaptation (TTA) mitigates domain shifts by using incoming test data to update a model on the fly. The majority of TTA methods require resourc

applicationsarxiv-cs-cv
7 Jul 2026
Research

Cross-Modal Fusion of OCT and OCT angiography enface for Improved Diagnostics of Diabetic Retinopathy

DGX agent

arXiv:2607.03959v1 Announce Type: cross Abstract: Diabetic retinopathy (DR) is a leading cause of vision impairment worldwide, highlighting the need for accurate and accessible screening tools. Optica

researcharxiv-cs-cv
7 Jul 2026
Model Releases

CTForensics: A Comprehensive Dataset and Method for AI-Generated CT Image Detection

DGX agent

arXiv:2603.01878v2 Announce Type: replace Abstract: Recent advances in generative AI have made synthetic Computed Tomography (CT) images increasingly realistic, enabling promising applications in medi

model-releasesarxiv-cs-cv
7 Jul 2026
Local Ai

Ctrl-Z Sampling: Scaling Diffusion Sampling with Controlled Random Zigzag Explorations

DGX agent

arXiv:2506.20294v5 Announce Type: replace Abstract: Diffusion models generate conditional samples by progressively denoising Gaussian noise, yet the denoising trajectory can stall at visually plausibl

local-aiarxiv-cs-cv
7 Jul 2026
Research

CURE: Controllable Unified Image Restoration for Complex Degradations

DGX agent

arXiv:2607.03044v1 Announce Type: new Abstract: The presence of composite degradations poses a significant challenge, since the underlying corruption factors exhibit complex and interdependent interac

researcharxiv-cs-cv
7 Jul 2026
Safety

CV-DCLR: Causal-Visual Dynamic Label Refinement for Robust Zero-Shot Learning

DGX agent

arXiv:2607.02601v1 Announce Type: new Abstract: Zero-Shot Learning (ZSL) facilitates knowledge transfer via shared semantic spaces. However, a critical bottleneck in this paradigm is Semantic Entangle

safetyarxiv-cs-cv
7 Jul 2026
Research

DC-Motion: Decoupling Structure and Details via Discrete-Continuous Tokens for Human Motion Generation

DGX agent

arXiv:2606.14721v2 Announce Type: replace-cross Abstract: Text-to-motion generation requires modeling both global action structure and fine-grained motion dynamics from natural language. Existing appr

researcharxiv-cs-cv
7 Jul 2026
Model Releases

Deep Learning-Based Characterization of Detonation-Cell Size Distributions in Soot-Foil Records

DGX agent

arXiv:2607.03764v1 Announce Type: cross Abstract: The geometric size and regularity of detonation cells are key physical parameters for characterizing detonation waves. Traditional manual measurement

model-releasesarxiv-cs-cv
7 Jul 2026
Applications

Deep Learning for Semen Analysis in Male Infertility: Computer Vision, Multimodal Fusion, and Clinical Translation

DGX agent

arXiv:2607.05311v1 Announce Type: new Abstract: Male infertility contributes substantially to the global infertility burden, and sperm analysis remains central to diagnosis, treatment planning, and as

applicationsarxiv-cs-cv
7 Jul 2026
Safety

Defending from GeoLocalization through Adversarial Road Trips

DGX agent

arXiv:2607.03277v1 Announce Type: new Abstract: Retrieval-based image geolocalization has emerged as a powerful technique for determining the location of a query image by matching it against a large,

safetyarxiv-cs-cv
7 Jul 2026
Model Releases

Deform360: A Massive Multi-view Visuotactile Dataset for Deformable World Models

DGX agent

arXiv:2607.05390v1 Announce Type: cross Abstract: Predicting object dynamics (i.e., world modeling) is a fundamental challenge for robotic manipulation, and modeling deformable objects presents a part

model-releasesarxiv-cs-cv
7 Jul 2026
Safety

DeGenseGS: Geometrically and Semantically Decoupled Surgical Scene Understanding in 4D Gaussian Splatting

DGX agent

arXiv:2607.04761v1 Announce Type: new Abstract: Real-time, text-promptable 4D reconstruction is indispensable for autonomous surgical interaction. Severe misalignment between semantic meaning and phys

safetyarxiv-cs-cv
7 Jul 2026
Tutorials

DGSeg: Dynamic Gating of Semantic-Spatial Guided Predictions for Reasoning Segmentation

DGX agent

arXiv:2607.04779v1 Announce Type: new Abstract: Reasoning segmentation aims to predict pixel-wise masks for targets given complex language queries. Existing approaches leverage Multimodal Large Langua

tutorialsarxiv-cs-cv
7 Jul 2026
Safety

DiCE-CIR: Direct Composition Learning for Efficient Zero-Shot Composed Image Retrieval

DGX agent

arXiv:2607.04665v1 Announce Type: new Abstract: Zero-shot composed image retrieval (ZS-CIR) aims to retrieve a target image from a multimodal query consisting of a reference image and an edit text des

safetyarxiv-cs-cv
7 Jul 2026
Safety

DICT: Data Injection and Contrastive Trajectory Refinement for Conditional Image Generation with Diffusion Models

DGX agent

arXiv:2607.03899v1 Announce Type: new Abstract: Diffusion models have become a dominant paradigm for conditional image generation, yet existing approaches generally follow two directions: task-specifi

safetyarxiv-cs-cv
7 Jul 2026
Model Releases

Diffusion Models are Open-World Affordance Learners: Leveraging Generative Priors for 3D Affordance Learning

DGX agent

arXiv:2508.01651v2 Announce Type: replace Abstract: 3D affordance grounding aims to understand how diverse objects can be manipulated, making it a cornerstone of embodied interaction. However, prior w

model-releasesarxiv-cs-cv
7 Jul 2026
Research

Direct Time-of-Flight Measurement Accuracy Improvement With Perimeter-Gated SPADs

DGX agent

arXiv:2607.02546v1 Announce Type: cross Abstract: Direct time of flight (dToF) measurements are susceptible to errors because of system-level and circuit-level timing jitters. In addition, device-leve

researcharxiv-cs-cv
7 Jul 2026
Safety

Displacement Preserving Relational Distillation for Robust Medical Segmentation

DGX agent

arXiv:2607.04599v1 Announce Type: new Abstract: Accurate 3D medical segmentation is limited by anatomical variability and high computational costs. While knowledge distillation (KD) offers a route for

safetyarxiv-cs-cv
7 Jul 2026
Research

DistillH-Mamba: A Hypergraph-Mamba-Based Knowledge Distillation Model for Efficient Impact Fall Detection

DGX agent

arXiv:2607.03156v1 Announce Type: new Abstract: Falls among the elderly represent a significant public health concern due to their prevalence, consequences, and societal burden. While deep learning ha

researcharxiv-cs-cv
7 Jul 2026
Research

Distribution Matching Distillation Meets Reinforcement Learning

DGX agent

arXiv:2511.13649v5 Announce Type: replace Abstract: Distribution Matching Distillation (DMD) facilitates efficient inference by distilling multi-step diffusion models into few-step variants. Concurren

researcharxiv-cs-cv
7 Jul 2026
Safety

Diverse Normal Prototypes-Guided Contrastive Reconstruction for Medical Anomaly Detection

DGX agent

arXiv:2508.19573v2 Announce Type: replace Abstract: Anomaly detection in medical images is challenging due to limited annotations and the domain gap. Existing reconstruction-based methods often rely o

safetyarxiv-cs-cv
7 Jul 2026
Research

Diversity-aware View Partitioning for Scalable VGGT

DGX agent

arXiv:2607.01885v2 Announce Type: replace Abstract: Geometry transformers such as VGGT achieve strong performance by jointly reasoning over multiple views with global attention. However, scaling them

researcharxiv-cs-cv
7 Jul 2026
Model Releases

Do Diabetic Foot Ulcer Segmentation Models Generalize? A Cross-Dataset Benchmark of CNN and Transformer Architectures

DGX agent

arXiv:2607.02555v1 Announce Type: new Abstract: Deep learning models for diabetic foot ulcer (DFU) segmentation routinely report high accuracy, but they are almost always trained and tested on the sam

model-releasesarxiv-cs-cv
7 Jul 2026
Research

Do Flat Minima Improve Sparse Novel View Synthesis?

DGX agent

arXiv:2511.17918v2 Announce Type: replace Abstract: Despite the success of recent novel view synthesis methods, they tend to struggle in sparse-view settings. This poor generalization to unseen viewpo

researcharxiv-cs-cv
7 Jul 2026
Model Releases

Do Medical Vision Language Models Actually See? A Counterfactual Grounding Framework and Hard-Negative Contrastive Training for Visually-Reliant Medical VLMs

DGX agent

arXiv:2607.03647v1 Announce Type: new Abstract: Large vision language models (VLMs) report strong accuracy on medical question-answering, yet it remains unclear whether they reason from visual evidenc

model-releasesarxiv-cs-cv
7 Jul 2026
Research

DriftST: One-Step Generative Inference of Spatial Transcriptomics from H&E Histology

DGX agent

arXiv:2607.04740v1 Announce Type: new Abstract: Spatial Transcriptomics (ST) measures gene expression while preserving spatial context, but its high cost and low throughput leave public datasets small

researcharxiv-cs-cv
7 Jul 2026
Applications

DS-SAC: Density Search for Sample Consensus

DGX agent

arXiv:2607.03972v1 Announce Type: new Abstract: Robust geometric model estimation is a fundamental problem in computer vision. RANSAC and its variants remain widely used for this task; however, they r

applicationsarxiv-cs-cv
7 Jul 2026
Model Releases

Dual-Adaptive SAM3: Hierarchical Routing over Low-Rank Expert Layers for Parameter-Efficient Medical Image Segmentation

DGX agent

arXiv:2607.02571v1 Announce Type: new Abstract: The Segment Anything Model with Concepts (SAM3) heralds a new paradigm for open-vocabulary segmentation through natural language interaction, offering s

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

DynaWM: A Base-VLA-Guided World Foundation Model for Moving-Object Manipulation

DGX agent

arXiv:2607.02604v1 Announce Type: new Abstract: Although vision-language-action (VLA) models have received widespread attention, many challenges remain in manipulating dynamic moving objects. In most

model-releasesarxiv-cs-cv
7 Jul 2026
Research

E-TraMamba: A New Paradigm for Efficient Long-Term 3D Feature Tracking with Event Cameras

DGX agent

arXiv:2607.02866v1 Announce Type: new Abstract: Event-based 3D tracking enables low-latency and high-speed perception, while existing CNN- and Transformer-based trackers struggle to capture long-range

researcharxiv-cs-cv
7 Jul 2026
Model Releases

EgoInertia-MI: A Multimodal Egocentric Vision and IMU Benchmark for Motor Impairment Assessment

DGX agent

arXiv:2607.03934v1 Announce Type: new Abstract: Motor impairments, including tremor, bradykinesia, gait abnormalities, and postural instability, are common across many neurological and movement-relate

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

EM3M: An Electron Micrograph Dataset for Microstructural Segmentation and Generation

DGX agent

arXiv:2508.16239v2 Announce Type: replace Abstract: Quantitative microstructural characterization is fundamental to materials science, and electron micrographs (EMs) provide indispensable high-resolut

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

EmoteGPT: 3D Human Facial Expressions from Natural Language Descriptions

DGX agent

arXiv:2607.02674v1 Announce Type: new Abstract: Precise control of 3D facial expressions from text is crucial for virtual avatars, animation, and human-computer interaction, yet existing text-to-3D me

model-releasesarxiv-cs-cv
7 Jul 2026
Tutorials

EMPURPLE: A Free Lunch for Diffusion Distillation based on the Information Bottleneck

DGX agent

arXiv:2607.04276v1 Announce Type: new Abstract: Diffusion models achieve impressive image-generation quality but remain expensive at inference time. Diffusion distillation reduces sampling steps, yet

tutorialsarxiv-cs-cv
7 Jul 2026
Safety

Enhancing Facial Expression Recognition in Head-Mounted Displays with Synthetic Data

DGX agent

arXiv:2607.04490v1 Announce Type: new Abstract: Facial expression recognition (FER) is crucial for social interaction in mixed reality environments that employ head-mounted displays (HMD). However, co

safetyarxiv-cs-cv
7 Jul 2026
Local Ai

Enhancing Large Multimodal Models in Key Information Extraction via Scene-Aware Document Synthesis

DGX agent

arXiv:2607.04636v1 Announce Type: new Abstract: Key Information Extraction (KIE) converts visually rich documents into structured data, but practical deployment remains challenging: strong performance

local-aiarxiv-cs-cv
7 Jul 2026
Safety

Enhancing Monocular 3D Hand Reconstruction with Learned Texture Priors

DGX agent

arXiv:2508.09629v2 Announce Type: replace Abstract: We revisit the role of texture in monocular 3D hand reconstruction, not as an afterthought for photorealism, but as a dense, spatially grounded cue

safetyarxiv-cs-cv
7 Jul 2026
← Previous
1…6364656667…263
Next →