AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,515 results
Safety

Who Gets Missed in the Tail? Thresholded Subgroup Underdiagnosis in Long-Tailed Chest X-ray Classification

DGX agent

arXiv:2607.07717v1 Announce Type: cross Abstract: In chest X-ray (CXR) classification, acceptable ranking performance can still leave rare-positive patients below threshold, especially within subgroup

safetyarxiv-cs-cv
10 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

XOV-Action: Towards Generalizable Open-Vocabulary Action Recognition

DGX agent

arXiv:2403.01560v3 Announce Type: replace Abstract: Inspired by the impressive success of image-text foundation models, recent works have proposed to adapt these foundation models to video data, leadi

model-releasesarxiv-cs-cv
10 Jul 2026
Research

ZipDepth: Bringing Lightweight Zero-Shot Monocular Depth Anywhere, on Any Device

DGX agent

arXiv:2607.08771v1 Announce Type: new Abstract: Monocular depth estimation has seen remarkable progress through foundation models achieving robust zero-shot generalization, yet their computational dem

researcharxiv-cs-cv
10 Jul 2026
Safety

Zoom-IQA: Image Quality Assessment with Reliable Region-Aware Reasoning

DGX agent

arXiv:2601.02918v3 Announce Type: replace Abstract: Image Quality Assessment (IQA) is a long-standing problem in computer vision. Previous methods typically focus on predicting numerical scores withou

safetyarxiv-cs-cv
10 Jul 2026
Research

A Good Initialization is All You Need for Faithful Visual Attribution

DGX agent

arXiv:2607.06726v1 Announce Type: new Abstract: Faithful visual attribution identifies which image regions support a model prediction. Search-based perturbation methods lead the insertion--deletion fa

researcharxiv-cs-cv
9 Jul 2026
Tutorials

A Theory of Contrastive Learning with Natural Images

DGX agent

arXiv:2607.07470v1 Announce Type: new Abstract: Why does contrastive learning with simple images and augmentations yield useful representations for downstream tasks? We address this question by analyt

tutorialsarxiv-cs-cv
9 Jul 2026
Local Ai

AA-ViT: Anatomically Aware Vision Transformer with Structural and Frequency Guidance for Contrast Enhanced Brain MRI Synthesis

DGX agent

arXiv:2607.07553v1 Announce Type: new Abstract: Accurate tumour localization and diagnosis is a critical component of clinical care for brain cancers. Magnetic Resonance Imaging (MRI) is the most comm

local-aiarxiv-cs-cv
9 Jul 2026
Local Ai

Activation Quantization of Vision Encoders Needs Prefixing Registers

DGX agent

arXiv:2510.04547v5 Announce Type: replace-cross Abstract: Large pretrained vision encoders are central to multimodal intelligence, powering applications from on-device vision processing to vision-lang

local-aiarxiv-cs-cv
9 Jul 2026
Safety

AI for Cultural Heritage Textiles: Fine-Tuned Latent Diffusion for Novel Ulos Motif Synthesis

DGX agent

arXiv:2607.06590v1 Announce Type: new Abstract: Preserving and revitalising traditional textiles such as Ulos, a cultural heritage of the Batak ethnic group in North Sumatra, Indonesia, requires balan

safetyarxiv-cs-cv
9 Jul 2026
Tutorials

An Edge-aware Prompt-enhanced SAM for Ultrasound Image Segmentation

DGX agent

arXiv:2607.07240v1 Announce Type: new Abstract: Ultrasound image segmentation is essential for delineating anatomical structures and lesions, providing the foundation for accurate diagnosis. While the

tutorialsarxiv-cs-cv
9 Jul 2026
Model Releases

ASFR-Net: Adversarial Alignment and Spatio-Frequency Refinement Network for Heterogeneous Remote Sensing Image Change Detection

DGX agent

arXiv:2607.07161v1 Announce Type: new Abstract: The core challenge of heterogeneous change detection in remote sensing imagery lies in effectively decoupling genuine land-cover changes from significan

model-releasesarxiv-cs-cv
9 Jul 2026
Safety

`Attention-Guided Cross-Temporal Clustering for Self-Supervised Video Object Segmentation

DGX agent

arXiv:2607.07230v1 Announce Type: new Abstract: Video object segmentation (VOS) is a fundamental task in video understanding, requiring accurate delineation and consistent tracking of objects across f

safetyarxiv-cs-cv
9 Jul 2026
Research

Attention in Geometry: Scalable Spatial Modeling via Adaptive Density Fields and FAISS-Accelerated Kernels

DGX agent

arXiv:2601.06135v3 Announce Type: replace-cross Abstract: Spatial computation in geographic systems increasingly requires query-conditioned, local, interpretable aggregation under metric constraints.

researcharxiv-cs-cv
9 Jul 2026
Safety

Automatic Echocardiography Segmentation via Transition Probability Correlation for Stable Semantic Extraction

DGX agent

arXiv:2607.07580v1 Announce Type: new Abstract: While echocardiography is essential for cardiovascular diagnosis, inherent speckle noise and low signal-to-noise ratio often lead to ambiguous semantic

safetyarxiv-cs-cv
9 Jul 2026
Research

Bi-PT: Bidirectional Cross-Attention Point Transformers for Four-Chamber Heart Reconstruction from Sparse Cardiac MRI Data

DGX agent

arXiv:2607.06923v1 Announce Type: new Abstract: We propose Bi-PT, a pipeline for reconstructing 3D four-chamber human heart meshes from clinical sparsely sampled cardiac magnetic resonance imaging (CM

researcharxiv-cs-cv
9 Jul 2026
Research

BUS: Brain-Inspired Unsupervised Self-Reflection for Advanced Multimodal Reasoning

DGX agent

arXiv:2607.07361v1 Announce Type: new Abstract: Current Vision-Language Models (VLMs) often struggle to handle complex visual tasks that require consistent and fine-grained reasoning. Recent methods a

researcharxiv-cs-cv
9 Jul 2026
Research

Cardiac MRI Through-Plane Super-Resolution Guided by Reference and Memory

DGX agent

arXiv:2607.07581v1 Announce Type: new Abstract: Clinical cardiac MRI is commonly acquired with high in-plane resolution but coarse through-plane resolution to reduce scan time and accommodate breath-h

researcharxiv-cs-cv
9 Jul 2026
Research

CEVAR: Centerline Embedding Extraction for Endovascular Aneurysm Repair

DGX agent

arXiv:2606.15667v2 Announce Type: replace Abstract: Long-term mortality rates after endovascular aneurysm repair (EVAR) remain elevated due to post-EVAR rupture caused by loss of seal in stent graft s

researcharxiv-cs-cv
9 Jul 2026
Research

CoFINN: Conservation Flux Informed Neural Networks for Physics Problems Governed by Conservation Laws

DGX agent

arXiv:2607.06587v1 Announce Type: new Abstract: We present CoFINN (Conservation Flux Informed Neural Networks), a physics-informed deep learning framework for predicting compressible flow fields gover

researcharxiv-cs-cv
9 Jul 2026
Research

ColorFM: An Optimization-to-Learning Framework for Color Transfer via Flow Matching

DGX agent

arXiv:2607.07119v1 Announce Type: new Abstract: Color transfer aims to align the color distribution of a source image with that of a reference image while preserving structural and semantic consistenc

researcharxiv-cs-cv
9 Jul 2026
Agents

CoMind: Understanding Collaborative Human Activity from Multiple Minds and Views

DGX agent

arXiv:2607.06691v1 Announce Type: new Abstract: Human-human collaboration is a fundamental aspect of everyday life, essential to success in a wide range of goal-directed activities from household task

agentsarxiv-cs-cv
9 Jul 2026
Model Releases

Comparative Study of Domain-adapted VLMs for General Document Visual Question Answering

DGX agent

arXiv:2607.07179v1 Announce Type: new Abstract: Document Visual Question Answering (DocVQA) presents a complex multimodal challenge, requiring models to exploit visual, textual, and layout information

model-releasesarxiv-cs-cv
9 Jul 2026
Research

Compass: Prostate Cancer Detection Needs Multi-View Context

DGX agent

arXiv:2607.06919v1 Announce Type: new Abstract: Artificial intelligence (AI) analysis of micro-ultrasound (muUS) has shown promise for prostate cancer (PCa) detection. However, most existing AI method

researcharxiv-cs-cv
9 Jul 2026
Local Ai

Context-Aware Slum Mapping in Sub-Saharan Africa Using Sentinel-1 Texture and Local Climate Zones

DGX agent

arXiv:2607.07532v1 Announce Type: new Abstract: Accurate mapping of informal settlements remains a major challenge in Sub-Saharan African (SSA) cities because optical imagery often fails to distinguis

local-aiarxiv-cs-cv
9 Jul 2026
Research

CRIS: Cross-Plane Self-Supervised Isotropic Restoration for Anisotropic Volumetric Imaging Across Modalities

DGX agent

arXiv:2606.15967v2 Announce Type: replace Abstract: Anisotropic volumetric acquisitions are common in clinical MRI and volume electron microscopy (vEM), where sparse through-plane sampling creates thi

researcharxiv-cs-cv
9 Jul 2026
Research

DiffCVE: Diffusion-based Compressed Video Enhancement

DGX agent

arXiv:2607.07195v1 Announce Type: new Abstract: Perceptual quality enhancement of severely compressed videos remains challenging due to complex artifact patterns and substantial information loss. Rece

researcharxiv-cs-cv
9 Jul 2026
Safety

Discovering Geometric Biases in 3D Face Reconstruction: A Curvature-Aware Spectral Framework for Fairness Evaluation

DGX agent

arXiv:2607.07486v1 Announce Type: new Abstract: 3D Morphable Models (3DMMs) remain the standard parametric shape priors for many state-of-the-art 3D face reconstruction algorithms. However, as these m

safetyarxiv-cs-cv
9 Jul 2026
Safety

Dual Latent Memory in Vision-Language-Action Models for Robotic Manipulation

DGX agent

arXiv:2607.07608v1 Announce Type: cross Abstract: Mainstream Vision-Language-Action (VLA) models predict actions primarily from the current observation under a Markovian assumption, thus struggling wi

safetyarxiv-cs-cv
9 Jul 2026
Hardware

DYNA-PRUNER: Input-Adaptive Data-Model Co-Pruning for Efficient and Scalable Spatio-Temporal Media Prediction

DGX agent

arXiv:2606.15346v2 Announce Type: replace Abstract: Spatio-temporal prediction supports radar/satellite nowcasting and city-scale traffic monitoring, but modern models are often too expensive for real

hardwarearxiv-cs-cv
9 Jul 2026
Applications

Dynamic Object Detection and Tracking in Construction: A Fisheye Camera and LiDAR Sensor Fusion Model

DGX agent

arXiv:2607.06896v1 Announce Type: cross Abstract: Robust dynamic object detection and tracking are essential for enabling robots to operate safely and effectively alongside humans in complex environme

applicationsarxiv-cs-cv
9 Jul 2026
Research

ECHO: Ego-Centric modeling of Human-Object interactions

DGX agent

arXiv:2508.21556v3 Announce Type: replace Abstract: Modeling human-object interactions (HOI) from an egocentric perspective is a critical yet challenging task, particularly when relying on sparse sign

researcharxiv-cs-cv
9 Jul 2026
Research

EdgeCompress: Coupling Multidimensional Model Compression and Dynamic Inference for EdgeAI

DGX agent

arXiv:2607.06982v1 Announce Type: new Abstract: Convolutional neural networks (CNNs) have demonstrated encouraging results in image classification tasks. However, the prohibitive computational cost of

researcharxiv-cs-cv
9 Jul 2026
Local Ai

EditVerse3D: High-Quality 3D Object Editing with Region-Aware Learning

DGX agent

arXiv:2607.07187v1 Announce Type: new Abstract: Local editing of 3D objects remains a long-standing challenge. When interacting with 3D content, humans naturally tend to specify a coarse region of int

local-aiarxiv-cs-cv
9 Jul 2026
Model Releases

Ego-Human Motion Prediction with 3D-Aware LLM

DGX agent

arXiv:2607.07001v1 Announce Type: new Abstract: Anticipating human motion from an egocentric perspective is fundamental for proactive assistance in AR/VR, human-robot collaboration, and embodied AI. W

model-releasesarxiv-cs-cv
9 Jul 2026
Safety

EmbodiedGen V2: An Agentic, Simulation-Ready 3D World Engine for Embodied AI

DGX agent

arXiv:2607.07459v1 Announce Type: cross Abstract: We present EmbodiedGen V2, a generative 3D world engine for building executable sim-ready environments for embodied intelligence. Sim-ready 3D asset g

safetyarxiv-cs-cv
9 Jul 2026
Tutorials

Ensemble Deep Learning Approaches for AI-Altered Video Detection

DGX agent

arXiv:2607.06872v1 Announce Type: new Abstract: The increasing accessibility of artificial intelligence has led to a rapid rise in AI-generated videos, making it more difficult to distinguish between

tutorialsarxiv-cs-cv
9 Jul 2026
Research

EventVGGT: Exploring Cross-Modal Distillation for Consistent Event-based Depth Estimation

DGX agent

arXiv:2603.09385v2 Announce Type: replace Abstract: Event cameras offer superior sensitivity to high-speed motion and extreme lighting, making event-based monocular depth estimation a promising approa

researcharxiv-cs-cv
9 Jul 2026
Applications

Face-trace: Open-Set Attribution and Progressive Discovery of Synthetic Face Generators

DGX agent

arXiv:2607.07545v1 Announce Type: new Abstract: Recent advances in generative Artificial Intelligence have made synthetic face images increasingly realistic, creating new challenges for multimedia for

applicationsarxiv-cs-cv
9 Jul 2026
Model Releases

FMMC: Harnessing the Power of Foundation Models for Accurate Material Classification

DGX agent

arXiv:2603.17390v2 Announce Type: replace Abstract: Material classification has emerged as a critical task in computer vision and graphics, supporting the assignment of accurate material properties to

model-releasesarxiv-cs-cv
9 Jul 2026
Local Ai

Format-Controlled Multi-Scale JPEG Compression Response Analysis for Image-Level Forgery Screening

DGX agent

arXiv:2607.06615v1 Announce Type: cross Abstract: Image forgery detection is a critical task in digital forensics, yet many deep-learning localization approaches are typically GPU-accelerated and comp

local-aiarxiv-cs-cv
9 Jul 2026
Research

From Data Completeness to Data Sufficiency: A Task-Driven Imaging Framework for Intraoperative CBCT under Quality-Time-Dose Trade-offs

DGX agent

arXiv:2607.07039v1 Announce Type: cross Abstract: Mobile C-arm cone-beam computed tomography (CBCT) has been widely used for real-time intraoperative 3D imaging. However, current practice often mechan

researcharxiv-cs-cv
9 Jul 2026
Model Releases

From My View to Yours: Learning Egocentric Cues from Exocentric Video using Privileged Egocentric Supervision

DGX agent

arXiv:2501.05711v4 Announce Type: replace Abstract: Vision Language Models (VLMs) have achieved strong performance across a wide range of video understanding tasks. However, their viewpoint-invariant

model-releasesarxiv-cs-cv
9 Jul 2026
Research

G-PROBE: Cross-FOV Place Recognition and Certainty-Coupled Localization for 3D Point Clouds

DGX agent

arXiv:2607.06782v1 Announce Type: cross Abstract: Global localization from 3D point clouds remains challenging under limited or asymmetric fields of view (FOV), which fail to provide the dense, symmet

researcharxiv-cs-cv
9 Jul 2026
Applications

G-ZAP: A Generalizable Zero-Shot Framework for Arbitrary-Scale Pansharpening

DGX agent

arXiv:2603.14412v2 Announce Type: replace Abstract: Pansharpening aims to fuse a high-resolution panchromatic (PAN) image and a low-resolution multispectral (LRMS) image to produce a high-resolution m

applicationsarxiv-cs-cv
9 Jul 2026
Safety

Gen4U: Unifying Video Generation and Understanding via Diffusion

DGX agent

arXiv:2607.06856v1 Announce Type: new Abstract: Prior work suggests that diffusion representations capture low-level geometry but struggle with high-level semantics. We demonstrate that state-of-the-a

safetyarxiv-cs-cv
9 Jul 2026
Applications

General Incomplete Multimodal Learning via Dynamic Quality Perception

DGX agent

arXiv:2607.06943v1 Announce Type: new Abstract: Multimodal learning robust to missing modalities is essential for real-world applications. Existing methods mainly focus on inter-modality missing, wher

applicationsarxiv-cs-cv
9 Jul 2026
Research

Geometric Collapse: When Vision Models Fail to Verify Physical Causality

DGX agent

arXiv:2607.06871v1 Announce Type: new Abstract: Recent progress in large-scale self-supervised learning has improved dense geometric prediction, but it remains unclear whether such scaling yields infe

researcharxiv-cs-cv
9 Jul 2026
Research

Geometry-Aware Single-Image 4D Synthesis via Dense Trajectory Generation

DGX agent

arXiv:2512.05044v2 Announce Type: replace Abstract: Generating interactive and dynamic 4D scenes from a single static image remains a core challenge. Most existing generate-then-reconstruct and recons

researcharxiv-cs-cv
9 Jul 2026
← Previous
1…5556575859…261
Next →