AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,515 results
Research

Unified Video Dense Prediction from Disjoint Data

DGX agent

arXiv:2607.21592v1 Announce Type: new Abstract: Scene understanding requires simultaneous prediction about geometry, appearance, and semantics. However, existing task-specific annotations are fragment

researcharxiv-cs-cv
24 Jul 2026
Research

Unsupervised Metal Artifact Reduction in Dental CBCT using Fine-tuned Cycle-Consistent Adversarial Networks

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2607.20977v1 Announce Type: new Abstract: Metal artifacts generated by dental implants significantly degrade cone-beam computed tomography (CBCT) volumes, obscuring critical anatomical structure

researcharxiv-cs-cv
24 Jul 2026
Model Releases

ViSTR-Bench: Can MLLMs Reason from Continuous Visual Cues in Dynamic Scenes?

DGX agent

arXiv:2607.20868v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable success across diverse expert-level tasks, but they still struggle with fundamental ab

model-releasesarxiv-cs-cv
24 Jul 2026
Research

WAT3R: Feedforward Underwater 3D Reconstruction

DGX agent

arXiv:2607.21023v1 Announce Type: new Abstract: Reliable feedforward underwater 3D reconstruction remains challenging due to severe light attenuation and backscattering, which degrade visual quality a

researcharxiv-cs-cv
24 Jul 2026
Model Releases

Webly Supervised Multi-Label Recognition: Evaluation Benchmark and Dual-Branch Multi-Label Contrastive Learning

DGX agent

arXiv:2607.20874v1 Announce Type: new Abstract: Training deep learning models with freely available web images can reduce their dependence on costly manual annotations. Although webly supervised learn

model-releasesarxiv-cs-cv
24 Jul 2026
Model Releases

WhereEdit: Mask-aware Local Latent Editing for One-Step Image Editing

DGX agent

arXiv:2607.20883v1 Announce Type: new Abstract: Recent one-step text-to-image (T2I) models enable efficient image synthesis and provide new opportunities for real-time image editing. However, existing

model-releasesarxiv-cs-cv
24 Jul 2026
Model Releases

4DGS360: 360{eg} Gaussian Reconstruction of Dynamic Objects from a Single Video

DGX agent

arXiv:2603.21618v2 Announce Type: replace Abstract: We introduce 4DGS360, a diffusion-free framework for 360^{irc} dynamic object reconstruction from casual monocular video. Existing methods often fai

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

A Unified Tokenization Framework for Pain Recognition using Heterogeneous 3D Modalities

DGX agent

arXiv:2607.19716v1 Announce Type: new Abstract: Pain is a complex and pervasive phenomenon affecting a large percentage of the population, and accurate assessment is essential for effective clinical m

model-releasesarxiv-cs-cv
23 Jul 2026
Research

A Unified Variational Framework for Deep Weakly Supervised Image Segmentation

DGX agent

arXiv:2607.19669v1 Announce Type: new Abstract: We propose a unified variational framework for image segmentation under sparse pixel-level supervision. Our method is based on a simplex-constrained Pot

researcharxiv-cs-cv
23 Jul 2026
Local Ai

ABot-World-0: Infinite Interactive World Rollout on a Single Desktop GPU

DGX agent

arXiv:2607.19191v1 Announce Type: new Abstract: We present ABot-World-0, an action-conditioned video world model for real-time, long-horizon closed-loop interaction, supported by a multi-source data i

local-aiarxiv-cs-cv
23 Jul 2026
Research

Adaptive Visual Autoregressive Acceleration via Dual-Linkage Entropy Analysis

DGX agent

arXiv:2602.01345v2 Announce Type: replace Abstract: Visual AutoRegressive modeling (VAR) suffers from substantial computational cost due to the massive token count involved. Failing to account for the

researcharxiv-cs-cv
23 Jul 2026
Tutorials

Advancing Multimodal Fusion on Heterogeneous Medical Data with Hybrid Geometry Attention

DGX agent

arXiv:2607.19086v1 Announce Type: new Abstract: Multimodal fusion learning (MFL) has shown great potential in the medical domain, where we are faced with disparate data modalities such as imaging, cli

tutorialsarxiv-cs-cv
23 Jul 2026
Model Releases

Aligned Stable Inpainting: Mitigating Unwanted Object Insertion and Preserving Color Consistency

DGX agent

arXiv:2601.15368v3 Announce Type: replace Abstract: Generative image inpainting can produce realistic results even with large, irregular masks, but existing methods still suffer from two common proble

model-releasesarxiv-cs-cv
23 Jul 2026
Local Ai

An Exploratory Analysis of Pain Localization via Explainable Computational Modeling

DGX agent

arXiv:2607.19726v1 Announce Type: new Abstract: Automatic pain localization, which involves identifying the anatomical origin of pain from peripheral physiological signals without patient self-report,

local-aiarxiv-cs-cv
23 Jul 2026
Hardware

Anatomy-Aware 3D Mesh Refinement of Pericardium Segmentations on Computed Tomography

DGX agent

arXiv:2607.19210v1 Announce Type: new Abstract: Accurate delineation of the pericardium in a cardiac CT scan is essential for quantifying epicardial adipose tissue, yet it remains one of the most chal

hardwarearxiv-cs-cv
23 Jul 2026
Applications

AniGS: Bridging Rendering and Diffusion Prior for 3D Scene Animation

DGX agent

arXiv:2607.18539v1 Announce Type: new Abstract: Novel view rendering of large and complex reconstructed scenes is becoming increasingly photorealistic. However, most reconstructions remain static and

applicationsarxiv-cs-cv
23 Jul 2026
Local Ai

Appearance Pointers -- Multimodal Region Control of Diffusion Transformers

DGX agent

arXiv:2607.19344v1 Announce Type: new Abstract: Controllable image generation remains challenging for creative professionals, who often require precise regional control over materials, object identiti

local-aiarxiv-cs-cv
23 Jul 2026
Hardware

ATSplat: Compact Feed-forward 3D Gaussian Splatting with Adaptive Token Expansion

DGX agent

arXiv:2607.20417v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) achieves high-quality novel-view synthesis by optimizing freely placed primitives in 3D and adaptively densifying them in u

hardwarearxiv-cs-cv
23 Jul 2026
Research

Attention Without Grounding: Causal Evaluation of Visual Explanations in Medical VLMs

DGX agent

arXiv:2607.18577v1 Announce Type: new Abstract: Attention and saliency heatmaps are widely used to explain medical Vision-Language Model (VLM) outputs on chest X-rays, yet whether they truly highlight

researcharxiv-cs-cv
23 Jul 2026
Research

Attributes Should Come from Images, Not Class Names: Distribution-Conditioned Attribute Selection for Vision-Language Models

DGX agent

arXiv:2607.18695v1 Announce Type: new Abstract: A popular route to interpretable zero-shot classification asks a large language model (LLM) to describe each class name and prompts CLIP with the result

researcharxiv-cs-cv
23 Jul 2026
Research

Benchmarking Deep Learning Approaches for AEC Engineering Drawing Layout Detection and Information Extraction

DGX agent

arXiv:2607.18997v1 Announce Type: new Abstract: Information Extraction (IE) from Architecture, Engineering, and Construction (AEC) drawings remains hindered by manual inefficiency, while Layout Detect

researcharxiv-cs-cv
23 Jul 2026
Applications

BLUE: Semantics-Preserving Video Compression for Efficient Vision-Language Surveillance Analytics

DGX agent

arXiv:2607.19515v1 Announce Type: cross Abstract: Continuous surveillance video creates a growing storage, transmission, and inference burden for enterprise video analytics systems. While modern codec

applicationsarxiv-cs-cv
23 Jul 2026
Research

Bounding Boxes to Improve Small Language Model Performance on Vision-Based Grading Tasks

DGX agent

arXiv:2607.18767v1 Announce Type: new Abstract: The deployment of Small Language Models (SLMs) in educational settings offers significant advantages in terms of privacy, cost, and scalability. However

researcharxiv-cs-cv
23 Jul 2026
Local Ai

Brewing Stronger Features: Dual-Teacher Distillation for Multispectral Earth Observation

DGX agent

arXiv:2602.19863v3 Announce Type: replace Abstract: Foundation models are transforming Earth Observation (EO), yet the diversity of EO sensors and modalities makes a single universal model unrealistic

local-aiarxiv-cs-cv
23 Jul 2026
Research

CANDOR: Chance-Calibrated Discordance in Frozen Foundation Encoders

DGX agent

arXiv:2607.18451v1 Announce Type: cross Abstract: Frozen encoders are chosen by how well a lightweight head reads a finding from their features, not whether the geometry separates it. Nearest-neighbor

researcharxiv-cs-cv
23 Jul 2026
Research

CGMap: A Geospatially Aware Deep Learning Framework for Crop Gap Mapping Using UAV

DGX agent

arXiv:2607.18779v1 Announce Type: new Abstract: In India, crop germination is primarily monitored by visual inspection and manual counting, which are prone to errors, despite their crucial role in det

researcharxiv-cs-cv
23 Jul 2026
Research

ChronoStitch: Training-Free Composition of Visual KV Memories for Long-Horizon Temporal Reasoning

DGX agent

arXiv:2607.19547v1 Announce Type: new Abstract: Long-video question answering requires a model to preserve visual evidence over time without repeatedly reprocessing the same video. A practical approac

researcharxiv-cs-cv
23 Jul 2026
Local Ai

CNS-Edit++: Category-Agnostic 3D Editing with Coupled Neural Shape Representation

DGX agent

arXiv:2607.16577v2 Announce Type: replace Abstract: This paper presents a latent-space 3D shape editing framework built upon a coupled neural shape (CNS) representation and a neural feature volume opt

local-aiarxiv-cs-cv
23 Jul 2026
Safety

Cognitive Dual-Process Planning for Autonomous Driving with Structured Scene Knowledge and Verifiable Reasoning-Action Consistency

DGX agent

arXiv:2607.19194v2 Announce Type: cross Abstract: High-level planning for autonomous driving is a knowledge-intensive engineering decision task that requires accurate scene understanding, timely infer

safetyarxiv-cs-cv
23 Jul 2026
Agents

CoGoal3D: Collaborative 3D Object Detection with 3D-Aware Fusion and Refinement

DGX agent

arXiv:2607.19036v1 Announce Type: new Abstract: V2X collaborative object detection features overcoming the limitations of single-vehicle systems by aggregating environmental features from multiple col

agentsarxiv-cs-cv
23 Jul 2026
Safety

Confidence-Gated Vision-Only Heading Alignment for UAV-UGV Cooperative Systems

DGX agent

arXiv:2607.18713v1 Announce Type: cross Abstract: Vision-based heading prediction is useful for UAV--UGV cooperation, but accurate prediction alone does not guarantee that every predicted heading shou

safetyarxiv-cs-cv
23 Jul 2026
Research

Context-structured Video Anomaly Detection with Large Vision-Language Models

DGX agent

arXiv:2607.19077v1 Announce Type: new Abstract: Training video anomaly detectors is challenging due to the difficulty and cost of annotating diverse and rare abnormal events. Although recent large vis

researcharxiv-cs-cv
23 Jul 2026
Model Releases

Continual Video-MLLM Adaptation over Evolving Domains

DGX agent

arXiv:2607.18716v1 Announce Type: new Abstract: Video multimodal large language models have shown strong capability in video understanding, yet their adaptation to sequentially evolving domains remain

model-releasesarxiv-cs-cv
23 Jul 2026
Safety

Contour Errors: Ego-Centric Matching for 3D Multi-Object Tracking Performance Evaluation

DGX agent

arXiv:2506.04122v4 Announce Type: replace Abstract: Open-loop performance evaluation of 3D multi-object tracking in autonomous driving requires matching criteria that effectively penalize translationa

safetyarxiv-cs-cv
23 Jul 2026
Safety

Contrastive On-Policy Distillation

DGX agent

arXiv:2607.19046v1 Announce Type: new Abstract: On-policy Distillation (OPD) supervises a student model on trajectories sampled from its own policy by minimizing the divergence between the output dist

safetyarxiv-cs-cv
23 Jul 2026
Model Releases

CR-Refiner: An Object-Centric Optimal Transport Reranker for Edit-Conditioned 3D Scene Retrieval

DGX agent

arXiv:2607.19115v1 Announce Type: new Abstract: Edit-conditioned 3D scene retrieval pairs a reference 3D room with a natural-language modification and retrieves rooms from a corpus that satisfy the ed

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

CreatiPoster: Towards Editable and Controllable Multi-Layer Graphic Design Generation

DGX agent

arXiv:2506.10890v2 Announce Type: replace Abstract: Graphic design plays a crucial role in both commercial and personal contexts, yet creating high-quality, editable, and aesthetically pleasing graphi

model-releasesarxiv-cs-cv
23 Jul 2026
Safety

Cross-Dataset Generalization in Breast MRI Tumor Classification via Class-Wise Dataset Mixing

DGX agent

arXiv:2607.18678v1 Announce Type: new Abstract: Breast MRI is highly sensitive for detecting breast tumors, but exams contain many slices and require substantial reading time. Deep learning models oft

safetyarxiv-cs-cv
23 Jul 2026
Model Releases

Cross-Modal UAV Object Tracking: State-Aware Representation Learning and A Unified Benchmark

DGX agent

arXiv:2607.18768v1 Announce Type: new Abstract: Unmanned Aerial Vehicle (UAV) object tracking has emerged as a popular research field with broad practical applications. Modern UAVs are increasingly eq

model-releasesarxiv-cs-cv
23 Jul 2026
Safety

Crowd4D: Scene-Aware Monocular 4D Crowd Reconstruction

DGX agent

arXiv:2607.19517v1 Announce Type: new Abstract: Recovering scene-consistent 4D crowd motion from monocular video in large-scale scenes remains challenging due to severe depth ambiguity and complex sce

safetyarxiv-cs-cv
23 Jul 2026
Model Releases

Current Injection Spiking Neural Network for Infrared and Visible Image Fusion

DGX agent

arXiv:2607.19879v1 Announce Type: new Abstract: Infrared and visible image fusion (IVIF) integrates the complementary information of two modalities into a single image with richer scene content. While

model-releasesarxiv-cs-cv
23 Jul 2026
Local Ai

Decoupled Pipeline with Proposal Reranking and Score Fusion for Positive-Unlabeled Marine Species Detection

DGX agent

arXiv:2607.18700v1 Announce Type: new Abstract: The FathomNetCLEF 2026 competition combines underwater object detection and fine-grained marine species classification under a positive-unlabeled evalua

local-aiarxiv-cs-cv
23 Jul 2026
Research

Deep Learning Estimation of Sex, Age, Height, and Weight from CT-derived Digitally Reconstructed Radiographs

DGX agent

arXiv:2607.18638v1 Announce Type: new Abstract: Purpose: To develop and validate a deep learning ensemble for estimating adult sex, age, height, and weight from coronal digitally reconstructed radiogr

researcharxiv-cs-cv
23 Jul 2026
Local Ai

DeforM: Reasoning-Guided Physics-Aware Video Generation via Spatial-Temporal Masking

DGX agent

arXiv:2607.18664v1 Announce Type: new Abstract: Video generation models achieve high visual quality but often struggle to generate physics-aware videos. Unlike rigid-body motion, which can be describe

local-aiarxiv-cs-cv
23 Jul 2026
Model Releases

Delineate Anything v2: A Global Foundation Model for Field Delineation

DGX agent

arXiv:2607.19069v1 Announce Type: new Abstract: Accurate agricultural field boundary delineation at large scale is a foundational task for food security, supply chain transparency, and carbon accounti

model-releasesarxiv-cs-cv
23 Jul 2026
Research

Denoising Monte Carlo Renders with Diffusion Models

DGX agent

arXiv:2404.00491v3 Announce Type: replace Abstract: Physically-based renderings contain Monte Carlo noise, with variance that increases as the number of rays per pixel decreases. This noise, while zer

researcharxiv-cs-cv
23 Jul 2026
Research

Detect Early, Escalate Rarely: Anytime Detection of AI-Generated Video from the Compressed Bitstream

DGX agent

arXiv:2607.19476v1 Announce Type: new Abstract: Detectors for AI-generated video are evaluated offline. A clip is decoded to pixels and scored once, increasingly by a large vision-language model. Dete

researcharxiv-cs-cv
23 Jul 2026
Local Ai

Development of an automated, reliable, and clinically meaningful artificial intelligence (AI) tool for diagnosing cardiac disease from conventional cardiovascular magnetic resonance (CMR) images

DGX agent

arXiv:2607.20087v1 Announce Type: new Abstract: Aims: Cardiovascular magnetic resonance (CMR) imaging enables non-invasive assessment of myocardial structure, function, and pathology, but requires sub

local-aiarxiv-cs-cv
23 Jul 2026
← Previous
1…4445464748…261
Next →