AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,515 results
Model Releases

TripVVT: A Large-Scale Triplet Dataset and a Coarse-Mask Baseline for In-the-Wild Video Virtual Try-On

DGX agent

arXiv:2604.27958v1 Announce Type: new Abstract: Due to the scarcity of large-scale in-the-wild triplet data and the improper use of masks, the performance of video virtual try-on models remains limite

model-releasesarxiv-cs-cv
1 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Tutorials

UHR-Net: An Uncertainty-Aware Hypergraph Refinement Network for Medical Image Segmentation

DGX agent

arXiv:2604.28095v1 Announce Type: new Abstract: Accurate lesion segmentation is crucial for clinical diagnosis and treatment planning. However, lesions often resemble surrounding tissues and exhibit i

tutorialsarxiv-cs-cv
1 May 2026
Research

Uncertainty Quantification Framework for Aerial and UAV Photogrammetry through Error Propagation

DGX agent

arXiv:2507.13486v2 Announce Type: replace Abstract: Uncertainty quantification of the photogrammetry process is essential for providing per-point accuracy credentials of the point clouds. Unlike airbo

researcharxiv-cs-cv
1 May 2026
Agents

Understanding Adversarial Transferability in Vision-Language Models for Autonomous Driving: A Cross-Architecture Analysis

DGX agent

arXiv:2604.27414v1 Announce Type: new Abstract: Vision-language models (VLMs) are increasingly used in autonomous driving because they combine visual perception with language-based reasoning, supporti

agentsarxiv-cs-cv
1 May 2026
Research

Uni-HOI:A Unified framework for Learning the Joint distribution of Text and Human-Object Interaction

DGX agent

arXiv:2604.27491v1 Announce Type: new Abstract: Modeling 4D human-object interaction (HOI) is a compelling challenge in computer vision and an essential technology powering virtual and mixed-reality a

researcharxiv-cs-cv
1 May 2026
Research

Unsupervised Machine Learning for Osteoporosis Diagnosis Using Singh Index Clustering on Hip Radiographs

DGX agent

arXiv:2411.15253v2 Announce Type: replace-cross Abstract: Osteoporosis, a prevalent condition among the aging population worldwide, is characterized by diminished bone mass and altered bone structure,

researcharxiv-cs-cv
1 May 2026
Model Releases

VeraRetouch: A Lightweight Fully Differentiable Framework for Multi-Task Reasoning Photo Retouching

DGX agent

arXiv:2604.27375v1 Announce Type: new Abstract: Reasoning photo retouching has gained significant traction, requiring models to analyze image defects, give reasoning processes, and execute precise ret

model-releasesarxiv-cs-cv
1 May 2026
Local Ai

VerteNet -- A Multi-Context Hybrid CNN Transformer for Accurate Vertebral Landmark Localization in Lateral Spine DXA Images

DGX agent

arXiv:2502.02097v3 Announce Type: replace Abstract: This aims to develop and validate a deep learning model that can accurately locate vertebral landmarks in lateral spine Dual energy X-ray Absorptiom

local-aiarxiv-cs-cv
1 May 2026
Model Releases

Visual Generation in the New Era: An Evolution from Atomic Mapping to Agentic World Modeling

DGX agent

arXiv:2604.28185v1 Announce Type: new Abstract: Recent visual generation models have made major progress in photorealism, typography, instruction following, and interactive editing, yet they still str

model-releasesarxiv-cs-cv
1 May 2026
Research

VTBench: A Multimodal Framework for Time-Series Classification with Chart-Based Representations

DGX agent

arXiv:2604.27259v1 Announce Type: new Abstract: Time-series classification (TSC) has advanced significantly with deep learning, yet most models rely solely on raw numerical inputs, overlooking alterna

researcharxiv-cs-cv
1 May 2026
Applications

World2Minecraft: Occupancy-Driven Simulated Scenes Construction

DGX agent

arXiv:2604.27578v1 Announce Type: new Abstract: Embodied intelligence requires high-fidelity simulation environments to support perception and decision-making, yet existing platforms often suffer from

applicationsarxiv-cs-cv
1 May 2026
Research

YOSE: You Only Select Essential Tokens for Efficient DiT-based Video Object Removal

DGX agent

arXiv:2604.27322v1 Announce Type: new Abstract: Recent advances in Diffusion Transformer (DiT)-based video generation technologies have shown impressive results for video object removal. However, thes

researcharxiv-cs-cv
1 May 2026
Agents

3D Generation for Embodied AI and Robotic Simulation: A Survey

DGX agent

arXiv:2604.26509v1 Announce Type: cross Abstract: Embodied AI and robotic systems increasingly depend on scalable, diverse, and physically grounded 3D content for simulation-based training and real-wo

agentsarxiv-cs-cv
30 Apr 2026
Model Releases

3D-LENS: A 3D Lifting-based Elevated Novel-view Synthesis method for Single-View Aerial-Ground Re-Identification

DGX agent

arXiv:2604.26520v1 Announce Type: new Abstract: Aerial-Ground Re-Identification (AG-ReID) is constrained by the viewpoint-domain gap, as drastic viewpoint disparities occlude or distort discriminative

model-releasesarxiv-cs-cv
30 Apr 2026
Research

A Diffeomorphism Groupoid and Algebroid Framework for Discontinuous Image Registration

DGX agent

arXiv:2603.11806v2 Announce Type: replace-cross Abstract: In this paper, we propose a novel mathematical framework for piecewise diffeomorphic image registration that involves discontinuous sliding mo

researcharxiv-cs-cv
30 Apr 2026
Research

A Multimodal Depth-Aware Method For Embodied Reference Understanding

DGX agent

arXiv:2510.08278v3 Announce Type: replace Abstract: Embodied Reference Understanding requires identifying a target object in a visual scene based on both language instructions and pointing cues. While

researcharxiv-cs-cv
30 Apr 2026
Safety

A Multimodal Pre-trained Network for Integrated EEG-Video Seizure Detection

DGX agent

arXiv:2604.26379v1 Announce Type: new Abstract: Reliable seizure detection in mouse models is essential for preclinical epilepsy research, yet manual review of synchronized video-EEG recordings is lab

safetyarxiv-cs-cv
30 Apr 2026
Model Releases

A Multistage Extraction Pipeline for Long Scanned Financial Documents: An Empirical Study in Industrial KYC Workflows

DGX agent

arXiv:2604.26462v1 Announce Type: new Abstract: Structured information extraction from long, multilingual scanned financial documents is a core requirement in industrial KYC and compliance workflows.

model-releasesarxiv-cs-cv
30 Apr 2026
Local Ai

Action Hints: Semantic Typicality and Context Uniqueness for Generalizable Skeleton-based Video Anomaly Detection

DGX agent

arXiv:2509.11058v2 Announce Type: replace Abstract: Zero-Shot Video Anomaly Detection (ZS-VAD) requires temporally localizing anomalies without target domain training data, which is a crucial task due

local-aiarxiv-cs-cv
30 Apr 2026
Research

Adaptive Transform Coding for Semantic Compression

DGX agent

arXiv:2604.26492v1 Announce Type: cross Abstract: Visual data compression is shifting from human-centered reconstruction to machine-oriented representation coding. In this setting, an image is often m

researcharxiv-cs-cv
30 Apr 2026
Model Releases

AirZoo: A Unified Large-Scale Dataset for Grounding Aerial Geometric 3D Vision

DGX agent

arXiv:2604.26567v1 Announce Type: new Abstract: Despite the rapid progress in data-driven 3D vision, aerial geometric 3D vision remains a formidable challenge due to the severe scarcity of large-scale

model-releasesarxiv-cs-cv
30 Apr 2026
Research

AnimateAnyMesh++: A Flexible 4D Foundation Model for High-Fidelity Text-Driven Mesh Animation

DGX agent

arXiv:2604.26917v1 Announce Type: new Abstract: Recent advances in 4D content generation have attracted increasing attention, yet creating high-quality animated 3D models remains challenging due to th

researcharxiv-cs-cv
30 Apr 2026
Research

Are Data Augmentation and Segmentation Always Necessary? Insights from COVID-19 X-Rays and a Methodology Thereof

DGX agent

arXiv:2604.26437v1 Announce Type: new Abstract: Purpose: Rapid and reliable diagnostic tools are crucial for managing respiratory diseases like COVID-19, where chest X-ray analysis coupled with artifi

researcharxiv-cs-cv
30 Apr 2026
Safety

Attribution-Guided Multimodal Deepfake Detection via Cross-Modal Forensic Fingerprints

DGX agent

arXiv:2604.26453v1 Announce Type: new Abstract: Audio-visual deepfakes have reached a level of realism that makes perceptual detection unreliable, threatening media integrity and biometric security. W

safetyarxiv-cs-cv
30 Apr 2026
Model Releases

Benchmarking Deep Learning and Vision Foundation Models for Atypical vs. Normal Mitosis Classification with Cross-Dataset Evaluation

DGX agent

arXiv:2506.21444v4 Announce Type: replace Abstract: Atypical mitosis marks a deviation in the cell division process that has been shown be an independent prognostic marker for tumor malignancy. Howeve

model-releasesarxiv-cs-cv
30 Apr 2026
Model Releases

Beyond Fixed Formulas: Data-Driven Linear Predictor for Efficient Diffusion Models

DGX agent

arXiv:2604.26365v1 Announce Type: new Abstract: To address the high sampling cost of Diffusion Transformers (DiTs), feature caching offers a training-free acceleration method. However, existing method

model-releasesarxiv-cs-cv
30 Apr 2026
Safety

Beyond Shortcuts: Mitigating Visual Illusions in Frozen VLMs via Qualitative Reasoning

DGX agent

arXiv:2604.26250v1 Announce Type: new Abstract: While Vision-Language Models (VLMs) have achieved state-of-the-art performance in general visual tasks, their perceptual robustness remains remarkably b

safetyarxiv-cs-cv
30 Apr 2026
Model Releases

Breaking the Rigid Prior: Towards Articulated 3D Anomaly Detection

DGX agent

arXiv:2604.26868v1 Announce Type: new Abstract: Existing 3D anomaly detection methods are built on a rigid prior: normal geometry is pose-invariant and can be canonicalized through registration or ali

model-releasesarxiv-cs-cv
30 Apr 2026
Model Releases

Bridge: Basis-Driven Causal Inference Marries VFMs for Domain Generalization

DGX agent

arXiv:2604.26820v1 Announce Type: new Abstract: Detectors often suffer from degraded performance, primarily due to the distributional gap between the source and target domains. This issue is especiall

model-releasesarxiv-cs-cv
30 Apr 2026
Research

Camera-RFID Fusion for Robust Asset Tracking in Forested Environments

DGX agent

arXiv:2604.26241v1 Announce Type: new Abstract: Passive RFID tags offer a cost-effective and scalable solution for tracking numerous deployed assets. However, in forested environments, signal attenuat

researcharxiv-cs-cv
30 Apr 2026
Research

ChartVerse: Scaling Chart Reasoning via Reliable Programmatic Synthesis from Scratch

DGX agent

arXiv:2601.13606v2 Announce Type: replace Abstract: Chart reasoning is a critical capability for Vision Language Models (VLMs). However, the development of open-source models is severely hindered by t

researcharxiv-cs-cv
30 Apr 2026
Research

Circular Phase Representation and Geometry-Aware Optimization for Ptychographic Image Reconstruction

DGX agent

arXiv:2604.26664v1 Announce Type: cross Abstract: Traditional iterative reconstruction methods are accurate but computationally expensive, limiting their use in high-throughput and real-time ptychogra

researcharxiv-cs-cv
30 Apr 2026
Tutorials

CO-EVO: Co-evolving Semantic Anchoring and Style Diversification for Federated DG-ReID

DGX agent

arXiv:2604.26363v1 Announce Type: new Abstract: Federated domain generalization for person re-identification (FedDG-ReID) aims to collaboratively train a pedestrian retrieval model across multiple dec

tutorialsarxiv-cs-cv
30 Apr 2026
Applications

Color-Encoded Illumination for High-Speed Volumetric Scene Reconstruction

DGX agent

arXiv:2604.26920v1 Announce Type: new Abstract: The task of capturing and rendering 3D dynamic scenes from 2D images has become increasingly popular in recent years. However, most conventional cameras

applicationsarxiv-cs-cv
30 Apr 2026
Research

COMMA: Coordinate-aware Modulated Mamba Network for 3D Dispersed Vessel Segmentation

DGX agent

arXiv:2503.02332v3 Announce Type: replace-cross Abstract: Accurate segmentation of 3D vascular structures is essential for various medical imaging applications. The dispersed nature of vascular struct

researcharxiv-cs-cv
30 Apr 2026
Research

Contrastive Heliophysical Image Pretraining for Solar Dynamics Observatory Records

DGX agent

arXiv:2511.22958v2 Announce Type: replace Abstract: Deep learning has revolutionized solar image analysis, yet most approaches train task-specific encoders from scratch or rely on natural-image pretra

researcharxiv-cs-cv
30 Apr 2026
Model Releases

COP-GEN: Latent Diffusion Transformer for Copernicus Earth Observation Data

DGX agent

arXiv:2603.03239v2 Announce Type: replace Abstract: Earth observation applications increasingly rely on data from multiple sensors, including optical, radar, elevation, and land-cover. Relationships b

model-releasesarxiv-cs-cv
30 Apr 2026
Model Releases

Cross-Domain Transfer of Hyperspectral Foundation Models

DGX agent

arXiv:2604.26478v1 Announce Type: new Abstract: Hyperspectral imaging (HSI) semantic segmentation typically relies on in-domain training, but limited data availability often restricts model performanc

model-releasesarxiv-cs-cv
30 Apr 2026
Model Releases

CurEvo: Curriculum-Guided Self-Evolution for Video Understanding

DGX agent

arXiv:2604.26707v1 Announce Type: new Abstract: Recent advances in self-evolution video understanding frameworks have demonstrated the potential of autonomous learning without human annotations. Howev

model-releasesarxiv-cs-cv
30 Apr 2026
Model Releases

Decoupled Prototype Matching with Vision Foundation Models for Few-Shot Industrial Object Detection

DGX agent

arXiv:2604.26404v1 Announce Type: new Abstract: Industrial object detection systems typically rely on large annotated datasets, which are expensive to collect and challenging to maintain in industrial

model-releasesarxiv-cs-cv
30 Apr 2026
Safety

Delta Score Matters! Spatial Adaptive Multi Guidance in Diffusion Models

DGX agent

arXiv:2604.26503v1 Announce Type: new Abstract: Diffusion models have achieved remarkable success in synthesizing complex static and temporal visuals, a breakthrough largely driven by Classifier-Free

safetyarxiv-cs-cv
30 Apr 2026
Model Releases

DenseStep2M: A Scalable, Training-Free Pipeline for Dense Instructional Video Annotation

DGX agent

arXiv:2604.26565v1 Announce Type: new Abstract: Long-term video understanding requires interpreting complex temporal events and reasoning over procedural activities. While instructional video corpora,

model-releasesarxiv-cs-cv
30 Apr 2026
Local Ai

Edge AI for Automotive Vulnerable Road User Safety: Deployable Detection via Knowledge Distillation

DGX agent

arXiv:2604.26857v1 Announce Type: new Abstract: Deploying accurate object detection for Vulnerable Road User (VRU) safety on edge hardware requires balancing model capacity against computational const

local-aiarxiv-cs-cv
30 Apr 2026
Local Ai

Efficient Zero-Shot Inpainting with Decoupled Diffusion Guidance

DGX agent

arXiv:2512.18365v2 Announce Type: replace Abstract: Diffusion models have emerged as powerful priors for image editing tasks such as inpainting and local modification, where the objective is to genera

local-aiarxiv-cs-cv
30 Apr 2026
Research

EnerGS: Energy-Based Gaussian Splatting with Partial Geometric Priors

DGX agent

arXiv:2604.26238v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) has been widely adopted for scene reconstruction, where training inherently constitutes a highly coupled and non-convex opt

researcharxiv-cs-cv
30 Apr 2026
Research

Event-based Liveness Detection using Temporal Ocular Dynamics: An Exploratory Approach

DGX agent

arXiv:2604.26285v1 Announce Type: new Abstract: Face liveness detection has been extensively studied using RGB cameras, achieving strong performance under controlled conditions but often failing to ge

researcharxiv-cs-cv
30 Apr 2026
Model Releases

ext{PKS}^4:Parallel Kinematic Selective State Space Scanners for Efficient Video Understanding

DGX agent

arXiv:2604.26461v1 Announce Type: new Abstract: Temporal modeling remains a fundamental challenge in video understanding, particularly as sequence lengths scale. Traditional video models relying on de

model-releasesarxiv-cs-cv
30 Apr 2026
Research

FASH-iCNN: Making Editorial Fashion Identity Inspectable Through Multimodal CNN Probing

DGX agent

arXiv:2604.26186v1 Announce Type: new Abstract: Fashion AI systems routinely encode the aesthetic logic of specific houses, editors, and historical moments without disclosing it. We present FASH-iCNN,

researcharxiv-cs-cv
30 Apr 2026
← Previous
1…205206207208209…261
Next →