AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,618 results
Safety

Cross-Modal Corroboration for Annotation-Free Wildlife Monitoring

DGX agent

arXiv:2606.21613v1 Announce Type: new Abstract: Scaling wildlife monitoring for real-world conservation deployments requires automated analysis of smart sensors that operate under severe annotation sc

safetyarxiv-cs-cv
23 Jun 2026
Research
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Cross-View Yaw Estimation in Location Uncertainty with Line-Aligning Yaw Scoring

DGX agent

arXiv:2606.22094v1 Announce Type: new Abstract: Accurate yaw estimation is a bottleneck in cross-view localization between ground view and Bird's Eye View (BEV). Existing methods couple yaw with trans

researcharxiv-cs-cv
23 Jun 2026
Local Ai

CuDi: Curve Distillation for Efficient and Controllable Exposure Adjustment

DGX agent

arXiv:2207.14273v2 Announce Type: replace Abstract: We present Curve Distillation, CuDi, for efficient and controllable exposure adjustment without the requirement of paired or unpaired data during tr

local-aiarxiv-cs-cv
23 Jun 2026
Research

Cultural Counterfactuals: Evaluating Cultural Biases in Large Vision-Language Models with Counterfactual Examples

DGX agent

arXiv:2603.02370v2 Announce Type: replace Abstract: Large Vision-Language Models (LVLMs) have grown increasingly powerful in recent years, but can also exhibit harmful biases. Prior studies investigat

researcharxiv-cs-cv
23 Jun 2026
Agents

Curvature-Adaptive Consistency Flow Matching: Autonomous Trajectory Optimization via Reinforcement Learning

DGX agent

arXiv:2606.22394v1 Announce Type: new Abstract: Consistency distillation has significantly accelerated the inference of diffusion models. In this work, we reveal an intriguing asymmetry: while Logit-N

agentsarxiv-cs-cv
23 Jun 2026
Model Releases

Curvature-aware 3D length estimation of greenhouse cucumbers using RGB-D imaging and cubic spline arc-length integration

DGX agent

arXiv:2606.22439v1 Announce Type: new Abstract: Commercial greenhouse cucumber production is graded by fruit length, which drives harvest scheduling, labour allocation, and logistics. Manual measureme

model-releasesarxiv-cs-cv
23 Jun 2026
Tutorials

CurvSegFlow: Time-Conditioned Flow Matching for Robust Segmentation of Curvilinear Structures in Noisy Biomedical Images

DGX agent

arXiv:2606.21608v1 Announce Type: new Abstract: Accurate segmentation of curvilinear structures remains challenging in biomedical imaging due to their thin geometry, complex topology, and sensitivity

tutorialsarxiv-cs-cv
23 Jun 2026
Safety

Customizing Video Portraits via Identity-ActionDecoupling

DGX agent

arXiv:2606.22347v1 Announce Type: new Abstract: Identity-Preserving Text-to-Video Generation (IPT2V) seeks to synthesize a temporally coherent video from a reference image and a textual description, w

safetyarxiv-cs-cv
23 Jun 2026
Model Releases

CVSBench: A Comprehensive Benchmark for Cross-view Spatial Reasoning and Dreaming

DGX agent

arXiv:2606.22476v1 Announce Type: new Abstract: Humans can effortlessly reason about scenes across different viewpoints, yet it remains unclear whether Vision-Language Models (VLMs) possess similar cr

model-releasesarxiv-cs-cv
23 Jun 2026
Safety

D2HDMap: Non-visible Driveline Map Prior for Online Vectorized HD Map Prediction

DGX agent

arXiv:2606.20725v1 Announce Type: new Abstract: Accurate, up-to-date representations of road structures are critical for the safe operation of autonomous vehicles. Existing systems rely either on cost

safetyarxiv-cs-cv
23 Jun 2026
Research

DamageArbiter: A Multimodal Arbitration Framework for Disaster Damage Assessment from Street-View Imagery

DGX agent

arXiv:2603.14837v2 Announce Type: replace Abstract: Analyzing street-view imagery with computer vision models offers a promising approach for rapid, hyperlocal disaster damage assessment, but existing

researcharxiv-cs-cv
23 Jun 2026
Safety

Data-Driven Image Registration and Deformation Modeling for Image-Guided Neurosurgery: A Systematic Review

DGX agent

arXiv:2602.10155v2 Announce Type: replace-cross Abstract: Accurate compensation of brain deformation is critical for reliable image-guided neurosurgery. Surgical manipulation and tumor resection induc

safetyarxiv-cs-cv
23 Jun 2026
Research

Data Selection Through Iterative Self-Filtering for Vision-Language Settings

DGX agent

arXiv:2606.23611v1 Announce Type: new Abstract: The availability of large amounts of clean data is paramount to training neural networks. However, at large scales, manual oversight is impractical, res

researcharxiv-cs-cv
23 Jun 2026
Safety

DBT-Bleed: Dual-Branch Temporal Modeling with Key-Frame Selection for Surgical Bleeding Detection

DGX agent

arXiv:2606.22829v1 Announce Type: new Abstract: Intraoperative Adverse Events (IAEs) detection is critical for improving surgical safety, with bleeding being among the most frequent events across many

safetyarxiv-cs-cv
23 Jun 2026
Research

DE-FIVE: Detecting Malicious Image Prompts via Fourier Features and Image Vector Embeddings

DGX agent

arXiv:2606.22779v1 Announce Type: cross Abstract: Vision language models (VLMs) employ both visual and textual modalities to enable advanced vision-language inference. However, incorporating visual mo

researcharxiv-cs-cv
23 Jun 2026
Model Releases

Decoupling the Declarative from the Procedural in Vision-Language-Action Models

DGX agent

arXiv:2606.21496v1 Announce Type: cross Abstract: Deploying generalist robotic agents in the real world requires transferable skills. Specifically, a policy trained to clone a behavior from object-spe

model-releasesarxiv-cs-cv
23 Jun 2026
Local Ai

Deep EM with Hierarchical Latent Label Modelling for Multi-Site Prostate Lesion Segmentation

DGX agent

arXiv:2603.14418v2 Announce Type: replace Abstract: Label variability is a major challenge for prostate lesion segmentation. In multi-site datasets, annotations often reflect centre-specific contourin

local-aiarxiv-cs-cv
23 Jun 2026
Tutorials

Deep Unrolled Networks in Representation Space Applied to MRI Reconstruction

DGX agent

arXiv:2606.21602v1 Announce Type: cross Abstract: Deep unrolled networks (DUNs) integrate physical forward models with learned regularization in cascaded network architectures, achieving exceptional p

tutorialsarxiv-cs-cv
23 Jun 2026
Safety

Delta-Diffusion: Modeling Longitudinal Brain Amyloid-PET Trajectories via Conditional Poisson Diffusion Bridge

DGX agent

arXiv:2606.22216v1 Announce Type: cross Abstract: While longitudinal brain PET imaging is the gold standard for quantifying the spatiotemporal accumulation of Beta-amyloid, its widespread clinical uti

safetyarxiv-cs-cv
23 Jun 2026
Agents

Democratizing and accelerating AI-driven pathology research through agentic intelligence

DGX agent

arXiv:2606.20677v1 Announce Type: cross Abstract: Computational pathology has advanced rapidly with the emergence of foundation models, yet widespread adoption remains limited by substantial technical

agentsarxiv-cs-cv
23 Jun 2026
Research

Denoising-Enhanced Coarse-to-Fine Infrared Small Target Detection with Attention Prior-Guided Knowledge Distillation

DGX agent

arXiv:2606.21956v1 Announce Type: new Abstract: Infrared small target detection (IRSTD) in high-resolution images is crucial for many practical applications, such as surveillance of unmanned aerial ve

researcharxiv-cs-cv
23 Jun 2026
Safety

Dense Reward for Multi-View 3D Reasoning with Global Maps and Local Views

DGX agent

arXiv:2606.23557v1 Announce Type: new Abstract: Multi-view 3D Visual Question Answering (MV3D-VQA) requires integrating partial observations into a coherent 3D scene representation and selecting infor

safetyarxiv-cs-cv
23 Jun 2026
Model Releases

Detail++: Training-Free Detail Enhancer for T2I Diffusion Models

DGX agent

arXiv:2507.17853v3 Announce Type: replace Abstract: Recent advances in text-to-image (T2I) generation have led to impressive visual results. However, these models still face significant challenges whe

model-releasesarxiv-cs-cv
23 Jun 2026
Tutorials

Diffusion Integrated Gradients: Controllable Path Generation for Flexible Feature Attribution

DGX agent

arXiv:2606.22314v1 Announce Type: cross Abstract: Path-based attribution methods such as Integrated Gradients (IG) are widely adopted for their strong axiomatic properties and effectiveness in attribu

tutorialsarxiv-cs-cv
23 Jun 2026
Applications

DIPBox: A Multi-scale Testing Framework for Tracking Dataset Regeneration

DGX agent

arXiv:2606.21240v1 Announce Type: cross Abstract: Training datasets have tremendous proprietary value and are vulnerable to unauthorized copying. Existing defenses mainly focus on tracking individual

applicationsarxiv-cs-cv
23 Jun 2026
Research

Discovering Latent Groups for Robust Classification

DGX agent

arXiv:2606.23609v1 Announce Type: cross Abstract: Machine learning models exploit spurious correlations, achieving high average accuracy but failing disproportionately on underrepresented subgroups. E

researcharxiv-cs-cv
23 Jun 2026
Research

DivCon-NeRF: Diverse and Consistent Ray Augmentation for Few-Shot NeRF

DGX agent

arXiv:2503.12947v3 Announce Type: replace Abstract: Neural Radiance Field (NeRF) has shown remarkable performance in novel view synthesis but requires numerous multi-view images, limiting its practica

researcharxiv-cs-cv
23 Jun 2026
Model Releases

Do Modern Video-LLMs Need to Listen? A Benchmark Audit and Scalable Remedy

DGX agent

arXiv:2509.17901v4 Announce Type: replace Abstract: Speech and audio encoders developed over years of community effort are routinely excluded from video understanding pipelines, not because they fail,

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

DR-Mamba: Automatic Inference-Time Domain Adaptation for Document Image Binarization via Sample-Conditioned Detail-Background Suppression

DGX agent

arXiv:2606.22625v1 Announce Type: new Abstract: Degraded document image binarization is sensitive to domain shifts caused by paper aging, bleed-through, stains, shadows, and uneven illumination, and t

model-releasesarxiv-cs-cv
23 Jun 2026
Applications

DreamUV: Unwrap Artist-like UV by End-to-End Flow Matching

DGX agent

arXiv:2606.22445v1 Announce Type: new Abstract: UV parameterization is a fundamental step in 3D content creation, yet producing production-ready UV layouts remains challenging due to the gap between g

applicationsarxiv-cs-cv
23 Jun 2026
Model Releases

DrivingVoxels: Compositional Sparse Voxel Rasterization for Dynamic Driving Scene Reconstruction

DGX agent

arXiv:2606.23031v1 Announce Type: new Abstract: Reconstructing dynamic urban scenes remains challenging due to the unbounded nature of driving environments and the presence of multiple dynamic objects

model-releasesarxiv-cs-cv
23 Jun 2026
Research

Dual-Stream EEG Decoding for 3D Visual Perception

DGX agent

arXiv:2606.22182v1 Announce Type: new Abstract: This paper explores a novel brain decoding model for 3D shape perception through a dual pathway architecture mirroring biological vision. Our bio-inspir

researcharxiv-cs-cv
23 Jun 2026
Model Releases

Each Judge Its Own Yardstick: Discovering Per-VLM Taxonomies for Physical Video Evaluation

DGX agent

arXiv:2606.22918v1 Announce Type: new Abstract: Maintaining physical consistency in video generators and world models increasingly relies on vision-language models (VLMs) as automated judges that prov

model-releasesarxiv-cs-cv
23 Jun 2026
Research

EchoingPixels: Aliasing-Resistant Joint Token Reduction for Audio-Visual LLMs

DGX agent

arXiv:2512.10324v2 Announce Type: replace Abstract: Audio-Visual Large Language Models (AV-LLMs) face prohibitive computational costs of processing massive, redundant audio-visual tokens. Existing uni

researcharxiv-cs-cv
23 Jun 2026
Local Ai

Efficient Document Tampering Localization with Multi-Level Discrepancy Features and Unified DCT-Quantization Embedding

DGX agent

arXiv:2606.22285v1 Announce Type: new Abstract: Localizing document tampering is extremely challenging, as manipulations are crafted to appear visually consistent and often leave only subtle traces th

local-aiarxiv-cs-cv
23 Jun 2026
Local Ai

Efficient Traffic State Prediction With Dynamic Joint Spatio-Temporal Relation Inference

DGX agent

arXiv:2504.08061v2 Announce Type: replace Abstract: Traffic prediction is difficult due to the complex interplay of temporal evolution, spatial interactions, and delayed spatio-temporal propagation ov

local-aiarxiv-cs-cv
23 Jun 2026
Model Releases

EgoExo-Con: Exploring View-Invariant Video Temporal Understanding

DGX agent

arXiv:2510.26113v2 Announce Type: replace Abstract: Do Video-LLMs have consistent temporal understanding when videos capture the same event from different viewpoints? To study this question, we introd

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

ELDiff: When Evidential Learning Meets Text-to-Image Diffusion

DGX agent

arXiv:2606.20924v1 Announce Type: new Abstract: In multi-object text-to-image (T2I) diffusion, ensuring semantic consistency between textual prompts and generated visual content is crucial for image s

model-releasesarxiv-cs-cv
23 Jun 2026
Safety

EmbodiedUS-FS: Fast Slow Intelligence for Ultrasound Robotics

DGX agent

arXiv:2606.22319v1 Announce Type: cross Abstract: Robotic ultrasound scanning in real clinical environments requires both high-level clinical workflow reasoning and low-level closed-loop execution. Ph

safetyarxiv-cs-cv
23 Jun 2026
Safety

Enhancing IMU-Based Online Handwriting Recognition via Contrastive Learning with Zero Inference Overhead

DGX agent

arXiv:2602.07049v2 Announce Type: replace Abstract: Online handwriting recognition using inertial measurement units opens up handwriting on paper as input for digital devices. Doing it on edge hardwar

safetyarxiv-cs-cv
23 Jun 2026
Safety

Enhancing Road Safety: An IoT-Based Accident Detection and Prevention Mechanism

DGX agent

arXiv:2606.22381v1 Announce Type: cross Abstract: Road traffic accidents remain a critical global crisis, consistently serving as a primary driver of preventable mortality and severe injury. These inc

safetyarxiv-cs-cv
23 Jun 2026
Local Ai

Enlight: Fast Low-Light Image Enhancement via Multi-Objective Optimization and Shadow-Aware Refinement

DGX agent

arXiv:2606.21674v1 Announce Type: new Abstract: We present ENLIGHT, a fast and training free framework for low-light image enhancement based on direct optimization of a perceptual objective. Unlike de

local-aiarxiv-cs-cv
23 Jun 2026
Local Ai

EnTrust: Modeling Inter-Modal Conflict for Trustworthy Multimodal Medical Image Analysis

DGX agent

arXiv:2606.21384v1 Announce Type: new Abstract: Multimodal medical imaging fuses complementary anatomical and functional information, yet modalities frequently disagree in pathologically heterogeneous

local-aiarxiv-cs-cv
23 Jun 2026
Model Releases

ENVS: Environment-Native Verified Search for Long-Horizon GUI Agents

DGX agent

arXiv:2606.22948v1 Announce Type: cross Abstract: As multimodal agents move from interface understanding to real software control, successful trajectory discovery in live desktop environments becomes

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Evaluating and Enhancing Negation Comprehension in Remote Sensing MLLMs

DGX agent

arXiv:2606.20177v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have demonstrated remarkable success in various Remote Sensing (RS) tasks. However, their ability to compre

model-releasesarxiv-cs-cv
23 Jun 2026
Research

Evaluating self-supervised echocardiographic representations across downstream extraction strategies for left-ventricular segmentation and ejection fraction estimation

DGX agent

arXiv:2606.22943v1 Announce Type: new Abstract: Self-supervised learning (SSL) is increasingly used in medical imaging to reduce annotation requirements, but representation quality is often judged usi

researcharxiv-cs-cv
23 Jun 2026
Model Releases

Evaluation of Medical Vision Language Models HuluMed and MedGemma, and general purpose chatbots Gemma 3, ChatGPT Plus, and Claude Pro on real previously unseen wound images

DGX agent

arXiv:2606.20723v1 Announce Type: new Abstract: Chronic wound assessment remains a clinically challenging task that requires accurate interpretation of wound morphology, tissue composition, vascular c

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Evo-RAD: Navigating Rare Retinal Disease Diagnosis via Self-Evolving Agentic Retrieval

DGX agent

arXiv:2606.22955v1 Announce Type: new Abstract: Large-scale pretrained foundation models have revolutionized general medical screening, but often falter on rare diseases because such conditions are un

model-releasesarxiv-cs-cv
23 Jun 2026
← Previous
1…96979899100…263
Next →