AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,515 results
Model Releases

AEGIS: Anchor-Enforced Gradient Isolation for Knowledge-Preserving Vision-Language-Action Fine-Tuning

DGX agent

arXiv:2604.16067v1 Announce Type: cross Abstract: Adapting pre-trained vision-language models (VLMs) for robotic control requires injecting high-magnitude continuous gradients from a flow-matching act

model-releasesarxiv-cs-cv
20 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Applications

AeroDeshadow: Physics-Guided Shadow Synthesis and Penumbra-Aware Deshadowing for Aerospace Imagery

DGX agent

arXiv:2604.15903v1 Announce Type: new Abstract: Shadows are prevalent in high-resolution aerospace imagery (ASI). They often cause spectral distortion and information loss, which degrade downstream in

applicationsarxiv-cs-cv
20 Apr 2026
Applications

AHS: Adaptive Head Synthesis via Synthetic Data Augmentations

DGX agent

arXiv:2604.15857v1 Announce Type: new Abstract: Recent digital media advancements have created increasing demands for sophisticated portrait manipulation techniques, particularly head swapping, where

applicationsarxiv-cs-cv
20 Apr 2026
Research

Aligning What Vision-Language Models See and Perceive with Adaptive Information Flow

DGX agent

arXiv:2604.15809v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have demonstrated strong capability in a wide range of tasks such as visual recognition, document parsing, and visual grou

researcharxiv-cs-cv
20 Apr 2026
Applications

An Empirical Study of Validating Synthetic Data for Text-Based Person Retrieval

DGX agent

arXiv:2503.22171v2 Announce Type: replace Abstract: Data plays a pivotal role in Text-Based Person Retrieval (TBPR) research. Mainstream research paradigm necessitates real-world person images with ma

applicationsarxiv-cs-cv
20 Apr 2026
Model Releases

APC: Transferable and Efficient Adversarial Point Counterattack for Robust 3D Point Cloud Recognition

DGX agent

arXiv:2604.15708v1 Announce Type: new Abstract: The advent of deep neural networks has led to remarkable progress in 3D point cloud recognition, but they remain vulnerable to adversarial attacks. Alth

model-releasesarxiv-cs-cv
20 Apr 2026
Model Releases

Art3D: Training-Free 3D Generation from Flat-Colored Illustration

DGX agent

arXiv:2504.10466v2 Announce Type: replace Abstract: Large-scale pre-trained image-to-3D generative models have exhibited remarkable capabilities in diverse shape generations. However, most of them str

model-releasesarxiv-cs-cv
20 Apr 2026
Agents

AstroVLM: Expert Multi-agent Collaborative Reasoning for Astronomical Imaging Quality Diagnosis

DGX agent

arXiv:2604.16024v1 Announce Type: cross Abstract: Vision Language Models (VLMs) have been applied to several specific domains and have shown strong problem-solving capabilities. However, astronomical

agentsarxiv-cs-cv
20 Apr 2026
Safety

AutoDrive-R^2: Incentivizing Reasoning and Self-Reflection Capacity for VLA Model in Autonomous Driving

DGX agent

arXiv:2509.01944v3 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models in autonomous driving systems have recently demonstrated transformative potential by integrating multimoda

safetyarxiv-cs-cv
20 Apr 2026
Research

Beyond Text Prompts: Precise Concept Erasure through Text-Image Collaboration

DGX agent

arXiv:2604.15829v1 Announce Type: new Abstract: Text-to-image generative models have achieved impressive fidelity and diversity, but can inadvertently produce unsafe or undesirable content due to impl

researcharxiv-cs-cv
20 Apr 2026
Research

Breakout-picker: Reducing false positives in deep learning-based borehole breakout characterization from acoustic image logs

DGX agent

arXiv:2604.16011v1 Announce Type: new Abstract: Borehole breakouts are stress-induced spalling on the borehole wall, which are identifiable in acoustic image logs as paired zones with near-symmetry az

researcharxiv-cs-cv
20 Apr 2026
Safety

CASR: A Robust Cyclic Framework for Arbitrary Large-Scale Super-Resolution with Distribution Alignment and Self-Similarity Awareness

DGX agent

arXiv:2602.22159v2 Announce Type: replace Abstract: Arbitrary-Scale SR (ASISR) remains fundamentally limited by cross-scale distribution shift: once the inference scale leaves the training range, nois

safetyarxiv-cs-cv
20 Apr 2026
Safety

Causal Bootstrapped Alignment for Unsupervised Video-Based Visible-Infrared Person Re-Identification

DGX agent

arXiv:2604.15631v1 Announce Type: new Abstract: VVI-ReID is a critical technique for all-day surveillance, where temporal information provides additional cues beyond static images. However, existing a

safetyarxiv-cs-cv
20 Apr 2026
Model Releases

ChatENV: An Interactive Vision-Language Model for Sensor-Guided Environmental Monitoring and Scenario Simulation

DGX agent

arXiv:2508.10635v3 Announce Type: replace Abstract: Understanding environmental changes from remote sensing imagery is vital for climate resilience, urban planning, and ecosystem monitoring. Yet, curr

model-releasesarxiv-cs-cv
20 Apr 2026
Research

CLOTH-HUGS: Cloth Aware Human Gaussian Splatting

DGX agent

arXiv:2604.15875v1 Announce Type: new Abstract: We present Cloth-HUGS, a Gaussian Splatting based neural rendering framework for photorealistic clothed human reconstruction that explicitly disentangle

researcharxiv-cs-cv
20 Apr 2026
Research

CollideNet: Hierarchical Multi-scale Video Representation Learning with Disentanglement for Time-To-Collision Forecasting

DGX agent

arXiv:2604.16240v1 Announce Type: new Abstract: Time-to-Collision (TTC) forecasting is a critical task in collision prevention, requiring precise temporal prediction and comprehending both local and g

researcharxiv-cs-cv
20 Apr 2026
Research

Comparison Study: Glacier Calving Front Delineation in Synthetic Aperture Radar Images With Deep Learning

DGX agent

arXiv:2501.05281v2 Announce Type: replace Abstract: Continuous monitoring of glacier calving fronts is essential for sea level rise projections. This study benchmarks Deep Learning systems for front d

researcharxiv-cs-cv
20 Apr 2026
Safety

Concept-wise Attention for Fine-grained Concept Bottleneck Models

DGX agent

arXiv:2604.15748v1 Announce Type: new Abstract: Recently impressive performance has been achieved in Concept Bottleneck Models (CBM) by utilizing the image-text alignment learned by a large pre-traine

safetyarxiv-cs-cv
20 Apr 2026
Local Ai

Continual Hand-Eye Calibration for Open-world Robotic Manipulation

DGX agent

arXiv:2604.15814v1 Announce Type: new Abstract: Hand-eye calibration through visual localization is a critical capability for robotic manipulation in open-world environments. However, most deep learni

local-aiarxiv-cs-cv
20 Apr 2026
Hardware

CPU Optimization of a Monocular 3D Biomechanics Pipeline for Low-Resource Deployment

DGX agent

arXiv:2604.15665v1 Announce Type: new Abstract: Markerless 3D movement analysis from monocular video enables accessible biomechanical assessment in clinical and sports settings. However, most research

hardwarearxiv-cs-cv
20 Apr 2026
Tutorials

Cross-modal learning for plankton recognition

DGX agent

arXiv:2603.16427v2 Announce Type: replace Abstract: This paper considers self-supervised cross-modal coordination as a strategy enabling utilization of multiple modalities and large volumes of unlabel

tutorialsarxiv-cs-cv
20 Apr 2026
Model Releases

CTSCAN: Evaluation Leakage in Chest CT Segmentation and a Reproducible Patient-Disjoint Benchmark

DGX agent

arXiv:2604.15561v1 Announce Type: cross Abstract: Reported chest CT segmentation performance can be strongly inflated when train and test partitions mix slices from the same study. We present CTSCAN,

model-releasesarxiv-cs-cv
20 Apr 2026
Model Releases

CXR-LT 2026 Challenge: Multi-Center Long-Tailed and Zero Shot Chest X-ray Classification

DGX agent

arXiv:2604.15555v1 Announce Type: new Abstract: Chest X-ray (CXR) interpretation is hindered by the long-tailed distribution of pathologies and the open-world nature of clinical environments. Existing

model-releasesarxiv-cs-cv
20 Apr 2026
Applications

DENALI: A Dataset Enabling Non-Line-of-Sight Spatial Reasoning with Low-Cost LiDARs

DGX agent

arXiv:2604.16201v1 Announce Type: cross Abstract: Consumer LiDARs in mobile devices and robots typically output a single depth value per pixel. Yet internally, they record full time-resolved histogram

applicationsarxiv-cs-cv
20 Apr 2026
Model Releases

DenTab: A Dataset for Table Recognition and Visual QA on Real-World Dental Estimates

DGX agent

arXiv:2604.16099v1 Announce Type: new Abstract: Tables condense key transactional and administrative information into compact layouts, but practical extraction requires more than text recognition: sys

model-releasesarxiv-cs-cv
20 Apr 2026
Research

Dental Panoramic Radiograph Analysis Using YOLO26 From Tooth Detection to Disease Diagnosis

DGX agent

arXiv:2604.16231v1 Announce Type: new Abstract: Panoramic radiography is a fundamental diagnostic tool in dentistry, offering a comprehensive view of the entire dentition with minimal radiation exposu

researcharxiv-cs-cv
20 Apr 2026
Local Ai

DINOv3 Beats Specialized Detectors: A Simple Foundation Model Baseline for Image Forensics

DGX agent

arXiv:2604.16083v1 Announce Type: new Abstract: With the rapid advancement of deep generative models, realistic fake images have become increasingly accessible, yet existing localization methods rely

local-aiarxiv-cs-cv
20 Apr 2026
Model Releases

DriveLaW:Unifying Planning and Video Generation in a Latent Driving World

DGX agent

arXiv:2512.23421v3 Announce Type: replace Abstract: World models have become crucial for autonomous driving, as they learn how scenarios evolve over time to address the long-tail challenges of the rea

model-releasesarxiv-cs-cv
20 Apr 2026
Model Releases

DualTrack: Sensorless 3D Ultrasound needs Local and Global Context

DGX agent

arXiv:2509.09530v2 Announce Type: replace Abstract: Three-dimensional ultrasound (US) offers many clinical advantages over conventional 2D imaging, yet its widespread adoption is limited by the cost a

model-releasesarxiv-cs-cv
20 Apr 2026
Research

DVP-MVS++: Synergize Depth-Normal-Edge and Harmonized Visibility Prior for Multi-View Stereo

DGX agent

arXiv:2506.13215v2 Announce Type: replace Abstract: Recently, patch deformation-based methods have demonstrated significant effectiveness in multi-view stereo due to their incorporation of deformable

researcharxiv-cs-cv
20 Apr 2026
Safety

DyTact: Capturing Dynamic Contacts in Hand-Object Manipulation

DGX agent

arXiv:2506.03103v2 Announce Type: replace Abstract: Reconstructing dynamic hand-object contacts is essential for realistic manipulation in AI character animation, XR, and robotics, yet it remains chal

safetyarxiv-cs-cv
20 Apr 2026
Research

EchoVLM: Dynamic Mixture-of-Experts Vision-Language Model for Universal Ultrasound Intelligence

DGX agent

arXiv:2509.14977v2 Announce Type: replace Abstract: Ultrasound imaging has become the preferred imaging modality for early cancer screening due to its advantages of non-ionizing radiation, low cost, a

researcharxiv-cs-cv
20 Apr 2026
Applications

Efficient Video Diffusion Models: Advancements and Challenges

DGX agent

arXiv:2604.15911v1 Announce Type: new Abstract: Video diffusion models have rapidly become the dominant paradigm for high-fidelity generative video synthesis, but their practical deployment remains co

applicationsarxiv-cs-cv
20 Apr 2026
Safety

Elucidating the SNR-t Bias of Diffusion Probabilistic Models

DGX agent

arXiv:2604.16044v1 Announce Type: new Abstract: Diffusion Probabilistic Models have demonstrated remarkable performance across a wide range of generative tasks. However, we have observed that these mo

safetyarxiv-cs-cv
20 Apr 2026
Research

Enhancing Hazy Wildlife Imagery: AnimalHaze3k and IncepDehazeGan

DGX agent

arXiv:2604.16284v1 Announce Type: new Abstract: Atmospheric haze significantly degrades wildlife imagery, impeding computer vision applications critical for conservation, such as animal detection, tra

researcharxiv-cs-cv
20 Apr 2026
Research

EventCrab: Harnessing Frame and Point Synergy for Event-based Action Recognition and Beyond

DGX agent

arXiv:2411.18328v2 Announce Type: replace Abstract: Event-based Action Recognition (EAR) possesses the advantages of high-temporal resolution capturing and privacy preservation compared with tradition

researcharxiv-cs-cv
20 Apr 2026
Local Ai

Fed3D: Federated 3D Object Detection

DGX agent

arXiv:2604.15795v1 Announce Type: new Abstract: 3D object detection models trained in one server plays an important role in autonomous driving, robotics manipulation, and augmented reality scenarios.

local-aiarxiv-cs-cv
20 Apr 2026
Model Releases

FETAL-GAUGE: A Benchmark for Assessing Vision-Language Models in Fetal Ultrasound

DGX agent

arXiv:2512.22278v2 Announce Type: replace Abstract: The growing demand for prenatal ultrasound imaging has intensified a global shortage of trained sonographers, creating barriers to essential fetal h

model-releasesarxiv-cs-cv
20 Apr 2026
Safety

Find, Fix, Reason: Context Repair for Video Reasoning

DGX agent

arXiv:2604.16243v1 Announce Type: new Abstract: Reinforcement learning has advanced video reasoning in large multi-modal models, yet dominant pipelines either rely on on-policy self-exploration, which

safetyarxiv-cs-cv
20 Apr 2026
Model Releases

FineCog-Nav: Integrating Fine-grained Cognitive Modules for Zero-shot Multimodal UAV Navigation

DGX agent

arXiv:2604.16298v1 Announce Type: new Abstract: UAV vision-language navigation (VLN) requires an agent to navigate complex 3D environments from an egocentric perspective while following ambiguous mult

model-releasesarxiv-cs-cv
20 Apr 2026
Model Releases

Frequency-Aware Flow Matching for High-Quality Image Generation

DGX agent

arXiv:2604.15521v1 Announce Type: new Abstract: Flow matching models have emerged as a powerful framework for realistic image generation by learning to reverse a corruption process that progressively

model-releasesarxiv-cs-cv
20 Apr 2026
Applications

From Articles to Canopies: Knowledge-Driven Pseudo-Labelling for Tree Species Classification using LLM Experts

DGX agent

arXiv:2604.16115v1 Announce Type: new Abstract: Hyperspectral tree species classification is challenging due to limited and imbalanced class labels, spectral mixing (overlapping light signatures from

applicationsarxiv-cs-cv
20 Apr 2026
Safety

From Competition to Coopetition: Coopetitive Training-Free Image Editing Based on Text Guidance

DGX agent

arXiv:2604.15948v1 Announce Type: new Abstract: Text-guided image editing, a pivotal task in modern multimedia content creation, has seen remarkable progress with training-free methods that eliminate

safetyarxiv-cs-cv
20 Apr 2026
Local Ai

From Limited Labels to Open Domains:An Efficient Learning Method for Drone-view Geo-Localization

DGX agent

arXiv:2503.07520v5 Announce Type: replace Abstract: Traditional supervised drone-view geo-localization (DVGL) methods heavily depend on paired training data and encounter difficulties in learning cros

local-aiarxiv-cs-cv
20 Apr 2026
Model Releases

From Zero to Detail: A Progressive Spectral Decoupling Paradigm for UHD Image Restoration with New Benchmark

DGX agent

arXiv:2604.15654v1 Announce Type: new Abstract: Ultra-high-definition (UHD) image restoration poses unique challenges due to the high spatial resolution, diverse content, and fine-grained structures p

model-releasesarxiv-cs-cv
20 Apr 2026
Tutorials

GaussianFlow SLAM: Monocular Gaussian Splatting SLAM Guided by GaussianFlow

DGX agent

arXiv:2604.15612v1 Announce Type: cross Abstract: Gaussian splatting has recently gained traction as a compelling map representation for SLAM systems, enabling dense and photo-realistic scene modeling

tutorialsarxiv-cs-cv
20 Apr 2026
Applications

GAViD: A Large-Scale Multimodal Dataset for Context-Aware Group Affect Recognition from Videos

DGX agent

arXiv:2604.16214v1 Announce Type: new Abstract: Understanding affective dynamics in real-world social systems is fundamental to modeling and analyzing human-human interactions in complex environments.

applicationsarxiv-cs-cv
20 Apr 2026
Research

GenHSI: Controllable Generation of Human-Scene Interaction Videos

DGX agent

arXiv:2506.19840v2 Announce Type: replace Abstract: Large-scale pre-trained video diffusion models have exhibited remarkable capabilities in diverse video generation. However, existing solutions face

researcharxiv-cs-cv
20 Apr 2026
← Previous
1…233234235236237…261
Next →