AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,618 results
Research

DiffSight-Former: Modeling Structural Differences and Temporal Dynamics for Glaucoma Progression Prediction

DGX agent

arXiv:2606.09140v1 Announce Type: new Abstract: Glaucoma is a leading cause of irreversible blindness worldwide, and early detection from fundus images is critical for effective disease management. Wh

researcharxiv-cs-cv
9 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Research

Diffusion Bridge or Flow Matching? A Unifying Framework and Comparative Analysis

DGX agent

arXiv:2509.24531v2 Announce Type: replace Abstract: Diffusion Bridge and Flow Matching have both demonstrated compelling empirical performance in transformation between arbitrary distributions. Howeve

researcharxiv-cs-cv
9 Jun 2026
Research

DIJIT: A Robotic Head for an Active Observer

DGX agent

arXiv:2512.07998v2 Announce Type: replace-cross Abstract: We present DIJIT, a novel binocular robotic head expressly designed for mobile agents that behave as active observers. DIJIT's unique breadth

researcharxiv-cs-cv
9 Jun 2026
Model Releases

DisCo: World Models with Discrete Camera Motion Control

DGX agent

arXiv:2606.07967v1 Announce Type: new Abstract: Controllable video world models target interactive world exploration, where models must faithfully execute explicit action commands while preserving vis

model-releasesarxiv-cs-cv
9 Jun 2026
Safety

Distant Object Localisation from Noisy Image Segmentation Sequences

DGX agent

arXiv:2509.20906v3 Announce Type: replace Abstract: 3D object localisation based on a sequence of camera measurements is essential for safety-critical surveillance tasks, such as drone-based wildfire

safetyarxiv-cs-cv
9 Jun 2026
Model Releases

Distortion-Aware PETR for BEV Object Detection with Mixed Pinhole-Fisheye Cameras

DGX agent

arXiv:2606.08680v1 Announce Type: new Abstract: Fisheye cameras are widely deployed in autonomous driving perception suites for their low cost and full-coverage field of view (FOV), yet their potentia

model-releasesarxiv-cs-cv
9 Jun 2026
Safety

Do VLMs See What Sensors Feel? A Scalable Expert-Guided Design for Wheelchair Accessibility Assessment from Street View

DGX agent

arXiv:2606.07642v1 Announce Type: new Abstract: Assessing built-environment interaction, such as wheelchair accessibility, is difficult because real-world mobility is shaped by distributed, context-de

safetyarxiv-cs-cv
9 Jun 2026
Safety

Dr. SHAP-AV: Decoding Relative Modality Contributions via Shapley Attribution in Audio-Visual Speech Recognition

DGX agent

arXiv:2603.12046v2 Announce Type: replace-cross Abstract: Audio-Visual Speech Recognition (AVSR) leverages both acoustic and visual information for robust recognition under noise. However, how models

safetyarxiv-cs-cv
9 Jun 2026
Model Releases

DriveReward: A Comprehensive Dataset and Generative Vision-Language Reward Model for Autonomous Driving

DGX agent

arXiv:2606.08525v1 Announce Type: new Abstract: Reward models play a pivotal role in reinforcement learning (RL) and multi-modal trajectory selection for autonomous driving. However, acquiring such re

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Driving Video Retrieval for Complex Queries with Structured Grounding

DGX agent

arXiv:2606.09109v1 Announce Type: new Abstract: Video retrieval at scale is central to data curation and safety validation in autonomous driving, where users want to find not only scenes but also dyna

model-releasesarxiv-cs-cv
9 Jun 2026
Applications

DroneDAR: Long-Range Drone Distance Estimation Using Monocular Vision and Bounding-Box Features

DGX agent

arXiv:2606.07756v1 Announce Type: new Abstract: Accurate distance estimation for small drones in long-range imagery is important for tracking and situational awareness, yet remains challenging due to

applicationsarxiv-cs-cv
9 Jun 2026
Safety

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning

DGX agent

arXiv:2606.08035v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has emerged as a leading paradigm for enhancing visual reasoning in Multimodal Large Language Mode

safetyarxiv-cs-cv
9 Jun 2026
Model Releases

Echo-DM: Ultrasound Marker Removal via Conditional Latent Diffusion and Region-Aware Fusion

DGX agent

arXiv:2606.09378v1 Announce Type: new Abstract: Clinical ultrasound images often contain artificial markers, such as measurement calipers and text, to assist diagnostic interpretation and comparison.

model-releasesarxiv-cs-cv
9 Jun 2026
Research

Echo-Memory: A Controlled Study of Memory in Action World Models

DGX agent

arXiv:2606.09803v1 Announce Type: new Abstract: We present extbf{Echo-Memory}, a controlled study of memory mechanisms in action-conditioned world models. These models generate multi-segment videos fr

researcharxiv-cs-cv
9 Jun 2026
Research

Edge-Constrained UAV Small-Object Detection with P2 Enhancement and Quantum-Inspired Lightweight Structure Search

DGX agent

arXiv:2606.09081v1 Announce Type: new Abstract: Unmanned aerial vehicle (UAV) object detection requires compact detectors that retain small-object details under onboard computation and memory constrai

researcharxiv-cs-cv
9 Jun 2026
Agents

EditSSC: Toward Editable Semantic Occupancy Scenes with Unconditional Diffusion Models

DGX agent

arXiv:2606.09273v1 Announce Type: new Abstract: 3D semantic scene generation is crucial for autonomous driving applications, yet most methods rely on complex 3D-specific architectures such as triplane

agentsarxiv-cs-cv
9 Jun 2026
Model Releases

Efficient Minimal Solvers for Relative Pose Estimation in Autonomous Driving Applications

DGX agent

arXiv:2606.09569v1 Announce Type: cross Abstract: With the advancement of visual sensing systems, computer vision is playing an increasingly important role in autonomous driving and robot navigation.

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Efficient Minimal Solvers for Visual-Inertial Relative Pose Estimation in Multi-Camera Systems

DGX agent

arXiv:2606.09477v1 Announce Type: new Abstract: Estimating the relative poses of multi-camera systems is a fundamental problem in computer vision, with critical applications in autonomous vehicles, mo

model-releasesarxiv-cs-cv
9 Jun 2026
Research

EgoPriMo: Egocentric Motion Generation for Interactive Humanoid Control

DGX agent

arXiv:2606.08495v1 Announce Type: cross Abstract: Humanoid robots require whole-body motions that adapt to scene context, task requirements, and user intent. Motion tracking reproduces specified traje

researcharxiv-cs-cv
9 Jun 2026
Research

Embedded Graph Convolutional Networks for Real-Time Event Data Processing on SoC FPGAs

DGX agent

arXiv:2406.07318v3 Announce Type: replace Abstract: The utilisation of event cameras represents an important and swiftly evolving trend aimed at addressing the constraints of traditional video systems

researcharxiv-cs-cv
9 Jun 2026
Local Ai

Empowering Feed-Forward Reconstruction Models with Metric Scale via Satellite Images

DGX agent

arXiv:2606.08205v1 Announce Type: new Abstract: Feed-forward 3D reconstruction models have recently shown strong generalization across diverse scenes, yet most of them recover geometry only up to an u

local-aiarxiv-cs-cv
9 Jun 2026
Research

End-to-End Optimization of Incoherent Imaging for Classification Under Detector-Limited Readout

DGX agent

arXiv:2606.09792v1 Announce Type: new Abstract: End-to-end co-optimization of optical front-ends (e.g. metasurfaces) and neural network back-ends has been widely applied to imaging tasks, yet a formal

researcharxiv-cs-cv
9 Jun 2026
Model Releases

Enhanced Detection of Tiny Objects in Aerial Images

DGX agent

arXiv:2509.17078v3 Announce Type: replace Abstract: While one-stage detectors like YOLOv8 offer fast training speed, they often under-perform on detecting small objects as a trade-off. This becomes ev

model-releasesarxiv-cs-cv
9 Jun 2026
Local Ai

Enhancing Adversarial Robustness with Signed Distance Fields for Harmonizing Geometric Invariance and Texture

DGX agent

arXiv:2602.05175v2 Announce Type: replace Abstract: Deep neural networks demonstrate impressive performance in visual recognition but remain highly vulnerable to imperceptible adversarial attacks. Exi

local-aiarxiv-cs-cv
9 Jun 2026
Research

EPS3D: End-to-End Feed-Forward 3D Panoptic Segmentation

DGX agent

arXiv:2606.08980v1 Announce Type: new Abstract: This paper introduces EPS3D, a new end-to-end feed-forward framework for open-vocabulary 3D panoptic segmentation. Unlike existing methods relying on ad

researcharxiv-cs-cv
9 Jun 2026
Research

Evaluating the Representation Space of Diffusion Models via Self-Supervised Principles

DGX agent

arXiv:2606.09718v1 Announce Type: cross Abstract: Diffusion models have demonstrated remarkable generative capabilities and have also emerged as powerful self-supervised representation learners, yet t

researcharxiv-cs-cv
9 Jun 2026
Local Ai

Event-driven dynamic trajectories reconstruction and measurement of mechanical parameters for fragments

DGX agent

arXiv:2606.09208v1 Announce Type: new Abstract: During warhead detonation, high-density, high-speed, and mutually occluded fragments are generated. Their mechanical parameters (position, velocity, kin

local-aiarxiv-cs-cv
9 Jun 2026
Research

ExDet: Open-Domain Open-Vocabulary Detection with Cross-modal Extrapolation and Rectification

DGX agent

arXiv:2606.09360v1 Announce Type: new Abstract: Open-domain open-vocabulary detection (ODOVD) requires detectors to generalize to both novel categories and unseen domains, making it more challenging t

researcharxiv-cs-cv
9 Jun 2026
Research

Facial Expression Recognition in the Deep Learning Era: A Systematic Multi-Criteria Review of Methods, Models, Datasets, Performance, Challenges, and Future Research Directions

DGX agent

arXiv:2606.08612v1 Announce Type: new Abstract: Facial Expression Recognition (FER) has advanced rapidly over the last decade, driven by the shift from handcrafted descriptors and shallow classifiers

researcharxiv-cs-cv
9 Jun 2026
Safety

FADRW: A Feature-Aware Modulated and Dynamically Reweighted Loss for Few-Shot Linguistic Steganalysis

DGX agent

arXiv:2606.07655v1 Announce Type: cross Abstract: The ubiquity of social media platforms facilitates malicious linguistic steganography, posing significant security risks. However, detection is severe

safetyarxiv-cs-cv
9 Jun 2026
Research

Feasibility to detect rapid change and disappearance of seagrass: Lessons from nearly 80 years of vegetation change in the Ako, Seto Inland Sea, Japan

DGX agent

arXiv:2606.07949v1 Announce Type: cross Abstract: This study analyses the Ako tidal flat in the Seto Inland Sea, Japan, where nearly all Zostera marina disappeared within a single year in 2025. Using

researcharxiv-cs-cv
9 Jun 2026
Safety

FlowLet: Conditional 3D Brain MRI Synthesis using Wavelet Flow Matching

DGX agent

arXiv:2601.05212v2 Announce Type: replace Abstract: Brain Magnetic Resonance Imaging (MRI) plays a central role in studying neurological development, aging, and diseases. One key application is Brain

safetyarxiv-cs-cv
9 Jun 2026
Model Releases

FMRFusion: Frequency-Aware Multi-View Representation Learning for Heterogeneous Image Fusion

DGX agent

arXiv:2606.07985v1 Announce Type: new Abstract: Infrared and visible image fusion aims to generate a composite image that retains significant target information and preserves detailed textures, integr

model-releasesarxiv-cs-cv
9 Jun 2026
Safety

Frankenstein in the Pipeline: Computational Epistemicide in Facial Recognition

DGX agent

arXiv:2606.07628v1 Announce Type: cross Abstract: While the eugenic roots of computer vision are well-documented in critical technology studies, less attention has been paid to the operational mechani

safetyarxiv-cs-cv
9 Jun 2026
Research

Frequency Decoupled Framework for Screen Content Image Super-Resolution

DGX agent

arXiv:2606.09029v1 Announce Type: new Abstract: Methods based on implicit neural representations have demonstrated superior performance in Screen Content Image Super-Resolution (SCISR) . However, they

researcharxiv-cs-cv
9 Jun 2026
Research

Frequency-Scale Saliency for Spectral Descriptor Analysis in 3D Shape Retrieval

DGX agent

arXiv:2606.07791v1 Announce Type: cross Abstract: Classical spectral descriptors such as the Heat Kernel Signature and Wave Kernel Signature are widely used for non-rigid 3D shape retrieval, yet their

researcharxiv-cs-cv
9 Jun 2026
Research

Fully Spiking Neural Networks with Target Awareness for Energy-Efficient UAV Tracking

DGX agent

arXiv:2603.27493v2 Announce Type: replace Abstract: Spiking Neural Networks (SNNs), characterized by their event-driven computation and low power consumption, have shown great potential for energy-eff

researcharxiv-cs-cv
9 Jun 2026
Applications

G2G: Exploiting Intra-Group Geometry for Inter-Group Pose Estimation

DGX agent

arXiv:2606.08284v1 Announce Type: new Abstract: Recovering the relative 6-DoF pose between two image groups underlies cross-sequence relocalization and multi-camera rig odometry. Each group carries kn

applicationsarxiv-cs-cv
9 Jun 2026
Model Releases

GD-MIL: Grade-Disentangled Multiple Instance Learning for Multimodal Biochemical Recurrence Prediction in Prostate Cancer

DGX agent

arXiv:2606.09453v1 Announce Type: new Abstract: Biochemical recurrence (BCR) after radical prostatectomy is a critical endpoint in prostate cancer, yet risk stratification relies almost entirely on va

model-releasesarxiv-cs-cv
9 Jun 2026
Research

Generalizing Geometry-Guided Mamba as a Plug-and-Play Context Module for CNN-based Semantic Segmentation

DGX agent

arXiv:2606.08866v1 Announce Type: new Abstract: CNN-based semantic segmentation networks usually rely on context heads such as ASPP, PPM, or attention modules to enlarge the receptive field. These hea

researcharxiv-cs-cv
9 Jun 2026
Local Ai

GenEyePose: Patient-Free, Knowledge-Based Saccadic Eye Movement Modeling for Digital Neurophysiologic Biomarker Development

DGX agent

arXiv:2606.09681v1 Announce Type: new Abstract: Eye movements, including saccades, are widely regarded as highly sensitive and objective biomarkers of neurophysiologic states. Detecting saccadic signa

local-aiarxiv-cs-cv
9 Jun 2026
Research

Geometric Analysis of Magnetic Labyrinthine Stripe Evolution via Deep Learning Segmentation

DGX agent

arXiv:2509.11485v3 Announce Type: replace-cross Abstract: Labyrinthine stripe patterns are common in many physical systems, yet their lack of long-range order makes quantitative characterization chall

researcharxiv-cs-cv
9 Jun 2026
Agents

Geometry-Aware Fisheye-LiDAR Fusion for Robust 3D Object Detection in Low-Overlap Setups

DGX agent

arXiv:2606.08844v1 Announce Type: new Abstract: As autonomous systems expand from capital-intensive robotaxis to cost-sensitive logistics, sensor configurations are increasingly optimized for coverage

agentsarxiv-cs-cv
9 Jun 2026
Research

Geometry-Driven Flow Analysis of Brain Sulcal Pattern

DGX agent

arXiv:2606.08404v1 Announce Type: new Abstract: Cortical folding reflects coordinated neurodevelopmental processes and is increasingly recognized as a sensitive marker of neurological disease. However

researcharxiv-cs-cv
9 Jun 2026
Applications

GimmBO: Interactive Generative Image Model Merging via Bayesian Optimization

DGX agent

arXiv:2601.18585v2 Announce Type: replace Abstract: Fine-tuning-based adaptation is widely used to customize diffusion-based image generation, leading to large collections of community-created adapter

applicationsarxiv-cs-cv
9 Jun 2026
Research

GraspFoM: Towards Reconstruction-Driven Robotic Grasping with 3D Foundation Priors

DGX agent

arXiv:2606.08440v1 Announce Type: cross Abstract: Robotic grasping is a fundamental capability in robotic manipulation. Yet grasping remains challenging under partial observations. Reliable grasping d

researcharxiv-cs-cv
9 Jun 2026
Research

Gravity-guided Contact Dynamics Estimation from 3D Human Motions

DGX agent

arXiv:2606.08133v1 Announce Type: new Abstract: Ground contact forces acting on the human body, are crucial for biomechanics studies or sport performance analysis. Prior methods rely on force plates o

researcharxiv-cs-cv
9 Jun 2026
Research

HACK++: Towards More Effective Head-Aware Key-Value Compression for Efficient Visual Autoregressive Modeling

DGX agent

arXiv:2606.08302v1 Announce Type: new Abstract: Visual Autoregressive (VAR) models adopt a next-scale prediction paradigm, offering high-quality generation with substantially fewer decoding steps. How

researcharxiv-cs-cv
9 Jun 2026
← Previous
1…110111112113114…263
Next →