AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,515 results
Model Releases

AGVBench: A Reliability-Oriented Benchmark of Data Augmentation for Vein Recognition

DGX agent

arXiv:2607.02271v2 Announce Type: replace Abstract: Vein recognition is a secure biometric technology often constrained by limited annotated data and imaging variations. While data augmentation mitiga

model-releasesarxiv-cs-cv
28 Jul 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Research

Ambient pressure compensation and robust position control of oil-filled electric joint systems for underwater manipulators

DGX agent

arXiv:2607.24384v1 Announce Type: new Abstract: Electric joint systems are significant elements of an underwater manipulator for its actuation, drive, and control. Working in an underwater environment

researcharxiv-cs-cv
28 Jul 2026
Model Releases

AptAvatar: Fast and Vivid Long-Form Audio-Driven Video Generation for Production-Ready Avatars

DGX agent

arXiv:2607.24013v1 Announce Type: new Abstract: Production-ready audio-driven avatar generation requires efficient inference without sacrificing fidelity or motion expressiveness. However, existing ac

model-releasesarxiv-cs-cv
28 Jul 2026
Research

ATCNet-CIAM for Multi-Session Motor Imagery EEG Signal Classification

DGX agent

arXiv:2607.23522v1 Announce Type: new Abstract: Motor imagery (MI)-based electroencephalography is widely used in non-invasive brain--computer interfaces (BCIs), but robust decoding remains challengin

researcharxiv-cs-cv
28 Jul 2026
Model Releases

BATON: A Multimodal Benchmark for Bidirectional Automation Transition Observation in Naturalistic Driving

DGX agent

arXiv:2604.07263v2 Announce Type: replace-cross Abstract: Existing driving automation (DA) systems on production vehicles rely on human drivers to decide when to engage DA while requiring them to rema

model-releasesarxiv-cs-cv
28 Jul 2026
Research

Benchmarking the Domain Gap: Model Selection Instability Under Domain Shift in Video Capsule Endoscopy

DGX agent

arXiv:2607.22736v1 Announce Type: new Abstract: Video capsule endoscopy (VCE) classification is typically evaluated within a single dataset, yet clinical deployment demands robustness across acquisiti

researcharxiv-cs-cv
28 Jul 2026
Model Releases

Beyond Appearance: A Multi-cue Framework and Large-scale Benchmark for Pedestrian Association and Tracking on Mobile Aerial-Ground Platforms

DGX agent

arXiv:2607.23803v1 Announce Type: new Abstract: Multi-view Multi-object Association and Tracking (MvMoAT) associates objects across camera views and tracks them over time, supporting identity persiste

model-releasesarxiv-cs-cv
28 Jul 2026
Research

Beyond Error-vs-Discard Characteristic: Toward Stable and Reliable Evaluation for Face Image Quality Assessment

DGX agent

arXiv:2607.22752v1 Announce Type: new Abstract: Face Image Quality Assessment (FIQA) aims to estimate the utility of facial images for reliable recognition. The evaluation of FIQA methods is predomina

researcharxiv-cs-cv
28 Jul 2026
Tutorials

BeyondFusion: Self-Aligned Latent Diffusion for Calibration-Free Infrared Super-Resolution and Infrared-Visible Fusion

DGX agent

arXiv:2607.24110v1 Announce Type: new Abstract: Mobile infrared-visible imaging typically pairs a compact infrared sensor with a high-resolution visible camera for complementary perception. While cros

tutorialsarxiv-cs-cv
28 Jul 2026
Safety

Breaking the Synthetic-Real Domain Shortcut for Training-Free Generative Replay-based Class Incremental Learning

DGX agent

arXiv:2607.22994v1 Announce Type: new Abstract: Class-incremental learning (CIL) requires models to continuously acquire new knowledge while avoiding catastrophic forgetting. While exemplar replay is

safetyarxiv-cs-cv
28 Jul 2026
Research

Calibration-Free 3D Multi-Camera People Tracking for Indoor Environment

DGX agent

arXiv:2607.22731v1 Announce Type: new Abstract: Multi-Camera People Tracking (MCPT) traditionally relies on precise intrinsic and extrinsic camera calibration to project 2D detections into a unified 3

researcharxiv-cs-cv
28 Jul 2026
Model Releases

CameraAnything: Refilming Videos with Arbitrary Camera Control

DGX agent

arXiv:2607.24591v1 Announce Type: new Abstract: We introduce CameraAnything, the first unified framework for camera controlled video editing that enables joint control of both intrinsic and extrinsic

model-releasesarxiv-cs-cv
28 Jul 2026
Research

Cascade Forgery Mining Network for Fingerprint Presentation Attack Detection

DGX agent

arXiv:2607.24090v1 Announce Type: new Abstract: Fingerprint Presentation Attack Detection (PAD) is a critical component of fingerprint identification systems, serving as a protective measure against u

researcharxiv-cs-cv
28 Jul 2026
Model Releases

Child-Oriented AIGC Video Risk Reviewing: A Benchmark and Knowledge-Supported Iterative Reasoning Framework

DGX agent

arXiv:2607.22715v1 Announce Type: new Abstract: The rapid growth of Artificial Intelligence-generated content (AIGC) is reshaping video production and circulation, exposing children to an increasing v

model-releasesarxiv-cs-cv
28 Jul 2026
Research

Codebook Capacity Governs Perceptual Quality Across Resolutions in Hierarchical Discrete Video Compression

DGX agent

arXiv:2607.23366v1 Announce Type: cross Abstract: Learned video codecs based on continuous latent representations typically require resolution-specific retraining or rate-distortion (RD) recalibration

researcharxiv-cs-cv
28 Jul 2026
Research

Color Fundus Photography Analysis: Co-evolution of Data, Preprocessing, and Modeling toward Multimodal AI

DGX agent

arXiv:2607.23972v1 Announce Type: new Abstract: Color Fundus Photography (CFP) is a primary non-invasive imaging modality for large-scale screening of ophthalmic and systemic diseases. Existing survey

researcharxiv-cs-cv
28 Jul 2026
Safety

ConFusion: Continuous Fusion Space Learning for Fine-Grained Controllable Infrared and Visible Image Fusion

DGX agent

arXiv:2607.23600v1 Announce Type: new Abstract: Controllable infrared-visible image fusion aims to integrate complementary thermal and structural information with flexible region-aware modulation, pro

safetyarxiv-cs-cv
28 Jul 2026
Model Releases

Consistent Evidence, Robust Recognition: Faithful Attribution Regularization under Geometric Transformations

DGX agent

arXiv:2607.23835v1 Announce Type: new Abstract: Attribution methods are widely used to characterize the evidence underlying model predictions, yet their potential to improve model behavior remains und

model-releasesarxiv-cs-cv
28 Jul 2026
Model Releases

Contrastive Parameter Disentanglement for Multi-modal Remote Sensing Image Generation

DGX agent

arXiv:2607.23673v1 Announce Type: new Abstract: Existing remote sensing image generation methods are largely confined to single-modality synthesis and therefore fail to exploit the complementary infor

model-releasesarxiv-cs-cv
28 Jul 2026
Model Releases

Controllable Diversity in Normalization-Based Implicit Ensembles via Softmax-Temperature Modulation

DGX agent

arXiv:2607.23860v1 Announce Type: cross Abstract: Deep ensembles provide the most reliable uncertainty estimates in deep learning, but their cost grows linearly with the number of members. Implicit en

model-releasesarxiv-cs-cv
28 Jul 2026
Tutorials

Counterfactual Motion Reliability Learning for Robust UAV Tracking

DGX agent

arXiv:2607.23209v1 Announce Type: new Abstract: Infrared unmanned aerial vehicle (UAV) tracking is challenging because the target is often small, low-contrast, and easily confused with thermal distrac

tutorialsarxiv-cs-cv
28 Jul 2026
Tutorials

CrossSpine: Multi-scale Cross-sequence Attention with Anatomical Priors for Automated Pfirrmann Grading

DGX agent

arXiv:2607.22728v1 Announce Type: new Abstract: Automated grading of Lumbar Disc Degeneration is essential for the objective quantification of structural changes associated with low back pain. Observi

tutorialsarxiv-cs-cv
28 Jul 2026
Model Releases

DailyBench: A Unified Benchmark for AI-Generated and Manipulated Images from Modern Generative Models

DGX agent

arXiv:2607.24016v1 Announce Type: new Abstract: Recent advances in generative models have shifted AI-generated image detection from identifying easily distinguishable, fully synthetic images to identi

model-releasesarxiv-cs-cv
28 Jul 2026
Model Releases

DAP-Pose: Deep Temporal Alignment and Physics-aware Cross-modal Sensor Fusion for Robust Pose Estimation

DGX agent

arXiv:2607.23755v1 Announce Type: new Abstract: Robust and accurate pose estimation with multi-modal sensors is fundamental for autonomous vehicles and mobile robotic systems in complex environments.

model-releasesarxiv-cs-cv
28 Jul 2026
Safety

Data Pyramid for Embodied Manipulation

DGX agent

arXiv:2607.24744v1 Announce Type: cross Abstract: Multimodal foundation models learned to see and to speak by consuming the whole internet. Embodied agents admit no such shortcut, since they require d

safetyarxiv-cs-cv
28 Jul 2026
Research

DDVT: Dynamic Dual-level Vision Transformer Fusion Network for Answer Grounding in Visual Question Answering

DGX agent

arXiv:2607.23921v1 Announce Type: new Abstract: Answer grounding in visual question answering aims to locate the region from a given natural language question associated with the visual content of an

researcharxiv-cs-cv
28 Jul 2026
Applications

Deblur-Avatar: Animatable Avatars from Motion-Blurred Monocular Videos

DGX agent

arXiv:2501.13335v4 Announce Type: replace Abstract: We introduce a novel framework for modeling high-fidelity, animatable 3D human avatars from motion-blurred monocular video inputs. Motion blur is pr

applicationsarxiv-cs-cv
28 Jul 2026
Research

DeCoRAG: Cognitive Decoupling and Semantic-Aware Cropping for Complex Document Understanding

DGX agent

arXiv:2607.24554v1 Announce Type: cross Abstract: Advancing multimodal retrieval-augmented generation (RAG) for complex document understanding presents a formidable dual dilemma of accuracy and effici

researcharxiv-cs-cv
28 Jul 2026
Safety

Dementia Etiology Diagnosis via Collaborative Meta Knowledge Enhancement

DGX agent

arXiv:2607.22770v1 Announce Type: cross Abstract: Although artificial intelligence (AI) has shown promising performance in several medical tasks, accurate dementia etiology diagnosis with AI remains c

safetyarxiv-cs-cv
28 Jul 2026
Research

Denoising 3D images: robustness of persistent homology measures

DGX agent

arXiv:2607.24579v1 Announce Type: cross Abstract: When computing sub/super-level-set persistent homology (PH), the effect of noise may introduce millions of (short-lived) topological generators, prese

researcharxiv-cs-cv
28 Jul 2026
Safety

DeVA: Decoupled Video-Action Model with physical guidance for robot policy learning

DGX agent

arXiv:2607.24159v1 Announce Type: cross Abstract: Generalizable robot manipulation requires policies that can anticipate how visual scenes evolve while executing language instructions. While recent Vi

safetyarxiv-cs-cv
28 Jul 2026
Local Ai

Development of Vision-Language Model-based GNSS Spoofing Detection for Autonomous Vehicle Navigation

DGX agent

arXiv:2607.23962v1 Announce Type: new Abstract: Autonomous vehicles (AVs) depend on Global Navigation Satellite Systems (GNSS) for localization and navigation, making them vulnerable to spoofing attac

local-aiarxiv-cs-cv
28 Jul 2026
Research

DINOv3-MIL: Per-Kidney Multi-Label Tumour and Cyst Detection from Foundation-Model Patch Tokens on KiTS23

DGX agent

arXiv:2607.22687v1 Announce Type: new Abstract: Foundation vision models trained on natural images transfer to medical tasks without domain pre-training, but volumetric classification requires aggrega

researcharxiv-cs-cv
28 Jul 2026
Applications

Direction-adaptive Mamba: Spatial-Frequency Dual-Domain Collaborative Learning for PolSAR Image Classification

DGX agent

arXiv:2607.23464v1 Announce Type: cross Abstract: Deep learning dominates polarimetric synthetic aperture radar (PolSAR) image classification, with Mamba architectures serving as favorable backbones d

applicationsarxiv-cs-cv
28 Jul 2026
Model Releases

DishSeg24k: A Large-Scale Benchmark for Food Segmentation with Stochastic Expert Decoding

DGX agent

arXiv:2607.23070v1 Announce Type: new Abstract: Food segmentation is essential for applications such as intelligent catering, dietary assessment, and recommendation. However, existing benchmarks fail

model-releasesarxiv-cs-cv
28 Jul 2026
Agents

DispatchRAG: Grounding Emergency Dispatch Decisions in Real-World Protocols from Traffic Accident Video

DGX agent

arXiv:2607.23132v1 Announce Type: new Abstract: Assessing the severity of a traffic accident scenario is important to decide which emergency service to dispatch. Missing an ambulance dispatch on a ped

agentsarxiv-cs-cv
28 Jul 2026
Research

DreamStyle3D: Efficient 3D Stylized Asset Generation via Dual-Attention Disentanglement

DGX agent

arXiv:2607.24721v1 Announce Type: new Abstract: With the growth of gaming, animation, and virtual reality industries, the demand for efficient generation of stylized 3D assets is rapidly increasing. H

researcharxiv-cs-cv
28 Jul 2026
Model Releases

DY-LUT: Depth-Aware YCbCr Lookup Tables for Real-Time Underwater Image Enhancement

DGX agent

arXiv:2607.22801v1 Announce Type: cross Abstract: Underwater image enhancement is challenged by spatially non-uniform, wavelength-dependent attenuation. Propagation distance and wavelength govern this

model-releasesarxiv-cs-cv
28 Jul 2026
Model Releases

EditCLEVR: A Paired-Scene Intervention Benchmark for Compositional Faithfulness of Object-Centric Representations

DGX agent

arXiv:2607.22705v1 Announce Type: new Abstract: Object-centric learning aims to represent scenes as objects whose properties can be reused in new combinations. Existing evaluations usually score segme

model-releasesarxiv-cs-cv
28 Jul 2026
Tutorials

Effect of User-Prompted Priors on Semi-Automated Cancer Lesion Segmentation in Whole-Body Computed Tomography

DGX agent

arXiv:2607.24210v1 Announce Type: new Abstract: In clinical oncology studies, metastatic cancer is commonly evaluated using 'Response Evaluation Criteria in Solid Tumors' (RECIST), in which the diamet

tutorialsarxiv-cs-cv
28 Jul 2026
Research

Effective Receptive Field Ordering Matters for Infrared Small Target Detection

DGX agent

arXiv:2607.23994v1 Announce Type: new Abstract: In this work, we investigate a previously unexplored architectural dimension for infrared small target detection: the organization of effective receptiv

researcharxiv-cs-cv
28 Jul 2026
Local Ai

Embeddings based Anomaly Detection for Cleaning Global Crop Type Reference Datasets

DGX agent

arXiv:2607.23908v1 Announce Type: new Abstract: High quality reference data remain a critical bottleneck for crop-type mapping at any spatial and temporal scale. Operational systems such as WorldCerea

local-aiarxiv-cs-cv
28 Jul 2026
Hardware

Energy Constrained Hierarchical Underwater Monitoring via Local Multi-Agent RAG

DGX agent

arXiv:2607.24313v1 Announce Type: cross Abstract: Marine life monitoring is limited by strict energy constraints, poor underwater connectivity, and the high cost of transmitting raw multimodal data fr

hardwarearxiv-cs-cv
28 Jul 2026
Model Releases

Evaluation of Blood Vessel Segmentation Methods on Hard-to-Detect Vascular Structures

DGX agent

arXiv:2406.13128v2 Announce Type: replace Abstract: Due to the intricate structure of vascular trees, minor segmentation errors can significantly alter connectivity patterns and increase variability i

model-releasesarxiv-cs-cv
28 Jul 2026
Research

Event Driven Clustering Algorithm

DGX agent

arXiv:2602.00115v2 Announce Type: replace Abstract: This paper introduces a novel asynchronous, event-driven algorithm for real-time detection of small event clusters in event camera data. Similar to

researcharxiv-cs-cv
28 Jul 2026
Safety

Face Age Verification Vulnerabilities Under Simple Appearance Manipulations

DGX agent

arXiv:2607.24194v1 Announce Type: new Abstract: Online platforms increasingly rely on automated age estimation systems to enforce minimum-age policies. Focusing on vision-based models designed for thi

safetyarxiv-cs-cv
28 Jul 2026
Applications

Farm-LightSeek: An Edge-centric Multimodal Agricultural IoT Data Analytics Framework with Lightweight LLMs

DGX agent

arXiv:2506.03168v2 Announce Type: replace Abstract: Amid the challenges posed by global population growth and climate change, traditional agricultural Internet of Things (IoT) systems is currently und

applicationsarxiv-cs-cv
28 Jul 2026
Research

Fast Fourier Convolutional GAN for 30 m Clear-Sky Land Surface Temperature Gap-Free Reconstruction

DGX agent

arXiv:2607.22734v1 Announce Type: new Abstract: Satellite-derived Land Surface Temperature (LST) provides spatially comprehensive data that ground stations cannot match. However, its utility is freque

researcharxiv-cs-cv
28 Jul 2026
← Previous
1…3637383940…261
Next →