AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,515 results
Model Releases

Group Orthogonal Low-Rank Adaptation for RGB-T Tracking

DGX agent

arXiv:2512.05359v2 Announce Type: replace Abstract: Parameter-efficient fine-tuning has emerged as a promising paradigm in RGB-T tracking, enabling downstream task adaptation by freezing pretrained pa

model-releasesarxiv-cs-cv
28 Apr 2026
Research
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

GS-DOT: Gaussian splatting-based image reconstruction for diffuse optical tomography

DGX agent

arXiv:2604.23675v1 Announce Type: cross Abstract: This work presents GS-DOT, a novel image reconstruction framework based on Gaussian Splatting (GS) for diffuse optical tomography (DOT). Inspired by G

researcharxiv-cs-cv
28 Apr 2026
Research

H-SemiS: Hierarchical Fusion of Semi and Self-Supervised Learning for Knee Osteoarthritis Severity Grading

DGX agent

arXiv:2604.23335v1 Announce Type: new Abstract: Knee osteoarthritis (KOA) is a degenerative joint disease that can lead to chronic pain, reduced mobility, and long-term disability. Automated severity

researcharxiv-cs-cv
28 Apr 2026
Model Releases

HAC: Parameter-Efficient Hyperbolic Adaptation of CLIP for Zero-Shot VQA

DGX agent

arXiv:2604.23665v1 Announce Type: new Abstract: Recent advances in representation learning have shown that hyperbolic geometry can offer a more expressive alternative to the Euclidean embeddings used

model-releasesarxiv-cs-cv
28 Apr 2026
Hardware

Hallo-Live: Real-Time Streaming Joint Audio-Video Avatar Generation with Asynchronous Dual-Stream and Human-Centric Preference Distillation

DGX agent

arXiv:2604.23632v1 Announce Type: new Abstract: Real-time text-driven joint audio-video avatar generation requires jointly synthesizing portrait video and speech with high fidelity and precise synchro

hardwarearxiv-cs-cv
28 Apr 2026
Safety

Hierarchical Prototype-based Domain Priors for Multiple Instance Learning in Multimodal Histopathology Analysis

DGX agent

arXiv:2604.23982v1 Announce Type: new Abstract: Digital pathology has fundamentally altered diagnostic workflows by enabling the computational analysis of gigapixel Whole Slide Images (WSIs), yet effe

safetyarxiv-cs-cv
28 Apr 2026
Model Releases

Hierarchical Spatio-Channel Clustering for Efficient Model Compression in Medical Image Analysis

DGX agent

arXiv:2604.23375v1 Announce Type: new Abstract: Convolutional neural networks (CNNs) have become increasingly difficult to deploy in resource-constrained environments due to their large memory and com

model-releasesarxiv-cs-cv
28 Apr 2026
Research

Identity-Decoupled Anonymization for Visual Evidence in Multi-modal Retrieval-Augmented Generation

DGX agent

arXiv:2604.23584v1 Announce Type: new Abstract: Multi-modal retrieval-augmented generation (MRAG) systems retrieve visual evidence from large image corpora to ground the responses of large multi-modal

researcharxiv-cs-cv
28 Apr 2026
Model Releases

ImmerIris: A Large-Scale Dataset and Benchmark for Off-Axis and Unconstrained Iris Recognition in Immersive Applications

DGX agent

arXiv:2510.10113v3 Announce Type: replace Abstract: Recently, iris recognition is regaining prominence in immersive applications such as extended reality as a means of seamless user identification. Th

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

Improving Vision-language Models with Perception-centric Process Reward Models

DGX agent

arXiv:2604.24583v1 Announce Type: new Abstract: Recent advancements in reinforcement learning with verifiable rewards (RLVR) have significantly improved the complex reasoning ability of vision-languag

model-releasesarxiv-cs-cv
28 Apr 2026
Research

Infrastructure-Guided Connectivity-Enhanced Road Crack Detection and Estimation

DGX agent

arXiv:2604.24616v1 Announce Type: new Abstract: In this paper, we report the world's first infrastructure-guided communication-enhanced road crack detection pipeline that is effective and implementabl

researcharxiv-cs-cv
28 Apr 2026
Model Releases

Insert In Style: A Zero-Shot Generative Framework for Harmonious Cross-Domain Object Composition

DGX agent

arXiv:2511.15197v2 Announce Type: replace Abstract: Reference-based object composition involves integrating foreground reference image with background scene to produce harmonious fused image. This tas

model-releasesarxiv-cs-cv
28 Apr 2026
Safety

INSIGHT: Indoor Scene Intelligence from Geometric-Semantic Hierarchy Transfer for Public~Safety

DGX agent

arXiv:2604.23095v1 Announce Type: new Abstract: Indoor environments lack the spatial intelligence infrastructure that GPS provides outdoors; first responders arriving at unfamiliar buildings typically

safetyarxiv-cs-cv
28 Apr 2026
Research

Instance Awareness of Multi-class Semantic Segmentation Loss Functions

DGX agent

arXiv:2604.24276v1 Announce Type: new Abstract: Instance-sensitive losses for semantic segmentation such as blob loss and CC loss were designed to address instance imbalance, ensuring small lesions ge

researcharxiv-cs-cv
28 Apr 2026
Research

Instruction-Free Tuning of Large Vision Language Models for Medical Instruction Following

DGX agent

arXiv:2603.19482v2 Announce Type: replace Abstract: Large vision language models (LVLMs) have demonstrated impressive performance across a wide range of tasks. These capabilities largely stem from vis

researcharxiv-cs-cv
28 Apr 2026
Applications

IoT-Enhanced CNN-Based Labelled Crack Detection for Additive Manufacturing Image Annotation in Industry 4.0

DGX agent

arXiv:2604.22857v1 Announce Type: new Abstract: This paper presents an IoT-enhanced deep learning framework for automated crack detection in Additive Manufacturing (AM) surfaces using convolutional ne

applicationsarxiv-cs-cv
28 Apr 2026
Research

ISExplore:Informative Segment Selection for Efficient Personalized 3D Talking Face Generation

DGX agent

arXiv:2511.07940v2 Announce Type: replace Abstract: Talking Face Generation (TFG) methods based on Neural Radiance Fields (NeRF) and 3D Gaussian Splatting (3DGS) have recently achieved impressive prog

researcharxiv-cs-cv
28 Apr 2026
Safety

iWatchRoad: Scalable Detection and Geospatial Visualization of Potholes for Smart Cities

DGX agent

arXiv:2508.10945v2 Announce Type: replace Abstract: Potholes on the roads are a serious hazard and maintenance burden. This poses a significant threat to road safety and vehicle longevity, especially

safetyarxiv-cs-cv
28 Apr 2026
Safety

JSSFF: A Joint Structural-Semantic Fusion Framework for Remote Sensing Image Captioning

DGX agent

arXiv:2604.24031v1 Announce Type: new Abstract: The encoder-decoder framework has become widely popular nowadays. In this model, the encoder extracts informative visual features from an input image, a

safetyarxiv-cs-cv
28 Apr 2026
Safety

KAConvNet: Kolmogorov-Arnold Convolutional Networks for Vision Recognition

DGX agent

arXiv:2604.23320v1 Announce Type: new Abstract: The Convolutional Neural Networks (CNNs) have been the dominant and effective approach for general computer vision tasks. Recently, Kolmogorov-Arnold ne

safetyarxiv-cs-cv
28 Apr 2026
Research

Keypoint-based Dynamic Object 6-DoF Pose Tracking via Event Camera

DGX agent

arXiv:2604.23387v1 Announce Type: new Abstract: Accurate 6-DoF pose estimation of objects is critical for robots to perform precise manipulation tasks. However, for dynamic object pose estimation, con

researcharxiv-cs-cv
28 Apr 2026
Hardware

Latent Inter-Frame Pruning: A Training-Free Method Bridging Traditional Video Compression and Modern Diffusion Transformers for Efficient Generation

DGX agent

arXiv:2604.23858v1 Announce Type: new Abstract: Video generation, while capable of generating realistic videos, is computationally expensive and slow, prohibiting real-time applications. In this paper

hardwarearxiv-cs-cv
28 Apr 2026
Research

LatentBurst: A Fast and Efficient Multi Frame Super-Resolution for Hexadeca-Bayer Pattern CIS images

DGX agent

arXiv:2604.23268v1 Announce Type: new Abstract: This paper introduces a novel multi frame super-resolution network (MFSR) for burst hexadeca Bayer pattern Contact Image Sensor (CIS) images, which incl

researcharxiv-cs-cv
28 Apr 2026
Applications

LatentStealth: Unnoticeable and Efficient Adversarial Attacks on Expressive Human Pose and Shape Estimation

DGX agent

arXiv:2505.12009v2 Announce Type: replace Abstract: Expressive human pose and shape estimation (EHPS) plays a central role in digital human generation, particularly in live-streaming applications. How

applicationsarxiv-cs-cv
28 Apr 2026
Safety

LAVA: Layered Audio-Visual Anti-tampering Watermarking for Robust Deepfake Detection and Localization

DGX agent

arXiv:2604.23957v1 Announce Type: new Abstract: Proactive watermarking offers a promising approach for deepfake tamper detection and localization in short-form videos. However, existing methods often

safetyarxiv-cs-cv
28 Apr 2026
Research

Learning an Image Editing Model without Image Editing Pairs

DGX agent

arXiv:2510.14978v2 Announce Type: replace Abstract: Recent image editing models have achieved impressive results while following natural language editing instructions, but they rely on supervised fine

researcharxiv-cs-cv
28 Apr 2026
Tutorials

Learning Binary Sampling Patterns for Single-Pixel Imaging using Bilevel Optimisation

DGX agent

arXiv:2508.19068v2 Announce Type: replace Abstract: Single-Pixel Imaging (SPI) enables the reconstruction of objects using a single detector through sequential illuminations with structured light patt

tutorialsarxiv-cs-cv
28 Apr 2026
Safety

Learning from Imperfect Text Guidance: Robust Long-Tail Visual Recognition with High-Noise Label

DGX agent

arXiv:2604.23125v1 Announce Type: new Abstract: Real-world data often exhibit long-tailed distributions with numerous noisy labels, substantially degrading the performance of deep models. While prior

safetyarxiv-cs-cv
28 Apr 2026
Local Ai

Learning from Noisy Prompts: Saliency-Guided Prompt Distillation for Robust Segmentation with SAM

DGX agent

arXiv:2604.23314v1 Announce Type: new Abstract: Segmentation is central to clinical diagnosis and monitoring, yet the reliability of modern foundation models in medical imaging still depends on the av

local-aiarxiv-cs-cv
28 Apr 2026
Tutorials

Learning Scene-Level Signed Directional Distance Function with Ellipsoidal Priors and Neural Residuals

DGX agent

arXiv:2503.20066v2 Announce Type: replace-cross Abstract: Dense reconstruction and differentiable rendering are fundamental tightly connected operations in 3D vision and computer graphics. Recent neur

tutorialsarxiv-cs-cv
28 Apr 2026
Applications

Learning to Decipher from Pixels -- A Case Study of Copiale

DGX agent

arXiv:2604.23683v1 Announce Type: new Abstract: Historical encrypted manuscripts require both paleographic interpretation of cipher symbols and cryptanalytic recovery of plaintext. Most existing compu

applicationsarxiv-cs-cv
28 Apr 2026
Agents

Learning to Identify Out-of-Distribution Objects for 3D LiDAR Anomaly Segmentation

DGX agent

arXiv:2604.23604v1 Announce Type: new Abstract: Understanding the surrounding environment is fundamental in autonomous driving and robotic perception. Distinguishing between known classes and previous

agentsarxiv-cs-cv
28 Apr 2026
Model Releases

Learning Under Low Illumination: A Dataset and Algorithm for Traffic Sign Recognition

DGX agent

arXiv:2511.17183v2 Announce Type: replace Abstract: Traffic signboards are vital for road safety and intelligent transportation systems, enabling navigation and autonomous driving. Yet, recognizing tr

model-releasesarxiv-cs-cv
28 Apr 2026
Safety

LearnPruner: Rethinking Attention-based Token Pruning in Vision Language Models

DGX agent

arXiv:2604.23950v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have recently demonstrated remarkable capabilities in visual understanding and reasoning, but they also impose significant

safetyarxiv-cs-cv
28 Apr 2026
Safety

Less is More in Semantic Space: Intrinsic Decoupling via Clifford-M for Fundus Image Classification

DGX agent

arXiv:2603.20806v2 Announce Type: replace Abstract: Multi-label fundus diagnosis requires features that capture both fine-grained lesions and large-scale retinal structure. Many multi-scale medical vi

safetyarxiv-cs-cv
28 Apr 2026
Research

Leveraging Spatial Transcriptomics as Alternative to Manual Annotations for Deep Learning-Based Nuclei Analysis

DGX agent

arXiv:2604.23481v1 Announce Type: new Abstract: Deep learning-based nuclei segmentation and classification in pathology images typically rely on large-scale pixel-level manual annotations, which are c

researcharxiv-cs-cv
28 Apr 2026
Research

Light 'em Up: Enabling Few-Shot Low-Light 3D Gaussian Splatting with Multi-Scale Explicit Retinex Illumination Decoupling

DGX agent

arXiv:2604.24053v1 Announce Type: new Abstract: Full 360^irc novel view synthesis under low-light conditions remains challenging. Insufficient illumination, noise amplification, and view-dependent pho

researcharxiv-cs-cv
28 Apr 2026
Research

LLMind: Bio-inspired Training-free Adaptive Visual Representations for Vision-Language Models

DGX agent

arXiv:2603.14882v2 Announce Type: replace Abstract: Vision-Language Models (VLMs) typically assume a uniform spatial fidelity across the entire field of view of visual inputs, dedicating equal precisi

researcharxiv-cs-cv
28 Apr 2026
Safety

LoGeR: Long-Context Geometric Reconstruction with Hybrid Memory

DGX agent

arXiv:2603.03269v2 Announce Type: replace Abstract: Feedforward geometric foundation models achieve strong short-window reconstruction, yet scaling them to minutes-long videos is bottlenecked by quadr

safetyarxiv-cs-cv
28 Apr 2026
Model Releases

Lost in the Vibrations: Vision Language Models Fail the Dynamic Gauges Test

DGX agent

arXiv:2604.22829v1 Announce Type: new Abstract: The digital transformation of industrial manufacturing increasingly relies on the ability of autonomous robots to interact with legacy infrastructure, p

model-releasesarxiv-cs-cv
28 Apr 2026
Tutorials

LunarDepthNet: Generation of Digital Elevation Models using Deep Learning and Monocular Satellite Images

DGX agent

arXiv:2604.22848v1 Announce Type: new Abstract: Recent times have seen an increase in demand of high quality Digital Elevation Models (DEMs) for the lunar surface, because they are highly important fo

tutorialsarxiv-cs-cv
28 Apr 2026
Model Releases

Majorization-Guided Test-Time Adaptation for Vision-Language Models under Modality-Specific Shift

DGX agent

arXiv:2604.24602v1 Announce Type: new Abstract: Vision-language models transfer well in zero-shot settings, but at deployment the visual and textual branches often shift asymmetrically. Under this con

model-releasesarxiv-cs-cv
28 Apr 2026
Research

Mammographic Lesion Segmentation with Lightweight Models: A Comparative Study

DGX agent

arXiv:2604.23899v1 Announce Type: new Abstract: Breast cancer is a leading cause of cancer-related mortality among women worldwide, with mammography as the primary screening tool. While deep learning

researcharxiv-cs-cv
28 Apr 2026
Research

MARRS: Masked Autoregressive Unit-based Reaction Synthesis

DGX agent

arXiv:2505.11334v4 Announce Type: replace Abstract: This work aims at a challenging task: human action-reaction synthesis, i.e., generating human reactions conditioned on the action sequence of anothe

researcharxiv-cs-cv
28 Apr 2026
Research

MeshLAM: Feed-Forward One-Shot Animatable Textured Mesh Avatar Reconstruction

DGX agent

arXiv:2604.22865v1 Announce Type: new Abstract: We introduce MeshLAM, a feed-forward framework for one-shot animatable mesh head reconstruction that generates high-fidelity, animatable 3D head avatars

researcharxiv-cs-cv
28 Apr 2026
Research

Micro-Expression-Aware Avatar Fingerprinting via Inter-Frame Feature Differencing

DGX agent

arXiv:2604.23247v1 Announce Type: new Abstract: Avatar fingerprinting, i.e., verifying who drives a synthetic talking-head video rather than whether it is real, is a critical safeguard for authorized

researcharxiv-cs-cv
28 Apr 2026
Safety

MIRAGE: A Micro-Interaction Relational Architecture for Grounded Exploration in Multi-Figure Artworks

DGX agent

arXiv:2604.23788v1 Announce Type: new Abstract: Appreciating multi-figure paintings requires understanding how characters relate through subtle cues like gaze alignment, gesture, and spatial arrangeme

safetyarxiv-cs-cv
28 Apr 2026
Research

Monocular Depth Estimation via Neural Network with Learnable Algebraic Group and Ring Structures

DGX agent

arXiv:2604.24328v1 Announce Type: new Abstract: Monocular depth estimation (MDE) has witnessed remarkable progress driven by Convolutional Neural Networks and transformer-based architectures. However,

researcharxiv-cs-cv
28 Apr 2026
← Previous
1…212213214215216…261
Next →