AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,515 results
Research

A B-Spline Function Based 3D Point Cloud Unwrapping Scheme for 3D Fingerprint Recognition and Identification

DGX agent

arXiv:2604.16546v1 Announce Type: new Abstract: Three-dimensional (3D) fingerprint recognition and identification offer several advantages over traditional two-dimensional (2D) recognition systems. Th

researcharxiv-cs-cv
21 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

A Benchmark Study of Segmentation Models and Adaptation Strategies for Landslide Detection from Satellite Imagery

DGX agent

arXiv:2604.16663v1 Announce Type: new Abstract: Landslide detection from high resolution satellite imagery is a critical task for disaster response and risk assessment, yet the relative effectiveness

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

A Comparative Evaluation of Geometric Accuracy in NeRF and Gaussian Splatting

DGX agent

arXiv:2604.18205v1 Announce Type: new Abstract: Recent advances in neural rendering have introduced numerous 3D scene representations. Although standard computer vision metrics evaluate the visual qua

model-releasesarxiv-cs-cv
21 Apr 2026
Research

A deep learning pipeline for PAM50 subtype classification using histopathology images and multi-objective patch selection

DGX agent

arXiv:2604.01798v2 Announce Type: replace Abstract: Breast cancer is a highly heterogeneous disease with diverse molecular profiles. The PAM50 gene signature is widely recognized as a standard for cla

researcharxiv-cs-cv
21 Apr 2026
Safety

A High-Accuracy Optical Music Recognition Method Based on Bottleneck Residual Convolutions

DGX agent

arXiv:2604.16446v1 Announce Type: new Abstract: Optical Music Recognition (OMR) aims to convert printed or handwritten music score images into editable symbolic representations. This paper presents an

safetyarxiv-cs-cv
21 Apr 2026
Local Ai

A Lightweight Transformer for Pain Recognition from Brain Activity

DGX agent

arXiv:2604.16491v1 Announce Type: new Abstract: Pain is a multifaceted and widespread phenomenon with substantial clinical and societal burden, making reliable automated assessment a critical objectiv

local-aiarxiv-cs-cv
21 Apr 2026
Safety

A Real-Time Bike-Pedestrian Safety System with Wide-Angle Perception and Evaluation Testbed for Urban Intersections

DGX agent

arXiv:2604.17046v1 Announce Type: new Abstract: Collisions between cyclists and pedestrians at urban intersections remain a persistent source of injuries, yet few systems attempt real-time warnings to

safetyarxiv-cs-cv
21 Apr 2026
Model Releases

A Survey of Spatial Memory Representations for Efficient Robot Navigation

DGX agent

arXiv:2604.16482v1 Announce Type: new Abstract: As vision-based robots navigate larger environments, their spatial memory grows without bound, eventually exhausting computational resources, particular

model-releasesarxiv-cs-cv
21 Apr 2026
Research

A Two-Stage Deep Learning Framework for Segmentation of Ten Gastrointestinal Organs from Coronal MR Enterography

DGX agent

arXiv:2604.17118v1 Announce Type: cross Abstract: Accurate segmentation of gastrointestinal (GI) organs in magnetic resonance enterography (MRE) is critical for diagnosing inflammatory bowel disease (

researcharxiv-cs-cv
21 Apr 2026
Research

A Two-Stage Multi-Modal MRI Framework for Lifespan Brain Age Prediction

DGX agent

arXiv:2604.16655v1 Announce Type: cross Abstract: The accurate quantification of brain age from MRI has emerged as an important biomarker of brain health. However, existing approaches are often restri

researcharxiv-cs-cv
21 Apr 2026
Applications

Active World-Model with 4D-informed Retrieval for Exploration and Awareness

DGX agent

arXiv:2604.16733v1 Announce Type: new Abstract: Physical awareness, especially in a large and dynamic environment, is shaped by sensing decisions that determine observability across space, time, and s

applicationsarxiv-cs-cv
21 Apr 2026
Hardware

AdaCluster: Adaptive Query-Key Clustering for Sparse Attention in Video Generation

DGX agent

arXiv:2604.18348v1 Announce Type: new Abstract: Video diffusion transformers (DiTs) suffer from prohibitive inference latency due to quadratic attention complexity. Existing sparse attention methods e

hardwarearxiv-cs-cv
21 Apr 2026
Model Releases

Adaptive Forensic Feature Refinement via Intrinsic Importance Perception

DGX agent

arXiv:2604.16879v1 Announce Type: new Abstract: With the rapid development of generative models and multimodal content editing technologies, the key challenge faced by synthetic image detection (SID)

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Adaptive Local Frequency Filtering for Fourier-Encoded Implicit Neural Representations

DGX agent

arXiv:2604.02846v2 Announce Type: replace Abstract: Fourier-encoded implicit neural representations (INRs) have shown strong capability in modeling continuous signals from discrete samples. However, c

model-releasesarxiv-cs-cv
21 Apr 2026
Agents

Adaptive Quantized Planetary Crater Detection System for Autonomous Space Exploration

DGX agent

arXiv:2508.18025v4 Announce Type: replace-cross Abstract: Autonomous planetary exploration demands real-time, high-fidelity environmental perception. Standard deep learning models require massive comp

agentsarxiv-cs-cv
21 Apr 2026
Tutorials

Adaptive receptive field-based spatial-frequency feature reconstruction network for few-shot fine-grained image classification

DGX agent

arXiv:2604.16936v1 Announce Type: new Abstract: Feature reconstruction techniques are widely applied for few-shot fine-grained image classification (FSFGIC). Our research indicates that one of the mai

tutorialsarxiv-cs-cv
21 Apr 2026
Research

Advancing Vision Transformer with Enhanced Spatial Priors

DGX agent

arXiv:2604.18549v1 Announce Type: new Abstract: In recent years, the Vision Transformer (ViT) has garnered significant attention within the computer vision community. However, the core component of Vi

researcharxiv-cs-cv
21 Apr 2026
Model Releases

Adverse-to-the-eXtreme Panoptic Segmentation: URVIS 2026 Study and Benchmark

DGX agent

arXiv:2604.16984v1 Announce Type: new Abstract: This paper presents the report of the URVIS 2026 challenge on adverse-to-extreme panoptic segmentation. As the first challenge of its kind, it attracted

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

AeroRAG: Structured Multimodal Retrieval-Augmented LLM for Fine-Grained Aerial Visual Reasoning

DGX agent

arXiv:2604.17889v1 Announce Type: new Abstract: Despite recent progress in multimodal large language models (MLLMs), reliable visual question answering in aerial scenes remains challenging. In such sc

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Agentic Large Language Models for Training-Free Neuro-Radiological Image Analysis

DGX agent

arXiv:2604.16729v1 Announce Type: new Abstract: State-of-the-art large language models (LLMs) show high performance in general visual question answering. However, a fundamental limitation remains: cur

model-releasesarxiv-cs-cv
21 Apr 2026
Research

AI Approach for MRI-only Full-Spine Vertebral Segmentation and 3D Reconstruction in Paediatric Scoliosis

DGX agent

arXiv:2604.17846v1 Announce Type: new Abstract: MRI is preferred over CT in paediatric imaging because it avoids ionising radiation, but its use in spine deformity assessment is largely limited by the

researcharxiv-cs-cv
21 Apr 2026
Applications

AI-based Waste Mapping for Addressing Climate-Exacerbated Flood Risk

DGX agent

arXiv:2604.18151v1 Announce Type: new Abstract: Urban flooding is a growing climate change-related hazard in rapidly expanding African cities, where inadequate waste management often blocks drainage s

applicationsarxiv-cs-cv
21 Apr 2026
Model Releases

AIM 2025 Rip Current Segmentation (RipSeg) Challenge Report

DGX agent

arXiv:2508.13401v3 Announce Type: replace Abstract: This report presents an overview of the AIM 2025 RipSeg Challenge, a competition designed to advance techniques for automatic rip current segmentati

model-releasesarxiv-cs-cv
21 Apr 2026
Tutorials

Aletheia: Physics-Conditioned Localized Artifact Attention (PhyLAA-X) for End-to-End Generalizable and Robust Deepfake Video Detection

DGX agent

arXiv:2604.16486v1 Announce Type: new Abstract: State-of-the-art deepfake detectors achieve near-perfect in-domain accuracy yet degrade under cross-generator shifts, heavy compression, and adversarial

tutorialsarxiv-cs-cv
21 Apr 2026
Research

Amortized Inverse Kinematics via Graph Attention for Real-Time Human Avatar Animation

DGX agent

arXiv:2604.16629v1 Announce Type: new Abstract: Inverse kinematics (IK) is a core operation in animation, robotics, and biomechanics: given Cartesian constraints, recover joint rotations under a known

researcharxiv-cs-cv
21 Apr 2026
Model Releases

An Uncertainty-Aware Loss Function Incorporating Fuzzy Logic: Application to MRI Brain Image Segmentation

DGX agent

arXiv:2604.16490v1 Announce Type: new Abstract: Accurate brain image segmentation, particularly for distinguishing various tissues from magnetic resonance imaging (MRI) images, plays a pivotal role in

model-releasesarxiv-cs-cv
21 Apr 2026
Local Ai

AnchorSeg: Language Grounded Query Banks for Reasoning Segmentation

DGX agent

arXiv:2604.18562v1 Announce Type: new Abstract: Reasoning segmentation requires models to ground complex, implicit textual queries into precise pixel-level masks. Existing approaches rely on a single

local-aiarxiv-cs-cv
21 Apr 2026
Research

AnyLift: Scaling Motion Reconstruction from Internet Videos via 2D Diffusion

DGX agent

arXiv:2604.17818v1 Announce Type: new Abstract: Reconstructing 3D human motion and human-object interactions (HOI) from Internet videos is a fundamental step toward building large-scale datasets of hu

researcharxiv-cs-cv
21 Apr 2026
Applications

Appearance-free Action Recognition: Zero-shot Generalization in Humans and a Two-Pathway Model

DGX agent

arXiv:2604.16675v1 Announce Type: new Abstract: Action recognition is a fundamental ability for social species. Yet, its underlying computations are not well understood. Classical psychophysical studi

applicationsarxiv-cs-cv
21 Apr 2026
Research

Applications of deep generative models to DNA reaction kinetics and to cryogenic electron microscopy

DGX agent

arXiv:2604.16851v1 Announce Type: cross Abstract: This dissertation explores how deep generative models can advance the analysis of challenging biological problems by integrating domain knowledge with

researcharxiv-cs-cv
21 Apr 2026
Model Releases

Are We Using the Right Benchmark: An Evaluation Framework for Visual Token Compression Methods

DGX agent

arXiv:2510.07143v3 Announce Type: replace Abstract: Recent efforts to accelerate inference in Multimodal Large Language Models (MLLMs) have largely focused on visual token compression. The effectivene

model-releasesarxiv-cs-cv
21 Apr 2026
Safety

Asset Harvester: Extracting 3D Assets from Autonomous Driving Logs for Simulation

DGX agent

arXiv:2604.18468v1 Announce Type: new Abstract: Closed-loop simulation is a core component of autonomous vehicle (AV) development, enabling scalable testing, training, and safety validation before rea

safetyarxiv-cs-cv
21 Apr 2026
Research

AstroSURE: Learning to Remove Noise from Astronomical Images Without Ground Truth Data

DGX agent

arXiv:2604.16793v1 Announce Type: cross Abstract: In astronomical imaging, the low photon count of exposures necessitates extensive post-processing steps, including contamination removal and denoising

researcharxiv-cs-cv
21 Apr 2026
Research

Attention Is not Everything: Efficient Alternatives for Vision

DGX agent

arXiv:2604.17439v1 Announce Type: new Abstract: Recently computer vision has seen advancements mainly thanks to Transformer-based models. However many non-Transformer methods are still doing well bein

researcharxiv-cs-cv
21 Apr 2026
Research

Attention-ResUNet for Automated Fetal Head Segmentation

DGX agent

arXiv:2604.18148v1 Announce Type: new Abstract: Automated fetal head segmentation in ultrasound images is critical for accurate biometric measurements in prenatal care. While existing deep learning ap

researcharxiv-cs-cv
21 Apr 2026
Safety

Attention-space Contrastive Guidance for Efficient Hallucination Mitigation in LVLMs

DGX agent

arXiv:2601.13707v2 Announce Type: replace Abstract: Hallucinations in large vision--language models (LVLMs) often arise when language priors dominate over visual evidence, leading to object misidentif

safetyarxiv-cs-cv
21 Apr 2026
Local Ai

Attraction, Repulsion, and Friction: Introducing DMF, a Friction-Augmented Drifting Model

DGX agent

arXiv:2604.18194v1 Announce Type: cross Abstract: Drifting Models [Deng et al., 2026] train a one-step generator by evolving samples under a kernel-based drift field, avoiding ODE integration at infer

local-aiarxiv-cs-cv
21 Apr 2026
Research

Authenticated Contradictions from Desynchronized Provenance and Watermarking

DGX agent

arXiv:2603.02378v2 Announce Type: replace-cross Abstract: Cryptographic provenance standards such as C2PA and invisible watermarking are positioned as complementary defenses for content authentication

researcharxiv-cs-cv
21 Apr 2026
Research

Automated Palynological Analysis System: Integrating Deep Metric Learning and U^{2}-Net Detection in Hinfty bright field microscopy

DGX agent

arXiv:2604.16743v1 Announce Type: new Abstract: Traditional melissopalynology is a time-consuming and subjective process, often taking 4-6 hours per sample. We present an automated, high-throughput mi

researcharxiv-cs-cv
21 Apr 2026
Local Ai

Automated Road Crack Localization to Guide Highway Maintenance

DGX agent

arXiv:2601.16737v2 Announce Type: replace Abstract: Highway networks are crucial for economic prosperity. Climate change-induced temperature fluctuations are exacerbating stress on road pavements, res

local-aiarxiv-cs-cv
21 Apr 2026
Agents

Autonomous Unmanned Aircraft Systems for Enhanced Search and Rescue of Drowning Swimmers: Image-Based Localization and Mission Simulation

DGX agent

arXiv:2604.18088v1 Announce Type: new Abstract: Drowning is an omnipresent risk associated with any activity on or in the water, and rescuing a drowning person is particularly challenging because of t

agentsarxiv-cs-cv
21 Apr 2026
Agents

AutoVQA-G: Self-Improving Agentic Framework for Automated Visual Question Answering and Grounding Annotation

DGX agent

arXiv:2604.17488v1 Announce Type: new Abstract: Manual annotation of high-quality visual question answering with grounding (VQA-G) datasets, which pair visual questions with evidential grounding, is c

agentsarxiv-cs-cv
21 Apr 2026
Research

AvatarPointillist: AutoRegressive 4D Gaussian Avatarization

DGX agent

arXiv:2604.04787v2 Announce Type: replace Abstract: We introduce AvatarPointillist, a novel framework for generating dynamic 4D Gaussian avatars from a single portrait image. At the core of our method

researcharxiv-cs-cv
21 Apr 2026
Model Releases

AVRT: Audio-Visual Reasoning Transfer through Single-Modality Teachers

DGX agent

arXiv:2604.16617v1 Announce Type: new Abstract: Recent advances in reasoning models have shown remarkable progress in text-based domains, but transferring those capabilities to multimodal settings, e.

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

AWPD: Frequency Shield Network for Agnostic Watermark Presence Detection

DGX agent

arXiv:2603.06723v3 Announce Type: replace Abstract: Invisible watermarks, as an essential technology for image copyright protection, have been widely deployed with the rapid development of social medi

model-releasesarxiv-cs-cv
21 Apr 2026
Safety

Back into Plato's Cave: Examining Cross-modal Representational Convergence at Scale

DGX agent

arXiv:2604.18572v1 Announce Type: new Abstract: The Platonic Representation Hypothesis suggests that neural networks trained on different modalities (e.g., text and images) align and eventually conver

safetyarxiv-cs-cv
21 Apr 2026
Research

BARD: Bridging AutoRegressive and Diffusion Vision-Language Models Via Highly Efficient Progressive Block Merging and Stage-Wise Distillation

DGX agent

arXiv:2604.16514v1 Announce Type: new Abstract: Autoregressive vision-language models (VLMs) deliver strong multimodal capability, but their token-by-token decoding imposes a fundamental inference bot

researcharxiv-cs-cv
21 Apr 2026
Model Releases

BasketHAR: A Multimodal Dataset for Human Activity Recognition and Sport Analysis in Basketball Training Scenarios

DGX agent

arXiv:2604.17065v1 Announce Type: new Abstract: Human Activity Recognition (HAR) involves the automatic identification of user activities and has gained significant research interest due to its broad

model-releasesarxiv-cs-cv
21 Apr 2026
← Previous
1…224225226227228…261
Next →