AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,515 results
Model Releases

Anatomy-Aware Unsupervised Detection and Localization of Retinal Abnormalities in Optical Coherence Tomography

DGX agent

arXiv:2604.22139v1 Announce Type: new Abstract: Reliable automated analysis of Optical Coherence Tomography (OCT) imaging is crucial for diagnosing retinal disorders but faces a critical barrier: the

model-releasesarxiv-cs-cv
27 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

ArchSym: Detecting 3D-Grounded Architectural Symmetries in the Wild

DGX agent

arXiv:2604.22202v1 Announce Type: new Abstract: Symmetry detection is a fundamental problem in computer vision, and symmetries serve as powerful priors for downstream tasks. However, existing learning

model-releasesarxiv-cs-cv
27 Apr 2026
Tutorials

Are Natural-Domain Foundation Models Effective for Accelerated Cardiac MRI Reconstruction?

DGX agent

arXiv:2604.22557v1 Announce Type: cross Abstract: The emergence of large-scale pretrained foundation models has transformed computer vision, enabling strong performance across diverse downstream tasks

tutorialsarxiv-cs-cv
27 Apr 2026
Research

Are Video Models Emerging as Zero-Shot Learners and Reasoners in Medical Imaging?

DGX agent

arXiv:2510.10254v2 Announce Type: replace Abstract: Recent advances in large generative models have shown that simple autoregressive formulations, when scaled appropriately, can exhibit strong zero-sh

researcharxiv-cs-cv
27 Apr 2026
Safety

Beyond Chain-of-Thought: Rewrite as a Universal Interface for Generative Multimodal Embeddings

DGX agent

arXiv:2604.22280v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have emerged as a promising foundation for universal multimodal embeddings. Recent studies have shown that reas

safetyarxiv-cs-cv
27 Apr 2026
Research

Breaking Watermarks in the Frequency Domain: A Modulated Diffusion Attack Framework

DGX agent

arXiv:2604.22220v1 Announce Type: new Abstract: Digital image watermarking has advanced rapidly for copyright protection of generative AI, yet the comparatively limited progress in watermark attack te

researcharxiv-cs-cv
27 Apr 2026
Research

CAGE-SGG: Counterfactual Active Graph Evidence for Open-Vocabulary Scene Graph Generation

DGX agent

arXiv:2604.22274v1 Announce Type: new Abstract: Open-vocabulary scene graph generation (SGG) aims to describe visual scenes with flexible and fine-grained relation phrases beyond a fixed predicate voc

researcharxiv-cs-cv
27 Apr 2026
Model Releases

CharTide: Data-Centric Chart-to-Code Generation via Tri-Perspective Tuning and Inquiry-Driven Evolution

DGX agent

arXiv:2604.22192v1 Announce Type: new Abstract: Chart-to-code generation demands strict visual precision and syntactic correctness from Vision-Language Models (VLMs). However, existing approaches are

model-releasesarxiv-cs-cv
27 Apr 2026
Safety

Conditional Diffusion Posterior Alignment for Sparse-View CT Reconstruction

DGX agent

arXiv:2604.21960v1 Announce Type: cross Abstract: Computed Tomography (CT) is a widely used imaging modality in medical and industrial applications. To limit radiation exposure and measurement time, t

safetyarxiv-cs-cv
27 Apr 2026
Applications

Contrastive Semantic Projection: Faithful Neuron Labeling with Contrastive Examples

DGX agent

arXiv:2604.22477v1 Announce Type: new Abstract: Neuron labeling assigns textual descriptions to internal units of deep networks. Existing approaches typically rely on highly activating examples, often

applicationsarxiv-cs-cv
27 Apr 2026
Research

Decomposed Attention Fusion in MLLMs for Training-Free Video Reasoning Segmentation

DGX agent

arXiv:2510.19592v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) demonstrate strong video understanding by attending to visual tokens relevant to textual queries. To direct

researcharxiv-cs-cv
27 Apr 2026
Local Ai

Depth-Aware Rover: A Study of Edge AI and Monocular Vision for Real-World Implementation

DGX agent

arXiv:2604.22331v1 Announce Type: new Abstract: This study analyses simulated and real-world implementations of depth-aware rover navigation, highlighting the transition from stereo vision to monocula

local-aiarxiv-cs-cv
27 Apr 2026
Model Releases

Different Strokes for Different Folks: Writer Identification for Historical Arabic Manuscripts

DGX agent

arXiv:2604.22515v1 Announce Type: new Abstract: Handwritten Arabic manuscripts preserve the Arab world's intellectual and cultural heritage, and writer identification supports provenance, authenticity

model-releasesarxiv-cs-cv
27 Apr 2026
Tutorials

Distilling Vision Transformers for Distortion-Robust Representation Learning

DGX agent

arXiv:2604.22529v1 Announce Type: new Abstract: Self-supervised learning has achieved remarkable success in learning visual representations from clean data, yet remains challenging when clean observat

tutorialsarxiv-cs-cv
27 Apr 2026
Research

DocPrune:Efficient Document Question Answering via Background, Question, and Comprehension-aware Token Pruning

DGX agent

arXiv:2604.22281v1 Announce Type: new Abstract: Recent advances in vision-language models have demonstrated remarkable performance across diverse multi-modal tasks, including document question answeri

researcharxiv-cs-cv
27 Apr 2026
Research

Edit-aware RAW Reconstruction

DGX agent

arXiv:2512.05859v2 Announce Type: replace Abstract: Users frequently edit camera images post-capture to achieve their preferred photofinishing style. While editing in the RAW domain provides greater a

researcharxiv-cs-cv
27 Apr 2026
Research

Efficient Diffusion Distillation via Embedding Loss

DGX agent

arXiv:2604.22379v1 Announce Type: new Abstract: Recent advances in distilling expensive diffusion models into efficient few-step generators show significant promise. However, these methods typically d

researcharxiv-cs-cv
27 Apr 2026
Model Releases

EV-CLIP: Efficient Visual Prompt Adaptation for CLIP in Few-shot Action Recognition under Visual Challenges

DGX agent

arXiv:2604.22595v1 Announce Type: new Abstract: CLIP has demonstrated strong generalization in visual domains through natural language supervision, even for video action recognition. However, most exi

model-releasesarxiv-cs-cv
27 Apr 2026
Agents

Evaluation of image simulation open source solutions for simulation of synthetic images in lunar environment

DGX agent

arXiv:2604.22296v1 Announce Type: new Abstract: Synthetic image generation is one of the crucial input for planetary missions. It enables researchers and engineers to visualize planned planetary missi

agentsarxiv-cs-cv
27 Apr 2026
Research

EvFlow-GS: Event Enhanced Motion Deblurring with Optical Flow for 3D Gaussian Splatting

DGX agent

arXiv:2604.22183v1 Announce Type: new Abstract: Achieving sharp 3D reconstruction from motion-blurred images alone becomes challenging, motivating recent methods to incorporate event cameras, benefiti

researcharxiv-cs-cv
27 Apr 2026
Research

Evolving Thematic Map Design in Academic Cartography: A Thirty-Year Study Based on Multilingual Journals

DGX agent

arXiv:2604.22539v1 Announce Type: new Abstract: Thematic maps play a central role in academic communication, yet their large-scale design evolution has rarely been examined empirically. This study pre

researcharxiv-cs-cv
27 Apr 2026
Research

FeudalNav: A Simple Framework for Visual Navigation

DGX agent

arXiv:2602.06974v2 Announce Type: replace-cross Abstract: Visual navigation for robotics is inspired by the human ability to navigate environments using visual cues and memory, eliminating the need fo

researcharxiv-cs-cv
27 Apr 2026
Model Releases

FILTR: Extracting Topological Features from Pretrained 3D Models

DGX agent

arXiv:2604.22334v1 Announce Type: new Abstract: Recent advances in pretraining 3D point cloud encoders (e.g., Point-BERT, Point-MAE) have produced powerful models, whose abilities are typically evalua

model-releasesarxiv-cs-cv
27 Apr 2026
Model Releases

FLARE-BO: Fused Luminance and Adaptive Retinex Enhancement via Bayesian Optimisation for Low-Light Robotic Vision

DGX agent

arXiv:2604.22093v1 Announce Type: new Abstract: Reliable visual perception under low illumination remains a core challenge for autonomous robotic systems, where degraded image quality directly comprom

model-releasesarxiv-cs-cv
27 Apr 2026
Local Ai

Flow4DGS-SLAM: Optical Flow-Guided 4D Gaussian Splatting SLAM

DGX agent

arXiv:2604.22339v1 Announce Type: new Abstract: Handling the dynamic environments is a significant research challenge in Visual Simultaneous Localization and Mapping (SLAM). Recent research combines 3

local-aiarxiv-cs-cv
27 Apr 2026
Safety

FlowAnchor: Stabilizing the Editing Signal for Inversion-Free Video Editing

DGX agent

arXiv:2604.22586v1 Announce Type: new Abstract: We propose FlowAnchor, a training-free framework for stable and efficient inversion-free, flow-based video editing. Inversion-free editing methods have

safetyarxiv-cs-cv
27 Apr 2026
Research

Forecasting Solar Energy Using a Single Image

DGX agent

arXiv:2604.21982v1 Announce Type: new Abstract: Solar panels are increasingly deployed in cities on rooftops, walls, and urban infrastructure. Although the panel costs have fallen in recent years, the

researcharxiv-cs-cv
27 Apr 2026
Research

Generative Modeling of Neurodegenerative Brain Anatomy with 4D Longitudinal Diffusion Model

DGX agent

arXiv:2604.22700v1 Announce Type: new Abstract: Understanding and predicting the progression of neurodegenerative diseases remains a major challenge in medical AI, with significant implications for ea

researcharxiv-cs-cv
27 Apr 2026
Tutorials

GOSPA and T-GOSPA quasi-metrics for evaluation of multi-object tracking algorithms

DGX agent

arXiv:2507.13706v2 Announce Type: replace Abstract: This paper introduces two quasi-metrics for performance assessment of multi-object tracking (MOT) algorithms. One quasi-metric is an extension of th

tutorialsarxiv-cs-cv
27 Apr 2026
Research

HFS-TriNet: A Three-Branch Collaborative Feature Learning Network for Prostate Cancer Classification from TRUS Videos

DGX agent

arXiv:2604.22388v1 Announce Type: new Abstract: Transrectal ultrasound (TRUS) imaging is a cost-effective and non-invasive modality widely used in the diagnosis of prostate cancer. The computer-aided

researcharxiv-cs-cv
27 Apr 2026
Model Releases

Holo360D: A Large-Scale Real-World Dataset with Continuous Trajectories for Advancing Panoramic 3D Reconstruction and Beyond

DGX agent

arXiv:2604.22482v1 Announce Type: new Abstract: While feed-forward 3D reconstruction models have advanced rapidly, they still exhibit degraded performance on panoramas due to spherical distortions. Mo

model-releasesarxiv-cs-cv
27 Apr 2026
Safety

How Many Visual Levers Drive Urban Perception? Interventional Counterfactuals via Multiple Localised Edits

DGX agent

arXiv:2604.22103v1 Announce Type: cross Abstract: Street-view perception models predict subjective attributes such as safety at scale, but remain correlational: they do not identify which localized vi

safetyarxiv-cs-cv
27 Apr 2026
Applications

ICPR 2026 Competition on Low-Resolution License Plate Recognition

DGX agent

arXiv:2604.22506v1 Announce Type: new Abstract: Low-Resolution License Plate Recognition (LRLPR) remains a challenging problem in real-world surveillance scenarios, where long capture distances, compr

applicationsarxiv-cs-cv
27 Apr 2026
Safety

Improving Driver Drowsiness Detection via Personalized EAR/MAR Thresholds and CNN-Based Classification

DGX agent

arXiv:2604.22479v1 Announce Type: new Abstract: Driver drowsiness is a major cause of traffic accidents worldwide, posing a serious threat to public safety. Vision-based driver monitoring systems ofte

safetyarxiv-cs-cv
27 Apr 2026
Research

Inter-Stance: A Dyadic Multimodal Corpus for Conversational Stance Analysis

DGX agent

arXiv:2604.22739v1 Announce Type: new Abstract: Social interactions dominate our perceptions of the world and shape our daily behavior by attaching social meaning to acts as simple and spontaneous as

researcharxiv-cs-cv
27 Apr 2026
Model Releases

Knowledge Visualization: A Benchmark and Method for Knowledge-Intensive Text-to-Image Generation

DGX agent

arXiv:2604.22302v1 Announce Type: new Abstract: Recent text-to-image (T2I) models have demonstrated impressive capabilities in photorealistic synthesis and instruction following. However, their reliab

model-releasesarxiv-cs-cv
27 Apr 2026
Agents

Learning Reactive Human Motion Generation from Paired Interaction Data Using Transformer-Based Models

DGX agent

arXiv:2604.22164v1 Announce Type: new Abstract: Recent advances in deep learning have enabled the generation of videos from textual descriptions as well as the prediction of future sequences from inpu

agentsarxiv-cs-cv
27 Apr 2026
Model Releases

Long-tail Internet photo reconstruction

DGX agent

arXiv:2604.22714v1 Announce Type: new Abstract: Internet photo collections exhibit an extremely long-tailed distribution: a few famous landmarks are densely photographed and easily reconstructed in 3D

model-releasesarxiv-cs-cv
27 Apr 2026
Model Releases

LTBs-KAN: Linear-Time B-splines Kolmogorov-Arnold Networks

DGX agent

arXiv:2604.22034v1 Announce Type: cross Abstract: Kolmogorov-Arnold Networks (KANs) are a recent neural network architecture offering an alternative to Multilayer Perceptrons (MLPs) with improved expl

model-releasesarxiv-cs-cv
27 Apr 2026
Model Releases

MTT-Bench: Predicting Social Dominance in Mice via Multimodal Large Language Models

DGX agent

arXiv:2604.22492v1 Announce Type: cross Abstract: Understanding social dominance in animal behavior is critical for neuroscience and behavioral studies. In this work, we explore the capability of Mult

model-releasesarxiv-cs-cv
27 Apr 2026
Tutorials

Multimodal Diffusion to Mutually Enhance Polarized Light and Low Resolution EBSD Data

DGX agent

arXiv:2604.22212v1 Announce Type: cross Abstract: In spite of the utility of 3-D electron back-scattered diffraction (EBSD) microscopy, the data collection process can be time-consuming with serial-se

tutorialsarxiv-cs-cv
27 Apr 2026
Research

Non-Minimal Sampling and Consensus for Prohibitively Large Datasets

DGX agent

arXiv:2604.22518v1 Announce Type: new Abstract: We introduce NONSAC (Non-Minimal Sampling and Consensus), a general framework for robust and scalable model estimation from arbitrarily large datasets c

researcharxiv-cs-cv
27 Apr 2026
Research

NRGS: Neural Regularization for Robust 3D Semantic Gaussian Splatting

DGX agent

arXiv:2604.22439v1 Announce Type: new Abstract: We propose a neural regularization method that refines the noisy 3D semantic field produced by lifting multi-view inconsistent 2D features, in order to

researcharxiv-cs-cv
27 Apr 2026
Applications

Nuclear Diffusion Models for Low-Rank Background Suppression in Videos

DGX agent

arXiv:2509.20886v2 Announce Type: replace Abstract: Video sequences often contain structured noise and background artifacts that obscure dynamic content, posing challenges for accurate analysis and re

applicationsarxiv-cs-cv
27 Apr 2026
Model Releases

OccDirector: Language-Guided Behavior and Interaction Generation in 4D Occupancy Space

DGX agent

arXiv:2604.22240v1 Announce Type: new Abstract: Generative world models increasingly rely on 4D occupancy for realistic autonomous driving simulation. However, existing generation frameworks depend on

model-releasesarxiv-cs-cv
27 Apr 2026
Tutorials

One Shot Learning for Edge Detection on Point Clouds

DGX agent

arXiv:2604.22354v1 Announce Type: new Abstract: Each scanner possesses its unique characteristics and exhibits its distinct sampling error distribution. Training a network on a dataset that includes d

tutorialsarxiv-cs-cv
27 Apr 2026
Research

PAGaS: Pixel-Aligned 1DoF Gaussian Splatting for Depth Refinement

DGX agent

arXiv:2604.22129v1 Announce Type: new Abstract: Gaussian Splatting (GS) has emerged as an efficient approach for high-quality novel view synthesis. While early GS variants struggled to accurately mode

researcharxiv-cs-cv
27 Apr 2026
Tutorials

PASR: Pose-Aware 3D Shape Retrieval from Occluded Single Views

DGX agent

arXiv:2604.22658v1 Announce Type: new Abstract: Single-view 3D shape retrieval is a fundamental yet challenging task that is increasingly important with the growth of available 3D data. Existing appro

tutorialsarxiv-cs-cv
27 Apr 2026
← Previous
1…215216217218219…261
Next →