AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,515 results
Research

Rethinking Low-Light Image Enhancement: A Log-Domain Intensity--Chromaticity Decoupling Perspective

DGX agent

arXiv:2605.02627v1 Announce Type: new Abstract: Explicit reconstruction constraints derived from the decoupled representation are further imposed to suppress abnormal channel amplification and chromat

researcharxiv-cs-cv
5 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Rethinking Model Selection in VLM Through the Lens of Gromov-Wasserstein Distance

DGX agent

arXiv:2605.01325v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have enhanced traditional LLMs with visual capabilities through the integration of vision encoders. While recent works hav

safetyarxiv-cs-cv
5 May 2026
Tutorials

Rethinking the Need for Source Models: Source-Free Domain Adaptation from Scratch Guided by a Vision-Language Model

DGX agent

arXiv:2605.02604v1 Announce Type: new Abstract: Source-Free Domain Adaptation (SFDA) adapts source models to target domains without accessing source data, addressing privacy and transmission issues. H

tutorialsarxiv-cs-cv
5 May 2026
Safety

Retrieval-Guided Generation for Safer Histopathology Image Captioning

DGX agent

arXiv:2605.00893v1 Announce Type: new Abstract: Generative vision-language models can produce fluent medical image captions but remain prone to hallucination, over-specific diagnostic claims, and fact

safetyarxiv-cs-cv
5 May 2026
Model Releases

Retrieving Any Relevant Moments: Benchmark and Models for Generalized Moment Retrieval

DGX agent

arXiv:2605.02623v1 Announce Type: new Abstract: Video Moment Retrieval (VMR) aims to localize temporal segments in videos that correspond to a natural language query, but typically assumes only a sing

model-releasesarxiv-cs-cv
5 May 2026
Research

Revisiting Map Relations for Unsupervised Non-Rigid Shape Matching

DGX agent

arXiv:2310.11420v2 Announce Type: replace Abstract: We propose a novel unsupervised learning approach for non-rigid 3D shape matching. Our approach improves upon recent state-of-the art deep functiona

researcharxiv-cs-cv
5 May 2026
Local Ai

Robust Cross-Domain WiFi Fall Detection via Physics-Driven Attention-Enhanced Transformers

DGX agent

arXiv:2605.00869v1 Announce Type: cross Abstract: Device-free fall detection utilizing WiFi Channel State Information (CSI) has emerged as a promising, privacy-preserving solution for elderly health m

local-aiarxiv-cs-cv
5 May 2026
Applications

Robust Fundamental Matrix Estimation from Single Image Motion Blur

DGX agent

arXiv:2605.01552v1 Announce Type: new Abstract: In this paper, we introduce a challenging task: extracting a fundamental matrix from a single motion blurred image. For a camera moving in 3D during exp

applicationsarxiv-cs-cv
5 May 2026
Research

Robustness of Transformer-Based Fluence Map Prediction Under Clinically Realistic Perturbations

DGX agent

arXiv:2605.00904v1 Announce Type: new Abstract: Learning-based fluence map prediction offers a fast alternative to iterative inverse planning in intensity-modulated radiation therapy (IMRT), but its r

researcharxiv-cs-cv
5 May 2026
Research

SAIL: Structure-Aware Interpretable Learning for Anatomy-Aligned Post-hoc Explanations in OCT

DGX agent

arXiv:2605.02707v1 Announce Type: new Abstract: Optical coherence tomography (OCT), a commonly used retinal imaging modality, plays a central role in retinal disease diagnosis by providing high-resolu

researcharxiv-cs-cv
5 May 2026
Research

SaLF: Sparse Local Fields for Multi-Sensor Rendering in Real-Time

DGX agent

arXiv:2507.18713v2 Announce Type: replace Abstract: High-fidelity sensor simulation of light-based sensors such as cameras and LiDARs is critical for safe and accurate autonomy testing. Neural radianc

researcharxiv-cs-cv
5 May 2026
Model Releases

SAMamba3D: adapting Segment Anything for generalizable 3D segmentation of multiphase pore-scale images

DGX agent

arXiv:2605.00916v1 Announce Type: new Abstract: Reliable segmentation of multiphase pore-scale X-ray images of rocks is necessary to quantify fluid saturation, connectivity, and interfacial geometry.

model-releasesarxiv-cs-cv
5 May 2026
Research

Sample-wise Adaptive Weighting for Transfer Consistency in Adversarial Distillation

DGX agent

arXiv:2512.10275v2 Announce Type: replace Abstract: Adversarial distillation in the standard min-max adversarial training framework aims to transfer adversarial robustness from a large, robust teacher

researcharxiv-cs-cv
5 May 2026
Research

Scaling Sequence-to-Sequence Generative Neural Rendering

DGX agent

arXiv:2510.04236v3 Announce Type: replace Abstract: We present Kaleido, a family of generative models designed for photorealistic, unified object- and scene-level neural rendering. Kaleido operates on

researcharxiv-cs-cv
5 May 2026
Model Releases

Scaling Vision Transformers for Functional MRI with Flat Maps

DGX agent

arXiv:2510.13768v2 Announce Type: replace Abstract: We study the problem of training self-supervised foundation models for functional MRI. Our main contributions are: (1) we introduce a new model fami

model-releasesarxiv-cs-cv
5 May 2026
Research

ScribbleEdit: Synthetic Data for Image Editing with Scribbles and Text

DGX agent

arXiv:2605.01135v1 Announce Type: new Abstract: Recent progress in generative models has significantly advanced image editing capabilities, yet precise and intuitive user control remains difficult. Sp

researcharxiv-cs-cv
5 May 2026
Model Releases

Seeing Realism from Simulation: Efficient Video Transfer for Vision-Language-Action Data Augmentation

DGX agent

arXiv:2605.02757v1 Announce Type: new Abstract: Vision-language-action (VLA) models typically rely on large-scale real-world videos, whereas simulated data, despite being inexpensive and highly parall

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

Seeing the Scene Matters: Revealing Forgetting in Video Understanding Models with a Scene-Aware Long-Video Benchmark

DGX agent

arXiv:2603.27259v2 Announce Type: replace Abstract: Long video understanding (LVU) remains a core challenge in multimodal learning. Although recent vision-language models (VLMs) have made notable prog

model-releasesarxiv-cs-cv
5 May 2026
Research

Selective Attention-Based Network for Robust Infrared Small Target Detection

DGX agent

arXiv:2605.00886v1 Announce Type: new Abstract: Infrared small target detection (IRSTD) plays a pivotal role in a broad spectrum of mission-critical applications, including maritime surveillance, mili

researcharxiv-cs-cv
5 May 2026
Applications

Selective Correlation Based Knowledge Distillation for Ground Reaction Force Estimation

DGX agent

arXiv:2605.00888v1 Announce Type: new Abstract: Wearable sensor-based human gait analysis holds great promise in healthcare, rehabilitation, clinical diagnosis and monitoring, and sports activities. S

applicationsarxiv-cs-cv
5 May 2026
Model Releases

Self-Supervised Learning for Multimodal Non-Rigid 3D Shape Matching

DGX agent

arXiv:2303.10971v2 Announce Type: replace Abstract: The matching of 3D shapes has been extensively studied for shapes represented as surface meshes, as well as for shapes represented as point clouds.

model-releasesarxiv-cs-cv
5 May 2026
Research

Self-Supervised Spatial And Zero-Shot Angular Super-Resolution by Spatial-Angular Implicit Representation For Rotating-View SNR-Efficient Diffusion MRI

DGX agent

arXiv:2605.02575v1 Announce Type: new Abstract: Rotating-view thick-slice acquisition is highly SNR-efficient for mesoscale diffusion MRI (dMRI) but requires numerous rotating views to satisfy Nyquist

researcharxiv-cs-cv
5 May 2026
Local Ai

Semantic Context-aware mOdality fUsion Transformer (SCOUT): A Context-Aware Multimodal Transformer for Concept-Grounded Pathology Report Generation

DGX agent

arXiv:2605.01144v1 Announce Type: new Abstract: Whole-slide images (WSIs) present a fundamental challenge for computational pathology due to their extreme resolution, multi-scale heterogeneity, and th

local-aiarxiv-cs-cv
5 May 2026
Model Releases

SF20K Competition 2025: Summary and findings

DGX agent

arXiv:2605.01496v1 Announce Type: new Abstract: This report presents the results and findings of the first edition of the Short-Films 20K (SF20K) Competition, held in conjunction with the SLoMO Worksh

model-releasesarxiv-cs-cv
5 May 2026
Research

SHARP: Spectrum-aware Highly-dynamic Adaptation for Resolution Promotion in Remote Sensing Synthesis

DGX agent

arXiv:2603.21783v2 Announce Type: replace Abstract: Text-to-image generation powered by Diffusion Transformers (DiTs) has made remarkable strides, yet remote sensing (RS) synthesis lags behind due to

researcharxiv-cs-cv
5 May 2026
Research

SIAM: Head and Brain MRI Segmentation from Few High-Quality Templates via Synthetic Training

DGX agent

arXiv:2605.02737v1 Announce Type: new Abstract: Synthetic training has recently advanced brain MRI segmentation by enabling contrast-agnostic models trained entirely on generated data. However, most e

researcharxiv-cs-cv
5 May 2026
Safety

SIFT-VTON: Geometric Correspondence Supervision on Cross-Attention for Virtual Try-On

DGX agent

arXiv:2605.01296v1 Announce Type: new Abstract: Diffusion-based virtual try-on methods achieve photorealistic synthesis through cross-attention mechanisms that transfer garment features to target body

safetyarxiv-cs-cv
5 May 2026
Research

SignMAE: Segmentation-Driven Self-Supervised Learning for Sign Language Recognition

DGX agent

arXiv:2605.02094v1 Announce Type: new Abstract: Subtle hand differences make sign language recognition challenging, yet many existing methods rely on encoders pretrained on generic action datasets tha

researcharxiv-cs-cv
5 May 2026
Agents

SimPB++: Simultaneously Detecting 2D and 3D Objects from Multiple Cameras

DGX agent

arXiv:2605.01924v1 Announce Type: new Abstract: Simultaneous perception of 2D objects in perspective view and 3D objects in Bird's Eye View (BEV) is challenging for multi-camera autonomous driving. Ex

agentsarxiv-cs-cv
5 May 2026
Applications

Single Image Defogging Using a Fourth-Order Telegraph PDE Guided by Physical Haze Modeling

DGX agent

arXiv:2605.00878v1 Announce Type: new Abstract: In real-world scenarios, image defogging is an inverse problem due to unknown scene depth, atmospheric scattering, and the common absence of ground trut

applicationsarxiv-cs-cv
5 May 2026
Applications

Skeleton-Based Posture Classification to Promote Safer Walker-Assisted Gait in Older Adults

DGX agent

arXiv:2605.00890v1 Announce Type: new Abstract: Falls among older adults are a significant public health concern, leading to severe injuries, loss of independence, and increased healthcare costs. This

applicationsarxiv-cs-cv
5 May 2026
Applications

SlimDiffSR: Toward Lightweight and Efficient Remote Sensing Image Super-Resolution via Diffusion Model Distillation

DGX agent

arXiv:2605.02198v1 Announce Type: new Abstract: Diffusion models have recently achieved remarkable performance in image super-resolution (SR), but their high computational cost limits practical deploy

applicationsarxiv-cs-cv
5 May 2026
Safety

Sonar-GPS Fusion for Seabed Mapping in Turbid Shallow Waters with an Autonomous Surface Vehicle

DGX agent

arXiv:2605.01949v1 Announce Type: cross Abstract: Accurate seabed mapping is essential for habitat monitoring and infrastructure inspection. In turbid, shallow coastal waters, such as shellfish aquacu

safetyarxiv-cs-cv
5 May 2026
Agents

Sound Source Localization for Spatial Mapping of Surgical Actions in Dynamic Scenes

DGX agent

arXiv:2510.24332v3 Announce Type: replace-cross Abstract: Purpose: Surgical scene understanding is key to advancing computer-aided and intelligent surgical systems. Current approaches predominantly re

agentsarxiv-cs-cv
5 May 2026
Applications

Space-Time Forecasting of Dynamic Scenes with Motion-aware Gaussian Grouping

DGX agent

arXiv:2602.21668v2 Announce Type: replace Abstract: Forecasting dynamic scenes remains a fundamental challenge in computer vision, as limited observations make it difficult to capture coherent object-

applicationsarxiv-cs-cv
5 May 2026
Research

Sparse Representation Learning for Vessels

DGX agent

arXiv:2605.01382v1 Announce Type: new Abstract: Analyzing human vasculature and vessel-like, tubular structures, such as airways, is crucial for disease diagnosis and treatment. Current methods often

researcharxiv-cs-cv
5 May 2026
Research

SparseContrast: Dynamic Sparse Attention for Efficient and Accurate Contrastive Learning in Medical Imaging

DGX agent

arXiv:2605.00887v1 Announce Type: new Abstract: We propose SparseContrast, a new framework that merges dynamic sparse attention with contrastive learning for medical imaging, with a focus on chest X-r

researcharxiv-cs-cv
5 May 2026
Applications

SPAT: A Semantic Port-Aware Adaptive-Rate Transmission Protocol for Semantic Communication

DGX agent

arXiv:2605.00897v1 Announce Type: cross Abstract: With the evolution of 6G, semantic communication has emerged as a promising paradigm by prioritizing the delivery of task-relevant meaning over strict

applicationsarxiv-cs-cv
5 May 2026
Model Releases

SpecEdit: Training-Free Acceleration for Diffusion based Image Editing via Semantic Locking

DGX agent

arXiv:2605.02152v1 Announce Type: new Abstract: Diffusion-based image editing offers strong semantic controllability, but remains computationally expensive due to iterative high-resolution denoising o

model-releasesarxiv-cs-cv
5 May 2026
Safety

SpectraDINO: Bridging the Spectral Gap in Vision Foundation Models via Lightweight Adapters

DGX agent

arXiv:2605.02258v1 Announce Type: new Abstract: Vision Foundation Models (VFMs) pretrained on large-scale RGB data have demonstrated remarkable representation quality, yet their applicability to multi

safetyarxiv-cs-cv
5 May 2026
Model Releases

SplAttN: Bridging 2D and 3D with Gaussian Soft Splatting and Attention for Point Cloud Completion

DGX agent

arXiv:2605.01466v1 Announce Type: new Abstract: Although multi-modal learning has advanced point cloud completion, the theoretical mechanisms remain unclear. Recent works attribute success to the conn

model-releasesarxiv-cs-cv
5 May 2026
Research

SRGAN-CKAN: Expressive Super-Resolution with Nonlinear Functional Operators under Minimal Resources

DGX agent

arXiv:2605.01459v1 Announce Type: new Abstract: Single-Image Super-Resolution (SISR) aims to reconstruct a High-Resolution (HR) image from a Low-Resolution (LR) observation, a fundamentally ill-posed

researcharxiv-cs-cv
5 May 2026
Safety

StableMind: Source-Free Cross-Subject fMRI Decoding with Regularized Adaptation

DGX agent

arXiv:2605.02586v1 Announce Type: new Abstract: Existing cross-subject fMRI decoding methods typically train a model on multiple scanned subjects and then adapt it to a new subject using substantial p

safetyarxiv-cs-cv
5 May 2026
Model Releases

SteeringDiffusion: A Bottlenecked Activation Control Interface for Diffusion Models

DGX agent

arXiv:2605.01653v1 Announce Type: new Abstract: We introduce SteeringDiffusion, a bottlenecked activation-level control interface for diffusion models that exposes a smooth, monotonic, and runtime-adj

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

StereoMamba: Real-time and Robust Intraoperative Stereo Disparity Estimation via Long-range Spatial Dependencies

DGX agent

arXiv:2504.17401v2 Announce Type: replace Abstract: Stereo disparity estimation is crucial for obtaining depth information in robot-assisted minimally invasive surgery (RAMIS). While current deep lear

model-releasesarxiv-cs-cv
5 May 2026
Research

Stylistic Attribute Control in Latent Diffusion Models

DGX agent

arXiv:2605.02583v1 Announce Type: new Abstract: Text-to-image diffusion models have revolutionized image synthesis and editing, but precise control over stylistic attributes remains a challenge, often

researcharxiv-cs-cv
5 May 2026
Local Ai

Super-resolution of airborne laser scanning point clouds for forest inventory

DGX agent

arXiv:2605.02201v1 Announce Type: new Abstract: Airborne Laser Scanning (ALS) can collect point clouds across large areas, enabling large-scale forest inventory. However, ALS point clouds are sparse a

local-aiarxiv-cs-cv
5 May 2026
Model Releases

SurgCheck: Do Vision-Language Models Really Look at Images in Surgical VQA?

DGX agent

arXiv:2605.01911v1 Announce Type: new Abstract: Purpose: Vision-language models (VLMs) have shown promising performance in surgical visual question answering (VQA). However, existing surgical VQA data

model-releasesarxiv-cs-cv
5 May 2026
← Previous
1…198199200201202…261
Next →