AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,618 results
Local Ai

Learning to Balance: Decoupled Siamese Diffusion Transformer for Reference-Based Remote Sensing Image Super-Resolution

DGX agent

arXiv:2605.17980v1 Announce Type: new Abstract: Diffusion-based methods demonstrate significant potential for remote sensing image super-resolution at large scaling factors, particularly in reference-

local-aiarxiv-cs-cv
19 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

LESSViT: Robust Hyperspectral Representation Learning under Spectral Configuration Shift

DGX agent

arXiv:2605.18541v1 Announce Type: new Abstract: Modeling hyperspectral imagery (HSI) across different sensors presents a fundamental challenge due to variations in wavelength coverage, band sampling,

model-releasesarxiv-cs-cv
19 May 2026
Tutorials

Leveraging Latent Visual Reasoning in Silence

DGX agent

arXiv:2605.18641v1 Announce Type: new Abstract: Latent visual reasoning involves visual evidence more directly in multimodal reasoning by inserting continuous latent tokens before textual generation.

tutorialsarxiv-cs-cv
19 May 2026
Research

Lightweight Physics-Aware Zero-Shot Ultrasound Plane-Wave Denoising

DGX agent

arXiv:2506.21499v2 Announce Type: replace-cross Abstract: Ultrasound Coherent Plane-Wave Compounding (CPWC) enhances image contrast by combining echoes from multiple steered transmissions. While incre

researcharxiv-cs-cv
19 May 2026
Applications

LiPS: Lightweight Panoptic Segmentation for Resource-Constrained Robotics

DGX agent

arXiv:2604.00634v2 Announce Type: replace-cross Abstract: Panoptic segmentation is a key enabler for robotic perception, as it unifies semantic understanding with object-level reasoning. However, the

applicationsarxiv-cs-cv
19 May 2026
Research

LISA: Language-guided Interference-aware Spatial-Frequency Attention for Driver Gaze Estimation

DGX agent

arXiv:2605.17287v1 Announce Type: new Abstract: Driver gaze estimation serves as a fundamental metric for evaluating driver attentiveness in modern monitoring systems. Beyond being vulnerable to sudde

researcharxiv-cs-cv
19 May 2026
Research

LiteFrame: Efficient Vision Encoders Unlock Frame Scaling in Video LLMs

DGX agent

arXiv:2605.17260v1 Announce Type: new Abstract: The fundamental challenge in scaling Video Large Language Models (Video LLMs) to long-form video lies in managing the explosion of visual-token context

researcharxiv-cs-cv
19 May 2026
Local Ai

LongDPM: Overlap-Aware 4D Reconstruction from Long Monocular Videos

DGX agent

arXiv:2605.17303v1 Announce Type: new Abstract: Recovering a dynamic 3D scene from a long monocular video is crucial for dense geometry, camera motion, and temporal correspondence to remain consistent

local-aiarxiv-cs-cv
19 May 2026
Hardware

LongLive-2.0: An NVFP4 Parallel Infrastructure for Long Video Generation

DGX agent

arXiv:2605.18739v1 Announce Type: new Abstract: We present LongLive-2.0, an NVFP4-based parallel infrastructure throughout the full training and inference workflow of long video generation, addressing

hardwarearxiv-cs-cv
19 May 2026
Research

Lost in the Folds: When Cross-Validation Is Not a Deep Ensemble for Uncertainty Estimation

DGX agent

arXiv:2605.18329v1 Announce Type: new Abstract: Ensemble disagreement is widely used as a proxy for epistemic uncertainty in medical image segmentation. In practice, many studies form ensembles via K-

researcharxiv-cs-cv
19 May 2026
Research

Lotus-2: Advancing Geometric Dense Prediction with Powerful Image Generative Model

DGX agent

arXiv:2512.01030v3 Announce Type: replace Abstract: Recovering pixel-wise geometric properties from a single image is fundamentally ill-posed due to appearance ambiguity and non-injective mappings bet

researcharxiv-cs-cv
19 May 2026
Applications

Low Latency Gaze Tracking via Latent Optical Sensing

DGX agent

arXiv:2605.17990v1 Announce Type: new Abstract: We present a real-time gaze tracking system that directly acquires task-relevant latent features using a fully passive optical encoder. Instead of formi

applicationsarxiv-cs-cv
19 May 2026
Research

LURE: Latent Space Unblocking for Multi-Concept Reawakening in Diffusion Models

DGX agent

arXiv:2601.14330v2 Announce Type: replace Abstract: Concept erasure aims to suppress sensitive content in diffusion models, but recent studies show that erased concepts can still be reawakened, reveal

researcharxiv-cs-cv
19 May 2026
Tutorials

Machine Learning Enabled Graph Analysis of Particulate Composites: Application to Solid-state Battery Cathodes

DGX agent

arXiv:2512.16085v2 Announce Type: replace-cross Abstract: Particulate composites underpin many solid-state chemical and electrochemical systems, where microstructural features such as multiphase bound

tutorialsarxiv-cs-cv
19 May 2026
Safety

Mamba-VGGT: Persistent Long-Sequence Video Geometry Grounded Transformer via External Sliding Window Mamba Memory

DGX agent

arXiv:2605.17478v1 Announce Type: new Abstract: Visual Geometry Grounded Transformers (VGGT) have set new benchmarks in high-fidelity 3D scene reconstruction. However, as the sequence length increases

safetyarxiv-cs-cv
19 May 2026
Research

Markerless Motion Capture for Biomechanical Whole-Body Kinematic Estimation in Infants

DGX agent

arXiv:2605.17120v1 Announce Type: new Abstract: arly identification of motor impairment in infancy relies on expert visual assessment of spontaneous movement, motivating the development of automated,

researcharxiv-cs-cv
19 May 2026
Research

MARQUIS: A Three-Stage Pipeline for Video Retrieval-Augmented Generation

DGX agent

arXiv:2605.17640v1 Announce Type: cross Abstract: Retrieval-augmented generation from videos requires systems to retrieve relevant audiovisual evidence from large corpora and synthesize it into cohere

researcharxiv-cs-cv
19 May 2026
Local Ai

MaskAttn-SDXL: Controllable Region-Level Text-To-Image Generation

DGX agent

arXiv:2509.15357v2 Announce Type: replace Abstract: Diffusion models have achieved strong results in text-to-image generation, but important limitations remain as prompts become more structured and mu

local-aiarxiv-cs-cv
19 May 2026
Safety

Meltdown: Circuits and Bifurcations in Point-Cloud-Conditioned 3D Diffusion Transformers

DGX agent

arXiv:2602.11130v2 Announce Type: replace-cross Abstract: Sparse point clouds are a common input modality for 3D surface reconstruction, including in safety-critical settings such as surgical navigati

safetyarxiv-cs-cv
19 May 2026
Local Ai

MementoGUI: Learning Agentic Multimodal Memory Control for Long-Horizon GUI Agents

DGX agent

arXiv:2605.18652v1 Announce Type: new Abstract: Recent GUI agents have made substantial progress in visual grounding and action prediction, yet they remain brittle in long-horizon tasks that require m

local-aiarxiv-cs-cv
19 May 2026
Research

Memory-Augmented Query Intent Understanding for Efficient Chat-based Image Retrieval

DGX agent

arXiv:2605.17365v1 Announce Type: new Abstract: Different from traditional text-to-image retrieval tasks, chat-based image retrieval allows the human-interactive system to iteratively clarify and refi

researcharxiv-cs-cv
19 May 2026
Tutorials

Meta-Learning Guided Pruning for Few-Shot Plant Pathology on Edge Devices

DGX agent

arXiv:2601.02353v3 Announce Type: replace Abstract: Farmers in remote areas need quick and reliable methods for identifying plant diseases, yet they often lack access to laboratories or high-performan

tutorialsarxiv-cs-cv
19 May 2026
Research

MetaLab: Few-Shot Game Changer for Image Recognition

DGX agent

arXiv:2507.22057v2 Announce Type: replace Abstract: Difficult few-shot image recognition has significant application prospects, yet remaining the substantial technical gaps with the conventional large

researcharxiv-cs-cv
19 May 2026
Tutorials

Mind the Gap: Learning Modality-Agnostic Representations with a Cross-Modality UNet

DGX agent

arXiv:2605.16887v1 Announce Type: new Abstract: Cross-modality recognition has many important applications in science, law enforcement and entertainment. Popular methods to bridge the modality gap inc

tutorialsarxiv-cs-cv
19 May 2026
Local Ai

Mining Forgery Traces from Reconstruction Error: A Weakly Supervised Framework for Multimodal Deepfake Temporal Localization

DGX agent

arXiv:2601.21458v2 Announce Type: replace Abstract: Modern deepfakes have evolved into localized and intermittent manipulations that require fine-grained temporal localization to mitigate severe digit

local-aiarxiv-cs-cv
19 May 2026
Model Releases

MIRAGE: Robust multi-modal architectures translate fMRI-to-image models from vision to mental imagery

DGX agent

arXiv:2605.17198v1 Announce Type: cross Abstract: To be useful for downstream applications, vision decoding models that are trained to reconstruct seen images from human brain activity must be able to

model-releasesarxiv-cs-cv
19 May 2026
Research

Mitigating 3D Prostate Biparametric MRI Data Scarcity through Domain Adaptation using Locally-Trained Latent Diffusion Models for Prostate Cancer Detection

DGX agent

arXiv:2507.06384v2 Announce Type: replace-cross Abstract: Objective: Latent diffusion models (LDMs) could mitigate data scarcity challenges affecting machine learning development for medical image int

researcharxiv-cs-cv
19 May 2026
Safety

MoASE++: Mixture of Activation Sparsity Experts with Domain-Adaptive On-policy Distillation for Continual Test Time Adaptation

DGX agent

arXiv:2605.17743v1 Announce Type: new Abstract: Continual test-time adaptation adapts a source-pretrained model to non-stationary, unlabeled target streams while retaining past competence, yet texture

safetyarxiv-cs-cv
19 May 2026
Research

MoCA3D: Monocular 3D Bounding Box Prediction in the Image Plane

DGX agent

arXiv:2603.19538v2 Announce Type: replace Abstract: Monocular 3D object understanding has largely been cast as a 2D RoI-to-3D box lifting problem. However, emerging downstream applications require ima

researcharxiv-cs-cv
19 May 2026
Local Ai

Mono-Hydra++: Real-Time Monocular Scene Graph Construction with Multi-Task Learning for 3D Indoor Mapping

DGX agent

arXiv:2605.17661v1 Announce Type: cross Abstract: Autonomous agile robots need more than metric geometry: they must understand objects, rooms, places, and spatial relations for search, inspection, exp

local-aiarxiv-cs-cv
19 May 2026
Research

Monocular Depth Perception Enhancement Based on Joint Shading/Contrast Model and Motion Parallax (JSM)

DGX agent

arXiv:2605.17252v1 Announce Type: new Abstract: Stereoscopic 3D displays adopt a binocular depth cue to provide depth perception. However, users should be equipped with expensive special devices to ap

researcharxiv-cs-cv
19 May 2026
Model Releases

Monocular Open Vocabulary Occupancy Prediction for Indoor Scenes

DGX agent

arXiv:2602.22667v2 Announce Type: replace Abstract: Open-vocabulary 3D occupancy is vital for embodied agents, which need to understand complex indoor environments where semantic categories are abunda

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

MorphSeek: Fine-grained Latent Representation-Level Policy Optimization for Deformable Image Registration

DGX agent

arXiv:2511.17392v3 Announce Type: replace Abstract: Deformable image registration (DIR) remains a fundamental yet challenging problem in medical image analysis, largely due to the prohibitively high-d

model-releasesarxiv-cs-cv
19 May 2026
Agents

Motion Cues from Image-based Point Tracking for LiDAR Scene Flow Estimation

DGX agent

arXiv:2605.16922v1 Announce Type: new Abstract: LiDAR scene flow estimation is essential for autonomous driving, as it provides 3D motion for each point. Self-supervised approaches use static-dynamic

agentsarxiv-cs-cv
19 May 2026
Safety

MSIQ: Moment-based Scale-Invariant Quality Measure for Single Image Super-Resolution

DGX agent

arXiv:2605.17588v1 Announce Type: new Abstract: Assessing the quality of single image super-resolution (SISR) results remains an open methodological problem. Common full-reference metrics (PSNR, SSIM,

safetyarxiv-cs-cv
19 May 2026
Research

Multi-hop Relational Contrastive Learning: Extending Spatial Contrastive Pre-training Beyond Pairwise Relations

DGX agent

arXiv:2605.16456v1 Announce Type: new Abstract: Understanding how objects relate to each other in space is fundamental to scene understanding, yet most contrastive pre-training approaches only model p

researcharxiv-cs-cv
19 May 2026
Safety

Multi-Order Matching Network for Alignment-Free Depth Super-Resolution

DGX agent

arXiv:2511.16361v3 Announce Type: replace Abstract: Recent guided depth super-resolution methods are premised on the assumption of strict spatial alignment between depth and RGB, achieving high-qualit

safetyarxiv-cs-cv
19 May 2026
Agents

NeRF-based Spacecraft Reconstruction from Close-Range Monocular Imagery Under Illumination Variability and Pose Uncertainty

DGX agent

arXiv:2605.18447v1 Announce Type: new Abstract: Autonomous rendezvous and proximity operations around uncooperative, unknown spacecraft are critical for active debris removal and on-orbit servicing mi

agentsarxiv-cs-cv
19 May 2026
Research

NERVE: A Neuromorphic Vision and Radar Ensemble for Multi-Sensor Fusion Research

DGX agent

arXiv:2605.16414v1 Announce Type: new Abstract: We present NERVE (Neuromorphic Vision and Radar Ensemble), a multi-sensor dataset comprising 257 minutes of synchronized recordings from five sensors: t

researcharxiv-cs-cv
19 May 2026
Applications

Network Knowledge Prior Guided Learning for Data-Efficient Surface Defect Detection

DGX agent

arXiv:2605.17780v1 Announce Type: new Abstract: Deep learning-based methods have become the de facto standard for industrial defect detection. However, their data-hungry nature and inherent 'black-box

applicationsarxiv-cs-cv
19 May 2026
Research

NeuroLiDAR: Adaptive Frame Rate Depth Sensing via Neuromorphic Event-LiDAR Fusion

DGX agent

arXiv:2605.16805v1 Announce Type: new Abstract: LiDARs are widely used for 3D depth reconstruction, but their performance is often limited by inherent hardware constraints that impose trade-offs betwe

researcharxiv-cs-cv
19 May 2026
Model Releases

Neuroscience-inspired Staged Representation Learning with Disentangled Coarse- and Fine-Grained Semantics for EEG Visual Decoding

DGX agent

arXiv:2605.16923v1 Announce Type: new Abstract: Decoding visual information from electroencephalography (EEG) signals remains a fundamental challenge in brain-computer interfaces and medical rehabilit

model-releasesarxiv-cs-cv
19 May 2026
Safety

NEWTON: Agentic Planning for Physically Grounded Video Generation

DGX agent

arXiv:2605.18396v1 Announce Type: new Abstract: Video generation models produce visually compelling results but systematically violate physical commonsense -- on VideoPhy-2, the best model achieves on

safetyarxiv-cs-cv
19 May 2026
Model Releases

Noise2Params: Unification and Parameter Determination from Noise via a Probabilistic Event Camera Model

DGX agent

arXiv:2605.16317v1 Announce Type: new Abstract: Accurate, unified models for event cameras (ECs) remain elusive, hampering calibration and algorithm design. We develop a foundational probabilistic mod

model-releasesarxiv-cs-cv
19 May 2026
Research

Non-Colliding Biometric Identities for Digital Entities: Geometry, Capacity, and Million-Scale Virtual Identity Provisioning

DGX agent

arXiv:2605.18238v1 Announce Type: new Abstract: Digital entities such as AI agents and humanoid robots increasingly operate alongside real humans, yet their identity infrastructure is based on credent

researcharxiv-cs-cv
19 May 2026
Research

Nonlinear Bipolar Compensation: Handling Outliers in Post-Training Quantization

DGX agent

arXiv:2605.16423v1 Announce Type: new Abstract: Network quantization has emerged as one of the most practical model compression techniques, which significantly reduces a model's memory and compute con

researcharxiv-cs-cv
19 May 2026
Research

Omni-Customizer: End-to-End MultiModal Customization for Joint Audio-Video Generation

DGX agent

arXiv:2605.17488v1 Announce Type: new Abstract: The landscape of joint audio and video generation has been fundamentally transformed by the advent of powerful foundation models. Despite these strides,

researcharxiv-cs-cv
19 May 2026
Model Releases

Omni-DuplexEval: Evaluating Real-time Duplex Omni-modal Interaction

DGX agent

arXiv:2605.17360v1 Announce Type: new Abstract: Real-time duplex interaction is essential for multimodal AI systems operating in real-world scenarios, where models must continuously process streaming

model-releasesarxiv-cs-cv
19 May 2026
← Previous
1…161162163164165…263
Next →