AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,515 results
Research

GP-4DGS: Probabilistic 4D Gaussian Splatting from Monocular Video via Variational Gaussian Processes

DGX agent

arXiv:2604.02915v2 Announce Type: replace Abstract: We present GP-4DGS, a novel framework that integrates Gaussian Processes (GPs) into 4D Gaussian Splatting (4DGS) for principled probabilistic modeli

researcharxiv-cs-cv
9 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Hardware-aware Graph Neural Networks prunning for embedded event-based vision

DGX agent

arXiv:2607.06739v1 Announce Type: new Abstract: Event-based cameras are gaining popularity as the sensor of choice for mobile robotics, due to their high performance in dynamic environments. However,

researcharxiv-cs-cv
9 Jul 2026
Safety

HART: High-Resolution Annotation-Free Reasoning Technique through a Closed-loop Framework

DGX agent

arXiv:2602.23615v3 Announce Type: replace Abstract: Current Large Multimodal Models (LMMs) struggle with high-resolution visual inputs during the reasoning process, as the number of image tokens incre

safetyarxiv-cs-cv
9 Jul 2026
Research

HPR-SAM: Hierarchical Probabilistic Representation Learning for Prompt-free SAM-based Medical Image Segmentation

DGX agent

arXiv:2607.06972v1 Announce Type: new Abstract: Prompt-free adaptation of the Segment Anything Model (SAM) has emerged as a promising paradigm for automatic medical image segmentation. Existing method

researcharxiv-cs-cv
9 Jul 2026
Model Releases

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation

DGX agent

arXiv:2506.08797v2 Announce Type: replace Abstract: To address key limitations in human-object interaction (HOI) video generation -- specifically the reliance on curated motion data, limited generaliz

model-releasesarxiv-cs-cv
9 Jul 2026
Hardware

Infinite Worlds with Versatile Interactions

DGX agent

arXiv:2607.07534v1 Announce Type: new Abstract: We present LingBot-World 2.0 (also known as LingBot-World-Infinity), an advanced iteration of LingBot-World featuring four distinct upgrades. (1) Our mo

hardwarearxiv-cs-cv
9 Jul 2026
Model Releases

InfraQR: Edge-Placed QR-Inspired Structured Patch Attacks on Infrared Vision-Language Models

DGX agent

arXiv:2607.07288v1 Announce Type: new Abstract: Infrared vision-language models are increasingly used for perception under low-light and adverse visual conditions, yet their robustness to localized st

model-releasesarxiv-cs-cv
9 Jul 2026
Hardware

Latency-Constrained DNN Architecture Learning for Edge Systems using Zerorized Batch Normalization

DGX agent

arXiv:2607.06922v1 Announce Type: cross Abstract: Deep learning applications have been widely adopted on edge devices, to mitigate the privacy and latency issues of accessing cloud servers. Deciding t

hardwarearxiv-cs-cv
9 Jul 2026
Research

Learning to Unify Deformable Shape and Texture Representations for Cardiac Video Classification

DGX agent

arXiv:2607.07518v1 Announce Type: new Abstract: Deformable shape representations have proven to be robust complements to texture features in cardiac image classification, offering geometric priors tha

researcharxiv-cs-cv
9 Jul 2026
Research

LEMUR 2: Unlocking Neural Network Diversity for AI

DGX agent

arXiv:2607.06839v1 Announce Type: cross Abstract: Existing NAS benchmarks (e.g., NAS-Bench, NATS-Bench) cover only narrow, task-specific regions of the architectural design space and lack cross-domain

researcharxiv-cs-cv
9 Jul 2026
Model Releases

MedPMC: A Systematic Framework for Scaling High-Fidelity Medical Multimodal Data for Foundation Models

DGX agent

arXiv:2607.07673v1 Announce Type: new Abstract: Medicine is inherently multimodal, requiring clinicians to synthesize information across diverse data streams. Yet the development of multimodal foundat

model-releasesarxiv-cs-cv
9 Jul 2026
Local Ai

MegaFlow: Zero-Shot Large Displacement Optical Flow

DGX agent

arXiv:2603.25739v2 Announce Type: replace Abstract: Accurate estimation of large displacement optical flow remains a critical challenge. Existing methods typically rely on iterative local search or/an

local-aiarxiv-cs-cv
9 Jul 2026
Agents

MMAgent-R^2: Learning to Rerank and Reject for Agentic mRAG

DGX agent

arXiv:2607.07383v1 Announce Type: new Abstract: Knowledge-based Visual Question Answering (KB-VQA) requires models to retrieve visual entities matching the query image from large-scale encyclopedic kn

agentsarxiv-cs-cv
9 Jul 2026
Research

MMDiff: Extending Diffusion Transformers for Multi-Modal Generation

DGX agent

arXiv:2606.16673v2 Announce Type: replace Abstract: Diffusion transformers have demonstrated remarkable generative capabilities, yet the rich perceptual representations computed across their denoising

researcharxiv-cs-cv
9 Jul 2026
Local Ai

Naming the Concepts Classifiers Rely On: Language-Anchored Decomposition for Faithful Explanation

DGX agent

arXiv:2607.07264v1 Announce Type: new Abstract: Deep neural networks are widely deployed in high-stakes visual applications where interpretability is critical, yet existing explanations face a trade-o

local-aiarxiv-cs-cv
9 Jul 2026
Research

NoDrift3R: Raymap-Guided Coupling for Drift-Robust Unposed Feed-Forward 3D Reconstruction

DGX agent

arXiv:2607.07168v1 Announce Type: new Abstract: Pose-Free Feed-forward 3D Gaussian Splatting (3DGS) has recently emerged as a powerful paradigm for fast scene reconstruction. However, its performance

researcharxiv-cs-cv
9 Jul 2026
Research

Physically Grounded Monocular Depth via Nanophotonic Wavefront Encoding

DGX agent

arXiv:2503.15770v3 Announce Type: replace-cross Abstract: Depth foundation models (DFMs) offer strong learned priors for 3D perception from single RGB images but lack physical depth cues, leading to a

researcharxiv-cs-cv
9 Jul 2026
Research

Pixel-Precise Explainable Stress Indexing: A Semantic Segmentation Framework for Disease Severity Quantification in Field Crops

DGX agent

arXiv:2607.06585v1 Announce Type: new Abstract: Plant diseases, resulting from both biotic and abiotic stresses, cause an estimated 20-40% loss in global agricultural yield annually, resulting in econ

researcharxiv-cs-cv
9 Jul 2026
Research

Prior-matched evaluation of operational Earth-observation classifiers: a three-number reporting method demonstrated on Sentinel-1 internal-wave detection

DGX agent

arXiv:2607.07146v1 Announce Type: cross Abstract: The Internal Waves Service screens the Sentinel-1 Wave-mode archive for internal solitary waves, routing detections to experts whose adjudication time

researcharxiv-cs-cv
9 Jul 2026
Tutorials

Prototype-Anchored Generalized Manifold Regression for Unknown-Domain Object Detection

DGX agent

arXiv:2607.07192v1 Announce Type: new Abstract: In this paper, we study Single-Domain Generalized Object Detection (Single-DGOD), which aims to transfer a detector trained on a single source domain to

tutorialsarxiv-cs-cv
9 Jul 2026
Research

PUF: Plug-and-Play Uncertainty-Aware Fusion for Online 3D Scene Graph Generation

DGX agent

arXiv:2607.07170v1 Announce Type: new Abstract: Online 3D scene graph generation builds a persistent, structured representation of a scene by incrementally fusing 2D observations into a global 3D grap

researcharxiv-cs-cv
9 Jul 2026
Research

Rail Track Extraction from Rasterized Classified Point Clouds Using a Full-Resolution, Fully Convolutional Recurrent Neural Network

DGX agent

arXiv:2607.06829v1 Announce Type: new Abstract: Rail track extraction is essential for effective railway asset management and maintenance, especially in automated inspection and mapping workflows. Thi

researcharxiv-cs-cv
9 Jul 2026
Model Releases

Retrieving and Refining Winning Noise Tickets for Diffusion-Based Motion Generation

DGX agent

arXiv:2607.06843v1 Announce Type: new Abstract: Diffusion-based text-to-motion models synthesize realistic human motions but often exhibit semantic drift from the input text. Motion is inherently temp

model-releasesarxiv-cs-cv
9 Jul 2026
Model Releases

ROAD-Waymo: A Large-Scale Action Awareness Dataset for Autonomous Driving

DGX agent

arXiv:2411.01683v3 Announce Type: replace Abstract: Autonomous Vehicle (AV) perception systems require more than simply seeing, via e.g., object detection or scene segmentation. They need a holistic u

model-releasesarxiv-cs-cv
9 Jul 2026
Safety

Scaling Mixture-of-Experts Video Pretraining for Embodied Intelligence

DGX agent

arXiv:2607.07675v1 Announce Type: new Abstract: Despite the recent promise in robot control, video generative models suffer from a domain mismatch due to their primary focus on content creation. For e

safetyarxiv-cs-cv
9 Jul 2026
Research

Scaling Quantum Machine Learning without Tricks: Full-Resolution and Diverse Image Generation

DGX agent

arXiv:2603.00233v2 Announce Type: replace-cross Abstract: Quantum generative modeling is a rapidly evolving discipline at the intersection of quantum computing and machine learning. Contemporary quant

researcharxiv-cs-cv
9 Jul 2026
Research

Seeing What Matters: Lesion-Aware High-Resolution Patch Discovery and Fusion for Chest X-ray Report Generation

DGX agent

arXiv:2607.06909v1 Announce Type: new Abstract: Despite rapid advances in chest X-ray (CXR) foundation models, most radiology report generation (RRG) systems still rely on heavily downsampled inputs (

researcharxiv-cs-cv
9 Jul 2026
Local Ai

Segmenting Low-Contrast XCTs of Concrete: An Unsupervised Approach

DGX agent

arXiv:2603.00127v2 Announce Type: replace Abstract: X-Ray Computed Tomography (XCT) is a compelling tool in experimental mechanics, capable of non-destructively extracting information pertaining to th

local-aiarxiv-cs-cv
9 Jul 2026
Safety

SHTA: Semantic Hard Token Correction and Center Alignment for Semi-Supervised Medical Image Segmentation

DGX agent

arXiv:2607.07019v1 Announce Type: new Abstract: Recent advances in semi-supervised medical image segmentation have achieved remarkable performance through prediction consistency, pseudo-label supervis

safetyarxiv-cs-cv
9 Jul 2026
Local Ai

Smart Scissor: Coupling Spatial Redundancy Reduction and CNN Compression for Embedded Hardware

DGX agent

arXiv:2607.06915v1 Announce Type: new Abstract: Scaling down the resolution of input images can greatly reduce the computational overhead of convolutional neural networks (CNNs), which is promising fo

local-aiarxiv-cs-cv
9 Jul 2026
Applications

SoccerNet 2026 Challenges Results

DGX agent

arXiv:2607.07320v1 Announce Type: new Abstract: The SoccerNet 2026 Challenges constitute the sixth annual edition of the SoccerNet open benchmarking effort, dedicated to advancing computer vision rese

applicationsarxiv-cs-cv
9 Jul 2026
Research

SonoRank: Towards Calibration-Free Real-Time Finger Flexion Detection from Forearm Ultrasound Sequences

DGX agent

arXiv:2607.07542v1 Announce Type: cross Abstract: Powered prosthetic hands are frequently abandoned, largely due to the limited functionality of current devices that rely on surface electromyography (

researcharxiv-cs-cv
9 Jul 2026
Local Ai

Sparse Attention for Dense Open-Vocabulary Prediction in CLIP

DGX agent

arXiv:2607.07135v1 Announce Type: new Abstract: Contrastive Language-Image Pre-training (CLIP) relies on softmax-based self-attention, a strictly positive distribution that assigns probability mass to

local-aiarxiv-cs-cv
9 Jul 2026
Research

SpiS-GAN: Spiral-Modulated Handwriting Synthesis with Star Operation

DGX agent

arXiv:2607.06949v1 Announce Type: new Abstract: Training robust handwriting recognition (HTR) systems requires massive amounts of annotated data, which is often difficult to acquire. While synthetic h

researcharxiv-cs-cv
9 Jul 2026
Model Releases

Stage-Aware Adaptation and Distribution Calibration for Subject-Driven Personalized Text-to-Image Generation

DGX agent

arXiv:2607.07173v1 Announce Type: new Abstract: Subject-driven personalized text-to-image generation requires a pretrained diffusion model to acquire a specific subject from a few reference images whi

model-releasesarxiv-cs-cv
9 Jul 2026
Model Releases

T^{3}S: Think in Thermal Time for Generalizable Crop Mapping from Satellite Image Time Series

DGX agent

arXiv:2506.12885v4 Announce Type: replace Abstract: Crop type classification from optical satellite time series remains limited in its ability to generalize across growing seasons, particularly when c

model-releasesarxiv-cs-cv
9 Jul 2026
Safety

TACoS: Weakly Supervised Learning of Two-Dimensional Materials from Scribble Annotations to Precise Segmentation

DGX agent

arXiv:2607.07169v1 Announce Type: new Abstract: The precise pixel-level localization of 2D material flakes is crucial for high-throughput screening. However, traditional fully supervised methods rely

safetyarxiv-cs-cv
9 Jul 2026
Safety

TIR-Agent: Training an Explorative and Efficient Agent for Image Restoration

DGX agent

arXiv:2603.27742v2 Announce Type: replace Abstract: Vision-language agents that orchestrate specialized tools for image restoration (IR) have emerged as a promising method, yet most existing framework

safetyarxiv-cs-cv
9 Jul 2026
Hardware

Towards Accurate and Fast Clinical Body Composition: A Resource-Efficient Hierarchical Segmentation Framework for Multi-Source CT

DGX agent

arXiv:2607.07177v1 Announce Type: cross Abstract: Background: Automated 3D segmentation of muscles and adipose tissue from CT is vital for body composition analysis, but multi-source data heterogeneit

hardwarearxiv-cs-cv
9 Jul 2026
Model Releases

TRACE-Seg3D: Counterfactual Context Auditing For Robust 3D Glioma Segmentation Under Institutional Shift

DGX agent

arXiv:2607.07038v1 Announce Type: new Abstract: Medical image segmentation models can achieve strong benchmark performance while remaining sensitive to scanner, protocol, and institutional variation.

model-releasesarxiv-cs-cv
9 Jul 2026
Research

Trexplorer Super: Topologically Correct Centerline Tree Tracking of Tubular Objects in CT Volumes

DGX agent

arXiv:2507.10881v2 Announce Type: replace Abstract: Tubular tree structures, such as blood vessels and airways, are essential in human anatomy and accurately tracking them while preserving their topol

researcharxiv-cs-cv
9 Jul 2026
Model Releases

Two-Stage Multi-Modal Fusion with Adaptive Alignment for Action Quality Assessment

DGX agent

arXiv:2607.07438v1 Announce Type: new Abstract: Action Quality Assessment (AQA) aims to evaluate how well a person performs a movement, which is essential in applications such as sports scoring, skill

model-releasesarxiv-cs-cv
9 Jul 2026
Model Releases

Unraveling Machine Behavior by Multi-Level Bias Analysis and Detection: Methodology and Application to Computer Vision

DGX agent

arXiv:2607.07236v1 Announce Type: new Abstract: This study investigates the presence and propagation of bias within Neural Networks through a comprehensive multi-level analysis spanning the learned la

model-releasesarxiv-cs-cv
9 Jul 2026
Local Ai

URS-Stereo: Uncertainty-Guided Residual Search for Real-Time Stereo Matching

DGX agent

arXiv:2607.06779v1 Announce Type: new Abstract: Real-time stereo matching is crucial for robotics, autonomous systems, and embedded vision applications, where both computational efficiency and dispari

local-aiarxiv-cs-cv
9 Jul 2026
Model Releases

VCDP: Variation-Conditioned Distributional Proxy Learning for Semi-Supervised Medical Image Segmentation

DGX agent

arXiv:2607.07416v1 Announce Type: new Abstract: Semi-supervised 3D medical image segmentation reduces the need for dense voxel-level annotations by exploiting unlabeled volumes. Although existing meth

model-releasesarxiv-cs-cv
9 Jul 2026
Safety

VFM-Loc: Training-Free Cross-View Geo-Localization via Aligning Discriminative Visual Hierarchies

DGX agent

arXiv:2603.13855v2 Announce Type: replace Abstract: Cross-View Geo-Localization (CVGL) in remote sensing aims to locate a drone-view query by matching it to geo-tagged satellite images. Although super

safetyarxiv-cs-cv
9 Jul 2026
Research

Video-Based Detection of squint and cataract for accessibility-aware adaptive web interface rendering

DGX agent

arXiv:2607.07099v1 Announce Type: new Abstract: Squint and cataract are major ocular disorders that majorly affect visual perception and interaction capability. This paper proposes a real-time video-b

researcharxiv-cs-cv
9 Jul 2026
Model Releases

Video2Reaction: Mapping Video to Audience Reaction Distribution in the Wild

DGX agent

arXiv:2607.06875v1 Announce Type: new Abstract: Understanding and forecasting audience reactions to video content are crucial for improving content creation, recommendation systems, and media analysis

model-releasesarxiv-cs-cv
9 Jul 2026
← Previous
1…5657585960…261
Next →