AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,515 results
Local Ai

Visual Relocalization from Sparse Views in Aliased and Low-Texture Environments via Novel View Synthesis

DGX agent

arXiv:2607.22147v1 Announce Type: new Abstract: Visual localization becomes extremely challenging in planetary-like terrains characterized by low texture, perceptual aliasing, harsh illumination, and

local-aiarxiv-cs-cv
27 Jul 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Tutorials

Visual Saliency Steering Distillation for Multimodal Chain-of-Thought Reasoning

DGX agent

arXiv:2607.22013v1 Announce Type: new Abstract: Multimodal chain-of-thought (CoT) reasoning integrates visual and textual cues through step-by-step inference. In small models with limited token budget

tutorialsarxiv-cs-cv
27 Jul 2026
Agents

VTM-Nav: Harnessing Cross-Episode Experience for Object-Goal Navigation with Hierarchical Visual-Topological Memory

DGX agent

arXiv:2607.14514v2 Announce Type: replace Abstract: Training-free ObjectNav agents increasingly use vision-language models (VLMs), yet typically discard acquired scene knowledge after each request. We

agentsarxiv-cs-cv
27 Jul 2026
Research

What Happens to Accuracy When Photo Lineups Contain Non-Mated Rank-One Images From Large Galleries?

DGX agent

arXiv:2607.21792v1 Announce Type: new Abstract: One-to-many facial identification is commonly used to match a probe image from surveillance video against a gallery of driver's licenses and/or booking

researcharxiv-cs-cv
27 Jul 2026
Research

3D-GIMP: When 3D Gaussian Inpainting Meets PatchMatch

DGX agent

arXiv:2607.20789v1 Announce Type: new Abstract: Recent advances in 3D scene editing have leveraged iterative diffusion models to update input views. However, this process is computationally expensive

researcharxiv-cs-cv
24 Jul 2026
Agents

A real-time RGB-D perception pipeline for autonomous impact hammers in mining: self-filtering, rock segmentation and rock-breaking poses generation

DGX agent

arXiv:2607.20748v1 Announce Type: cross Abstract: Impact hammers, also known as rock-breakers, are essential machines in mining operations, where they perform secondary reduction. In underground minin

agentsarxiv-cs-cv
24 Jul 2026
Model Releases

Achieving Text-based Person Retrieval with Any Granularity

DGX agent

arXiv:2607.21057v1 Announce Type: new Abstract: Text-based person retrieval faces a critical but under-explored challenge: the inherent uncertainty of query granularity in real-world scenarios. This p

model-releasesarxiv-cs-cv
24 Jul 2026
Model Releases

Agentic Designer: Progressive Multi-Agent Collaboration for Structure-Aware Interior Layout Generation

DGX agent

arXiv:2607.20866v1 Announce Type: new Abstract: Generating realistic interior furniture layouts that strictly adhere to architectural constraints (e.g., walls, doors, and windows) remains a fundamenta

model-releasesarxiv-cs-cv
24 Jul 2026
Safety

ASTRA-Net: Anatomy-Specific Transfer and Representation Alignment for Drug-Induced Sleep Endoscopy Segmentation

DGX agent

arXiv:2607.21370v1 Announce Type: new Abstract: Quantitative drug-induced sleep endoscopy (DISE) requires reliable airway boundaries at specific anatomical levels. Pixel-level DISE annotations are sca

safetyarxiv-cs-cv
24 Jul 2026
Research

AUCH-Net: Action Unit-Based Consistency-Aware Hypergraph Network for Cross-Domain Few-Shot Facial Expression Recognition

DGX agent

arXiv:2607.21004v1 Announce Type: new Abstract: Recently, cross-domain few-shot facial expression recognition (CF-FER) has received considerable attention. However, the performance of existing CF-FER

researcharxiv-cs-cv
24 Jul 2026
Safety

Axolotl3D: a Unified Framework for Faithful 3D Shape Completion

DGX agent

arXiv:2607.20660v1 Announce Type: new Abstract: Recent 3D generative models produce high-quality geometry from a single image using large-scale priors and diffusion architectures. However, they assume

safetyarxiv-cs-cv
24 Jul 2026
Agents

Boosting Robustness for All-Weather Self-Supervised Depth Estimation in Autonomous Driving

DGX agent

arXiv:2607.21526v1 Announce Type: new Abstract: Self-supervised depth estimation is challenging for safe autonomous driving under various adverse weather conditions due to sensor perception degradatio

agentsarxiv-cs-cv
24 Jul 2026
Tutorials

C-PTQ: Fisher-weighted Channel-wise Sensitivity for Post-training Quantization of MLLMs

DGX agent

arXiv:2607.21076v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) require huge memory and computational costs, which limits their practical deployment. Post-training quantizatio

tutorialsarxiv-cs-cv
24 Jul 2026
Agents

Causal-AgentIR: Self-Evolving Causal Memory for Adaptive Image Restoration Agents

DGX agent

arXiv:2607.21125v1 Announce Type: new Abstract: Image restoration agents have recently emerged as a flexible paradigm for handling diverse and unpredictable degradations in real-world scenarios. Exist

agentsarxiv-cs-cv
24 Jul 2026
Research

CLUIE: Clustering-Aware Recurrent Propagation with Local Structural Compensation for Underwater Image Enhancement

DGX agent

arXiv:2607.21467v1 Announce Type: new Abstract: Underwater image enhancement remains challenging due to wavelength-dependent light absorption, scattering, and backscattering, which jointly cause color

researcharxiv-cs-cv
24 Jul 2026
Local Ai

Counterfactual Explainability Framework With CycleGAN And Counterfactual-Classifier Alignnment Score for Retinal Disease Classification

DGX agent

arXiv:2607.21068v1 Announce Type: cross Abstract: Automated detection of vision impairing retina-based ocular conditions from fundus images is important for early screening, timely referral and reduci

local-aiarxiv-cs-cv
24 Jul 2026
Model Releases

CT-Merging: Consensus Directions and Task-Level Scaling for LoRA Adapter Merging

DGX agent

arXiv:2607.20561v1 Announce Type: cross Abstract: LoRA adapters provide an efficient way to specialize a pretrained model for many downstream tasks, but deploying one adapter per task requires adapter

model-releasesarxiv-cs-cv
24 Jul 2026
Agents

DAPM: UAV Monocular Depth Estimation from Any Height, Pitch, Roll and FOV

DGX agent

arXiv:2607.21438v1 Announce Type: new Abstract: Monocular depth estimation is a fundamental prerequisite for 3D reconstruction and autonomous navigation in Unmanned Aerial Vehicles (UAVs). In practica

agentsarxiv-cs-cv
24 Jul 2026
Tutorials

DART: A Degradation-Aware Recurrent Transformer for Archival Film Restoration

DGX agent

arXiv:2607.21219v1 Announce Type: new Abstract: Archival film restoration is a challenging problem because historical footage contains compound degradations such as scratches, dust, blur, noise, flick

tutorialsarxiv-cs-cv
24 Jul 2026
Tutorials

DCVC-MV: Deep Contextual Multiview Video Compression with Efficient Inter-View Prediction

DGX agent

arXiv:2509.03922v2 Announce Type: replace Abstract: Multiview video is a key format for 3D applications such as free-viewpoint broadcasting and virtual reality, yet its large data volume poses signifi

tutorialsarxiv-cs-cv
24 Jul 2026
Local Ai

Decoupling Cross-Modality Manifold Discrepancy: Leveraging Visible Diffusion Priors for Infrared Super-Resolution

DGX agent

arXiv:2607.21174v1 Announce Type: new Abstract: Infrared image super-resolution (IISR) mitigates the limitations imposed by low spatial resolution. Existing methods have recognized that IISR should pr

local-aiarxiv-cs-cv
24 Jul 2026
Model Releases

Detecting Neural Network Failures through Spectral Analysis of Internal Activations

DGX agent

arXiv:2607.20590v1 Announce Type: cross Abstract: Neural network misclassifications exhibit characteristic spectral instability in internal activations that is invisible at the output layer. This phen

model-releasesarxiv-cs-cv
24 Jul 2026
Safety

Detectors Learn the Wrong Thing: Shortcut-Resistant Adversarial Training Against Physically Realizable Attacks

DGX agent

arXiv:2607.21243v1 Announce Type: new Abstract: AI-enabled visual perception systems are increasingly deployed in intelligent transportation infrastructure and autonomous vehicle related applications.

safetyarxiv-cs-cv
24 Jul 2026
Model Releases

DINO-VPT: Hierarchical Visual Prompt Tuning for Joint Physical-Digital Face Anti-Spoofing

DGX agent

arXiv:2607.20900v1 Announce Type: new Abstract: With the increasing diversity of spoofing attacks, there is a growing demand for unified Face Anti-Spoofing (FAS) models capable of detecting both physi

model-releasesarxiv-cs-cv
24 Jul 2026
Safety

Distribution-Alignment Bridge for Uncertainty-Aware Text-to-Video Retrieval

DGX agent

arXiv:2607.20984v1 Announce Type: new Abstract: This paper proposes the Distribution-Alignment Bridge (DAB), a framework that reconceptualizes text-to-video retrieval as a distribution alignment task

safetyarxiv-cs-cv
24 Jul 2026
Model Releases

Do Pathology Vision-Language Models Truly See Pathology?

DGX agent

arXiv:2607.21065v1 Announce Type: new Abstract: Pathology vision-language models (VLMs) have recently progressed rapidly and are commonly evaluated by answer accuracy on pathology VQA benchmarks. Howe

model-releasesarxiv-cs-cv
24 Jul 2026
Safety

DTIF: Robust Loop Closure Detection via Delaunay Triangle Topology in Complex Forests

DGX agent

arXiv:2607.21138v1 Announce Type: new Abstract: Accurate forest inventory and large-scale mapping are essential for ecosystem monitoring and sustainable forest management. Multiple low-cost edge platf

safetyarxiv-cs-cv
24 Jul 2026
Safety

EmoSpace: Immersive Affective Image Generation Guided by Fine-Grained Emotion Prototypes

DGX agent

arXiv:2602.11658v2 Announce Type: replace Abstract: Immersive affective content generation aims to create visually compelling VR imagery with controllable emotional nuance, yet existing methods typica

safetyarxiv-cs-cv
24 Jul 2026
Model Releases

Engine-Native Editable 3D World Reconstruction with Objects and Lighting

DGX agent

arXiv:2607.20889v1 Announce Type: new Abstract: Editable 3D scene creation requires object instances and lights that can be inspected, moved, and imported into standard engines, yet existing single-im

model-releasesarxiv-cs-cv
24 Jul 2026
Research

Evaluation and Prognostic Validation of Deep Regression Models for WSI-Based Gene-Expression Prediction

DGX agent

arXiv:2410.00945v2 Announce Type: replace-cross Abstract: Gene-expression profiling is widely used in research and central to many areas of precision oncology, but remains costly and not universally a

researcharxiv-cs-cv
24 Jul 2026
Model Releases

Explainable Deepfake Detection Challenge

DGX agent

arXiv:2607.21007v1 Announce Type: new Abstract: Deepfake detection is moving beyond binary classification decisions toward systems that can also explain the visual evidence supporting those decisions.

model-releasesarxiv-cs-cv
24 Jul 2026
Safety

Explainable graph attention network for stress recognition (StressGAT) via differential action units

DGX agent

arXiv:2607.20819v1 Announce Type: new Abstract: Stress is a dynamic process characterized by significant individual variability in facial expression. Traditional architectures, such as Recurrent Neura

safetyarxiv-cs-cv
24 Jul 2026
Research

FA-LAM: Focus-Aware Large Avatar Model for One-Shot 4D Animatable Gaussian Head

DGX agent

arXiv:2607.20922v1 Announce Type: new Abstract: We propose FA-LAM, a Focus-Aware Large Avatar Model for one-shot animatable Gaussian head creation, while simultaneously enabling static 3D and dynamic

researcharxiv-cs-cv
24 Jul 2026
Research

Fitting Generalized Power Diagrams to 3D Image Data: A Prerequisite for Virtual Materials Testing

DGX agent

arXiv:2507.14268v2 Announce Type: replace Abstract: This paper reviews algorithmic and modeling approaches for fitting generalized power diagrams to three-dimensional image data, a key step in virtual

researcharxiv-cs-cv
24 Jul 2026
Model Releases

Flash EQ-Linear: Accelerating Equivariant Linear Layers via Group-wise Discrete Fourier Transform

DGX agent

arXiv:2607.21271v1 Announce Type: new Abstract: Equivariant networks embed geometric symmetries as structural priors through weight sharing, achieving remarkable parameter efficiency across vision tas

model-releasesarxiv-cs-cv
24 Jul 2026
Tutorials

Focus on What Matters: Constraining Spatial-Temporal Attention via Action-Units for Noise-Resilient AQA

DGX agent

arXiv:2511.05611v2 Announce Type: replace Abstract: The core challenge in Action Quality Assessment (AQA) lies in extracting fine-grained motion features from redundant and complex video backgrounds.

tutorialsarxiv-cs-cv
24 Jul 2026
Research

FSB-Net: Frequency-Spatial Boundary Network for Brain Stroke Lesion Segmentation in Non-Contrast CT

DGX agent

arXiv:2607.20955v1 Announce Type: new Abstract: Accurate segmentation of brain stroke lesions in non-contrast computed tomography (NCCT) scans is critical for rapid clinical decision-making, yet remai

researcharxiv-cs-cv
24 Jul 2026
Model Releases

Future Rendering neq Future Surface: A Benchmark and Dataset for Dynamic Surface Reconstruction Beyond the Observed Window

DGX agent

arXiv:2607.21471v1 Announce Type: new Abstract: Dynamic-scene reconstruction is almost always evaluated inside the observed time window, yet deployment settings such as AR overlays, robot interaction,

model-releasesarxiv-cs-cv
24 Jul 2026
Research

Geo3R: Mitigating Spatial Reasoning Hallucination in Multimodal Large Language Models

DGX agent

arXiv:2607.21085v1 Announce Type: new Abstract: Despite remarkable progress in visual understanding, Multimodal Large Language Models (MLLMs) remain prone to hallucinations when reasoning about spatia

researcharxiv-cs-cv
24 Jul 2026
Research

GeoThreat: Transferable Targeted Adversarial Attacks on Large Vision-Language Models for Remote Sensing Image Interpretation

DGX agent

arXiv:2607.21036v1 Announce Type: new Abstract: Adversarial attacks against large vision-language models (LVLMs) serve as an effective means of assessing their robustness in cross-modal semantic under

researcharxiv-cs-cv
24 Jul 2026
Safety

GLAM-SLAM: Real-time Gaussian Large-scale Mapping via Flow Densification and Spatial Decomposition

DGX agent

arXiv:2607.21416v1 Announce Type: cross Abstract: Existing Gaussian-splatting-based monocular Simultaneous Localization and Mapping (SLAM) systems are either tailored to short sequences, are not real-

safetyarxiv-cs-cv
24 Jul 2026
Model Releases

GrainGS: Gradient-Decoupled Gaussian Splatting for Efficient Dynamic Novel View Synthesis

DGX agent

arXiv:2607.21448v1 Announce Type: new Abstract: Dynamic scene reconstruction with 3D Gaussian Splatting requires a balance between fine-grained motion modeling, structural stability, and compact repre

model-releasesarxiv-cs-cv
24 Jul 2026
Safety

GroupVideo: Multi-Identity Customized Text-to-Video Generation

DGX agent

arXiv:2607.21027v1 Announce Type: new Abstract: Current identity customized video generation methodologies are predominantly limited to single-identity scenarios, as the lack of explicit identity sepa

safetyarxiv-cs-cv
24 Jul 2026
Model Releases

HalluScope: Fine-grained Hallucination Diagnosis for Multimodal Large Language Models

DGX agent

arXiv:2607.21105v1 Announce Type: new Abstract: Although Multimodal Large Language Models have achieved strong performance across a wide range of vision-language tasks, they still suffer from hallucin

model-releasesarxiv-cs-cv
24 Jul 2026
Agents

HGeo-TopoMap: Boosting Topological Mapping with Hierarchical Geometric Priors

DGX agent

arXiv:2607.21281v1 Announce Type: new Abstract: Topological maps are key outputs of autonomous driving perception systems, delivering essential road information for path planning. They identify instan

agentsarxiv-cs-cv
24 Jul 2026
Model Releases

HyperImageNet: A Large-Scale High-Spatial Resolution Hyperspectral Imagery Classification Benchmark

DGX agent

arXiv:2607.21050v1 Announce Type: new Abstract: We present HyperImageNet, a large-scale benchmark for fine-grained hyperspectral land-cover understanding. The dataset contains 26,084 airborne hyperspe

model-releasesarxiv-cs-cv
24 Jul 2026
Research

Incremental Optimal Assignment for Real-Time Crowd Tracking

DGX agent

arXiv:2607.21368v1 Announce Type: new Abstract: Multi-object tracking in dense crowds requires solving a bipartite assignment problem between detections and trajectories at every video frame. The clas

researcharxiv-cs-cv
24 Jul 2026
Safety

Inference-Time Scaling of Diffusion Models via Progressive Seed Pruning

DGX agent

arXiv:2607.21591v1 Announce Type: new Abstract: Diffusion and flow-matching models dominate conditional image generation, yet inference-time scaling for these models is far less developed than for aut

safetyarxiv-cs-cv
24 Jul 2026
← Previous
1…4243444546…261
Next →