AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,515 results
Safety

TRACE: High-Fidelity 3D Scene Editing via Tangible Reconstruction and Geometry-Aligned Contextual Video Masking

DGX agent

arXiv:2604.01207v2 Announce Type: replace Abstract: Existing 3D Gaussian Splatting (3DGS) editing methods primarily focus on appearance modification and often struggle to support flexible geometry edi

safetyarxiv-cs-cv
3 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Training-Free Entity-Level Few-Shot Segmentation of Remote Sensing Images with Advection Refinement

DGX agent

arXiv:2607.29278v1 Announce Type: new Abstract: Existing cross-domain few-shot segmentation approaches suffer from high training costs due to source-domain episodic training and pixel-wise dense predi

researcharxiv-cs-cv
3 Aug 2026
Agents

UltraSAM3: A Concept-Driven Foundation Model for Universal Ultrasound Image Segmentation

DGX agent

arXiv:2607.29200v1 Announce Type: new Abstract: Ultrasound imaging has become increasingly widespread in clinical practice due to its portability, low cost and real-time capability, making ultrasound

agentsarxiv-cs-cv
3 Aug 2026
Research

Uncertainty-Aware Deepfake Detection via Multi-View Structural Learning

DGX agent

arXiv:2607.28769v1 Announce Type: new Abstract: Security-critical biometric and forensic applications require accurate predictions and reliable confidence estimates, particularly under distribution sh

researcharxiv-cs-cv
3 Aug 2026
Safety

VFAD: Variational Semantic Prompting Meets Frequency-Adaptive Representation Learning for Zero-Shot Anomaly Detection

DGX agent

arXiv:2607.29370v1 Announce Type: new Abstract: Zero-shot anomaly detection (ZSAD) aims to detect and localize anomalies in unseen categories without access to target-specific training data. Although

safetyarxiv-cs-cv
3 Aug 2026
Research

Visual Distribution Anchoring for Efficient Prompt Tuning

DGX agent

arXiv:2607.28967v1 Announce Type: new Abstract: Prompt tuning adapts vision--language models with few trainable parameters, but existing approaches trade off efficiency and adaptation: static textual

researcharxiv-cs-cv
3 Aug 2026
Safety

WaMo: Wavelet-Enhanced Multi-Frequency Trajectory Analysis for Fine-Grained Text-Motion Retrieval

DGX agent

arXiv:2508.03343v2 Announce Type: replace Abstract: Text-Motion Retrieval (TMR) aims to retrieve 3D motion sequences semantically relevant to text descriptions. However, matching 3D motions with text

safetyarxiv-cs-cv
3 Aug 2026
Research

Weight-Space Mixture-of-Experts for Implicit Neural Representation Classification

DGX agent

arXiv:2607.29463v1 Announce Type: new Abstract: Implicit Neural Representations (INRs) encode signals as the weights of a coordinate-based neural network and have recently been proposed as an alternat

researcharxiv-cs-cv
3 Aug 2026
Research

4DHumanDiff: Direct Text-to-4DGS Generation for Consistent 360-Degree Dynamic Humans

DGX agent

arXiv:2607.27634v1 Announce Type: new Abstract: Generating high-quality 360-degree dynamic human assets from text prompts is challenging. Existing methods usually synthesize monocular or multi-view vi

researcharxiv-cs-cv
31 Jul 2026
Model Releases

ACE-Data-0: Human-Centric Ambient Capture as Embodied Data Engine

DGX agent

arXiv:2607.28625v1 Announce Type: new Abstract: Embodied intelligence faces a fundamental data bottleneck. Models must capture how first-person perception, whole-body motion, dexterous manipulation, o

model-releasesarxiv-cs-cv
31 Jul 2026
Research

AdaAnchor4D: Anchor-Conditioned Spatiotemporal Feature Aggregation for Monocular UAV 4D Reconstruction

DGX agent

arXiv:2607.28320v1 Announce Type: new Abstract: Monocular UAV videos provide valuable observations for dynamic reconstruction of complex urban scenes. However, such scenes exhibit pronounced spatiotem

researcharxiv-cs-cv
31 Jul 2026
Model Releases

ARD-REFSM: Enhancing Reflection Symmetry Detection with Asymmetric Denoising and Rotation Equivariance

DGX agent

arXiv:2607.27927v1 Announce Type: new Abstract: Reflection symmetry detection remains challenging due to interference from asymmetric regions and arbitrary orientations of symmetric patterns. Asymmetr

model-releasesarxiv-cs-cv
31 Jul 2026
Research

Articulated Object Reconstruction from Rest-State Observation

DGX agent

arXiv:2607.27749v1 Announce Type: new Abstract: Building interactive digital twins requires recovering both 3D geometry and the kinematic structures that govern how objects articulate. Yet existing me

researcharxiv-cs-cv
31 Jul 2026
Tutorials

AuricularWorld: Hierarchical Action-Guided World Modeling for Fine-Grained Auricular Structure Segmentation from CT Scans

DGX agent

arXiv:2607.28487v1 Announce Type: new Abstract: Fine-grained segmentation of auricular structures in CT is challenging because the ear occupies a small image region, cartilage boundaries are highly ir

tutorialsarxiv-cs-cv
31 Jul 2026
Tutorials

Backbone-Agnostic Stochastic Perturbation Learning for End-to-End Real-World Image Dehazing

DGX agent

arXiv:2607.11623v3 Announce Type: replace Abstract: Real-world paired image dehazing remains challenging because haze degradation is spatially non-uniform, illumination-dependent, and physically ambig

tutorialsarxiv-cs-cv
31 Jul 2026
Model Releases

BCNet: Bronchus Classification via Structure Guided Representation Learning

DGX agent

arXiv:2205.06947v3 Announce Type: replace-cross Abstract: CT-based bronchial tree analysis is essential for diagnosing lung and airway diseases, yet automatic bronchus classification remains challengi

model-releasesarxiv-cs-cv
31 Jul 2026
Agents

Beacon: Knowing When and How to Perform Agentic Visual Reasoning

DGX agent

arXiv:2607.28595v1 Announce Type: new Abstract: The fundamental goal of agentic visual reasoning is to improve the success rate of multimodal large language models (MLLMs) on complex tasks, rather tha

agentsarxiv-cs-cv
31 Jul 2026
Model Releases

Benchmarking Foundation and Large Language Models for Few-Shot Medical Image Segmentation

DGX agent

arXiv:2607.27856v1 Announce Type: new Abstract: Few-shot medical image segmentation (FS-MIS) aims to segment novel regions of interest (ROIs) from a few annotated support examples. Despite rapid progr

model-releasesarxiv-cs-cv
31 Jul 2026
Research

Beyond Classification: Pathology Foundation Models as Detection Encoders for Mitotic Figures

DGX agent

arXiv:2607.28007v1 Announce Type: new Abstract: Pathology foundation models (FMs) are models trained on vast amounts of typically unlabeled data and have been shown to yield regularized latent spaces

researcharxiv-cs-cv
31 Jul 2026
Model Releases

Beyond Frame Selection: Generative Latent Evidence Aggregation for Long-Video Understanding

DGX agent

arXiv:2607.28516v1 Announce Type: new Abstract: Long-video understanding commonly compresses videos into a small set of frames or visual tokens for answer generation. Existing compact pipelines focus

model-releasesarxiv-cs-cv
31 Jul 2026
Research

Beyond Visual Ambiguity: Guiding Robust Monocular Depth Estimation in Challenging Scenarios via Detailed Long Captions

DGX agent

arXiv:2607.28285v1 Announce Type: new Abstract: Monocular depth estimation (MDE) faces challenges with non-Lambertian surfaces and adverse weather conditions due to the visual ambiguities inherent in

researcharxiv-cs-cv
31 Jul 2026
Applications

BladeYOLO: Wind Turbine Blade Defect Detection with Limited Annotations and Weak-Saliency Awareness

DGX agent

arXiv:2607.28065v1 Announce Type: new Abstract: Wind turbine blade defect detection remains highly challenging in real-world inspection scenarios due to limited on-site data and the subtle visual char

applicationsarxiv-cs-cv
31 Jul 2026
Model Releases

BlindPSNR: A No-Reference Fidelity Predictor for Low-Light Image Enhancement

DGX agent

arXiv:2607.27628v1 Announce Type: new Abstract: Low-light image enhancement (LLIE) methods involve tunable parameters that are typically fixed, often leading to performance degradation when applied ac

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

Bunraku: Turning a Single Illustration into an Editable Live2D Character

DGX agent

arXiv:2607.27348v1 Announce Type: new Abstract: Live2D is the dominant 2D character-animation format for anime characters and virtual avatars, representing each character as a stack of RGBA layers dri

model-releasesarxiv-cs-cv
31 Jul 2026
Research

Calibrate Before Reason: Robust Visual Token Reduction against Semantic Drift in VLMs

DGX agent

arXiv:2607.27700v1 Announce Type: new Abstract: Large Vision-Language Models (VLMs) suffer from prohibitive inference overhead due to long sequences of visual tokens. However, existing visual token re

researcharxiv-cs-cv
31 Jul 2026
Local Ai

Can Vision-Language Models Reason about AI Edits in Images?

DGX agent

arXiv:2607.28464v1 Announce Type: new Abstract: Detection and localization of AI-tampered images are critical for trustworthy AI, yet modern generative models have made such manipulations increasingly

local-aiarxiv-cs-cv
31 Jul 2026
Local Ai

Capturing Token Tendencies for Training-Free Token Pruning in Multimodal Large Language Models

DGX agent

arXiv:2607.28341v1 Announce Type: new Abstract: While visual token pruning is essential for efficient Multimodal Large Language Models (MLLMs), existing training-free methods suffer from a critical li

local-aiarxiv-cs-cv
31 Jul 2026
Model Releases

Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers

DGX agent

arXiv:2607.28611v1 Announce Type: new Abstract: Visual generation increasingly requires high-resolution images, long videos, and multimodal context, making the quadratic cost of full attention prohibi

model-releasesarxiv-cs-cv
31 Jul 2026
Research

Collaborative Feature Aggregation for Face Super-Resolution and Robust Re-Identification

DGX agent

arXiv:2607.28130v1 Announce Type: new Abstract: We propose a novel collaborative approach for face super-resolution (SR) and robust person re-identification from sequential or multi-view facial images

researcharxiv-cs-cv
31 Jul 2026
Research

Continual Learning with Vision-Language Models via Semantic-Geometry Preservation

DGX agent

arXiv:2603.12055v3 Announce Type: replace Abstract: Continual learning of pretrained vision-language models (VLMs) is prone to catastrophic forgetting, yet current approaches adapt to new tasks withou

researcharxiv-cs-cv
31 Jul 2026
Research

Convolutional Neural Shading for High-Quality 3D Reconstruction from Multi-View Images

DGX agent

arXiv:2607.28132v1 Announce Type: new Abstract: We propose a convolutional neural shading (CNS), a novel pipeline to reconstruct high-quality 3D shapes from multi-view images. Several recent studies h

researcharxiv-cs-cv
31 Jul 2026
Model Releases

CoRE-UIR: Prior-guided common and residual experts for efficient all-in-one remote sensing image restoration

DGX agent

arXiv:2607.27898v1 Announce Type: new Abstract: Remote sensing images acquired by unmanned aerial vehicles (UAVs) and satellites are often degraded by adverse weather, illumination variation, and imag

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

Cross-Embodiment Transfer via Behavior-Aligned Representations

DGX agent

arXiv:2607.27549v1 Announce Type: cross Abstract: Recent progress in large-scale imitation learning for robot manipulation has been driven by leveraging datasets across a wide range of robot embodimen

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

CXR-Retrieve: Compositional Text-to-Image Retrieval in Chest Radiography

DGX agent

arXiv:2607.27779v1 Announce Type: new Abstract: Large chest radiography archives are difficult to search because most studies are paired only with free-text reports rather than structured clinical ann

model-releasesarxiv-cs-cv
31 Jul 2026
Safety

DAS-PMVC: A Framework for Partial Multi-View Clustering via Dual Alignment and Structure Enhancement

DGX agent

arXiv:2607.27761v1 Announce Type: cross Abstract: In recent years, multi-view clustering has attracted widespread research interest. However, due to limitations in data collection devices, data across

safetyarxiv-cs-cv
31 Jul 2026
Safety

DECODE: Tackling Representation and Decision Degradation in Continual AI-Generated Image Detection

DGX agent

arXiv:2607.27882v1 Announce Type: new Abstract: As generative models continue to evolve, AI-generated image detectors must incrementally adapt to emerging generative domains while preserving knowledge

safetyarxiv-cs-cv
31 Jul 2026
Research

Deep learning-based hierarchical insect classification using camera trap imagery

DGX agent

arXiv:2607.28005v1 Announce Type: new Abstract: Declining insect populations make reliable biodiversity monitoring increasingly urgent, yet monitoring of insect biodiversity is hampered by a lack of s

researcharxiv-cs-cv
31 Jul 2026
Local Ai

DinoLizer: Separating VAE and Diffusion Artifacts in Generative Inpainting Localization

DGX agent

arXiv:2511.20722v2 Announce Type: replace Abstract: We introduce DinoLizer, a DINOv2-based localizer of manipulated areas in generative inpainting. The model is trained to focus on semantically altere

local-aiarxiv-cs-cv
31 Jul 2026
Applications

Drawing-Recode: Annotation Grounding for Parametric CAD Code Generation from Raster 2D CAD Drawings

DGX agent

arXiv:2607.27558v1 Announce Type: new Abstract: Recovering Parametric CAD sequences from raster-format 2D Computer-Aided Design (CAD) drawings accumulated prior to digital transformation is important

applicationsarxiv-cs-cv
31 Jul 2026
Model Releases

DS@GT ARC at ImageCLEFmedical 2026: Architectural Diversity for Concept Detection and Foundation-Model Scaling for Caption Prediction in Medical Image Analysis

DGX agent

arXiv:2607.27763v1 Announce Type: new Abstract: We describe the DS@GT submissions to the ImageCLEFmedical Caption 2026 challenge, which continues a long-running benchmark on the ROCOv2 dataset with tw

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

EEG-EditBench: Probing Visual Information in EEG-Image Retrieval Models with Controlled Image Edits

DGX agent

arXiv:2607.27857v1 Announce Type: new Abstract: Recent EEG-to-image retrieval models have achieved strong performance in identifying viewed images from semantically diverse candidates. Yet such succes

model-releasesarxiv-cs-cv
31 Jul 2026
Safety

EgoGenesis: Egocentric World-Action Modeling with Online Anchored Projective Memory and Action-3D RoPE

DGX agent

arXiv:2607.28243v1 Announce Type: new Abstract: Egocentric video offers rich manipulation experience for embodied AI, yet collecting diverse egocentric data across scenes, objects, motions, and embodi

safetyarxiv-cs-cv
31 Jul 2026
Model Releases

EgoGVAE: Ego-body Mesh Reconstruction via Guided Variational Autoencoder

DGX agent

arXiv:2607.27755v1 Announce Type: new Abstract: We address the problem of recovering the full-body mesh from only the head pose. This task has become essential for various applications based on head-m

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

EHGCN: Hierarchical Euclidean-Hyperbolic Fusion via Motion-Aware GCN for Hybrid Event Stream Perception

DGX agent

arXiv:2504.16616v4 Announce Type: replace Abstract: Event cameras, characterized by microsecond temporal resolution and very High Dynamic Range (HDR), emit high-speed event streams for perception task

model-releasesarxiv-cs-cv
31 Jul 2026
Safety

EmoFeedback^2: Reinforcement of Continuous Emotional Image Generation via LVLM-based Reward and Textual Feedback

DGX agent

arXiv:2511.19982v3 Announce Type: replace Abstract: Continuous emotional image content generation (C-EICG) is emerging rapidly due to its ability to produce images aligned with both user descriptions

safetyarxiv-cs-cv
31 Jul 2026
Research

ENCORE: Event-Assisted Complementary Motion Refinement for Learned Video Compression

DGX agent

arXiv:2607.28020v1 Announce Type: new Abstract: Learned video compression relies on accurate temporal modeling to remove redundancy between adjacent frames. However, most existing codecs infer motion

researcharxiv-cs-cv
31 Jul 2026
Research

Endo-NeRF++: Uncertainty-Aware Neural Rendering with Multi-Resolution Hash Encoding for Dynamic Surgical Scene Reconstruction

DGX agent

arXiv:2607.27825v1 Announce Type: cross Abstract: Reconstructing dynamic surgical scenes is crucial for robot-assisted minimally invasive surgery; however, it continues to be difficult because of tiss

researcharxiv-cs-cv
31 Jul 2026
Research

Energy-Driven Adaptive Visual Token Pruning for Efficient Vision-Language Models

DGX agent

arXiv:2603.05950v2 Announce Type: replace Abstract: Visual token reduction is critical for accelerating Vision-Language Models (VLMs), since visual inputs are represented as token sequences that intro

researcharxiv-cs-cv
31 Jul 2026
← Previous
1…2829303132…261
Next →