AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,414 results
Research

GaNI: Global and Near Field Illumination Aware Neural Inverse Rendering

DGX agent

arXiv:2403.15651v5 Announce Type: replace Abstract: In this paper, we present GaNI, a Global and Near-field Illumination-aware neural inverse rendering technique that can reconstruct geometry, albedo,

researcharxiv-cs-cv
14 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

GazeVaLM: A Multi-Observer Eye-Tracking Benchmark for Evaluating Clinical Realism in AI-Generated X-Rays

DGX agent

arXiv:2604.11653v1 Announce Type: new Abstract: We introduce GazeVaLM, a public eye-tracking dataset for studying clinical perception during chest radiograph authenticity assessment. The dataset compr

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Generalizable Deepfake Detection Based on Forgery-aware Layer Masking and Multi-artifact Subspace Decomposition

DGX agent

arXiv:2601.01041v3 Announce Type: replace Abstract: Deepfake detection remains highly challenging, particularly in cross-dataset scenarios and complex real-world settings. This challenge mainly arises

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

GeoArena: Evaluating Open-World Geographic Reasoning in Large Vision-Language Models

DGX agent

arXiv:2509.04334v4 Announce Type: replace Abstract: Geographic reasoning is a fundamental cognitive capability that requires models to infer plausible locations by synthesizing visual evidence with sp

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

GeoFormer: A Lightweight Swin Transformer for Joint Building Height and Footprint Estimation from Sentinel Imagery

DGX agent

arXiv:2602.09932v2 Announce Type: replace Abstract: Building height (BH) and footprint (BF) are fundamental urban morphological parameters required by climate modelling, disaster-risk assessment, and

model-releasesarxiv-cs-cv
14 Apr 2026
Research

GeomPrompt: Geometric Prompt Learning for RGB-D Semantic Segmentation Under Missing and Degraded Depth

DGX agent

arXiv:2604.11585v1 Announce Type: new Abstract: Multimodal perception systems for robotics and embodied AI often assume reliable RGB-D sensing, but in practice, depth is frequently missing, noisy, or

researcharxiv-cs-cv
14 Apr 2026
Applications

Geoparsing: Diagram Parsing for Plane and Solid Geometry with a Unified Formal Language

DGX agent

arXiv:2604.11600v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable progress but continue to struggle with geometric reasoning, primarily due to the perce

applicationsarxiv-cs-cv
14 Apr 2026
Tutorials

GIF: A Conditional Multimodal Generative Framework for IR Drop Imaging in Chip Layouts

DGX agent

arXiv:2604.09999v1 Announce Type: new Abstract: IR drop analysis is essential in physical chip design to ensure the power integrity of on-chip power delivery networks. Traditional Electronic Design Au

tutorialsarxiv-cs-cv
14 Apr 2026
Local Ai

Global monitoring of methane point sources using deep learning on hyperspectral radiance measurements from EMIT

DGX agent

arXiv:2604.10094v1 Announce Type: new Abstract: Anthropogenic methane (CH4) point sources drive near-term climate forcing, safety hazards, and system inefficiencies. Space-based imaging spectroscopy i

local-aiarxiv-cs-cv
14 Apr 2026
Research

GrOCE:Graph-Guided Online Concept Erasure for Text-to-Image Diffusion Models

DGX agent

arXiv:2511.12968v2 Announce Type: replace Abstract: Concept erasure aims to remove harmful, inappropriate, or copyrighted content from text-to-image diffusion models while preserving non-target semant

researcharxiv-cs-cv
14 Apr 2026
Research

Grounded Forcing: Bridging Time-Independent Semantics and Proximal Dynamics in Autoregressive Video Synthesis

DGX agent

arXiv:2604.06939v2 Announce Type: replace Abstract: Autoregressive video synthesis offers a promising pathway for infinite-horizon generation but is fundamentally hindered by three intertwined challen

researcharxiv-cs-cv
14 Apr 2026
Research

GS4City: Hierarchical Semantic Gaussian Splatting via City-Model Priors

DGX agent

arXiv:2604.11401v1 Announce Type: new Abstract: Recent semantic 3D Gaussian Splatting (3DGS) methods primarily rely on 2D foundation models, often yielding ambiguous boundaries and limited support for

researcharxiv-cs-cv
14 Apr 2026
Safety

GTASA: Ground Truth Annotations for Spatiotemporal Analysis, Evaluation and Training of Video Models

DGX agent

arXiv:2604.10385v1 Announce Type: new Abstract: Generating complex multi-actor scenario videos remains difficult even for state-of-the-art neural generators, while evaluating them is hard due to the l

safetyarxiv-cs-cv
14 Apr 2026
Research

H-SPAM: Hierarchical Superpixel Anything Model

DGX agent

arXiv:2604.11218v1 Announce Type: new Abstract: Superpixels offer a compact image representation by grouping pixels into coherent regions. Recent methods have reached a plateau in terms of segmentatio

researcharxiv-cs-cv
14 Apr 2026
Model Releases

HDR 3D Gaussian Splatting via Luminance-Chromaticity Decomposition

DGX agent

arXiv:2511.12895v2 Announce Type: replace Abstract: High Dynamic Range (HDR) 3D reconstruction is pivotal for professional content creation in filmmaking and virtual production. Existing methods typic

model-releasesarxiv-cs-cv
14 Apr 2026
Safety

HDR Video Generation via Latent Alignment with Logarithmic Encoding

DGX agent

arXiv:2604.11788v1 Announce Type: new Abstract: High dynamic range (HDR) imagery offers a rich and faithful representation of scene radiance, but remains challenging for generative models due to its m

safetyarxiv-cs-cv
14 Apr 2026
Research

HFI: A unified framework for training-free detection and implicit watermarking of latent diffusion model generated images

DGX agent

arXiv:2412.20704v2 Announce Type: replace Abstract: Dramatic advances in the quality of the latent diffusion models (LDMs) also led to the malicious use of AI-generated images. While current AI-genera

researcharxiv-cs-cv
14 Apr 2026
Model Releases

HG-Lane: High-Fidelity Generation of Lane Scenes under Adverse Weather and Lighting Conditions without Re-annotation

DGX agent

arXiv:2603.10128v2 Announce Type: replace Abstract: Lane detection is a crucial task in autonomous driving, as it helps ensure the safe operation of vehicles. However, existing datasets such as CULane

model-releasesarxiv-cs-cv
14 Apr 2026
Tutorials

HiddenObjects: Scalable Diffusion-Distilled Spatial Priors for Object Placement

DGX agent

arXiv:2604.10675v1 Announce Type: new Abstract: We propose a method to learn explicit, class-conditioned spatial priors for object placement in natural scenes by distilling the implicit placement know

tutorialsarxiv-cs-cv
14 Apr 2026
Research

Hide-and-Seek Attribution: Weakly Supervised Segmentation of Vertebral Metastases in CT

DGX agent

arXiv:2512.06849v2 Announce Type: replace Abstract: Accurate segmentation of vertebral metastasis in CT is clinically important yet difficult to scale, as voxel-level annotations are scarce and both l

researcharxiv-cs-cv
14 Apr 2026
Tutorials

HO-Flow: Generalizable Hand-Object Interaction Generation with Latent Flow Matching

DGX agent

arXiv:2604.10836v1 Announce Type: new Abstract: Generating realistic 3D hand-object interactions (HOI) is a fundamental challenge in computer vision and robotics, requiring both temporal coherence and

tutorialsarxiv-cs-cv
14 Apr 2026
Research

HOG-Layout: Hierarchical 3D Scene Generation, Optimization and Editing via Vision-Language Models

DGX agent

arXiv:2604.10772v1 Announce Type: new Abstract: 3D layout generation and editing play a crucial role in Embodied AI and immersive VR interaction. However, manual creation requires tedious labor, while

researcharxiv-cs-cv
14 Apr 2026
Tutorials

How to Design a Compact High-Throughput Video Camera?

DGX agent

arXiv:2604.10619v1 Announce Type: new Abstract: High throughput video acquisition is a challenging problem and has been drawing increasing attention. Existing high throughput imaging systems splice hu

tutorialsarxiv-cs-cv
14 Apr 2026
Tutorials

How to Spin an Object: First, Get the Shape Right

DGX agent

arXiv:2412.10273v3 Announce Type: replace Abstract: Image-to-3D models increasingly rely on hierarchical generation to disentangle geometry and texture. However, the design choices underlying these tw

tutorialsarxiv-cs-cv
14 Apr 2026
Research

HuiYanEarth-SAR: A Foundation Model for High-Fidelity and Low-Cost Global Remote Sensing Imagery Generation

DGX agent

arXiv:2604.11444v1 Announce Type: new Abstract: Synthetic Aperture Radar (SAR) imagery generation is essential for deepening the study of scattering mechanisms, establishing trustworthy electromagneti

researcharxiv-cs-cv
14 Apr 2026
Research

Immune2V: Image Immunization Against Dual-Stream Image-to-Video Generation

DGX agent

arXiv:2604.10837v1 Announce Type: new Abstract: Image-to-video (I2V) generation has the potential for societal harm because it enables the unauthorized animation of static images to create realistic d

researcharxiv-cs-cv
14 Apr 2026
Model Releases

Immunizing 3D Gaussian Generative Models Against Unauthorized Fine-Tuning via Attribute-Space Traps

DGX agent

arXiv:2604.09688v1 Announce Type: new Abstract: Recent large-scale generative models enable high-quality 3D synthesis. However, the public accessibility of pre-trained weights introduces a critical vu

model-releasesarxiv-cs-cv
14 Apr 2026
Local Ai

IMPLICITSTAINER: Resolution Agnostic Data-Efficient Virtual Staining Using Neural Implicit Functions

DGX agent

arXiv:2505.09831v2 Announce Type: replace-cross Abstract: Hematoxylin and eosin (H&E)-stained slides are central to cancer diagnosis and monitoring, visualizing tissue architecture and cellular morpho

local-aiarxiv-cs-cv
14 Apr 2026
Research

Improving Deep Learning-Based Target Volume Auto-Delineation for Adaptive MR-Guided Radiotherapy in Head and Neck Cancer: Impact of a Volume-Aware Dice Loss

DGX agent

arXiv:2604.10130v1 Announce Type: new Abstract: Background: Manual delineation of target volumes in head and neck cancer (HNC) remains a significant bottleneck in radiotherapy planning, characterized

researcharxiv-cs-cv
14 Apr 2026
Agents

Improving Layout Representation Learning Across Inconsistently Annotated Datasets via Agentic Harmonization

DGX agent

arXiv:2604.11042v1 Announce Type: new Abstract: Fine-tuning object detection (OD) models on combined datasets assumes annotation compatibility, yet datasets often encode conflicting spatial definition

agentsarxiv-cs-cv
14 Apr 2026
Tutorials

Improving the Reasoning of Multi-Image Grounding in MLLMs via Reinforcement Learning

DGX agent

arXiv:2507.00748v3 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) perform well in single-image visual grounding but struggle with real-world tasks that demand cross-image re

tutorialsarxiv-cs-cv
14 Apr 2026
Applications

Inferring Dynamic Physical Properties from Video Foundation Models

DGX agent

arXiv:2510.02311v2 Announce Type: replace Abstract: We study the task of predicting dynamic physical properties from videos. More specifically, we consider physical properties that require temporal in

applicationsarxiv-cs-cv
14 Apr 2026
Model Releases

INSPATIO-WORLD: A Real-Time 4D World Simulator via Spatiotemporal Autoregressive Modeling

DGX agent

arXiv:2604.07209v2 Announce Type: replace Abstract: Building world models with spatial consistency and real-time interactivity remains a fundamental challenge in computer vision. Current video generat

model-releasesarxiv-cs-cv
14 Apr 2026
Local Ai

Intelligent bear deterrence system based on computer vision: Reducing human bear conflicts in remote areas

DGX agent

arXiv:2503.23178v2 Announce Type: replace Abstract: Conflicts between humans and bears on the Tibetan Plateau present substantial threats to local communities and hinder wildlife preservation initiati

local-aiarxiv-cs-cv
14 Apr 2026
Applications

Interactive Interface For Semantic Segmentation Dataset Synthesis

DGX agent

arXiv:2506.23470v2 Announce Type: replace Abstract: The rapid advancement of AI and computer vision has significantly increased the demand for high-quality annotated datasets, particularly for semanti

applicationsarxiv-cs-cv
14 Apr 2026
Research

Intra-finger Variability of Diffusion-based Latent Fingerprint Generation

DGX agent

arXiv:2604.10040v1 Announce Type: new Abstract: The primary goal of this work is to systematically evaluate the intra-finger variability of synthetic fingerprints (particularly latent prints) generate

researcharxiv-cs-cv
14 Apr 2026
Model Releases

Investigating Bias and Fairness in Appearance-based Gaze Estimation

DGX agent

arXiv:2604.10707v1 Announce Type: new Abstract: While appearance-based gaze estimation has achieved significant improvements in accuracy and domain adaptation, the fairness of these systems across dif

model-releasesarxiv-cs-cv
14 Apr 2026
Research

Iterative Inference-time Scaling with Adaptive Frequency Steering for Image Super-Resolution

DGX agent

arXiv:2512.23532v2 Announce Type: replace Abstract: Diffusion models have become a leading paradigm for image super-resolution (SR), but existing methods struggle to guarantee both the high-frequency

researcharxiv-cs-cv
14 Apr 2026
Model Releases

ITIScore: An Image-to-Text-to-Image Rating Framework for the Image Captioning Ability of MLLMs

DGX agent

arXiv:2604.03765v2 Announce Type: replace Abstract: Recent advances in multimodal large language models (MLLMs) have greatly improved image understanding and captioning capabilities. However, existing

model-releasesarxiv-cs-cv
14 Apr 2026
Research

K-STEMIT: Knowledge-Informed Spatio-Temporal Efficient Multi-Branch Graph Neural Network for Subsurface Stratigraphy Thickness Estimation from Radar Data

DGX agent

arXiv:2604.09922v1 Announce Type: cross Abstract: Subsurface stratigraphy contains important spatio-temporal information about accumulation, deformation, and layer formation in polar ice sheets. In pa

researcharxiv-cs-cv
14 Apr 2026
Research

KiseKloset for Fashion Retrieval and Recommendation

DGX agent

arXiv:2506.23471v2 Announce Type: replace-cross Abstract: The global fashion e-commerce industry has become integral to people's daily lives, leveraging technological advancements to offer personalize

researcharxiv-cs-cv
14 Apr 2026
Model Releases

Language Prompt vs. Image Enhancement: Boosting Object Detection With CLIP in Hazy Environments

DGX agent

arXiv:2604.10637v1 Announce Type: new Abstract: Object detection in hazy environments is challenging because degraded objects are nearly invisible and their semantics are weakened by environmental noi

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

LARY: A Latent Action Representation Yielding Benchmark for Generalizable Vision-to-Action Alignment

DGX agent

arXiv:2604.11689v1 Announce Type: new Abstract: While the shortage of explicit action data limits Vision-Language-Action (VLA) models, human action videos offer a scalable yet unlabeled data source. A

model-releasesarxiv-cs-cv
14 Apr 2026
Research

LDEPrompt: Layer-importance guided Dual Expandable Prompt Pool for Pre-trained Model-based Class-Incremental Learning

DGX agent

arXiv:2604.11091v1 Announce Type: new Abstract: Prompt-based class-incremental learning methods typically construct a prompt pool consisting of multiple trainable key-prompts and perform instance-leve

researcharxiv-cs-cv
14 Apr 2026
Model Releases

LEADER: Learning Reliable Local-to-Global Correspondences for LiDAR Relocalization

DGX agent

arXiv:2604.11355v1 Announce Type: new Abstract: LiDAR relocalization has attracted increasing attention as it can deliver accurate 6-DoF pose estimation in complex 3D environments. Recent learning-bas

model-releasesarxiv-cs-cv
14 Apr 2026
Applications

Learnable Motion-Focused Tokenization for Effective and Efficient Video Unsupervised Domain Adaptation

DGX agent

arXiv:2604.09955v1 Announce Type: new Abstract: Video Unsupervised Domain Adaptation (VUDA) poses a significant challenge in action recognition, requiring the adaptation of a model from a labeled sour

applicationsarxiv-cs-cv
14 Apr 2026
Research

Learning 3D Representations for Spatial Intelligence from Unposed Multi-View Images

DGX agent

arXiv:2604.10573v1 Announce Type: new Abstract: Robust 3D representation learning forms the perceptual foundation of spatial intelligence, enabling downstream tasks in scene understanding and embodied

researcharxiv-cs-cv
14 Apr 2026
Tutorials

Learning Long-term Motion Embeddings for Efficient Kinematics Generation

DGX agent

arXiv:2604.11737v1 Announce Type: new Abstract: Understanding and predicting motion is a fundamental component of visual intelligence. Although modern video models exhibit strong comprehension of scen

tutorialsarxiv-cs-cv
14 Apr 2026
← Previous
1…245246247248249…259
Next →