AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,618 results
Agents

Physical Adversarial Clothing Evades Visible-Thermal Detectors via Non-Overlapping RGB-T Pattern

DGX agent

arXiv:2605.04675v1 Announce Type: new Abstract: Visible-thermal (RGB-T) object detection is a crucial technology for applications such as autonomous driving, where multimodal fusion enhances performan

agentsarxiv-cs-cv
7 May 2026
Research
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Physics-Guided Regime Unmixing

DGX agent

arXiv:2605.04247v1 Announce Type: new Abstract: The Linear Mixing Model (LMM) dominates spectral unmixing for its simplicity, but fails under multiple scattering; existing nonlinear models compensate

researcharxiv-cs-cv
7 May 2026
Safety

POMA-3D: The Point Map Way to 3D Scene Understanding

DGX agent

arXiv:2511.16567v3 Announce Type: replace Abstract: In this paper, we introduce POMA-3D, the first self-supervised 3D representation model learned from point maps. Point maps encode explicit 3D coordi

safetyarxiv-cs-cv
7 May 2026
Research

PRISM: Color-Stratified Point Cloud Sampling

DGX agent

arXiv:2601.06839v2 Announce Type: replace Abstract: We present PRISM, a novel color-guided stratified sampling method for RGB-LiDAR point clouds. Our approach is motivated by the observation that uniq

researcharxiv-cs-cv
7 May 2026
Model Releases

Privacy-Preserving Empathy Detection in Video Interactions

DGX agent

arXiv:2504.10808v3 Announce Type: replace Abstract: Detecting empathy from video interactions has emerging applications, yet raw videos that could be used for training AI models are rarely available d

model-releasesarxiv-cs-cv
7 May 2026
Research

Progressive J-Invariant Self-supervised Learning for Low-Dose CT Denoising

DGX agent

arXiv:2601.14180v3 Announce Type: replace Abstract: Self-supervised learning has been increasingly investigated for low-dose computed tomography (LDCT) image denoising, as it alleviates the dependence

researcharxiv-cs-cv
7 May 2026
Model Releases

Prompt-Anchored Vision-Text Distillation for Lifelong Person Re-identification

DGX agent

arXiv:2605.05027v1 Announce Type: new Abstract: Lifelong person re-identification (LReID) aims to train a generalizable model with sequentially collected data. However, such models often suffer from s

model-releasesarxiv-cs-cv
7 May 2026
Research

QuadBox: Accelerating 3D Gaussian Splatting with Geometry-Aware Boxes

DGX agent

arXiv:2605.04844v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) has emerged as an advanced technique for real-time novel view synthesis by representing scene geometry and appearance using

researcharxiv-cs-cv
7 May 2026
Research

RealLiFe: Real-Time Light Field Reconstruction via Hierarchical Sparse Gradient Descent

DGX agent

arXiv:2307.03017v5 Announce Type: replace Abstract: With the rise of Extended Reality (XR) technology, there is a growing need for real-time light field reconstruction from sparse view inputs. Existin

researcharxiv-cs-cv
7 May 2026
Research

Reduced-order Neural Modeling with Differentiable Simulation for High-Detail Tactile Perception

DGX agent

arXiv:2605.05053v1 Announce Type: cross Abstract: Tactile perception is key to dexterous manipulation, yet simulating high-resolution elastomer deformation remains computationally prohibitive. Finite

researcharxiv-cs-cv
7 May 2026
Tutorials

Reference-based Category Discovery: Unsupervised Object Detection with Category Awareness

DGX agent

arXiv:2605.04606v1 Announce Type: new Abstract: Traditional one-shot detection methods have addressed the closed-set problem in object detection, but the high cost of data annotation remains a critica

tutorialsarxiv-cs-cv
7 May 2026
Agents

RemoteZero: Geospatial Reasoning with Zero Human Annotations

DGX agent

arXiv:2605.04451v1 Announce Type: new Abstract: Geospatial reasoning requires models to resolve complex spatial semantics and user intent into precise target locations for Earth observation. Recent pr

agentsarxiv-cs-cv
7 May 2026
Applications

RetimeGS: Continuous-Time Reconstruction of 4D Gaussian Splatting

DGX agent

arXiv:2603.13783v2 Announce Type: replace Abstract: Temporal retiming, the ability to reconstruct and render dynamic scenes at arbitrary timestamps, is crucial for applications such as slow-motion pla

applicationsarxiv-cs-cv
7 May 2026
Research

Reward-Guided Semantic Evolution for Test-time Adaptive Object Detection

DGX agent

arXiv:2605.04531v1 Announce Type: new Abstract: Open-vocabulary object detection with vision-language models (VLMs) such as Grounding DINO suffers from performance degradation under test-time distribu

researcharxiv-cs-cv
7 May 2026
Model Releases

RoDyGS: Robust Dynamic Gaussian Splatting for Casual Videos

DGX agent

arXiv:2412.03077v2 Announce Type: replace Abstract: 4D reconstruction from casually captured monocular videos is challenging due to inherent ambiguity in reconstructing dynamic 3D geometry. To address

model-releasesarxiv-cs-cv
7 May 2026
Safety

S1-MMAlign: A Large-Scale, Multi-Disciplinary Dataset for Scientific Figure-Text Understanding

DGX agent

arXiv:2601.00264v2 Announce Type: replace Abstract: Multimodal learning has revolutionized general domain tasks, yet its application in scientific discovery is hindered by the profound semantic gap be

safetyarxiv-cs-cv
7 May 2026
Safety

SafeRedir: Prompt Embedding Redirection for Robust Unlearning in Image Generation Models

DGX agent

arXiv:2601.08623v2 Announce Type: replace Abstract: Image generation models (IGMs), while capable of producing impressive and creative content, often memorize a wide range of undesirable concepts from

safetyarxiv-cs-cv
7 May 2026
Research

SAMIC: A Lightweight Semantic-Aware Mamba for Efficient Perceptual Image Compression

DGX agent

arXiv:2605.04560v1 Announce Type: new Abstract: Perceptual image compression focuses on preserving high visual quality under low-bitrate constraints. Most existing approaches to perceptual compression

researcharxiv-cs-cv
7 May 2026
Model Releases

Scalable Object Detection in the Car Interior With Vision Foundation Models

DGX agent

arXiv:2508.19651v2 Announce Type: replace Abstract: AI tasks in the car interior like identifying and localizing externally introduced objects is crucial for response quality of personal assistants. H

model-releasesarxiv-cs-cv
7 May 2026
Research

ScriptHOI: Learning Scripted State Transitions for Open-Vocabulary Human-Object Interaction Detection

DGX agent

arXiv:2605.05057v1 Announce Type: new Abstract: Open-vocabulary human-object interaction (HOI) detection requires recognizing interaction phrases that may not appear as annotated categories during tra

researcharxiv-cs-cv
7 May 2026
Research

Segmenting proto-halos with vision transformers

DGX agent

arXiv:2508.00049v2 Announce Type: cross Abstract: The formation of dark-matter halos from small cosmological perturbations generated in the early universe is a highly non-linear process typically mode

researcharxiv-cs-cv
7 May 2026
Applications

Shape2Animal: Creative Animal Generation from Natural Silhouettes

DGX agent

arXiv:2506.20616v3 Announce Type: replace Abstract: Humans possess a unique ability to perceive meaningful patterns in ambiguous stimuli, a cognitive phenomenon known as pareidolia. This paper introdu

applicationsarxiv-cs-cv
7 May 2026
Model Releases

SlotVLA: Towards Modeling of Object-Relation Representations in Robotic Manipulation

DGX agent

arXiv:2511.06754v3 Announce Type: replace-cross Abstract: Inspired by how humans reason over discrete objects and their relationships, we explore whether compact object-centric and object-relation rep

model-releasesarxiv-cs-cv
7 May 2026
Model Releases

StableI2I: Spotting Unintended Changes in Image-to-Image Transition

DGX agent

arXiv:2605.04453v1 Announce Type: new Abstract: In most real-world image-to-image (I2I) scenarios, existing evaluations primarily focus on instruction following and the perceptual quality or aesthetic

model-releasesarxiv-cs-cv
7 May 2026
Tutorials

Stream-T1: Test-Time Scaling for Streaming Video Generation

DGX agent

arXiv:2605.04461v1 Announce Type: new Abstract: While Test-Time Scaling (TTS) offers a promising direction to enhance video generation without the surging costs of training, current test-time video ge

tutorialsarxiv-cs-cv
7 May 2026
Research

Structured 3D Latents Are Surprisingly Powerful: Unleashing Generalizable Style with 2D Diffusion

DGX agent

arXiv:2605.04412v1 Announce Type: new Abstract: 3D asset generation plays a pivotal role in fields such as gaming and virtual reality, enabling the rapid synthesis of high-fidelity 3D objects from a s

researcharxiv-cs-cv
7 May 2026
Tutorials

SV-GS: Sparse View 4D Reconstruction with Skeleton-Driven Gaussian Splatting

DGX agent

arXiv:2601.00285v2 Announce Type: replace Abstract: Reconstructing a dynamic target moving over a large area is challenging. Standard approaches for dynamic object reconstruction require dense coverag

tutorialsarxiv-cs-cv
7 May 2026
Research

Syn4D: A Multiview Synthetic 4D Dataset

DGX agent

arXiv:2605.05207v1 Announce Type: new Abstract: Dense 3D reconstruction and tracking of dynamic scenes from monocular video remains an important open challenge in computer vision. Progress in this are

researcharxiv-cs-cv
7 May 2026
Research

Taming Outlier Tokens in Diffusion Transformers

DGX agent

arXiv:2605.05206v1 Announce Type: new Abstract: We study outlier tokens in Diffusion Transformers (DiTs) for image generation. Prior work has shown that Vision Transformers (ViTs) can produce a small

researcharxiv-cs-cv
7 May 2026
Safety

Temporal Structure Matters for Efficient Test-Time Adaptation in Wearable Human Activity Recognition

DGX agent

arXiv:2605.04617v1 Announce Type: new Abstract: Wearable human activity recognition (WHAR) models often suffer from performance degradation under real-world cross-user distribution shifts. Test-time a

safetyarxiv-cs-cv
7 May 2026
Applications

Topology-Constrained Quantized nnUNet for Efficient and Anatomically Accurate 3D Tooth Segmentation

DGX agent

arXiv:2605.04201v1 Announce Type: new Abstract: We propose a topology-constrained quantized nnUNet framework for efficient and anatomically accurate 3D tooth segmentation, addressing the challenges of

applicationsarxiv-cs-cv
7 May 2026
Research

Topology-Preserving Data Augmentation for Ring-Type Polygon Annotations

DGX agent

arXiv:2603.14764v3 Announce Type: replace Abstract: Geometric data augmentation is widely used in segmentation workflows, but polygon annotations are often assumed to remain valid after transformation

researcharxiv-cs-cv
7 May 2026
Safety

Towards General Preference Alignment: Diffusion Models at Nash Equilibrium

DGX agent

arXiv:2605.04494v1 Announce Type: cross Abstract: Reinforcement learning from human feedback (RLHF) has been popular for aligning text-to-image (T2I) diffusion models with human preferences. As a main

safetyarxiv-cs-cv
7 May 2026
Local Ai

Towards Generative Location Awareness for Disaster Response: A Probabilistic Cross-view Geolocalization Approach

DGX agent

arXiv:2512.20056v2 Announce Type: replace-cross Abstract: As Earth's climate changes, it is impacting disasters and extreme weather events across the planet. Record-breaking heat waves, drenching rain

local-aiarxiv-cs-cv
7 May 2026
Model Releases

UAV as Urban Construction Change Monitor: A New Benchmark and Change Captioning Model

DGX agent

arXiv:2605.04409v1 Announce Type: new Abstract: Remote Sensing Image Change Captioning (RSICC) aims to generate spatially grounded natural language descriptions of scene evolution from bi-temporal ima

model-releasesarxiv-cs-cv
7 May 2026
Safety

UAV-VL-R1: Generalizing Vision-Language Models via Supervised Fine-Tuning and Multi-Stage GRPO for UAV Visual Reasoning

DGX agent

arXiv:2508.11196v2 Announce Type: replace Abstract: Recent advances in vision-language models (VLMs) have demonstrated strong generalization in natural image tasks. However, their performance often de

safetyarxiv-cs-cv
7 May 2026
Safety

UI2Code^N: UI-to-Code Generation as Interactive Visual Optimization

DGX agent

arXiv:2511.08195v3 Announce Type: replace Abstract: UI-to-code aims to translate UI screenshots into executable front-end code. Despite progress with vision-language models (VLMs), most existing metho

safetyarxiv-cs-cv
7 May 2026
Safety

ULF-Loc: Unbiased Landmark Feature for Robust Visual Localization with 3D Gaussian Splatting

DGX agent

arXiv:2605.04730v1 Announce Type: new Abstract: Visual localization is a core technology for augmented reality and autonomous navigation. Recent methods combine the efficient rendering of 3D Gaussian

safetyarxiv-cs-cv
7 May 2026
Safety

UniMoCo: Unified Modality Completion for Robust Multi-Modal Embeddings

DGX agent

arXiv:2505.11815v2 Announce Type: replace Abstract: Current vision-language models have been explored for multi-modal embedding tasks like information retrieval. However, they face significant challen

safetyarxiv-cs-cv
7 May 2026
Research

UniPCB: A Generation-Assisted Detection Framework for PCB Defect Inspection

DGX agent

arXiv:2605.04635v1 Announce Type: new Abstract: Printed Circuit Board (PCB) defect inspection faces two compounding challenges: scarce and imbalanced defect samples that limit model training, and insu

researcharxiv-cs-cv
7 May 2026
Research

VC-FeS: Viewpoint-Conditioned Feature Selection for Vehicle Re-identification in Thermal Vision

DGX agent

arXiv:2605.04750v1 Announce Type: new Abstract: Identification of less-articulated objects using single-channel images, such as thermal images, is important in many applications, such as surveillance.

researcharxiv-cs-cv
7 May 2026
Tutorials

Velox: Learning Representations of 4D Geometry and Appearance

DGX agent

arXiv:2605.04527v1 Announce Type: new Abstract: We introduce a framework for learning latent representations of 4D objects which are descriptive, faithfully capturing object geometry and appearance; c

tutorialsarxiv-cs-cv
7 May 2026
Model Releases

Vision-EKIPL: External Knowledge-Infused Policy Learning for Visual Reasoning

DGX agent

arXiv:2506.06856v3 Announce Type: replace Abstract: Visual reasoning is crucial for understanding complex multimodal data and advancing Artificial General Intelligence. Existing methods enhance the re

model-releasesarxiv-cs-cv
7 May 2026
Model Releases

VL-UniTrack: A Unified Framework with Visual-Language Prompts for UAV-Ground Visual Tracking

DGX agent

arXiv:2605.04574v1 Announce Type: new Abstract: UAV-ground visual tracking (UGVT) aims to simultaneously track the same object from both the UAV and the ground view. However, existing two-stream metho

model-releasesarxiv-cs-cv
7 May 2026
Model Releases

VTAgent: Agentic Keyframe Anchoring for Evidence-Aware Video TextVQA

DGX agent

arXiv:2605.04870v1 Announce Type: new Abstract: Video text-based visual question answering (Video TextVQA) aims to answer questions by reasoning over visual textual content appearing in videos. Despit

model-releasesarxiv-cs-cv
7 May 2026
Model Releases

Wasserstein-Aligned Localisation for VLM-Based Distributional OOD Detection in Medical Imaging

DGX agent

arXiv:2605.05161v1 Announce Type: new Abstract: Zero-shot anomaly localisation via vision-language models (VLMs) offers a compelling approach for rare pathology detection, yet its performance is funda

model-releasesarxiv-cs-cv
7 May 2026
Local Ai

What Matters in Practical Learned Image Compression

DGX agent

arXiv:2605.05148v1 Announce Type: new Abstract: One of the major differentiators unlocked by learned codecs relative to their hard-coded traditional counterparts is their ability to be optimized direc

local-aiarxiv-cs-cv
7 May 2026
Research

3D Human Face Reconstruction with 3DMM face model from RGB image

DGX agent

arXiv:2605.03996v1 Announce Type: new Abstract: Nowadays as convolution neural networks demonstrate its powerful problem-solving ability in the area of image processing, efforts have been made to reco

researcharxiv-cs-cv
6 May 2026
← Previous
1…192193194195196…263
Next →