AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,618 results
Model Releases

OmniPro: A Comprehensive Benchmark for Omni-Proactive Streaming Video Understanding

DGX agent

arXiv:2605.18577v1 Announce Type: new Abstract: Omni-proactive streaming video understanding, i.e., autonomously deciding when to speak and what to say from continuous audio-visual streams, is an emer

model-releasesarxiv-cs-cv
19 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Research

OmniSelect: Dynamic Modality-Aware Token Compression for Efficient Omni-modal Large Language Models

DGX agent

arXiv:2605.18041v1 Announce Type: new Abstract: Omnimodal large language models (OmniLLMs) have recently gained increasing attention for unified audio-video understanding. However, processing long mul

researcharxiv-cs-cv
19 May 2026
Research

On Applicability of Synthetic Datasets for Facial Expression Recognition

DGX agent

arXiv:2605.17483v1 Announce Type: new Abstract: Facial Expression Recognition faces two core challenges. The first is class imbalance in public datasets, which skews the learning process and weakens g

researcharxiv-cs-cv
19 May 2026
Applications

Open Set Face Forgery Detection via Dual-Level Evidence Collection

DGX agent

arXiv:2512.04331v2 Announce Type: replace Abstract: The surge in face forgeries has increasingly undermined confidence in the authenticity of online content. As generation algorithms rapidly evolve, n

applicationsarxiv-cs-cv
19 May 2026
Research

OpenGaFF: Open-Vocabulary Gaussian Feature Field with Codebook Attention

DGX agent

arXiv:2605.06088v2 Announce Type: replace Abstract: Understanding open-vocabulary 3D scenes with Gaussian-based representations remains challenging due to fragmented and spatially inconsistent semanti

researcharxiv-cs-cv
19 May 2026
Research

OPTNet: Ordering Point Transformer Network for Post-disaster 3D Semantic Segmentation

DGX agent

arXiv:2605.17197v1 Announce Type: cross Abstract: Post-disaster damage assessment requires rapid and accurate semantic segmentation of 3D point clouds to identify critical infrastructure such as damag

researcharxiv-cs-cv
19 May 2026
Agents

P2GS: Physical Prior-guided Gaussian Splatting for Photometrically Consistent Urban Reconstruction

DGX agent

arXiv:2605.16925v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) has recently emerged as a powerful explicit representation enabling fast, high-fidelity rendering, making it a promising fo

agentsarxiv-cs-cv
19 May 2026
Research

PanoWorld: A Generative Spatial World Model for Consistent Whole-House Panorama Synthesis

DGX agent

arXiv:2605.17916v1 Announce Type: new Abstract: Generating a consistent whole-house VR tour from a floorplan and style reference requires both photorealistic panoramas and cross-view spatial coherence

researcharxiv-cs-cv
19 May 2026
Applications

PartDiffuser: Part-wise 3D Mesh Generation via Discrete Diffusion

DGX agent

arXiv:2511.18801v3 Announce Type: replace Abstract: Existing autoregressive (AR) methods for generating artist-designed meshes struggle to balance global structural consistency with high-fidelity loca

applicationsarxiv-cs-cv
19 May 2026
Safety

Patch Ensembles for Robust Salmon Re-Identification with Weak Trajectory Labels

DGX agent

arXiv:2605.18038v1 Announce Type: new Abstract: Salmon re-identification in commercial net-pens is challenging due to large populations, which impose strict accuracy requirements and make large-scale

safetyarxiv-cs-cv
19 May 2026
Research

Patch-MoE Mamba: A Patch-Ordered Mixture-of-Experts State Space Architecture for Medical Image Segmentation

DGX agent

arXiv:2605.17719v1 Announce Type: new Abstract: CNN- and Transformer-based architectures have achieved strong performance in medical image segmentation, but CNNs are limited in modeling long-range dep

researcharxiv-cs-cv
19 May 2026
Research

Patchwork: A compact representation for 3D polygonal shapes

DGX agent

arXiv:2605.16266v1 Announce Type: cross Abstract: We introduce Patchwork, a new general-purpose shape representation capable of modeling 2D and 3D geometry with a small number of parameters. Patchwork

researcharxiv-cs-cv
19 May 2026
Model Releases

PERL: Parameter Efficient Reasoning in CLIP Latent Space

DGX agent

arXiv:2605.18464v1 Announce Type: new Abstract: Contrastively trained vision-language models such as CLIP provide strong zero-shot transfer by aligning images and text in a shared embedding space. How

model-releasesarxiv-cs-cv
19 May 2026
Research

PFlow-T: A Persistence-Driven Forward Process for Topology-Controlled Generation

DGX agent

arXiv:2605.17555v1 Announce Type: cross Abstract: Current topology aware diffusion models face an architectural mismatch by using Gaussian noise for corruption while recovering structural features thr

researcharxiv-cs-cv
19 May 2026
Tutorials

PhysSkin: Real-Time and Generalizable Physics-Based Animation via Self-Supervised Neural Skinning

DGX agent

arXiv:2603.23194v2 Announce Type: replace-cross Abstract: Achieving real-time physics-based animation that generalizes across diverse 3D shapes and discretizations remains a fundamental challenge. We

tutorialsarxiv-cs-cv
19 May 2026
Research

PIXLRelight: Controllable Relighting via Intrinsic Conditioning

DGX agent

arXiv:2605.18735v1 Announce Type: new Abstract: We present PIXLRelight, a feed-forward approach for physically controllable single-image relighting. Existing methods either provide limited lighting co

researcharxiv-cs-cv
19 May 2026
Applications

PlantPose: Universal Plant Skeleton Estimation via Tree-constrained Graph Generation

DGX agent

arXiv:2605.17773v1 Announce Type: new Abstract: Accurate estimation of plant skeletal structures (e.g., branching structures) from images is essential for smart agriculture and plant science. Unlike h

applicationsarxiv-cs-cv
19 May 2026
Safety

Position: Age Estimation Models Do Not Process Biometric Data

DGX agent

arXiv:2605.17347v1 Announce Type: cross Abstract: When a neural network estimates someone's age from a photograph, does it process biometric data? The answer depends on whether identity-discriminative

safetyarxiv-cs-cv
19 May 2026
Research

Principal Component Analysis for Lunar Crater Detection

DGX agent

arXiv:2605.17125v1 Announce Type: new Abstract: Optical navigation is a critical component for lunar orbiter and lander missions. Image-based crater identification has emerged as a promising technolog

researcharxiv-cs-cv
19 May 2026
Research

ProtoFlow: Mitigating Forgetting in Class-Incremental Remote Sensing Segmentation via Low-Curvature Prototype Flow

DGX agent

arXiv:2604.03212v2 Announce Type: replace Abstract: Remote sensing segmentation in real deployment is inherently continual: new semantic categories emerge, and acquisition conditions shift across seas

researcharxiv-cs-cv
19 May 2026
Local Ai

PySIFT: GPU-Resident Deterministic SIFT for Deep Learning Vision Pipelines

DGX agent

arXiv:2605.17869v1 Announce Type: new Abstract: A widespread assumption in local feature research holds that classical handcrafted descriptors are accuracy-limited relics best replaced by learned alte

local-aiarxiv-cs-cv
19 May 2026
Applications

QuadLink: Autoregressive Quad-Dominant Mesh Generation via Point-Relation Learning

DGX agent

arXiv:2605.16813v1 Announce Type: cross Abstract: The generation of production-ready quad-dominant meshes is a cornerstone of modern 3D content creation. Generating anisotropic quad-dominant meshes fr

applicationsarxiv-cs-cv
19 May 2026
Safety

Rad-VLSM: A Cross-Modal Framework with Semantics-Assisted Prompting for Medical Segmentation and Diagnosis

DGX agent

arXiv:2605.18130v1 Announce Type: new Abstract: Medical image segmentation is more clinically valuable when it supports diagnosis rather than merely producing lesion masks. However, diagnostically rel

safetyarxiv-cs-cv
19 May 2026
Research

RadGenome-Anatomy: A Large-Scale Anatomy-Labeled Chest Radiograph Dataset via Physically Grounded Volumetric Projection

DGX agent

arXiv:2605.17368v1 Announce Type: new Abstract: Anatomical structure labels for chest radiographs are essential for medical image segmentation and a broad range of downstream diagnostic tasks. However

researcharxiv-cs-cv
19 May 2026
Model Releases

Radial-Angular Geometry for Reliable Update Diagnosis in Noisy-Label Learning

DGX agent

arXiv:2605.17429v1 Announce Type: cross Abstract: Noisy-label methods often estimate sample reliability from forward-space signals such as loss, confidence, or entropy. These signals indicate whether

model-releasesarxiv-cs-cv
19 May 2026
Tutorials

RadJEPA: Radiology Encoder for Chest X-Rays via Joint Embedding Predictive Architecture

DGX agent

arXiv:2601.15891v2 Announce Type: replace Abstract: Recent advances in medical vision language models guide the learning of visual representations; however, this form of supervision is constrained by

tutorialsarxiv-cs-cv
19 May 2026
Safety

RAVE: Re-Allocating Visual Attention in Large Multimodal Models

DGX agent

arXiv:2605.18359v1 Announce Type: new Abstract: Large multimodal models (LMMs) inherit the self-attention mechanism of pretrained language backbones, yet standard attention can exhibit suboptimal allo

safetyarxiv-cs-cv
19 May 2026
Research

Real-Time Neural Hair Denoising

DGX agent

arXiv:2605.17557v1 Announce Type: cross Abstract: We propose a lightweight real-time method for reconstructing strand-based hair G-Buffers from severely undersampled rasterized inputs. Our pipeline fi

researcharxiv-cs-cv
19 May 2026
Model Releases

ReBaR: Reference-Based Reasoning for Robust Pose Estimation from Monocular Images

DGX agent

arXiv:2303.11675v3 Announce Type: replace Abstract: R}easoning for Robust Human Pose and Shape Estimation), designed to estimate human body shape and pose from single-view images. ReBaR effectively ad

model-releasesarxiv-cs-cv
19 May 2026
Safety

REC-RL: Referring expression counting via Gaussian and range-based reward optimization

DGX agent

arXiv:2605.16460v1 Announce Type: new Abstract: Referring expression counting (REC) is an intention-driven task that requires context-aware visual reasoning. While recent vision-language models incorp

safetyarxiv-cs-cv
19 May 2026
Safety

Resolving Representation Ambiguity in Feedforward Novel View Synthesis Transformer via Semantic-Spatial Decoupling

DGX agent

arXiv:2605.18599v1 Announce Type: new Abstract: Transformer-based models have advanced feedforward novel view synthesis (NVS). Current architectures such as GS-LRM and LVSM mix semantic information (e

safetyarxiv-cs-cv
19 May 2026
Research

Rethinking Generative Image Pretraining: How Far Are We From Scaling Up Next-Pixel Prediction?

DGX agent

arXiv:2511.08704v2 Announce Type: replace Abstract: This paper investigates the scaling properties of autoregressive next-pixel prediction, a simple, end-to-end yet under-explored framework for unifie

researcharxiv-cs-cv
19 May 2026
Local Ai

Rethinking Point Clouds as Sequences: A Causal Next-Token Predictive Learning Framework

DGX agent

arXiv:2605.17566v1 Announce Type: new Abstract: With the rapid progress of multimodal foundation models and predictive pre-training, an important open question is how to equip 3D point clouds with a p

local-aiarxiv-cs-cv
19 May 2026
Research

Rethinking the State Update Gate for Long-Sequence Recurrent 3D Reconstruction

DGX agent

arXiv:2605.16981v1 Announce Type: new Abstract: Streaming 3D reconstruction under a strict constant-memory budget hinges on how the recurrent state is updated as the stream evolves. We profile TTT3R-s

researcharxiv-cs-cv
19 May 2026
Applications

RHINO: Reconstructing Human Interactions with Novel Objects from Monocular Videos

DGX agent

arXiv:2605.17014v1 Announce Type: new Abstract: Reconstructing people, objects, and their interactions in 3D is a long-standing goal for intelligent systems. Often the input is RGB video from a moving

applicationsarxiv-cs-cv
19 May 2026
Safety

Right Predictions, Misleading Explanations: On the Vulnerability of Vision-Language Model Explanations

DGX agent

arXiv:2605.16651v1 Announce Type: new Abstract: Explanation mechanisms are increasingly used to support transparency and trust in vision-language models (VLMs), particularly in settings where model de

safetyarxiv-cs-cv
19 May 2026
Local Ai

Robo-Cortex: A Self-Evolving Embodied Agent via Dual-Grain Cognitive Memory and Autonomous Knowledge Induction

DGX agent

arXiv:2605.18729v1 Announce Type: cross Abstract: The ability to navigate and interact with complex environments is central to real-world embodied agents, yet navigation in unseen environments remains

local-aiarxiv-cs-cv
19 May 2026
Model Releases

ROVR-Open-Dataset: A Large-Scale Depth Dataset for Autonomous Driving

DGX agent

arXiv:2508.13977v3 Announce Type: replace Abstract: Depth estimation is a fundamental component of spatial perception for autonomous driving and other unmanned systems operating in open urban environm

model-releasesarxiv-cs-cv
19 May 2026
Research

RSEdit: Text-Guided Image Editing for Remote Sensing

DGX agent

arXiv:2603.13708v2 Announce Type: replace Abstract: In this paper, we explore text-guided image editing in the remote sensing domain using generative modeling. We propose rsedit, a collection of model

researcharxiv-cs-cv
19 May 2026
Research

RT-Splatting: Joint Reflection-Transmission Modeling with Gaussian Splatting

DGX agent

arXiv:2605.18263v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) enables real-time novel view synthesis with high visual quality. However, existing methods struggle with semi-transparent s

researcharxiv-cs-cv
19 May 2026
Safety

SafeDiffusion-R1: Online Reward Steering for Safe Diffusion Post-Training

DGX agent

arXiv:2605.18719v1 Announce Type: new Abstract: Diffusion models have been widely studied for removing unsafe content learned during pre-training. Existing methods require expensive supervised data, e

safetyarxiv-cs-cv
19 May 2026
Model Releases

SAM 2++: Tracking Anything at Any Granularity

DGX agent

arXiv:2510.18822v4 Announce Type: replace Abstract: Due to the varying granularity of target states across different tasks, most existing trackers are tailored to a single task, which specificity limi

model-releasesarxiv-cs-cv
19 May 2026
Local Ai

SAMRI: Segment Any MRI

DGX agent

arXiv:2510.26635v3 Announce Type: replace-cross Abstract: Summary: SAMRI is an MRI-specialized adaptation of the Segment Anything Model achieving superior whole-body MRI segmentation, particularly for

local-aiarxiv-cs-cv
19 May 2026
Research

SCAR: Self-Supervised Continuous Action Representation Learning

DGX agent

arXiv:2605.16412v1 Announce Type: cross Abstract: Despite the central role of action in embodied intelligence, learning transferable action representations from visual transitions remains a fundamenta

researcharxiv-cs-cv
19 May 2026
Model Releases

SCARED-C: Corrected Camera Poses for Endoscopic Depth Estimation

DGX agent

arXiv:2605.16628v1 Announce Type: new Abstract: The SCARED dataset is a widely used benchmark for endoscopic depth estimation, offering ground-truth 3D reconstructions captured with a structured light

model-releasesarxiv-cs-cv
19 May 2026
Local Ai

SeamCam: Quantifying Seamless Camouflage via Multi-Cue Visual Detectability

DGX agent

arXiv:2605.16515v1 Announce Type: new Abstract: Animals are described as effectively camouflaged when they blend seamlessly with their surrounding, yet no standardized quantitative measure of this sea

local-aiarxiv-cs-cv
19 May 2026
Research

See Silhouettes in Motion with Neuromorphic Vision

DGX agent

arXiv:2605.17984v1 Announce Type: cross Abstract: Quasi-bimodal objects, such as text, road signs, and barcodes, play a basic yet vital role in daily visual communication. By boiling these down to cle

researcharxiv-cs-cv
19 May 2026
Model Releases

Seeing Together:Multi-Robot Cooperative Egocentric Spatial Reasoning with Multimodal Large Language Models

DGX agent

arXiv:2605.18431v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have made substantial progress in egocentric video understanding, but their ability to reason cooperatively fro

model-releasesarxiv-cs-cv
19 May 2026
← Previous
1…162163164165166…263
Next →