AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,515 results
Agents

DreamSat-Pose: Spacecraft Pose Estimation from Single-View 3D Reconstructions and Learned 2D-3D Feature Matching

DGX agent

arXiv:2607.13449v1 Announce Type: new Abstract: 6-DoF pose estimation is a critical task in autonomous rendezvous and proximity operations. In the case of an unknown target, this task becomes challeng

agentsarxiv-cs-cv
16 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

DriveFace: A Cross-Spectral Through-Glass Face Dataset for On-the-Move Vehicular Border Control

DGX agent

arXiv:2607.13515v1 Announce Type: new Abstract: The continuous growth in cross-border mobility places increasing pressure on existing border control infrastructures, motivating on-the-move biometric a

researcharxiv-cs-cv
16 Jul 2026
Research

Efficient Computing for Medical Image Acquisition and Reconstruction

DGX agent

arXiv:2607.13204v1 Announce Type: cross Abstract: Medical imaging systems such as CT, MRI, PET, and SPECT do not directly acquire images. Instead, they measure physical signals that encode anatomical

researcharxiv-cs-cv
16 Jul 2026
Applications

Efficient LiDAR Reflectance Compression via Scanning Serialization

DGX agent

arXiv:2505.09433v3 Announce Type: replace Abstract: Reflectance attributes in LiDAR point clouds provide essential information for downstream tasks but remain underexplored in neural compression metho

applicationsarxiv-cs-cv
16 Jul 2026
Model Releases

EgoHTR: Egocentric 4D Demonstrations of Human Terrain Traversal

DGX agent

arXiv:2607.13472v1 Announce Type: cross Abstract: Deploying humanoid robots in unstructured terrain remains an open problem. While classic reinforcement learning struggles with the sheer complexity of

model-releasesarxiv-cs-cv
16 Jul 2026
Model Releases

EgoProceVQA: A Novel Egocentric Procedural Understanding Task with Self-Skill-Exploration Agent

DGX agent

arXiv:2607.13792v1 Announce Type: new Abstract: Most daily activities are inherently procedural. However, existing evaluations for egocentric video understanding seldom address procedural understandin

model-releasesarxiv-cs-cv
16 Jul 2026
Model Releases

Evaluating Vision Foundation Models for Pixel and Object Classification in Microscopy

DGX agent

arXiv:2603.19802v2 Announce Type: replace Abstract: Deep learning underlies most modern approaches and tools in computer vision, including biomedical imaging. However, for interactive semantic segment

model-releasesarxiv-cs-cv
16 Jul 2026
Model Releases

Exploratory, Communicative, and Deployable: Vision-Driven Embodied Agents for Open-World Mobile Manipulation

DGX agent

arXiv:2607.13653v1 Announce Type: new Abstract: Real-world deployment of embodied agents requires active exploration, visual grounding, and interactive intent disambiguation. However, existing framewo

model-releasesarxiv-cs-cv
16 Jul 2026
Model Releases

FastCentNN: Accelerating Centroid Neural Network with Entropy Proxy

DGX agent

arXiv:2607.13613v1 Announce Type: cross Abstract: Centroid neural network (CentNN) is an unsupervised competitive learning algorithm in which centroid splitting is triggered only after strict local st

model-releasesarxiv-cs-cv
16 Jul 2026
Model Releases

Fine-grained CLIP fine-tuning with self-annotated region alignment

DGX agent

arXiv:2607.13661v1 Announce Type: new Abstract: Contrastive Language-Image Pre-training (CLIP) has been shown to have limitations in its fine-grained dense feature representation, due to its pre-train

model-releasesarxiv-cs-cv
16 Jul 2026
Safety

Fine-Grained Vision-Language Pretraining with Organ-Conditioned Pattern Tokens for CT Understanding

DGX agent

arXiv:2607.13892v1 Announce Type: new Abstract: Computed tomography (CT) vision-language pretraining from paired volumes and radiology reports is a scalable yet challenging task. Existing methods comm

safetyarxiv-cs-cv
16 Jul 2026
Model Releases

FM^2: Unified Federated Foundation Models for Heterogeneous Multimodal Medical Imaging

DGX agent

arXiv:2607.13386v1 Announce Type: new Abstract: Building foundation models for medical imaging requires pooling data across institutions, yet privacy regulations prohibit centralized aggregation. Exis

model-releasesarxiv-cs-cv
16 Jul 2026
Tutorials

FOLIO: Focused Semantic Memory for Streaming Video Understanding

DGX agent

arXiv:2607.13298v1 Announce Type: new Abstract: In online streaming video understanding, a video stream continues to arrive and queries may be issued at any time. Because streaming frames grow without

tutorialsarxiv-cs-cv
16 Jul 2026
Safety

FreeLit: Paired-Free Indoor Relighting via Physics-Guided Diffusion

DGX agent

arXiv:2607.13656v1 Announce Type: new Abstract: Image-based indoor scene relighting remains challenging due to the complex interplay between cluttered geometry and local illumination, requiring precis

safetyarxiv-cs-cv
16 Jul 2026
Research

From Pixels to States: Rethinking Interactive World Models as Game Engines

DGX agent

arXiv:2607.14076v1 Announce Type: new Abstract: Building interactive worlds that respond coherently to player actions has long been a shared goal of computer graphics, games, and artificial intelligen

researcharxiv-cs-cv
16 Jul 2026
Model Releases

From Surface Forecasting to Observability Forecasting: A Latent World Model for Cloud-Aware EO Monitoring

DGX agent

arXiv:2607.13651v1 Announce Type: new Abstract: The bottleneck of Earth Observation processing chains is not the arrival of new imagery but whether the surface is actually visible when the image arriv

model-releasesarxiv-cs-cv
16 Jul 2026
Safety

Generalizable VLA Finetuning via Representation Anchoring and Language-Action Alignment

DGX agent

arXiv:2607.13429v1 Announce Type: cross Abstract: Finetuning a pretrained vision-language model (VLM) on robot demonstrations via behavior cloning (BC) has become the standard recipe for vision-langua

safetyarxiv-cs-cv
16 Jul 2026
Model Releases

GPOcc++: Unified Sparse Gaussian Occupancy Prediction with Visual Geometry Priors

DGX agent

arXiv:2607.13481v1 Announce Type: new Abstract: Accurate 3D scene understanding is fundamental to embodied intelligence and autonomous driving, where 3D occupancy provides a unified representation of

model-releasesarxiv-cs-cv
16 Jul 2026
Research

HIVE-3D: Hierarchical Voxel Enhancement for High-Quality 3D Scene Generation

DGX agent

arXiv:2607.13468v1 Announce Type: new Abstract: Recently, a line of works can generate impressive 3D objects from a single image, but they are limited by restricted representation resolution, making t

researcharxiv-cs-cv
16 Jul 2026
Research

Improving Medical Image Generative Models with Frechet Distance Loss

DGX agent

arXiv:2607.13300v1 Announce Type: new Abstract: Diffusion generative models have demonstrated immense potential for synthetic medical image generation. However, these models often struggle to capture

researcharxiv-cs-cv
16 Jul 2026
Research

Just-In-Time Scene Graph Growth: Combating Perceptual Saturation in Long-Horizon Robotics

DGX agent

arXiv:2607.13245v1 Announce Type: new Abstract: While 3D Scene Graphs (3DSGs) provide crucial structured representations for embodied agents, conventional Ahead-of-Time, build-everything-then-filter p

researcharxiv-cs-cv
16 Jul 2026
Research

Kepler-Encoder-v0.1: Towards a Multimodal Embedding Model for Robots

DGX agent

arXiv:2607.13522v1 Announce Type: cross Abstract: A robot must understand the state of its own body, but a camera sees only part of it. Force and contact leave almost no trace in a single frame, and r

researcharxiv-cs-cv
16 Jul 2026
Research

LaME: Learning to Think in Latent Space for Multimodal Embedding via Information Bottleneck

DGX agent

arXiv:2606.13061v2 Announce Type: replace Abstract: Reasoning-driven universal multimodal embedding has advanced rapidly by introducing Chain-of-Thought (CoT) reasoning into the embedding pipeline. De

researcharxiv-cs-cv
16 Jul 2026
Applications

Learning Speaker Identity Beyond Language and Modality Constraints: Insights from the POLY-SIM 2026 Challenge

DGX agent

arXiv:2607.13669v1 Announce Type: new Abstract: Multimodal speaker identification systems typically assume the availability of complete and homogeneous audio-visual modalities during both training and

applicationsarxiv-cs-cv
16 Jul 2026
Model Releases

Look Again Before You Abstain:Budgeted Conformal Evidence Acquisition for Reliable Vision-Language Model

DGX agent

arXiv:2606.16667v2 Announce Type: replace Abstract: Large vision-language models (LVLMs) hallucinate: they assert visual details that the image does not support. A principled remedy is selective predi

model-releasesarxiv-cs-cv
16 Jul 2026
Local Ai

Low-latency Event-based Object Detection with Spatially-Sparse Linear Attention

DGX agent

arXiv:2603.06228v2 Announce Type: replace Abstract: Event cameras provide sequential visual data with spatial sparsity and high temporal resolution, making them attractive for low-latency object detec

local-aiarxiv-cs-cv
16 Jul 2026
Applications

LPM: Industrial-Scale Generative Video Restoration

DGX agent

arXiv:2607.13460v1 Announce Type: new Abstract: We present the Large Processing Model (LPM), a diffusion-based generative framework for photorealistic video restoration under complex, in-the-wild degr

applicationsarxiv-cs-cv
16 Jul 2026
Applications

M2P-AD: Memory-to-Prototype Learning with Boundary-aware Score Refinement for 3D Anomaly Detection

DGX agent

arXiv:2607.13499v1 Announce Type: new Abstract: 3D anomaly detection has recently emerged as an important research topic in computer vision. Although existing methods have achieved high performance, e

applicationsarxiv-cs-cv
16 Jul 2026
Safety

Marker-free deformable registration and fusion for augmented reality-guided positive margin localization during tumor resection surgery

DGX agent

arXiv:2607.13343v1 Announce Type: new Abstract: Positive margins in head and neck oncologic surgery require mapping specimen-side pathology findings to the patient resection bed. This is challenging b

safetyarxiv-cs-cv
16 Jul 2026
Agents

M^ext{4}World: A Multi-view Multimodal Driving World Model for Interactive Object Manipulation and Minute-long Streaming

DGX agent

arXiv:2607.14005v1 Announce Type: new Abstract: Driving-world generation has emerged as a core capability for scalable autonomous-driving simulation, yet existing methods remain limited in object-leve

agentsarxiv-cs-cv
16 Jul 2026
Research

MGFace: Mask-Gated Face Matching via Conditional Similarity Routing

DGX agent

arXiv:2607.13187v1 Announce Type: new Abstract: Face identification has achieved remarkable performance under normal conditions. Yet, its accuracy often degrades significantly when query faces are par

researcharxiv-cs-cv
16 Jul 2026
Applications

Multi-view Hand Reconstruction with a Point-Embedded Transformer

DGX agent

arXiv:2408.10581v3 Announce Type: replace Abstract: This work introduces a novel and generalizable multi-view Hand Mesh Reconstruction (HMR) model, named POEM, designed for practical use in real-world

applicationsarxiv-cs-cv
16 Jul 2026
Research

MultiAnimate: A Unified Framework for Controllable Multi-Character Animation

DGX agent

arXiv:2607.13415v1 Announce Type: new Abstract: Recent advances in generative models and technological innovations have significantly addressed the fundamental challenges of character image animation.

researcharxiv-cs-cv
16 Jul 2026
Hardware

NanoGS: Training-Free Gaussian Splat Simplification

DGX agent

arXiv:2603.16103v2 Announce Type: replace Abstract: 3D Gaussian Splat (3DGS) enables high-fidelity, real-time novel view synthesis by representing scenes with large sets of anisotropic primitives, but

hardwarearxiv-cs-cv
16 Jul 2026
Model Releases

NeMo: Needle in a Montage for Video-Language Understanding

DGX agent

arXiv:2509.24563v3 Announce Type: replace Abstract: Recent advances in video large language models (VideoLLMs) call for new evaluation protocols and benchmarks for video-language understanding. Inspir

model-releasesarxiv-cs-cv
16 Jul 2026
Research

Nexus: Native Mesh Generation with Diffusion

DGX agent

arXiv:2607.13563v1 Announce Type: new Abstract: Generating high-quality triangle meshes is essential for film, gaming, and interactive 3D applications. Mainstream methods rely on mesh serialization an

researcharxiv-cs-cv
16 Jul 2026
Model Releases

OccTrack360: 4D Panoptic Occupancy Tracking from Surround-View Fisheye Cameras

DGX agent

arXiv:2603.08521v2 Announce Type: replace Abstract: Understanding dynamic 3D environments in a spatially continuous and temporally consistent manner is fundamental for robotics and autonomous driving.

model-releasesarxiv-cs-cv
16 Jul 2026
Model Releases

Peak-End-Net: A Peak-End Rule Inspired Framework for Generalizable Video Aesthetic Assessment

DGX agent

arXiv:2607.13941v1 Announce Type: new Abstract: Video aesthetic assessment (VAA) aims to predict how aesthetically pleasing a video is, yet remains far less explored than other visual assessment tasks

model-releasesarxiv-cs-cv
16 Jul 2026
Model Releases

PiVoT: A Variational Solution for Real-time Large-scale Multi-object Detection and Tracking under Heavy Clutter

DGX agent

arXiv:2607.13891v1 Announce Type: cross Abstract: Multi-object detection and tracking from noisy point clouds remain challenging in many data-scarce radar applications. Current Bayesian trackers based

model-releasesarxiv-cs-cv
16 Jul 2026
Research

PlumeQuant: Uncertainty-aware consistency assessment of methane plume masks and emission-rate estimates

DGX agent

arXiv:2607.13945v1 Announce Type: new Abstract: Imaging spectrometers increasingly distribute source-resolved methane plume products in which the plume mask, integrated mass enhancement (IME), plume l

researcharxiv-cs-cv
16 Jul 2026
Research

Prospective clinical indication, post-hoc report leakage, and fusion design in multi-image chest radiograph classification: a patient-clustered evaluation

DGX agent

arXiv:2607.13800v1 Announce Type: cross Abstract: Chest radiograph datasets often combine multiple images with Clinical Indication, Findings, and Impression, although these inputs are produced at diff

researcharxiv-cs-cv
16 Jul 2026
Model Releases

Quantum Circuit Vision: Cost-Aware Evaluation of Visual AI Agents for Quantum Code Generation

DGX agent

arXiv:2607.10057v1 Announce Type: cross Abstract: Can AI agents visually comprehend quantum circuit diagrams and generate verified executable code--and at what cost? We present Quantum Circuit Vision,

model-releasesarxiv-cs-cv
16 Jul 2026
Research

RainDancer: RGB-Event Video Deraining with Rain-Oriented Spiking Dynamics

DGX agent

arXiv:2607.13802v1 Announce Type: new Abstract: Video deraining aims to recover clean visual content from rainy videos for reliable perception under adverse weather. Existing methods mainly rely on RG

researcharxiv-cs-cv
16 Jul 2026
Agents

Recursive ArUco Markers: A Scalable Fiducial Marker Design for Unmanned Aerial Vehicle Landing Pads

DGX agent

arXiv:2607.13830v1 Announce Type: cross Abstract: Unmanned Aerial Vehicles (UAVs) increasingly rely on visual fiducial markers for autonomous navigation and precision landing. However, standard marker

agentsarxiv-cs-cv
16 Jul 2026
Tutorials

Reflecting Process Expertise in Procedural Material Generation

DGX agent

arXiv:2607.13318v1 Announce Type: new Abstract: Procedural material creation underpins applications in digital content creation, visual effects, and 3D asset design. Achieving high-quality results req

tutorialsarxiv-cs-cv
16 Jul 2026
Local Ai

RoughNet: Mapping Arctic Sea Ice Roughness Using Diffusion-Based Super-Resolution of Satellite Imagery

DGX agent

arXiv:2607.13371v1 Announce Type: new Abstract: Accurate estimation of landfast sea ice roughness is critical for climate modeling and safe Arctic over-ice travel, yet existing approaches rely on cost

local-aiarxiv-cs-cv
16 Jul 2026
Research

SalientGS: Unified SfM-to-3DGS with Importance-Guided MCMC Gaussian Allocation

DGX agent

arXiv:2607.11285v2 Announce Type: replace Abstract: Reconstructing 3D scenes from unordered images remains bottlenecked by expensive Structure-from-Motion (SfM) preprocessing and frozen pose interface

researcharxiv-cs-cv
16 Jul 2026
Safety

SARFA: Segment Anything with Radiomic Feature Alignment

DGX agent

arXiv:2607.13323v1 Announce Type: new Abstract: The Segment Anything Model (SAM) has demonstrated strong generalizability across a variety of segmentation tasks. However, SAM often struggles in situat

safetyarxiv-cs-cv
16 Jul 2026
← Previous
1…4950515253…261
Next →