AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlog
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,618 results
Research

Multi-Channel Uncertainty-Weighted Score Matching for Conditional Diffusion in Medical UDA

DGX agent

arXiv:2509.22476v2 Announce Type: replace Abstract: Robust medical image segmentation across modalities remains challenging due to severe domain shifts and the lack of target-domain labels. While diff

researcharxiv-cs-cv
1 Jul 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Multimodal Benchmark for Safety Assessment in Industrial Inspection Scenarios

DGX agent

arXiv:2601.21173v2 Announce Type: replace-cross Abstract: With the rapid development of industrial intelligence and unmanned inspection, reliable perception and safety assessment for AI systems in com

model-releasesarxiv-cs-cv
1 Jul 2026
Applications

MuSViT: A Foundation Vision Model for Sheet Music Representation

DGX agent

arXiv:2606.31811v1 Announce Type: new Abstract: Foundation models have transformed vision and language processing by providing rich, reusable representations that transfer across diverse tasks. Sheet

applicationsarxiv-cs-cv
1 Jul 2026
Research

MV-GEL: Language-Driven Multi-View Geometric Entity Localization on Meshes

DGX agent

arXiv:2606.31533v1 Announce Type: new Abstract: Identifying and grounding precise geometric entities, such as edges, planar regions, and curved surfaces within 3D objects, is foundational to computer-

researcharxiv-cs-cv
1 Jul 2026
Safety

No Adaptation Without Observation: Observability-Constrained Test-Time Prompt Tuning for LiDAR Semantic Segmentation

DGX agent

arXiv:2606.30937v1 Announce Type: new Abstract: LiDAR semantic segmentation often degrades under real-world deployment due to evolving sensing conditions, while collecting new annotations for retraini

safetyarxiv-cs-cv
1 Jul 2026
Model Releases

No Place to Hide: Benchmarking Video Hallucination with Background-Controlled Pairs

DGX agent

arXiv:2606.31933v1 Announce Type: new Abstract: We introduce VidPair-Halluc, a new benchmark for evaluating video hallucination in large video models (LVMs) under rigorous and controlled conditions. U

model-releasesarxiv-cs-cv
1 Jul 2026
Research

No Prompt, No Leaks: A Robust Generative Steganography Framework via Prompt-Free Diffusion

DGX agent

arXiv:2606.31427v1 Announce Type: new Abstract: Generative image steganography synthesizes stego images directly from secret information to achieve inherent security advantages. Latent Diffusion Model

researcharxiv-cs-cv
1 Jul 2026
Model Releases

NURBS Splatting: A Unified Differentiable Rendering Framework for Vector Graphics

DGX agent

arXiv:2606.31764v1 Announce Type: cross Abstract: Differentiable rendering of planar rational splines remains largely underexplored, despite their widespread use in vector graphics and design. Existin

model-releasesarxiv-cs-cv
1 Jul 2026
Safety

Off the Rails: Hijacking the Scoring Head in Generative End-to-End Driving Planners with Safety-Violating Adversarial Perturbations

DGX agent

arXiv:2606.30807v1 Announce Type: cross Abstract: Generative models have recently seen rapid adoption in End-to-End (E2E) autonomous driving (AD), with diffusion-based denoising and vocabulary-based r

safetyarxiv-cs-cv
1 Jul 2026
Research

On the Role of Rotation Equivariance in Monocular 2D-to-3D Human Pose Lifting

DGX agent

arXiv:2601.13913v2 Announce Type: replace Abstract: Estimating 3D from 2D is one of the central tasks in computer vision. In this work, we consider the monocular setting, i.e. single-view input, for 3

researcharxiv-cs-cv
1 Jul 2026
Model Releases

One Video, One World: Turning Monocular Video into Physical 4D Scenes

DGX agent

arXiv:2606.31388v1 Announce Type: new Abstract: We introduce extbf{OVOW}, the first training-free system that reconstructs instance-level, simulation-ready 4D mesh scenes from a single monocular video

model-releasesarxiv-cs-cv
1 Jul 2026
Research

Online TT-ALS for Streaming Tensor Decomposition with Incremental Orthogonalization

DGX agent

arXiv:2606.31061v1 Announce Type: cross Abstract: Tensor Train (TT) decomposition is a powerful technique for analyzing high-dimensional data. Existing algorithms for computing TT decompositions can b

researcharxiv-cs-cv
1 Jul 2026
Safety

PA-VAD: Diffusion-Based Pseudo-Only Video Anomaly Detection via Domain-Aligned Memory Updates

DGX agent

arXiv:2512.06845v2 Announce Type: replace Abstract: Deploying video anomaly detection (VAD) in the real world is often constrained by the scarcity, privacy, and cost of collecting real abnormal footag

safetyarxiv-cs-cv
1 Jul 2026
Research

Pano3D: Unified 3D Reconstruction and Panoptic Segmentation

DGX agent

arXiv:2606.14307v2 Announce Type: replace Abstract: Recent advances in 3D feedforward reconstruction neural networks have achieved remarkable success in dense reconstruction from images without any ca

researcharxiv-cs-cv
1 Jul 2026
Research

Patient-Level Elbow Abnormality Detection: Leakage-Aware Evaluation of Learned Preprocessing, Calibration, and Triage-Oriented Operating Points

DGX agent

arXiv:2606.31348v1 Announce Type: new Abstract: In this study, we examine learned preprocessing pipelines in the context of triage-oriented orthopedic abnormality detection task using elbow radiograph

researcharxiv-cs-cv
1 Jul 2026
Tutorials

Phantom: A Unified Face-Swap Deepfake Protection Framework with Latent and Spatial Constraints

DGX agent

arXiv:2606.31703v1 Announce Type: new Abstract: Face-swapping deepfakes pose an escalating threat to personal privacy by enabling unauthorized identity manipulation. While adversarial approaches have

tutorialsarxiv-cs-cv
1 Jul 2026
Research

Phase-Aligned RoPE for Mixed-Resolution Diffusion Transformer

DGX agent

arXiv:2511.19778v2 Announce Type: replace Abstract: Rotary positional embeddings (RoPE) are widely used in diffusion transformers (DiTs) to encode spatial relationships, yet their behavior with mixed-

researcharxiv-cs-cv
1 Jul 2026
Research

{Phi}eat: Physically Grounded Material Feature Representation

DGX agent

arXiv:2511.11270v2 Announce Type: replace Abstract: While foundation models have emerged as general-purpose visual backbones, their representations are primarily optimized for semantics and lack expli

researcharxiv-cs-cv
1 Jul 2026
Research

PhotoQuilt: Training-Free Arbitrary-Resolution Photomosaics via Bootstrapped Tiled Denoising

DGX agent

arXiv:2606.30968v1 Announce Type: new Abstract: Photomosaics are large images whose local regions are seen as independent tiles while their overall arrangement forms a coherent scene. Generating them

researcharxiv-cs-cv
1 Jul 2026
Safety

PiLoT v2: Pixel-to-Orthogonal Map Alignment for Free-view UAV Geo-localization

DGX agent

arXiv:2606.31098v1 Announce Type: new Abstract: Real-time, drift-free UAV geo-localization is essential for autonomous missions in GNSS-denied environments. The pioneering system, PiLoT, achieves high

safetyarxiv-cs-cv
1 Jul 2026
Model Releases

Planar-SfM: Camera Pose Estimation via Homography Graph Embeddings

DGX agent

arXiv:2606.31979v1 Announce Type: new Abstract: Structure from Motion (SfM) systems traditionally struggle with planar scenes, where standard epipolar geometry-based methods become degenerate. Rather

model-releasesarxiv-cs-cv
1 Jul 2026
Tutorials

PointSplat: Compact Gaussian Splatting via Human-Centric Prediction

DGX agent

arXiv:2606.32036v1 Announce Type: new Abstract: Producing 3D human representations from input views on the fly is essential for immersive live streaming systems, where representation compactness is as

tutorialsarxiv-cs-cv
1 Jul 2026
Model Releases

PolyFlow: Continuous Topology Embedding Flow Matching for Artist-style Mesh Generation

DGX agent

arXiv:2606.30673v1 Announce Type: cross Abstract: Autoregressive Transformers dominate high-quality mesh generation by producing artist-worthy topologies, yet their inherent sequential decoding induce

model-releasesarxiv-cs-cv
1 Jul 2026
Research

PoseGravity: Pose Estimation from Points and Lines with Axis Prior

DGX agent

arXiv:2405.12646v3 Announce Type: replace Abstract: This paper presents a new algorithm to estimate absolute camera pose given an axis of the camera's rotation matrix. Current algorithms solve the pro

researcharxiv-cs-cv
1 Jul 2026
Research

Practical High-Fidelity Novel-View Synthesis of Mounted Lepidoptera

DGX agent

arXiv:2606.31679v1 Announce Type: cross Abstract: Mounted butterflies are among the most striking objects in natural history collections. However, their beauty is notoriously hard to digitize in 3D: t

researcharxiv-cs-cv
1 Jul 2026
Model Releases

PriorEye: Geospatial Visual Priors for End-to-End Autonomous Driving

DGX agent

arXiv:2606.31830v1 Announce Type: new Abstract: Most end-to-end autonomous driving methods rely solely on instantaneous sensor observations, limiting them to reactive behavior without the anticipatory

model-releasesarxiv-cs-cv
1 Jul 2026
Model Releases

PrISM-IQA: Image Quality Assessment Made Practical for Smartphone Photography

DGX agent

arXiv:2606.31626v1 Announce Type: new Abstract: Existing smartphone image quality assessment (IQA) methods commonly reduce perceptual quality to a single score. However, this scalar formulation is poo

model-releasesarxiv-cs-cv
1 Jul 2026
Research

PRISM: Latent Composition Consistency for Single-Image Reflection Removal

DGX agent

arXiv:2606.31513v1 Announce Type: new Abstract: Single-image reflection removal (SIRR) seeks to recover the transmission layer from a mixture corrupted by reflections -- a severely ill-posed problem.

researcharxiv-cs-cv
1 Jul 2026
Research

Proxy-GS: Unified Occlusion Priors for Training and Inference in Structured 3D Gaussian Splatting

DGX agent

arXiv:2509.24421v5 Announce Type: replace Abstract: 3D Gaussian Splatting (3DGS) has emerged as an efficient approach for achieving photorealistic rendering. Recent MLP-based variants further improve

researcharxiv-cs-cv
1 Jul 2026
Local Ai

PSHuman: Photorealistic Single-image 3D Human Reconstruction using Cross-Scale Multiview Diffusion and Explicit Remeshing

DGX agent

arXiv:2409.10141v3 Announce Type: replace Abstract: Detailed and photorealistic 3D human modeling is essential for various applications and has seen tremendous progress. However, full-body reconstruct

local-aiarxiv-cs-cv
1 Jul 2026
Research

RASR: Retrieval-Augmented Semantic Reasoning for Fake News Video Detection

DGX agent

arXiv:2604.06687v2 Announce Type: replace Abstract: Multimodal fake news video detection is a crucial research direction for maintaining the credibility of online information. Existing studies primari

researcharxiv-cs-cv
1 Jul 2026
Applications

RCL-Mamba: A Dual-domain State Space Model for Measurement-oriented Image Restoration in Rotational Sparse-View Scanning Computed Laminography

DGX agent

arXiv:2606.31353v1 Announce Type: new Abstract: Rotational Scanning Computed Laminography (RCL) is widely utilized for the Non-Destructive Testing(NDT) of large planar components. However, to facilita

applicationsarxiv-cs-cv
1 Jul 2026
Agents

Reasoning-aware Speculative Decoding for Efficient Vision-Language-Action Models in Autonomous Driving

DGX agent

arXiv:2606.31160v1 Announce Type: new Abstract: Modern Vision-Language-Action (VLA) planners for autonomous driving emit a chain-of-causation (CoC) reasoning step before producing a trajectory. The re

agentsarxiv-cs-cv
1 Jul 2026
Local Ai

Reasoning in machine vision by learning fast and slow thinking

DGX agent

arXiv:2506.22075v2 Announce Type: replace Abstract: Reasoning is a hallmark of human intelligence, enabling adaptive decision-making in complex unfamiliar scenarios. In contrast, machine intelligence

local-aiarxiv-cs-cv
1 Jul 2026
Research

REDI: Corpus Aware Patch Ranking for DINOv3 Token Reduction

DGX agent

arXiv:2606.31676v1 Announce Type: new Abstract: Most token reduction methods for Vision Transformers seek favorable tradeoffs between accuracy and efficiency by pruning, merging, or pooling patch toke

researcharxiv-cs-cv
1 Jul 2026
Model Releases

Reference-Free Image Quality Assessment for Virtual Try-On via Human Feedback

DGX agent

arXiv:2603.13057v2 Announce Type: replace Abstract: As virtual try-on (VTON) systems become increasingly important in fashion e-commerce, there is a growing need for reliable reference-free evaluation

model-releasesarxiv-cs-cv
1 Jul 2026
Tutorials

Region-Aware Multimodal Large Language Model via SlowFast Tokenization and Pseudo-Mask Guidance for 3D CT Report Generation

DGX agent

arXiv:2506.23102v2 Announce Type: replace-cross Abstract: Current CT report generation frameworks predominantly rely on global feature representations, often failing to capture region-specific details

tutorialsarxiv-cs-cv
1 Jul 2026
Applications

Registering the 4D Millimeter Wave Radar Point Clouds Via Generalized Method of Moments

DGX agent

arXiv:2508.02187v3 Announce Type: replace-cross Abstract: 4D millimeter wave radars (4D radars) are new emerging sensors that provide point clouds of objects with both position and radial velocity mea

applicationsarxiv-cs-cv
1 Jul 2026
Model Releases

RESOLVE: A Multi-Resolution and Multi-Modal Dataset for Roadside Cooperative Perception

DGX agent

arXiv:2606.31895v1 Announce Type: new Abstract: LiDAR has increasingly been integrated into traffic cameras to expand coverage and mitigate occlusion in roadside cooperative perception. However, how u

model-releasesarxiv-cs-cv
1 Jul 2026
Research

Rethinking Foundation Model Collaboration: Enhancing Specialized Models through Proxy Task Reasoning

DGX agent

arXiv:2606.31157v1 Announce Type: new Abstract: Foundation models are increasingly integrated into embodied intelligence systems, but directly assigning them structured prediction tasks requires preci

researcharxiv-cs-cv
1 Jul 2026
Model Releases

Rethinking the Role of Feature Engineering and Learning Strategies in Few-Shot Hidden Emotion Recognition

DGX agent

arXiv:2606.31249v1 Announce Type: new Abstract: In this paper, we present the solution developed by our team, XInsight Lab, which achieved first place in Track 3 of the 4th EI-MIGA-IJCAI Challenge wit

model-releasesarxiv-cs-cv
1 Jul 2026
Model Releases

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios

DGX agent

arXiv:2512.24561v2 Announce Type: replace Abstract: Visual grounding (VG) localizes target objects in an image from natural-language expressions. In real-world perception, RGB cues often degrade under

model-releasesarxiv-cs-cv
1 Jul 2026
Safety

Rhythm-Structured Predictive Learning for Remote Photoplethysmography

DGX agent

arXiv:2606.31736v1 Announce Type: new Abstract: Remote photoplethysmography (rPPG) estimates physiological signals from facial videos by analyzing subtle pulse induced skin color variations. Despite r

safetyarxiv-cs-cv
1 Jul 2026
Agents

Robust Autonomous UAV Landing on Maritime Platforms via Multimodal Agentic AI and Active Wave Compensation

DGX agent

arXiv:2606.31613v1 Announce Type: new Abstract: Autonomous aerial inspection of marine infrastructure is frequently compromised by stochastic sea states, introducing risks of high-kinetic impacts, pos

agentsarxiv-cs-cv
1 Jul 2026
Research

SAMBA: A Scatter-Guided Masked Bidirectional Mamba Foundation Model for SAR Target Recognition

DGX agent

arXiv:2606.31668v1 Announce Type: new Abstract: Synthetic aperture radar automatic target recognition (SAR ATR) is critical for Earth observation and defense, but its practical deployment is constrain

researcharxiv-cs-cv
1 Jul 2026
Research

Seeing Through the Weights: Privacy Leakage in Scene Coordinate Regression

DGX agent

arXiv:2606.31164v1 Announce Type: new Abstract: Scene Coordinate Regression (SCR) methods are increasingly adopted for visual localization. In these approaches, the scene is implicitly encoded within

researcharxiv-cs-cv
1 Jul 2026
Research

Self-Supervised Temporal Regularization for Landmark-Based Cardiac Segmentation with Automatic AHA Regional Mapping

DGX agent

arXiv:2606.31785v1 Announce Type: new Abstract: Graph-based cardiac segmentation with implicit anatomical correspondences provides topological guarantees and population-level analysis capabilities, bu

researcharxiv-cs-cv
1 Jul 2026
Research

Semantic-Aware Multiple Access via Spatial Redundancy Exploitation for Uplink-Dominant 6G Use Cases

DGX agent

arXiv:2606.31715v1 Announce Type: cross Abstract: Emerging uplink-dominant 6G use cases, such as cooperative vehicular streaming, require efficient transmission of high-volume visual data over limited

researcharxiv-cs-cv
1 Jul 2026
← Previous
1…7576777879…263
Next →