AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,618 results
Research

Optimizing Few-Step Generation with Adaptive Matching Distillation

DGX agent

arXiv:2602.07345v2 Announce Type: replace Abstract: Distribution Matching Distillation (DMD) is a powerful acceleration paradigm, yet its stability is often compromised in Forbidden Zone, regions wher

researcharxiv-cs-cv
9 Jun 2026
Safety
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

OrderDP: A Theoretically Guaranteed Lossless Dynamic Data Pruning Framework

DGX agent

arXiv:2606.08574v1 Announce Type: cross Abstract: Data pruning (DP), as an oft-stated strategy to alleviate heavy training burdens, reduces the volume of training samples according to a well-defined p

safetyarxiv-cs-cv
9 Jun 2026
Safety

PairWise Image Finder: An Open-source Tool for Finding Visually Aligned Street-Level Image Pairs for Urban Perception Studies

DGX agent

arXiv:2606.08795v1 Announce Type: new Abstract: Change detection and scene recognition techniques have been widely applied to Street View Imagery (SVI) to understand changes in scenes across the years

safetyarxiv-cs-cv
9 Jun 2026
Research

Pantheon360: Taming Digital Twin Generation via 3D-Aware 360{eg} Video Diffusion

DGX agent

arXiv:2605.25449v2 Announce Type: replace Abstract: Generating complete digital twins from videos requires precise camera control, global scene coverage, and strict spatial-temporal consistency constr

researcharxiv-cs-cv
9 Jun 2026
Model Releases

PEDRA: Evaluating the Realism of Pedestrian Dynamics in Video Generation

DGX agent

arXiv:2510.20182v2 Announce Type: replace Abstract: Pedestrian simulation traditionally relies on expert-tuned, hand-crafted models that limit scalability and generalization. Meanwhile, large-scale vi

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

PereStruct: Multimodal Semantic Assembly for Robust Historical Document Parsing

DGX agent

arXiv:2606.07661v1 Announce Type: new Abstract: Parsing historical documents with complex, non-standard layouts remains a fundamental bottleneck in large-scale archival digitization. Unlike modern typ

model-releasesarxiv-cs-cv
9 Jun 2026
Research

Phase Marginalization for Patch-Grid Instability in Vision Transformers

DGX agent

arXiv:2606.08132v1 Announce Type: new Abstract: Vision Transformers operate on fixed patch grids, which can introduce phase-dependent instability for dense prediction: changing the patch partition can

researcharxiv-cs-cv
9 Jun 2026
Local Ai

PhysAgent: Automating Physics-Based 4D Synthesis via Trajectory-Grounded Multi-Agent Feedback

DGX agent

arXiv:2606.08688v1 Announce Type: cross Abstract: Achieving fully automated, physically plausible 3D motion synthesis is a core objective in graphics and generative AI. However, configuring complex en

local-aiarxiv-cs-cv
9 Jun 2026
Applications

PhysGraph: A Physics-aware 3D Scene Graph for Perception and Reasoning

DGX agent

arXiv:2606.08655v1 Announce Type: cross Abstract: To perform a wide range of daily tasks, robots need to construct a 3D representation that is semantically rich, physically grounded, and structured en

applicationsarxiv-cs-cv
9 Jun 2026
Local Ai

PicoSAM3: Real-Time In-Sensor Region-of-Interest Segmentation

DGX agent

arXiv:2603.11917v2 Announce Type: replace Abstract: Real-time, on-device segmentation is critical for latency-sensitive and privacy-aware applications such as smart glasses and Internet-of-Things devi

local-aiarxiv-cs-cv
9 Jun 2026
Safety

Polaffini: A feature-based approach for robust affine and polyaffine image registration

DGX agent

arXiv:2602.17337v2 Announce Type: replace Abstract: In this work we present Polaffini, a robust and versatile framework for anatomically grounded registration. Medical image registration is dominated

safetyarxiv-cs-cv
9 Jun 2026
Model Releases

POTATR: A Lightweight Image-to-Graph Model for Page-Level Table Extraction

DGX agent

arXiv:2606.09788v1 Announce Type: new Abstract: Large-scale document processing requires contextually aware table extraction (TE) that is both accurate and efficient. Yet current approaches require bi

model-releasesarxiv-cs-cv
9 Jun 2026
Safety

Prisma-World: Camera-Controllable Multi-Agent Video World Model

DGX agent

arXiv:2606.09507v1 Announce Type: new Abstract: Video world models have made rapid progress in generating controllable visual experiences, but most of them still simulate the world from a single obser

safetyarxiv-cs-cv
9 Jun 2026
Model Releases

Programmable Silicon Retina on Pixel Processor Array

DGX agent

arXiv:2606.08370v1 Announce Type: cross Abstract: Standard dynamic vision sensors approximate retinal processing by detecting temporal contrast changes, offering high speed and high dynamic range. In

model-releasesarxiv-cs-cv
9 Jun 2026
Safety

Property-Informed Diffusion-Based Text-to-Microstructure Generation

DGX agent

arXiv:2606.08150v1 Announce Type: new Abstract: Designing 3D metamaterial microstructures that meet the intended functions remains a major challenge, as it typically requires domain expertise, iterati

safetyarxiv-cs-cv
9 Jun 2026
Safety

PRPO: Perception-Reinforced Policy Optimization via Token-Level Dynamic Advantage Reshaping

DGX agent

arXiv:2606.08708v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has become an effective paradigm for improving the reasoning capability of Large Vision-Language M

safetyarxiv-cs-cv
9 Jun 2026
Tutorials

Quantifying Noise of Dynamic Vision Sensor

DGX agent

arXiv:2404.01948v3 Announce Type: replace Abstract: Dynamic visual sensors (DVS) are characterized by a large amount of background activity (BA) noise, which it is mixed with the original (cleaned) se

tutorialsarxiv-cs-cv
9 Jun 2026
Research

QuoVLA: Quotient Space for Vision-Language-Action Models

DGX agent

arXiv:2605.24890v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) models commonly adapt pretrained Vision-Language Models (VLMs) to robot control by mapping visual observations and lang

researcharxiv-cs-cv
9 Jun 2026
Model Releases

RAD: A Dataset and Benchmark for Real-Life Anomaly Detection with Robotic Observations

DGX agent

arXiv:2410.00713v4 Announce Type: replace Abstract: Anomaly detection is a core capability for robotic perception and industrial inspection, yet most existing benchmarks are collected under controlled

model-releasesarxiv-cs-cv
9 Jun 2026
Research

REACT 2026: The Fourth Multiple Appropriate Facial Reaction Generation Challenge: Personalised MAFRG and Appropriate EEG Reaction Prediction

DGX agent

arXiv:2606.07935v1 Announce Type: new Abstract: In dyadic interactions, various human facial reactions could be appropriate for responding to each human speaker behaviour. Following the successful org

researcharxiv-cs-cv
9 Jun 2026
Model Releases

Readable Yet Unpredictable: Rotated-Outcome Prediction in Vision-Language Models

DGX agent

arXiv:2606.07641v1 Announce Type: new Abstract: Can vision-language models predict what a 180{eg} rotation would reveal from the original image alone? We study this ability through Rotated-Outcome Pre

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Real-Time Industrial Defect Detection on Edge Hardware Using Fine-Tuned YOLOv8: A Systematic Benchmark on the NEU Surface Defect Database and MVTec AD with Automotive & Battery Manufacturing Extensions

DGX agent

arXiv:2606.07659v1 Announce Type: new Abstract: Automated surface defect detection is critical for ensuring rigorous quality control in high-speed manufacturing environments. While deep learning model

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Reason Twice: Segmentation via Candidate Discovery and Comparative Reasoning

DGX agent

arXiv:2606.09303v1 Announce Type: new Abstract: The rapid development of pretrained foundation models has enabled more general image segmentation. Multimodal large language models (MLLMs) have been wi

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

REFINE: Super-efficient 3D Gaussian Splatting Pruning via Rendering-Free Primitive Importance

DGX agent

arXiv:2606.09074v1 Announce Type: new Abstract: Existing pruning methods for 3D Gaussian splatting (3DGS) suffer from either severe quality degradation or prohibitive computational overhead. In this p

model-releasesarxiv-cs-cv
9 Jun 2026
Safety

Region-Wise Correspondence Prediction between Manga Line Art Images

DGX agent

arXiv:2509.09501v4 Announce Type: replace Abstract: Understanding region-wise correspondences between manga line art images is fundamental for high-level manga processing, supporting downstream tasks

safetyarxiv-cs-cv
9 Jun 2026
Safety

Reinforcing Temporal Answer Grounding in Instructional Video via Candidate-Aware Causal Reasoning

DGX agent

arXiv:2606.08436v1 Announce Type: new Abstract: The task of temporal answer grounding in instructional video (TAGV), which aims to locate precise video segments that respond to natural language querie

safetyarxiv-cs-cv
9 Jun 2026
Local Ai

Relational Epipolar Graphs for Robust Relative Camera Pose Estimation

DGX agent

arXiv:2604.04554v2 Announce Type: replace Abstract: A key component of Visual Simultaneous Localization and Mapping (VSLAM) is estimating relative camera poses using matched keypoints. Accurate estima

local-aiarxiv-cs-cv
9 Jun 2026
Model Releases

Remember with Confidence: Uncertainty Quantification for Spatio-temporal Memory with Probabilistic Guarantees

DGX agent

arXiv:2606.08277v1 Announce Type: new Abstract: Long-horizon robot operation requires spatio-temporal memory to record the environment state and recall it for downstream reasoning. Scene graphs and re

model-releasesarxiv-cs-cv
9 Jun 2026
Research

Rethinking 3D Shape Generation: Diffusion over Superquadrics

DGX agent

arXiv:2606.08957v1 Announce Type: new Abstract: Diffusion models have advanced 3D shape generation, yet most methods still denoise in high-cardinality spaces (e.g., voxel/SDF grids, meshes, or point c

researcharxiv-cs-cv
9 Jun 2026
Safety

Revisiting Articulated Parts Perception in Robot Manipulation

DGX agent

arXiv:2606.08103v1 Announce Type: cross Abstract: We are surrounded by various objects with movable, articulated parts, e.g., box, handle, door. An accurate and generalizable perception of articulated

safetyarxiv-cs-cv
9 Jun 2026
Tutorials

RGB-S: Image-Aligned Tactile Saliency for Robust Dexterous Manipulation

DGX agent

arXiv:2606.08765v1 Announce Type: cross Abstract: Effective visuo-tactile integration is critical for robotic dexterous manipulation, especially when visual observations are unreliable or occluded. Ho

tutorialsarxiv-cs-cv
9 Jun 2026
Applications

RT-SDGOD: Real-Time Single-Domain Generalized Object Detection

DGX agent

arXiv:2606.09367v1 Announce Type: new Abstract: In real-world deployment under strict real-time constraints, weather and imaging variations induce significant distribution shifts, severely degrading d

applicationsarxiv-cs-cv
9 Jun 2026
Model Releases

Scaling by Diversified Experience for Vision-Language-Action Models

DGX agent

arXiv:2606.09009v1 Announce Type: new Abstract: Vision-Language-Action models face significant challenges in real-world deployment due to the entanglement of high-level reasoning with low-level contro

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

SciFlow-Bench: Evaluating Structure-Aware Scientific Diagram Generation via Inverse Parsing

DGX agent

arXiv:2602.09809v2 Announce Type: replace Abstract: Scientific diagrams convey explicit structural information, yet modern text-to-image models often produce visually plausible but structurally incorr

model-releasesarxiv-cs-cv
9 Jun 2026
Research

SDTrack: A Baseline for Event-based Tracking via Spiking Neural Networks

DGX agent

arXiv:2503.08703v4 Announce Type: replace-cross Abstract: Event cameras provide superior temporal resolution, dynamic range, energy efficiency, and pixel bandwidth. Spiking Neural Networks (SNNs) natu

researcharxiv-cs-cv
9 Jun 2026
Research

Securing Self-supervised Data Curation for Foundation Models Robustness

DGX agent

arXiv:2606.09511v1 Announce Type: new Abstract: Self-supervised data curation provides a pathway to scaling and improving the generalization capabilities of machine learning models. By leveraging self

researcharxiv-cs-cv
9 Jun 2026
Safety

See More, Match Better: Multi-Source Feature Fusion for Two-View Correspondence Learning

DGX agent

arXiv:2606.09262v1 Announce Type: new Abstract: Two-view correspondence learning aims to distinguish true correspondences (inliers) from false ones (outliers) in image pairs by leveraging their underl

safetyarxiv-cs-cv
9 Jun 2026
Model Releases

SegmentAnyTreeV2: Scaling Transformer-Based Tree Instance Segmentation Across Sensors, Platforms, and Forests

DGX agent

arXiv:2606.08206v1 Announce Type: new Abstract: We present SegmentAnyTreeV2, a sensor- and platform-agnostic framework for semantic and instance segmentation of forest point clouds. The model combines

model-releasesarxiv-cs-cv
9 Jun 2026
Research

Segmentation-Assisted Brain MRI Synthesis with Cross-Image Multi-Contrast Feature Memory Bank Retrieval Augmentation

DGX agent

arXiv:2606.08421v1 Announce Type: new Abstract: Multi-contrast brain MRI provide complementary soft-tissue characteristics that aid in the screening and diagnosis of diseases. However, limited scannin

researcharxiv-cs-cv
9 Jun 2026
Research

Self-supervised Learning Matters: A Simple Ensemble Solution for Micro-Gesture Recognition

DGX agent

arXiv:2606.09261v1 Announce Type: new Abstract: In this paper, we present XInsight Lab's solution to the micro-gesture classification track of the 4th MiGA Challenge at IJCAI 2026, in which our soluti

researcharxiv-cs-cv
9 Jun 2026
Safety

Self-Supervised Learning with a Multi-Task Latent Space Objective

DGX agent

arXiv:2602.05845v2 Announce Type: replace Abstract: We propose a multi-task formulation of self-predictive Siamese SSL in which each spatial transformation defines a distinct latent-space alignment ta

safetyarxiv-cs-cv
9 Jun 2026
Safety

SemDINO: A DINOv3-Driven Network for Cross-Temporal Semantic Alignment in Change Detection

DGX agent

arXiv:2606.09772v1 Announce Type: new Abstract: Semantic change detection (SCD) aims to simultaneously locate land-cover changes and identify semantic categories before and after transition. However,

safetyarxiv-cs-cv
9 Jun 2026
Model Releases

Semi-supervised Source Detection in Astronomical Images: New Benchmark and Strong Baseline

DGX agent

arXiv:2606.09219v1 Announce Type: new Abstract: Source detection in modern observational astronomy is a cornerstone for localizing and identifying stellar sources accurately. It is crucial for studies

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Shift-Dependent Asymmetry: Orthogonal Inverse Low-Rank Adaptation for Federated Medical Segmentation

DGX agent

arXiv:2606.08687v1 Announce Type: new Abstract: Low-Rank Adaptation (LoRA) enables efficient federated fine-tuning of segmentation foundation models for medical imaging. However, most federated LoRA m

model-releasesarxiv-cs-cv
9 Jun 2026
Applications

Simultaneous hyperkinetic movement disorders phenotyping: a cross-cohort pediatric transfer study using routine videos, markerless pose estimation and a tabular foundation model

DGX agent

arXiv:2606.07674v1 Announce Type: new Abstract: Objective: To develop and externally test a video-based framework for simultaneous detection of hyperkinetic MDs phenomenologies: dystonia, tremor, myoc

applicationsarxiv-cs-cv
9 Jun 2026
Safety

SMI: Efficient Self-Supervised Learning via Mutual-Information-Inspired Dependency Optimization

DGX agent

arXiv:2606.08332v1 Announce Type: new Abstract: Self-supervised learning (SSL) has achieved remarkable representation learning performance, but many existing methods rely on large batch sizes, memory

safetyarxiv-cs-cv
9 Jun 2026
Hardware

SoccerNet 2026 Player-Centric Ball-Action Spotting:Retraining and Post-Processing Extensions to the FOOTPASS Baselines

DGX agent

arXiv:2606.09679v1 Announce Type: new Abstract: We describe our system for the SoccerNet 2026 Player-Centric Ball-Action Spotting Challenge, which requires predicting who performs which action and whe

hardwarearxiv-cs-cv
9 Jun 2026
Research

SOMA: From Surface Observations to Muscle Anatomy

DGX agent

arXiv:2606.09246v1 Announce Type: new Abstract: With the growing demand for realistic virtual humans, parametric body models have become a cornerstone of modern medicine, sports, and entertainment app

researcharxiv-cs-cv
9 Jun 2026
← Previous
1…112113114115116…263
Next →