AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,515 results
Safety

Eddeep: a deep-learning framework for fast eddy-current distortion correction in diffusion MRI

DGX agent

arXiv:2607.26292v1 Announce Type: new Abstract: Diffusion MRI (dMRI) relies on diffusion-weighted echo-planar imaging, which is highly susceptible to eddy-current-induced geometric distortions. These

safetyarxiv-cs-cv
30 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

EgoSafe: A First-Person Mobile-Captured Benchmark for Visual Safety Understanding

DGX agent

arXiv:2607.26518v1 Announce Type: new Abstract: Reliable visual safety understanding in real-world scenarios demands more than just object recognition; it requires causal reasoning under epistemic unc

model-releasesarxiv-cs-cv
30 Jul 2026
Safety

Explainable and Resource-Efficient Spatial Reasoning in Multimodal LLMs for Decision-Critical Applications

DGX agent

arXiv:2607.27145v1 Announce Type: new Abstract: As Multimodal Large Language Models (MLLMs) are increasingly deployed in decision-critical pipelines such as robotics, embodied AI, and safety monitorin

safetyarxiv-cs-cv
30 Jul 2026
Local Ai

FakeIDet3-DB: Refining Digital Attacks and Patch Extraction for Secure ID Benchmarking

DGX agent

arXiv:2607.26641v1 Announce Type: new Abstract: Identity document (ID) authentication relies on the structural integrity of complex, high-frequency security patterns. However, advanced Generative AI m

local-aiarxiv-cs-cv
30 Jul 2026
Model Releases

FAS-R1: A Unified Multi-Task MLLM for Reasoning Face Anti-Spoofing

DGX agent

arXiv:2607.26432v1 Announce Type: new Abstract: Face anti-spoofing (FAS) is increasingly expected to provide not only bona fide/spoof decisions, but also attack semantics and image-grounded evidence f

model-releasesarxiv-cs-cv
30 Jul 2026
Safety

FPSGen: Flexible Point Cloud Scene Generation with BEV-Supported Transport Flows

DGX agent

arXiv:2607.26645v1 Announce Type: new Abstract: Existing point-based generative methods for outdoor scenes primarily focus on LiDAR-conditioned completion. During training, noisy point clouds are cons

safetyarxiv-cs-cv
30 Jul 2026
Research

FreeShadow: Training-Free Shadow Removal via Illumination Transfer and Selective Content Preservation in Diffusion Models

DGX agent

arXiv:2607.26715v1 Announce Type: new Abstract: Existing supervised and unsupervised shadow removal methods often suffer from limited generalization due to the insufficient diversity of available trai

researcharxiv-cs-cv
30 Jul 2026
Research

FreqForcing: Autoregressive Long Video Generation via Spectral Self-Anchoring

DGX agent

arXiv:2607.27110v1 Announce Type: new Abstract: Autoregressive video diffusion models enable real-time streaming video generation. However, errors introduced during self-rollout accumulate over long h

researcharxiv-cs-cv
30 Jul 2026
Local Ai

From Keypoints to Predictive Distributions: Post-Hoc Uncertainty for YOLO-Pose Models

DGX agent

arXiv:2607.26921v1 Announce Type: new Abstract: YOLO-Pose models provide efficient keypoint localization, but do not quantify the associated spatial uncertainty. We introduce a lightweight post-hoc pr

local-aiarxiv-cs-cv
30 Jul 2026
Research

From Spatial Semantics to Temporal Context: Leveraging Gaze Trajectory for Weakly Supervised Medical Image Segmentation

DGX agent

arXiv:2607.26542v1 Announce Type: new Abstract: Medical image segmentation heavily depends on labor-intensive and time-consuming pixel-level annotations. Eye tracking offers a cost-effective solution

researcharxiv-cs-cv
30 Jul 2026
Local Ai

From Uncertainty to Determinism: Coarse-to-Fine Visual Floorplan Localization without Ray Matching

DGX agent

arXiv:2607.26817v1 Announce Type: cross Abstract: Visual Floorplan Localization (FLoc) has emerged as a promising solution for indoor localization by matching egocentric images against minimalist stru

local-aiarxiv-cs-cv
30 Jul 2026
Hardware

Genie Sim PanoWorld: An Infinite Indoor 3D World Generation Pipeline via Panoramic Scene Modeling and Simulation

DGX agent

arXiv:2607.26646v1 Announce Type: new Abstract: We address the problem of reconstructing a high-fidelity, freely navigable 3D scene from a single 360^irc panorama, without per-scene optimization or mu

hardwarearxiv-cs-cv
30 Jul 2026
Research

HERMES: A Hybrid Ensemble for Head-and-Neck Tumor Segmentation, TN Staging, and Recurrence-Free Survival on PET/CT

DGX agent

arXiv:2607.26498v1 Announce Type: new Abstract: We present HERMES (Hybrid Ensemble for Radiotherapy-target segmentation, Malignancy staging, and Event-free Survival), a single containerized algorithm

researcharxiv-cs-cv
30 Jul 2026
Agents

HeteroPROPMT: A Real-time and Privacy-Preserving Heterogeneous Collaborative Perception Framework

DGX agent

arXiv:2607.26283v1 Announce Type: new Abstract: Collaborative Perception (CP) improves autonomous systems' awareness of their surroundings by sharing sensor data, intermediate features, and detection

agentsarxiv-cs-cv
30 Jul 2026
Model Releases

HumanCLAW: Can Vision-Language Models Act Through a Body?

DGX agent

arXiv:2607.27180v1 Announce Type: new Abstract: Evaluating whether a vision-language model (VLM) can act through a physical body is challenging. The outcome of an action couples the VLM's decision wit

model-releasesarxiv-cs-cv
30 Jul 2026
Model Releases

ICDAR 2026 Competition on Information Extraction from Atomic Layer Deposition/Etching (ALD/E) Scientific Figures

DGX agent

arXiv:2607.26848v1 Announce Type: new Abstract: Scientific figure comprehension and reasoning using multimodal AI requires integrating visual perception with domain-specific reasoning to extract meani

model-releasesarxiv-cs-cv
30 Jul 2026
Research

Improving Knowledge Distillation Under Unknown Covariate Shift Through Confidence-Guided Data Augmentation

DGX agent

arXiv:2506.02294v3 Announce Type: replace Abstract: Large foundation models trained on extensive datasets demonstrate strong zero-shot capabilities in various domains. Knowledge distillation has becom

researcharxiv-cs-cv
30 Jul 2026
Research

InkShield: Writing Style Protection Against Unauthorized Handwriting Mimicry

DGX agent

arXiv:2607.26976v1 Announce Type: cross Abstract: Recent handwritten text generators can reproduce a writer's style from publicly available references, posing risks of document forgery and identity mi

researcharxiv-cs-cv
30 Jul 2026
Research

Interpretable Image-Level Acne Severity Grading via EfficientNet-B0 Transfer Learning and Grad-CAM

DGX agent

arXiv:2607.26461v1 Announce Type: new Abstract: Acne vulgaris affects most adolescents and many adults. Accurate severity grading guides treatment, monitoring, and clinical trial endpoints, but manual

researcharxiv-cs-cv
30 Jul 2026
Model Releases

JEPADepth: Masked Predictive Representation Learning for Self-Supervised Monocular Depth Estimation

DGX agent

arXiv:2607.26600v1 Announce Type: new Abstract: Self-supervised monocular depth estimation typically relies on photometric reconstruction losses that couple depth, pose, and appearance assumptions. In

model-releasesarxiv-cs-cv
30 Jul 2026
Research

Kinetic Mining in Context: Few-Shot Action Synthesis via Text-to-Motion Distillation

DGX agent

arXiv:2512.11654v3 Announce Type: replace Abstract: The acquisition cost for large, annotated motion datasets remains a critical bottleneck for skeletal-based Human Activity Recognition (HAR). Althoug

researcharxiv-cs-cv
30 Jul 2026
Research

Knowledge-guided Disentanglement with Atomic Actions for Action Recognition

DGX agent

arXiv:2607.26097v1 Announce Type: new Abstract: Action recognition in complex scenes often involves multiple concurrent fine-grained actions, making it challenging to model internal action structures.

researcharxiv-cs-cv
30 Jul 2026
Safety

Lag-aware cross-hand alignment for dual-hand action segmentation

DGX agent

arXiv:2607.26215v1 Announce Type: new Abstract: Dual-hand action segmentation commonly fuses left- and right-hand representations at identical temporal indices, although coordinated hand transitions m

safetyarxiv-cs-cv
30 Jul 2026
Model Releases

Level, Sharpness, and Corpus: Why Zero-Shot OOD Detector Rankings Do Not Transfer

DGX agent

arXiv:2607.26582v1 Announce Type: new Abstract: Selecting a zero-shot out-of-distribution (OOD) detector for a new deployment is typically based on benchmark rankings, implicitly assuming that the hig

model-releasesarxiv-cs-cv
30 Jul 2026
Safety

LiDARDraft: Generating LiDAR Point Cloud from Versatile Inputs

DGX agent

arXiv:2512.20105v2 Announce Type: replace Abstract: Generating realistic and diverse LiDAR point clouds is crucial for autonomous driving simulation. Although previous methods achieve LiDAR point clou

safetyarxiv-cs-cv
30 Jul 2026
Hardware

Lightweight Image Classification of Raptor Species for Edge Devices: Rare-Species Dataset Expansion via Video Frame Extraction, Knowledge Distillation, and TensorRT Deployment

DGX agent

arXiv:2607.26238v1 Announce Type: new Abstract: We investigate lightweight raptor-species classification for real-time edge deployment in wind-turbine collision mitigation. Using DINOv2-L (304M parame

hardwarearxiv-cs-cv
30 Jul 2026
Research

LISA-3D: Lifting Language-Image Segmentation to 3D via Multi-View Consistency

DGX agent

arXiv:2512.01008v2 Announce Type: replace Abstract: Text-driven 3D reconstruction requires masks that understand free-form instructions and remain stable under viewpoint changes. We present LISA-3D, a

researcharxiv-cs-cv
30 Jul 2026
Research

LLM-Grounded Dynamic Task Planning with Hierarchical Temporal Logic for Human-Aware Multi-Robot Handover

DGX agent

arXiv:2602.09472v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) enable non-experts to specify open-world multi-robot tasks, but the generated plans are often kinematically infea

researcharxiv-cs-cv
30 Jul 2026
Safety

Long-Tailed 3D Point Cloud Dataset Distillation

DGX agent

arXiv:2607.26763v1 Announce Type: new Abstract: Dataset distillation compresses large-scale datasets into compact synthetic sets while preserving their training utility, enabling efficient 3D point cl

safetyarxiv-cs-cv
30 Jul 2026
Safety

Lottery Tickets Are Not Deployment Tickets

DGX agent

arXiv:2607.27031v1 Announce Type: cross Abstract: Reports on how sparsification, compression, and lottery tickets change model behavior have been mixed in the prior literature, with beneficial effects

safetyarxiv-cs-cv
30 Jul 2026
Research

LumaGuide: Distribution Shaping for Training-Free HDR Generation in Diffusion Models

DGX agent

arXiv:2607.26237v1 Announce Type: new Abstract: Pretrained diffusion models generate realistic images but are constrained by the statistical biases of their training data, limiting their ability to pr

researcharxiv-cs-cv
30 Jul 2026
Local Ai

MedARC: Training-Free Adaptive Redundancy Compression of Visual Tokens for 3D Medical Vision-Language Models

DGX agent

arXiv:2607.26554v1 Announce Type: new Abstract: Integrating 3D medical images with vision-language models (VLMs) holds substantial promise for computer-aided diagnosis. However, volumetric images gene

local-aiarxiv-cs-cv
30 Jul 2026
Agents

Mitigating Compounding Error via Video Representation Regularization

DGX agent

arXiv:2607.27036v1 Announce Type: new Abstract: Video diffusion-based world models enable long autoregressive video generation for robotics, autonomous driving and simulation tasks, yet sliding-window

agentsarxiv-cs-cv
30 Jul 2026
Local Ai

MoSAIC: Aligned Intervention Supervision for Part-Local Motion Style Transfer

DGX agent

arXiv:2607.26304v1 Announce Type: new Abstract: Editing character motion often requires transferring a gesture or gait from one or more reference motions while preserving the source action, timing, ro

local-aiarxiv-cs-cv
30 Jul 2026
Research

Multimodal fusion of visual and morphometric features for avian bone classification

DGX agent

arXiv:2607.26743v1 Announce Type: new Abstract: Artificial intelligence has shown considerable potential for archaeological applications, yet its use in zooarchaeology remains limited, particularly fo

researcharxiv-cs-cv
30 Jul 2026
Research

Neural Network Assisted Lifting Steps For Improved Fully Scalable Lossy Image Compression in JPEG 2000

DGX agent

arXiv:2403.01647v2 Announce Type: replace Abstract: This work proposes to augment the lifting steps of the conventional wavelet transform with additional neural network assisted lifting steps. These a

researcharxiv-cs-cv
30 Jul 2026
Applications

Neural Radiance Fields for the Real World: A Survey

DGX agent

arXiv:2501.13104v3 Announce Type: replace Abstract: Neural Radiance Fields (NeRFs) have remodeled 3D scene representation since release. NeRFs can effectively reconstruct complex 3D scenes from 2D ima

applicationsarxiv-cs-cv
30 Jul 2026
Local Ai

Object Detection for Autonomous Driving in Chinese Rural Scenes: An Experimental Study on Real-Synthetic Data Mixing and Model Evaluation

DGX agent

arXiv:2607.27058v1 Announce Type: new Abstract: Currently, autonomous driving object detection models face significant data scarcity and generalization challenges when navigating complex Chinese rural

local-aiarxiv-cs-cv
30 Jul 2026
Model Releases

OmniAD: Detect and Understand Industrial Anomaly via Multimodal Reasoning

DGX agent

arXiv:2505.22039v2 Announce Type: replace Abstract: While anomaly detection has made significant progress, generating detailed analyses that incorporate industrial knowledge remains a challenge. To ad

model-releasesarxiv-cs-cv
30 Jul 2026
Safety

One-Frame Calibration with Siamese Network in Facial Action Unit Recognition

DGX agent

arXiv:2409.00240v2 Announce Type: replace Abstract: Automatic facial action unit (AU) recognition is used widely in facial expression analysis. Most existing AU recognition systems aim for cross-parti

safetyarxiv-cs-cv
30 Jul 2026
Model Releases

Online Handwriting Trajectory Reconstruction from Kinematic Sensors using Temporal Convolutional Network

DGX agent

arXiv:2607.26733v1 Announce Type: new Abstract: Handwriting with digital pens is a common way to facilitate human-computer interaction through the use of Online Handwriting (OH) trajectory reconstruct

model-releasesarxiv-cs-cv
30 Jul 2026
Research

Particle-Filtering-based Latent Diffusion for Inverse Problems

DGX agent

arXiv:2408.13868v2 Announce Type: replace Abstract: Current strategies for solving image-based inverse problems apply latent diffusion models to perform posterior sampling.However, almost all approach

researcharxiv-cs-cv
30 Jul 2026
Model Releases

PatchDenoiser: Parameter-efficient multi-scale patch learning and fusion denoiser for Low-dose CT imaging

DGX agent

arXiv:2602.21987v3 Announce Type: replace Abstract: Low-dose CT images are essential for reducing radiation exposure in cancer screening, pediatric imaging, and longitudinal monitoring protocols, but

model-releasesarxiv-cs-cv
30 Jul 2026
Safety

Physically Real-time Infrared Attack against Optical Flow Estimation Networks

DGX agent

arXiv:2607.26651v1 Announce Type: new Abstract: With the promising performance of deep neural networks on image-based tasks, different real-world applications such as autonomous driving and motion det

safetyarxiv-cs-cv
30 Jul 2026
Research

Prior Directions: Why GUI Grounding Gets Locked in the Past

DGX agent

arXiv:2607.26913v1 Announce Type: new Abstract: Vision-language models often use descriptions of earlier visual states to make decisions about the current scene. When the scene changes, stale language

researcharxiv-cs-cv
30 Jul 2026
Research

PRISM-Net: Patient-specific reference-guided inter-breast symmetry matching for three-class breast DCE-MRI classification

DGX agent

arXiv:2607.26799v1 Announce Type: new Abstract: Breast DCE-MRI AI is increasingly being explored for breast-level classification of no-lesion, benign, and malignant findings, beyond conventional lesio

researcharxiv-cs-cv
30 Jul 2026
Model Releases

Progressive Multimodal Alignment for Continual Instruction Tuning

DGX agent

arXiv:2607.26947v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) rely on a projector to align visual representations with the language embedding space, making it central to cro

model-releasesarxiv-cs-cv
30 Jul 2026
Safety

R-SLPR: Region-based Small-to-Large Point-cloud Registration with Contrastive Learning

DGX agent

arXiv:2607.26583v1 Announce Type: new Abstract: Point-cloud (PC) registration is fundamental to three-dimensional (3D) perception in robotic systems. However, classic registration algorithms falter wh

safetyarxiv-cs-cv
30 Jul 2026
← Previous
1…3233343536…261
Next →