AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,618 results
Research

ZeroGVC: Zero-Shot Generative Video Compression with Autoregressive Diffusion Priors

DGX agent

arXiv:2606.22371v1 Announce Type: cross Abstract: Recent generative video compression methods leverage powerful generative priors to achieve perceptually pleasing reconstructions. However, most existi

researcharxiv-cs-cv
23 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

3D-CBM: A Framework for Concept-Based Interpretability in Generative 3D Modeling

DGX agent

arXiv:2606.11446v1 Announce Type: new Abstract: This research introduces a framework for incorporating Concept Bottleneck Models (CBMs) into 3D generative architectures to address the inherent 'semant

safetyarxiv-cs-cv
11 Jun 2026
Model Releases

4DP-QA: Scalable QA for 4D Perception in Vision Language Models

DGX agent

arXiv:2606.11568v1 Announce Type: new Abstract: Despite recent advances, Vision Language Models (VLMs) still struggle to grasp the dynamics of the world. We note that the ability to reason about a 4D

model-releasesarxiv-cs-cv
11 Jun 2026
Model Releases

A Comprehensive Ecosystem for Open-Domain Customized Video Generation

DGX agent

arXiv:2606.11783v1 Announce Type: new Abstract: Recent progress in video generation has shown impressive visual synthesis capabilities. However, open-domain customized video generation remains limited

model-releasesarxiv-cs-cv
11 Jun 2026
Hardware

A Scalable PyTorch Abstraction for Multi-GPU Gaussian Splatting

DGX agent

arXiv:2606.11390v1 Announce Type: new Abstract: Gaussian splatting methods have become increasingly popular for neural reconstruction of the real world. However, they are often limited in scale and re

hardwarearxiv-cs-cv
11 Jun 2026
Local Ai

A Turbo-Inference Strategy for Object Detection and Instance Segmentation

DGX agent

arXiv:2606.12371v1 Announce Type: new Abstract: Object detection and instance segmentation tasks are closely related. Existing top-down instance segmentation methods usually follow a detect-then-segme

local-aiarxiv-cs-cv
11 Jun 2026
Research

A2SG:Adaptive and Asymmetric Surrogate Gradients for Training Deep Spiking Neural Networks

DGX agent

arXiv:2606.11236v1 Announce Type: cross Abstract: Training deep spiking neural networks (SNNs) remains challenging due to sharp loss landscapes and temporal inconsistency caused by surrogate gradients

researcharxiv-cs-cv
11 Jun 2026
Safety

Adapting Vision-Language Models from Iconic to Inclusive for Multi-Label Recognition Without Labels

DGX agent

arXiv:2606.11626v1 Announce Type: new Abstract: Understanding multi-label images remains a challenging task in computer vision. With the rapid progress of vision-language multimodal learning, vision-l

safetyarxiv-cs-cv
11 Jun 2026
Safety

Adv-TGD: Adversarial Text-Guided Diffusion for Face Recognition Impersonation Attacks

DGX agent

arXiv:2606.11615v1 Announce Type: new Abstract: The widespread adoption of face recognition (FR) technologies raises serious privacy concerns, as facial data can be exploited without consent. To addre

safetyarxiv-cs-cv
11 Jun 2026
Safety

AerialClaw: An Open-Source Framework for LLM-Driven Autonomous Aerial Agents

DGX agent

arXiv:2606.12142v1 Announce Type: cross Abstract: Unmanned aerial vehicles (UAVs) are increasingly used in inspection, search and rescue, environmental monitoring, and emergency response. However, mos

safetyarxiv-cs-cv
11 Jun 2026
Tutorials

AGE-MIL: Anchor-Guided Evidence Learning for Patient-Level Prediction

DGX agent

arXiv:2606.12126v1 Announce Type: new Abstract: Existing computational pathology methods predominantly operate within whole-slide image (WSI)-level multiple instance learning (MIL) paradigms, while pa

tutorialsarxiv-cs-cv
11 Jun 2026
Model Releases

An Electric Potential-Augmented Benchmark Dataset for Physics-Guided Image Reconstruction of Electrical Capacitance Tomography

DGX agent

arXiv:2606.12226v1 Announce Type: new Abstract: While deep learning has significantly advanced image reconstruction of Electrical Capacitance Tomography (ECT), most data-driven methods map directly be

model-releasesarxiv-cs-cv
11 Jun 2026
Research

Anatomically Conditioned Recurrent Refinement for Topology-Aware Circle of Willis Segmentation

DGX agent

arXiv:2606.12319v1 Announce Type: new Abstract: Segmenting the Circle of Willis (CoW) from Magnetic Resonance Angiography (MRA) is challenging due to complex topology and thin vascular structures that

researcharxiv-cs-cv
11 Jun 2026
Safety

Auditing Demographic Bias in Facial Landmark Detection for Fair Human-Robot Interaction

DGX agent

arXiv:2604.06961v2 Announce Type: replace Abstract: Fairness in human-robot interaction critically depends on the reliability of the perceptual models that enable robots to interpret human behavior. W

safetyarxiv-cs-cv
11 Jun 2026
Research

Battery detection of XRay images using transfer learning

DGX agent

arXiv:2606.11779v1 Announce Type: new Abstract: The need for detecting and sorting batteries is drastically increasing for many applications. This study proves the potential of transfer learning in pr

researcharxiv-cs-cv
11 Jun 2026
Model Releases

Benchmarking Cross-Domain Audio-Visual Deception Detection

DGX agent

arXiv:2405.06995v4 Announce Type: replace-cross Abstract: Automated deception detection is crucial for assisting humans in accurately assessing truthfulness and identifying deceptive behavior. Convent

model-releasesarxiv-cs-cv
11 Jun 2026
Research

Beyond Dark Knowledge: Mixup-Based Distillation for Reliable Predictions

DGX agent

arXiv:2606.12171v1 Announce Type: new Abstract: Knowledge Distillation (KD) and mixup have proven effective at inducing smoothness in class boundaries; KD captures inherent class relationships in prob

researcharxiv-cs-cv
11 Jun 2026
Safety

Bridging Day and Night: Unsupervised Cross-Domain Re-Identification with Synergistic Prompt and Prototype Learning

DGX agent

arXiv:2606.12258v1 Announce Type: new Abstract: Cross-domain day-night re-identification (ReID) is fundamentally challenged by the substantial visual appearance discrepancies between daytime and night

safetyarxiv-cs-cv
11 Jun 2026
Applications

Bridging the Modality Gap in Forensic Image Retrieval

DGX agent

arXiv:2606.12294v1 Announce Type: new Abstract: Automated image retrieval plays an increasingly critical role in modern forensic analysis, supporting investigative workflows that rely on efficient com

applicationsarxiv-cs-cv
11 Jun 2026
Tutorials

Causal Clothes-Invariant Feature Learning for Cloth-Changing Person Re-ID

DGX agent

arXiv:2305.06145v2 Announce Type: replace Abstract: In cloth-changing person re-identification (CCReID), it is critical to learn clothes-invariant feature, which can provide discriminative ID features

tutorialsarxiv-cs-cv
11 Jun 2026
Research

CellNet -- Localizing Cells using Sparse and Noisy Point Annotations

DGX agent

arXiv:2606.12286v1 Announce Type: new Abstract: Counting living cells is an important step in many biological research workflows. Our collaborators at the Wellcome Sanger Institute study vital genes i

researcharxiv-cs-cv
11 Jun 2026
Model Releases

CFCamo: A Counterfactual Detect-or-Abstain Framework for Camouflaged Object Detection

DGX agent

arXiv:2606.11231v1 Announce Type: new Abstract: Vision-language reinforcement learning has recently shown strong target-present localization for camouflaged object detection (COD). Yet localization is

model-releasesarxiv-cs-cv
11 Jun 2026
Applications

Contactless 3D Human Body Measurement Using Depth Cameras for Smart Health Monitoring

DGX agent

arXiv:2606.11578v1 Announce Type: new Abstract: Contactless body measurement technologies are becoming increasingly significant for smart health monitoring, digital health applications, and remote pat

applicationsarxiv-cs-cv
11 Jun 2026
Research

Continual Learning with Support Boundary Experience Blending

DGX agent

arXiv:2507.23534v3 Announce Type: replace-cross Abstract: Continual learning (CL) seeks to mitigate catastrophic forgetting when models are trained with sequential tasks. A common approach, experience

researcharxiv-cs-cv
11 Jun 2026
Safety

Corpus Augmentation for Sign Language Translation via LLM-Guided Video Stitching

DGX agent

arXiv:2606.11925v1 Announce Type: new Abstract: Sign language translation (SLT) converts sign language video into spoken language text and holds significant promise for improving accessibility and ena

safetyarxiv-cs-cv
11 Jun 2026
Research

CountZES: Counting via Zero-Shot Exemplar Selection

DGX agent

arXiv:2512.16415v3 Announce Type: replace Abstract: Object counting in complex scenes is particularly challenging in the zero-shot (ZS) setting, where instances of unseen categories are counted using

researcharxiv-cs-cv
11 Jun 2026
Model Releases

CoVR-R:Reason-Aware Composed Video Retrieval

DGX agent

arXiv:2603.20190v2 Announce Type: replace Abstract: Composed Video Retrieval (CoVR) aims to find a target video given a reference video and a textual modification. Prior work assumes the modification

model-releasesarxiv-cs-cv
11 Jun 2026
Research

Cross-Domain Multi-Person Human Activity Recognition via Near-Field Wi-Fi Sensing

DGX agent

arXiv:2510.17816v2 Announce Type: replace-cross Abstract: Wi-Fi-based human activity recognition (HAR) provides substantial convenience and has emerged as a thriving research field, yet the coarse spa

researcharxiv-cs-cv
11 Jun 2026
Model Releases

Cross-Modal Benchmarking for Robotic Perception in Natural Environments

DGX agent

arXiv:2606.11563v1 Announce Type: new Abstract: Natural environments present a complex challenge to robotics perception systems. Current models, particularly vision foundation models, are largely trai

model-releasesarxiv-cs-cv
11 Jun 2026
Applications

DAM-VLA: Decoupled Asynchronous Multimodal Vision Language Action model

DGX agent

arXiv:2606.12105v1 Announce Type: cross Abstract: Vision-language-action (VLA) models inherit a shared synchronous clock from vision-language pretraining, processing every input at one rate. This is m

applicationsarxiv-cs-cv
11 Jun 2026
Model Releases

Damage-TriageFormer: A Foundation-Model Framework for Typology-Based Building Damage Assessment from Mono-Temporal Imagery

DGX agent

arXiv:2606.12248v1 Announce Type: new Abstract: Decision-relevant building damage assessment is critical for prioritizing resources and recovery after a disaster, yet most automated methods either fla

model-releasesarxiv-cs-cv
11 Jun 2026
Research

DarkVGGT: Seeing Through Darkness Using Thermal Geometry without Daylight Tax

DGX agent

arXiv:2606.11326v1 Announce Type: new Abstract: Recent feed-forward 3D reconstruction methods have demonstrated strong performance and flexibility in efficient end-to-end scene geometry estimation fro

researcharxiv-cs-cv
11 Jun 2026
Applications

DeceptionX: Explainable Deception Detection with Multimodal Large Language Models

DGX agent

arXiv:2606.11385v1 Announce Type: new Abstract: Deception detection is a critical and highly challenging task within affective computing and behavioral analysis. Existing deep learning methods typical

applicationsarxiv-cs-cv
11 Jun 2026
Tutorials

DepthMaster: Unified Monocular Depth Estimation for Perspective and Panoramic Images

DGX agent

arXiv:2606.12368v1 Announce Type: new Abstract: While monocular depth estimation has achieved significant progress, achieving generalized metric depth estimation for both narrow field-of-view (FoV) pe

tutorialsarxiv-cs-cv
11 Jun 2026
Agents

DrivingAgent: Design and Scheduling Agents for Autonomous Driving Systems

DGX agent

arXiv:2606.12236v1 Announce Type: cross Abstract: Many autonomous driving systems are increasingly incorporating foundation models to improve generalization and handle long-tail scenarios. However, th

agentsarxiv-cs-cv
11 Jun 2026
Model Releases

DroneShield-AI: A Multi-Modal Sensor Fusion Framework for Real-Time Autonomous Drone Threat Detection, Behavioral Intent Classification, and Swarm Intelligence in Contested Airspace

DGX agent

arXiv:2606.11687v1 Announce Type: new Abstract: Unmanned Aerial Vehicle (UAV) threats have emerged as a defining security challenge of the 21st century. This paper presents DroneShield-AI, a unified o

model-releasesarxiv-cs-cv
11 Jun 2026
Research

DynaTok: Token-Based 4D Reconstruction from Partial Point Clouds

DGX agent

arXiv:2606.12189v1 Announce Type: new Abstract: We address 4D reconstruction from partial point cloud sequences, where depth-sensor observations are incomplete, unordered, and lack explicit temporal c

researcharxiv-cs-cv
11 Jun 2026
Hardware

Echoes of the Prior: A Computational Phenomenology of Forgetting

DGX agent

arXiv:2606.12340v1 Announce Type: new Abstract: Memory is not merely the storage of data; it is the scaffolding of reality. When biological memory fades, the world does not simply turn black; it regre

hardwarearxiv-cs-cv
11 Jun 2026
Research

ERN-Net : Evolving Reason Node-Net for Document Binarization

DGX agent

arXiv:2606.11710v1 Announce Type: new Abstract: This paper presents ERN-Net, an Evolving Reason Node-Net for efficient document image binarization. ERN-Net enhances degradation-sensitive regions, such

researcharxiv-cs-cv
11 Jun 2026
Research

EventRadar: Long-Range Visual UAV Discovery through Spatiotemporal Event Sensing

DGX agent

arXiv:2606.11285v1 Announce Type: new Abstract: Unauthorized unmanned aerial vehicle (UAV) activity around airports, public venues, and other sensitive sites has made protected-airspace monitoring inc

researcharxiv-cs-cv
11 Jun 2026
Research

EvoLMM: Self-Evolving Large Multimodal Models with Continuous Rewards

DGX agent

arXiv:2511.16672v4 Announce Type: replace Abstract: Recent advances in large multimodal models (LMMs) have enabled impressive reasoning and perception abilities, yet most existing training pipelines s

researcharxiv-cs-cv
11 Jun 2026
Research

Exploring Adaptive Masked Reconstruction for Self-Supervised Skeleton-Based Action Recognition

DGX agent

arXiv:2606.11450v1 Announce Type: new Abstract: Recently, masked skeleton reconstruction models have emerged as strong action representation learners, driving significant progress in self-supervised s

researcharxiv-cs-cv
11 Jun 2026
Agents

Feature extraction for plant growth estimation

DGX agent

arXiv:2606.11966v1 Announce Type: new Abstract: Precision agriculture requires the estimation of plant growth stages in real-time. When the plant growth stage is known, the wastage of resources in cul

agentsarxiv-cs-cv
11 Jun 2026
Research

Finding Sparse Subnetworks in One Training Cycle via Progressive Magnitude-Based Pruning

DGX agent

arXiv:2606.12278v1 Announce Type: new Abstract: Neural network pruning reduces model size by removing less important parameters while aiming to preserve predictive performance. Although the Lottery Ti

researcharxiv-cs-cv
11 Jun 2026
Tutorials

FitVTON: Fit-aware Virtual Try-On via Body-Garment Size Control

DGX agent

arXiv:2606.12012v1 Announce Type: new Abstract: While diffusion-based virtual try-on has achieved impressive visual realism, most methods treat the task as 2D inpainting, prioritizing texture preserva

tutorialsarxiv-cs-cv
11 Jun 2026
Safety

FreqKD: Frequency-Decoupled Cross-Modal Knowledge Distillation for Infrared Object Detection

DGX agent

arXiv:2606.11572v1 Announce Type: new Abstract: Transfer learning from large-scale RGB foundation models to infrared (IR) imagery through knowledge distillation (KD) remains challenging due to fundame

safetyarxiv-cs-cv
11 Jun 2026
Research

From 2D Grids to 1D Tokens: Reforming Shared Representations for Multimodal Image Fusion

DGX agent

arXiv:2606.12303v1 Announce Type: new Abstract: Multimodal image fusion aims to integrate complementary information from different modalities into a fused image that preserves rich local details while

researcharxiv-cs-cv
11 Jun 2026
Model Releases

From Content to Knowledge: Lightning Fast Long-Video Understanding with Neural Knowledge Representations

DGX agent

arXiv:2606.11913v1 Announce Type: new Abstract: We propose a new paradigm for long video understanding by treating a long video as a Neural Knowledge Representation (NKR). NKR represents video content

model-releasesarxiv-cs-cv
11 Jun 2026
← Previous
1…104105106107108…263
Next →