AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,618 results
Applications

Circular Quasiconformal Deturbulence: Geometry-Based Restoration from Multiple Turbulent Frames

DGX agent

arXiv:2504.13432v3 Announce Type: replace Abstract: Imaging through inhomogeneous media often results in severe distortions, posing significant challenges to downstream image-processing tasks. The lac

applicationsarxiv-cs-cv
26 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Coarse-to-Fine: A Hybrid Self-Supervised Method for Non-rigid 3D Shape Matching

DGX agent

arXiv:2606.26557v1 Announce Type: new Abstract: Non-rigid 3D shape matching is a fundamental task in computer vision and graphics. In this paper, we propose a hybrid self-supervised method based on a

researcharxiv-cs-cv
26 Jun 2026
Agents

CogAD: Cognitive-Hierarchy Guided End-to-End Autonomous Driving

DGX agent

arXiv:2505.21581v4 Announce Type: replace-cross Abstract: While end-to-end autonomous driving has advanced significantly, prevailing methods remain fundamentally misaligned with human cognitive princi

agentsarxiv-cs-cv
26 Jun 2026
Research

Computer Vision for MOBA Analytics: A Dataset and Baseline for Visibility Analysis in Dota 2

DGX agent

arXiv:2606.26970v1 Announce Type: new Abstract: Introduction: Most Multiplayer Online Battle Arena (MOBA) analytics studies rely on structured data, which does not directly capture what each team coul

researcharxiv-cs-cv
26 Jun 2026
Research

Controllable Egocentric Video Generation via Occlusion-Aware Sparse 3D Hand Joints

DGX agent

arXiv:2603.11755v2 Announce Type: replace Abstract: Controllable video generation for complex hand-object interactions is a critical step toward building visual world models. However, existing methods

researcharxiv-cs-cv
26 Jun 2026
Model Releases

CORTEX: A Structured Reasoning Benchmark for Trustworthy 3D Chest CT MLLMs

DGX agent

arXiv:2606.27264v1 Announce Type: new Abstract: Reasoning in multimodal large language models (MLLMs) has shown strong promise in medical imaging. However, this reasoning is usually free-form text jud

model-releasesarxiv-cs-cv
26 Jun 2026
Research

DanceDuo: Bridging Human Movement and AI Choreography

DGX agent

arXiv:2606.26507v1 Announce Type: cross Abstract: In recent years, advancements in deep learning and generative models have revolutionized music-driven dance generation. This paper introduces a novel

researcharxiv-cs-cv
26 Jun 2026
Model Releases

DeCoFlow: Structural Decomposition of Normalizing Flows for Continual Anomaly Detection

DGX agent

arXiv:2606.26687v1 Announce Type: new Abstract: In industrial environments, new product categories arrive sequentially, requiring continual anomaly detection without access to past data. Normalizing F

model-releasesarxiv-cs-cv
26 Jun 2026
Safety

Depth-Semantic Alignment and Affinity-Guided Fusion for Structured Radar Point Cloud Generation

DGX agent

arXiv:2606.26743v1 Announce Type: new Abstract: Point clouds are an important carrier of three-dimensional spatial information, and their quality directly affects the performance of downstream percept

safetyarxiv-cs-cv
26 Jun 2026
Model Releases

DiCoBench: Benchmarking Multi-Image Fine-Grained Perception via Differential and Commonality Visual Cues

DGX agent

arXiv:2606.26602v1 Announce Type: new Abstract: Recent advancements in Multimodal Large Language Models (MLLMs) have demonstrated impressive fine-grained perception capabilities. However, existing ben

model-releasesarxiv-cs-cv
26 Jun 2026
Research

DinoLink: A Token-Centric Representation Compression Framework for Bandwidth-Constrained Collaborative V2X Perception

DGX agent

arXiv:2606.26398v1 Announce Type: new Abstract: High-precision remote perception is often hindered by the severe bandwidth constraints of Vehicle-to-Everything (V2X) networks. We propose extit{DinoLin

researcharxiv-cs-cv
26 Jun 2026
Research

DnA: Denoising Attention for Visual Tasks

DGX agent

arXiv:2606.27372v1 Announce Type: new Abstract: The softmax activation in multihead attention (MHA) is the de facto standard for attention-based models in visual perception tasks. However, standard so

researcharxiv-cs-cv
26 Jun 2026
Model Releases

Do Image Editing Models Understand Lighting?

DGX agent

arXiv:2606.26738v1 Announce Type: new Abstract: While recent advancements in generative image editing models have achieved stunning visual fidelity, it remains an open question whether these systems p

model-releasesarxiv-cs-cv
26 Jun 2026
Safety

DocArena: Turning Raw Documents into Controllable Training Environments for Document Search Agents

DGX agent

arXiv:2606.26122v1 Announce Type: new Abstract: Recent methods train search agents via reinforcement learning from (question, answer, evidence) tuples without requiring expert trajectories. The tuples

safetyarxiv-cs-cv
26 Jun 2026
Safety

Don't Settle at the Mode! Mitigating Diversity Collapse in Pretrained Flow Models via Feature Self-Guidance

DGX agent

arXiv:2606.27371v1 Announce Type: new Abstract: State-of-the-art flow models generate stunning images from text or image prompts. However, they suffer from diversity collapse when generating multiple

safetyarxiv-cs-cv
26 Jun 2026
Model Releases

Dual-Prior Guided Null-Space Learning with Mixture-of-Splines for Arbitrary Medical Slice Super-Resolution

DGX agent

arXiv:2606.26716v1 Announce Type: cross Abstract: Arbitrary slice super-resolution reconstructs isotropic volumes from anisotropic clinical acquisitions by synthesizing intermediate slices at arbitrar

model-releasesarxiv-cs-cv
26 Jun 2026
Tutorials

DynFS-MoE: Dynamic Functional-Structural Mixture-of-Experts for Post-Traumatic Epilepsy Diagnosis

DGX agent

arXiv:2606.16203v3 Announce Type: replace Abstract: Post-traumatic epilepsy (PTE) is a severe complication of traumatic brain injury (TBI). Yet, early identification remains challenging due to the com

tutorialsarxiv-cs-cv
26 Jun 2026
Local Ai

EndoUFM: Utilizing Foundation Models for Monocular depth estimation of endoscopic images

DGX agent

arXiv:2508.17916v2 Announce Type: replace Abstract: Depth estimation is a foundational component for 3D reconstruction in minimally invasive endoscopic surgeries. However, existing monocular depth est

local-aiarxiv-cs-cv
26 Jun 2026
Hardware

Event-based Gaze Control System for Accurate Real-time Spin Estimation in Professional Ball Games

DGX agent

arXiv:2606.26780v1 Announce Type: new Abstract: Spin plays a crucial role in many ball sports due to its effect on the trajectory of the ball. Vision-based estimation of the ball's spin during a game

hardwarearxiv-cs-cv
26 Jun 2026
Research

Exact and Deterministic Patch Descriptor Retrieval via Hierarchical Normalization

DGX agent

arXiv:2606.27280v1 Announce Type: new Abstract: We present a patch descriptor retrieval method that returns the exact nearest neighbour -- provably identical to exhaustive full-vector search -- while

researcharxiv-cs-cv
26 Jun 2026
Research

Extracting Neural Materials from Multi-view Images

DGX agent

arXiv:2606.26715v1 Announce Type: new Abstract: Neural materials can represent complex specular reflections and scattering effects in a compact, universal basis. However, acquiring and authoring such

researcharxiv-cs-cv
26 Jun 2026
Local Ai

Fast LeWorldModel

DGX agent

arXiv:2606.26217v1 Announce Type: cross Abstract: Joint-Embedding Predictive Architectures (JEPAs), including recent LeWorldModel (LeWM), have become a promising foundation for reconstruction-free vis

local-aiarxiv-cs-cv
26 Jun 2026
Model Releases

FlameVQA: A Physically-Grounded UAV Wildfire VQA Benchmark with Radiometric Thermal Supervision

DGX agent

arXiv:2606.27128v1 Announce Type: new Abstract: Wildfire monitoring from UAVs requires reliable reasoning over complex aerial scenes, where smoke, scale variation, and occlusions often limit RGB-only

model-releasesarxiv-cs-cv
26 Jun 2026
Research

Focusing on What Matters: Saliency-Harnessing Accurate Routing for Diffusion MoE

DGX agent

arXiv:2606.26938v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) architectures have emerged as a powerful paradigm for scaling diffusion models in visual generation. Recent advancements have f

researcharxiv-cs-cv
26 Jun 2026
Tutorials

Forget, Anticipate and Adapt: Test Time Training for Long Videos

DGX agent

arXiv:2606.26515v1 Announce Type: new Abstract: Test Time Training (TTT) is a mechanism in which a model adapts to an incoming test-sample by performing some self-supervised (SSL) task and updating it

tutorialsarxiv-cs-cv
26 Jun 2026
Research

FracEvent: Event-Camera Simulation via Fractional-Relaxation Pixel Dynamics

DGX agent

arXiv:2606.26636v1 Announce Type: new Abstract: Event cameras asynchronously report brightness changes with microsecond-level temporal resolution, but real event data remain difficult to collect at sc

researcharxiv-cs-cv
26 Jun 2026
Research

Full spectrum Unlearnable Examples via Spectral Equalization

DGX agent

arXiv:2606.26719v1 Announce Type: new Abstract: Unlearnable examples (UEs) protect training data by injecting imperceptible perturbations so that models fail to extract exploitable representations. In

researcharxiv-cs-cv
26 Jun 2026
Model Releases

GeMoE: Gating Entropy is All You Need for Uncertainty-aware Adaptive Routing in MoE-based Large Vision-Language Models

DGX agent

arXiv:2606.26287v1 Announce Type: new Abstract: With the increase in model parameters and training data, the instruction following and generalization capabilities of Large VisionLanguage Models (LVLMs

model-releasesarxiv-cs-cv
26 Jun 2026
Tutorials

Generating a Paracosm for Training-Free Zero-Shot Composed Image Retrieval

DGX agent

arXiv:2602.00813v5 Announce Type: replace Abstract: Composed Image Retrieval (CIR) is the task of retrieving a target image from a database using a multimodal query, which consists of a reference imag

tutorialsarxiv-cs-cv
26 Jun 2026
Research

Geometric Gradient Rectification for Safe Open-Set Semi-Supervised Learning

DGX agent

arXiv:2606.26973v1 Announce Type: new Abstract: Open-set semi-supervised learning aims to leverage unlabeled data that may contain out-of-distribution outliers while maintaining performance on in-dist

researcharxiv-cs-cv
26 Jun 2026
Research

Geometry-Aware Superpixel Graph Transformer with Metadata for Skin Lesion Classification

DGX agent

arXiv:2606.20390v2 Announce Type: replace Abstract: Automated skin cancer classification from dermoscopic images remains challenging due to heterogeneous lesion structure, strong intra-class variabili

researcharxiv-cs-cv
26 Jun 2026
Model Releases

Hallucination in World Models is Predictable and Preventable

DGX agent

arXiv:2606.27326v1 Announce Type: cross Abstract: Modern generative world models render increasingly realistic action-controllable futures, yet they frequently hallucinate: rollouts remain visually fl

model-releasesarxiv-cs-cv
26 Jun 2026
Applications

Identifying the Unknown: Prompt-Free Open Vocabulary Anomaly Recognition for Robot-Object Interaction

DGX agent

arXiv:2606.26829v1 Announce Type: new Abstract: Robots operating in real-world environments must in general be able to recognize previously unseen objects. As robotic systems move toward open-world au

applicationsarxiv-cs-cv
26 Jun 2026
Safety

Improving Vision-Language-Action Model Fine-Tuning with Structured Stage and Keyframe Supervision

DGX agent

arXiv:2606.26801v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have shown strong potential for generalizable robotic manipulation. During fine-tuning, however, action supervisio

safetyarxiv-cs-cv
26 Jun 2026
Research

Intracranial Aneurysm Classification and Segmentation via Tri-Axial ROI and Multi-Task Learning

DGX agent

arXiv:2606.26706v1 Announce Type: new Abstract: Intracranial aneurysms are often asymptomatic until rupture, which carries high mortality. Rupture risk assessment and treatment planning depend on both

researcharxiv-cs-cv
26 Jun 2026
Model Releases

Layer-Specific Prompt Fusion Discovery via Differentiable Search in Vision Foundation Models

DGX agent

arXiv:2606.26379v1 Announce Type: new Abstract: Visual prompt tuning has emerged as a parameter-efficient fine-tuning approach for adapting large-scale Vision Transformers (ViTs) to downstream tasks.

model-releasesarxiv-cs-cv
26 Jun 2026
Safety

LayersReg: A Layer-by-Layer Progressive Regressor for Reliable Intraoperative 3D/2D Registration

DGX agent

arXiv:2606.26647v1 Announce Type: new Abstract: 3D/2D registration serves as a cornerstone technique in surgical navigation. Traditional iterative optimization algorithms suffer from low efficiency an

safetyarxiv-cs-cv
26 Jun 2026
Research

LearniBridge: Learnable Calibration of Feature Caching for Diffusion Models Acceleration

DGX agent

arXiv:2606.26778v1 Announce Type: new Abstract: Diffusion Transformers (DiTs) have driven substantial progress in image and video generation but suffer from prohibitive computational costs. Feature ca

researcharxiv-cs-cv
26 Jun 2026
Local Ai

Learning Adversarial Augmentation Policies for Robust Garlic Seedling Detection

DGX agent

arXiv:2606.26828v1 Announce Type: new Abstract: Accurate seedling detection during early growth stages is essential for timely replanting and effective crop management in precision agriculture. Howeve

local-aiarxiv-cs-cv
26 Jun 2026
Model Releases

Learning Language-Driven Sequence-Level Modal-Invariant Representations for Video-Based Visible-Infrared Person Re-Identification

DGX agent

arXiv:2601.12062v2 Announce Type: replace Abstract: The core of video-based visible-infrared person re-identification (VVI-ReID) lies in learning sequence-level modal-invariant representations across

model-releasesarxiv-cs-cv
26 Jun 2026
Model Releases

Liquid Fusion of Heterogeneous Representations Towards General Salient Object Detection

DGX agent

arXiv:2606.26849v1 Announce Type: new Abstract: General Salient Object Detection (SOD) aims to identify and segment visually interesting objects from uni-modality or multi-modality scenes, recently ad

model-releasesarxiv-cs-cv
26 Jun 2026
Safety

LISA: Likelihood Score Alignment for Visual-condition Controllable Generation

DGX agent

arXiv:2606.27192v1 Announce Type: new Abstract: The prevalent dual-branch paradigm, i.e., training a side network to encode visual conditions and fusing its intermediate-layer features to a frozen pre

safetyarxiv-cs-cv
26 Jun 2026
Model Releases

LiveEdit: Towards Real-Time Diffusion-Based Streaming Video Editing

DGX agent

arXiv:2606.26740v1 Announce Type: new Abstract: Streaming video editing has made rapid progress, yet practical deployment is still limited by two core issues: maintaining stable backgrounds and non-ed

model-releasesarxiv-cs-cv
26 Jun 2026
Research

LogicIR: Logic Gate Networks for Image Restoration

DGX agent

arXiv:2606.26609v1 Announce Type: new Abstract: Image restoration aims to reconstruct high-quality images from degraded low-quality inputs. As the computational demands of image restoration models con

researcharxiv-cs-cv
26 Jun 2026
Model Releases

Mask to Concept: Auto-Promptable SAM3 via Efficient Test-Time Concept Embedding Search for Few-Shot Annotation

DGX agent

arXiv:2606.26711v1 Announce Type: new Abstract: Transforming foundation segmentation models from human-prompted tools into auto-promptable annotators is critical for scalable medical data annotation.

model-releasesarxiv-cs-cv
26 Jun 2026
Research

MAVFusion: Efficient Infrared and Visible Video Fusion via Motion-Aware Sparse Interaction

DGX agent

arXiv:2604.01958v2 Announce Type: replace Abstract: Infrared and visible video fusion combines the object saliency from infrared images with the texture details from visible images to produce semantic

researcharxiv-cs-cv
26 Jun 2026
Applications

Methane-Plume Segmentation From Hyperspectral Satellite Imagery Via Multimodal Deep Learning

DGX agent

arXiv:2606.26416v1 Announce Type: new Abstract: Efficient detection of methane plumes is crucial for understanding and mitigating global warming, as accurately identifying and segmenting them in earth

applicationsarxiv-cs-cv
26 Jun 2026
Tutorials

Modeling Local, Global, and Cross-Modal Context in Multimodal 3D MRI

DGX agent

arXiv:2606.26894v1 Announce Type: new Abstract: Brain MRI poses a fundamental challenge for machine learning: models must learn from high-dimensional 3D data spanning multiple co-registered modalities

tutorialsarxiv-cs-cv
26 Jun 2026
← Previous
1…8687888990…263
Next →