AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,618 results
Applications

Hestia: Voxel-Face-Aware Hierarchical Next-Best-View Acquisition for Efficient 3D Reconstruction

DGX agent

arXiv:2508.01014v4 Announce Type: replace-cross Abstract: Advances in 3D reconstruction and novel view synthesis have enabled efficient and photorealistic rendering. However, images for reconstruction

applicationsarxiv-cs-cv
18 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Hierarchical and Holistic Open-Vocabulary Functional 3D Scene Graphs for Indoor Spaces

DGX agent

arXiv:2605.15753v1 Announce Type: cross Abstract: Functional 3D scene graphs offer a versatile and flexible representation for 3D scene understanding and robotic manipulation, defined by object nodes,

model-releasesarxiv-cs-cv
18 May 2026
Research

Highly Detailed and Generalizable Broadleaf Tree Crown Instance Segmentation from UAV Imagery

DGX agent

arXiv:2605.15673v1 Announce Type: cross Abstract: We present a highly detailed instance segmentation model for delineating individual tree crowns in natural broadleaf forests using aerial imagery acqu

researcharxiv-cs-cv
18 May 2026
Tutorials

How to Choose Your Teacher for Fine Grained Image Recognition

DGX agent

arXiv:2605.15689v1 Announce Type: new Abstract: Fine-grained image recognition classifies subcategories such as bird species or car models. While state-of-the-art (SOTA) models are accurate, they are

tutorialsarxiv-cs-cv
18 May 2026
Safety

HyperDiT: Hyper-Connected Transformers for High-Fidelity Pixel-Space Diffusion

DGX agent

arXiv:2605.15741v1 Announce Type: new Abstract: Pixel-space diffusion models bypass the reconstruction bottleneck of Variational Autoencoders (VAEs) but face a fundamental 'granularity dilemma': captu

safetyarxiv-cs-cv
18 May 2026
Research

IHF-Harmony: Multi-Modality Magnetic Resonance Images Harmonization using Invertible Hierarchy Flow Model

DGX agent

arXiv:2602.21536v2 Announce Type: replace Abstract: Retrospective MRI harmonization is limited by poor scalability across modalities and reliance on traveling subject datasets. To address these challe

researcharxiv-cs-cv
18 May 2026
Applications

Inevitable Encounters: Backdoor Attacks Involving Lossy Compression

DGX agent

arXiv:2603.13864v2 Announce Type: replace-cross Abstract: Real-world backdoor attacks often require poisoned datasets to be stored and transmitted before being used to compromise deep learning systems

applicationsarxiv-cs-cv
18 May 2026
Tutorials

Integrating chemical structures as treatments improves representations of microscopy images for morphological profiling

DGX agent

arXiv:2504.09544v3 Announce Type: replace-cross Abstract: Recent advances in self-supervised deep learning have improved our ability to quantify cellular morphological changes in high-throughput micro

tutorialsarxiv-cs-cv
18 May 2026
Tutorials

Invaria: Learning Scale and Density Invariance in Point Clouds via Next-Resolution Prediction

DGX agent

arXiv:2605.15923v1 Announce Type: new Abstract: Modern image encoders achieve high generalization by decoupling semantic meaning from resolution, an ability yet to be fully realized in the 3D domain.

tutorialsarxiv-cs-cv
18 May 2026
Local Ai

JAM-Flow: Joint Audio-Motion Synthesis with Flow Matching

DGX agent

arXiv:2506.23552v2 Announce Type: replace Abstract: The intrinsic link between facial motion and speech is often overlooked in generative modeling, where talking head synthesis and text-to-speech (TTS

local-aiarxiv-cs-cv
18 May 2026
Applications

LAPS: Improving Incremental LiDAR Mapping using Active Pooling and Sampling for Neural Distance Fields

DGX agent

arXiv:2605.15496v1 Announce Type: cross Abstract: Neural distance fields offer a compact and continuous representation of 3D geometry, making them attractive for incremental LiDAR mapping. However, th

applicationsarxiv-cs-cv
18 May 2026
Research

Layer Selection in Feature-Based Losses Affects Image Quality and Microstructural Consistency in Deep Learning Super-Resolution of Brain Diffusion MRI

DGX agent

arXiv:2605.15895v1 Announce Type: cross Abstract: Clinical application of high-resolution diffusion MRI is hindered by hardware limitations and prohibitive scan times, motivating computational super-r

researcharxiv-cs-cv
18 May 2026
Tutorials

LDGuid: A Framework for Robust Change Detection via Latent Difference Guidance

DGX agent

arXiv:2605.15582v1 Announce Type: new Abstract: Modern deep learning models for change detection (CD) often struggle to explicitly represent task-relevant semantic differences. This paper proposes the

tutorialsarxiv-cs-cv
18 May 2026
Model Releases

Learn2Splat: Extending the Horizon of Learned 3DGS Optimization

DGX agent

arXiv:2605.15760v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) optimization is most commonly performed using standard optimizers (Adam, SGD). While stable across diverse scenes, standard

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

Learning Disentangled Representations for Generalized Multi-view Clustering

DGX agent

arXiv:2605.15640v1 Announce Type: new Abstract: Multi-View Clustering (MVC) has gained significant attention for its ability to leverage complementary information across diverse views. However, existi

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

Learning Dynamic Structural Specialization for Underwater Salient Object Detection

DGX agent

arXiv:2605.15535v1 Announce Type: new Abstract: Underwater salient object detection (USOD) has attracted increasing attention for underwater visual scene understanding and vision-guided robotic applic

model-releasesarxiv-cs-cv
18 May 2026
Research

Learning Normalized Energy Models for Linear Inverse Problems

DGX agent

arXiv:2605.15487v1 Announce Type: cross Abstract: Generative diffusion models can provide powerful prior probability models for inverse problems in imaging, but existing implementations suffer from tw

researcharxiv-cs-cv
18 May 2026
Safety

LRCP: Low-Rank Compressibility Guided Visual Token Pruning for Efficient LVLMs

DGX agent

arXiv:2605.15621v1 Announce Type: new Abstract: Large vision-language models (LVLMs) achieve strong multimodal understanding, but their inference cost grows rapidly with the number of visual tokens, e

safetyarxiv-cs-cv
18 May 2026
Safety

LUIVITON: Learned Universal Interoperable VIrtual Try-ON

DGX agent

arXiv:2509.05030v2 Announce Type: replace Abstract: To enable large-scale reuse of real-world 3D assets, where garments and characters rarely share skeletons, templates, or dense correspondences, we p

safetyarxiv-cs-cv
18 May 2026
Safety

MAgSeg: Segmentation of Agricultural Landscapes in High-Resolution Satellite Imagery using Multimodal Large Language Models

DGX agent

arXiv:2605.16179v1 Announce Type: new Abstract: Agricultural landscape segmentation in the Global South is challenging as it is characterized by fragmented plots, high intra-class variance, and a scar

safetyarxiv-cs-cv
18 May 2026
Model Releases

Mask-Morph Graph U-Net: A Generalisable Mesh-Based Surrogate for Crashworthiness Field Prediction under Large Geometric Variation

DGX agent

arXiv:2605.15231v1 Announce Type: cross Abstract: Nonlinear finite element crash simulations are accurate but computationally expensive, limiting their use in iterative design optimisation. Machine-le

model-releasesarxiv-cs-cv
18 May 2026
Safety

MaTe: Images Are All You Need for Material Transfer via Diffusion Transformer

DGX agent

arXiv:2605.15660v1 Announce Type: new Abstract: Recent diffusion-based methods for material transfer rely on image fine-tuning or complex architectures with assistive networks, but face challenges inc

safetyarxiv-cs-cv
18 May 2026
Model Releases

MI-CXR: A Benchmark for Longitudinal Reasoning over Multi-Interval Chest X-rays

DGX agent

arXiv:2605.15574v1 Announce Type: new Abstract: Longitudinal chest X-ray (CXR) interpretation requires reasoning over disease evolution across multiple patient visits, yet most existing medical VQA be

model-releasesarxiv-cs-cv
18 May 2026
Local Ai

MIND: Decoupling Model-Induced Label Noise via Latent Manifold Disentanglement

DGX agent

arXiv:2605.16081v1 Announce Type: cross Abstract: The paradigm of learning from automatic annotations driven by pre-trained experts and Foundation Models dominates data-hungry applications. However, i

local-aiarxiv-cs-cv
18 May 2026
Model Releases

Minerva-Ego: Spatiotemporal Hints for Egocentric Video Understanding

DGX agent

arXiv:2605.15342v1 Announce Type: new Abstract: Video reasoning models are a core component of egocentric and embodied agents. However, standard benchmarks for assessing models provide only evaluation

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

MorphoHELM: A Comprehensive Benchmark for Evaluating Representations for Microscopy-Based Morphology Assays

DGX agent

arXiv:2605.15383v1 Announce Type: new Abstract: Microscopy images contain rich information about how cells respond to perturbations, making them essential to applications like drug screening. To quant

model-releasesarxiv-cs-cv
18 May 2026
Research

mRadNet: A Compact Radar Object Detector with MetaFormer

DGX agent

arXiv:2509.16223v3 Announce Type: replace-cross Abstract: Frequency-modulated continuous wave radars have gained increasing popularity in the automotive industry. Their robustness against adverse weat

researcharxiv-cs-cv
18 May 2026
Research

Multimodal Object Detection Under Sparse Forest-Canopy Occlusion

DGX agent

arXiv:2605.15326v1 Announce Type: new Abstract: Reliable detection of humans beneath forest canopy remains a difficult remote-sensing challenge due to sparse, structured, and viewpoint-dependent occlu

researcharxiv-cs-cv
18 May 2026
Model Releases

Navigating the Challenges of AI-Generated Image Detection in the Wild: What Truly Matters?

DGX agent

arXiv:2507.10236v2 Announce Type: replace Abstract: As generative Artificial Intelligence (AI) advances, the realism of AI generated imagery has reached a threshold capable of deceiving even vigilant

model-releasesarxiv-cs-cv
18 May 2026
Research

Neurosymbolic Object-Centric Learning with Distant Supervision

DGX agent

arXiv:2506.16129v2 Announce Type: replace Abstract: Neurosymbolic learning can use symbolic rules to provide supervision for latent concepts from weak labels, but it commonly assumes that the entities

researcharxiv-cs-cv
18 May 2026
Safety

Neutral-Reference Prompting for Vision-Language Models

DGX agent

arXiv:2605.15615v1 Announce Type: new Abstract: Efficient transfer learning of vision-language models (VLMs) commonly suffers from a Base-New Trade-off (BNT): improving performance on unseen (new) cla

safetyarxiv-cs-cv
18 May 2026
Local Ai

Not All Tasks Quantize Equally: Fisher-Guided Quantization for Visual Geometry Transformer

DGX agent

arXiv:2605.15828v1 Announce Type: new Abstract: Feed-forward 3D reconstruction models, represented by Visual Geometry Grounded Transformer (VGGT), jointly predict multiple visual geometry tasks such a

local-aiarxiv-cs-cv
18 May 2026
Model Releases

On RGB-TIR Stereo Calibration under Extreme Resolution Asymmetry

DGX agent

arXiv:2605.15860v1 Announce Type: new Abstract: Accurate geometric calibration of RGB-thermal infrared (TIR) stereo camera systems is essential for multimodal building envelope analysis, yet remains c

model-releasesarxiv-cs-cv
18 May 2026
Research

One Pass Is Not Enough: Recursive Latent Refinement for Generative Models

DGX agent

arXiv:2605.15309v1 Announce Type: new Abstract: Despite remarkable progress, image generation is far from solved. The dominant metric, FID, conflates sample fidelity with mode coverage and is close to

researcharxiv-cs-cv
18 May 2026
Safety

OpenFrontier: General Navigation with Visual-Language Grounded Frontiers

DGX agent

arXiv:2603.05377v2 Announce Type: replace-cross Abstract: Open-world navigation requires robots to make decisions in complex everyday environments while adapting to flexible task requirements. Convent

safetyarxiv-cs-cv
18 May 2026
Research

Overlap-aware segmentation for topological reconstruction of obscured objects

DGX agent

arXiv:2510.06194v2 Announce Type: replace-cross Abstract: The separation of overlapping objects presents a significant challenge in scientific imaging. While deep learning segmentation-regression algo

researcharxiv-cs-cv
18 May 2026
Model Releases

PhyDetEx: Detecting and Explaining the Physical Plausibility of T2V Models

DGX agent

arXiv:2512.01843v2 Announce Type: replace Abstract: Driven by the growing capacity and training scale, Text-to-Video (T2V) generation models have recently achieved substantial progress in video qualit

model-releasesarxiv-cs-cv
18 May 2026
Applications

Preprocessing Algorithm Leveraging Geometric Modeling for Scale Correction in Hyperspectral Images for Improved Unmixing Performance

DGX agent

arXiv:2508.08431v3 Announce Type: replace-cross Abstract: Spectral variability significantly impacts the accuracy and convergence of hyperspectral unmixing algorithms. Many methods address complex spe

applicationsarxiv-cs-cv
18 May 2026
Model Releases

Probabilistic Dating of Historical Manuscripts via Evidential Deep Regression on Visual Script Features

DGX agent

arXiv:2605.06475v1 Announce Type: cross Abstract: We introduce a probabilistic approach for dating historical manuscript pages from visual features alone. Instead of aggregating centuries into classes

model-releasesarxiv-cs-cv
18 May 2026
Safety

ReactiveGWM: Steering NPC in Reactive Game World Models

DGX agent

arXiv:2605.15256v1 Announce Type: new Abstract: Current game world models simulate environments from a subjective, player-centric perspective. However, by treating the Non-Player Character (NPC) merel

safetyarxiv-cs-cv
18 May 2026
Safety

ReAlign: Generalizable Image Forgery Detection via Reasoning-Aligned Representation

DGX agent

arXiv:2605.16080v1 Announce Type: new Abstract: The rise of AI-generated images (AIGIs) poses growing challenges for digital authenticity, prompting the need for efficient, generalizable image forgery

safetyarxiv-cs-cv
18 May 2026
Applications

RealRep: Generalized SDR-to-HDR Conversion via Attribute-Disentangled Representation Learning

DGX agent

arXiv:2505.07322v4 Announce Type: replace Abstract: High-Dynamic-Range Wide-Color-Gamut (HDR-WCG) technology is becoming increasingly widespread, driving a growing need for converting Standard Dynamic

applicationsarxiv-cs-cv
18 May 2026
Model Releases

Registers Matter for Pixel-Space Diffusion Transformers

DGX agent

arXiv:2605.16147v1 Announce Type: new Abstract: Vision Transformers (ViTs) are known to exhibit high-norm patch-token outliers that degrade feature map quality, a problem effectively mitigated by exti

model-releasesarxiv-cs-cv
18 May 2026
Safety

Res^2CLIP: Few-Shot Generalist Anomaly Detection with Residual-to-Residual Alignment

DGX agent

arXiv:2605.16171v1 Announce Type: new Abstract: Few-shot Generalist Anomaly Detection requires models to generalize to novel categories without retraining, posing significant challenges in real-world

safetyarxiv-cs-cv
18 May 2026
Safety

Reversing the Flow: Generation-to-Understanding Synergy in Large Multimodal Models

DGX agent

arXiv:2605.15792v1 Announce Type: new Abstract: The long-standing goal of multimodal AI is to build unified models in which visual understanding and visual generation mutually enhance one another. Des

safetyarxiv-cs-cv
18 May 2026
Research

RoiMAM: Region-of-Interest Medical Attention Model for Efficient Vision-Language Understanding

DGX agent

arXiv:2605.15561v1 Announce Type: new Abstract: Vision-Language Models (VLMs) facilitate medical visual question answering (MedVQA) by jointly interpreting images and text. However, existing models ty

researcharxiv-cs-cv
18 May 2026
Safety

Seeing What Matters: Visual Preference Policy Optimization for Visual Generation

DGX agent

arXiv:2511.18719v4 Announce Type: replace Abstract: Reinforcement learning (RL) has become a powerful tool for post-training visual generative models, with Group Relative Policy Optimization (GRPO) in

safetyarxiv-cs-cv
18 May 2026
Local Ai

Segmentation, Detection and Explanation: A Unified Framework for CT Appearance Reasoning

DGX agent

arXiv:2605.15997v1 Announce Type: new Abstract: Recent progress in deep learning has significantly advanced CT image analysis, particularly for segmentation tasks. However, these advances are largely

local-aiarxiv-cs-cv
18 May 2026
← Previous
1…166167168169170…263
Next →