AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,515 results
28 Jul 2026

Which Workloads Belong in Orbit? A Workload-First Framework for Orbital Data Centers Using Semantic Abstraction

ResearchDGX agent

arXiv:2603.20317v2 Announce Type: replace Abstract: Space-based compute is becoming plausible as launch costs fall and data-intensive AI workloads grow. This paper proposes a workload-centric framewor

XMatchAD: A Cross-Modal Matching Perspective on Reconstruction-based Anomaly Detection

Model ReleasesDGX agent

arXiv:2607.23658v1 Announce Type: new Abstract: The remarkable success of reconstruction-based methods in Unsupervised Anomaly Detection (UAD) lies in their ability to identify and localize anomalies

27 Jul 2026

A Dual Path Framework with Hotspot Guided Fusion for Three Dimensional CT to PET Synthesis in Head and Neck Cancer


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
ResearchDGX agent

arXiv:2607.21800v1 Announce Type: cross Abstract: 18F-FDG PET/CT plays a central role in staging, treatment planning, and response assessment for head and neck cancer by providing functional informati

A Framework for Individual Tree Growth Reconstruction Using Multi-Platform Laser Scanning

ResearchDGX agent

arXiv:2607.22129v1 Announce Type: new Abstract: Accurate tree-level forest monitoring using laser scanning data requires reliable tree delineation, consistent tree correspondence across multitemporal

A Smooth Phase-Separation Model for Weak-Boundary Segmentation of Homogeneous Structures

ResearchDGX agent

arXiv:2607.22053v1 Announce Type: new Abstract: Segmentation of adjacent structures with similar intensity distributions remains a challenging problem in image analysis, particularly when object bound

Active few-shot segmentation by reinforcing data selection

AgentsDGX agent

arXiv:2607.22371v1 Announce Type: new Abstract: Few-shot learning enables medical image segmentation models to adapt to new tasks using only a small number of labelled examples. However, adaptation pe

Adaptive Contrast Enhancement and Optimised Feature Matching for RootSIFT-Based Palm-Vein Recognition

Model ReleasesDGX agent

arXiv:2607.16077v2 Announce Type: replace Abstract: Palm-vein recognition is a highly secure biometric modality due to the uniqueness and subcutaneous nature of vein patterns. However, low contrast in

AgentHOI: Multi-Agent Reasoning for Human-Object-Interaction Video Generation via Implicit Representation Alignment

SafetyDGX agent

arXiv:2607.22241v1 Announce Type: new Abstract: Recent advances in video diffusion models have spurred interest in human-object interaction (HOI) video generation, which demands fine-grained control o

Alleviating Regional Shortcuts for Few-Shot Class-Incremental Learning

TutorialsDGX agent

arXiv:2607.22072v1 Announce Type: new Abstract: Few-shot class-incremental learning (FSCIL) aims to incrementally learn novel classes with only a few samples while avoiding forgetting base classes. Ho

Atlas 2 -- Foundation models for clinical deployment

ResearchDGX agent

arXiv:2601.05148v2 Announce Type: replace Abstract: Pathology foundation models substantially advanced the possibilities in computational pathology --- yet tradeoffs in terms of performance, robustnes

Automatic Map Density Selection for Locally-Performant Visual Place Recognition

ResearchDGX agent

arXiv:2602.21473v3 Announce Type: replace Abstract: A key challenge in translating Visual Place Recognition (VPR) from the lab to long-term deployment is ensuring a priori that a system can meet user-

Be Consistent! Enhancing Robust Visual Reasoning in LVLMs with Consistency Constraints

Model ReleasesDGX agent

arXiv:2607.21722v1 Announce Type: new Abstract: While Large Vision-Language Models (LVLMs) exhibit strong perceptual capabilities, they remain vulnerable in visual reasoning tasks. Existing benchmarks

Bowel Obstruction Detection and Localization on Abdominal CT with Deep Learning

Local AiDGX agent

arXiv:2607.22173v1 Announce Type: new Abstract: Bowel obstruction is a common and potentially life-threatening gastrointestinal condition. In the face of rising diagnostic workloads, the automated dia

CARA: Concept-Aware Risk Attention for Interpretable Collision Anticipation

AgentsDGX agent

arXiv:2607.22494v1 Announce Type: cross Abstract: Collision anticipation in autonomous driving requires not only accurate early warnings but also interpretable reasoning about what risk factors are be

CARDIAG: A Dense Segment Classification Benchmark of Deep Learning Architectures for Coronary Angiography

Model ReleasesDGX agent

arXiv:2607.22139v1 Announce Type: new Abstract: Accurate pixel-level classification of coronary angiograms is critical for cardiovascular disease assessment, yet the field lacks standardized evaluatio

CARE: Anti-entanglement Ultrasound Image Segmentation via Channel-Aware Region Extrication

ResearchDGX agent

arXiv:2508.13899v2 Announce Type: replace Abstract: Accurate ultrasound image segmentation is fundamentally challenged by target-context entanglement, where lesion cues are easily mixed with surroundi

Class-Balanced Softmax: A Bayes Theory-Based Method for Long-Tailed Recognition

ResearchDGX agent

arXiv:2607.22258v1 Announce Type: cross Abstract: Deep learning models using traditional softmax classifiers have achieved remarkable success in various classification tasks. However, their performanc

Closing the Loop: Training-Free Revisit Consistency for Autoregressive Generative Rendering

ApplicationsDGX agent

arXiv:2607.21848v1 Announce Type: new Abstract: Recent conditional video generation models have shown promising potentials to transform 3D engine renderings, such as depth maps and untextured geometry

Color Me Correctly: Bridging Perceptual Color Spaces and Text Embeddings for Improved Diffusion Generation

SafetyDGX agent

arXiv:2509.10058v2 Announce Type: replace Abstract: Accurate color alignment in text-to-image (T2I) generation is critical for applications such as fashion, product visualization, and interior design,

CommandLM: Data driven behavior level descriptor for ego vehicles

SafetyDGX agent

arXiv:2607.22078v1 Announce Type: new Abstract: As autonomous driving systems move toward real-world deployment, interpretable, behavior-level decision-making is essential for safety, trust, and regul

Correlation-Aware and Gaussianity-Preserving Robust Latent Angular Watermarking for Diffusion Models

ResearchDGX agent

arXiv:2607.22386v1 Announce Type: new Abstract: Latent domain watermarking for diffusion models embeds watermarks directly into the latent prior, enjoying non-intrusiveness to model parameters and sea

CorVS+: Correspondence-Driven Association of Video Trajectories and Sensors for Identity-Aware Person Localization in Warehouses

Local AiDGX agent

arXiv:2510.26369v2 Announce Type: replace-cross Abstract: Logistics warehouses have struggled with labor shortages, but the inbound processes remain particularly human-powered. Worker location data is

Deep Convolutional Large-Margin ell_p-SVDD for Visual Anomaly Detection

TutorialsDGX agent

arXiv:2607.22212v1 Announce Type: new Abstract: Visual anomaly detection requires adaptive representations and reliable decision boundaries, particularly when anomalous training samples are scarce and

DeepUrban: Interaction-Aware Trajectory Prediction and Planning for Automated Driving by Aerial Imagery

AgentsDGX agent

arXiv:2601.10554v3 Announce Type: replace Abstract: The efficacy of autonomous driving systems hinges critically on robust prediction and planning capabilities. However, current benchmarks are impeded

Deformable Triangle Splatting: Flexible Primitives for Real-Time Radiance Field Rendering

SafetyDGX agent

arXiv:2607.22446v1 Announce Type: new Abstract: Recent radiance field methods represent scenes with 2D primitives that offer surface alignment and efficient rasterization, from Gaussian disks to trian

DM3D: Dynamic Mamba via Offset-Guided Feature Resampling for Point Cloud Understanding

Model ReleasesDGX agent

arXiv:2512.03424v4 Announce Type: replace Abstract: State Space Models (SSMs) model long token sequences of point cloud with linear complexity, but require an unordered point cloud to be serialized. E

dRAE: Representation Autoencoder with Hyper-Spherical Codes

ResearchDGX agent

arXiv:2607.22148v1 Announce Type: new Abstract: In this work, we aim to discretize the high-dimensional visual representations to bridge the gap with language models - a non-trivial challenge, as exis

EVL-MCoT: Enhanced Vision-Language Multi-CoT for Harmful Meme Detection

Model ReleasesDGX agent

arXiv:2607.22016v1 Announce Type: new Abstract: MEMEs are widely used on the internet and often carry strong elements of sarcasm or irony. Understanding their hidden meanings typically requires a join

FAIR: Feature-Augmented Implicit Regularization for AI-generated Fake Image Detection

ResearchDGX agent

arXiv:2607.22087v1 Announce Type: new Abstract: Generalization remains a critical bottleneck in AI-generated image detection. Because many modern generators are proprietary or adversarially modified,

Farmland Extent and Visible Boundary Mapping from 1 m NAIP Imagery Using Residual U-Net and Text-Prompted SAM 3 Refinement

ApplicationsDGX agent

arXiv:2607.21881v1 Announce Type: new Abstract: Agricultural field maps are often proprietary, incomplete, or outdated, yet they provide the spatial framework for crop monitoring, production accountin

Filling Before Advancing: Capability-Gap-Driven Post-Training for Scenario-Specialized Remote Sensing MLLMs

Model ReleasesDGX agent

arXiv:2607.22205v1 Announce Type: new Abstract: Remote sensing multimodal large language models (RS-MLLMs) have improved general aerial-image understanding. However, Earth observation applications req

fMRI2Face: A Full-HD fMRI-Video Dataset and Geometry-Guided Neural Decoding Framework for Dynamic Human Face Reconstruction

Model ReleasesDGX agent

arXiv:2607.22302v1 Announce Type: new Abstract: Reconstructing dynamic human faces from brain activity provides a powerful way to study how the mind perceives identity, expression, and facial motion.

Forensics Adapter: Unleashing CLIP for Generalizable Face Forgery Detection

Model ReleasesDGX agent

arXiv:2411.19715v4 Announce Type: replace Abstract: We describe Forensics Adapter, an adapter network designed to transform CLIP into an effective and generalizable face forgery detector. Although CLI

From level set evolution to threshold optimization: A grayscale level set framework for image segmentation

ResearchDGX agent

arXiv:2607.22255v1 Announce Type: new Abstract: The segmentation of multiple degradations has been a challenging problem in the field of image segmentation. Existing level set approaches commonly adop

GeoDiff-SAR: A Geometric Prior Guided Diffusion Model for SAR Image Generation

TutorialsDGX agent

arXiv:2601.03499v2 Announce Type: replace-cross Abstract: Synthetic aperture radar (SAR) image generation can mitigate data scarcity, but controllablegeneration under sparse observation angles remains

Geometric 2D Scene Graph Generation

ApplicationsDGX agent

arXiv:2607.22325v1 Announce Type: new Abstract: In production processes for consumer products, assembly instructions are essential not only for planning but also for executing the production process.

Geometry-Guided Representations for Coherent Lane and Traffic Topology Reasoning in Driving Scenes

AgentsDGX agent

arXiv:2506.13553v4 Announce Type: replace Abstract: Road topology reasoning is fundamental for autonomous driving, requiring both accurate perception of road elements and understanding of their comple

GLI-AL: A Multi-Modal Glioma MRI Label Resource with Unified Anatomy-Lesion Labels

Model ReleasesDGX agent

arXiv:2607.22135v1 Announce Type: new Abstract: Existing BraTS-GLI datasets provide a widely used benchmark for adult glioma MRI segmentation, but their task definition focuses on tumor subregions and

Hash-QNeRF: Multiresolution Hash Encoding for Quantum Neural Radiance Fields

ResearchDGX agent

arXiv:2607.21675v1 Announce Type: cross Abstract: Neural Radiance Fields (NeRF) have revolutionized novel view synthesis, yet their classical implementations remain computationally intensive for high-

Hiding Faces in Plain Sight: Defending DeepFakes by Disrupting Face Detection

ResearchDGX agent

arXiv:2412.01101v2 Announce Type: replace Abstract: Face-swapping DeepFakes have become an escalating societal concern, attracting increasing attention in recent years. To counter this, we investigate

Improving Large Vision-Language Models' Understanding for Flow Field Data

Model ReleasesDGX agent

arXiv:2507.18311v3 Announce Type: replace Abstract: Large Vision-Language Models (LVLMs) have shown impressive capabilities across a range of tasks that integrate visual and textual understanding, suc

InnoText: A Unified Model for Visual Text Generation and Editing

ResearchDGX agent

arXiv:2607.22101v1 Announce Type: new Abstract: Diffusion models have recently achieved remarkable success in high-fidelity image synthesis, yet their application to visual text generation and editing

IR275K: A Benchmark for Infrared Multi-Frame Super-Resolution Toward Efficient Remote Sensing

Model ReleasesDGX agent

arXiv:2607.22380v1 Announce Type: new Abstract: Efficient processing is becoming increasingly important in infrared remote sensing, where satellite constellations produce large volumes of observations

ISPCloak: Weaponizing ISP for Optimization-Free Physical Camouflage against Deepfake Detectors

ResearchDGX agent

arXiv:2607.21897v1 Announce Type: new Abstract: The rapid advancement of generative models has spurred the critical need to evaluate the worst-case robustness of deepfake detectors. In this paper, we

Joint Lossless Compression and Steganography for Medical Images via Large Language Models

ResearchDGX agent

arXiv:2508.01782v4 Announce Type: replace-cross Abstract: Recently, large language models (LLMs) have driven promising progress in lossless image compression. However, directly adopting existing parad

JustDepth: Real-Time Radar-Camera Depth Estimation with Single-Scan LiDAR Supervision

AgentsDGX agent

arXiv:2607.22172v1 Announce Type: new Abstract: Accurate yet low-latency depth is essential for radar-camera perception in autonomous systems. Cameras provide rich appearance but lack metric scale, wh

Latent Interpolation Learning Using Diffusion Models for Cardiac Volume Reconstruction

ResearchDGX agent

arXiv:2508.13826v4 Announce Type: replace-cross Abstract: Cardiac Magnetic Resonance (CMR) imaging is a critical tool for diagnosing and managing cardiovascular disease, yet its utility is often limit

LayoutLite: Token-Level Implicit Layout Analysis for Efficient Document OCR

SafetyDGX agent

arXiv:2607.22200v1 Announce Type: new Abstract: End-to-end OCR systems based on vision-language models have achieved strong performance in complex document OCR, but their efficiency is limited by the

Learning Adaptive Semantic Gaussian Allocation for 3D Occupancy

TutorialsDGX agent

arXiv:2607.21896v1 Announce Type: new Abstract: Semantic 3D Gaussians provide a compact representation for 3D semantic occupancy prediction by rendering semantic primitives into a voxel volume under v

LLM-Based Visual Explanation Evaluation Framework for Assessing the Explainability of Facial Skin Disease Classification Models

Model ReleasesDGX agent

arXiv:2606.16794v2 Announce Type: replace Abstract: This study proposes a domain-specific LLM-based Visual Explanation Evaluation Framework for assessing visual attention explanations in facial skin d

Local Synaptic Rules Can Implement a SIGReg Gradient Without Backpropagation

ResearchDGX agent

arXiv:2607.21622v1 Announce Type: cross Abstract: We prove that two canonical local synaptic learning rules, the potentiation arm of spike-timing-dependent plasticity (STDP^+) and homeostatic plastici

Low-Altitude Channel Multipath Prediction via Panoramic Perception and Vision-Language Model

ResearchDGX agent

arXiv:2607.21953v1 Announce Type: new Abstract: Unmanned aerial vehicle (UAV) communication is expected to support a wide range of low-altitude applications in 6G mobile networks. However, traditional

Medical-Checklist: Assessing the Comprehension of Medical Images by Multimodal Models

Model ReleasesDGX agent

arXiv:2607.21998v1 Announce Type: new Abstract: This paper introduces a new benchmark test, Medical-Checklist, for assessing medical multimodal models. The recent advancements in multimodal models hav

Moving Beyond Diversity: Visual Token Pruning as Subspace Reconstruction for Efficient VLMs

ResearchDGX agent

arXiv:2606.18681v2 Announce Type: replace Abstract: Despite their remarkable performance, Vision Language Models (VLMs) incur substantial computational overhead due to the large number of visual token

NWaaS: A Non-Intrusive and Privacy-Preserving Watermarking-as-a-Service System with Adaptive Resource Scheduling

Model ReleasesDGX agent

arXiv:2507.18036v2 Announce Type: replace-cross Abstract: Securing intellectual property (IP) in Machine Learning as a Service is critical yet challenging. While deep neural network watermarking serve

OpenNavMap: Multi-Session Appearance-Based Topometric Mapping for Scalable Visual Navigation

Model ReleasesDGX agent

arXiv:2601.12291v2 Announce Type: replace-cross Abstract: Scalable and maintainable maps are fundamental to large-scale navigation and the long-term deployment of robots in real-world environments. Ho

Optimal Transport Image Representation and Deep Covariance Alignment (CORAL) for Control Valve Stiction Detection

Model ReleasesDGX agent

arXiv:2607.22486v1 Announce Type: new Abstract: Control valve stiction is a common cause of unwanted oscillations and poor control-loop performance in industrial processes. Data-driven methods can aut

Oxygen-TryOn: Fashion-Native Foundation Model for Any-item Virtual Try-On

ResearchDGX agent

arXiv:2607.21694v1 Announce Type: new Abstract: We present Oxygen-TryOn, a unified foundation model for any-item virtual try-on. Rather than repurposing a general-purpose image editor, Oxygen-TryOn is

Physiological Signals as a Forensic Modality for Talking-Face Deepfake Detection

ResearchDGX agent

arXiv:2607.21776v1 Announce Type: cross Abstract: Talking-face (TF) deepfake generation synthesizes photore- alistic facial video from a static source image and an au- dio signal, producing forgeries

Pixels for Programs? A Cross-Provider Case Study of Input-Token Accounting for Source Code as Text and Images

Model ReleasesDGX agent

arXiv:2607.21672v1 Announce Type: cross Abstract: Long source-code contexts consume many text tokens, motivating the proposal to render code as images for vision-language models. Recent work asks whet

← Previous
1…3233343536…209
Next →