AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,515 results
Applications

VEMamba: Efficient Isotropic Reconstruction of Volume Electron Microscopy with Axial-Lateral Consistent Mamba

DGX agent

arXiv:2603.00887v2 Announce Type: replace Abstract: Volume Electron Microscopy (VEM) is crucial for 3D tissue imaging but often produces anisotropic data with poor axial resolution, hindering visualiz

applicationsarxiv-cs-cv
28 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

ViDS: Video Diffusion Shader using 3D Face Tracking

DGX agent

arXiv:2607.24124v1 Announce Type: new Abstract: We introduce ViDS, a Video Diffusion Shader that leverages 3D face tracking for expressive and identity-preserving portrait animation. We first reconstr

researcharxiv-cs-cv
28 Jul 2026
Research

VIP: Finding Important People in Images

DGX agent

arXiv:1502.05678v3 Announce Type: replace Abstract: People preserve memories of events such as birthdays, weddings, or vacations by capturing photos, often depicting groups of people. Invariably, some

researcharxiv-cs-cv
28 Jul 2026
Tutorials

VIPER: Visual In-Context Physics Reasoning for Physically Plausible Video Generation

DGX agent

arXiv:2607.23472v1 Announce Type: new Abstract: Modern video generation models can synthesize visually compelling and temporally coherent clips, yet controlling their physical behavior remains difficu

tutorialsarxiv-cs-cv
28 Jul 2026
Applications

Visual Information Extraction from Documents via Classification-Guided Large Vision-Language Models

DGX agent

arXiv:2607.22723v1 Announce Type: new Abstract: Visual information extraction (VIE) from visually rich documents remains challenging due to high layout variability and real-world impairments. Existing

applicationsarxiv-cs-cv
28 Jul 2026
Safety

Visual Token Compression Enhances Robustness of MLLMs

DGX agent

arXiv:2607.22716v1 Announce Type: new Abstract: In this paper, we show for the first time that visual token pruning enhances the robustness of Multimodal Large Language Models (MLLMs), mitigating vuln

safetyarxiv-cs-cv
28 Jul 2026
Research

WaveZip: Wavelet-Driven Space-Time Decoupling for Video Token Condensation

DGX agent

arXiv:2607.23265v1 Announce Type: new Abstract: Existing Large Vision-Language Models (LVLMs) struggle with long-form video understanding due to the quadratic computational cost of visual tokens. Whil

researcharxiv-cs-cv
28 Jul 2026
Research

Weakly Supervised Instance-Level Gleason Pattern Estimation Using Primary and Secondary Labels

DGX agent

arXiv:2607.23594v1 Announce Type: new Abstract: In prostate cancer histopathology, the Gleason Score is determined by the most frequent (Primary) and second most frequent (Secondary) Gleason patterns

researcharxiv-cs-cv
28 Jul 2026
Local Ai

WGDnet: Wishart-guided Geometric-aware Deep Network for PolSAR Image Classification

DGX agent

arXiv:2607.23638v1 Announce Type: new Abstract: Polarimetric Synthetic Aperture Radar (PolSAR) classification underpins all-weather Earth observation. Conventional Wishart methods depend on rigid hand

local-aiarxiv-cs-cv
28 Jul 2026
Agents

What Can I Edit? Open-Ended Strategy Discovery and the Emotion Editability Landscape

DGX agent

arXiv:2607.23920v1 Announce Type: new Abstract: Emotional image editing requires more than applying affective filters or modifying predefined visual factors: an effective edit must identify what a par

agentsarxiv-cs-cv
28 Jul 2026
Model Releases

When Less Is More: A Controlled Benchmark of Lightweight CNNs for Satellite Land-Cover Segmentation on DeepGlobe

DGX agent

arXiv:2607.23024v1 Announce Type: new Abstract: High-resolution satellite imagery is the backbone of good land-cover classification, and without that, environmental monitoring, urban planning, and sus

model-releasesarxiv-cs-cv
28 Jul 2026
Model Releases

When Low CER is Not Enough: An Analysis of Hallucinations in Vision-Language OCR Systems on Historical Uruguayan Documents

DGX agent

arXiv:2607.24077v1 Announce Type: new Abstract: Optical Character Recognition (OCR) is a key component in the digitization of historical archives. Recently, Vision-Language Models (VLMs) have emerged

model-releasesarxiv-cs-cv
28 Jul 2026
Research

Which Workloads Belong in Orbit? A Workload-First Framework for Orbital Data Centers Using Semantic Abstraction

DGX agent

arXiv:2603.20317v2 Announce Type: replace Abstract: Space-based compute is becoming plausible as launch costs fall and data-intensive AI workloads grow. This paper proposes a workload-centric framewor

researcharxiv-cs-cv
28 Jul 2026
Model Releases

XMatchAD: A Cross-Modal Matching Perspective on Reconstruction-based Anomaly Detection

DGX agent

arXiv:2607.23658v1 Announce Type: new Abstract: The remarkable success of reconstruction-based methods in Unsupervised Anomaly Detection (UAD) lies in their ability to identify and localize anomalies

model-releasesarxiv-cs-cv
28 Jul 2026
Research

A Dual Path Framework with Hotspot Guided Fusion for Three Dimensional CT to PET Synthesis in Head and Neck Cancer

DGX agent

arXiv:2607.21800v1 Announce Type: cross Abstract: 18F-FDG PET/CT plays a central role in staging, treatment planning, and response assessment for head and neck cancer by providing functional informati

researcharxiv-cs-cv
27 Jul 2026
Research

A Framework for Individual Tree Growth Reconstruction Using Multi-Platform Laser Scanning

DGX agent

arXiv:2607.22129v1 Announce Type: new Abstract: Accurate tree-level forest monitoring using laser scanning data requires reliable tree delineation, consistent tree correspondence across multitemporal

researcharxiv-cs-cv
27 Jul 2026
Research

A Smooth Phase-Separation Model for Weak-Boundary Segmentation of Homogeneous Structures

DGX agent

arXiv:2607.22053v1 Announce Type: new Abstract: Segmentation of adjacent structures with similar intensity distributions remains a challenging problem in image analysis, particularly when object bound

researcharxiv-cs-cv
27 Jul 2026
Agents

Active few-shot segmentation by reinforcing data selection

DGX agent

arXiv:2607.22371v1 Announce Type: new Abstract: Few-shot learning enables medical image segmentation models to adapt to new tasks using only a small number of labelled examples. However, adaptation pe

agentsarxiv-cs-cv
27 Jul 2026
Model Releases

Adaptive Contrast Enhancement and Optimised Feature Matching for RootSIFT-Based Palm-Vein Recognition

DGX agent

arXiv:2607.16077v2 Announce Type: replace Abstract: Palm-vein recognition is a highly secure biometric modality due to the uniqueness and subcutaneous nature of vein patterns. However, low contrast in

model-releasesarxiv-cs-cv
27 Jul 2026
Safety

AgentHOI: Multi-Agent Reasoning for Human-Object-Interaction Video Generation via Implicit Representation Alignment

DGX agent

arXiv:2607.22241v1 Announce Type: new Abstract: Recent advances in video diffusion models have spurred interest in human-object interaction (HOI) video generation, which demands fine-grained control o

safetyarxiv-cs-cv
27 Jul 2026
Tutorials

Alleviating Regional Shortcuts for Few-Shot Class-Incremental Learning

DGX agent

arXiv:2607.22072v1 Announce Type: new Abstract: Few-shot class-incremental learning (FSCIL) aims to incrementally learn novel classes with only a few samples while avoiding forgetting base classes. Ho

tutorialsarxiv-cs-cv
27 Jul 2026
Research

Atlas 2 -- Foundation models for clinical deployment

DGX agent

arXiv:2601.05148v2 Announce Type: replace Abstract: Pathology foundation models substantially advanced the possibilities in computational pathology --- yet tradeoffs in terms of performance, robustnes

researcharxiv-cs-cv
27 Jul 2026
Research

Automatic Map Density Selection for Locally-Performant Visual Place Recognition

DGX agent

arXiv:2602.21473v3 Announce Type: replace Abstract: A key challenge in translating Visual Place Recognition (VPR) from the lab to long-term deployment is ensuring a priori that a system can meet user-

researcharxiv-cs-cv
27 Jul 2026
Model Releases

Be Consistent! Enhancing Robust Visual Reasoning in LVLMs with Consistency Constraints

DGX agent

arXiv:2607.21722v1 Announce Type: new Abstract: While Large Vision-Language Models (LVLMs) exhibit strong perceptual capabilities, they remain vulnerable in visual reasoning tasks. Existing benchmarks

model-releasesarxiv-cs-cv
27 Jul 2026
Local Ai

Bowel Obstruction Detection and Localization on Abdominal CT with Deep Learning

DGX agent

arXiv:2607.22173v1 Announce Type: new Abstract: Bowel obstruction is a common and potentially life-threatening gastrointestinal condition. In the face of rising diagnostic workloads, the automated dia

local-aiarxiv-cs-cv
27 Jul 2026
Agents

CARA: Concept-Aware Risk Attention for Interpretable Collision Anticipation

DGX agent

arXiv:2607.22494v1 Announce Type: cross Abstract: Collision anticipation in autonomous driving requires not only accurate early warnings but also interpretable reasoning about what risk factors are be

agentsarxiv-cs-cv
27 Jul 2026
Model Releases

CARDIAG: A Dense Segment Classification Benchmark of Deep Learning Architectures for Coronary Angiography

DGX agent

arXiv:2607.22139v1 Announce Type: new Abstract: Accurate pixel-level classification of coronary angiograms is critical for cardiovascular disease assessment, yet the field lacks standardized evaluatio

model-releasesarxiv-cs-cv
27 Jul 2026
Research

CARE: Anti-entanglement Ultrasound Image Segmentation via Channel-Aware Region Extrication

DGX agent

arXiv:2508.13899v2 Announce Type: replace Abstract: Accurate ultrasound image segmentation is fundamentally challenged by target-context entanglement, where lesion cues are easily mixed with surroundi

researcharxiv-cs-cv
27 Jul 2026
Research

Class-Balanced Softmax: A Bayes Theory-Based Method for Long-Tailed Recognition

DGX agent

arXiv:2607.22258v1 Announce Type: cross Abstract: Deep learning models using traditional softmax classifiers have achieved remarkable success in various classification tasks. However, their performanc

researcharxiv-cs-cv
27 Jul 2026
Applications

Closing the Loop: Training-Free Revisit Consistency for Autoregressive Generative Rendering

DGX agent

arXiv:2607.21848v1 Announce Type: new Abstract: Recent conditional video generation models have shown promising potentials to transform 3D engine renderings, such as depth maps and untextured geometry

applicationsarxiv-cs-cv
27 Jul 2026
Safety

Color Me Correctly: Bridging Perceptual Color Spaces and Text Embeddings for Improved Diffusion Generation

DGX agent

arXiv:2509.10058v2 Announce Type: replace Abstract: Accurate color alignment in text-to-image (T2I) generation is critical for applications such as fashion, product visualization, and interior design,

safetyarxiv-cs-cv
27 Jul 2026
Safety

CommandLM: Data driven behavior level descriptor for ego vehicles

DGX agent

arXiv:2607.22078v1 Announce Type: new Abstract: As autonomous driving systems move toward real-world deployment, interpretable, behavior-level decision-making is essential for safety, trust, and regul

safetyarxiv-cs-cv
27 Jul 2026
Research

Correlation-Aware and Gaussianity-Preserving Robust Latent Angular Watermarking for Diffusion Models

DGX agent

arXiv:2607.22386v1 Announce Type: new Abstract: Latent domain watermarking for diffusion models embeds watermarks directly into the latent prior, enjoying non-intrusiveness to model parameters and sea

researcharxiv-cs-cv
27 Jul 2026
Local Ai

CorVS+: Correspondence-Driven Association of Video Trajectories and Sensors for Identity-Aware Person Localization in Warehouses

DGX agent

arXiv:2510.26369v2 Announce Type: replace-cross Abstract: Logistics warehouses have struggled with labor shortages, but the inbound processes remain particularly human-powered. Worker location data is

local-aiarxiv-cs-cv
27 Jul 2026
Tutorials

Deep Convolutional Large-Margin ell_p-SVDD for Visual Anomaly Detection

DGX agent

arXiv:2607.22212v1 Announce Type: new Abstract: Visual anomaly detection requires adaptive representations and reliable decision boundaries, particularly when anomalous training samples are scarce and

tutorialsarxiv-cs-cv
27 Jul 2026
Agents

DeepUrban: Interaction-Aware Trajectory Prediction and Planning for Automated Driving by Aerial Imagery

DGX agent

arXiv:2601.10554v3 Announce Type: replace Abstract: The efficacy of autonomous driving systems hinges critically on robust prediction and planning capabilities. However, current benchmarks are impeded

agentsarxiv-cs-cv
27 Jul 2026
Safety

Deformable Triangle Splatting: Flexible Primitives for Real-Time Radiance Field Rendering

DGX agent

arXiv:2607.22446v1 Announce Type: new Abstract: Recent radiance field methods represent scenes with 2D primitives that offer surface alignment and efficient rasterization, from Gaussian disks to trian

safetyarxiv-cs-cv
27 Jul 2026
Model Releases

DM3D: Dynamic Mamba via Offset-Guided Feature Resampling for Point Cloud Understanding

DGX agent

arXiv:2512.03424v4 Announce Type: replace Abstract: State Space Models (SSMs) model long token sequences of point cloud with linear complexity, but require an unordered point cloud to be serialized. E

model-releasesarxiv-cs-cv
27 Jul 2026
Research

dRAE: Representation Autoencoder with Hyper-Spherical Codes

DGX agent

arXiv:2607.22148v1 Announce Type: new Abstract: In this work, we aim to discretize the high-dimensional visual representations to bridge the gap with language models - a non-trivial challenge, as exis

researcharxiv-cs-cv
27 Jul 2026
Model Releases

EVL-MCoT: Enhanced Vision-Language Multi-CoT for Harmful Meme Detection

DGX agent

arXiv:2607.22016v1 Announce Type: new Abstract: MEMEs are widely used on the internet and often carry strong elements of sarcasm or irony. Understanding their hidden meanings typically requires a join

model-releasesarxiv-cs-cv
27 Jul 2026
Research

FAIR: Feature-Augmented Implicit Regularization for AI-generated Fake Image Detection

DGX agent

arXiv:2607.22087v1 Announce Type: new Abstract: Generalization remains a critical bottleneck in AI-generated image detection. Because many modern generators are proprietary or adversarially modified,

researcharxiv-cs-cv
27 Jul 2026
Applications

Farmland Extent and Visible Boundary Mapping from 1 m NAIP Imagery Using Residual U-Net and Text-Prompted SAM 3 Refinement

DGX agent

arXiv:2607.21881v1 Announce Type: new Abstract: Agricultural field maps are often proprietary, incomplete, or outdated, yet they provide the spatial framework for crop monitoring, production accountin

applicationsarxiv-cs-cv
27 Jul 2026
Model Releases

Filling Before Advancing: Capability-Gap-Driven Post-Training for Scenario-Specialized Remote Sensing MLLMs

DGX agent

arXiv:2607.22205v1 Announce Type: new Abstract: Remote sensing multimodal large language models (RS-MLLMs) have improved general aerial-image understanding. However, Earth observation applications req

model-releasesarxiv-cs-cv
27 Jul 2026
Model Releases

fMRI2Face: A Full-HD fMRI-Video Dataset and Geometry-Guided Neural Decoding Framework for Dynamic Human Face Reconstruction

DGX agent

arXiv:2607.22302v1 Announce Type: new Abstract: Reconstructing dynamic human faces from brain activity provides a powerful way to study how the mind perceives identity, expression, and facial motion.

model-releasesarxiv-cs-cv
27 Jul 2026
Model Releases

Forensics Adapter: Unleashing CLIP for Generalizable Face Forgery Detection

DGX agent

arXiv:2411.19715v4 Announce Type: replace Abstract: We describe Forensics Adapter, an adapter network designed to transform CLIP into an effective and generalizable face forgery detector. Although CLI

model-releasesarxiv-cs-cv
27 Jul 2026
Research

From level set evolution to threshold optimization: A grayscale level set framework for image segmentation

DGX agent

arXiv:2607.22255v1 Announce Type: new Abstract: The segmentation of multiple degradations has been a challenging problem in the field of image segmentation. Existing level set approaches commonly adop

researcharxiv-cs-cv
27 Jul 2026
Tutorials

GeoDiff-SAR: A Geometric Prior Guided Diffusion Model for SAR Image Generation

DGX agent

arXiv:2601.03499v2 Announce Type: replace-cross Abstract: Synthetic aperture radar (SAR) image generation can mitigate data scarcity, but controllablegeneration under sparse observation angles remains

tutorialsarxiv-cs-cv
27 Jul 2026
Applications

Geometric 2D Scene Graph Generation

DGX agent

arXiv:2607.22325v1 Announce Type: new Abstract: In production processes for consumer products, assembly instructions are essential not only for planning but also for executing the production process.

applicationsarxiv-cs-cv
27 Jul 2026
← Previous
1…4041424344…261
Next →