AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlog
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,618 results
Local Ai

Biological Spatial Priors Regularize Foundation Model Representations for Cross-Site MSI Generalization in Colorectal Cancer

DGX agent

arXiv:2605.02660v1 Announce Type: cross Abstract: Predicting microsatellite instability (MSI) status from routine hematoxylin and eosin (H&E) whole slide images (WSIs) offers a practical alternative t

local-aiarxiv-cs-cv
5 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Research

BoostDream: Efficient Refining for High-Quality Text-to-3D Generation from Multi-View Diffusion

DGX agent

arXiv:2401.16764v5 Announce Type: replace Abstract: Witnessing the evolution of text-to-image diffusion models, significant strides have been made in text-to-3D generation. Currently, two primary para

researcharxiv-cs-cv
5 May 2026
Model Releases

Boosting Multimodal Remote Sensing Image Classification with Transformer-based Heterogeneously Salient Graph Representation

DGX agent

arXiv:2311.10320v3 Announce Type: replace Abstract: Data collected by different modalities can provide a wealth of complementary information, such as hyperspectral image (HSI) to offer rich spectral-s

model-releasesarxiv-cs-cv
5 May 2026
Research

Breaking the Resolution Barrier: Arbitrary-resolution Deep Image Steganography Framework

DGX agent

arXiv:2601.15739v2 Announce Type: replace Abstract: Deep image steganography (DIS) has achieved significant results in capacity and invisibility. However, current paradigms enforce the secret image to

researcharxiv-cs-cv
5 May 2026
Model Releases

BRITE: A Benchmark for Reliable and Interpretable T2V Evaluation on Implausible Scenarios

DGX agent

arXiv:2605.00873v1 Announce Type: cross Abstract: The rapid advancement of photorealistic Text-to-Video (T2V) generation brings in an urgent need for up-to-date evaluation methods. Existing benchmarks

model-releasesarxiv-cs-cv
5 May 2026
Applications

CADFit: Precise Mesh-to-CAD Program Generation with Hybrid Optimization

DGX agent

arXiv:2605.01171v1 Announce Type: new Abstract: Despite recent progress, recovering parametric CAD construction sequences from geometric input, such as meshes or point clouds, is a key challenge for d

applicationsarxiv-cs-cv
5 May 2026
Applications

CADFS: A Big CAD Program Dataset and Framework for Computer-Aided Design with Large Language Models

DGX agent

arXiv:2605.01925v1 Announce Type: new Abstract: We introduce CADFS, a data-centric framework that enables large vision-language models to generate complex CAD design histories. Existing generative CAD

applicationsarxiv-cs-cv
5 May 2026
Research

Certified vs. Empirical Adversarial Robust-ness via Hybrid Convolutions with Attention Stochasticity

DGX agent

arXiv:2605.01519v1 Announce Type: new Abstract: We introduce Hybrid Convolutions with Attention Stochasticity (HyCAS), an adversarial defense that narrows the long-standing gap between provable robust

researcharxiv-cs-cv
5 May 2026
Model Releases

CEZSAR: A Contrastive Embedding Method for Zero-Shot Action Recognition

DGX agent

arXiv:2605.01165v1 Announce Type: new Abstract: This paper proposes a novel Zero-Shot Action Recognition~(ZSAR) method based on contrastive learning. In ZSAR, we aim to classify examples from classes

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

CGFformer: Cluster-Guidance Frequency Transformer for Pansharpening

DGX agent

arXiv:2605.01490v1 Announce Type: new Abstract: Pansharpening aims to generate high-resolution multispectral (HRMS) images by fusing low-resolution multispectral (LRMS) images with high-resolution pan

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

Channel-Level Relation to Attentive Aggregation with Neighborhood-Homogeneity Constraint for Point Cloud Analysis

DGX agent

arXiv:2605.02357v1 Announce Type: new Abstract: In 3D point cloud understanding, the core challenge lies in accurately capturing discriminative features within complex neighborhoods, which directly af

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

Chart-FR1: Visual Focus-Driven Fine-Grained Reasoning on Dense Charts

DGX agent

arXiv:2605.01882v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have shown considerable potential in chart understanding and reasoning tasks. However, they still struggle with

model-releasesarxiv-cs-cv
5 May 2026
Local Ai

CHASE: Competing Hypotheses for Ambiguity-Aware Selective Prediction

DGX agent

arXiv:2605.01346v1 Announce Type: new Abstract: Standard selective prediction methods typically estimate uncertainty from the output of a single predictive branch. While effective for general uncertai

local-aiarxiv-cs-cv
5 May 2026
Model Releases

Checkerboard: A Simple, Effective, Efficient and Learning-free Clean Label Backdoor Attack with Low Poisoning Budget

DGX agent

arXiv:2605.01298v1 Announce Type: cross Abstract: Backdoor attacks threaten the deep learning supply chain by poisoning a small fraction of the training data so that a model behaves normally on clean

model-releasesarxiv-cs-cv
5 May 2026
Tutorials

Closed-Loop LLM Discovery of Non-Standard Channel Priors in Vision Models

DGX agent

arXiv:2601.08517v2 Announce Type: replace Abstract: Channel-configuration search, the optimization of layer specifications such as channel widths in deep neural networks, presents a combinatorial chal

tutorialsarxiv-cs-cv
5 May 2026
Model Releases

CNN-based Multi-In-Multi-Out Model for Efficient Spatiotemporal Prediction

DGX agent

arXiv:2605.01277v1 Announce Type: new Abstract: Recently, Convolutional Neural Network (CNN) or Transformer architecture based models have been proposed to overcome the limitations of Recurrent Neural

model-releasesarxiv-cs-cv
5 May 2026
Safety

Colinearity Decay: Training Quantization-Friendly ViTs with Outlier Decay

DGX agent

arXiv:2605.01330v1 Announce Type: new Abstract: Low-bit quantization is a practical route for efficiently deploying vision Transformers, yet activation outliers complicate fully quantized deployment.

safetyarxiv-cs-cv
5 May 2026
Safety

Combining Facial Videos and Biosignals for Stress Estimation During Driving

DGX agent

arXiv:2601.04376v3 Announce Type: replace Abstract: Reliable stress recognition is critical in applications such as medical monitoring and safety-critical systems, including real-world driving. While

safetyarxiv-cs-cv
5 May 2026
Research

Comparative Evaluation of Convolutional and Transformer-Based Detectors for Automated Weed Detection in Precision Agriculture

DGX agent

arXiv:2605.00908v1 Announce Type: new Abstract: This paper presents a comparative evaluation of convolutional and transformer-based object detection architectures for early weed detection in realistic

researcharxiv-cs-cv
5 May 2026
Research

Compression as Adaptation: Implicit Visual Representation with Diffusion Foundation Models

DGX agent

arXiv:2603.07615v2 Announce Type: replace-cross Abstract: Modern visual generative models acquire rich visual knowledge through large-scale training, yet existing visual representations (such as pixel

researcharxiv-cs-cv
5 May 2026
Research

Continual Few-shot Adaptation for Synthetic Fingerprint Detection

DGX agent

arXiv:2603.14632v2 Announce Type: replace Abstract: The quality and realism of synthetically generated fingerprint images have increased significantly over the past decade fueled by advancements in ge

researcharxiv-cs-cv
5 May 2026
Research

Continuous quantification of viral plaque dynamics using ultra-large-area label-free imaging enables rapid antiviral susceptibility testing

DGX agent

arXiv:2605.01738v1 Announce Type: cross Abstract: The plaque reduction assay (PRA) remains the gold standard for antiviral susceptibility testing, evaluating drug potency by measuring reductions in pl

researcharxiv-cs-cv
5 May 2026
Research

Correlates of Image Memorability in Vision Encoders: Activations, Attention Entropy, Patch Uniformity and Autoencoder Losses

DGX agent

arXiv:2509.01453v2 Announce Type: replace Abstract: Images vary in how memorable they are to humans. Inspired by findings from cognitive science and computer vision, we explore correlates of image mem

researcharxiv-cs-cv
5 May 2026
Research

Cross-Domain Adversarial Augmentation: Stabilizing GANs for Medical and Handwriting Data Scarcity

DGX agent

arXiv:2605.01815v1 Announce Type: new Abstract: Generative Adversarial Networks (GANs) offer a pragmatic route to mitigate data scarcity in vision tasks. We study generative augmentation across two lo

researcharxiv-cs-cv
5 May 2026
Model Releases

Cross-Language Learning within Arabic Script for Low-Resource HTR

DGX agent

arXiv:2605.02089v1 Announce Type: new Abstract: Handwritten Text Recognition (HTR) under limited labeled data remains a challenging problem, particularly for Arabic-script languages. Although modern s

model-releasesarxiv-cs-cv
5 May 2026
Research

Cross-Polarization Fusion of VV AND VH SAR Observations for Improved Flood Mapping

DGX agent

arXiv:2605.02153v1 Announce Type: new Abstract: Synthetic Aperture Radar (SAR) imagery is widely used for flood monitoring due to its all-weather and day-night imaging capability. However, flood mappi

researcharxiv-cs-cv
5 May 2026
Tutorials

Cross-Scale Pretraining: Enhancing Self-Supervised Learning for Low-Resolution Satellite Imagery for Semantic Segmentation

DGX agent

arXiv:2601.12964v2 Announce Type: replace Abstract: Self-supervised pretraining in remote sensing is mostly done using mid-spatial resolution (MR) image datasets due to their high availability. Given

tutorialsarxiv-cs-cv
5 May 2026
Research

CSGuard: Toward Forgery-Resistant Watermarking in Diffusion Models via Compressed Sensing Constraint

DGX agent

arXiv:2605.01479v1 Announce Type: new Abstract: Latent-based diffusion model watermarking embeds watermarks into generated images' latent space to enable content attribution, offering a training-free

researcharxiv-cs-cv
5 May 2026
Safety

CUE: Concept-Aware Multi-Label Expansion to Mitigate Concept Confusion in Long-Tailed Learning

DGX agent

arXiv:2605.01309v1 Announce Type: new Abstract: Long-tailed distributions are common in real-world recognition tasks, where a few head classes have many samples while most tail classes have very few.

safetyarxiv-cs-cv
5 May 2026
Safety

Decision Boundary-aware Generation for Long-tailed Learning

DGX agent

arXiv:2605.01468v1 Announce Type: new Abstract: Long-tailed data bias decision boundaries toward head classes and degrade tail class accuracy. Diffusion-based generative augmentation address this prob

safetyarxiv-cs-cv
5 May 2026
Model Releases

Decompose and Recompose: Reasoning New Skills from Existing Abilities for Cross-Task Robotic Manipulation

DGX agent

arXiv:2605.01448v1 Announce Type: cross Abstract: Cross-task generalization is a core challenge in open-world robotic manipulation, and the key lies in extracting transferable manipulation knowledge f

model-releasesarxiv-cs-cv
5 May 2026
Tutorials

Decouple and Cache: KV Cache Construction for Streaming Video Understanding

DGX agent

arXiv:2605.01858v1 Announce Type: new Abstract: Streaming video understanding requires processing unbounded video streams with limited memory and computation, posing two key challenges. First, continu

tutorialsarxiv-cs-cv
5 May 2026
Model Releases

Deep neural networks with Fisher vector encoding for medical image classification

DGX agent

arXiv:2605.01667v1 Announce Type: new Abstract: Orderless encoding methods have shown to improve Convolutional Neural Networks (CNNs) for image classification in the context of limited availability of

model-releasesarxiv-cs-cv
5 May 2026
Tutorials

Degradation-Aware Adaptive Context Gating for Unified Image Restoration

DGX agent

arXiv:2605.01236v1 Announce Type: new Abstract: Unified image restoration using a single model often faces task interference due to diverse degradations. To address this, we propose DACG-IR (Degradati

tutorialsarxiv-cs-cv
5 May 2026
Model Releases

Developing a Strong Pre-Trained Base Model for Plant Leaf Disease Classification

DGX agent

arXiv:2605.01283v1 Announce Type: new Abstract: Plants, crops and their yields are essential to our very existence, but diseases and pests cause large losses every year. As such it is vital to ensure

model-releasesarxiv-cs-cv
5 May 2026
Research

DGS-Net: Distillation-Guided Gradient Surgery for CLIP Fine-Tuning in AI-Generated Image Detection

DGX agent

arXiv:2511.13108v3 Announce Type: replace Abstract: The rapid progress of generative models such as GANs and diffusion models has led to the widespread proliferation of AI-generated images, raising co

researcharxiv-cs-cv
5 May 2026
Research

Dino-NestedUNet: Unlocking Foundation Vision Encoders for Pathology Tumor Bulk Segmentation via Dense Decoding

DGX agent

arXiv:2605.00894v1 Announce Type: new Abstract: Vision foundation models (VFMs), such as DINOv3, provide rich semantic representations that are promising for computational pathology. However, many cur

researcharxiv-cs-cv
5 May 2026
Applications

DIPLI: Deep Image Prior Lucky Imaging for Blind Astronomical Image Restoration

DGX agent

arXiv:2503.15984v3 Announce Type: replace Abstract: Modern image restoration and super-resolution methods utilize deep learning due to its superior performance compared to traditional algorithms. Howe

applicationsarxiv-cs-cv
5 May 2026
Research

DirectEdit: Step-Level Accurate Inversion for Flow-Based Image Editing

DGX agent

arXiv:2605.02417v1 Announce Type: new Abstract: With recent advancements in large-scale pre-trained text-to-image (T2I) models, training-free image editing methods have demonstrated remarkable success

researcharxiv-cs-cv
5 May 2026
Safety

Disciplined Diffusion: Text-to-Image Diffusion Model against NSFW Generation

DGX agent

arXiv:2605.01113v1 Announce Type: new Abstract: Text-to-image (T2I) diffusion models have the ability to build high-quality pictures from text prompts, but they pose safety concerns because they can g

safetyarxiv-cs-cv
5 May 2026
Research

Disentangled Anatomy-Disease Diffusion (DADD) for Controllable Ulcerative Colitis Progression Synthesis

DGX agent

arXiv:2605.01848v1 Announce Type: new Abstract: Synthesizing longitudinal medical images at controllable disease stages while preserving patient-specific anatomy is hindered by the entanglement of pat

researcharxiv-cs-cv
5 May 2026
Research

DissolveStereo: Coarse Depth Injection for Zero-Shot Stereo Video Generation

DGX agent

arXiv:2411.14295v3 Announce Type: replace Abstract: Generating high-quality stereo videos requires consistent depth perception and temporal coherence across frames. Despite advances in image and video

researcharxiv-cs-cv
5 May 2026
Research

Divergence is Uncertainty: A Closed-Form Posterior Covariance for Flow Matching

DGX agent

arXiv:2605.00941v1 Announce Type: cross Abstract: Flow matching has become a leading framework for generative modeling, but quantifying the uncertainty of its samples remains an open problem. Existing

researcharxiv-cs-cv
5 May 2026
Safety

Divide and Conquer: Decoupled Representation Alignment for Multimodal World Models

DGX agent

arXiv:2605.01896v1 Announce Type: new Abstract: Emerging multi-modal world models attempt to jointly generate videos across diverse modalities (e.g., RGB, depth, and mask), yet they fail to fully expl

safetyarxiv-cs-cv
5 May 2026
Applications

Does it Really Count? Assessing Semantic Grounding in Text-Guided Class-Agnostic Counting

DGX agent

arXiv:2605.02752v1 Announce Type: new Abstract: Open-world text-guided class-agnostic counting (CAC) has emerged as a flexible paradigm for counting arbitrary object classes by using natural language

applicationsarxiv-cs-cv
5 May 2026
Research

DP-SfM: Dual-Pixel Structure-from-Motion without Scale Ambiguity

DGX agent

arXiv:2605.01852v1 Announce Type: new Abstract: Multi-view 3D reconstruction, namely, structure-from-motion followed by multi-view stereo, is a fundamental component of 3D computer vision. In general,

researcharxiv-cs-cv
5 May 2026
Model Releases

Dual-branch Robust Unlearnable Examples

DGX agent

arXiv:2605.01718v1 Announce Type: new Abstract: Unlearnable examples (UEs) aim to compromise model training by injecting imperceptible perturbations to clean samples. However, existing UE schemes exhi

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

DuFal: Dual-Frequency-Aware Learning for High-Fidelity Extremely Sparse-view CBCT Reconstruction

DGX agent

arXiv:2601.15416v2 Announce Type: replace Abstract: Sparse-view Cone-Beam Computed Tomography reconstruction from limited X-ray projections remains a challenging problem in medical imaging due to the

model-releasesarxiv-cs-cv
5 May 2026
← Previous
1…196197198199200…263
Next →