AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,515 results
Model Releases

Personalization Toolkit: Training Free Personalization of Large Vision Language Models

DGX agent

arXiv:2502.02452v4 Announce Type: replace Abstract: Personalization of Large Vision-Language Models (LVLMs) involves customizing models to recognize specific users or object instances and to generate

model-releasesarxiv-cs-cv
29 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Tutorials

Personalized Cross-Modal Emotional Correlation Learning for Speech-Preserving Facial Expression Manipulation

DGX agent

arXiv:2604.25255v1 Announce Type: new Abstract: Speech-preserving facial expression manipulation (SPFEM) aims to enhance human expressiveness without altering mouth movements tied to the original spee

tutorialsarxiv-cs-cv
29 Apr 2026
Research

PhyloSDF: Phylogenetically-Conditioned Neural Generation of 3D Skull Morphology via Residual Flow Matching

DGX agent

arXiv:2604.25371v1 Announce Type: cross Abstract: Generating novel, biologically plausible three-dimensional morphological structures is a fundamental challenge in computational evolutionary biology,

researcharxiv-cs-cv
29 Apr 2026
Model Releases

PortraVec: Image-Based Portrait Vectorization with Text-Guided Manipulation

DGX agent

arXiv:2410.04182v2 Announce Type: replace Abstract: While portrait sketch generation is a special task in sketch synthesis, most existing methods are pixel-based, limiting their interpretability and e

model-releasesarxiv-cs-cv
29 Apr 2026
Research

Power Foam: Unifying Real-Time Differentiable Ray Tracing and Rasterization

DGX agent

arXiv:2604.24994v1 Announce Type: cross Abstract: We introduce a differentiable 3D representation that unifies the ray tracing capabilities of foam-based ray tracing with the efficiency of modern rast

researcharxiv-cs-cv
29 Apr 2026
Hardware

Practical exposure correction via compensation

DGX agent

arXiv:2212.14245v2 Announce Type: replace Abstract: In computer vision, correcting the exposure level is a fundamental task for enhancing the visual quality of observations with inappropriate lightnes

hardwarearxiv-cs-cv
29 Apr 2026
Research

Prefill-Time Intervention for Mitigating Hallucination in Large Vision-Language Models

DGX agent

arXiv:2604.25642v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) have achieved remarkable progress in visual-textual understanding, yet their reliability is critically undermined b

researcharxiv-cs-cv
29 Apr 2026
Model Releases

QB-LIF: Learnable-Scale Quantized Burst Neurons for Efficient SNNs

DGX agent

arXiv:2604.25688v1 Announce Type: new Abstract: Binary spike coding enables sparse and event-driven computation in spiking neural networks (SNNs), yet its 1-bit-per-timestep representation fundamental

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

QCalEval: Benchmarking Vision-Language Models for Quantum Calibration Plot Understanding

DGX agent

arXiv:2604.25884v1 Announce Type: cross Abstract: Quantum computing calibration depends on interpreting experimental data, and calibration plots provide the most universal human-readable representatio

model-releasesarxiv-cs-cv
29 Apr 2026
Research

Quantum-Inspired Robust and Scalable SAR Object Classification

DGX agent

arXiv:2604.25755v1 Announce Type: cross Abstract: SAR image classification naturally has to deal with huge noise and a high dynamic range particularly requiring robust classification models. Additiona

researcharxiv-cs-cv
29 Apr 2026
Local Ai

RABC-Net: Reliability-Aware Annotation-Free Skin Lesion Segmentation for Low-Resource Dermoscopy

DGX agent

arXiv:2604.05594v2 Announce Type: replace Abstract: Pixel-level annotation is costly in low-resource dermoscopy. We present RABC-Net, a reliability-aware annotation-free segmentation system that combi

local-aiarxiv-cs-cv
29 Apr 2026
Research

Rapid tracking through strongly scattering media with physics-informed neuromorphic speckle analysis

DGX agent

arXiv:2604.25310v1 Announce Type: new Abstract: This work addresses the critical problem of tracking fast-moving objects through strongly scattering media in a low-light environment. Different from ex

researcharxiv-cs-cv
29 Apr 2026
Safety

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models

DGX agent

arXiv:2604.25636v1 Announce Type: new Abstract: Unified multimodal models (UMMs) integrate visual understanding and generation within a single framework. For text-to-image (T2I) tasks, this unified ca

safetyarxiv-cs-cv
29 Apr 2026
Model Releases

Representation Paradigms in AI-based 3D Radiological Image Reconstruction: A Systematic Review

DGX agent

arXiv:2504.11349v3 Announce Type: replace Abstract: The demand for high-quality medical imaging in clinical practice and assisted diagnosis has made 3D image reconstruction in radiological imaging a k

model-releasesarxiv-cs-cv
29 Apr 2026
Safety

ResetEdit: Precise Text-guided Editing of Generated Image via Resettable Starting Latent

DGX agent

arXiv:2604.25128v1 Announce Type: new Abstract: Recent advances in diffusion models have enabled high-quality image generation, leading to increasing demand for post-generation editing that modifies l

safetyarxiv-cs-cv
29 Apr 2026
Safety

ReSim: Reliable World Simulation for Autonomous Driving

DGX agent

arXiv:2506.09981v2 Announce Type: replace Abstract: How can we reliably simulate future driving scenarios under a wide range of ego driving behaviors? Recent driving world models, developed exclusivel

safetyarxiv-cs-cv
29 Apr 2026
Applications

Robust Deepfake Detection: Mitigating Spatial Attention Drift via Calibrated Complementary Ensembles

DGX agent

arXiv:2604.25889v1 Announce Type: new Abstract: Current deepfake detection models achieve state-of-the-art performance on pristine academic datasets but suffer severe spatial attention drift under rea

applicationsarxiv-cs-cv
29 Apr 2026
Applications

Robustness Evaluation of a Foundation Segmentation Model Under Simulated Domain Shifts in Abdominal CT: Implications for Health Digital Twin Deployment

DGX agent

arXiv:2604.25685v1 Announce Type: cross Abstract: Foundation segmentation models such as the Segment Anything Model (SAM) have demonstrated strong generalization across natural images; however, their

applicationsarxiv-cs-cv
29 Apr 2026
Research

SaliencyDecor: Enhancing Neural Network Interpretability through Feature Decorrelation

DGX agent

arXiv:2604.25315v1 Announce Type: new Abstract: Gradient-based saliency methods are widely used to interpret deep neural networks, yet they often produce noisy and unstable explanations that poorly al

researcharxiv-cs-cv
29 Apr 2026
Local Ai

SAMe: A Semantic Anatomy Mapping Engine for Robotic Ultrasound

DGX agent

arXiv:2604.25646v1 Announce Type: new Abstract: Robotic ultrasound has advanced local image-driven control, contact regulation, and view optimization, yet current systems lack the anatomical understan

local-aiarxiv-cs-cv
29 Apr 2026
Model Releases

SARE: Sample-wise Adaptive Reasoning for Training-free Fine-grained Visual Recognition

DGX agent

arXiv:2603.17729v3 Announce Type: replace Abstract: Recent advances in Large Vision-Language Models (LVLMs) have enabled training-free Fine-Grained Visual Recognition (FGVR). However, effectively expl

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

SARU: A Shadow-Aware and Removal Unified Framework for Remote Sensing Images with New Benchmarks

DGX agent

arXiv:2604.25432v1 Announce Type: new Abstract: Shadows are a prevalent problem in remote sensing imagery (RSI), degrading visual quality and severely limiting the performance of downstream tasks like

model-releasesarxiv-cs-cv
29 Apr 2026
Local Ai

Scalable Secure Biometric Authentication without Auxiliary Identifiers

DGX agent

arXiv:2604.25071v1 Announce Type: cross Abstract: The prevalence of biometric authentication has been on the rise due to its ease of use and elimination of weak passwords. To date, most biometric auth

local-aiarxiv-cs-cv
29 Apr 2026
Model Releases

SecureScan: An AI-Driven Multi-Layer Framework for Malware and Phishing Detection Using Logistic Regression and Threat Intelligence Integration

DGX agent

arXiv:2602.10750v2 Announce Type: replace-cross Abstract: The growing sophistication of modern malware and phishing campaigns has diminished the effectiveness of traditional signature-based intrusion

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

Self-DACE++: Robust Low-Light Enhancement via Efficient Adaptive Curve Estimation

DGX agent

arXiv:2604.25367v1 Announce Type: new Abstract: In this paper, we present Self-DACE++, an improved unsupervised and lightweight framework for Low-Light Image Enhancement (LLIE), building upon our prev

model-releasesarxiv-cs-cv
29 Apr 2026
Research

Semantic-aware Random Convolution and Source Matching for Domain Generalization in Medical Image Segmentation

DGX agent

arXiv:2512.01510v3 Announce Type: replace Abstract: We tackle the challenging problem of single-source domain generalization (DG) for medical image segmentation, where we train a network on one domain

researcharxiv-cs-cv
29 Apr 2026
Research

ShapeY: A Principled Framework for Measuring Shape Recognition Capacity via Nearest-Neighbor Matching

DGX agent

arXiv:2604.25065v1 Announce Type: new Abstract: Object recognition (OR) in humans relies heavily on shape cues and the ability to recognize objects across varying 3D viewpoints. Unlike humans, deep ne

researcharxiv-cs-cv
29 Apr 2026
Model Releases

SIEVES: Selective Prediction Generalizes through Visual Evidence Scoring

DGX agent

arXiv:2604.25855v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) achieve ever-stronger performance on visual-language tasks. Even as traditional visual question answering bench

model-releasesarxiv-cs-cv
29 Apr 2026
Research

Sketch2Arti: Sketch-based Articulation Modeling of CAD Objects

DGX agent

arXiv:2604.25781v1 Announce Type: new Abstract: Articulation modeling aims to infer movable parts and their motion parameters for a 3D object, enabling interactive animation, simulation, and shape edi

researcharxiv-cs-cv
29 Apr 2026
Model Releases

Soft-TransFormers for Continual Learning

DGX agent

arXiv:2411.16073v3 Announce Type: replace-cross Abstract: Inspired by the Well-initialized Lottery Ticket Hypothesis (WLTH), we introduce Soft-Transformer (Soft-TF), a parameter-efficient framework fo

model-releasesarxiv-cs-cv
29 Apr 2026
Research

Splatent: Splatting Diffusion Latents for Novel View Synthesis

DGX agent

arXiv:2512.09923v2 Announce Type: replace Abstract: Radiance field representations have recently been explored in the latent space of VAEs that are commonly used by diffusion models. This direction of

researcharxiv-cs-cv
29 Apr 2026
Model Releases

Subjective Portrait Region Cropping in Landscape Videos with Temporal Annotation Smoothing

DGX agent

arXiv:2604.24947v1 Announce Type: new Abstract: With the rise of mobile video consumption on diverse handheld display resolutions and orientation modes, altering videos to aspect ratios poses challeng

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

SynMotion: Semantic-Visual Adaptation for Motion Customized Video Generation

DGX agent

arXiv:2506.23690v2 Announce Type: replace Abstract: Diffusion-based video motion customization facilitates the acquisition of human motion representations from a few video samples, while achieving arb

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

Task-Driven Prompt Learning: A Joint Framework for Multi-modal Cloud Removal and Segmentation

DGX agent

arXiv:2601.12052v2 Announce Type: replace Abstract: Optical remote sensing imagery is indispensable for Earth observation, yet persistent cloud occlusion limits its downstream utility. Most cloud remo

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

The Forensic Cost of Watermark Removal

DGX agent

arXiv:2604.25491v1 Announce Type: new Abstract: Current watermark removal methods are evaluated on two axes: attack success rate and perceptual quality. We show this is insufficient. While state-of-th

model-releasesarxiv-cs-cv
29 Apr 2026
Tutorials

The Surprising Effectiveness of Canonical Knowledge Distillation for Semantic Segmentation

DGX agent

arXiv:2604.25530v1 Announce Type: new Abstract: Recent knowledge distillation (KD) methods for semantic segmentation introduce increasingly complex hand-crafted objectives, yet are typically evaluated

tutorialsarxiv-cs-cv
29 Apr 2026
Model Releases

The Thinking Pixel: Recursive Sparse Reasoning in Multimodal Diffusion Latents

DGX agent

arXiv:2604.25299v1 Announce Type: new Abstract: Diffusion models have achieved success in high-fidelity data synthesis, yet their capacity for more complex, structured reasoning like text following ta

model-releasesarxiv-cs-cv
29 Apr 2026
Research

TopoMamba: Topology-Aware Scanning and Fusion for Segmenting Heterogeneous Medical Visual Media

DGX agent

arXiv:2604.25545v1 Announce Type: new Abstract: Visual state-space models (SSMs) have shown strong potential for medical image segmentation, yet their effectiveness is often limited by two practical i

researcharxiv-cs-cv
29 Apr 2026
Research

Towards Robust Deep Learning-based Rumex Obtusifolius Detection from Drone Images

DGX agent

arXiv:2604.25316v1 Announce Type: new Abstract: Domain adaptation (DA) addresses the challenge of transferring a machine learning model trained on a source domain to a target domain with a different d

researcharxiv-cs-cv
29 Apr 2026
Tutorials

Towards Seamless Lunar Mosaics: Deep Radiometric Normalization for Cross-Sensor Orbital Imagery Using Chandrayaan-2 TMC Data

DGX agent

arXiv:2604.25208v1 Announce Type: new Abstract: Radiometric inconsistencies remain a major challenge in generating seamless lunar mosaics from multi-mission orbital imagery due to variability in illum

tutorialsarxiv-cs-cv
29 Apr 2026
Hardware

UltraGS: Real-Time Physically-Decoupled Gaussian Splatting for Ultrasound Novel View Synthesis

DGX agent

arXiv:2511.07743v3 Announce Type: replace Abstract: Ultrasound imaging is a cornerstone of non-invasive clinical diagnostics, yet its limited field of view poses challenges for novel view synthesis. W

hardwarearxiv-cs-cv
29 Apr 2026
Tutorials

UniSER: A Foundation Model for Unified Soft Effects Removal

DGX agent

arXiv:2511.14183v3 Announce Type: replace Abstract: Digital images are often degraded by soft effects such as lens flare, haze, shadows, and reflections, which reduce aesthetics even though the underl

tutorialsarxiv-cs-cv
29 Apr 2026
Applications

VibeToken: Scaling 1D Image Tokenizers and Autoregressive Models for Dynamic Resolution Generations

DGX agent

arXiv:2604.24885v1 Announce Type: new Abstract: We introduce an efficient, resolution-agnostic autoregressive (AR) image synthesis approach that generalizes to arbitrary resolutions and aspect ratios,

applicationsarxiv-cs-cv
29 Apr 2026
Tutorials

ViPO: Visual Preference Optimization at Scale

DGX agent

arXiv:2604.24953v1 Announce Type: new Abstract: While preference optimization is crucial for improving visual generative models, how to effectively scale this paradigm remains largely unexplored. Curr

tutorialsarxiv-cs-cv
29 Apr 2026
Safety

VISION-SLS: Safe Perception-Based Control from Learned Visual Representations via System Level Synthesis

DGX agent

arXiv:2604.24894v1 Announce Type: cross Abstract: We propose VISION-SLS, a method for nonlinear output-feedback control from high-resolution RGB images which provides robust constraint satisfaction gu

safetyarxiv-cs-cv
29 Apr 2026
Research

Vision SmolMamba: Spike-Guided Token Pruning for Energy-Efficient Spiking State-Space Vision Models

DGX agent

arXiv:2604.25570v1 Announce Type: new Abstract: Spiking Transformers have shown strong potential for long-range visual modeling through spike-driven self-attention. However, their quadratic token inte

researcharxiv-cs-cv
29 Apr 2026
Model Releases

When the Forger Is the Judge: GPT-Image-2 Cannot Recognize Its Own Faked Documents

DGX agent

arXiv:2604.25213v1 Announce Type: new Abstract: OpenAI's GPT-Image-2 has effectively erased the visual boundary between authentic and AI-edited document images: a single number on a receipt can be rep

model-releasesarxiv-cs-cv
29 Apr 2026
Tutorials

2D Pre-Training for 3D Pose Estimation

DGX agent

arXiv:2604.22830v1 Announce Type: new Abstract: Pre-training is a general method that is used in a range of deep learning tasks. By first training a model on one task, and then further training on the

tutorialsarxiv-cs-cv
28 Apr 2026
← Previous
1…209210211212213…261
Next →