AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,618 results
Research

Handwriting Extraction and Analysis of Signature Lists in Swiss Popular Initiatives

DGX agent

arXiv:2606.05018v1 Announce Type: new Abstract: Popular initiatives and referendums are central to Swiss democracy, yet the validation of handwritten signature lists remains a labor-intensive manual p

researcharxiv-cs-cv
4 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

HD-DinoMoE: A Class-Aware Hierarchical Dual Mixture-of-Experts Network for Scleral Anomaly Segmentation in Complex Acquisition Scenarios

DGX agent

arXiv:2606.04888v1 Announce Type: new Abstract: Traditional Chinese Medicine (TCM) ocular inspection provides empirical cues for assessing scleral surface anomalies, but its clinical use remains subje

model-releasesarxiv-cs-cv
4 Jun 2026
Model Releases

Hierarchical Self-Supervised Adversarial Training for Robust Vision Models in Histopathology

DGX agent

arXiv:2503.10629v2 Announce Type: replace Abstract: Adversarial attacks pose significant challenges for vision models in critical fields like healthcare, where reliability is essential. Although adver

model-releasesarxiv-cs-cv
4 Jun 2026
Research

Hierarchical Space Partition for Surface Reconstruction

DGX agent

arXiv:2606.04891v1 Announce Type: new Abstract: Generating compact polygonal models from point clouds is a key problem in 3D vision and computer graphics. However, due to inherent limitations of LiDAR

researcharxiv-cs-cv
4 Jun 2026
Model Releases

Hyper-ICL: Attention Calibration with Hyperbolic Anchor Distillation for Multimodal In-Context Learning

DGX agent

arXiv:2606.04434v1 Announce Type: new Abstract: Multimodal In-Context Learning (ICL) has emerged as a practical inference paradigm for Multimodal Large Language Models, where a small set of interleave

model-releasesarxiv-cs-cv
4 Jun 2026
Local Ai

Identifying Gems from Roman RAPIDly

DGX agent

arXiv:2606.05103v1 Announce Type: cross Abstract: The Nancy Grace Roman Space Telescope (Roman), set for launch as early as September 2026, will conduct wide-field infrared imaging surveys with unprec

local-aiarxiv-cs-cv
4 Jun 2026
Model Releases

Imagine Before You Draw: Visual Prompt Engineering for Image Generation

DGX agent

arXiv:2606.04457v1 Announce Type: new Abstract: Incorporating visual semantic representations as an intermediate step before image generation can reduce the modeling difficulty between text and images

model-releasesarxiv-cs-cv
4 Jun 2026
Applications

Implicit Fuzzification via Bounded Noise Injection for Robust Medical Image Segmentation

DGX agent

arXiv:2606.04427v1 Announce Type: new Abstract: Image segmentation remains fundamentally limited by boundary ambiguity arising from sampling-induced information loss and inherent uncertainty in pixel-

applicationsarxiv-cs-cv
4 Jun 2026
Research

IMPose: Interactive Multi-person Pose Estimation with Dynamic Correction Propagation

DGX agent

arXiv:2606.04480v1 Announce Type: new Abstract: High-quality dynamic human pose annotation equips AI with precise motion kinematics to enable human behavior mastery, yet remains labor-intensive and ti

researcharxiv-cs-cv
4 Jun 2026
Model Releases

Impostor: An Agent-Curated Benchmark for Realistic AIGC Manipulation Localization

DGX agent

arXiv:2606.04545v1 Announce Type: new Abstract: Recent advances in generative image editing have improved the realism and controllability of localized image manipulation, raising new challenges for im

model-releasesarxiv-cs-cv
4 Jun 2026
Research

Improving Semantic Uncertainty Quantification in LVLMs with Semantic Gaussian Processes

DGX agent

arXiv:2512.14177v3 Announce Type: replace Abstract: Large Vision-Language Models (LVLMs) often produce plausible but unreliable outputs, making robust uncertainty estimation essential. Recent work on

researcharxiv-cs-cv
4 Jun 2026
Model Releases

InstantRetouch: Efficient and High-Fidelity Instruction-Guided Image Retouching with Bilateral Space

DGX agent

arXiv:2606.05071v1 Announce Type: new Abstract: Language-guided photo retouching aims to adjust color and tone while preserving geometry and texture. Recently, diffusion-based retouching shows a super

model-releasesarxiv-cs-cv
4 Jun 2026
Agents

INTACT: Ego-Guided Typed Sparse Evidence Retrieval for Heterogeneous Collaborative Perception

DGX agent

arXiv:2606.04437v1 Announce Type: new Abstract: Collaborative perception extends the perceptual range of autonomous vehicles by sharing information across agents, but heterogeneous sensors and percept

agentsarxiv-cs-cv
4 Jun 2026
Research

Intra-Modal Neighbors Never Lie: Rectifying Inter-Modal Noisy Correspondence via Graph-Based Intra-Modal Reasoning

DGX agent

arXiv:2606.04061v1 Announce Type: new Abstract: Large-scale web-harvested datasets have fueled the progress of cross-modal retrieval but inevitably suffer from noisy correspondence, which severely deg

researcharxiv-cs-cv
4 Jun 2026
Research

IRIS-GAN: Staged Specialist Detection of Deepfake Faces

DGX agent

arXiv:2606.04863v1 Announce Type: new Abstract: We introduce IRIS-GAN, a specialist forensic detector for synthetic face images under cross-generator shift. Rather than addressing universal synthetic-

researcharxiv-cs-cv
4 Jun 2026
Tutorials

J-RAS: Mutual Adaptation for Medical Image Segmentation via Contrastive Retrieval-Augmented Joint Optimization

DGX agent

arXiv:2510.09953v3 Announce Type: replace Abstract: Manual medical image segmentation by clinicians, though accurate, is time-consuming and variable across experts, whereas AI-based models automate th

tutorialsarxiv-cs-cv
4 Jun 2026
Research

Label-Efficient 3D Forest Mapping: Self-Supervised and Transfer Learning for Instance Segmentation, Semantic Segmentation, and Species Classification

DGX agent

arXiv:2511.06331v2 Announce Type: replace Abstract: Detailed structural and species information on individual tree level is increasingly important to support precision forestry, biodiversity conservat

researcharxiv-cs-cv
4 Jun 2026
Tutorials

Learning Association via Track-Detection Matching for Multi-Object Tracking

DGX agent

arXiv:2512.22105v2 Announce Type: replace Abstract: Multi-object tracking aims to maintain object identities over time by associating detections across video frames. Two dominant paradigms exist in li

tutorialsarxiv-cs-cv
4 Jun 2026
Research

MaCo-GAN: Manifold-Contrastive Adversarial Learning for Single Image Super-Resolution

DGX agent

arXiv:2606.05068v1 Announce Type: new Abstract: Conventional Generative Adversarial Networks (GANs) for Single Image Super-Resolution (SISR) often struggle with hallucinated artifacts, largely because

researcharxiv-cs-cv
4 Jun 2026
Research

MAOAM: Unified Object and Material Selection with Vision-Language Models

DGX agent

arXiv:2606.04880v1 Announce Type: new Abstract: Selection is a core operation in interactive image editing. To be practical, a user should be able to specify and disambiguate the desired selection reg

researcharxiv-cs-cv
4 Jun 2026
Safety

MATCH: Multi-faceted Adaptive Topo-Consistency for Semi-Supervised Histopathology Segmentation

DGX agent

arXiv:2510.01532v2 Announce Type: replace Abstract: In semi-supervised segmentation, capturing meaningful semantic structures from unlabeled data is essential. This is particularly challenging in hist

safetyarxiv-cs-cv
4 Jun 2026
Safety

Measuring Model Robustness via Fisher Information: Spectral Bounds, Theoretical Guarantees, and Practical Algorithms

DGX agent

arXiv:2606.04767v1 Announce Type: cross Abstract: The robustness of deep neural networks is crucial for safety-critical deployments, yet existing evaluation methods are often attack-dependent and lack

safetyarxiv-cs-cv
4 Jun 2026
Research

Med-Banana: Learning Quality-Controlled Medical Image Editing from Success-and-Failure Trajectories

DGX agent

arXiv:2511.00801v4 Announce Type: replace Abstract: Text-guided medical image editing must satisfy the requested pathology while preserving anatomy, modality-specific appearance, and clinical plausibi

researcharxiv-cs-cv
4 Jun 2026
Research

MeshFlow: Efficient Artistic Mesh Generation via MeshVAE and Flow-based Diffusion Transformer

DGX agent

arXiv:2606.04621v1 Announce Type: new Abstract: We present MeshFlow, a new method for generating artist-like 3D meshes. Current mesh generators often adopt Auto-Regressive (AR) next-token prediction,

researcharxiv-cs-cv
4 Jun 2026
Local Ai

MeshWeaver: Sparse-Voxel-Guided Surface Weaving for Autoregressive Mesh Generation

DGX agent

arXiv:2606.04688v1 Announce Type: new Abstract: Autoregressive mesh generation has gained attention by tokenizing meshes into sequences and training models in a language-modeling fashion. However, exi

local-aiarxiv-cs-cv
4 Jun 2026
Agents

MetaPoint: Unlocking Precise Spatial Control in Agentic Visual Generation

DGX agent

arXiv:2606.05031v1 Announce Type: new Abstract: Generative visual models fundamentally struggle with precise spatial control. This arises from a core disconnect: models can process textual description

agentsarxiv-cs-cv
4 Jun 2026
Tutorials

Motion-Guided Causal Disentanglement for Robust Multi-View Cine Cardiac MRI Diagnosis

DGX agent

arXiv:2606.04414v1 Announce Type: new Abstract: Multi-view cardiac magnetic resonance (CMR) imaging provides complementary anatomical information and is widely used for noninvasive disease assessment.

tutorialsarxiv-cs-cv
4 Jun 2026
Applications

Multi-Camera AR Guidance System for Surgical Instrument Handling and Assembly: Investigating Workload and Efficiency

DGX agent

arXiv:2606.04992v1 Announce Type: new Abstract: The handling and assembly of instruments during surgery imposes high cognitive demands on scrub nurses, particularly when instruments are unfamiliar. We

applicationsarxiv-cs-cv
4 Jun 2026
Safety

Optimal Transport Flow Matching by Design

DGX agent

arXiv:2606.04092v1 Announce Type: new Abstract: Flow matching models learn to transport samples from a simple prior distribution to a complex data distribution. When prior-data pairs are coupled via o

safetyarxiv-cs-cv
4 Jun 2026
Model Releases

Physics-Informed Video Generation via Mixture-of-Experts Latent Alignment

DGX agent

arXiv:2606.04737v1 Announce Type: new Abstract: Large-scale video generation models have made remarkable progress in semantic consistency and visual quality, producing videos that are increasingly coh

model-releasesarxiv-cs-cv
4 Jun 2026
Research

Pinpoint: Grounded Worldwide Image Geolocation via Cross-Source Retrieval and Reranking

DGX agent

arXiv:2606.04133v1 Announce Type: new Abstract: Image geolocation aims to estimate where a photograph was taken from its visual content. At worldwide scale, this remains challenging because visual evi

researcharxiv-cs-cv
4 Jun 2026
Safety

Plug-and-Play Diffusion Meets ADMM: Dual-Variable Coupling for Robust Medical Image Reconstruction

DGX agent

arXiv:2602.23214v2 Announce Type: replace Abstract: Plug-and-Play diffusion prior (PnPDP) frameworks have emerged as a powerful paradigm for solving imaging inverse problems by treating pretrained gen

safetyarxiv-cs-cv
4 Jun 2026
Tutorials

Prospective Dynamic 3D MRI Reconstruction via Latent-Space Motion Tracking from Single Measurement

DGX agent

arXiv:2606.04249v1 Announce Type: new Abstract: Prospective reconstruction is crucial in many clinical applications such as MRI-guided radiotherapy, which demands accurate image reconstruction and fas

tutorialsarxiv-cs-cv
4 Jun 2026
Research

PureLight: Learning Complex Luminaires with Light Tracing

DGX agent

arXiv:2606.04319v1 Announce Type: cross Abstract: We propose a neural formulation for estimating the appearance of complex luminaires. We focus on challenging luminaires with complex light transport (

researcharxiv-cs-cv
4 Jun 2026
Research

Radiomic Feature Selection Using Gradient Loss of Deep Neural Network for Lung Cancer Stage Detection

DGX agent

arXiv:2606.04453v1 Announce Type: new Abstract: Radiomics enables extraction of quantitative imaging biomarkers from medical images and has become an important tool for computer-aided cancer diagnosis

researcharxiv-cs-cv
4 Jun 2026
Agents

Recent Advances and Trends in Learning-based 3D Representations

DGX agent

arXiv:2606.04871v1 Announce Type: new Abstract: The selection of an appropriate 3D representation is a fundamental design decision that dictates the efficiency, quality, and capabilities of modern com

agentsarxiv-cs-cv
4 Jun 2026
Research

ReConFuse: Reconstruction-Error Guided Semantic Fusion for AI-Generated Video Detection

DGX agent

arXiv:2606.04706v1 Announce Type: new Abstract: AI-generated videos are becoming increasingly realistic, raising serious concerns about misinformation, content authenticity, and media trust. Reliable

researcharxiv-cs-cv
4 Jun 2026
Applications

Reflection Separation from a Single Image via Joint Latent Diffusion

DGX agent

arXiv:2606.04107v1 Announce Type: new Abstract: Single-image reflection separation is highly challenging under extreme conditions like glare or weak reflections. Existing methods often struggle to rec

applicationsarxiv-cs-cv
4 Jun 2026
Safety

Robust-LLaVA: On the Effectiveness of Large-Scale Robust Image Encoders for Multi-modal Large Language Models

DGX agent

arXiv:2502.01576v2 Announce Type: replace Abstract: Multi-modal Large Language Models (MLLMs) excel in vision-language tasks but remain vulnerable to visual adversarial perturbations that can induce h

safetyarxiv-cs-cv
4 Jun 2026
Model Releases

Robust Multi-view Clustering against Imperfect Information

DGX agent

arXiv:2606.04343v1 Announce Type: new Abstract: Real-world multi-view data always suffer from imperfect information problem, where the view-specific observations are absent (i.e., Incomplete Views, IV

model-releasesarxiv-cs-cv
4 Jun 2026
Local Ai

SBP-Net: Learning Thin Structure Reconstruction with Sliding-Box Projections

DGX agent

arXiv:2606.04251v1 Announce Type: new Abstract: Reconstructing thin 3D structures is challenging due to their sparsity, scale variation, and complex geometry. Such structures arise in a wide range of

local-aiarxiv-cs-cv
4 Jun 2026
Research

Scalable Event Cloud Network for Event-based Classification

DGX agent

arXiv:2412.20803v2 Announce Type: replace Abstract: Event cameras are biologically inspired sensors garnering significant attention from both industry and academia. Mainstream methods favor frame and

researcharxiv-cs-cv
4 Jun 2026
Model Releases

Scene-Centric Unsupervised Video Panoptic Segmentation

DGX agent

arXiv:2606.04925v1 Announce Type: new Abstract: Video panoptic segmentation (VPS) aims to jointly detect, segment, and track all objects while partitioning the video into semantically consistent regio

model-releasesarxiv-cs-cv
4 Jun 2026
Research

SharpNet: Enhancing MLPs to Represent Functions with Controlled Non-differentiability

DGX agent

arXiv:2601.19683v2 Announce Type: replace Abstract: Multi-layer perceptrons (MLPs) are a standard tool for learning and function approximation, but they inherently produce globally smooth outputs. Con

researcharxiv-cs-cv
4 Jun 2026
Model Releases

Shifting the Breaking Point of Flow Matching for Multi-Instance Editing

DGX agent

arXiv:2602.08749v3 Announce Type: replace Abstract: Flow matching models have recently emerged as an efficient alternative to diffusion, especially for text-guided image generation and editing, offeri

model-releasesarxiv-cs-cv
4 Jun 2026
Research

Spatial Artifact Coherence Determines Codec Robustness in Patch-Based rPPG

DGX agent

arXiv:2606.04198v1 Announce Type: new Abstract: Remote photoplethysmography (rPPG) achieves low heart-rate error on uncompressed benchmarks yet is deployed over compressed video channels in telehealth

researcharxiv-cs-cv
4 Jun 2026
Research

Spatially Grounded Concept Bottleneck Models via Part-Factorized Attention

DGX agent

arXiv:2606.04364v1 Announce Type: new Abstract: Concept bottleneck models (CBMs) predict a layer of human-named attributes before predicting a class, which makes their decisions auditable. On fine-gra

researcharxiv-cs-cv
4 Jun 2026
Applications

StrokeTimer: Robust Representation Learning for Ischemic Stroke Onset-Time Estimation from Non-contrast CT

DGX agent

arXiv:2606.04722v1 Announce Type: new Abstract: Ischemic stroke is a major global disease. Treatment decisions are highly time-sensitive, as eligibility for reperfusion therapies relies on the interva

applicationsarxiv-cs-cv
4 Jun 2026
← Previous
1…120121122123124…263
Next →