AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,515 results
Model Releases

Think Sparse, Predict Dense: Continuous Thought Machines for Image Super-Resolution

DGX agent

arXiv:2607.18856v1 Announce Type: new Abstract: Continuous Thought Machines introduce an internal temporal dimension in which neuron-level histories and synchronization-derived representations evolve

model-releasesarxiv-cs-cv
23 Jul 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Research

Timeripple: Accelerating vDiTs by Understanding the Spatio-Temporal Correlations in Latent Space

DGX agent

arXiv:2511.12035v2 Announce Type: replace-cross Abstract: The recent surge in video generation has shown the growing demand for high-quality video synthesis using large vision models. Existing video g

researcharxiv-cs-cv
23 Jul 2026
Model Releases

Toward a Vision-Language Foundation Model for Medical Data: Multimodal Dataset and Benchmarks for Vietnamese PET/CT Report Generation

DGX agent

arXiv:2509.24739v4 Announce Type: replace Abstract: Vision-Language Foundation Models (VLMs), trained on large-scale multimodal datasets, have driven significant advances in Artificial Intelligence (A

model-releasesarxiv-cs-cv
23 Jul 2026
Research

Toward Seasonal Guidelines for Robust Deep-Learning Sentinel-2 Building Detection in Different Area Types

DGX agent

arXiv:2607.19994v2 Announce Type: new Abstract: Sentinel-2 imagery offers open access, global coverage, and frequent revisit times, making it attractive for practical building mapping at scale; howeve

researcharxiv-cs-cv
23 Jul 2026
Research

Trace: A Taxonomy-Guided Environment for Multidomain Visual Reasoning

DGX agent

arXiv:2607.19790v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has substantially improved language-model reasoning, yet its extension to vision-language models r

researcharxiv-cs-cv
23 Jul 2026
Model Releases

Trusted Multi-View Deep Learning Classification of Fetal Congenital Heart Disease with Feature-level and Decision-level Fusion

DGX agent

arXiv:2606.15265v2 Announce Type: replace Abstract: Congenital heart disease (CHD) refers to the abnormal anatomical structure caused by the abnormal development of the heart and great vessels during

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

Universality Reconsidered: Rethinking the Validation of Foundation Models for General-Purpose 3D Medical Segmentation

DGX agent

arXiv:2602.07643v2 Announce Type: replace Abstract: Foundation models have emerged as a transformative paradigm in 3D medical imaging, with the promise of unified quantitative analysis across diverse

model-releasesarxiv-cs-cv
23 Jul 2026
Research

UVFaceFusion: Fast Multi-view Topologically Consistent Face Reconstruction in the Wild via UV-space Neural Fusion

DGX agent

arXiv:2607.18798v1 Announce Type: new Abstract: Reconstructing high-fidelity facial geometry with an assigned topology is essential for digital avatar creation and animation, yet existing automated me

researcharxiv-cs-cv
23 Jul 2026
Research

Vera: Identity-Faithful Human Subject-to-Video Generation

DGX agent

arXiv:2607.20247v1 Announce Type: new Abstract: Subject-to-video (S2V) generation has made substantial progress in preserving reference subjects across diverse categories, yet generic subject consiste

researcharxiv-cs-cv
23 Jul 2026
Research

VQ-Transplant: Efficient VQ-Module Integration for Pre-trained Visual Tokenizers

DGX agent

arXiv:2607.19575v1 Announce Type: new Abstract: Vector Quantization (VQ) underpins modern discrete visual tokenization. However, training quantization modules for state-of-the-art VQ-based models requ

researcharxiv-cs-cv
23 Jul 2026
Research

WanSong v1.0 Technical Report

DGX agent

arXiv:2607.14749v3 Announce Type: replace-cross Abstract: Music generation foundation models have recently attracted significant industry attention. However, achieving efficient generation and high-fi

researcharxiv-cs-cv
23 Jul 2026
Agents

WASABI: Whole-graph Assignment-based Stabilizer for lAne topology By Inter-frame tracking

DGX agent

arXiv:2607.19781v1 Announce Type: new Abstract: Autonomous driving requires understanding the road as a graph of drivable lanes and their connectivity, beyond the ego lane alone, to follow routes thro

agentsarxiv-cs-cv
23 Jul 2026
Safety

Wave2Body: Rethinking mmWave Human Pose Estimation as Radar-to-Body Token Translation

DGX agent

arXiv:2607.18875v1 Announce Type: new Abstract: Millimeter-wave (mmWave) radar enables privacy-friendly human sensing, but its sparse point clouds are physical measurements of view-dependent electroma

safetyarxiv-cs-cv
23 Jul 2026
Research

Wavefront Parallelization for Efficient Learned Image Compression

DGX agent

arXiv:2607.19082v1 Announce Type: cross Abstract: Autoregressive context models are foundational for learned image compression,but they suffer from slow serial inference. Existing acceleration methods

researcharxiv-cs-cv
23 Jul 2026
Tutorials

Weakly Supervised Pathology-Informed Representation Learning for PET-Based Content Retrieval of Intra-Tumour Heterogeneity

DGX agent

arXiv:2607.18762v1 Announce Type: new Abstract: We propose a weakly supervised 18FFDG PET representation-learning framework for content based medical image retrieval, using H&E derived information dur

tutorialsarxiv-cs-cv
23 Jul 2026
Safety

WearWow: Native 2K Multi-Garment Virtual Try-On via Adaptive Token Packing and Preference Alignment

DGX agent

arXiv:2607.19923v1 Announce Type: new Abstract: Synthesizing native 2K multi-garment virtual try-on is a formidable frontier in digital fashion, critically bottlenecked by two fundamental limitations:

safetyarxiv-cs-cv
23 Jul 2026
Model Releases

WHU-PCPR: A cross-platform heterogeneous point cloud dataset for place recognition in complex urban scenes

DGX agent

arXiv:2601.06442v2 Announce Type: replace Abstract: Point Cloud-based Place Recognition (PCPR) demonstrates considerable potential in applications such as autonomous driving, robot localization and na

model-releasesarxiv-cs-cv
23 Jul 2026
Applications

ZeroSplat: Generalized Referring Segmentation in 3D Gaussian Splatting

DGX agent

arXiv:2607.18801v1 Announce Type: new Abstract: Recent advancements in 3D Gaussian Splatting (3DGS) have enabled language-guided scene understanding. However, existing Referring 3D Gaussian Splatting

applicationsarxiv-cs-cv
23 Jul 2026
Model Releases

2D Rotary Position Embedding for Scene Text Recognition with Transformers

DGX agent

arXiv:2607.13458v1 Announce Type: new Abstract: Scene Text Recognition (STR) remains challenging due to the diversity of text appearances, including curvature, rotation, and perspective distortion. Re

model-releasesarxiv-cs-cv
16 Jul 2026
Model Releases

A Comparative Evaluation of Large Vision-Language Models for 2D Object Detection under SOTIF Conditions

DGX agent

arXiv:2601.22830v2 Announce Type: replace Abstract: Reliable environmental perception remains one of the main obstacles for safe operation of automated vehicles. Safety of the Intended Functionality (

model-releasesarxiv-cs-cv
16 Jul 2026
Tutorials

A Masked Autoencoder Approach to Unsupervised Steel Surface Defect Recognition

DGX agent

arXiv:2607.13178v1 Announce Type: new Abstract: Automated visual inspection of steel surface defects is a recurring quality control task in which labeled defect data is scarce and costly to obtain, wh

tutorialsarxiv-cs-cv
16 Jul 2026
Research

A novel unsupervised machine learning strategy to handle multimodal cardiac PET/MRI data

DGX agent

arXiv:2607.13936v1 Announce Type: new Abstract: Arrhythmogenic left ventricular cardiomyopathy is a genetic myocardial disease difficult to diagnose due to the lack of gold standard criteria. Simultan

researcharxiv-cs-cv
16 Jul 2026
Research

A Space-Time Transformer for Precipitation Nowcasting

DGX agent

arXiv:2511.11090v3 Announce Type: replace Abstract: Until recently, numerical weather prediction (NWP) models have stood rivalless in operational forecasting despite a few limitations. Namely, physica

researcharxiv-cs-cv
16 Jul 2026
Local Ai

Active Learning for Efficient Annotation of Surgical Videos with Weak Supervision

DGX agent

arXiv:2607.13237v1 Announce Type: new Abstract: Precise spatial-temporal annotation of laparoscopic videos is time-consuming and requires expert knowledge. We propose a human-in-the-loop knowledge acq

local-aiarxiv-cs-cv
16 Jul 2026
Research

AffectFlow-DINO: Uncertainty-Aware Multi-Task Affect Estimation via Conditional Rectified Flow

DGX agent

arXiv:2607.13250v1 Announce Type: new Abstract: We present extbf{AffectFlow-DINO}, a multi-task learning system for the 11th ABAW challenge that extends a standard deterministic architecture with a co

researcharxiv-cs-cv
16 Jul 2026
Model Releases

AnomExpert: Identifying and Selecting Anatomical Planes for Prenatal Ultrasound Anomaly Diagnosis

DGX agent

arXiv:2607.13409v1 Announce Type: new Abstract: Life-limiting congenital anomalies require accurate prenatal diagnosis for appropriate clinical decision-making. Prenatal ultrasound (US) examinations i

model-releasesarxiv-cs-cv
16 Jul 2026
Safety

AspectCLIP: Optimizing CLIP Representation Space via Aspect-Guided Consistency Regularization

DGX agent

arXiv:2607.13805v1 Announce Type: new Abstract: Contrastive Language-Image Pretraining learns a shared representation space through large-scale contrastive learning. However, existing methods that enf

safetyarxiv-cs-cv
16 Jul 2026
Applications

Attention-Free and Lightweight Token Reduction for Efficient Vision-Language Models

DGX agent

arXiv:2607.13500v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have achieved strong performance in multimodal understanding, yet remain challenging to deploy on resource-constrained edg

applicationsarxiv-cs-cv
16 Jul 2026
Research

Attentive multilayer fusion for vision transformers

DGX agent

arXiv:2601.09322v2 Announce Type: replace Abstract: With the rise of large-scale foundation models, efficiently adapting them to downstream tasks remains a central challenge. Linear probing, which fre

researcharxiv-cs-cv
16 Jul 2026
Research

Audio-Text Cross-Attention with Psycholinguistic Support Features for Ambivalence/Hesitancy Recognition

DGX agent

arXiv:2607.13345v1 Announce Type: new Abstract: We present an audio-text system for the Ambivalence/Hesitancy Video Recognition Challenge of the 11th ABAW Competition. The method excludes visual frame

researcharxiv-cs-cv
16 Jul 2026
Hardware

Bake It Till You Make It: Ultrafast Spatial Texture-Atlas Splatting

DGX agent

arXiv:2607.13808v1 Announce Type: new Abstract: Recent extensions of 3D Gaussian Splatting (3DGS) capture fine color details using hash-grid-based appearance parameterization but incur high computatio

hardwarearxiv-cs-cv
16 Jul 2026
Model Releases

BenthiCat: An opti-acoustic dataset for advancing benthic classification and habitat mapping

DGX agent

arXiv:2510.04876v3 Announce Type: replace Abstract: Benthic habitat mapping is fundamental for understanding marine ecosystems, guiding conservation efforts, and supporting sustainable resource manage

model-releasesarxiv-cs-cv
16 Jul 2026
Model Releases

Beyond Description: Cognitively Benchmarking Fine-Grained Action for Embodied Agents

DGX agent

arXiv:2511.18685v4 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) show promising results as decision-making engines for embodied agents operating in complex, physical enviro

model-releasesarxiv-cs-cv
16 Jul 2026
Local Ai

BiLoG-Net: A Bi-Context Location-Guided Network for Breast Mass Segmentation and Malignancy Classification in Mammography

DGX agent

arXiv:2607.10188v2 Announce Type: replace Abstract: Breast cancer remains the most commonly diagnosed malignancy among women worldwide, yet accurate detection and characterization of breast masses in

local-aiarxiv-cs-cv
16 Jul 2026
Research

Bring Music The Horizon: Music-Driven 360^irc Video Generation

DGX agent

arXiv:2607.13471v1 Announce Type: new Abstract: Music visualization offers a powerful way to enhance listeners' understanding and experience of music by translating auditory signals into visual forms.

researcharxiv-cs-cv
16 Jul 2026
Safety

C-Norm: Cell-Distribution Normalization Enables Precision Recognition of Medical-Cell Image

DGX agent

arXiv:2607.13116v1 Announce Type: new Abstract: ThinPrep Cytologic Test (TCT) enables early cervical cancer screening, but manual reading is time-consuming and yields inconsistent diagnostic results a

safetyarxiv-cs-cv
16 Jul 2026
Model Releases

Calibrated Closed-Form Uncertainty for Radiative Gaussian Splatting in Sparse-View CT

DGX agent

arXiv:2607.13682v1 Announce Type: new Abstract: Radiative Gaussian splatting has made sparse-view CT reconstruction fast, but existing methods output point estimates with no notion of where the recons

model-releasesarxiv-cs-cv
16 Jul 2026
Model Releases

CASA-SDF: Curriculum-Aware Spatial Adaptation with Curvature-Guided Density for Neural Implicit Surface Reconstruction

DGX agent

arXiv:2607.13492v1 Announce Type: new Abstract: Neural implicit representations have emerged as a powerful paradigm for 3D reconstruction. However, high-fidelity indoor surface reconstruction remains

model-releasesarxiv-cs-cv
16 Jul 2026
Research

CF-Net: Conflict Fusion with Speaker Normalisation and Certainty Weighting for Ambivalence/Hesitancy Recognition

DGX agent

arXiv:2607.13976v1 Announce Type: new Abstract: Detecting ambivalence and hesitancy (AH) in unconstrained video is challenging because the target signal is inherently ambiguous and expressed through s

researcharxiv-cs-cv
16 Jul 2026
Research

CLIP-Guided Label-Free Discriminative Region Scoring for Fine-Grained Classification

DGX agent

arXiv:2607.13437v1 Announce Type: new Abstract: Recent vision models such as CLIP and SAM enable training-free segmentation and semantic encoding for fine-grained classification. A common approach is

researcharxiv-cs-cv
16 Jul 2026
Research

ClusIR: Towards Cluster-Guided All-in-One Image Restoration

DGX agent

arXiv:2512.10948v2 Announce Type: replace Abstract: All-in-One Image Restoration (AiOIR) aims to recover high-quality images from diverse degradations within a unified framework. However, existing met

researcharxiv-cs-cv
16 Jul 2026
Agents

Cyclone: Diffusion Model for Cycle-Consistent Weather Editing from Unpaired Driving Data

DGX agent

arXiv:2607.13927v1 Announce Type: new Abstract: Reliable perception under diverse weather conditions remains a major challenge for autonomous driving systems. A common strategy to improve robustness i

agentsarxiv-cs-cv
16 Jul 2026
Research

Delving into the Temporal Challenges of Unified Video Protection Against Image-to-Video and Fine-Tuning-based Customization

DGX agent

arXiv:2607.13336v1 Announce Type: new Abstract: Recent diffusion-based video generation models have enabled high-quality personalized video customization through both tuning-based pipelines, which fin

researcharxiv-cs-cv
16 Jul 2026
Model Releases

Detector Confidence Signals Presence Rather Than Occlusion in Cluttered Manipulation

DGX agent

arXiv:2607.13361v1 Announce Type: new Abstract: Occlude a named object until about an eighth of it remains visible, and an open-vocabulary detector's confidence that the object is present barely chang

model-releasesarxiv-cs-cv
16 Jul 2026
Local Ai

Differentiable Polarized Path Tracing

DGX agent

arXiv:2607.13265v1 Announce Type: new Abstract: Physically based differentiable rendering has proven to be a powerful tool for inverse rendering problems (e.g., 3D reconstruction, reflectance estimati

local-aiarxiv-cs-cv
16 Jul 2026
Research

DiffGI: Differentiable Geometry Images for High-Fidelity Thin-Shell 3D Generation

DGX agent

arXiv:2607.13365v1 Announce Type: new Abstract: Existing 3D generative models predominantly rely on implicit volumetric representations, which enforce watertight topology and struggle to represent thi

researcharxiv-cs-cv
16 Jul 2026
Model Releases

DNA: Dual-stage Native Attribution for Generated Image Source Tracing

DGX agent

arXiv:2607.13685v1 Announce Type: new Abstract: The rapid evolution of image generation has produced numerous within-family variants, making source-model attribution of suspect images increasingly imp

model-releasesarxiv-cs-cv
16 Jul 2026
Research

DP-BOA: Dirichlet-Process Birth-or-Assign for On-the-Fly Category Discovery

DGX agent

arXiv:2607.13504v1 Announce Type: new Abstract: On-the-fly category discovery requires deciding for each incoming test sample whether to assign it to an existing category or spawn a new one. Existing

researcharxiv-cs-cv
16 Jul 2026
← Previous
1…4849505152…261
Next →