AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,515 results
Research

Pixel-Space Diffusion Transformers

DGX agent

arXiv:2607.17585v2 Announce Type: replace Abstract: Latent diffusion models (LDMs) enable efficient high-resolution image synthesis by denoising in a VAE-compressed latent space. However, fixed visual

researcharxiv-cs-cv
23 Jul 2026
Model Releases

Point Ladder Tuning: Parameter-Efficient Hierarchical Adaptation for 3D Point Cloud Understanding

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2607.19171v1 Announce Type: new Abstract: Fine-tuning pre-trained point-cloud backbones typically updates all parameters, resulting in substantial computation and memory overhead. More important

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

Point-Selection Fine-Tuning Framework for Robust Point Cloud Classification

DGX agent

arXiv:2607.19711v1 Announce Type: new Abstract: Noisy and corrupted points can substantially degrade point cloud recognition performance, especially under challenging corruption settings. In particula

model-releasesarxiv-cs-cv
23 Jul 2026
Research

Pointing-Based Object Recognition

DGX agent

arXiv:2603.15403v2 Announce Type: replace Abstract: This paper presents a comprehensive pipeline for recognizing objects targeted by human pointing gestures using RGB images. As human-robot interactio

researcharxiv-cs-cv
23 Jul 2026
Local Ai

PoseIDON: 6DoF Pose Estimation with Foundation Model Features for Marine Sediment Burial Mapping

DGX agent

arXiv:2506.10386v2 Announce Type: replace Abstract: The burial state of anthropogenic objects on the seafloor provides insight into localized sedimentation dynamics and is also critical for assessing

local-aiarxiv-cs-cv
23 Jul 2026
Applications

Posterior Samplings are Missing Modalities Generators for Medical Image Translation

DGX agent

arXiv:2607.18763v1 Announce Type: new Abstract: Magnetic resonance imaging comes in various modality contrasts that provide complementary anatomical and pathological information. Complete multimodal a

applicationsarxiv-cs-cv
23 Jul 2026
Model Releases

PRiSM: Prototype Regularization for Few-Shot VLMs

DGX agent

arXiv:2607.17820v2 Announce Type: replace Abstract: Training-free few-shot adaptation methods have gained significant attention recently in the context of Vision-language Models (VLMs). Yet, current b

model-releasesarxiv-cs-cv
23 Jul 2026
Safety

Privileged Lesion-Context Relational Distillation for Mask-Free Skin Lesion Classification

DGX agent

arXiv:2607.18773v1 Announce Type: new Abstract: Accurate skin lesion classification can benefit from lesion segmentation masks, but requiring masks or an auxiliary segmentation model during inference

safetyarxiv-cs-cv
23 Jul 2026
Model Releases

Rarity-Aware Discrete Diffusion with Spatially Consistent Decoding for Photo-Realistic Image Super-Resolution

DGX agent

arXiv:2607.17612v2 Announce Type: replace Abstract: Continuous diffusion models have become the dominant paradigm for photo-realistic image Super-Resolution (SR), but they typically formulate reconstr

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

Read or Ignore? A Unified Benchmark for Typographic-Attack Robustness and Text Recognition in Vision-Language Models

DGX agent

arXiv:2512.11899v2 Announce Type: replace Abstract: Large vision-language models (LVLMs) are vulnerable to typographic attacks, where misleading text inserted into an image can override visual underst

model-releasesarxiv-cs-cv
23 Jul 2026
Local Ai

Real-Time EEG Cap Electrode Detection for Guided Point-of-Care Placement

DGX agent

arXiv:2607.20142v1 Announce Type: new Abstract: We present a two-stage vision system that detects EEG cap electrodes in a live webcam stream and validates their anatomical placement in real time. A si

local-aiarxiv-cs-cv
23 Jul 2026
Model Releases

Recti-Q: Feature-Space Rectification for Out-of-Distribution-Robust Quantized Perception in Edge Robotics

DGX agent

arXiv:2607.18540v1 Announce Type: new Abstract: Robotic perception pipelines increasingly rely on large vision backbones deployed on SWaP-constrained edge platforms, making post-training quantization

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

ReFace: Reorganizing Facial Spatiotemporal Representations for Improved Pain Assessment

DGX agent

arXiv:2607.19722v1 Announce Type: new Abstract: Automatic pain assessment from facial video remains challenging due to the spatial heterogeneity of pain-related facial cues. This study proposes ReFace

model-releasesarxiv-cs-cv
23 Jul 2026
Applications

Reliability-Aware 3D Geometric Injection for Universal Person Re-identification

DGX agent

arXiv:2607.18863v1 Announce Type: new Abstract: Universal person re-identification (ReID) aims to retrieve pedestrian identities across diverse real-world scenarios, including severe occlusions, cloth

applicationsarxiv-cs-cv
23 Jul 2026
Local Ai

RIM: A Retrieval-In-Matching Framework for Cross-Domain Global Visual Localization of UAVs

DGX agent

arXiv:2607.20116v1 Announce Type: new Abstract: Global visual localization of unmanned aerial vehicles (UAVs) using remote-sensing reference maps has attracted increasing attention. However, acquisiti

local-aiarxiv-cs-cv
23 Jul 2026
Local Ai

Robust Activation Map Rectification for Weakly Supervised Volumetric Segmentation: Temporal Coherence as a Free Lunch

DGX agent

arXiv:2607.19877v1 Announce Type: new Abstract: Weakly supervised segmentation relies heavily on class activation maps (CAMs) to initially localize target regions. However, CAMs are often noisy and pr

local-aiarxiv-cs-cv
23 Jul 2026
Research

Robust Multi-View Classification under Noisy Supervision via Global Anchor Consensus

DGX agent

arXiv:2607.18561v1 Announce Type: cross Abstract: In recent years, multi-view learning has attracted increasing attention, as it integrates the complementary information of heterogeneous views. Most e

researcharxiv-cs-cv
23 Jul 2026
Model Releases

ROMS-IMLE: A Minimalist Approach to Competitive Single-Step Generative Modelling

DGX agent

arXiv:2607.19332v1 Announce Type: cross Abstract: Generative models have undergone many generations of evolution, from VAEs/GANs to diffusion/flow matching. Along the way, the underlying techniques ha

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

RS-RIE-Bench: Benchmarking Reasoning-Guided Remote Sensing Image Editing

DGX agent

arXiv:2607.20197v1 Announce Type: new Abstract: Remote sensing image editing aims to modify remote sensing images according to natural language instructions while preserving geographic rules and senso

model-releasesarxiv-cs-cv
23 Jul 2026
Safety

SafeGen: Goal-Conditioned Video Diffusion of Safety-Critical Scenarios for VLM-Based Autonomous Driving

DGX agent

arXiv:2607.19701v1 Announce Type: new Abstract: VLMs are increasingly deployed in AD systems, creating an urgent need for rigorous safety evaluation under rare yet safety-critical scenarios. Among the

safetyarxiv-cs-cv
23 Jul 2026
Agents

Sarus: Privacy-Preserving Multi-Vendor Perception Fusion via Homomorphic Encryption

DGX agent

arXiv:2607.19146v1 Announce Type: cross Abstract: Cooperative perception enables autonomous vehicles (AVs) to improve situational awareness by aggregating detection outputs from multiple agents and se

agentsarxiv-cs-cv
23 Jul 2026
Model Releases

Seeing Before Generating: Object Perception Enhances Single-View 3D Reconstruction

DGX agent

arXiv:2607.18630v1 Announce Type: new Abstract: The relationship between object perception and reconstruction is well established in human vision, yet remains underexplored in computer vision. In this

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

Self Gradient Forcing: Native Long Video Extrapolation

DGX agent

arXiv:2607.20368v1 Announce Type: new Abstract: Recent autoregressive video diffusion methods are increasingly built upon Self Forcing, where the student is trained on histories produced by its own ro

model-releasesarxiv-cs-cv
23 Jul 2026
Research

Semantic Richness or Geometric Reasoning? The Fragility of VLM's Visual Invariance

DGX agent

arXiv:2604.01848v4 Announce Type: replace Abstract: This work investigates the fundamental fragility of state-of-the-art Vision-Language Models (VLMs) under basic geometric transformations. While mode

researcharxiv-cs-cv
23 Jul 2026
Safety

SemICP: Semantic Non-Rigid Point Cloud Registration with Elastic Energy Regularization

DGX agent

arXiv:2503.00972v4 Announce Type: replace Abstract: Purpose: Accurate point cloud registration is essential in computer-aided interventions (CAI) to align multi-modal medical images for intraoperative

safetyarxiv-cs-cv
23 Jul 2026
Local Ai

SHFormer: Dynamic Spectral Filtering Convolutional Neural Network and High-pass Kernel Generation Transformer for Adaptive MRI Reconstruction

DGX agent

arXiv:2607.20159v1 Announce Type: new Abstract: Attention Mechanism (AM) selectively focuses on essential information for imaging tasks and captures relationships between distant pixel neighborhoods t

local-aiarxiv-cs-cv
23 Jul 2026
Research

Signed Rectified Flow: Negativity-Controlled Generation

DGX agent

arXiv:2607.18516v1 Announce Type: cross Abstract: We introduce Signed Rectified Flow (Signed RF), a generalization of Rectified Flow that targets the signed measure pi^{sign} = (1+alpha)pi^+ - alphapi

researcharxiv-cs-cv
23 Jul 2026
Research

SIINR: Structurally Informed Implicit Neural Representations for super-resolution with uncertainty quantification of clinical quality diffusion MRI datasets

DGX agent

arXiv:2607.19943v1 Announce Type: new Abstract: Diffusion Magnetic Resonance Imaging (dMRI) is a powerful tool for probing brain microstructure, but clinical acquisitions are often limited by low out-

researcharxiv-cs-cv
23 Jul 2026
Model Releases

Single-Teacher View Augmentation: Enhancing Knowledge Distillation with Student-Guided Perturbations

DGX agent

arXiv:2607.11557v2 Announce Type: replace Abstract: Knowledge distillation (KD) typically relies on the fixed perspective of a single teacher, limiting the diversity of supervisory signals. While mult

model-releasesarxiv-cs-cv
23 Jul 2026
Applications

SkyEV: RGB-Event UAV detection and tracking dataset and baseline

DGX agent

arXiv:2607.18747v1 Announce Type: new Abstract: Detecting UAVs in air spaces has become increasingly important due to UAVs widespread availability and easy usage. However, due to their small size, the

applicationsarxiv-cs-cv
23 Jul 2026
Safety

STEREOFLOW: Progressive Stereo Matching with StereoDiT and Transition Flow Matching

DGX agent

arXiv:2607.19986v1 Announce Type: new Abstract: Stereo matching is a fundamental task in 3D reconstruction. Despite remarkable advances, the prevailing paradigms formulate stereo matching as a determi

safetyarxiv-cs-cv
23 Jul 2026
Model Releases

Strength-Parity Ensembling with Parameter-Isolated Experts for Multi-Task Affect Recognition

DGX agent

arXiv:2607.16290v2 Announce Type: replace Abstract: Leading entries on the multi-task track of the 11th ABAW challenge rely on heavy ensembling, yet which member is worth adding to an already strong e

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

StrokeSeg2: Stroke Lesion Segmentation in Clinical Research Workflows

DGX agent

arXiv:2607.19901v1 Announce Type: new Abstract: Deep learning frameworks like nnU-Net achieve state-of-theart brain lesion segmentation performance but remain difficult to deploy in clinical research

model-releasesarxiv-cs-cv
23 Jul 2026
Research

Strong Gravitational Lensing Posterior Sampling in Pixel-Space Using Diffusion Models and Recurrent Inference Machines

DGX agent

arXiv:2607.19459v1 Announce Type: cross Abstract: Modeling galaxy-galaxy strong gravitational lenses to infer the brightness of the source galaxy and the mass distribution of the foreground galaxy is

researcharxiv-cs-cv
23 Jul 2026
Applications

STS-NET: Spatio-Temporal Stress Network for Self-Supervised Crop Stress Detection using Satellite Image Time Series

DGX agent

arXiv:2607.18791v1 Announce Type: new Abstract: Early and accurate detection of crop stress is essential to improve agricultural productivity and ensure global food security. However, collecting a lar

applicationsarxiv-cs-cv
23 Jul 2026
Safety

Style over Substance: A Shortcut Audit of Emotion-Description Preference Evaluation

DGX agent

arXiv:2607.18508v1 Announce Type: new Abstract: Preference over model-generated emotion descriptions is emerging as a standard evaluation metric for multimodal emotion understanding, exemplified by th

safetyarxiv-cs-cv
23 Jul 2026
Research

SUPER Module for Detail-Sensitive and Cost-Efficient U-Net Variant Decoders

DGX agent

arXiv:2511.11015v2 Announce Type: replace Abstract: Skip-connected U-Net variants are widely used for dense inverse problems, yet their decoders commonly recover resolution through spatial upscaling,

researcharxiv-cs-cv
23 Jul 2026
Research

Surprise Forcing: What to Remember, When to Skip in Long Video Generation

DGX agent

arXiv:2607.18436v1 Announce Type: new Abstract: Streaming autoregressive diffusion makes minute-scale video synthesis practical, but its bounded context and fixed denoising schedule allocate resources

researcharxiv-cs-cv
23 Jul 2026
Tutorials

SWITi: Quantifying and Reducing Tiling Artifacts with Sliding Window Inner Tiling

DGX agent

arXiv:2607.18990v1 Announce Type: new Abstract: SWITi is a test-time method for reducing artifacts in tiled predictions, particularly for neural networks that learn posterior distributions from which

tutorialsarxiv-cs-cv
23 Jul 2026
Model Releases

SynGallery: A Synthetic Gallery of Real Paintings for Instance-Level Artwork Recognition

DGX agent

arXiv:2607.18907v1 Announce Type: new Abstract: Instance-level artwork recognition requires matching a handheld visitor photograph to a specific work in a large museum collection. This is challenging

model-releasesarxiv-cs-cv
23 Jul 2026
Research

Synthetic and Derived Training Images for Campus Waste Detection: A Multi-Seed Evaluation with YOLOv8n

DGX agent

arXiv:2607.19535v2 Announce Type: new Abstract: Incorrect disposal can contaminate campus recycling streams, and a bin-mounted camera could provide feedback as an item is discarded. We evaluated wheth

researcharxiv-cs-cv
23 Jul 2026
Safety

TAP-RAG: Task-Aware Policy Control for Long-Document Multimodal Question Answering

DGX agent

arXiv:2607.18917v1 Announce Type: new Abstract: Long-document multimodal question answering requires more than retrieving relevant chunks from a large document. Different queries require different evi

safetyarxiv-cs-cv
23 Jul 2026
Research

Team RAS in 11th ABAW Competition: Multimodal Ambivalence Recognition Approach

DGX agent

arXiv:2607.14702v2 Announce Type: replace Abstract: Automatic recognition of ambivalence and hesitancy is challenging because these states may be expressed through inconsistent linguistic, acoustic, f

researcharxiv-cs-cv
23 Jul 2026
Safety

Test-Time Registers as Global Priors for Tokenized Image Generation

DGX agent

arXiv:2607.16824v2 Announce Type: replace Abstract: Attention-based models often develop attention sinks, where a small number of tokens repeatedly attract attention and accumulate unusually large act

safetyarxiv-cs-cv
23 Jul 2026
Applications

Text-conditioned Segmentation for Tomato Phenotyping via Procedural Synthetic Data

DGX agent

arXiv:2607.18576v1 Announce Type: new Abstract: Vision-based automation is an excellent candidate for reducing manual labor in greenhouse crop production and phenotyping. However, progress is constrai

applicationsarxiv-cs-cv
23 Jul 2026
Research

Text Template Tokens Are Implicit Semantic Registers in Diffusion Transformers

DGX agent

arXiv:2607.19139v1 Announce Type: new Abstract: Text-to-image diffusion transformers (DiTs) jointly process text and image tokens, yet their internal computation during denoising remains poorly unders

researcharxiv-cs-cv
23 Jul 2026
Research

The JEPA Predictor: A Transferable Operator for Occluded Feature Completion

DGX agent

arXiv:2607.16274v2 Announce Type: replace Abstract: Joint-Embedding Predictive Architectures (JEPAs) train a predictor jointly with their encoder, but downstream deployment discards the predictor and

researcharxiv-cs-cv
23 Jul 2026
Research

The PAR dataset: Prostate biopsy whole slide images from an underrepresented Middle Eastern population

DGX agent

arXiv:2512.03854v2 Announce Type: replace Abstract: Artificial intelligence (AI) is increasingly used in digital pathology. Publicly available histopathology datasets remain scarce, and those that do

researcharxiv-cs-cv
23 Jul 2026
← Previous
1…4748495051…261
Next →