AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,515 results
Research

InterPartAbility: Text-Guided Part Matching for Interpretable Person Re-Identification

DGX agent

arXiv:2604.27122v1 Announce Type: new Abstract: Text-to-image person re-identification (TI-ReID) relies on natural-language text description to retrieve top matching individuals from a large gallery o

researcharxiv-cs-cv
1 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Iterative Definition Refinement for Zero-Shot Classification via LLM-Based Semantic Prototype Optimization

DGX agent

arXiv:2604.27335v1 Announce Type: new Abstract: Web filtering systems rely on accurate web content classification to block cyber threats, prevent data exfiltration, and ensure compliance. However, cla

model-releasesarxiv-cs-cv
1 May 2026
Model Releases

JI-ADF: Joint-Individual Learning with Adaptive Decision Fusion for Multimodal Skin Lesion Classification

DGX agent

arXiv:2604.27343v1 Announce Type: new Abstract: Skin lesion classification is essential for early dermatological diagnosis, yet many existing computer-aided systems rely primarily on dermoscopic image

model-releasesarxiv-cs-cv
1 May 2026
Model Releases

Judge, Then Drive: A Critic-Centric Vision Language Action Framework for Autonomous Driving

DGX agent

arXiv:2604.27366v1 Announce Type: new Abstract: Recent advances in vision language action (VLA) models have shown remarkable potential for autonomous driving by directly mapping multimodal inputs to c

model-releasesarxiv-cs-cv
1 May 2026
Safety

LA-Pose: Latent Action Pretraining Meets Pose Estimation

DGX agent

arXiv:2604.27448v1 Announce Type: new Abstract: This paper revisits camera pose estimation through the lens of self-supervised pretraining, focusing on inverse-dynamics pretraining as a scalable alter

safetyarxiv-cs-cv
1 May 2026
Model Releases

LaST-R1: Reinforcing Action via Adaptive Physical Latent Reasoning for VLA Models

DGX agent

arXiv:2604.28192v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have increasingly incorporated reasoning mechanisms for complex robotic manipulation. However, existing approaches

model-releasesarxiv-cs-cv
1 May 2026
Research

Leveraging Quantum-Based Architectures for Robust Diagnostics

DGX agent

arXiv:2511.12386v2 Announce Type: replace Abstract: Quantum machine learning has emerged as a promising approach for medical image analysis, particularly in settings where compact models and expressiv

researcharxiv-cs-cv
1 May 2026
Research

Leveraging Verifier-Based Reinforcement Learning in Image Editing

DGX agent

arXiv:2604.27505v1 Announce Type: new Abstract: While Reinforcement Learning from Human Feedback (RLHF) has become a pivotal paradigm for text-to-image generation, its application to image editing rem

researcharxiv-cs-cv
1 May 2026
Research

LM-CartSeg: Automated Segmentation of Lateral and Medial Cartilage and Subchondral Bone for Radiomics Analysis

DGX agent

arXiv:2512.03449v3 Announce Type: replace Abstract: Background and Objective: Radiomics of knee MRI requires robust, anatomically meaningful regions of interest (ROIs) that jointly capture cartilage a

researcharxiv-cs-cv
1 May 2026
Applications

Machine Unlearning for Class Removal through SISA-based Deep Neural Network Architectures

DGX agent

arXiv:2604.27804v1 Announce Type: new Abstract: The rapid proliferation of image generation models and other artificial intelligence (AI) systems has intensified concerns regarding data privacy and us

applicationsarxiv-cs-cv
1 May 2026
Research

MoCapAnything: Unified 3D Motion Capture for Arbitrary Skeletons from Monocular Videos

DGX agent

arXiv:2512.10881v2 Announce Type: replace Abstract: Motion capture now underpins content creation far beyond digital humans, yet most existing pipelines remain species- or template-specific. We formal

researcharxiv-cs-cv
1 May 2026
Local Ai

MoCapAnything V2: End-to-End Motion Capture for Arbitrary Skeletons

DGX agent

arXiv:2604.28130v1 Announce Type: new Abstract: Recent methods for arbitrary-skeleton motion capture from monocular video follow a factorized pipeline, where a Video-to-Pose network predicts joint pos

local-aiarxiv-cs-cv
1 May 2026
Safety

MSR:Hybrid Field Modeling for CT-MRI Rigid-Deformable Registration of the Cervical Spine with an Annotated Dataset

DGX agent

arXiv:2604.27654v1 Announce Type: new Abstract: Accurate CT-MRI registration of the cervical spine is essential for preoperative planning because this region is anatomically complex,highly variable,an

safetyarxiv-cs-cv
1 May 2026
Tutorials

Noise2Map: End-to-End Diffusion Model for Semantic Segmentation and Change Detection

DGX agent

arXiv:2604.27889v1 Announce Type: new Abstract: Semantic segmentation and change detection are two fundamental challenges in remote sensing, requiring models to capture either spatial semantics or tem

tutorialsarxiv-cs-cv
1 May 2026
Tutorials

Omni-Attribute: Open-vocabulary Attribute Encoder for Visual Concept Personalization

DGX agent

arXiv:2512.10955v2 Announce Type: replace Abstract: Visual concept personalization aims to transfer only specific image attributes, such as identity, expression, lighting, and style, into unseen conte

tutorialsarxiv-cs-cv
1 May 2026
Safety

OmniRobotHome: A Multi-Camera Platform for Real-Time Multiadic Human-Robot Interaction

DGX agent

arXiv:2604.28197v1 Announce Type: cross Abstract: Human-robot collaboration has been studied primarily in dyadic or sequential settings. However, real homes require multiadic collaboration, where mult

safetyarxiv-cs-cv
1 May 2026
Model Releases

Parameter-Efficient Architectural Modifications for Translation-Invariant CNNs

DGX agent

arXiv:2604.27870v1 Announce Type: new Abstract: Convolutional Neural Networks (CNNs) are widely assumed to be translation-invariant, yet standard architectures exhibit a startling fragility: even a si

model-releasesarxiv-cs-cv
1 May 2026
Research

PhotIQA: A photoacoustic image data set with image quality ratings

DGX agent

arXiv:2507.03478v2 Announce Type: replace-cross Abstract: Image quality assessment (IQA) is crucial in the evaluation stage of novel algorithms operating on images, including traditional and machine l

researcharxiv-cs-cv
1 May 2026
Research

Physically-Informed Fuzzy Clustering of Vertical Sounding Ionograms

DGX agent

arXiv:2604.27721v1 Announce Type: cross Abstract: This paper presents a physically-informed fuzzy clustering of vertical sounding ionograms for automatically separating the ionogram into tracks suitab

researcharxiv-cs-cv
1 May 2026
Research

PINN-Cast: Exploring the Role of Continuous-Depth NODE in Transformers and Physics Informed Loss as Soft Physical Constraints in Short-term Weather Forecasting

DGX agent

arXiv:2604.27313v1 Announce Type: cross Abstract: Operational weather prediction has long relied on physics-based numerical weather prediction (NWP), whose accuracy comes at the cost of substantial co

researcharxiv-cs-cv
1 May 2026
Research

Primus: Enforcing Attention Usage for 3D Medical Image Segmentation

DGX agent

arXiv:2503.01835v2 Announce Type: replace Abstract: Transformers have achieved remarkable success across multiple fields, yet their impact on 3D medical image segmentation remains limited with convolu

researcharxiv-cs-cv
1 May 2026
Model Releases

PVeRA: Probabilistic Vector-Based Random Matrix Adaptation

DGX agent

arXiv:2512.07703v2 Announce Type: replace Abstract: Large foundation models have emerged in the last years and are pushing performance boundaries for a variety of tasks. Training or even finetuning su

model-releasesarxiv-cs-cv
1 May 2026
Applications

RayFormer: Modeling Inter- and Intra-Ray Similarity for NeRF-Based Video Snapshot Compressive Imaging

DGX agent

arXiv:2604.27702v1 Announce Type: new Abstract: Video snapshot compressive imaging (SCI) enables the reconstruction of dynamic scenes from a single snapshot measurement. Recently, NeRF-based methods h

applicationsarxiv-cs-cv
1 May 2026
Research

Representation Frechet Loss for Visual Generation

DGX agent

arXiv:2604.28190v1 Announce Type: new Abstract: We show that Frechet Distance (FD), long considered impractical as a training objective, can in fact be effectively optimized in the representation spac

researcharxiv-cs-cv
1 May 2026
Model Releases

Representative Spectral Correlation Network for Multi-source Remote Sensing Image Classification

DGX agent

arXiv:2604.27323v1 Announce Type: cross Abstract: Hyperspectral image (HSI) and SAR/LiDAR data offer complementary spectral and structural information for land-cover classification. However, their eff

model-releasesarxiv-cs-cv
1 May 2026
Safety

Residual Gaussian Splatting for Ultra Sparse-View CBCT Reconstruction

DGX agent

arXiv:2604.27552v1 Announce Type: new Abstract: While 3D Gaussian splatting (3DGS) offers explicit and efficient scene representations for cone-beam computed tomography reconstruction, conventional ph

safetyarxiv-cs-cv
1 May 2026
Applications

ResiHMR: Residual-Limb Aware Single-Image 3D Human Mesh Recovery for Individuals with Limb Loss

DGX agent

arXiv:2604.28025v1 Announce Type: new Abstract: Single-image human mesh recovery provides a compact 3D, person-centric representation that supports analysis, animation, AR and VR, rehabilitation, and

applicationsarxiv-cs-cv
1 May 2026
Research

Rethinking Pulmonary Embolism Segmentation: A Study of Current Approaches and Challenges with an Open Weight Model

DGX agent

arXiv:2509.18308v3 Announce Type: replace Abstract: Pulmonary Embolism (PE) is a life-threatening condition for which accurate and timely detection is critical to patient care. However, our systematic

researcharxiv-cs-cv
1 May 2026
Research

Revealing the Impact of Visual Text Style on Attribute-based Descriptions Produced by Large Visual Language Models

DGX agent

arXiv:2604.27553v1 Announce Type: new Abstract: When the visual style of text is considered, a wide variety can be observed in font, color, and size. However, when a word is read, its meaning is indep

researcharxiv-cs-cv
1 May 2026
Research

REVIVE 3D: Refinement via Encoded Voluminous Inflated prior for Volume Enhancement

DGX agent

arXiv:2604.27504v1 Announce Type: new Abstract: Recent generative models have shown strong performance in generating diverse 3D assets from 2D images, a fundamental research topic in computer vision a

researcharxiv-cs-cv
1 May 2026
Safety

Robot Learning from Human Videos: A Survey

DGX agent

arXiv:2604.27621v1 Announce Type: cross Abstract: A critical bottleneck hindering further advancement in embodied AI and robotics is the challenge of scaling robot data. To address this, the field of

safetyarxiv-cs-cv
1 May 2026
Safety

Sample-efficient evidence estimation of score based priors for model selection

DGX agent

arXiv:2602.20549v2 Announce Type: replace-cross Abstract: The choice of prior is central to solving ill-posed imaging inverse problems, making it essential to select one consistent with the measuremen

safetyarxiv-cs-cv
1 May 2026
Research

SECOS: Semantic Capture for Rigorous Classification in Open-World Semi-Supervised Learning

DGX agent

arXiv:2604.27596v1 Announce Type: new Abstract: In open-world semi-supervised learning (OWSSL), a model learns from labeled data and unlabeled data containing both known and novel classes. In practica

researcharxiv-cs-cv
1 May 2026
Research

Self-Supervised Learning of Plant Image Representations

DGX agent

arXiv:2604.27538v1 Announce Type: new Abstract: Automated plant recognition plays a crucial role in biodiversity monitoring and conservation, yet current approaches rely heavily on supervised learning

researcharxiv-cs-cv
1 May 2026
Model Releases

Softmax-GS: Generalized Gaussians Learning When to Blend or Bound

DGX agent

arXiv:2604.27437v1 Announce Type: new Abstract: 3D Gaussian Splatting (3D GS) is widely adopted for novel view synthesis due to its high training and rendering efficiency. However, its efficiency reli

model-releasesarxiv-cs-cv
1 May 2026
Agents

SpaAct: Spatially-Activated Transition Learning with Curriculum Adaptation for Vision-Language Navigation

DGX agent

arXiv:2604.27620v1 Announce Type: new Abstract: Vision-and-Language Navigation (VLN) aims to enable an embodied agent to follow natural-language instructions and navigate to a target location in unsee

agentsarxiv-cs-cv
1 May 2026
Applications

Sparse-View 3D Gaussian Splatting in the Wild

DGX agent

arXiv:2604.27422v1 Announce Type: new Abstract: We propose a 3D novel sparse-view synthesis framework for unconstrained real-world scenarios that contain distractors. Unlike existing methods that prim

applicationsarxiv-cs-cv
1 May 2026
Model Releases

Spectral Dynamic Attention Network for Hyperspectral Image Super-Resolution

DGX agent

arXiv:2604.27326v1 Announce Type: cross Abstract: Hyperspectral image super-resolution is essential for enhancing the spatial fidelity of HSI data, yet existing deep learning methods often struggle wi

model-releasesarxiv-cs-cv
1 May 2026
Research

SQuadGen: Generating Simple Quad Layouts via Chart Distance Fields

DGX agent

arXiv:2604.27329v1 Announce Type: cross Abstract: 3D shapes from scanning, reconstruction, or AI-generated content often lack simple quad mesh layouts -- critical for efficient editing and modeling. E

researcharxiv-cs-cv
1 May 2026
Research

Stop Holding Your Breath: CT-Informed Gaussian Splatting for Dynamic Bronchoscopy

DGX agent

arXiv:2604.28179v1 Announce Type: new Abstract: Bronchoscopic navigation relies on registering endoscopic video to a preoperative CT scan, but respiratory motion deforms the airway by 5-20 mm, creatin

researcharxiv-cs-cv
1 May 2026
Research

Student Classroom Behavior Recognition Based on Improved YOLOv8s

DGX agent

arXiv:2604.27293v1 Announce Type: new Abstract: In classroom teaching, student behavior can reflect their learning state and classroom participation, which is of great significance for teaching qualit

researcharxiv-cs-cv
1 May 2026
Research

TAFA-GSGC: Group-wise Scalable Point Cloud Geometry Compression with Progressive Residual Refinement

DGX agent

arXiv:2604.28045v1 Announce Type: new Abstract: Scalable compression is essential for bandwidth-adaptive transmission, yet most learned codecs are optimized for a fixed rate-distortion point, making r

researcharxiv-cs-cv
1 May 2026
Research

Taming Noise-Induced Prototype Degradation for Privacy-Preserving Personalized Federated Fine-Tuning

DGX agent

arXiv:2604.27833v1 Announce Type: new Abstract: Prototype-based Personalized Federated Learning (ProtoPFL) enables efficient multi-domain adaptation by communicating compact class prototypes, but dire

researcharxiv-cs-cv
1 May 2026
Local Ai

TeD-Loc: Text Distillation for Weakly Supervised Object Localization

DGX agent

arXiv:2501.12632v2 Announce Type: replace Abstract: Weakly supervised object localization (WSOL) models are trained using only image-level class labels. They can predict both the object class and spat

local-aiarxiv-cs-cv
1 May 2026
Safety

Test-Time Distillation for Continual Model Adaptation

DGX agent

arXiv:2506.02671v3 Announce Type: replace Abstract: Deep neural networks often suffer performance degradation upon deployment due to distribution shifts. Continual Test-Time Adaptation (CTTA) aims to

safetyarxiv-cs-cv
1 May 2026
Model Releases

Towards All-Day Perception for Off-Road Driving: A Large-Scale Multispectral Dataset and Comprehensive Benchmark

DGX agent

arXiv:2604.27499v1 Announce Type: new Abstract: Off-road nighttime autonomous driving suffers from unreliable visible-light perception, making infrared modality crucial for accurate freespace detectio

model-releasesarxiv-cs-cv
1 May 2026
Research

Towards Generalizable Mapping of Hedges and Linear Woody Features from Earth Observation Data: a national Product for Germany

DGX agent

arXiv:2604.27247v1 Announce Type: new Abstract: Hedges and other linear woody features provide valuable ecosystem services, particularly within intensively managed agricultural landscapes. They are ke

researcharxiv-cs-cv
1 May 2026
Applications

TranSplat: Instant Object Relighting in Gaussian Splatting via Spherical Harmonic Radiance Transfer

DGX agent

arXiv:2503.22676v5 Announce Type: replace Abstract: We present TranSplat, a method for instant, accurate object relighting within the Gaussian Splatting (GS) framework. Rather than relying on costly i

applicationsarxiv-cs-cv
1 May 2026
← Previous
1…204205206207208…261
Next →