AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,515 results
Research

Coarse-to-Real: Generative Rendering for Populated Dynamic Scenes

DGX agent

arXiv:2601.22301v2 Announce Type: replace Abstract: Traditional rendering pipelines rely on complex assets, accurate materials and lighting, and substantial computational resources to produce realisti

researcharxiv-cs-cv
28 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Comparative Study of Weighted and Coupled Second- and Fourth-Order PDEs for Image Despeckling in Grayscale, Color, SAR, and Ultrasound

DGX agent

arXiv:2604.23612v1 Announce Type: new Abstract: Partial Differential Equation (PDE)-based approaches have gained significant attention in image despeckling due to their strong capability to preserve s

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

Complexity of Linear Regions in Self-supervised Deep ReLU Networks

DGX agent

arXiv:2604.24393v1 Announce Type: cross Abstract: There has been growing interest in studying the complexity of Rectified Linear Unit (ReLU) based activation networks. Recent work investigates the evo

model-releasesarxiv-cs-cv
28 Apr 2026
Safety

Computer Vision-Based Early Detection of Container Loss at Sea

DGX agent

arXiv:2604.24193v1 Announce Type: new Abstract: Containerised shipping underpins global trade, yet container loss at sea remains a persistent safety, environmental, and economic challenge. Despite com

safetyarxiv-cs-cv
28 Apr 2026
Applications

Contrastive Learning for Multimodal Human Activity Recognition with Limited Labeled Data

DGX agent

arXiv:2604.23281v1 Announce Type: cross Abstract: Human activity recognition serves as the foundation for various emerging applications. In recent years, researchers have used collaborative sensing of

applicationsarxiv-cs-cv
28 Apr 2026
Research

Decoupling Wavelet Sub-bands for Single Source Domain Generalization in Fundus Image Segmentation

DGX agent

arXiv:2603.28463v2 Announce Type: replace Abstract: Domain generalization in fundus imaging is challenging due to variations in acquisition conditions across devices and clinical settings. The inabili

researcharxiv-cs-cv
28 Apr 2026
Model Releases

Deploy DINO with Many-to-Many Association

DGX agent

arXiv:2604.23670v1 Announce Type: new Abstract: Motivated by the limited generalization of supervised image matching models to unseen image domains, we explore the zero-shot deployment of DINO feature

model-releasesarxiv-cs-cv
28 Apr 2026
Safety

Designing Instance-Level Sampling Schedules via REINFORCE with James-Stein Shrinkage

DGX agent

arXiv:2511.22177v2 Announce Type: replace-cross Abstract: Most post-training methods for text-to-image samplers focus on model weights: either fine-tuning the backbone for alignment or distilling it f

safetyarxiv-cs-cv
28 Apr 2026
Model Releases

Detecting and Evaluating Medical Hallucinations in Large Vision Language Models

DGX agent

arXiv:2406.10185v2 Announce Type: replace Abstract: Large Vision Language Models (LVLMs) are increasingly integral to healthcare applications, including medical visual question answering and imaging r

model-releasesarxiv-cs-cv
28 Apr 2026
Safety

DextER: Language-driven Dexterous Grasp Generation with Embodied Reasoning

DGX agent

arXiv:2601.16046v2 Announce Type: replace-cross Abstract: Language-driven dexterous grasp generation requires the models to understand task semantics, 3D geometry, and complex hand-object interactions

safetyarxiv-cs-cv
28 Apr 2026
Model Releases

DGHMesh: A Large-scale Dual-radar mmWave Dataset and Generalization-focused Benchmark for Human Mesh Reconstruction

DGX agent

arXiv:2604.22827v1 Announce Type: new Abstract: Millimeter-wave (mmWave) radar has shown great potential for contactless, privacy-preserving, and robust human sensing, yet existing mmWave-based human

model-releasesarxiv-cs-cv
28 Apr 2026
Research

DiffuSAM: Diffusion-Based Prompt-Free SAM2 for Few-Shot and Source-Free Medical Image Segmentation

DGX agent

arXiv:2604.24719v1 Announce Type: new Abstract: Segmentation models such as Segment Anything Model (SAM) and SAM2 achieve strong prompt-driven zero-shot performance. However, their training on natural

researcharxiv-cs-cv
28 Apr 2026
Research

Diffusion Model as a Generalist Segmentation Learner

DGX agent

arXiv:2604.24575v1 Announce Type: new Abstract: Diffusion models are primarily trained for image synthesis, yet their denoising trajectories encode rich, spatially aligned visual priors. In this paper

researcharxiv-cs-cv
28 Apr 2026
Research

Discriminator-Guided Adaptive Diffusion for Source-Free Test-Time Adaptation under Image Corruptions

DGX agent

arXiv:2604.23636v1 Announce Type: new Abstract: In this work, we study Source-Free Unsupervised Domain Adaptation under corruption-induced domain shifts, where performance degradation is caused by nat

researcharxiv-cs-cv
28 Apr 2026
Applications

Do Protective Perturbations Really Protect Portrait Privacy under Real-world Image Transformations?

DGX agent

arXiv:2604.23688v1 Announce Type: new Abstract: Proactive defense methods protect portrait images from unauthorized editing or talking face generation (TFG) by introducing pixel-level protective pertu

applicationsarxiv-cs-cv
28 Apr 2026
Model Releases

Don't Pause! Every prediction matters in a streaming video

DGX agent

arXiv:2604.24317v1 Announce Type: new Abstract: Streaming video models should respond the moment an event unfolds, not after the moment has passed. Yet existing online VideoQA benchmarks remain largel

model-releasesarxiv-cs-cv
28 Apr 2026
Research

Dream-Cubed: Controllable Generative Modeling in Minecraft by Training on Billions of Cubes

DGX agent

arXiv:2604.22847v1 Announce Type: new Abstract: We introduce Dream-Cubed, a large-scale dataset of Minecraft worlds at voxel resolution, and a family of models using cubes as powerful compositional un

researcharxiv-cs-cv
28 Apr 2026
Model Releases

DRIFT: Transferring Reasoning Priors for Efficient MLLM Fine-Tuning

DGX agent

arXiv:2510.15050v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) have made rapid progress, yet their reasoning ability often lags behind strong text-only LLMs. Bridging thi

model-releasesarxiv-cs-cv
28 Apr 2026
Local Ai

DYMAPIA: A Multi-Domain Framework for Detecting AI-based Video Manipulation

DGX agent

arXiv:2604.24426v1 Announce Type: new Abstract: AI-generated media are advancing rapidly, raising pressing concerns for content authenticity and digital trust. We introduce DYMAPIA, a multi-domain Dee

local-aiarxiv-cs-cv
28 Apr 2026
Model Releases

DynProto: Dynamic Prototype Evolution for Out-of-Distribution Detection

DGX agent

arXiv:2604.23729v1 Announce Type: new Abstract: Recent studies show that using potential out-of-distribution (OOD) labels from large corpora as auxiliary information can improve OOD detection in visio

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

EAGLE: Expert-Augmented Attention Guidance for Tuning-Free Industrial Anomaly Detection in Multimodal Large Language Models

DGX agent

arXiv:2602.17419v3 Announce Type: replace Abstract: Multimodal large language models (MLLMs) can enrich industrial anomaly detection with semantic descriptions and anomaly reasoning, but they still la

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

Easy Ensemble: Simple Deep Ensemble Learning for Sensor-Based Human Activity Recognition

DGX agent

arXiv:2203.04153v2 Announce Type: replace Abstract: Sensor-based human activity recognition (HAR) is a paramount technology in the Internet of Things services. HAR using representation learning, which

model-releasesarxiv-cs-cv
28 Apr 2026
Research

Edit Where You Mean: Region-Aware Adapter Injection for Mask-Free Local Image Editing

DGX agent

arXiv:2604.23763v1 Announce Type: new Abstract: Large diffusion transformers (DiTs) follow global editing instructions well but consistently leak local edits into unrelated regions, because joint-atte

researcharxiv-cs-cv
28 Apr 2026
Research

Efficient Image Annotation via Semi-Supervised Object Segmentation with Label Propagation

DGX agent

arXiv:2604.22992v1 Announce Type: new Abstract: Reliable object perception is necessary for general-purpose service robots. Open-vocabulary detectors struggle to generalize beyond a few classes and fu

researcharxiv-cs-cv
28 Apr 2026
Model Releases

ELSA: Exact Linear-Scan Attention for Fast and Memory-Light Vision Transformers

DGX agent

arXiv:2604.23798v1 Announce Type: cross Abstract: Existing attention accelerators often trade exact softmax semantics, depend on fused Tensor Core kernels, or incur sequential depth that limits FP32 t

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

EMCompress: Video-LLMs with Endomorphic Multimodal Compression

DGX agent

arXiv:2508.21094v3 Announce Type: replace Abstract: Video-LLMs face a fundamental tension in long-video reasoning: static, sparse frame sampling either dilutes evidence across task-irrelevant segments

model-releasesarxiv-cs-cv
28 Apr 2026
Research

Enhanced Privacy and Communication Efficiency in Non-IID Federated Learning with Adaptive Quantization and Differential Privacy

DGX agent

arXiv:2604.23426v1 Announce Type: new Abstract: Federated learning (FL) is a distributed machine learning method where multiple devices collaboratively train a model under the management of a central

researcharxiv-cs-cv
28 Apr 2026
Research

Evaluating Remote Sensing Image Captions Beyond Metric Biases

DGX agent

arXiv:2604.22855v1 Announce Type: new Abstract: The core objective of image captioning is to achieve lossless semantic compression from visual signals into textual modalities. However, the reliance on

researcharxiv-cs-cv
28 Apr 2026
Model Releases

EX-FIQA: Leveraging Intermediate Early eXit Representations from Vision Transformers for Face Image Quality Assessment

DGX agent

arXiv:2604.22842v1 Announce Type: new Abstract: Face Image Quality Assessment is crucial for reliable face recognition systems, yet existing Vision Transformer-based approaches rely exclusively on fin

model-releasesarxiv-cs-cv
28 Apr 2026
Local Ai

EXACT: an explainable anomaly-aware vision foundation model for analysis of 3D chest CT

DGX agent

arXiv:2604.24146v1 Announce Type: new Abstract: Chest computed tomography (CT) is central to the detection and management of thoracic disease, yet the growing scale and complexity of volumetric imagin

local-aiarxiv-cs-cv
28 Apr 2026
Model Releases

Exploring Hierarchical Consistency and Unbiased Objectness for Open-Vocabulary Object Detection

DGX agent

arXiv:2604.23344v1 Announce Type: new Abstract: Conventional object detectors typically operate under a closed-set assumption, limiting recognition to a predefined set of base classes seen during trai

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

FastAT Benchmark: A Comprehensive Framework for Fair Evaluation of Fast Adversarial Training Methods

DGX agent

arXiv:2604.22853v1 Announce Type: new Abstract: Fast Adversarial Training (FastAT) seeks to achieve adversarial robustness at a fraction of the computational cost incurred by standard multi-step metho

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

FDIM: A Feature-distance-based Generic Video Quality Metric for Versatile Codecs

DGX agent

arXiv:2604.24123v1 Announce Type: new Abstract: Video technology is advancing toward Ultra High Definition (UHD) and High Dynamic Range (HDR), which intensifies the need for higher compression efficie

model-releasesarxiv-cs-cv
28 Apr 2026
Research

FlashOverlap: Minimizing Tail Latency in Communication Overlap for Distributed LLM Training

DGX agent

arXiv:2604.24013v1 Announce Type: cross Abstract: The rapid growth in the size of large language models has necessitated the partitioning of computational workloads across accelerators such as GPUs, T

researcharxiv-cs-cv
28 Apr 2026
Model Releases

Franca: Nested Matryoshka Clustering for Scalable Visual Representation Learning

DGX agent

arXiv:2507.14137v4 Announce Type: replace Abstract: We present Franca (pronounced Fran-ka): free one; the first fully open-source (data, code, weights) vision foundation model that matches and in many

model-releasesarxiv-cs-cv
28 Apr 2026
Research

From Edges to Depth: Probing the Spatial Hierarchy in Vision Transformers

DGX agent

arXiv:2604.23452v1 Announce Type: new Abstract: Vision Transformers trained only on image classification routinely transfer to tasks that demand spatial understanding, yet they receive no spatial supe

researcharxiv-cs-cv
28 Apr 2026
Research

FunRec: Reconstructing Functional 3D Scenes from Egocentric Interaction Videos

DGX agent

arXiv:2604.05621v2 Announce Type: replace Abstract: We present FunRec, a method for reconstructing functional 3D digital twins of indoor scenes directly from egocentric RGB-D interaction videos. Unlik

researcharxiv-cs-cv
28 Apr 2026
Research

FUSER: Feed-Forward MUltiview 3D Registration Transformer and SE(3)^N Diffusion Refinement

DGX agent

arXiv:2512.09373v2 Announce Type: replace Abstract: Registration of multiview point clouds conventionally relies on extensive pairwise matching to build a pose graph for global synchronization, which

researcharxiv-cs-cv
28 Apr 2026
Research

GA2-CLIP: Generic Attribute Anchor for Efficient Prompt Tuningin Video-Language Models

DGX agent

arXiv:2511.22125v2 Announce Type: replace Abstract: Visual and textual soft prompt tuning can effectively improve the adaptability of Vision-Language Models (VLMs) in downstream tasks. However, fine-t

researcharxiv-cs-cv
28 Apr 2026
Model Releases

GCP: Guarded Collaborative Perception with Spatial-Temporal Aware Malicious Agent Detection

DGX agent

arXiv:2501.02450v2 Announce Type: replace Abstract: Collaborative perception significantly enhances autonomous driving safety by extending each vehicle's perception range through message sharing among

model-releasesarxiv-cs-cv
28 Apr 2026
Research

GenAssets: Generating in-the-wild 3D Assets in Latent Space

DGX agent

arXiv:2604.23010v1 Announce Type: new Abstract: High-quality 3D assets for traffic participants are critical for multi-sensor simulation, which is essential for the safe end-to-end development of auto

researcharxiv-cs-cv
28 Apr 2026
Model Releases

Generalising maximum mean discrepancy: kernelised functional Bregman divergences

DGX agent

arXiv:2604.24047v1 Announce Type: cross Abstract: Bregman divergences play a pivotal role in statistics, machine learning and computational information geometry. Particularly in the context of machine

model-releasesarxiv-cs-cv
28 Apr 2026
Research

Generalizable CT-Free PET Attenuation and Scatter Correction for Pediatric Patients

DGX agent

arXiv:2604.22894v1 Announce Type: cross Abstract: Computed tomography (CT)-based attenuation and scatter correction improves quantitative PET but adds radiation exposure that is particularly undesirab

researcharxiv-cs-cv
28 Apr 2026
Research

Geometric Analysis of Self-Supervised Vision Representations for Semantic Image Retrieval

DGX agent

arXiv:2604.24469v1 Announce Type: cross Abstract: Content-based image retrieval (CBIR) systems enable users to search images based on visual content instead of relying on metadata. The text domain has

researcharxiv-cs-cv
28 Apr 2026
Research

Geometry-Conditioned Diffusion for Occlusion-Robust In-Bed Pose Estimation

DGX agent

arXiv:2604.23651v1 Announce Type: new Abstract: Robust in-bed human pose estimation under blanket occlusion remains challenging due to the scarcity of reliable labeled training data for heavily covere

researcharxiv-cs-cv
28 Apr 2026
Model Releases

GoClick: Lightweight Element Grounding Model for Autonomous GUI Interaction

DGX agent

arXiv:2604.23941v1 Announce Type: new Abstract: Graphical User Interface (GUI) element grounding (precisely locating elements on screenshots based on natural language instructions) is fundamental for

model-releasesarxiv-cs-cv
28 Apr 2026
Applications

Gradient-Guided Exploration of Generative Model's Latent Space for Controlled Iris Image Augmentations

DGX agent

arXiv:2511.09749v2 Announce Type: replace Abstract: Developing reliable iris recognition and presentation attack detection methods requires diverse datasets that capture realistic variations in iris f

applicationsarxiv-cs-cv
28 Apr 2026
Model Releases

Graph-augmented Segmentation of Complex Shapes in Laser Powder bed Fusion for Enhanced In Situ Inspection

DGX agent

arXiv:2604.24234v1 Announce Type: new Abstract: The technological maturity of in situ inspection and monitoring methods in additive manufacturing is steadily increasing, enabling more efficient and pr

model-releasesarxiv-cs-cv
28 Apr 2026
← Previous
1…211212213214215…261
Next →