AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,618 results
Safety

When Policy Entropy Constraint Fails: Preserving Diversity in Flow-based RLHF via Perceptual Entropy

DGX agent

arXiv:2605.12112v1 Announce Type: new Abstract: RLHF is widely used to align flow-matching text-to-image models with human preferences, but often leads to severe diversity collapse after fine-tuning.

safetyarxiv-cs-cv
13 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

WildRelight: A Real-World Benchmark and Physics-Guided Adaptation for Single-Image Relighting

DGX agent

arXiv:2605.11696v1 Announce Type: new Abstract: Recent single-image relighting methods, powered by advanced generative models, have achieved impressive photorealism on synthetic benchmarks. However, t

model-releasesarxiv-cs-cv
13 May 2026
Research

WorldComp2D: Spatio-semantic Representations of Object Identity and Location from Local Views

DGX agent

arXiv:2605.11743v1 Announce Type: new Abstract: Learning latent representations that capture both semantic and spatial information is central to efficient spatio-semantic reasoning. However, many exis

researcharxiv-cs-cv
13 May 2026
Model Releases

XWOD: A Real-World Benchmark for Object Detection under Extreme Weather Conditions

DGX agent

arXiv:2605.11521v1 Announce Type: new Abstract: Autonomous driving and intelligent transportation systems remain vulnerable under extreme weather. The U.S. Federal Highway Administration reports that

model-releasesarxiv-cs-cv
13 May 2026
Research

Zero-shot Human Pose Estimation using Diffusion-based Inverse solvers

DGX agent

arXiv:2510.02043v2 Announce Type: replace Abstract: Pose estimation refers to tracking a human's full body posture, including their head, torso, arms, and legs. The problem is challenging in practical

researcharxiv-cs-cv
13 May 2026
Safety

ZeroIDIR: Zero-Reference Illumination Degradation Image Restoration with Perturbed Consistency Diffusion Models

DGX agent

arXiv:2605.11435v1 Announce Type: new Abstract: In this paper, we propose a zero-reference diffusion-based framework, named ZeroIDIR, for illumination degradation image restoration, which decouples th

safetyarxiv-cs-cv
13 May 2026
Model Releases

3DReflecNet: A Large-Scale Dataset for 3D Reconstruction of Reflective, Transparent, and Low-Texture Objects

DGX agent

arXiv:2605.10204v1 Announce Type: new Abstract: Accurate 3D reconstruction of objects with reflective, transparent, or low-texture surfaces still remains notoriously challenging. Such materials often

model-releasesarxiv-cs-cv
12 May 2026
Research

4D Neural Voxel Splatting: Dynamic Scene Rendering with Voxelized Guassian Splatting

DGX agent

arXiv:2511.00560v2 Announce Type: replace Abstract: Although 3D Gaussian Splatting (3D-GS) achieves efficient rendering for novel view synthesis, extending it to dynamic scenes still results in substa

researcharxiv-cs-cv
12 May 2026
Applications

A Breast Vision Pathology Foundation Model for Real-world Clinical Utility

DGX agent

arXiv:2605.08207v1 Announce Type: new Abstract: Pathology foundation models have shown strong retrospective performance, but whether such systems can support clinically relevant use remains unclear. T

applicationsarxiv-cs-cv
12 May 2026
Model Releases

A Deep Risk Estimator for Known Operator Learning

DGX agent

arXiv:2605.08517v1 Announce Type: cross Abstract: We describe an approach for estimating the statistical risk of deep networks that contain a mix of learned and known operators. Building on the maxima

model-releasesarxiv-cs-cv
12 May 2026
Applications

A Real-Calibrated Synthetic-First Data Engine

DGX agent

arXiv:2605.09699v1 Announce Type: cross Abstract: Modern computer vision systems increasingly encounter performance limitations in data-scarce domains, where collecting large-scale, high-quality label

applicationsarxiv-cs-cv
12 May 2026
Research

A Two-Stage Motion-Aware Framework for mmWave-based Human Mesh Recovery

DGX agent

arXiv:2605.08530v1 Announce Type: new Abstract: Millimeter-wave (mmWave) radar has emerged as a promising sensing modality for human perception due to its robustness under challenging environmental co

researcharxiv-cs-cv
12 May 2026
Model Releases

Action-Guided Attention for Video Action Anticipation

DGX agent

arXiv:2603.01743v2 Announce Type: replace Abstract: Anticipating future actions in videos is challenging, as the observed frames provide only evidence of past activities, requiring the inference of la

model-releasesarxiv-cs-cv
12 May 2026
Applications

Active-SAOOD: Active Sparsely Annotated Oriented Object Detection in Remote Sensing Images

DGX agent

arXiv:2605.10162v1 Announce Type: new Abstract: Reducing the annotation cost of oriented object detection in remote sensing remains a major challenge. Recently, sparse annotation has gained attention

applicationsarxiv-cs-cv
12 May 2026
Model Releases

ACWM-Phys: Investigating Generalized Physical Interaction in Action-Conditioned Video World Models

DGX agent

arXiv:2605.08567v1 Announce Type: new Abstract: Action-conditioned world models (ACWMs) have shown strong promise for video prediction and decision-making. However, existing benchmarks are largely res

model-releasesarxiv-cs-cv
12 May 2026
Research

Adaptive 3D Convolution for Remote Sensing Image Fusion

DGX agent

arXiv:2605.09455v1 Announce Type: new Abstract: Remote sensing image fusion aims to create a high-resolution multi/hyper-spectral image from a high-resolution image with limited spectral information a

researcharxiv-cs-cv
12 May 2026
Safety

Adaptive Context Matters: Towards Provable Multi-Modality Guidance for Super-Resolution

DGX agent

arXiv:2605.10470v1 Announce Type: new Abstract: Super-resolution (SR) is a severely ill-posed problem with inherent ambiguity, as widely recognized in both empirical and theoretical studies. Although

safetyarxiv-cs-cv
12 May 2026
Research

AdaptSplat: Adapting Vision Foundation Models for Feed-Forward 3D Gaussian Splatting

DGX agent

arXiv:2605.10239v1 Announce Type: new Abstract: This work explores a simple yet powerful lightweight adapter design for feed-forward 3D Gaussian Splatting (3DGS). Existing methods typically apply comp

researcharxiv-cs-cv
12 May 2026
Research

Advanced Tumor Segmentation in PET/CT Imaging: A Training Strategy Study with nnU-Net for AutoPET III

DGX agent

arXiv:2605.08161v1 Announce Type: new Abstract: Tumor segmentation in whole-body PET/CT imaging is crucial for precise disease evaluation and treatment planning. However, it remains challenging due to

researcharxiv-cs-cv
12 May 2026
Safety

Adversarial Attacks Against MLLMs via Progressive Resolution Processing and Adaptive Feature Alignment

DGX agent

arXiv:2605.09902v1 Announce Type: new Abstract: Adversarial perturbations can mislead Multimodal Large Language Models (MLLMs) recognize a benign image as a specific target object, posing serious risk

safetyarxiv-cs-cv
12 May 2026
Applications

AI- Enhanced Stethoscope in Remote Diagnostics for Cardiopulmonary Diseases

DGX agent

arXiv:2505.18184v2 Announce Type: replace-cross Abstract: The increase in cardiac and pulmonary diseases presents an alarming and pervasive health challenge on a global scale responsible for unexpecte

applicationsarxiv-cs-cv
12 May 2026
Model Releases

Alice v1: Distillation-Enhanced Video Generation Surpassing Closed-Source Models

DGX agent

arXiv:2605.08115v1 Announce Type: cross Abstract: Wepresent Alice v1, a 14-billion parameter open-source video generation model that achieves state-of-the-art quality through consistency distillation

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

AlignDrive: Aligned Lateral-Longitudinal Planning for End-to-End Autonomous Driving

DGX agent

arXiv:2601.01762v2 Announce Type: replace-cross Abstract: Practical autonomous driving requires models that generalize by reasoning through spatial-temporal possibilities to exclude unsafe outcomes. W

model-releasesarxiv-cs-cv
12 May 2026
Research

An Efficient Token Compression Framework for Visual Object Tracking

DGX agent

arXiv:2605.08329v1 Announce Type: new Abstract: Refining visual representations by eliminating their internal feature-level redundancy is crucial for simultaneously optimizing the performance and comp

researcharxiv-cs-cv
12 May 2026
Research

An Elastic Shape Variational Autoencoder for Skeleton Pose Trajectories

DGX agent

arXiv:2605.09231v1 Announce Type: new Abstract: Deep generative models provide flexible frameworks for modeling complex, structured data such as images, videos, 3D objects, and texts. However, when ap

researcharxiv-cs-cv
12 May 2026
Research

Anchoring the Eigengap: Cross-Modal Spectral Stabilization for Sample-Efficient Representation Learning

DGX agent

arXiv:2605.08764v1 Announce Type: cross Abstract: Deep vision models degrade sharply in low-data regimes, particularly in medical imaging where labeled samples are scarce. We show this arises not mere

researcharxiv-cs-cv
12 May 2026
Research

Annotation-free deep learning for detection and segmentation of fetal germinal matrix-intraventricular hemorrhage in brain MRI

DGX agent

arXiv:2605.09575v1 Announce Type: cross Abstract: Background: Prenatal germinal matrix-intraventricular hemorrhage (GMH-IVH) is a leading cause of infant mortality and neurodevelopmental impairment. M

researcharxiv-cs-cv
12 May 2026
Model Releases

AnyDepth-DETR/-YOLO: Any-depth object detection with a single network

DGX agent

arXiv:2605.09407v1 Announce Type: new Abstract: Modern object detectors are static, fixed-depth networks optimized for a single operating point, requiring separate models for different deployment scen

model-releasesarxiv-cs-cv
12 May 2026
Research

AQMP: Image compression through Adaptive Quadtree Refinement and Matching Pursuit with Hyperparameter Optimization

DGX agent

arXiv:2605.09190v1 Announce Type: new Abstract: We present AQMP, a novel image codec combining Adaptive Quadtree Refinement with Matching Pursuit. Unlike conventional Matching Pursuit methods that ope

researcharxiv-cs-cv
12 May 2026
Model Releases

Are vision-language models ready to zero-shot replace supervised classification models in agriculture?

DGX agent

arXiv:2512.15977v3 Announce Type: replace Abstract: Vision-language models (VLMs) are increasingly proposed as general-purpose solutions for visual recognition tasks, yet their reliability for agricul

model-releasesarxiv-cs-cv
12 May 2026
Agents

AstroSplat: Physics-Based Gaussian Splatting for Rendering and Reconstruction of Small Celestial Bodies

DGX agent

arXiv:2603.11969v2 Announce Type: replace Abstract: Image-based surface reconstruction and characterization are crucial for missions to small celestial bodies (e.g., asteroids), as it informs mission

agentsarxiv-cs-cv
12 May 2026
Research

Attention Itself Could Retrieve.RetrieveVGGT: Training-Free Long Context Streaming 3D Reconstruction via Query-Key Similarity Retrieval

DGX agent

arXiv:2605.09644v1 Announce Type: new Abstract: Visual Geometry Grounded Transformer (VGGT) advances 3D reconstruction via scalable Transformer architecture, but the quadratic complexity of global att

researcharxiv-cs-cv
12 May 2026
Safety

Attention Sinks in Diffusion Transformers: A Causal Analysis

DGX agent

arXiv:2605.09313v1 Announce Type: new Abstract: Attention sinks -- tokens that receive disproportionate attention mass -- are assumed to be functionally important in autoregressive language models, bu

safetyarxiv-cs-cv
12 May 2026
Research

Augmented Equivariant Mesh Networks for Anatomical Segmentation

DGX agent

arXiv:2605.08172v1 Announce Type: new Abstract: Anatomical mesh segmentation requires models that operate directly on irregular surface geometry while remaining robust to arbitrary patient pose and me

researcharxiv-cs-cv
12 May 2026
Model Releases

AUHead: Realistic Emotional Talking Head Generation via Action Units Control

DGX agent

arXiv:2602.09534v2 Announce Type: replace Abstract: Realistic talking-head video generation is critical for virtual avatars, film production, and interactive systems. Current methods struggle with nua

model-releasesarxiv-cs-cv
12 May 2026
Research

Automated Detection of Abnormalities in Zebrafish Development

DGX agent

arXiv:2605.10464v1 Announce Type: new Abstract: Zebrafish embryos are a valuable model for drug discovery due to their optical transparency and genetic similarity to humans. However, current evaluatio

researcharxiv-cs-cv
12 May 2026
Research

Automated high-frequency quantification of fish communities and biomass using computer vision

DGX agent

arXiv:2605.10449v1 Announce Type: new Abstract: Quantifying fish community structure is essential for understanding biodiversity and ecosystem responses in a changing environment, yet existing survey

researcharxiv-cs-cv
12 May 2026
Research

Automated Robotic Moisture Monitoring in Agricultural Fields

DGX agent

arXiv:2605.09050v1 Announce Type: cross Abstract: Monitoring moisture level of land in a large-scale plantation is tedious. The main objective of this project is to use a robotic kit in collaboration

researcharxiv-cs-cv
12 May 2026
Safety

BathyFacto: Refraction-Aware Two-Media Neural Radiance Fields for Bathymetry

DGX agent

arXiv:2605.10174v1 Announce Type: new Abstract: Through-water photogrammetry based on UAV imagery enables shallow-water bathymetry, but refraction at the air-water interface violates the straight-ray

safetyarxiv-cs-cv
12 May 2026
Research

BEA-GS: BEyond RAdiance Supervision in 3DGS for Precise Object Extraction

DGX agent

arXiv:2605.09662v1 Announce Type: new Abstract: Most Gaussian Splatting techniques that provide a 3D semantic representation of the scene do not optimize the underlying 3D geometry, making object-leve

researcharxiv-cs-cv
12 May 2026
Applications

Behavior-Centric Extraction of Scenarios from Highway Traffic Data and their Domain-Knowledge-Guided Clustering using CVQ-VAE

DGX agent

arXiv:2603.16964v2 Announce Type: replace Abstract: Approval of ADS depends on evaluating its behavior within representative real-world traffic scenarios. A common way to obtain such scenarios is to e

applicationsarxiv-cs-cv
12 May 2026
Model Releases

BenchHAR: Benchmarking Self-Supervised Learning for Generalizable Sensor-based Activity Recognition

DGX agent

arXiv:2605.08296v1 Announce Type: new Abstract: Human Activity Recognition (HAR) from wearable sensors supports broad healthcare and behavior science applications. However, data heterogeneity and the

model-releasesarxiv-cs-cv
12 May 2026
Local Ai

Beyond Bag-of-Patches: Learning Global Layout via Textual Supervision for Late-Interaction Visual Document Retrieval

DGX agent

arXiv:2605.08421v1 Announce Type: new Abstract: Visual Document Retrieval (VDR) models mostly rely on late interaction architectures, in which documents are represented by a set of local patch embeddi

local-aiarxiv-cs-cv
12 May 2026
Research

Beyond Nearest Neighbor Interpolation in Data Augmentation

DGX agent

arXiv:2504.01527v2 Announce Type: replace Abstract: Avoiding the risk of undefined categorical labels using nearest neighbor interpolation overlooks the risk of exacerbating pixel level annotation err

researcharxiv-cs-cv
12 May 2026
Research

Beyond Spatial Compression: Interface-Centric Generative States for Open-World 3D Structure

DGX agent

arXiv:2605.10438v1 Announce Type: cross Abstract: Current 3D tokenizers largely treat representation as spatial compression: compact codes reconstruct surface geometry, but leave component ownership a

researcharxiv-cs-cv
12 May 2026
Research

Beyond Thinking: Imagining in 360^irc for Humanoid Visual Search

DGX agent

arXiv:2605.09146v1 Announce Type: new Abstract: Humanoid Visual Search (HVS) requires agents to actively explore immersive 360^irc environments. While prior methods treat this as a monolithic task rel

researcharxiv-cs-cv
12 May 2026
Model Releases

Beyond Toy Benchmarks: A Systematic Evaluation of OOD Detection Methods For Plant Pathology Classification

DGX agent

arXiv:2605.08618v1 Announce Type: new Abstract: Out-of-distribution (OOD) detection is essential for reliable deployment of deep learning systems, yet the majority of existing methods are evaluated on

model-releasesarxiv-cs-cv
12 May 2026
Local Ai

Beyond ViT Tokens: Masked-Diffusion Pretrained Convolutional Pathology Foundation Model for Cell-Level Dense Prediction

DGX agent

arXiv:2605.08276v1 Announce Type: new Abstract: Cell-level dense prediction is central to computational pathology, but remains challenging due to fine-grained histological structures, strong domain sh

local-aiarxiv-cs-cv
12 May 2026
← Previous
1…179180181182183…263
Next →