AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,515 results
Local Ai

MApLe: Multi-instance Alignment of Diagnostic Reports and Large Medical Images

DGX agent

arXiv:2604.13970v1 Announce Type: new Abstract: In diagnostic reports, experts encode complex imaging data into clinically actionable information. They describe subtle pathological findings that are m

local-aiarxiv-cs-cv
16 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Med-CAM: Minimal Evidence for Explaining Medical Decision Making

DGX agent

arXiv:2604.13695v1 Announce Type: new Abstract: Reliable and interpretable decision-making is essential in medical imaging, where diagnostic outcomes directly influence patient care. Despite advances

safetyarxiv-cs-cv
16 Apr 2026
Applications

MSGS: Multispectral 3D Gaussian Splatting

DGX agent

arXiv:2604.13340v1 Announce Type: new Abstract: We present a multispectral extension to 3D Gaussian Splatting (3DGS) for wavelength-aware view synthesis. Each Gaussian is augmented with spectral radia

applicationsarxiv-cs-cv
16 Apr 2026
Local Ai

Multi-Agent Object Detection Framework Based on Raspberry Pi YOLO Detector and Slack-Ollama Natural Language Interface

DGX agent

arXiv:2604.13345v1 Announce Type: new Abstract: The paper presents design and prototype implementation of an edge based object detection system within the new paradigm of AI agents orchestration. It g

local-aiarxiv-cs-cv
16 Apr 2026
Safety

Multi-Dimensional Knowledge Profiling with Large-Scale Literature Database and Hierarchical Retrieval

DGX agent

arXiv:2601.15170v2 Announce Type: replace Abstract: The rapid expansion of research across machine learning, vision, and language has produced a volume of publications that is increasingly difficult t

safetyarxiv-cs-cv
16 Apr 2026
Research

Multi-modal panoramic 3D outdoor datasets for place categorization

DGX agent

arXiv:2604.13142v1 Announce Type: cross Abstract: We present two multi-modal panoramic 3D outdoor (MPO) datasets for semantic place categorization with six categories: forest, coast, residential area,

researcharxiv-cs-cv
16 Apr 2026
Tutorials

Multitasking Embedding for Embryo Blastocyst Grading Prediction (MEmEBG)

DGX agent

arXiv:2604.13217v1 Announce Type: new Abstract: Reliable evaluation of blastocyst quality is critical for the success of in vitro fertilization (IVF) treatments. Current embryo grading practices prima

tutorialsarxiv-cs-cv
16 Apr 2026
Research

MyoVision: A Mobile Research Tool and NEATBoost-Attention Ensemble Framework for Real Time Chicken Breast Myopathy Detection

DGX agent

arXiv:2604.13456v1 Announce Type: cross Abstract: Woody Breast (WB) and Spaghetti Meat (SM) myopathies significantly impact poultry meat quality, yet current detection methods rely either on subjectiv

researcharxiv-cs-cv
16 Apr 2026
Research

Neural 3D Reconstruction of Planetary Surfaces from Descent-Phase Wide-Angle Imagery

DGX agent

arXiv:2604.13235v1 Announce Type: new Abstract: Digital elevation modeling of planetary surfaces is essential for studying past and ongoing geological processes. Wide-angle imagery acquired during spa

researcharxiv-cs-cv
16 Apr 2026
Local Ai

One Token per Highly Selective Frame: Towards Extreme Compression for Long Video Understanding

DGX agent

arXiv:2604.14149v1 Announce Type: new Abstract: Long video understanding is inherently challenging for vision-language models (VLMs) because of the extensive number of frames. With each video frame ty

local-aiarxiv-cs-cv
16 Apr 2026
Research

OneHOI: Unifying Human-Object Interaction Generation and Editing

DGX agent

arXiv:2604.14062v1 Announce Type: new Abstract: Human-Object Interaction (HOI) modelling captures how humans act upon and relate to objects, typically expressed as triplets. Existing approaches split

researcharxiv-cs-cv
16 Apr 2026
Model Releases

OPTED: Open Preprocessed Trachoma Eye Dataset Using Zero-Shot SAM 3 Segmentation

DGX agent

arXiv:2603.06885v2 Announce Type: replace Abstract: Trachoma remains the leading infectious cause of blindness worldwide, with Sub-Saharan Africa bearing over 85% of the global burden and Ethiopia alo

model-releasesarxiv-cs-cv
16 Apr 2026
Tutorials

PartNerFace: Part-based Neural Radiance Fields for Animatable Facial Avatar Reconstruction

DGX agent

arXiv:2604.13918v1 Announce Type: new Abstract: We present PartNerFace, a part-based neural radiance fields approach, for reconstructing animatable facial avatar from monocular RGB videos. Existing so

tutorialsarxiv-cs-cv
16 Apr 2026
Research

PAT-VCM: Plug-and-Play Auxiliary Tokens for Video Coding for Machines

DGX agent

arXiv:2604.13294v1 Announce Type: new Abstract: Existing video coding for machines is often trained for a specific downstream task and model. As a result, the compressed representation becomes tightly

researcharxiv-cs-cv
16 Apr 2026
Model Releases

PatchPoison: Poisoning Multi-View Datasets to Degrade 3D Reconstruction

DGX agent

arXiv:2604.13153v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) has recently enabled highly photorealistic 3D reconstruction from casually captured multi-view images. However, this access

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

PBE-UNet: A light weight Progressive Boundary-Enhanced U-Net with Scale-Aware Aggregation for Ultrasound Image Segmentation

DGX agent

arXiv:2604.13791v1 Announce Type: new Abstract: Accurate lesion segmentation in ultrasound images is essential for preventive screening and clinical diagnosis, yet remains challenging due to low contr

model-releasesarxiv-cs-cv
16 Apr 2026
Research

Person Re-Identification via Generalized Class Prototypes

DGX agent

arXiv:2510.17043v2 Announce Type: replace Abstract: Advanced feature extraction methods have significantly contributed to enhancing the task of person re-identification. In addition, modifications to

researcharxiv-cs-cv
16 Apr 2026
Research

Physically-Guided Optical Inversion Enable Non-Contact Side-Channel Attack on Isolated Screens

DGX agent

arXiv:2604.13419v1 Announce Type: new Abstract: Noncontact exfiltration of electronic screen content poses a security challenge, with side-channel incursions as the principal vector. We introduce an o

researcharxiv-cs-cv
16 Apr 2026
Agents

POINTS-Seeker: Towards Training a Multimodal Agentic Search Model from Scratch

DGX agent

arXiv:2604.14029v1 Announce Type: new Abstract: While Large Multimodal Models (LMMs) demonstrate impressive visual perception, they remain epistemically constrained by their static parametric knowledg

agentsarxiv-cs-cv
16 Apr 2026
Applications

PostureObjectstitch: Anomaly Image Generation Considering Assembly Relationships in Industrial Scenarios

DGX agent

arXiv:2604.13863v1 Announce Type: new Abstract: Image generation technology can synthesize condition-specific images to supplement real-world industrial anomaly data and enhance anomaly detection mode

applicationsarxiv-cs-cv
16 Apr 2026
Applications

Radar-Informed 3D Multi-Object Tracking under Adverse Conditions

DGX agent

arXiv:2604.13571v1 Announce Type: new Abstract: The challenge of 3D multi-object tracking (3D MOT) is achieving robustness in real-world applications, for example under adverse conditions and maintain

applicationsarxiv-cs-cv
16 Apr 2026
Research

RadarSplat-RIO: Indoor Radar-Inertial Odometry with Gaussian Splatting-Based Radar Bundle Adjustment

DGX agent

arXiv:2604.13492v1 Announce Type: cross Abstract: Radar is more resilient to adverse weather and lighting conditions than visual and Lidar simultaneous localization and mapping (SLAM). However, most r

researcharxiv-cs-cv
16 Apr 2026
Research

Reconstruction of a 3D wireframe from a single line drawing via generative depth estimation

DGX agent

arXiv:2604.13549v1 Announce Type: new Abstract: The conversion of 2D freehand sketches into 3D models remains a pivotal challenge in computer vision, bridging the gap between human creativity and digi

researcharxiv-cs-cv
16 Apr 2026
Model Releases

ReConText3D: Replay-based Continual Text-to-3D Generation

DGX agent

arXiv:2604.13730v1 Announce Type: new Abstract: Continual learning enables models to acquire new knowledge over time while retaining previously learned capabilities. However, its application to text-t

model-releasesarxiv-cs-cv
16 Apr 2026
Local Ai

Remote Sensing Image Super-Resolution for Imbalanced Textures: A Texture-Aware Diffusion Framework

DGX agent

arXiv:2604.13994v1 Announce Type: new Abstract: Generative diffusion priors have recently achieved state-of-the-art performance in natural image super-resolution, demonstrating a powerful capability t

local-aiarxiv-cs-cv
16 Apr 2026
Local Ai

Rethinking Image-to-3D Generation with Sparse Queries: Efficiency, Capacity, and Input-View Bias

DGX agent

arXiv:2604.13905v1 Announce Type: new Abstract: We present SparseGen, a novel framework for efficient image-to-3D generation, which exhibits low input-view bias while being significantly faster. Unlik

local-aiarxiv-cs-cv
16 Apr 2026
Safety

Rethinking Uncertainty in Segmentation: From Estimation to Decision

DGX agent

arXiv:2604.13262v1 Announce Type: new Abstract: In medical image segmentation, uncertainty estimates are often reported but rarely used to guide decisions. We study the missing step: how uncertainty m

safetyarxiv-cs-cv
16 Apr 2026
Research

Right Regions, Wrong Labels: Semantic Label Flips in Segmentation under Correlation Shift

DGX agent

arXiv:2604.13326v1 Announce Type: new Abstract: The robustness of machine learning models can be compromised by spurious correlations between non-causal features in the input data and target labels. A

researcharxiv-cs-cv
16 Apr 2026
Safety

RoboTAG: End-to-end Robot Configuration Estimation via Topological Alignment Graph

DGX agent

arXiv:2511.07717v2 Announce Type: replace-cross Abstract: Estimating robot pose from a monocular RGB image is a challenge in robotics and computer vision. Existing methods typically build networks on

safetyarxiv-cs-cv
16 Apr 2026
Research

RobotPan: A 360^irc Surround-View Robotic Vision System for Embodied Perception

DGX agent

arXiv:2604.13476v1 Announce Type: cross Abstract: Surround-view perception is increasingly important for robotic navigation and loco-manipulation, especially in human-in-the-loop settings such as tele

researcharxiv-cs-cv
16 Apr 2026
Model Releases

ROSE: Retrieval-Oriented Segmentation Enhancement

DGX agent

arXiv:2604.14147v1 Announce Type: new Abstract: Existing segmentation models based on multimodal large language models (MLLMs), such as LISA, often struggle with novel or emerging entities due to thei

model-releasesarxiv-cs-cv
16 Apr 2026
Research

SceneGlue: Scene-Aware Transformer for Feature Matching without Scene-Level Annotation

DGX agent

arXiv:2604.13941v1 Announce Type: new Abstract: Local feature matching plays a critical role in understanding the correspondence between cross-view images. However, traditional methods are constrained

researcharxiv-cs-cv
16 Apr 2026
Research

SEDTalker: Emotion-Aware 3D Facial Animation Using Frame-Level Speech Emotion Diarization

DGX agent

arXiv:2604.13335v1 Announce Type: new Abstract: We introduce SEDTalker, an emotion-aware framework for speech-driven 3D facial animation that leverages frame-level speech emotion diarization to achiev

researcharxiv-cs-cv
16 Apr 2026
Model Releases

Seedance 2.0: Advancing Video Generation for World Complexity

DGX agent

arXiv:2604.14148v1 Announce Type: new Abstract: Seedance 2.0 is a new native multi-modal audio-video generation model, officially released in China in early February 2026. Compared with its predecesso

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

Seek-and-Solve: Benchmarking MLLMs for Visual Clue-Driven Reasoning in Daily Scenarios

DGX agent

arXiv:2604.14041v1 Announce Type: new Abstract: Daily scenarios are characterized by visual richness, requiring Multimodal Large Language Models (MLLMs) to filter noise and identify decisive visual cl

model-releasesarxiv-cs-cv
16 Apr 2026
Safety

See&Say: Vision Language Guided Safe Zone Detection for Autonomous Package Delivery Drones

DGX agent

arXiv:2604.13292v1 Announce Type: new Abstract: Autonomous drone delivery systems are rapidly advancing, but ensuring safe and reliable package drop-offs remains highly challenging in cluttered urban

safetyarxiv-cs-cv
16 Apr 2026
Model Releases

SemAttNet: Towards Attention-based Semantic Aware Guided Depth Completion

DGX agent

arXiv:2204.13635v2 Announce Type: replace Abstract: Depth completion involves recovering a dense depth map from a sparse map and an RGB image. Recent approaches focus on utilizing color images as guid

model-releasesarxiv-cs-cv
16 Apr 2026
Hardware

SemiFA: An Agentic Multi-Modal Framework for Autonomous Semiconductor Failure Analysis Report Generation

DGX agent

arXiv:2604.13236v1 Announce Type: new Abstract: Semiconductor failure analysis (FA) requires engineers to examine inspection images, correlate equipment telemetry, consult historical defect records, a

hardwarearxiv-cs-cv
16 Apr 2026
Research

SiLVR: A Simple Language-based Video Reasoning Framework

DGX agent

arXiv:2505.24869v3 Announce Type: replace Abstract: Recent advances in test-time optimization have led to remarkable reasoning capabilities in Large Language Models (LLMs), enabling them to solve high

researcharxiv-cs-cv
16 Apr 2026
Local Ai

Simplicity Prevails: The Emergence of Generalizable AIGI Detection in Visual Foundation Models

DGX agent

arXiv:2602.01738v2 Announce Type: replace Abstract: While specialized detectors for AI-Generated Images (AIGI) achieve near-perfect accuracy on curated benchmarks, they suffer from a dramatic performa

local-aiarxiv-cs-cv
16 Apr 2026
Model Releases

SLQ: Bridging Modalities via Shared Latent Queries for Retrieval with Frozen MLLMs

DGX agent

arXiv:2604.13710v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) exhibit strong reasoning and world knowledge, yet adapting them for retrieval remains challenging. Existing app

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

SocialMirror: Reconstructing 3D Human Interaction Behaviors from Monocular Videos with Semantic and Geometric Guidance

DGX agent

arXiv:2604.13581v1 Announce Type: new Abstract: Accurately reconstructing human behavior in close-interaction scenarios is crucial for enabling realistic virtual interactions in augmented reality, pre

model-releasesarxiv-cs-cv
16 Apr 2026
Research

SSD-GS: Scattering and Shadow Decomposition for Relightable 3D Gaussian Splatting

DGX agent

arXiv:2604.13333v1 Announce Type: new Abstract: We present SSD-GS, a physically-based relighting framework built upon 3D Gaussian Splatting (3DGS) that achieves high-quality reconstruction and photore

researcharxiv-cs-cv
16 Apr 2026
Model Releases

Target-Bench: Can Video World Models Achieve Mapless Path Planning with Semantic Targets?

DGX agent

arXiv:2511.17792v2 Announce Type: replace Abstract: While recent video world models can generate highly realistic videos, their ability to perform semantic reasoning and planning remains unclear and u

model-releasesarxiv-cs-cv
16 Apr 2026
Safety

Temporally Consistent Long-Term Memory for 3D Single Object Tracking

DGX agent

arXiv:2604.13789v1 Announce Type: new Abstract: 3D Single Object Tracking (3D-SOT) aims to localize a target object across a sequence of LiDAR point clouds, given its 3D bounding box in the first fram

safetyarxiv-cs-cv
16 Apr 2026
Research

The Gaussian Latent Machine: Efficient Prior and Posterior Sampling for Inverse Problems

DGX agent

arXiv:2505.12836v2 Announce Type: replace-cross Abstract: We consider the problem of sampling from a product-of-experts-type model that encompasses many standard prior and posterior distributions comm

researcharxiv-cs-cv
16 Apr 2026
Research

The Spectrascapes Dataset: Street-view imagery beyond the visible captured using a mobile platform

DGX agent

arXiv:2604.13315v1 Announce Type: new Abstract: High-resolution data in spatial and temporal contexts is imperative for developing climate resilient cities. Current datasets for monitoring urban param

researcharxiv-cs-cv
16 Apr 2026
Tutorials

Tokenizing Semantic Segmentation with Run Length Encoding

DGX agent

arXiv:2602.21627v3 Announce Type: replace Abstract: This paper presents a new unified approach to semantic segmentation in both images and videos by using language modeling to output the masks as sequ

tutorialsarxiv-cs-cv
16 Apr 2026
← Previous
1…240241242243244…261
Next →