AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,515 results
Applications

Amodal SAM: A Unified Amodal Segmentation Framework with Generalization

DGX agent

arXiv:2604.20748v1 Announce Type: new Abstract: Amodal segmentation is a challenging task that aims to predict the complete geometric shape of objects, including their occluded regions. Although exist

applicationsarxiv-cs-cv
23 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Automated Description Generation of Cytologic Findings for Lung Cytological Images Using a Pretrained Vision Model and Dual Text Decoders: Preliminary Study

DGX agent

arXiv:2403.18151v2 Announce Type: replace-cross Abstract: Objective: Cytology plays a crucial role in lung cancer diagnosis. Pulmonary cytology involves cell morphological characterization in the spec

researcharxiv-cs-cv
23 Apr 2026
Model Releases

Benchmarking ResNet for Short-Term Hypoglycemia Classification with DiaData

DGX agent

arXiv:2511.02849v2 Announce Type: replace-cross Abstract: Individualized therapy is driven forward by medical data analysis, which provides insight into the patient's context. In particular, for Type

model-releasesarxiv-cs-cv
23 Apr 2026
Research

Bio-inspired Color Constancy: From Gray Anchoring Theory to Gray Pixel Methods

DGX agent

arXiv:2604.20243v1 Announce Type: new Abstract: Color constancy is a fundamental ability of many biological visual systems and a crucial step in computer imaging systems. Bio-inspired modeling offers

researcharxiv-cs-cv
23 Apr 2026
Tutorials

Camera Control for Text-to-Image Generation via Learning Viewpoint Tokens

DGX agent

arXiv:2604.19954v1 Announce Type: new Abstract: Current text-to-image models struggle to provide precise camera control using natural language alone. In this work, we present a framework for precise c

tutorialsarxiv-cs-cv
23 Apr 2026
Model Releases

CCTVBench: Contrastive Consistency Traffic VideoQA Benchmark for Multimodal LLMs

DGX agent

arXiv:2604.20460v1 Announce Type: new Abstract: Safety-critical traffic reasoning requires contrastive consistency: models must detect true hazards when an accident occurs, and reliably reject plausib

model-releasesarxiv-cs-cv
23 Apr 2026
Safety

CLIP-RD: Relative Distillation for Efficient CLIP Knowledge Distillation

DGX agent

arXiv:2603.25383v3 Announce Type: replace Abstract: CLIP aligns image and text embeddings via contrastive learning and demonstrates strong zero-shot generalization. Its large-scale architecture requir

safetyarxiv-cs-cv
23 Apr 2026
Model Releases

ConeSep: Cone-based Robust Noise-Unlearning Compositional Network for Composed Image Retrieval

DGX agent

arXiv:2604.20358v1 Announce Type: new Abstract: The Composed Image Retrieval (CIR) task provides a flexible retrieval paradigm via a reference image and modification text, but it heavily relies on exp

model-releasesarxiv-cs-cv
23 Apr 2026
Research

Confidence-Based Mesh Extraction from 3D Gaussians

DGX agent

arXiv:2603.24725v2 Announce Type: replace Abstract: Recently, 3D Gaussian Splatting (3DGS) greatly accelerated mesh extraction from posed images due to its explicit representation and fast software ra

researcharxiv-cs-cv
23 Apr 2026
Safety

CoRe: Joint Optimization with Contrastive Learning for Medical Image Registration

DGX agent

arXiv:2603.23694v2 Announce Type: replace Abstract: Medical image registration is a fundamental task in medical image analysis, enabling the alignment of images from different modalities or time point

safetyarxiv-cs-cv
23 Apr 2026
Research

CrackForward: Context-Aware Severity Stage Crack Synthesis for Data Augmentation

DGX agent

arXiv:2604.19941v1 Announce Type: new Abstract: Reliable crack detection and segmentation are vital for structural health monitoring, yet the scarcity of well-annotated data constitutes a major challe

researcharxiv-cs-cv
23 Apr 2026
Research

CXR-LanIC: Language-Grounded Interpretable Classifier for Chest X-Ray Diagnosis

DGX agent

arXiv:2510.21464v2 Announce Type: replace Abstract: Deep learning models have achieved remarkable accuracy in chest X-ray diagnosis, yet their widespread clinical adoption remains limited by the black

researcharxiv-cs-cv
23 Apr 2026
Agents

DeVI: Physics-based Dexterous Human-Object Interaction via Synthetic Video Imitation

DGX agent

arXiv:2604.20841v1 Announce Type: new Abstract: Recent advances in video generative models enable the synthesis of realistic human-object interaction videos across a wide range of scenarios and object

agentsarxiv-cs-cv
23 Apr 2026
Research

Diagnosing Urban Street Vitality via a Visual-Semantic and Spatiotemporal Framework for Street-Level Economics

DGX agent

arXiv:2604.19798v1 Announce Type: cross Abstract: Micro-scale street-level economic assessment is fundamental for precision spatial resource allocation. While Street View Imagery (SVI) advances urban

researcharxiv-cs-cv
23 Apr 2026
Local Ai

DynamicRad: Content-Adaptive Sparse Attention for Long Video Diffusion

DGX agent

arXiv:2604.20470v1 Announce Type: new Abstract: Leveraging the natural spatiotemporal energy decay in video diffusion offers a path to efficiency, yet relying solely on rigid static masks risks losing

local-aiarxiv-cs-cv
23 Apr 2026
Applications

Efficient INT8 Single-Image Super-Resolution via Deployment-Aware Quantization and Teacher-Guided Training

DGX agent

arXiv:2604.20291v1 Announce Type: new Abstract: Efficient single-image super-resolution (SISR) requires balancing reconstruction fidelity, model compactness, and robustness under low-bit deployment, w

applicationsarxiv-cs-cv
23 Apr 2026
Applications

Energy-Based Open-Set Active Learning for Object Classification

DGX agent

arXiv:2604.20083v1 Announce Type: cross Abstract: Active learning (AL) has emerged as a crucial methodology for minimizing labeling costs in deep learning by selecting the most valuable samples from a

applicationsarxiv-cs-cv
23 Apr 2026
Research

Excretion Detection in Pigsties Using Convolutional and Transformerbased Deep Neural Networks

DGX agent

arXiv:2412.00256v3 Announce Type: replace Abstract: Animal excretions in form of urine puddles and feces are a significant source of emissions in livestock farming. Automated detection of soiled floor

researcharxiv-cs-cv
23 Apr 2026
Tutorials

Exploring High-Order Self-Similarity for Video Understanding

DGX agent

arXiv:2604.20760v1 Announce Type: new Abstract: Space-time self-similarity (STSS), which captures visual correspondences across frames, provides an effective way to represent temporal dynamics for vid

tutorialsarxiv-cs-cv
23 Apr 2026
Model Releases

Exploring Spatial Intelligence from a Generative Perspective

DGX agent

arXiv:2604.20570v1 Announce Type: new Abstract: Spatial intelligence is essential for multimodal large language models, yet current benchmarks largely assess it only from an understanding perspective.

model-releasesarxiv-cs-cv
23 Apr 2026
Local Ai

FA-Seg: A Fast and Accurate Diffusion-Based Method for Open-Vocabulary Segmentation

DGX agent

arXiv:2506.23323v5 Announce Type: replace Abstract: Open-vocabulary semantic segmentation (OVSS) aims to segment objects from arbitrary text categories without requiring densely annotated datasets. Al

local-aiarxiv-cs-cv
23 Apr 2026
Research

Fast Amortized Fitting of Scientific Signals Across Time and Ensembles via Transferable Neural Fields

DGX agent

arXiv:2604.19979v1 Announce Type: cross Abstract: Neural fields, also known as implicit neural representations (INRs), offer a powerful framework for modeling continuous geometry, but their effectiven

researcharxiv-cs-cv
23 Apr 2026
Model Releases

Fast-then-Fine: A Two-Stage Framework with Multi-Granular Representation for Cross-Modal Retrieval in Remote Sensing

DGX agent

arXiv:2604.20429v1 Announce Type: new Abstract: Remote sensing (RS) image-text retrieval plays a critical role in understanding massive RS imagery. However, the dense multi-object distribution and com

model-releasesarxiv-cs-cv
23 Apr 2026
Safety

FluSplat: Sparse-View 3D Editing without Test-Time Optimization

DGX agent

arXiv:2604.20038v1 Announce Type: new Abstract: Recent advances in text-guided image editing and 3D Gaussian Splatting (3DGS) have enabled high-quality 3D scene manipulation. However, existing pipelin

safetyarxiv-cs-cv
23 Apr 2026
Research

Fourier Series Coder: A Novel Perspective on Angle Boundary Discontinuity Problem for Oriented Object Detection

DGX agent

arXiv:2604.20281v1 Announce Type: new Abstract: With the rapid advancement of intelligent driving and remote sensing, oriented object detection has gained widespread attention. However, achieving high

researcharxiv-cs-cv
23 Apr 2026
Research

From Competition to Synergy: Unlocking Reinforcement Learning for Subject-Driven Image Generation

DGX agent

arXiv:2510.18263v2 Announce Type: replace-cross Abstract: Subject-driven image generation models face a fundamental trade-off between identity preservation (fidelity) and prompt adherence (editability

researcharxiv-cs-cv
23 Apr 2026
Research

From Diffusion to Flow: Efficient Motion Generation in MotionGPT3

DGX agent

arXiv:2603.26747v2 Announce Type: replace Abstract: Recent text-driven motion generation methods span both discrete token-based approaches and continuous-latent formulations. MotionGPT3 exemplifies th

researcharxiv-cs-cv
23 Apr 2026
Tutorials

From Ideal to Real: Stable Video Object Removal under Imperfect Conditions

DGX agent

arXiv:2603.09283v2 Announce Type: replace Abstract: Removing objects from videos remains difficult in the presence of real-world imperfections such as shadows, abrupt motion, and defective masks. Exis

tutorialsarxiv-cs-cv
23 Apr 2026
Research

From Image to Music Language: A Two-Stage Structure Decoding Approach for Complex Polyphonic OMR

DGX agent

arXiv:2604.20522v1 Announce Type: cross Abstract: We propose a new approach for the second stage of a practical two-stage Optical Music Recognition (OMR) pipeline. Given symbol and event candidates fr

researcharxiv-cs-cv
23 Apr 2026
Safety

FurnSet: Exploiting Repeats for 3D Scene Reconstruction

DGX agent

arXiv:2604.20093v1 Announce Type: new Abstract: Single-view 3D scene reconstruction involves inferring both object geometry and spatial layout. Existing methods typically reconstruct objects independe

safetyarxiv-cs-cv
23 Apr 2026
Hardware

Gaussians on a Diet: High-Quality Memory-Bounded 3D Gaussian Splatting Training

DGX agent

arXiv:2604.20046v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) has revolutionized novel view synthesis with high-quality rendering through continuous aggregations of millions of 3D Gauss

hardwarearxiv-cs-cv
23 Apr 2026
Research

Generative Prior-Guided Neural Interface Reconstruction for 3D Electrical Impedance Tomography

DGX agent

arXiv:2505.16487v3 Announce Type: replace-cross Abstract: Reconstructing complex 3D interfaces from indirect measurements remains a grand challenge in scientific computing, particularly for ill-posed

researcharxiv-cs-cv
23 Apr 2026
Research

GeoRect4D: Geometry-Compatible Generative Rectification for Dynamic Sparse-View 3D Reconstruction

DGX agent

arXiv:2604.20784v1 Announce Type: new Abstract: Reconstructing dynamic 3D scenes from sparse multi-view videos is highly ill-posed, often leading to geometric collapse, trajectory drift, and floating

researcharxiv-cs-cv
23 Apr 2026
Research

GeoRelight: Learning Joint Geometrical Relighting and Reconstruction with Flexible Multi-Modal Diffusion Transformers

DGX agent

arXiv:2604.20715v1 Announce Type: new Abstract: Relighting a person from a single photo is an attractive but ill-posed task, as a 2D image ambiguously entangles 3D geometry, intrinsic appearance, and

researcharxiv-cs-cv
23 Apr 2026
Model Releases

Global Offshore Wind Infrastructure: Deployment and Operational Dynamics from Dense Sentinel-1 Time Series

DGX agent

arXiv:2604.20822v1 Announce Type: new Abstract: The offshore wind energy sector is expanding rapidly, increasing the need for independent, high-temporal-resolution monitoring of infrastructure deploym

model-releasesarxiv-cs-cv
23 Apr 2026
Research

GSCompleter: A Distillation-Free Plugin for Metric-Aware 3D Gaussian Splatting Completion in Seconds

DGX agent

arXiv:2604.20155v1 Announce Type: new Abstract: While 3D Gaussian Splatting (3DGS) has revolutionized real-time rendering, its performance degrades significantly under sparse-view extrapolation, manif

researcharxiv-cs-cv
23 Apr 2026
Research

Hallucination Early Detection in Diffusion Models

DGX agent

arXiv:2604.20354v1 Announce Type: new Abstract: Text-to-Image generation has seen significant advancements in output realism with the advent of diffusion models. However, diffusion models encounter di

researcharxiv-cs-cv
23 Apr 2026
Applications

Human-like Content Analysis for Generative AI with Language-Grounded Sparse Encoders

DGX agent

arXiv:2508.18236v4 Announce Type: replace Abstract: The rapid development of generative AI has transformed content creation, communication, and human development. However, this technology raises profo

applicationsarxiv-cs-cv
23 Apr 2026
Research

HumanScore: Benchmarking Human Motions in Generated Videos

DGX agent

arXiv:2604.20157v1 Announce Type: new Abstract: Recent advances in model architectures, compute, and data scale have driven rapid progress in video generation, producing increasingly realistic content

researcharxiv-cs-cv
23 Apr 2026
Safety

Hybrid Latent Reasoning with Decoupled Policy Optimization

DGX agent

arXiv:2604.20328v1 Announce Type: new Abstract: Chain-of-Thought (CoT) reasoning significantly elevates the complex problem-solving capabilities of multimodal large language models (MLLMs). However, a

safetyarxiv-cs-cv
23 Apr 2026
Safety

i-WiViG: Interpretable Window Vision GNN

DGX agent

arXiv:2503.08321v2 Announce Type: replace Abstract: Vision graph neural networks have emerged as a popular approach for modeling the global and spatial context for image recognition. However, a signif

safetyarxiv-cs-cv
23 Apr 2026
Research

Improving Facial Emotion Recognition through Dataset Merging and Balanced Training Strategies

DGX agent

arXiv:2604.20307v1 Announce Type: new Abstract: In this paper, a deep learning framework is proposed for automatic facial emotion based on deep convolutional networks. In order to increase the general

researcharxiv-cs-cv
23 Apr 2026
Research

Integrated AI Nodule Detection and Diagnosis for Lung Cancer Screening Beyond Size and Growth-Based Standards Compared with Radiologists and Leading Models

DGX agent

arXiv:2512.00281v2 Announce Type: replace Abstract: Early detection of malignant lung nodules remains limited by reliance on size- and growth-based screening criteria, which can delay diagnosis. We pr

researcharxiv-cs-cv
23 Apr 2026
Research

Investigation of cardinality classification for bacterial colony counting using explainable artificial intelligence

DGX agent

arXiv:2604.20026v1 Announce Type: new Abstract: Automatic bacterial colony counting is a highly sought-after technology in modern biological laboratories because it eliminates manual counting effort.

researcharxiv-cs-cv
23 Apr 2026
Research

KD-Judge: A Knowledge-Driven Automated Judge Framework for Functional Fitness Movements on Edge Devices

DGX agent

arXiv:2604.19834v1 Announce Type: new Abstract: Functional fitness movements are widely used in training, competition, and health-oriented exercise programs, yet consistently enforcing repetition (rep

researcharxiv-cs-cv
23 Apr 2026
Applications

Learn2Synth: Learning Optimal Data Synthesis Using Hypergradients for Brain Image Segmentation

DGX agent

arXiv:2411.16719v4 Announce Type: replace Abstract: Domain randomization through synthesis is a powerful strategy to train networks that are unbiased with respect to the domain of the input images. Ra

applicationsarxiv-cs-cv
23 Apr 2026
Local Ai

Learning Spatial-Temporal Coherent Correlations for Speech-Preserving Facial Expression Manipulation

DGX agent

arXiv:2604.20226v1 Announce Type: new Abstract: Speech-preserving facial expression manipulation (SPFEM) aims to modify facial emotions while meticulously maintaining the mouth animation associated wi

local-aiarxiv-cs-cv
23 Apr 2026
Safety

Learning to count small and clustered objects with application to bacterial colonies

DGX agent

arXiv:2604.20030v1 Announce Type: new Abstract: Automated bacterial colony counting from images is an important technique to obtain data required for the development of vaccines and antibiotics. Howev

safetyarxiv-cs-cv
23 Apr 2026
← Previous
1…219220221222223…261
Next →