AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,515 results
Research

Edge-preserving noise for diffusion models

DGX agent

arXiv:2410.01540v4 Announce Type: replace Abstract: Classical diffusion models typically rely on isotropic Gaussian noise, treating all regions uniformly and overlooking structural information importa

researcharxiv-cs-cv
17 Apr 2026
Research

Efficient closed-form approaches for pose estimation using Sylvester forms

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2604.14747v1 Announce Type: new Abstract: Solving non-linear least-squares problem for pose estimation (rotation and translation) is often a time consuming yet fundamental problem in several rea

researcharxiv-cs-cv
17 Apr 2026
Research

Efficient Search of Implantable Adaptive Cells for Medical Image Segmentation

DGX agent

arXiv:2604.14849v1 Announce Type: new Abstract: Purpose: Adaptive skip modules can improve medical image segmentation, but searching for them is computationally costly. Implantable Adaptive Cells (IAC

researcharxiv-cs-cv
17 Apr 2026
Research

Enhancing LLM-Based Neural Network Generation: Few-Shot Prompting and Efficient Validation for Automated Architecture Design

DGX agent

arXiv:2512.24120v2 Announce Type: replace Abstract: Automated neural network architecture design remains a significant challenge in computer vision. Task diversity and computational constraints requir

researcharxiv-cs-cv
17 Apr 2026
Research

Estimating the Diameter at Breast Height of Trees in a Forest from RGB

DGX agent

arXiv:2505.03093v3 Announce Type: replace Abstract: Forest inventories rely on accurate measurements of the diameter at breast height (DBH) for ecological monitoring, resource management, and carbon a

researcharxiv-cs-cv
17 Apr 2026
Research

FADPNet: Frequency-Aware Dual-Path Network for Face Super-Resolution

DGX agent

arXiv:2506.14121v2 Announce Type: replace Abstract: Face super-resolution (FSR) under limited computational budgets remains challenging. Existing methods often treat all facial pixels equally, leading

researcharxiv-cs-cv
17 Apr 2026
Model Releases

FAIR Universe Weak Lensing ML Uncertainty Challenge: Handling Uncertainties and Distribution Shifts for Precision Cosmology

DGX agent

arXiv:2604.14451v1 Announce Type: cross Abstract: Weak gravitational lensing, the correlated distortion of background galaxy shapes by foreground structures, is a powerful probe of the matter distribu

model-releasesarxiv-cs-cv
17 Apr 2026
Tutorials

Feature Extraction in the Remote Sensing Data Value Chain: A Systematic Review of Methods and Applications

DGX agent

arXiv:2510.18935v3 Announce Type: replace Abstract: Earth observation involves collecting, analyzing, and processing an ever-growing mass of data. This planetary data is crucial for addressing relevan

tutorialsarxiv-cs-cv
17 Apr 2026
Research

Federated Breast Cancer Detection Enhanced by Synthetic Ultrasound Image Augmentation

DGX agent

arXiv:2506.23334v3 Announce Type: replace-cross Abstract: Federated learning enables collaborative training of deep learning models across institutions without sharing sensitive patient data. However,

researcharxiv-cs-cv
17 Apr 2026
Local Ai

Few-Shot Left Atrial Wall Segmentation in 3D LGE MRI via Meta-Learning

DGX agent

arXiv:2603.24985v2 Announce Type: replace Abstract: Segmenting the left atrial wall from late gadolinium enhancement magnetic resonance images (MRI) is challenging due to the wall's thin geometry, low

local-aiarxiv-cs-cv
17 Apr 2026
Research

Find the Differences: Differential Morphing Attack Detection vs Face Recognition

DGX agent

arXiv:2604.14734v1 Announce Type: new Abstract: Morphing is a challenge to face recognition (FR) for which several morphing attack detection solutions have been proposed. We argue that face recognitio

researcharxiv-cs-cv
17 Apr 2026
Research

Flow of Truth: Proactive Temporal Forensics for Image-to-Video Generation

DGX agent

arXiv:2604.15003v1 Announce Type: new Abstract: The rapid rise of image-to-video (I2V) generation enables realistic videos to be created from a single image but also brings new forensic demands. Unlik

researcharxiv-cs-cv
17 Apr 2026
Model Releases

FoodSense: A Multisensory Food Dataset and Benchmark for Predicting Taste, Smell, Texture, and Sound from Images

DGX agent

arXiv:2604.14388v1 Announce Type: new Abstract: Humans routinely infer taste, smell, texture, and even sound from food images a phenomenon well studied in cognitive science. However, prior vision lang

model-releasesarxiv-cs-cv
17 Apr 2026
Model Releases

Frame forecasting in cine MRI using the PCA respiratory motion model: comparing recurrent neural networks trained online and transformers

DGX agent

arXiv:2410.05882v3 Announce Type: replace-cross Abstract: Respiratory motion complicates accurate irradiation of thoraco-abdominal tumors during radiotherapy, as treatment-system latency entails targe

model-releasesarxiv-cs-cv
17 Apr 2026
Model Releases

FreqTrack: Frequency Learning based Vision Transformer for RGB-Event Object Tracking

DGX agent

arXiv:2604.14526v1 Announce Type: new Abstract: Existing single-modal RGB trackers often face performance bottlenecks in complex dynamic scenes, while the introduction of event sensors offers new pote

model-releasesarxiv-cs-cv
17 Apr 2026
Model Releases

Frequency-Enhanced Dual-Subspace Networks for Few-Shot Fine-Grained Image Classification

DGX agent

arXiv:2604.14958v1 Announce Type: new Abstract: Few-shot fine-grained image classification aims to recognize subcategories with high visual similarity using only a limited number of annotated samples.

model-releasesarxiv-cs-cv
17 Apr 2026
Safety

From Boundaries to Semantics: Prompt-Guided Multi-Task Learning for Petrographic Thin-section Segmentation

DGX agent

arXiv:2604.14805v1 Announce Type: new Abstract: Grain-edge segmentation (GES) and lithology semantic segmentation (LSS) are two pivotal tasks for quantifying rock fabric and composition. However, thes

safetyarxiv-cs-cv
17 Apr 2026
Model Releases

From Memorization to Creativity: LLM as a Designer of Novel Neural Architectures

DGX agent

arXiv:2601.02997v2 Announce Type: replace-cross Abstract: Large language models (LLMs) excel in program synthesis, yet their capacity for neural architecture design -- balancing syntactic reliability,

model-releasesarxiv-cs-cv
17 Apr 2026
Research

FSDETR: Frequency-Spatial Feature Enhancement for Small Object Detection

DGX agent

arXiv:2604.14884v1 Announce Type: new Abstract: Small object detection remains a significant challenge due to feature degradation from downsampling, mutual occlusion in dense clusters, and complex bac

researcharxiv-cs-cv
17 Apr 2026
Research

G-MIXER: Geodesic Mixup-based Implicit Semantic Expansion and Explicit Semantic Re-ranking for Zero-Shot Composed Image Retrieval

DGX agent

arXiv:2604.14710v1 Announce Type: new Abstract: Composed Image Retrieval (CIR) aims to retrieve target images by integrating a reference image with a corresponding modification text. CIR requires join

researcharxiv-cs-cv
17 Apr 2026
Research

Generative Data Augmentation for Skeleton Action Recognition

DGX agent

arXiv:2604.14933v1 Announce Type: new Abstract: Skeleton-based human action recognition is a powerful approach for understanding human behaviour from pose data, but collecting large-scale, diverse, an

researcharxiv-cs-cv
17 Apr 2026
Research

Generative Modeling of Complex-Valued Brain MRI Data

DGX agent

arXiv:2604.14800v1 Announce Type: cross Abstract: Objective. Standard Magnetic Resonance Imaging (MRI) reconstruction pipelines discard phase information captured during acquisition, despite evidence

researcharxiv-cs-cv
17 Apr 2026
Research

Geometrically Consistent Multi-View Scene Generation from Freehand Sketches

DGX agent

arXiv:2604.14302v1 Announce Type: new Abstract: We tackle a new problem: generating geometrically consistent multi-view scenes from a single freehand sketch. Freehand sketches are the most geometrical

researcharxiv-cs-cv
17 Apr 2026
Research

Giving Faces Their Feelings Back: Explicit Emotion Control for Feedforward Single-Image 3D Head Avatars

DGX agent

arXiv:2604.14541v1 Announce Type: new Abstract: We present a framework for explicit emotion control in feed-forward, single-image 3D head avatar reconstruction. Unlike existing pipelines where emotion

researcharxiv-cs-cv
17 Apr 2026
Local Ai

GlobalSplat: Efficient Feed-Forward 3D Gaussian Splatting via Global Scene Tokens

DGX agent

arXiv:2604.15284v1 Announce Type: new Abstract: The efficient spatial allocation of primitives serves as the foundation of 3D Gaussian Splatting, as it directly dictates the synergy between representa

local-aiarxiv-cs-cv
17 Apr 2026
Safety

GSNR: Graph Smooth Null-Space Representation for Inverse Problems

DGX agent

arXiv:2602.20328v2 Announce Type: replace Abstract: Inverse problems in imaging are ill-posed, leading to infinitely many solutions consistent with the measurements due to the non-trivial null-space o

safetyarxiv-cs-cv
17 Apr 2026
Model Releases

H2VLR: Heterogeneous Hypergraph Vision-Language Reasoning for Few-Shot Anomaly Detection

DGX agent

arXiv:2604.14507v1 Announce Type: new Abstract: As a classic vision task, anomaly detection has been widely applied in industrial inspection and medical imaging. In this task, data scarcity is often a

model-releasesarxiv-cs-cv
17 Apr 2026
Research

HAMSA: Scanning-Free Vision State Space Models via SpectralPulseNet

DGX agent

arXiv:2604.14724v1 Announce Type: new Abstract: Vision State Space Models (SSMs) like Vim, VMamba, and SiMBA rely on complex scanning strategies to adapt sequential SSMs to process 2D images, introduc

researcharxiv-cs-cv
17 Apr 2026
Research

High-Speed Full-Color HDR Imaging via Unwrapping Modulo-Encoded Spike Streams

DGX agent

arXiv:2604.14632v1 Announce Type: new Abstract: Conventional RGB-based high dynamic range (HDR) imaging faces a fundamental trade-off between motion artifacts in multi-exposure captures and irreversib

researcharxiv-cs-cv
17 Apr 2026
Tutorials

How to Correctly Make Mistakes: A Framework for Constructing and Benchmarking Mistake Aware Egocentric Procedural Videos

DGX agent

arXiv:2604.15134v1 Announce Type: new Abstract: Reliable procedural monitoring in video requires exposure to naturally occurring human errors and the recoveries that follow. In egocentric recordings,

tutorialsarxiv-cs-cv
17 Apr 2026
Model Releases

HRDexDB: A Large-Scale Dataset of Dexterous Human and Robotic Hand Grasps

DGX agent

arXiv:2604.14944v1 Announce Type: cross Abstract: We present HRDexDB, a large-scale, multi-modal dataset of high-fidelity dexterous grasping sequences featuring both human and diverse robotic hands. U

model-releasesarxiv-cs-cv
17 Apr 2026
Research

HY-World 2.0: A Multi-Modal World Model for Reconstructing, Generating, and Simulating 3D Worlds

DGX agent

arXiv:2604.14268v1 Announce Type: new Abstract: We introduce HY-World 2.0, a multi-modal world model framework that advances our prior project HY-World 1.0. HY-World 2.0 accommodates diverse input mod

researcharxiv-cs-cv
17 Apr 2026
Safety

Hybrid Latents -- Geometry-Appearance-Aware Surfel Splatting

DGX agent

arXiv:2604.14928v1 Announce Type: new Abstract: We introduce a hybrid Gaussian-hash-grid radiance representation for reconstructing 2D Gaussian scene models from multi-view images. Similar to NeST spl

safetyarxiv-cs-cv
17 Apr 2026
Safety

Implicit Neural Representations: A Signal Processing Perspective

DGX agent

arXiv:2604.15047v1 Announce Type: new Abstract: Implicit neural representations (INRs) mark a fundamental shift in signal modeling, moving from discrete sampled data to continuous functional represent

safetyarxiv-cs-cv
17 Apr 2026
Research

Improved Multiscale Structural Mapping with Supervertex Vision Transformer for the Detection of Alzheimer's Disease Neurodegeneration

DGX agent

arXiv:2604.14837v1 Announce Type: new Abstract: Alzheimer's disease (AD) confirmation often relies on positron emission tomography (PET) or cerebrospinal fluid (CSF) analysis, which are costly and inv

researcharxiv-cs-cv
17 Apr 2026
Research

Improving Prostate Gland Segmentation Using Transformer based Architectures

DGX agent

arXiv:2506.14844v2 Announce Type: replace-cross Abstract: Inter reader variability and cross site domain shift challenge the automatic segmentation of prostate anatomy using T2 weighted MRI images. Th

researcharxiv-cs-cv
17 Apr 2026
Safety

Integrating Object Detection, LiDAR-Enhanced Depth Estimation, and Segmentation Models for Railway Environments

DGX agent

arXiv:2604.14781v1 Announce Type: new Abstract: Obstacle detection in railway environments is crucial for ensuring safety. However, very few studies address the problem using a complete, modular, and

safetyarxiv-cs-cv
17 Apr 2026
Local Ai

Interpretable Human Activity Recognition for Subtle Robbery Detection in Surveillance Videos

DGX agent

arXiv:2604.14329v1 Announce Type: new Abstract: Non-violent street robberies (snatch-and-run) are difficult to detect automatically because they are brief, subtle, and often indistinguishable from ben

local-aiarxiv-cs-cv
17 Apr 2026
Research

JoyVASA: Portrait and Animal Image Animation with Diffusion-Based Audio-Driven Facial Dynamics and Head Motion Generation

DGX agent

arXiv:2411.09209v5 Announce Type: replace Abstract: Audio-driven portrait animation has made significant advances with diffusion-based models, improving video quality and lipsync accuracy. However, th

researcharxiv-cs-cv
17 Apr 2026
Research

KVNN: Learnable Multi-Kernel Volterra Neural Networks

DGX agent

arXiv:2604.15141v1 Announce Type: new Abstract: Higher-order learning is fundamentally rooted in exploiting compositional features. It clearly hinges on enriching the representation by more elaborate

researcharxiv-cs-cv
17 Apr 2026
Model Releases

Label-efficient underwater species classification with logistic regression on frozen foundation model embeddings

DGX agent

arXiv:2604.00313v2 Announce Type: replace Abstract: Automated species classification from underwater imagery is bottlenecked by the cost of expert annotation, and supervised models trained on one data

model-releasesarxiv-cs-cv
17 Apr 2026
Local Ai

Large Vision Model-Guided Masked Low-Rank Approximation for Ground-Roll Attenuation

DGX agent

arXiv:2604.00998v2 Announce Type: replace Abstract: Ground roll is a common type of coherent noise in seismic records, and its attenuation remains challenging due to its substantial overlap with usefu

local-aiarxiv-cs-cv
17 Apr 2026
Research

Latent Wavelet Diffusion For Ultra-High-Resolution Image Synthesis

DGX agent

arXiv:2506.00433v4 Announce Type: replace Abstract: High-resolution image synthesis remains a core challenge in generative modeling, particularly in balancing computational efficiency with the preserv

researcharxiv-cs-cv
17 Apr 2026
Safety

LeapAlign: Post-Training Flow Matching Models at Any Generation Step by Building Two-Step Trajectories

DGX agent

arXiv:2604.15311v1 Announce Type: new Abstract: This paper focuses on the alignment of flow matching models with human preferences. A promising way is fine-tuning by directly backpropagating reward gr

safetyarxiv-cs-cv
17 Apr 2026
Model Releases

Learning Where to Embed: Noise-Aware Positional Embedding for Query Retrieval in Small-Object Detection

DGX agent

arXiv:2604.15065v1 Announce Type: new Abstract: Transformer-based detectors have advanced small-object detection, but they often remain inefficient and vulnerable to background-induced query noise, wh

model-releasesarxiv-cs-cv
17 Apr 2026
Model Releases

LLMOrbit: A Circular Taxonomy of Large Language Models -From Scaling Walls to Agentic AI Systems

DGX agent

arXiv:2601.14053v2 Announce Type: replace-cross Abstract: The field of artificial intelligence has undergone a revolution from foundational Transformer architectures to reasoning-capable systems appro

model-releasesarxiv-cs-cv
17 Apr 2026
Research

M3D-Net: Multi-Modal 3D Facial Feature Reconstruction Network for Deepfake Detection

DGX agent

arXiv:2604.14574v1 Announce Type: new Abstract: With the rapid advancement of deep learning in image generation, facial forgery techniques have achieved unprecedented realism, posing serious threats t

researcharxiv-cs-cv
17 Apr 2026
Research

MapSR: Prompt-Driven Land Cover Map Super-Resolution via Vision Foundation Models

DGX agent

arXiv:2604.14582v1 Announce Type: new Abstract: High-resolution (HR) land-cover mapping is often constrained by the high cost of dense HR annotations. We revisit this problem from the perspective of m

researcharxiv-cs-cv
17 Apr 2026
← Previous
1…236237238239240…261
Next →