AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,618 results
Tutorials

NGPS: Structure-Preserving Self-Supervised Denoising via Neighbor-Guided Patch Sampling

DGX agent

arXiv:2606.23200v1 Announce Type: cross Abstract: Neighboring-slice self-supervised denoising is attractive for volumetric medical imaging, yet inter-slice misalignment breaks anatomical correspondenc

tutorialsarxiv-cs-cv
23 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Research

NoduLoCC2026: Lung Nodule Localization and Classification Contest from Chest X-Ray Images

DGX agent

arXiv:2606.21290v1 Announce Type: new Abstract: We propose NoduLoCC2026, a challenge on lung nodule detection and localization in chest X-ray images. We have provided a dataset for both tasks and rece

researcharxiv-cs-cv
23 Jun 2026
Agents

Non-line-of-sight imaging with arbitrary relay surface geometries via 3D Gaussian Transient Rendering

DGX agent

arXiv:2606.21270v1 Announce Type: cross Abstract: Imaging objects hidden outside the direct line of sight expands the effective field of view and is critical for applications such as autonomous drivin

agentsarxiv-cs-cv
23 Jun 2026
Research

Null-Space Diffusion Distillation Unlocks Speed, Fidelity and Realism in Lensless Imaging

DGX agent

arXiv:2511.12024v3 Announce Type: replace Abstract: Lensless imaging reconstructs scenes from highly multiplexed measurements, resulting in a severely ill-posed inverse problem. In this work, we ident

researcharxiv-cs-cv
23 Jun 2026
Tutorials

NullFlow: One-Step Generative Reconstruction

DGX agent

arXiv:2606.22696v1 Announce Type: new Abstract: We propose NullFlow, a principled framework for one-step generative image reconstruction. Our key idea is to confine the generative flow to a measuremen

tutorialsarxiv-cs-cv
23 Jun 2026
Research

Object-Centric Dataset Resources for Constrained-Data Image Generation and Augmentation

DGX agent

arXiv:2606.21113v1 Announce Type: new Abstract: Object-centric image generation is important in settings with few labeled examples, including pedestrian analysis in smart-city scenes, traffic-sign ins

researcharxiv-cs-cv
23 Jun 2026
Research

Ocean4D: Generative Underwater 4D Reconstruction via Medium-Aware Video Diffusion

DGX agent

arXiv:2606.23298v1 Announce Type: new Abstract: Underwater 4D reconstruction remains challenging due to the coupling between degraded light transport in participating media and dynamic water variation

researcharxiv-cs-cv
23 Jun 2026
Research

Odoriko: A Shape-Aware Multimodal Diffusion Framework for Human Motion

DGX agent

arXiv:2606.21135v1 Announce Type: new Abstract: Human motion generation has been widely studied across diverse input modalities, text, music, and video, and recent efforts have unified these into sing

researcharxiv-cs-cv
23 Jun 2026
Safety

OmniNWM: Omniscient Driving Navigation World Models

DGX agent

arXiv:2510.18313v5 Announce Type: replace Abstract: Autonomous driving world models are expected to work effectively across three core dimensions: state, action, and reward. However, existing methods

safetyarxiv-cs-cv
23 Jun 2026
Agents

OmniSpace: Efficient Geometry Awareness for Autonomous Vehicles MLLMs

DGX agent

arXiv:2606.22617v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable performance on 2D visual tasks, yet enhancing their spatial intelligence for real-worl

agentsarxiv-cs-cv
23 Jun 2026
Research

On-Manifold Variational Learning with Heat-Kernel Priors

DGX agent

arXiv:2606.18658v2 Announce Type: replace Abstract: Learning unsupervised representations of medical imaging cohorts can reveal clinically meaningful prototypes without expert labels, which are often

researcharxiv-cs-cv
23 Jun 2026
Safety

One Image is All You Need: Agentic One-Shot Image Generation via Text-Based World Models for Long-Tail Spatial Perception

DGX agent

arXiv:2606.20764v1 Announce Type: new Abstract: Reliable spatial decision automation, such as autonomous driving and maritime surveillance, critically depends on robust visual perception. However, rea

safetyarxiv-cs-cv
23 Jun 2026
Local Ai

One-Shot Data Selection for Medical Image Classification via Graph Coverage

DGX agent

arXiv:2606.22002v1 Announce Type: new Abstract: Training medical image classifiers on entire datasets is wasteful when annotation budgets are limited: not all samples contribute equally, yet acquiring

local-aiarxiv-cs-cv
23 Jun 2026
Model Releases

Open Annotations and Synthetic Data for Field Localisation in Indian Bank Cheques

DGX agent

arXiv:2606.20682v1 Announce Type: new Abstract: Automated cheque processing requires localising key fields (date, legal amount, IFSC code, account number, signature, and payee name) before any recogni

model-releasesarxiv-cs-cv
23 Jun 2026
Research

OphthaDT: Generative Digital Twins for Forecasting Visual Acuity Trajectories in Ophthalmology

DGX agent

arXiv:2606.22101v1 Announce Type: cross Abstract: Precision medicine in ophthalmology requires accurate longitudinal predictions, but the fragmented nature of multimodal clinical data remains a barrie

researcharxiv-cs-cv
23 Jun 2026
Model Releases

Oracle-RLAIF: An Improved Fine-Tuning Framework for Multi-modal Video Models using Reinforcement Learning from Ranking Feedback

DGX agent

arXiv:2510.02561v2 Announce Type: replace Abstract: Recent advances in large video-language models (VLMs) rely on extensive fine-tuning techniques that strengthen alignment between textual and visual

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

ORBIT: Training-Free Multi-Attribute Behavioral Steering via Orthogonal Subspace Rotation

DGX agent

arXiv:2606.22357v1 Announce Type: cross Abstract: Language models are widely used in assistant settings, where controlling behavioral attributes is often essential. Activation steering modifies hidden

model-releasesarxiv-cs-cv
23 Jun 2026
Research

OrthoMotion:Disentangling Camera and Subject Motion via Geometry Semantics Orthogonal Attention

DGX agent

arXiv:2606.22835v1 Announce Type: new Abstract: Controllable video generation demands independent command of the camera and the subject, yet 2D conditioning entangles them: camera- and object-induced

researcharxiv-cs-cv
23 Jun 2026
Applications

OSOG: A Differentiable, Physics-Informed Synthetic Data Engine for Micro-Optical Environments

DGX agent

arXiv:2606.21381v1 Announce Type: new Abstract: Deep learning in computational microscopy is severely constrained by the scarcity of densely annotated datasets. While synthetic data generation has bri

applicationsarxiv-cs-cv
23 Jun 2026
Local Ai

P-JEPA: Procedural Video Representation Learning via Joint Embedding Predictive Architecture

DGX agent

arXiv:2606.23256v1 Announce Type: new Abstract: The increasing maturity of embodied AI platforms has driven a growing interest in procedural video representation learning to support intelligent assist

local-aiarxiv-cs-cv
23 Jun 2026
Research

PaaF: Raising the perceived quality of INR-Based Image Compression

DGX agent

arXiv:2606.21655v1 Announce Type: cross Abstract: Implicit Neural Representations (INRs) have recently emerged as a promising paradigm for image compression, offering a fundamentally different approac

researcharxiv-cs-cv
23 Jun 2026
Safety

PG-MAP: Joint MAP Optimization for Inference-Time Alignment of Diffusion and Flow-Matching Models

DGX agent

arXiv:2606.22958v1 Announce Type: cross Abstract: Inference-time alignment of pretrained text-to-image models is typically performed along a single control axis, such as classifier-free guidance, atte

safetyarxiv-cs-cv
23 Jun 2026
Research

PHAST-Net: Attention-Guided, Physics-Informed Network for Unified Estimation of Ideal Time-Frequency Representations

DGX agent

arXiv:2606.23665v1 Announce Type: cross Abstract: We introduce PHAST-Net, an attention-guided, physics-informed network for unified estimation of Ideal Time-Frequency Representations (ITFRs), spanning

researcharxiv-cs-cv
23 Jun 2026
Safety

phi-Scene: Physically Grounded Image-to-3D Scene Reconstruction

DGX agent

arXiv:2606.21596v1 Announce Type: new Abstract: Reconstructing compositional 3D scenes from a single image is a fundamental challenge in 3D world modeling. Recent methods can recover high-fidelity, co

safetyarxiv-cs-cv
23 Jun 2026
Model Releases

PHOEBI: An Open-World Benchmark for Bacterial Identification in Phase-Contrast Microscopy

DGX agent

arXiv:2606.22890v1 Announce Type: new Abstract: Optical microscopy enables rapid, label-free imaging of live bacteria and is the standard instrument for species identification across clinical, environ

model-releasesarxiv-cs-cv
23 Jun 2026
Tutorials

PhyGDPO: Physics-Aware Groupwise Direct Preference Optimization for Physically Consistent Text-to-Video Generation

DGX agent

arXiv:2512.24551v4 Announce Type: replace Abstract: Recent advances in text-to-video (T2V) generation have achieved good visual quality, yet synthesizing videos that faithfully follow physical laws re

tutorialsarxiv-cs-cv
23 Jun 2026
Model Releases

PhysFlow: Frequency Decoupled with Dual-Field Rectified Flow for Remote Photoplethysmography

DGX agent

arXiv:2606.23226v1 Announce Type: new Abstract: Remote Photoplethysmography (rPPG) enables contactless pulse estimation from facial videos, serving as a vital tool for health monitoring. However, curr

model-releasesarxiv-cs-cv
23 Jun 2026
Safety

Physically-guided Image Generation for Multi-Projection Mapping

DGX agent

arXiv:2606.22477v1 Announce Type: new Abstract: Projection Mapping (PM) enables seamless superimposition of digital content onto real-world 3D objects, serving as a fundamental technique for immersive

safetyarxiv-cs-cv
23 Jun 2026
Research

Physics-Guided Spatiotemporal State Space Modeling for Lookahead Molten Pool Segmentation in Laser Wire-Feed Welding

DGX agent

arXiv:2606.23028v1 Announce Type: new Abstract: Real-time weld-pool perception is critical for closed-loop control in laser wire-feed welding, where sensing, computation, and actuator response introdu

researcharxiv-cs-cv
23 Jun 2026
Research

PIAvatar: Physically Interactive Avatars via Deformation Gradient Decoupling

DGX agent

arXiv:2606.21162v1 Announce Type: cross Abstract: 3D human avatars have shown impressive visual fidelity driven by pose-conditioned models, yet they still lack the physical ability required for intera

researcharxiv-cs-cv
23 Jun 2026
Safety

PISCES: Annotation-free Text-to-Video Post-Training via Optimal Transport-Aligned Rewards

DGX agent

arXiv:2602.01624v2 Announce Type: replace Abstract: Text-to-video (T2V) generation aims to synthesize videos with high visual quality and temporal consistency that are semantically aligned with input

safetyarxiv-cs-cv
23 Jun 2026
Tutorials

Poisson2Gaussian: Noise Gaussianization to Enhance Image Denoising

DGX agent

arXiv:2606.23098v1 Announce Type: new Abstract: The quantum nature of light determines the inherent Poisson stochasticity of photon detection, which is ubiquitous in photography, microscopy, and astro

tutorialsarxiv-cs-cv
23 Jun 2026
Safety

Policy-as-Data: Learning Generalizable HOI Diffusion Models from Simulated Physics

DGX agent

arXiv:2606.22806v1 Announce Type: new Abstract: Synthesizing realistic Human-Object Interactions (HOI) is critical for creating embodied avatars and functional virtual environments. However, current d

safetyarxiv-cs-cv
23 Jun 2026
Safety

PolicyTrim: Boosting Intrinsic Policy Efficiency of Vision-Language-Action Models

DGX agent

arXiv:2606.22540v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models provide a unified paradigm for robotic manipulation, yet their real-world deployment is often bottlenecked by execut

safetyarxiv-cs-cv
23 Jun 2026
Model Releases

Polycepta: Object-Centric Appearance Estimation for Multi-Object Tracking

DGX agent

arXiv:2606.23604v1 Announce Type: new Abstract: The tracking-by-detection paradigm in multi-object tracking (MOT) typically relies on static appearance descriptors to complement motion estimation. How

model-releasesarxiv-cs-cv
23 Jun 2026
Research

Polynomial Dice Loss for Medical Image Segmentation

DGX agent

arXiv:2606.23373v1 Announce Type: new Abstract: Medical image segmentation is a fundamental task for medical image processing and computer-assisted intervention, yet data imbalance and small lesion de

researcharxiv-cs-cv
23 Jun 2026
Safety

Pose Anything Anywhere:Model-free Object Poses from Arbitrary References

DGX agent

arXiv:2606.23634v1 Announce Type: new Abstract: Estimating the 6D pose of unseen objects is a fundamental yet challenging problem for open-world robotics and embodied perception. Model-based methods a

safetyarxiv-cs-cv
23 Jun 2026
Model Releases

Precision Recall Controllable Radiology Report Generation via Hybrid Natural Language and Clinical Reward Learning

DGX agent

arXiv:2606.21447v1 Announce Type: cross Abstract: Automated radiology report generation (RRG) has gained increasing attention because it can reduce the heavy workload of clinical report writing. Howev

model-releasesarxiv-cs-cv
23 Jun 2026
Research

Predicting Immune Biomarkers with MultiModal Mixture-of-Expert Pathology Foundation Models Empowers Precision Oncology

DGX agent

arXiv:2606.18123v2 Announce Type: replace Abstract: Predicting immune biomarkers associated with the tumor immune microenvironment (TIME) is critical for advancing precision oncology, yet existing app

researcharxiv-cs-cv
23 Jun 2026
Research

Privacy-Preserving Person Re-Identification from Temporal Sequences with Transformer and Hungarian Optimization

DGX agent

arXiv:2606.23230v1 Announce Type: new Abstract: Person re-identification (Re-ID) is a crucial task in surveillance and human behavior analysis, often used in public spaces such as transport hubs. Trad

researcharxiv-cs-cv
23 Jun 2026
Research

Projection-Volume Fidelity Divergence: Diagnosing and Controlling Optimization Drift in Sparse-View 3D Gaussian Tomography

DGX agent

arXiv:2606.22525v1 Announce Type: new Abstract: Sparse-view computed tomography is a severely ill-posed inverse problem, where recent 3D Gaussian Splatting methods offer an efficient explicit represen

researcharxiv-cs-cv
23 Jun 2026
Research

Prompt-Calibrated SAM 3 for Open-Vocabulary Remote Sensing Semantic Segmentation

DGX agent

arXiv:2606.21863v1 Announce Type: new Abstract: Open-vocabulary semantic segmentation (OVSS) in remote sensing images aims to segment categories beyond a fixed label space. Recent SAM 3-based methods

researcharxiv-cs-cv
23 Jun 2026
Research

Prompting Diffusion Models for Zero-Shot Instance Segmentation

DGX agent

arXiv:2606.22660v1 Announce Type: new Abstract: Several disruptive research directions have recently emerged in computer vision, including foundation models achieving previously unseen zero-shot perfo

researcharxiv-cs-cv
23 Jun 2026
Model Releases

PROTON: Prototype-Based Test-Time Online OOD Detection for Medical VLMs

DGX agent

arXiv:2606.20913v1 Announce Type: new Abstract: Medical vision-language models (VLMs) enable zero-shot clinical image classification, yet reliably detecting out-of-distribution (OOD) inputs at deploym

model-releasesarxiv-cs-cv
23 Jun 2026
Research

Quantile Adaptive Temperature Scaling for Confidence Calibration

DGX agent

arXiv:2606.21749v1 Announce Type: new Abstract: Deep neural networks often produce poorly calibrated confidence estimates, overstating their certainty even when predictions are incorrect. Temperature

researcharxiv-cs-cv
23 Jun 2026
Research

Quantum Visual Fields with Neural Amplitude Encoding

DGX agent

arXiv:2508.10900v2 Announce Type: replace Abstract: Quantum Implicit Neural Representations (QINRs) have emerged as a promising paradigm that leverages parametrised quantum circuits to encode and proc

researcharxiv-cs-cv
23 Jun 2026
Research

Radial Basis Function Networks as Projection Heads in Self-Supervised Learning

DGX agent

arXiv:2606.21590v1 Announce Type: new Abstract: Self-supervised learning (SSL) typically relies on a backbone encoder followed by a small multilayer perceptron (MLP) projection head, which is conventi

researcharxiv-cs-cv
23 Jun 2026
Agents

RAPID: A Reproducible Multi-Agent Pipeline for Interpretable Disaster Damage Assessment from Satellite and Street-View Imagery

DGX agent

arXiv:2606.21819v1 Announce Type: new Abstract: Due to the increasing frequency and intensity of extreme climate events, there is a clear demand for intelligent, scalable, and autonomous approaches to

agentsarxiv-cs-cv
23 Jun 2026
← Previous
1…100101102103104…263
Next →