AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlog
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,618 results
Model Releases

A Physics-Grounded Benchmark for Multi-Agent Dynamics in World Models

DGX agent

arXiv:2606.28757v1 Announce Type: new Abstract: Generative world models hold immense promise as scalable simulators for autonomous systems, particularly for synthesizing rare but safety-critical multi

model-releasesarxiv-cs-cv
30 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Research

A Point Cloud Transformer for Remote Monitoring and Automated Assessment of Physical Rehabilitation Exercises

DGX agent

arXiv:2606.30309v1 Announce Type: new Abstract: Rehabilitation exercises are essential in restoring lost physical functions of patients suffering from various diseases (e.g., Parkinson's, back pain).

researcharxiv-cs-cv
30 Jun 2026
Research

A Self-Supervised Learning Framework for Video Encoding Complexity Clustering

DGX agent

arXiv:2606.29166v1 Announce Type: cross Abstract: Adaptive video streaming is a widely used technique for delivering video content over the internet. One of the key challenges is determining the optim

researcharxiv-cs-cv
30 Jun 2026
Research

A Zero-Shot Deep Image Prior Framework for Denoising and Deconvolution in Fluorescence Microscopy

DGX agent

arXiv:2606.28431v1 Announce Type: cross Abstract: Fluorescence microscopy images are degraded by noise and diffraction-induced blur, which compromise structural fidelity and limit quantitative analysi

researcharxiv-cs-cv
30 Jun 2026
Safety

AccelAes: Accelerating Diffusion Transformers for Training-Free Aesthetic-Enhanced Image Generation

DGX agent

arXiv:2603.12575v2 Announce Type: replace Abstract: Diffusion Transformers (DiTs) are a dominant backbone for high-fidelity text-to-image generation due to strong scalability and alignment at high res

safetyarxiv-cs-cv
30 Jun 2026
Safety

Accurate Recognition of Pneumonia and COVID-19 by Geometric Shape Normalization of Lung Region using Automatic Landmark Detection and Piecewise Affine Warping

DGX agent

arXiv:2606.29715v1 Announce Type: new Abstract: This paper presents an automatic system for recognizing pulmonary diseases in chest X-rays using geometric normalization of the lung region. The method

safetyarxiv-cs-cv
30 Jun 2026
Research

AD-DAE: Alzheimer's Disease Progression Modeling with Unpaired Longitudinal MRI using Diffusion Auto-Encoders

DGX agent

arXiv:2511.05934v2 Announce Type: replace Abstract: Generative modeling frameworks have emerged as an effective approach to capture high-dimensional image distributions from large datasets without req

researcharxiv-cs-cv
30 Jun 2026
Research

Adaptive Spectrum-Aware Feature Disentangled Network for Small Object Detection

DGX agent

arXiv:2606.29029v1 Announce Type: new Abstract: Small Object Detection (SOD) is a fundamental yet challenging problem in computer vision due to its limited spatial resolution and weak visual cues. Alt

researcharxiv-cs-cv
30 Jun 2026
Applications

AEGIR: Modeling Area Emitters for Indoor Inverse Rendering using Gaussian Splatting

DGX agent

arXiv:2606.28635v1 Announce Type: new Abstract: Inverse rendering requires separating illumination from surface materials, which is highly ambiguous due to their tight coupling in observed images. Whi

applicationsarxiv-cs-cv
30 Jun 2026
Model Releases

AerialMetric: Benchmarking and Adapting UAV Monocular Metric Depth Estimation in the Real World

DGX agent

arXiv:2606.29716v1 Announce Type: new Abstract: This paper addresses the problem of monocular metric depth estimation in aerial UAV imagery. Although recent data-driven methods have achieved remarkabl

model-releasesarxiv-cs-cv
30 Jun 2026
Research

Again-Pose: Anchor-Guided Adaptive Inter-Frame Motion Cues Propagating for High-quality Human Pose Reconstruction

DGX agent

arXiv:2606.29230v1 Announce Type: new Abstract: Reconstructing continuous 3D human poses from unconstrained videos is challenging, especially in extreme motion scenarios involving severe motion blur a

researcharxiv-cs-cv
30 Jun 2026
Applications

AHOY! Animatable Humans under Occlusion from YouTube Videos with Gaussian Splatting and Video Diffusion Priors

DGX agent

arXiv:2603.17975v2 Announce Type: replace Abstract: We present AHOY, a method for reconstructing complete, animatable 3D Gaussian avatars from in-the-wild monocular video despite heavy occlusion. Exis

applicationsarxiv-cs-cv
30 Jun 2026
Research

Anatomy-Grounded Synthetic Coronary Angiography for Geometry-Informed Multi-View Matching

DGX agent

arXiv:2606.28474v1 Announce Type: cross Abstract: Accurate correspondence matching across multiple angiographic views is the prerequisite for 3D coronary reconstruction and interventional guidance. Ho

researcharxiv-cs-cv
30 Jun 2026
Research

APRIL-MedSeg: A Modular Medical Image Segmentation Toolbox Embracing Modern Paradigms

DGX agent

arXiv:2606.30577v1 Announce Type: new Abstract: We present APRIL-MedSeg, a YAML-driven modular framework for 2D medical image segmentation. It provides a unified and extensible ecosystem that decompos

researcharxiv-cs-cv
30 Jun 2026
Model Releases

Argus: Metric Panoramic 3D Reconstruction for Indoor Scenes

DGX agent

arXiv:2606.30047v1 Announce Type: new Abstract: Metric feed-forward 3D reconstruction for panoramic data remains under-explored due to the lack of large-scale panoramic RGB-D training data. We present

model-releasesarxiv-cs-cv
30 Jun 2026
Applications

Articulating then Matching: Zero-Shot Shape Matching for Uncurated Data

DGX agent

arXiv:2606.29167v1 Announce Type: new Abstract: Finding dense correspondences between 3D shapes is a fundamental yet unresolved challenge, especially in real-world environments. These environments pre

applicationsarxiv-cs-cv
30 Jun 2026
Agents

ASTAD: Asymmetric Style Transfer for Synthetic-to-Real Adaptation in Autonomous Driving

DGX agent

arXiv:2606.29286v1 Announce Type: new Abstract: Synthetic data mitigates the data scarcity problem in autonomous driving perception. However, the synthetic-to-real gap leads to performance degradation

agentsarxiv-cs-cv
30 Jun 2026
Research

AsyncMDE: Real-Time Monocular Depth Estimation via Asynchronous Spatial Memory

DGX agent

arXiv:2603.10438v2 Announce Type: replace-cross Abstract: Foundation-model-based monocular depth estimation offers a viable alternative to active sensors for robot perception, yet its computational co

researcharxiv-cs-cv
30 Jun 2026
Agents

BackTranslation2.0 -- A Linguistically Motivated Metric to Assess Sign Language Production

DGX agent

arXiv:2606.28673v1 Announce Type: new Abstract: Sign Languages (SLs) are the primary means of communication for millions of deaf individuals, yet existing evaluation metrics for generated SL remain si

agentsarxiv-cs-cv
30 Jun 2026
Model Releases

Benchmark AUC Is Not Deployable Reliability: A Cross-Dataset Audit of Off-the-Shelf Features for Surveillance Video Anomaly Detection

DGX agent

arXiv:2606.29506v1 Announce Type: new Abstract: Automated 'suspicious behavior' flagging is a headline promise of AI surveillance, and the field reports high frame-level ROC-AUC on standard video anom

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

Benchmarking Geospatial Foundation Models for Agriculture Applications

DGX agent

arXiv:2606.29664v1 Announce Type: new Abstract: Geospatial foundation models pretrained on satellite imagery promise broad generalization across remote sensing tasks and regions, but their geographic

model-releasesarxiv-cs-cv
30 Jun 2026
Safety

Beyond Backscatter: AlphaEarth Land-Cover Priors for Rapid SAR Flood Segmentation Across Foundation Backbones

DGX agent

arXiv:2606.29134v1 Announce Type: new Abstract: Rapid flood mapping is critical for emergency response, yet optical imagery is often unusable during major flooding and single-temporal SAR is ambiguous

safetyarxiv-cs-cv
30 Jun 2026
Model Releases

Beyond Trajectory Matching: Reflow with Marginal Distribution Alignment

DGX agent

arXiv:2606.29287v1 Announce Type: cross Abstract: Diffusion and continuous-flow generative models achieve high-quality generation, and their deterministic sampling can be formulated as solving learned

model-releasesarxiv-cs-cv
30 Jun 2026
Research

Bit-ViP: Leveraging Bit-planes to Preserve Visual Privacy in Images through Obfuscation

DGX agent

arXiv:2606.29417v1 Announce Type: new Abstract: The unprecedented growth of computer vision applications, such as surveillance systems and social media, raises security and visual privacy concerns, es

researcharxiv-cs-cv
30 Jun 2026
Research

BLUE: A Stale-Pixel Optical-Flow Compositor for Entropy-Efficient Surveillance Video Encoding

DGX agent

arXiv:2606.28753v1 Announce Type: cross Abstract: Continuous-recording surveillance systems face a storage problem that codec tuning alone cannot fully solve: even at aggressive CRF settings, a static

researcharxiv-cs-cv
30 Jun 2026
Safety

BrainJanus: A Unified Model for Understanding and Generation across Brain, Vision, and Language

DGX agent

arXiv:2606.30319v1 Announce Type: new Abstract: Modeling the bidirectional correspondence between external sensory stimuli and internal neural activity has emerged as a critical frontier in neuroscien

safetyarxiv-cs-cv
30 Jun 2026
Model Releases

BrepLLM: Enabling Large Language Models to Understand Boundary Representations

DGX agent

arXiv:2512.16413v2 Announce Type: replace Abstract: Current token-sequence-based Large Language Models (LLMs) struggle to directly process 3D Boundary Representation (B-rep) models that contain comple

model-releasesarxiv-cs-cv
30 Jun 2026
Research

Bricker to BRACE: A Bracket Exposure RAW Dataset and Restoration Model for Flicker-Banding

DGX agent

arXiv:2606.29845v1 Announce Type: new Abstract: Flicker-banding (FB), arises from temporal aliasing between a camera's rolling shutter and a display's brightness modulation, degrading screen-captured

researcharxiv-cs-cv
30 Jun 2026
Model Releases

Bridging the Gap Between Image Restoration and Navigational Safety in Hazy Conditions: A New Visibility Estimation Metric for Maritime Surveillance

DGX agent

arXiv:2606.30049v1 Announce Type: new Abstract: Visibility distance is critical to maritime navigational safety because it determines the effective observation range of shipborne and shore-based monit

model-releasesarxiv-cs-cv
30 Jun 2026
Tutorials

Building artificial intelligence virtual tissue (AIVT) for tissue state representation, feature prediction, and dynamic simulation

DGX agent

arXiv:2606.29883v1 Announce Type: new Abstract: Modeling tissue states and their transitions is essential for understanding tissue homeostasis in health and pathological remodeling in disease. However

tutorialsarxiv-cs-cv
30 Jun 2026
Model Releases

Can AI Draw Science? A Benchmark for Evaluating Scientific Figure Generation by Text-to-Image and Multimodal Models

DGX agent

arXiv:2606.28406v1 Announce Type: cross Abstract: Text-to-image and multimodal generative models are increasingly used to produce scientific figures such as mechanism diagrams, experimental-design sch

model-releasesarxiv-cs-cv
30 Jun 2026
Research

CellDETR: A Detection-Guided Framework for Scalable Cell Representation Learning from Histopathology Images

DGX agent

arXiv:2606.29463v1 Announce Type: new Abstract: Recent advances in pathology foundation models have substantially improved patch and slide level representation learning from whole-slide images (WSIs).

researcharxiv-cs-cv
30 Jun 2026
Applications

Character Recognition of Nepali Number Plate

DGX agent

arXiv:2606.28946v1 Announce Type: new Abstract: This paper presents a robust Automatic Number Plate Recognition (ANPR) system tailored for Nepali license plates written in Devanagari script. In this p

applicationsarxiv-cs-cv
30 Jun 2026
Hardware

CLEAR-MoE: Shared-Basis Expert Extraction from Frozen Vision Transformers via Calibration-Driven Layer Selection

DGX agent

arXiv:2606.28516v1 Announce Type: new Abstract: We present CLEAR-MoE, a four-phase post-training pipeline that converts a frozen pretrained Vision Transformer (ViT) into a sparse Mixture-of-Experts (M

hardwarearxiv-cs-cv
30 Jun 2026
Safety

Clearer Sight, Fewer Lies: Oriented Pickup Preference Optimization for Multimodal Hallucination Mitigation

DGX agent

arXiv:2606.29805v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) are prone to hallucination as their generation preferences are insufficiently calibrated to visual evidence, ca

safetyarxiv-cs-cv
30 Jun 2026
Research

CLIMP: Contrastive Language-Image Mamba Pretraining

DGX agent

arXiv:2601.06891v2 Announce Type: replace Abstract: Contrastive Language-Image Pre-training (CLIP) relies on Vision Transformers whose attention mechanism is susceptible to spurious correlations, and

researcharxiv-cs-cv
30 Jun 2026
Research

Clinical Risk-Aware Multi-Level Grading for Coronary Artery Stenosis through Curved Feature Reconstruction

DGX agent

arXiv:2606.30082v1 Announce Type: new Abstract: Developing a multi-level grading model for coronary artery stenosis holds great clinical significance for the diagnosis of coronary artery disease. Howe

researcharxiv-cs-cv
30 Jun 2026
Local Ai

ClusterStyle: Modeling Intra-Style Diversity with Prototypical Clustering for Stylized Motion Generation

DGX agent

arXiv:2512.02453v2 Announce Type: replace Abstract: Existing stylized motion generation models have shown their remarkable ability to understand specific style information from the style motion, and i

local-aiarxiv-cs-cv
30 Jun 2026
Model Releases

CoGS: Compositional Dynamic Human-Object Scenes Gaussian Splatting from Monocular Video

DGX agent

arXiv:2606.28820v1 Announce Type: new Abstract: Reconstructing dynamic human--object interaction scenes from monocular video is difficult because the human, manipulated object, and background obey dif

model-releasesarxiv-cs-cv
30 Jun 2026
Applications

CogSENet: Blind Image Deblurring with Blur-Conditioned Semantic Routing and Explicit Frequency Fusion

DGX agent

arXiv:2606.30030v1 Announce Type: new Abstract: Blind image deblurring demands the recovery of high-fidelity details and coherent structures from complex, unknown degradations. Current blind image deb

applicationsarxiv-cs-cv
30 Jun 2026
Local Ai

CollabOD: Collaborative Multi-Backbone with Cross-scale Vision for UAV Small Object Detection

DGX agent

arXiv:2603.05905v2 Announce Type: replace Abstract: Small object detection in unmanned aerial vehicle (UAV) imagery is challenging because high-altitude viewpoints produce severe scale variation, weak

local-aiarxiv-cs-cv
30 Jun 2026
Research

CoLR-Det: Collaborative Latent Restoration for Small Object Detection in Low-Resolution Remote Sensing Images

DGX agent

arXiv:2601.12507v2 Announce Type: replace Abstract: Low-resolution remote sensing small object detection is limited by both missing visual details and the ambiguity of how details serve detection. Exi

researcharxiv-cs-cv
30 Jun 2026
Research

Complete virtual unwrapping and reading of a rolled Herculaneum papyrus

DGX agent

arXiv:2606.29085v1 Announce Type: cross Abstract: The carbonized papyri from Herculaneum preserve the only large-scale library to survive from classical antiquity, but many unopened rolls remain unrea

researcharxiv-cs-cv
30 Jun 2026
Safety

Concept Removal Guidance: Evidence-Calibrated Negative Guidance for Safe Diffusion Sampling

DGX agent

arXiv:2606.29801v1 Announce Type: new Abstract: Text-to-image diffusion models remain vulnerable to adversarial prompts that elicit disallowed content, motivating reliable inference-time controls. A p

safetyarxiv-cs-cv
30 Jun 2026
Research

Consensus Clustering of Free-Viewing Gaze Data: New Insights into Human-Information Interaction

DGX agent

arXiv:2606.30035v1 Announce Type: new Abstract: Free-viewing gaze data provides a rich, task-free window into human visual attention. Conventional exploratory data analysis of the data provides user a

researcharxiv-cs-cv
30 Jun 2026
Safety

Consistency as Inductive Bias: Learning Cross-View Invariance for Robust Multimodal Reasoning

DGX agent

arXiv:2606.29812v1 Announce Type: new Abstract: Inductive biases steer learning toward generalizable solutions by encoding task structure. In this work, we identify a crucial missing bias in MLLMs: cr

safetyarxiv-cs-cv
30 Jun 2026
Tutorials

Contrastive vision-language learning with paraphrasing and negation

DGX agent

arXiv:2511.16527v2 Announce Type: replace Abstract: Contrastive vision-language models continue to be the dominant approach for image-text retrieval. Contrastive Language-Image Pre-training (CLIP) tra

tutorialsarxiv-cs-cv
30 Jun 2026
Model Releases

CoSPlan: Corrective Sequential Planning via Scene Graph Incremental Updates

DGX agent

arXiv:2512.10342v3 Announce Type: replace Abstract: Vision Language Models (VLMs) have shown promising planning capabilities, yet their success remains confined to the text domain, leaving visual deci

model-releasesarxiv-cs-cv
30 Jun 2026
← Previous
1…7778798081…263
Next →