AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,618 results
Research

Lightweight Polyp Segmentation via a Gain-Aware Prediction-Space Recursive Controller

DGX agent

arXiv:2607.03062v1 Announce Type: new Abstract: While lightweight polyp segmentation is highly desirable for low-cost deployment, reported performance gains often stem from upgraded backbone encoders,

researcharxiv-cs-cv
7 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

LILAC: Layer-Wise Independent LoRAs and Cascaded Conditioning for Multi-Concept Customization of Diffusion Models

DGX agent

arXiv:2607.04801v1 Announce Type: new Abstract: Personalizing text-to-image diffusion models to render several specific subjects in a coherent image remains challenging: the model must preserve each s

model-releasesarxiv-cs-cv
7 Jul 2026
Research

Lipschitz-Based Robustness Certification Under Floating-Point Execution

DGX agent

arXiv:2603.13334v4 Announce Type: replace-cross Abstract: Lipschitz-based robustness certification bounds a network's sensitivity through concrete numerical computation rather than symbolic reasoning,

researcharxiv-cs-cv
7 Jul 2026
Safety

LivingWorld: Interactive 4D World Generation with Environmental Dynamics

DGX agent

arXiv:2604.01641v2 Announce Type: replace Abstract: We introduce LivingWorld, an interactive framework for generating 4D worlds with environmental dynamics from a single image. While recent advances i

safetyarxiv-cs-cv
7 Jul 2026
Local Ai

LoMa: Local Feature Matching Revisited

DGX agent

arXiv:2604.04931v2 Announce Type: replace Abstract: Local feature matching has long been a fundamental component of 3D vision systems such as Structure-from-Motion (SfM), yet progress has lagged behin

local-aiarxiv-cs-cv
7 Jul 2026
Applications

MACRO: Training-free Multi-plane Attention for Closeup Render Optimization

DGX agent

arXiv:2607.03875v1 Announce Type: new Abstract: Close-up rendering, zooming into a scene well beyond any training camera, is important for virtual production and interactive 3D content, yet remains an

applicationsarxiv-cs-cv
7 Jul 2026
Research

MACS: Measurement-Aware Consistency Sampling for Inverse Problems

DGX agent

arXiv:2510.02208v3 Announce Type: replace-cross Abstract: Diffusion models have emerged as powerful generative priors for solving inverse imaging problems. However, their practical deployment is hinde

researcharxiv-cs-cv
7 Jul 2026
Safety

MAGE: View-guided Point Cloud Completion with Efficient Modality Alignment and Adaptive Geometry Enhancement

DGX agent

arXiv:2607.02568v1 Announce Type: new Abstract: View-based point cloud completion aims to recover a complete 3D shape from a partial point cloud, guided by a single-view image. However, existing appro

safetyarxiv-cs-cv
7 Jul 2026
Research

MambaRefine-CD: MambaVision with Region-Boundary Temporal Refinement

DGX agent

arXiv:2607.04403v1 Announce Type: cross Abstract: Binary change detection in remote sensing requires both complete changed-region localization and accurate boundary delineation. We present MambaRefine

researcharxiv-cs-cv
7 Jul 2026
Research

Measuring 3D Spatial Geometric Consistency in Dynamic Video Generation

DGX agent

arXiv:2603.19048v2 Announce Type: replace Abstract: Recent generative models can produce high-fidelity videos, yet they often exhibit 3D spatial geometric inconsistencies. Existing evaluation methods

researcharxiv-cs-cv
7 Jul 2026
Local Ai

MedMambaLite: Hardware-Aware Mamba for Medical Image Classification

DGX agent

arXiv:2508.05049v1 Announce Type: cross Abstract: AI-powered medical devices have driven the need for real-time, on-device inference such as biomedical image classification. Deployment of deep learnin

local-aiarxiv-cs-cv
7 Jul 2026
Research

MergeSurv: Merging-Based Continual Learning for Survival Analysis on Whole-Slide Images

DGX agent

arXiv:2607.04747v1 Announce Type: new Abstract: Survival analysis on Whole Slide Images (WSIs) is important in computational pathology for prognosis estimation and treatment planning. However, existin

researcharxiv-cs-cv
7 Jul 2026
Research

MetaMax: Improved Open-Set Deep Neural Networks via Weibull Calibration

DGX agent

arXiv:2211.10872v2 Announce Type: replace Abstract: Open-set recognition refers to the problem in which classes that were not seen during training appear at inference time. This requires the ability t

researcharxiv-cs-cv
7 Jul 2026
Research

MGCA-Net: Multi-Grained Category-Aware Network for Open-Vocabulary Temporal Action Localization

DGX agent

arXiv:2511.13039v2 Announce Type: replace Abstract: Open-Vocabulary Temporal Action Localization (OV-TAL) aims to recognize and localize instances of any desired action categories in videos without ex

researcharxiv-cs-cv
7 Jul 2026
Safety

Mitigating Covariate Shift in Imitation Learning for Autonomous Vehicles Using Latent Space Generative World Models

DGX agent

arXiv:2409.16663v5 Announce Type: replace-cross Abstract: We propose the use of latent space generative world models to address the covariate shift problem in autonomous driving. A world model is a ne

safetyarxiv-cs-cv
7 Jul 2026
Safety

Mixture-of-Gaussians-Guided Schedule Design for Brownian Bridge Diffusion Models

DGX agent

arXiv:2607.03517v1 Announce Type: cross Abstract: Brownian Bridge Diffusion Models (BBDM) offer an appealing framework for image restoration and inverse problems by constructing a stochastic bridge fr

safetyarxiv-cs-cv
7 Jul 2026
Model Releases

MMEarth-Bench: Global Model Adaptation via Multimodal Test-Time Training

DGX agent

arXiv:2602.06285v2 Announce Type: replace Abstract: Recent research in geospatial machine learning has demonstrated that models pretrained with self-supervised learning on Earth observation data can p

model-releasesarxiv-cs-cv
7 Jul 2026
Research

Model Confidence-Guided Multi-Image Fusion of Fundus Images for Diabetic Retinopathy Diagnosis

DGX agent

arXiv:2607.03643v1 Announce Type: cross Abstract: Purpose: Early screening for eye diseases is critical in low- and middle-income countries where access to care is limited. We investigate whether a co

researcharxiv-cs-cv
7 Jul 2026
Research

Motion Estimation Techniques for Volumetric Video Attribute Compression

DGX agent

arXiv:2607.03576v1 Announce Type: cross Abstract: Point cloud compression relies on techniques to compress both geometry and attributes. Motion-based approaches for dynamic solid point cloud geometry

researcharxiv-cs-cv
7 Jul 2026
Model Releases

MUSON: A Reasoning-oriented Multimodal Dataset for Socially Compliant Navigation in Urban Environments

DGX agent

arXiv:2512.22867v2 Announce Type: replace Abstract: Socially compliant navigation requires structured reasoning about dynamic pedestrians and physical constraints to ensure safe and interpretable deci

model-releasesarxiv-cs-cv
7 Jul 2026
Safety

MV-Forcing: Long Multi-View Video Generation via 4D-Grounded Spatio-Temporal Self-Forcing

DGX agent

arXiv:2607.05376v1 Announce Type: new Abstract: Recent advances in video diffusion models have enabled either long single-view generation through temporal autoregression, or short multi-view synthesis

safetyarxiv-cs-cv
7 Jul 2026
Research

NABLA: Neighborhood Adaptive Block-Level Attention

DGX agent

arXiv:2507.13546v2 Announce Type: replace Abstract: Recent progress in transformer-based architectures has demonstrated remarkable success in video generation tasks. However, the quadratic complexity

researcharxiv-cs-cv
7 Jul 2026
Model Releases

Natural Language Camera Movement Understanding

DGX agent

arXiv:2607.03043v1 Announce Type: new Abstract: Understanding camera movement in natural language is critical for training and evaluating video generation models, among other applications. However, we

model-releasesarxiv-cs-cv
7 Jul 2026
Safety

NavEYE: Vision-Centered Multi-Sensor Fusion-Based Situational Awareness System for Intelligent Surface Vehicles

DGX agent

arXiv:2607.03915v1 Announce Type: new Abstract: With the rapid development of sensor and artificial intelligence (AI) technologies, intelligent surface vehicles (ISVs) have gained increasing attention

safetyarxiv-cs-cv
7 Jul 2026
Research

Observable- and Positional-Encoding-Dependent Symmetry Readout from Neural Network Weights

DGX agent

arXiv:2607.03108v1 Announce Type: cross Abstract: Post-hoc analysis of trained neural network weights often seeks to recover geometric structure directly from the parameters. We show that, for positio

researcharxiv-cs-cv
7 Jul 2026
Research

Omni-Diffusion: Unified Multimodal Understanding and Generation with Masked Discrete Diffusion

DGX agent

arXiv:2603.06577v2 Announce Type: replace Abstract: While recent multimodal large language models (MLLMs) have made impressive strides, they predominantly employ a conventional autoregressive architec

researcharxiv-cs-cv
7 Jul 2026
Safety

OmniDS: Dual-Stream Context Fusion for Omnidirectional Depth from Fisheye Cameras

DGX agent

arXiv:2607.03038v1 Announce Type: new Abstract: Omnidirectional depth estimation from multi-fisheye camera rigs is complicated by visibility conflicts: wide baselines cause different cameras to observ

safetyarxiv-cs-cv
7 Jul 2026
Model Releases

OmniLayout: A Schematic-Coupled Multimodal Benchmark for Constraint-Aware Geometric Reasoning in PCB Layout

DGX agent

arXiv:2607.03261v1 Announce Type: new Abstract: Recent large language models (LLMs) have demonstrated remarkable progress in 3D spatial reasoning, spatial grounding, and fine-grained geometric underst

model-releasesarxiv-cs-cv
7 Jul 2026
Safety

Open-Attribute Person Retrieval: Finding People Through Distinctive and Novel Attributes

DGX agent

arXiv:2508.01389v3 Announce Type: replace Abstract: Person retrieval in surveillance videos often depends on attributes described by witnesses or operators. However, the most useful cues in practice a

safetyarxiv-cs-cv
7 Jul 2026
Safety

Overloading Large Vision-Language Models for Jailbreaking

DGX agent

arXiv:2607.02961v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) exhibit remarkable vision-language capabilities and are increasingly deployed in real-world applications such as pe

safetyarxiv-cs-cv
7 Jul 2026
Research

PAGE: Towards Practical Human-level Gaze Target Estimation

DGX agent

arXiv:2607.04860v1 Announce Type: new Abstract: Gaze target estimation, the task of predicting where a person is looking in a scene, is crucial to understanding human attention and intent. It is a cha

researcharxiv-cs-cv
7 Jul 2026
Research

Paired Uterine Whole-Slide Images and Pathology Reports for Multimodal Computational Pathology

DGX agent

arXiv:2607.04020v1 Announce Type: new Abstract: Uterine diseases represent an important category of gynecologic pathology and require accurate histopathological assessment for diagnosis and treatment

researcharxiv-cs-cv
7 Jul 2026
Model Releases

Perceiving Better Moments: Cover Frame Reselection and Enhancement for Live Photos with the Live2K Dataset

DGX agent

arXiv:2607.04151v1 Announce Type: new Abstract: Modern smartphones capture Live Photos, short video bursts surrounding a still image, offering a dynamic and engaging photographic experience. However,

model-releasesarxiv-cs-cv
7 Jul 2026
Research

Perceptual Flow Matching for Few-Step Generative Modeling

DGX agent

arXiv:2607.03524v1 Announce Type: new Abstract: We propose Perceptual Flow Matching (PFM), a simple yet effective framework for few-step generation in flow-matching models. Rather than performing velo

researcharxiv-cs-cv
7 Jul 2026
Model Releases

Phi-SegNet: Phase-Integrated Supervision for Medical Image Segmentation

DGX agent

arXiv:2601.16064v2 Announce Type: replace-cross Abstract: Deep learning has substantially advanced medical image segmentation, yet achieving robust generalization across diverse imaging modalities and

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

PhysMirror: Physics-Aware Mirror Object Generation

DGX agent

arXiv:2607.03470v1 Announce Type: new Abstract: Synthesizing physically accurate mirror reflections remains a fundamental challenge for modern text-to-image diffusion models, which are increasingly cr

model-releasesarxiv-cs-cv
7 Jul 2026
Research

Piecewise Dynamic Diffusion Regularization for Reconstruction of Cardiac Cine MRI

DGX agent

arXiv:2607.03299v1 Announce Type: cross Abstract: Real-time cardiac cine MRI enables visualization of the beating heart during free breathing, but severe undersampling and motion make reconstruction h

researcharxiv-cs-cv
7 Jul 2026
Safety

PixelPilot: Scalable Vision-Language-Action Models for End-to-End Autonomous Driving

DGX agent

arXiv:2607.04637v1 Announce Type: new Abstract: Vision-Language-Action Models (VLAs), which leverage the advanced reasoning capabilities of Vision-Language Models (VLMs), show promising generalization

safetyarxiv-cs-cv
7 Jul 2026
Research

PixWorld: Unifying 3D Scene Generation and Reconstruction in Pixel Space

DGX agent

arXiv:2607.05373v1 Announce Type: new Abstract: 3D reconstruction and generation are commonly tackled by separate paradigms: pixel-based regression for reconstruction, and latent diffusion for generat

researcharxiv-cs-cv
7 Jul 2026
Applications

Present but Not Remembered: Auditing How Frozen VLAs Encode, Deploy, and Steer Visual History

DGX agent

arXiv:2607.03372v1 Announce Type: new Abstract: A frozen vision-language-action model (VLA) receives recent observations at every decision step, yet prior work has focused on adding memory rather than

applicationsarxiv-cs-cv
7 Jul 2026
Research

PreSIST: Vision-Language-Informed Object Persistence Prediction in Open-World Scenes

DGX agent

arXiv:2607.04057v1 Announce Type: new Abstract: Robots deployed over long periods must reason about environments that change over time. Existing long-term perception systems often address object chang

researcharxiv-cs-cv
7 Jul 2026
Research

Pretreatment MRI reveals a latent, molecular-subtype-independent structural phenotype that organizes treatment trajectories and recurrence risk

DGX agent

arXiv:2607.02768v1 Announce Type: cross Abstract: Pathologic complete response and tumor shrinkage measure whether breast cancer responds to neoadjuvant therapy, but not whether that response was stru

researcharxiv-cs-cv
7 Jul 2026
Safety

PRIMA: Pre-training with Risk-integrated Image-Metadata Alignment for Medical Diagnosis via LLM

DGX agent

arXiv:2602.23297v2 Announce Type: replace Abstract: Medical diagnosis requires the effective synthesis of visual manifestations and clinical metadata. However, existing methods often treat metadata as

safetyarxiv-cs-cv
7 Jul 2026
Model Releases

Prior Bias in Vision Language Models on UML Diagram Interpretation

DGX agent

arXiv:2607.02853v1 Announce Type: new Abstract: Vision Language Models (VLMs) are increasingly applied to software engineering artifacts, especially UML class diagrams whose meaning depends on visual

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

PRISM3D: Probabilistic Refinement and Robust Initialization for Physically Consistent Scene Modeling under Extreme Motion Blur

DGX agent

arXiv:2607.03855v1 Announce Type: new Abstract: We address the inverse problem of blind 3D scene reconstruction from extremely motion-blurred images, a scenario where traditional Structure-from-Motion

model-releasesarxiv-cs-cv
7 Jul 2026
Research

Privacy-Preserving Industrial Ergonomics: mmWave-Based Automated REBA Scoring and Pose Estimation

DGX agent

arXiv:2607.02611v1 Announce Type: new Abstract: Work-related Musculoskeletal Disorders (WMSDs) require continuous ergonomic assessments. While Rapid Entire Body Assessment (REBA) is a gold-standard ob

researcharxiv-cs-cv
7 Jul 2026
Safety

Probabilistic Robustness in Medical Image Classification

DGX agent

arXiv:2607.03797v1 Announce Type: new Abstract: Deep learning (DL) has shown strong performance in medical image classification, but its trustworthy deployment remains challenging in safety-critical c

safetyarxiv-cs-cv
7 Jul 2026
Research

Probe-EM: Targeted Neuron Tracing via Training-Free Semantic Verification

DGX agent

arXiv:2607.04696v1 Announce Type: new Abstract: Establishing large-scale, high-resolution neural connectivity maps is fundamental to elucidating the structural basis of brain function. However, when p

researcharxiv-cs-cv
7 Jul 2026
← Previous
1…6667686970…263
Next →