AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,515 results
Research

Medical Image Segmentation based on Deep Active Contour and Mean Curvature Loss Function

DGX agent

arXiv:2607.12586v1 Announce Type: cross Abstract: Medical image segmentation is a crucial task in the field of clinical analysis and applications. Though deep learning techniques recently play a cruci

researcharxiv-cs-cv
15 Jul 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Research

MetaView: Monocular Novel View Synthesis with Scale-Aware Implicit Geometry Priors

DGX agent

arXiv:2607.12000v1 Announce Type: new Abstract: Current visual generation models are capable of producing high-quality content, yet they lack a coherent perception of the spatial structure. Existing g

researcharxiv-cs-cv
15 Jul 2026
Model Releases

Metric-Guided Synthetic Image Data Rendering for Deep Learning compatible with Agentic AI

DGX agent

arXiv:2607.12874v1 Announce Type: new Abstract: Deep learning computer vision for scientific applications requires collecting and annotating large datasets in a laborious, expensive and error-prone pr

model-releasesarxiv-cs-cv
15 Jul 2026
Research

MobileSAM2: Lightweight Segment Anything for Spatial Intelligence

DGX agent

arXiv:2607.12297v1 Announce Type: new Abstract: The recent large video foundation model, SAM2, enables segment anything in both images and videos, serving as a powerful base model for various applicat

researcharxiv-cs-cv
15 Jul 2026
Local Ai

More Than Where You Are: Learning Semantics, Structure, and Geometry from Cross-View Localization

DGX agent

arXiv:2607.12429v1 Announce Type: new Abstract: Consistent cross-view understanding under extreme viewpoint changes is essential for spatial intelligence, as it enables models to recognize the same sc

local-aiarxiv-cs-cv
15 Jul 2026
Model Releases

MQAdapter: Multi-Modal Quantum Adapter for Coarse-to-Fine VLM Fine-tuning

DGX agent

arXiv:2607.12418v1 Announce Type: new Abstract: Large-scale Vision-Language Models have demonstrated impressive transfer learning capabilities across a wide range of tasks. For few-shot classification

model-releasesarxiv-cs-cv
15 Jul 2026
Model Releases

NEEDL-Bench: Dataset for Swiss Needle Cast and Stomata Detection in Microscopy Images

DGX agent

arXiv:2607.12076v1 Announce Type: new Abstract: We present NEEDL-Bench, a microscopy detection benchmark for Swiss Needle Cast (SNC), a fungal disease of Douglas-fir trees. Douglas-fir is a keystone s

model-releasesarxiv-cs-cv
15 Jul 2026
Local Ai

Open-KNEAD: Knowledge-grounded Nutrition Estimation via Agentic Decomposition

DGX agent

arXiv:2607.12911v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) are increasingly used for dietary assessment from meal images, where retrieval-augmented grounding was shown to

local-aiarxiv-cs-cv
15 Jul 2026
Research

Overview of Cross-Component In-loop Filters in Video Coding Standards

DGX agent

arXiv:2607.12186v1 Announce Type: new Abstract: In-loop filters have been comprehensively explored during the development of video coding standards due to their remarkable noise-reduction capability.

researcharxiv-cs-cv
15 Jul 2026
Research

Physically Aware Radiomics Without Interpolation: Disentangling Voxel Geometry and Signal Modification in CT and MRI

DGX agent

arXiv:2607.12399v1 Announce Type: new Abstract: Objective: Radiomic texture features are usually computed in voxel-index neighborhoods, implicitly assuming isotropic spatial relationships. In anisotro

researcharxiv-cs-cv
15 Jul 2026
Local Ai

Pixel-Level Pavement Distress Assessment Using Instance Segmentation

DGX agent

arXiv:2605.26095v2 Announce Type: replace Abstract: Automated pavement distress assessment requires more than image-level classification or coarse bounding box detection, demanding precise localizatio

local-aiarxiv-cs-cv
15 Jul 2026
Agents

Point Tracking in Surgery--The 2025 Surgical Tattoos in Infrared Challenge (STIRC2025)

DGX agent

arXiv:2607.12939v1 Announce Type: new Abstract: Point tracking in surgery is crucial to enable applications in downstream tasks such as segmentation, 3D reconstruction, virtual tissue landmarking, aut

agentsarxiv-cs-cv
15 Jul 2026
Safety

PoseAlign: Sculpting Pose-Consistent Meshes via Text-Guided Deformation

DGX agent

arXiv:2607.10560v2 Announce Type: replace-cross Abstract: Mesh deformation, the process of altering the vertex positions of a 3D mesh while preserving its topological structure, is a cornerstone of co

safetyarxiv-cs-cv
15 Jul 2026
Research

ProtoPointNet: Prototype-Based Interpretable Classification of 3D Dental Point Clouds with Verifiable Spatial Activations

DGX agent

arXiv:2607.12335v1 Announce Type: new Abstract: Prototype-based networks provide inherently interpretable classification by linking predictions to learned exemplars, but their use in 3D point clouds a

researcharxiv-cs-cv
15 Jul 2026
Model Releases

Rank-1 Identity Consensus Predicts Gallery Enrollment in 1:N Face Matching More Accurately than Score Thresholding

DGX agent

arXiv:2607.12903v1 Announce Type: new Abstract: In operational 1:N face identification, a crucial question arises for each probe: is this person enrolled in the gallery or not? The stakes are high and

model-releasesarxiv-cs-cv
15 Jul 2026
Model Releases

RealSkin: Spatio-Spectral Partial Neural Adjoint Maps for Image-to-3D Attribute Transfer

DGX agent

arXiv:2607.12495v1 Announce Type: new Abstract: Creating photorealistic 3D assets requires bridging the appearance gap between real-world observations and synthetic models. A promising approach is to

model-releasesarxiv-cs-cv
15 Jul 2026
Agents

ReflectVLN: Training Vision-Language Navigation Agents with Reflective Reasoning

DGX agent

arXiv:2607.12680v1 Announce Type: new Abstract: Existing vision-language navigation methods often couple a VLM with waypoint decoders to produce multi-step action plans, but they typically lack an exp

agentsarxiv-cs-cv
15 Jul 2026
Local Ai

RegHead: Non-Humanoid Head Blendshapes via Feed-Forward Registration

DGX agent

arXiv:2607.12206v1 Announce Type: new Abstract: We present RegHead, a framework for constructing semantic blendshape sets for animatable non-humanoid head avatars. With a fixed expression vocabulary,

local-aiarxiv-cs-cv
15 Jul 2026
Research

RFMSR: Residual Flow Matching for Image Super-Resolution

DGX agent

arXiv:2607.12753v1 Announce Type: new Abstract: Image super-resolution (ISR) has witnessed remarkable progress with diffusion models and flow matching. The dominant text-to-image (T2I) based approache

researcharxiv-cs-cv
15 Jul 2026
Research

Rough Path Signature-Guided Geometry Augmentation for Few-Shot Industrial Surface Defect Detection

DGX agent

arXiv:2607.12245v1 Announce Type: new Abstract: Few-shot industrial defect detection remains difficult for standard supervised detectors, which achieve poor performance on boundary-dominated industria

researcharxiv-cs-cv
15 Jul 2026
Research

Same Compression Principle, Different Geometry: Rate-Distortion Signatures Dissociate Biological and Artificial Visual Systems

DGX agent

arXiv:2603.01568v2 Announce Type: replace-cross Abstract: Efficient coding theory predicts that biological perceptual systems compress sensory input optimally under resource constraints, with the syst

researcharxiv-cs-cv
15 Jul 2026
Safety

Sat2RealCity: Geometry-Aware and Appearance-Controllable 3D Urban Generation from Satellite Imagery

DGX agent

arXiv:2511.11470v3 Announce Type: replace Abstract: 3D urban generation from satellite imagery is an important task for scalable digital twins and real-world simulation environments. Existing approach

safetyarxiv-cs-cv
15 Jul 2026
Local Ai

SeamGen: Artist-Aligned UV Seam Generation via Graph Flow Matching

DGX agent

arXiv:2607.12379v1 Announce Type: new Abstract: UV seam placement is a critical yet labor-intensive step in 3D content creation, requiring artists to balance chart shape, seam concealment, and alignme

local-aiarxiv-cs-cv
15 Jul 2026
Local Ai

Seeing Globally, Refining Locally: Global Visual Guidance and Local Ultrasound Cues for Robust Freehand 3-D Ultrasound Reconstruction

DGX agent

arXiv:2607.12398v1 Announce Type: new Abstract: Freehand 3-D ultrasound (US) imaging has attracted increasing attention owing to its intuitive volumetric visualization, ease of use, and low cost. Howe

local-aiarxiv-cs-cv
15 Jul 2026
Model Releases

Self in Space: Benchmarking Self-Awareness and Spatial Cognition in UAV Embodied Intelligence

DGX agent

arXiv:2607.12477v1 Announce Type: new Abstract: Autonomous UAV systems increasingly rely on multimodal large language models (MLLMs) to operate in complex real-world environments. Such embodied scenar

model-releasesarxiv-cs-cv
15 Jul 2026
Research

Semantic-Edge Response Decoding of SAM3 for Zero-Shot Crack Segmentation

DGX agent

arXiv:2607.12292v1 Announce Type: new Abstract: Crack segmentation is essential for infrastructure inspection and structural health assessment, but existing high-performance methods typically require

researcharxiv-cs-cv
15 Jul 2026
Model Releases

Spherical-GOF: Geometry-Aware Panoramic Gaussian Opacity Fields for 3D Scene Reconstruction

DGX agent

arXiv:2603.08503v2 Announce Type: replace Abstract: Omnidirectional images are increasingly used in robotics and vision due to their wide field of view. However, extending 3D Gaussian Splatting (3DGS)

model-releasesarxiv-cs-cv
15 Jul 2026
Research

SpikeDS: Dual Sparsity Spikformer for Perineural Invasion Prediction in 3D MRI

DGX agent

arXiv:2607.11986v1 Announce Type: new Abstract: Perineural invasion (PNI) is associated with poor prognosis in cholangiocarcinoma (CCA). However, its detection from 3D MRI remains challenging due to t

researcharxiv-cs-cv
15 Jul 2026
Research

Statistical Non-linear Reconstruction Loss for Image Anomaly Detection

DGX agent

arXiv:2607.12866v1 Announce Type: new Abstract: Reconstruction-based methods are a cornerstone of unsupervised image anomaly detection, but they remain vulnerable to outlier leakage, where standard me

researcharxiv-cs-cv
15 Jul 2026
Research

Steering Diffusion Models via Class-Contrastive Influence for Few-Shot Medical Classification

DGX agent

arXiv:2607.12464v1 Announce Type: new Abstract: When labeled data are scarce, off-the-shelf diffusion models can augment training sets for few-shot medical image classification, but not all generated

researcharxiv-cs-cv
15 Jul 2026
Safety

Structure-Semantic Co-optimized Latent Diffusion Model for Fast Visual Anagram Synthesis

DGX agent

arXiv:2606.16241v3 Announce Type: replace Abstract: Visual anagram is an intriguing form of art creation wherein a single image presents different conceptual interpretations under transformations such

safetyarxiv-cs-cv
15 Jul 2026
Agents

SymbOmni: Evolving Agentic Omni Models via Symbolic Concept Learning

DGX agent

arXiv:2607.12042v1 Announce Type: new Abstract: Visual generation is increasingly ubiquitous in diverse domains, from text-to-image/video synthesis to multimodal interactive creation. Yet prevailing m

agentsarxiv-cs-cv
15 Jul 2026
Model Releases

TerraLogic: A Benchmark for Hierarchical Geospatial Reasoning in Earth Observation

DGX agent

arXiv:2607.12497v1 Announce Type: new Abstract: Beyond perception, reasoning is essential in remote sensing for advanced interpretation, inference, and decision-making. Recent advances in large langua

model-releasesarxiv-cs-cv
15 Jul 2026
Agents

The GEST-Engine: From Event Graphs to Synthetic Video. A Full Technical Report

DGX agent

arXiv:2607.12231v1 Announce Type: new Abstract: We present the GEST-Engine, a complete system that goes from natural-language text to fully-annotated multi-actor video. At its core is an explicit worl

agentsarxiv-cs-cv
15 Jul 2026
Research

The Seriality Gap in Video Diffusion Models

DGX agent

arXiv:2607.13031v1 Announce Type: cross Abstract: When one ball strikes another, then another, video models should predict the consequences of each bounce. In controlled experiments on multi-ball hard

researcharxiv-cs-cv
15 Jul 2026
Model Releases

The TopCoW Challenge -- Topology-Aware Circle of Willis Segmentation for CT and MR Angiography

DGX agent

arXiv:2312.17670v5 Announce Type: replace Abstract: The Circle of Willis (CoW) is an important network of arteries connecting major circulations of the brain. Its vascular architecture is believed to

model-releasesarxiv-cs-cv
15 Jul 2026
Safety

Together, Then Apart: Balancing Alignment and Distinctiveness for Multimodal Survival Analysis

DGX agent

arXiv:2511.18089v2 Announce Type: replace Abstract: Multimodal survival analysis aims to improve cancer prognosis using heterogeneous biomedical data, such as histopathology images and genomic profile

safetyarxiv-cs-cv
15 Jul 2026
Research

Towards Vision-Free CIR: Attribute-Augmented Scoring and LLM-Based Reranking for Zero-Shot Composed Image Retrieval

DGX agent

arXiv:2607.12621v1 Announce Type: new Abstract: Recent work has shown that 'Vision-Free'' approaches (representing images as text) can be effective for standard image retrieval tasks. However, it rema

researcharxiv-cs-cv
15 Jul 2026
Agents

Traj-VLN: Learning Pixel-Space Interaction via Autoregressive Trajectory Generation

DGX agent

arXiv:2607.10744v2 Announce Type: replace Abstract: Benefiting from the powerful priors embedded in large-scale pre-training data and the emerging commonsense reasoning ability, large language models

agentsarxiv-cs-cv
15 Jul 2026
Local Ai

TSCA-Net: Temporal-Spatial Clique Attention for Interpretable Multimodal Pedestrian Trajectory Prediction

DGX agent

arXiv:2607.11939v1 Announce Type: new Abstract: Accurate pedestrian trajectory prediction in crowded environments remains challenging due to the multimodal uncertainty of human motion and the variable

local-aiarxiv-cs-cv
15 Jul 2026
Tutorials

UMSS: Towards Unsupervised Multi-modal Semantic Segmentation

DGX agent

arXiv:2607.12372v1 Announce Type: new Abstract: Multimodal semantic segmentation (MSS) is essential for robust perception in complex environments, yet its potential remains largely untapped because of

tutorialsarxiv-cs-cv
15 Jul 2026
Research

Uncertainty-Aware Multi-Source Retinal Fluid Segmentation in OCT

DGX agent

arXiv:2607.12212v1 Announce Type: cross Abstract: Measuring retinal fluid from optical coherence tomography (OCT) drives treatment decisions in macular disease, but manual annotation is slow and segme

researcharxiv-cs-cv
15 Jul 2026
Research

UniMedSeg: Unified In-Context Learning for Multi-Paradigm 2D/3D Medical Image Segmentation

DGX agent

arXiv:2607.12896v1 Announce Type: new Abstract: Medical image segmentation foundation models are expected to generalize across diverse clinical scenarios, yet existing universal methods remain fragmen

researcharxiv-cs-cv
15 Jul 2026
Model Releases

UniVR: Thinking in Visual Space for Unified Visual Reasoning

DGX agent

arXiv:2607.12800v1 Announce Type: new Abstract: Learning broad world knowledge directly from raw visual data is a fundamental capability of intelligence. We introduce UniVR, the first investigation in

model-releasesarxiv-cs-cv
15 Jul 2026
Model Releases

VanillaBench: The Hidden Accuracy Cost of Adversarial Robustness

DGX agent

arXiv:2607.12545v1 Announce Type: cross Abstract: Adversarial robustness research has produced hundreds of defended models over the past decade, yet the literature almost universally reports robustnes

model-releasesarxiv-cs-cv
15 Jul 2026
Agents

ViCo3D: Empowering LiDAR-based Collaborative 3D Object Detection with Vision Foundation Models

DGX agent

arXiv:2607.12959v1 Announce Type: new Abstract: LiDAR-based collaborative 3D perception in Vehicle-to-Everything (V2X) systems typically relies on fusing bird's-eye-view (BEV) features across agents.

agentsarxiv-cs-cv
15 Jul 2026
Research

Virtual Chromoendscopy with Tunable Visibility Enhancement

DGX agent

arXiv:2607.12416v1 Announce Type: new Abstract: Chromoendoscopy (CE) is a common clinical practice that sprays indigo carmine blue dye onto the gastric surface to improve the visibility of gastric les

researcharxiv-cs-cv
15 Jul 2026
Agents

VISA: VLM-Guided Instance Semantic Auditing for 3D Occupancy World Models

DGX agent

arXiv:2606.13460v2 Announce Type: replace Abstract: Semantic 3D occupancy provides a voxelized world state for autonomous driving and robot decision making, but object and rare-class errors can affect

agentsarxiv-cs-cv
15 Jul 2026
← Previous
1…5253545556…261
Next →