AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,515 results
Research

Leveraging Semantic Maps for City-Scale Cross-View Localization

DGX agent

arXiv:2607.25215v1 Announce Type: cross Abstract: We want robots to localize in previously untraversed environments against commonly available prior data. Rich semantic data available from OpenStreetM

researcharxiv-cs-cv
29 Jul 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

LGFNet: A CTC-Guided Local-Global Fusion Framework for Single-Channel Sleep Staging

DGX agent

arXiv:2607.25197v1 Announce Type: new Abstract: Sleep staging remains challenging due to long-range temporal dependencies, ambiguous stage transitions-particularly in N1-and substantial distribution s

safetyarxiv-cs-cv
29 Jul 2026
Safety

Med-SegLens: Latent-Level Model Diffing for Interpretable Medical Image Segmentation

DGX agent

arXiv:2602.10508v2 Announce Type: replace Abstract: Modern segmentation models achieve strong predictive performance but remain largely opaque, limiting our ability to diagnose failures, understand da

safetyarxiv-cs-cv
29 Jul 2026
Applications

MEDIC-AD: Towards Medical Vision-Language Model's Clinical Intelligence

DGX agent

arXiv:2603.27176v2 Announce Type: replace Abstract: Lesion detection, symptom tracking, and visual explainability are central to real-world medical image analysis, yet current medical Vision-Language

applicationsarxiv-cs-cv
29 Jul 2026
Safety

Medical world models in healthcare: foundations, applications, and challenges for trustworthy clinical translation

DGX agent

arXiv:2607.25242v1 Announce Type: new Abstract: Medical world models offer a framework for extending medical artificial intelligence beyond static prediction by representing evolving patient states an

safetyarxiv-cs-cv
29 Jul 2026
Model Releases

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing

DGX agent

arXiv:2607.25300v1 Announce Type: new Abstract: Video editing is fundamentally message-driven: even from the same source footage, the selected shots change depending on the narrative the editor wishes

model-releasesarxiv-cs-cv
29 Jul 2026
Local Ai

Mondrian: On-Device High-Performance Video Analytics with Compressive Packed Inference

DGX agent

arXiv:2403.07598v2 Announce Type: replace Abstract: In this paper, we present Mondrian, an edge system that enables high-performance object detection on high-resolution video streams. Many lightweight

local-aiarxiv-cs-cv
29 Jul 2026
Model Releases

MorphUNet: Alpha-Controlled Biometric Transport for Diffusion-Based Face Morphing Attacks

DGX agent

arXiv:2607.25092v1 Announce Type: new Abstract: Face morphing attacks create synthetic images verifiable against multiple identities, threatening border control and identity verification systems. We i

model-releasesarxiv-cs-cv
29 Jul 2026
Safety

NEXT: Reasoning-Driven Video Recommendation via a Vision-Language Model

DGX agent

arXiv:2607.24789v1 Announce Type: cross Abstract: We present NEXT (Next-interest EXploration Transformer), a reasoning-driven video recommendation framework that reasons over the video a user has just

safetyarxiv-cs-cv
29 Jul 2026
Applications

Noise-Free One-Step LoRA for Task-Driven Image Restoration with Diffusion Priors

DGX agent

arXiv:2607.25390v1 Announce Type: new Abstract: Degraded images not only reduce visual quality but also impair downstream high-level vision tasks. Task-driven image restoration (TDIR) addresses this i

applicationsarxiv-cs-cv
29 Jul 2026
Model Releases

ObliCity: A Benchmark and Baseline for Roof-to-Ground Projection Displacement Correction

DGX agent

arXiv:2607.25210v1 Announce Type: new Abstract: Oblique-view urban remote sensing imagery inevitably exhibits geometric projection displacements between building roofs and footprints, leading to signi

model-releasesarxiv-cs-cv
29 Jul 2026
Safety

On the Use of Synthetic Data for Threshold Calibration in Face Recognition: Performance and Security Implications for Border Control Systems

DGX agent

arXiv:2607.25990v1 Announce Type: new Abstract: The recently deployed Entry/Exit System (EES) introduces large-scale biometric verification into European border control, requiring face recognition sys

safetyarxiv-cs-cv
29 Jul 2026
Research

Open-Ended CT Volume Segmentation with Weak Supervision from Language

DGX agent

arXiv:2607.25860v1 Announce Type: new Abstract: We introduce a method for training a text-conditioned segmentation model for CT scans, which combines voxel-level supervision with coarse but scalable s

researcharxiv-cs-cv
29 Jul 2026
Model Releases

OpenPVMapper: A Multi-source, Nationwide Database of Rooftop Photovoltaic Systems in France

DGX agent

arXiv:2607.25153v1 Announce Type: new Abstract: Rooftop photovoltaic (PV) systems account for the vast majority of PV grid connections, yet no open, comprehensive, installation-level dataset of these

model-releasesarxiv-cs-cv
29 Jul 2026
Model Releases

OrthKD: Extracting Generalized Clinical Knowledge from Heterogeneous Teachers for Lightweight Deployment

DGX agent

arXiv:2607.25545v1 Announce Type: cross Abstract: Deploying diabetic retinopathy (DR) screening models in primary care requires edge-efficient systems that remain accurate, safe, and reliable under do

model-releasesarxiv-cs-cv
29 Jul 2026
Model Releases

PanoLess: Environment Reconstruction from Partial Reflective Views

DGX agent

arXiv:2607.25362v1 Announce Type: new Abstract: Reflections from shiny objects and glass facades naturally extend the field of view of a camera, capturing the surrounding environment without the need

model-releasesarxiv-cs-cv
29 Jul 2026
Model Releases

Parallel Decoding Distillation for Fast Image and Video Generation

DGX agent

arXiv:2607.26004v1 Announce Type: new Abstract: Generation in video diffusion or flow models is computationally expensive due to the slow and iterative sampling process. Current state-of-the-art (SOTA

model-releasesarxiv-cs-cv
29 Jul 2026
Model Releases

PerceptionBench: Evaluating Atomic Visual Perception in Multimodal Large Language Models

DGX agent

arXiv:2607.24957v1 Announce Type: new Abstract: We introduce PerceptionBench, a benchmark specifically designed to evaluate the atomic visual perception capabilities of Multimodal Large Language Model

model-releasesarxiv-cs-cv
29 Jul 2026
Hardware

Quasi-SVD: Learning a Lie-constrained matrix factorisation for real-time imaging

DGX agent

arXiv:2607.25967v1 Announce Type: new Abstract: Singular Value Decomposition (SVD) underlies matrix factorisation tasks across computational imaging, with medical applications increasingly demanding r

hardwarearxiv-cs-cv
29 Jul 2026
Model Releases

RDVSv2: A Large-scale Benchmark for RGB-D Video Salient Object Detection

DGX agent

arXiv:2607.25392v1 Announce Type: new Abstract: We introduce RDVSv2, a large-scale benchmark for RGB-D video salient object detection (RGB-D VSOD) with dense frame-level annotations. Existing datasets

model-releasesarxiv-cs-cv
29 Jul 2026
Research

Reading Legends on Ancient Coins: An Object Detection Approach for Character Recognition on a Novel Roman Republican Dataset

DGX agent

arXiv:2607.25455v1 Announce Type: new Abstract: When it comes to the proper classification of ancient coins with respect to their time and issuer, the textual inscriptions on these coins, also known a

researcharxiv-cs-cv
29 Jul 2026
Model Releases

ReDesign: Recovering Editable Design Structures from Images via Agentic Decomposition

DGX agent

arXiv:2607.25565v1 Announce Type: new Abstract: Recovering an editable design file from a raster image is a common and costly bottleneck in modern design workflows, yet remains challenging since edita

model-releasesarxiv-cs-cv
29 Jul 2026
Safety

Safety-Aware Cascaded Inference for Crop Damage Assessment with Controlled Error Trade-offs

DGX agent

arXiv:2607.25468v1 Announce Type: new Abstract: In picture-based agricultural insurance for smallholder farmers, missed damage detections carry substantially higher cost than false alarms: a farmer wh

safetyarxiv-cs-cv
29 Jul 2026
Model Releases

SAM-MI: A Mask-Injected Framework for Enhancing Open-Vocabulary Semantic Segmentation with SAM

DGX agent

arXiv:2511.20027v2 Announce Type: replace Abstract: Open-vocabulary semantic segmentation (OVSS) aims to segment and recognize objects universally. Trained on extensive high-quality segmentation data,

model-releasesarxiv-cs-cv
29 Jul 2026
Research

Schrodinger's Cat: Probabilistic Representation and Prediction of Potential Scene Kinematics

DGX agent

arXiv:2607.25984v1 Announce Type: new Abstract: Predicting how a scene may evolve from partial observations requires reasoning about multiple possible futures rather than committing to a single trajec

researcharxiv-cs-cv
29 Jul 2026
Model Releases

ScoreShield: Differentially Private Release of Similarity Scores

DGX agent

arXiv:2607.25041v1 Announce Type: cross Abstract: A growing number of applications, such as biometrics and retrieval-augmented generation (RAG), rely on cosine similarity scores computed between vecto

model-releasesarxiv-cs-cv
29 Jul 2026
Model Releases

Sense it with your eyes: Sensation Generation and Understanding for Advertisements

DGX agent

arXiv:2607.25314v1 Announce Type: new Abstract: Sensory advertising evokes human senses through visual cues, enabling audiences to mentally simulate experiences and increasing persuasive impact. Despi

model-releasesarxiv-cs-cv
29 Jul 2026
Research

SepPrune:A Separator-based Pruning Framework for Efficient Multimodal Large Language Models

DGX agent

arXiv:2607.25818v1 Announce Type: new Abstract: Recent multimodal large language models (MLLMs), such as Qwen2.5-VL and InternVL3, generate large numbers of vision tokens for high-resolution inputs, l

researcharxiv-cs-cv
29 Jul 2026
Model Releases

SurgSLOT: Segment Anything in Surgical Videos via Semantic Long-term Tracking

DGX agent

arXiv:2511.16618v2 Announce Type: replace Abstract: Surgical scene understanding demands temporally consistent tracking of instruments and tissues. For clinical use, such tracking should generalize to

model-releasesarxiv-cs-cv
29 Jul 2026
Research

TIGA: Trajectory-Injected Generative Attack against Black-box AIGC Detectors

DGX agent

arXiv:2607.25894v1 Announce Type: new Abstract: Recent diffusion models have achieved remarkable realism in facial image synthesis, posing growing challenges to artificial intelligence-generated conte

researcharxiv-cs-cv
29 Jul 2026
Model Releases

Towards Faithful Sentimental Image Captioning via Evidence-Aware Multi-Agent Reasoning

DGX agent

arXiv:2607.25789v1 Announce Type: new Abstract: Sentimental Image Captioning (SIC) requires balancing emotional expression with visual fidelity. Existing methods often struggle with this trade-off, le

model-releasesarxiv-cs-cv
29 Jul 2026
Research

Towards Reliable Stain Transfer: An Iterative Data-Model Co-Optimization Framework Based on Multimodal Expert-Guided Assessment

DGX agent

arXiv:2607.25393v1 Announce Type: new Abstract: Histopathological examination primarily relies on hematoxylin and eosin (H&E) and immunohistochemistry (IHC) staining. Although IHC provides critical mo

researcharxiv-cs-cv
29 Jul 2026
Model Releases

Track-Leakage-Free Hold-Out Self-Validation for Photogrammetric Reconstruction: Protocol, Sensitivity, and Limits

DGX agent

arXiv:2607.24852v1 Announce Type: new Abstract: Automated photogrammetric inspection emits metric measurements from a 3D reconstruction whose own correctness is normally unknown without an external su

model-releasesarxiv-cs-cv
29 Jul 2026
Tutorials

Unifying Active Learning and Semi-Supervised Learning for Medical Image Segmentation

DGX agent

arXiv:2607.25014v1 Announce Type: new Abstract: In practical settings, medical image segmentation models are often developed with limited annotated data rather than fully labeled datasets. Training fr

tutorialsarxiv-cs-cv
29 Jul 2026
Model Releases

Universal Pansharpening Model

DGX agent

arXiv:2603.03831v2 Announce Type: replace Abstract: Pansharpening generates the high-resolution multi-spectral (MS) image by integrating spatial details from a texture-rich panchromatic (PAN) image an

model-releasesarxiv-cs-cv
29 Jul 2026
Safety

VetClaw: An Edge-Cloud Multimodal Agentic System for Veterinary Disease Screening

DGX agent

arXiv:2607.26042v1 Announce Type: new Abstract: We present VetClaw, an edge-cloud multimodal agentic system for early veterinary disease screening. VetClaw uses a camera module as an edge sensing devi

safetyarxiv-cs-cv
29 Jul 2026
Research

WHTMix: Efficient Stereo Depth Estimation via Walsh-Hadamard Token Mixing

DGX agent

arXiv:2607.25234v1 Announce Type: new Abstract: Stereo depth estimation for driving, robotics and augmented reality must run at high resolution under tight latency budgets, yet in transformer-based ma

researcharxiv-cs-cv
29 Jul 2026
Research

Wonder: Video World Model Done Better

DGX agent

arXiv:2607.26037v1 Announce Type: new Abstract: We present Wonder, a general-purpose video world model for real-time, camera-controllable world exploration. Given an image or a conditional video, Wond

researcharxiv-cs-cv
29 Jul 2026
Model Releases

A Controlled Visual-Backbone Benchmark for Multimodal Short-Term Solar Irradiance Forecasting

DGX agent

arXiv:2607.23633v1 Announce Type: cross Abstract: Sky-image irradiance studies often compare forecasting systems in which the image encoder, temporal model, fusion block, target definition, and traini

model-releasesarxiv-cs-cv
28 Jul 2026
Research

A Diagnostic Gap Framework for Evaluating Reconstruction Fidelity in Weakly Supervised Mammography

DGX agent

arXiv:2607.22740v1 Announce Type: new Abstract: Weakly supervised pipelines for medical imaging have become increasingly popular over the years. These systems often include multiple stages and compone

researcharxiv-cs-cv
28 Jul 2026
Research

A Modern ConvNet for Solar Filament Detection

DGX agent

arXiv:2607.24525v1 Announce Type: cross Abstract: Automated solar filament detection using deep learning faces several challenges. Semantic segmentation of solar filaments is a complicated multiscale

researcharxiv-cs-cv
28 Jul 2026
Research

A Reconstruction-Based Framework for Caption Evaluation Beyond Reference Captions

DGX agent

arXiv:2607.23235v1 Announce Type: new Abstract: Image captioning is a primary task in vision--language research, yet assessing how faithfully a caption preserves image semantics without relying on ref

researcharxiv-cs-cv
28 Jul 2026
Research

A Reference-Free Framework for Evaluating Single-Frame ISP Pipelines

DGX agent

arXiv:2607.23321v1 Announce Type: cross Abstract: Evaluating camera image signal processing (ISP) pipelines requires measuring low-level artifacts introduced by operations such as denoising, demosaici

researcharxiv-cs-cv
28 Jul 2026
Model Releases

A Scale-adaptive Vision Model Links C. elegans Neuronal Morphology to Behavior for Neurotoxicity Assessment

DGX agent

arXiv:2607.23183v1 Announce Type: cross Abstract: Neurological disorders are a leading cause of global disability and are increasingly linked to environmental chemical exposures. Yet neurotoxicity ass

model-releasesarxiv-cs-cv
28 Jul 2026
Research

A Unified Stereo Geometry Estimation Framework for Disparity and Surface Normal

DGX agent

arXiv:2607.24024v1 Announce Type: new Abstract: Stereo matching and surface normal estimation are fundamental tasks in 3D vision. However, existing feed-forward stereo methods still struggle to produc

researcharxiv-cs-cv
28 Jul 2026
Agents

Accuracy potential of visual localization exploiting high-end street-level imagery

DGX agent

arXiv:2607.24409v1 Announce Type: new Abstract: Accurate and reliable pose information with respect to a reference frame is increasingly demanded across applications such as autonomous navigation, sur

agentsarxiv-cs-cv
28 Jul 2026
Research

Act, Think or Abstain: Complexity-Aware Adaptive Inference for Vision-Language-Action Models

DGX agent

arXiv:2603.05147v2 Announce Type: replace Abstract: Current research on Vision-Language-Action (VLA) models predominantly focuses on enhancing generalization through reasoning techniques. While effect

researcharxiv-cs-cv
28 Jul 2026
Research

AdaKAN: A dual-branch adaptive Kolmogorov-Arnold network for medical image segmentation

DGX agent

arXiv:2607.22891v1 Announce Type: new Abstract: Medical image segmentation is a fundamental task in computer-aided diagnosis, yet it remains challenging due to the complexity of anatomical structures

researcharxiv-cs-cv
28 Jul 2026
← Previous
1…3536373839…261
Next →