AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,515 results
Model Releases

Towards Generalized Synapse Detection Across Invertebrate Species

DGX agent

arXiv:2509.17041v2 Announce Type: replace Abstract: Behavioural differences across organisms, whether healthy or pathological, are closely tied to the structure of their neural circuits. Yet, the fine

model-releasesarxiv-cs-cv
31 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Towards Practical Algorithm Selection for Unsupervised Domain Adaptation in Medical Imaging

DGX agent

arXiv:2607.28125v1 Announce Type: new Abstract: Numerous unsupervised domain adaptation (UDA) algori-thms exist, but for clinical practice, selecting the best-suited one along with proper hyperparamet

researcharxiv-cs-cv
31 Jul 2026
Safety

Towards Real-Time PixOOD: Efficient Anomaly Segmentation for Autonomous Vehicles

DGX agent

arXiv:2607.28483v1 Announce Type: new Abstract: Real-time anomaly segmentation is essential for the safety of autonomous systems. Although recent approaches offer high accuracy, their computational co

safetyarxiv-cs-cv
31 Jul 2026
Tutorials

Towards Robust Monocular Depth Estimation in Non-Lambertian Surfaces

DGX agent

arXiv:2408.06083v2 Announce Type: replace Abstract: In the field of monocular depth estimation (MDE), many models with excellent zero-shot performance in general scenes emerge recently. However, these

tutorialsarxiv-cs-cv
31 Jul 2026
Model Releases

Towards Unified Multimodal Misinformation Detection in Social Media: A Benchmark Dataset and Baseline

DGX agent

arXiv:2509.25991v3 Announce Type: replace-cross Abstract: Detecting deceptive multimodal content on social media has become an increasingly important problem. Two major types of deception dominate: hu

model-releasesarxiv-cs-cv
31 Jul 2026
Research

TSOG: A Format For Temporally And Spatially Ordered Gaussians

DGX agent

arXiv:2607.28049v1 Announce Type: cross Abstract: We propose Temporally and Spatially Ordered Gaussians (TSOG), a format for efficient representation of 4D Gaussian Splatting (4DGS) content. TSOG exte

researcharxiv-cs-cv
31 Jul 2026
Model Releases

Tycho: Active Abstraction with Programmatic World Models for ARC-AGI-3

DGX agent

arXiv:2607.28287v1 Announce Type: cross Abstract: ARC-AGI-3 turns abstraction into an interactive problem of skill acquisition. A player must infer an unfamiliar game's rules, hidden state, and goal w

model-releasesarxiv-cs-cv
31 Jul 2026
Applications

Uncertainty-Aware Multimodal Fusion for Oral Lesion Classification

DGX agent

arXiv:2511.12268v3 Announce Type: replace-cross Abstract: Early detection of oral cancer and potentially malignant diseases is a major challenge in low-resource settings due to the scarcity of annotat

applicationsarxiv-cs-cv
31 Jul 2026
Research

Understanding Submodular Information Measure Based Objectives for Representation Learning: A Variance and Separation Perspective

DGX agent

arXiv:2607.27660v1 Announce Type: cross Abstract: Submodular Information Measures (SIMs) have recently emerged as a powerful framework for representation learning and multimodal learning. In particula

researcharxiv-cs-cv
31 Jul 2026
Safety

UniCross: Unified Cross-Skill Dexterous Manipulation Synthesis

DGX agent

arXiv:2607.28198v1 Announce Type: cross Abstract: Many dexterous manipulation tasks require the object to remain securely held throughout the interaction. From the perspective of hand-object relationa

safetyarxiv-cs-cv
31 Jul 2026
Safety

Unifying Adversarially Robust Model Experts in Vision-Language Models

DGX agent

arXiv:2607.27897v1 Announce Type: new Abstract: Vision-language models (VLMs), such as CLIP, are vulnerable to adversarial attacks, posing a serious problem for real-life applications and deployment.

safetyarxiv-cs-cv
31 Jul 2026
Local Ai

VCP-DCN: Beyond Visual Concealed Property via Depth Collaborative Network for Camouflaged Object Detection

DGX agent

arXiv:2607.27843v1 Announce Type: new Abstract: Camouflaged Object Detection (COD) aims to identify and segment camouflaged objects in complex environments, which are often concealed because their col

local-aiarxiv-cs-cv
31 Jul 2026
Research

VETO: Towards Protecting Images From Frontier AI Editing

DGX agent

arXiv:2607.27292v1 Announce Type: new Abstract: The rise of powerful, accessible image-editing models such as FLUX.2 has brought high-fidelity editing within broad reach. Their capabilities now extend

researcharxiv-cs-cv
31 Jul 2026
Agents

VideoCoCo: Code-as-CoT for Physically-Consistent Video Generation via an Agentic Dual-Engine System

DGX agent

arXiv:2607.27380v1 Announce Type: new Abstract: Text-to-video models have achieved remarkable visual quality, yet they still struggle to generate physically consistent dynamics because the temporal ev

agentsarxiv-cs-cv
31 Jul 2026
Applications

ViewMind3D: Modular View-Aware Inference for Training-Free 3D-QA

DGX agent

arXiv:2607.28442v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) and vision-language models (VLMs) have enabled new possibilities for 3D question answering (3D-QA), a ke

applicationsarxiv-cs-cv
31 Jul 2026
Research

ViP-Rig: Visual-Prompted Controllable Rigging

DGX agent

arXiv:2607.27982v1 Announce Type: new Abstract: Rigging is inherently task-dependent because the same mesh may require different skeletons and deformation behaviors across animation tasks. In practice

researcharxiv-cs-cv
31 Jul 2026
Research

VisualRouter: Query-Grounded Visual Sampling for Long Video Understanding

DGX agent

arXiv:2607.28463v1 Announce Type: new Abstract: Large vision-language models (LVLMs) have achieved significant progress in video understanding, yet understanding long videos remains challenging due to

researcharxiv-cs-cv
31 Jul 2026
Model Releases

What Makes Deep Learning Work for Traditional Chinese Medicine Tongue Diagnosis? A Comprehensive Ablation Study

DGX agent

arXiv:2607.28148v1 Announce Type: new Abstract: Deep learning has shown promise for automated tongue diagnosis in traditional Chinese medicine (TCM), yet the design space remains underexplored. We con

model-releasesarxiv-cs-cv
31 Jul 2026
Research

What to Remove, What to Preserve: Dual-Ambiguity Rectification for All-in-One Image Restoration

DGX agent

arXiv:2607.28526v1 Announce Type: new Abstract: All-in-one image restoration aims to handle diverse degradations within a unified framework. Existing methods commonly encode heterogeneous degradation

researcharxiv-cs-cv
31 Jul 2026
Research

Witness Evidence Portfolios: Single-Prefill Risk Detection for Closed Multimodal Answers

DGX agent

arXiv:2607.27667v1 Announce Type: new Abstract: Reliable deployment of multimodal large language models (MLLMs) requires deciding whether a confident visual answer should be trusted, reviewed, or rout

researcharxiv-cs-cv
31 Jul 2026
Research

You Only Look Omni Gradient Backpropagation for Moving Infrared Small Target Detection

DGX agent

arXiv:2511.13013v2 Announce Type: replace Abstract: Moving infrared small target detection is a key component of infrared search and tracking systems, yet it remains extremely challenging due to low s

researcharxiv-cs-cv
31 Jul 2026
Research

ZMIS-SAM: Segment Anything Model Enhanced with Wavelet Transform for Zooplankton Microscopy Image Instance Segmentation

DGX agent

arXiv:2607.27585v1 Announce Type: new Abstract: As primary consumers in the marine food chain, zooplankton play a crucial role in maintaining marine ecological balance. However, the Segment Anything M

researcharxiv-cs-cv
31 Jul 2026
Local Ai

3DGBGS: 3D Granular Ball Gaussian Splatting for Compact Novel View Synthesis

DGX agent

arXiv:2607.26578v1 Announce Type: new Abstract: Three-dimensional Gaussian Splatting (3DGS) enables high-quality real-time novel-view synthesis through explicit Gaussian primitives and differentiable

local-aiarxiv-cs-cv
30 Jul 2026
Research

A Closer Look at Dynamic Scene Graph Generation In the Era of Multimodal Large Language Models

DGX agent

arXiv:2503.15846v2 Announce Type: replace Abstract: Dynamic Scene Graph Generation (DSGG) aims to capture objects and their evolving relations in videos. Despite recent progress, the practicality and

researcharxiv-cs-cv
30 Jul 2026
Safety

A Picture Says Thousands of Words - Harnessing Dermal Exposure Data from Images through Hybrid Deep Learning for Enhanced Safety Assessment

DGX agent

arXiv:2607.26170v1 Announce Type: new Abstract: This study developed a hybrid computer vision method to quantify exposed skin from images for dermal exposure assessment. Using 170 indoor-painting imag

safetyarxiv-cs-cv
30 Jul 2026
Safety

Anatomy Contextualized Adaption of CT Foundation Models

DGX agent

arXiv:2607.27154v1 Announce Type: new Abstract: CT vision-language foundation models have demonstrated promising performance across downstream tasks, but are typically trained with whole-volume repres

safetyarxiv-cs-cv
30 Jul 2026
Safety

Anchoring and Steering Diffusion: Enhancing the Faithfulness of Text-to-Image Generation at Inference Time

DGX agent

arXiv:2607.26647v1 Announce Type: new Abstract: While text-to-image diffusion models achieve impressive visual quality, they frequently struggle to maintain precise alignment with complex compositiona

safetyarxiv-cs-cv
30 Jul 2026
Hardware

BATS: Resource-Efficient Volumetric Segmentation with Boundary-Aware Mixed-Resolution Tokens

DGX agent

arXiv:2607.26829v1 Announce Type: new Abstract: Many high-performing volumetric segmentation models maintain dense multi-scale feature maps, leading to high activation memory and inference cost. We pr

hardwarearxiv-cs-cv
30 Jul 2026
Model Releases

BG-REAL: A Public Real-Data Anchored Benchmark for Background Manipulation Detection and Localization

DGX agent

arXiv:2607.26232v1 Announce Type: new Abstract: Background manipulation is a practical but under-specified image-forensics setting: the manipulated evidence can sit outside the salient foreground obje

model-releasesarxiv-cs-cv
30 Jul 2026
Model Releases

BrainG3N: A Dual-Purpose Tokenizer for Controllable 3D Brain MRI Generation

DGX agent

arXiv:2606.19651v2 Announce Type: replace-cross Abstract: Three-dimensional (3D) brain MRI is central to clinical neurology and neuro-oncology, where generative models could augment under-represented

model-releasesarxiv-cs-cv
30 Jul 2026
Tutorials

Breaking the Stealth-Potency Trade-off in Clean-Image Backdoors with Generative Trigger Optimization

DGX agent

arXiv:2511.07210v3 Announce Type: replace Abstract: Clean-image backdoor attacks, which use only label manipulation in training datasets to compromise deep neural networks, pose a significant threat t

tutorialsarxiv-cs-cv
30 Jul 2026
Model Releases

Calibri: Enhancing Diffusion Transformers via Parameter-Efficient Calibration

DGX agent

arXiv:2603.24800v2 Announce Type: replace Abstract: In this paper, we uncover the hidden potential of Diffusion Transformers (DiTs) to significantly enhance generative tasks. Through an in-depth analy

model-releasesarxiv-cs-cv
30 Jul 2026
Safety

CASIAL: Geometric Distortion Robust Image Watermarking

DGX agent

arXiv:2607.26729v1 Announce Type: new Abstract: Deep learning-based watermarking has shown strong robustness against non-geometric distortions, yet its performance under geometric transformations rema

safetyarxiv-cs-cv
30 Jul 2026
Safety

CG-World: A Large-Scale World-State Dataset and Protocol for World Models

DGX agent

arXiv:2607.26452v1 Announce Type: cross Abstract: World models must learn the joint dynamics of states, actions, events, and observations, yet existing video, robotics, and simulation datasets usually

safetyarxiv-cs-cv
30 Jul 2026
Safety

CinemaTraj: Composing Atomic Camera Trajectories for 3D Scenes with LLM Agents

DGX agent

arXiv:2607.26910v1 Announce Type: new Abstract: Automatically generating cinematically expressive camera trajectories through 3D scenes from natural language descriptions is a challenging task of high

safetyarxiv-cs-cv
30 Jul 2026
Safety

CineWeaver: Training-Free Reference-Controllable Multi-Shot Long Video Generation for Cinematic Storytelling

DGX agent

arXiv:2607.26529v1 Announce Type: new Abstract: Cinematic video generation is challenging for text-to-video diffusion models due to concurrent requirements on multi-shot generation, fine-grained contr

safetyarxiv-cs-cv
30 Jul 2026
Research

Classification of Disease from Lungs X-ray Images using VGG16, VGG19 and ResNet50 Models

DGX agent

arXiv:2607.26580v1 Announce Type: new Abstract: With the increase in the number of cases related to respiratory diseases, there is an urgent need to detect them early and diagnose them accurately. Con

researcharxiv-cs-cv
30 Jul 2026
Research

Clinical Graph-Mediated Distillation for Unpaired MRI-to-CFI Hypertension Prediction

DGX agent

arXiv:2603.21809v2 Announce Type: replace Abstract: Retinal fundus imaging enables low-cost and scalable hypertension (HTN) screening, but HTN-related retinal cues are subtle, yielding high-variance p

researcharxiv-cs-cv
30 Jul 2026
Research

Comparing the Performance of Foundation Model Derived Embeddings with Traditional Approaches for Distant Metastasis Prediction in Head and Neck Cancer

DGX agent

arXiv:2607.26276v1 Announce Type: new Abstract: Background: Early prediction of distant metastasis (DM) risk in head and neck cancer (HNC) can enable timely interventions that may improve treatment ou

researcharxiv-cs-cv
30 Jul 2026
Applications

ContactFlow: A video action conditioning that transfers across embodiments

DGX agent

arXiv:2607.26579v1 Announce Type: cross Abstract: World models offer a promising route toward robot planning by enabling agents to imagine and verify the consequences of actions before execution. Howe

applicationsarxiv-cs-cv
30 Jul 2026
Research

Context-measure: Contextualizing Metric for Camouflage

DGX agent

arXiv:2512.07076v4 Announce Type: replace Abstract: Camouflage relies heavily on context, but current metrics used in camouflaged object segmentation ignore contextual cues. We identify two major draw

researcharxiv-cs-cv
30 Jul 2026
Model Releases

Decoupled Visual Processing: Efficient Multimodal Adaptation via Modality-Specific Transformer Substitution

DGX agent

arXiv:2607.26596v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have demonstrated remarkable capabilities by integrating visual and textual understanding within a unified tran

model-releasesarxiv-cs-cv
30 Jul 2026
Research

Deep Expert Injection for Anchoring Retinal VLMs with Domain-Specific Knowledge

DGX agent

arXiv:2603.07131v4 Announce Type: replace Abstract: Large Vision Language Models (LVLMs) show immense potential for automated ophthalmic diagnosis. However, their clinical deployment is severely hinde

researcharxiv-cs-cv
30 Jul 2026
Safety

DistillAlign: Coordinating Mode Covering and Mode Seeking in Autoregressive Video Distillation

DGX agent

arXiv:2607.26811v1 Announce Type: new Abstract: Existing autoregressive video distillation methods commonly adopt a Distribution Matching Distillation (DMD)-based multi-stage pipeline. However, they t

safetyarxiv-cs-cv
30 Jul 2026
Local Ai

DLAM: Distributional Latent Actions with Temporal Constraints

DGX agent

arXiv:2607.27138v1 Announce Type: cross Abstract: Vision-language-action (VLA) models remain constrained by scarce action-labeled robot data, whereas action-free videos offer abundant observations of

local-aiarxiv-cs-cv
30 Jul 2026
Safety

Do Unified Multimodal Models Think in One Space? A Lens Through Cross-Branch Steering

DGX agent

arXiv:2607.26411v1 Announce Type: new Abstract: Unified multimodal models (UMMs) aim to integrate understanding and generation within a single architecture, yet it remains unclear whether these capabi

safetyarxiv-cs-cv
30 Jul 2026
Safety

Dual Inversion for Text-to-Image Diffusion Models: From Both Prompt and Noise Perspectives

DGX agent

arXiv:2607.26735v1 Announce Type: new Abstract: Prompt inversion, as a typical reverse engineering technique, enables text-to-image (T2I) diffusion models to generate the desired target images without

safetyarxiv-cs-cv
30 Jul 2026
Agents

DVPSFormer: Efficient Online Depth-aware Video Panoptic Segmentation for Autonomous Driving

DGX agent

arXiv:2607.26165v1 Announce Type: new Abstract: Safe autonomous navigation requires a holistic understanding of dynamic environments, necessitating the simultaneous estimation of metric depth, semanti

agentsarxiv-cs-cv
30 Jul 2026
← Previous
1…3132333435…261
Next →