AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,515 results
Research

Event Detection in Videos: A Framework for the Development of New Methods

DGX agent

arXiv:2607.04372v1 Announce Type: new Abstract: Event detection tasks in videos, the most important aspect of video surveillance, aim to detect events either at the pixel-level, frame-level, or clip-l

researcharxiv-cs-cv
7 Jul 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Explainable Flood Segmentation on Sentinel-1 SAR1 Imagery Using CNN and Transformer Architectures

DGX agent

arXiv:2606.16302v2 Announce Type: replace Abstract: Rapid and accurate flood prediction is essential for disaster response and mitigation planning. Synthetic Aperture Radar (SAR) sensors in satellites

model-releasesarxiv-cs-cv
7 Jul 2026
Applications

Exploring SAM Supervision for Fine-Grained UAV Target Segmentation under Data Scarcity

DGX agent

arXiv:2607.03754v1 Announce Type: new Abstract: Unmanned aerial vehicle (UAV) target segmentation remains challenging due to the small size of objects, appearance variations, cluttered backgrounds, an

applicationsarxiv-cs-cv
7 Jul 2026
Model Releases

ExpoMotion: A Large-Scale Benchmark and A Householder Projection Network for Multi-Exposure Fusion

DGX agent

arXiv:2607.03110v1 Announce Type: new Abstract: Multi-Exposure Fusion (MEF) effectively extends dynamic range, but practical deployment is hindered by motion-induced ghosting and the scarcity of high-

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

FairFlow: Demystifying and Mitigating Stereotype Bias in Text-to-Image Diffusion Transformers

DGX agent

arXiv:2607.03180v1 Announce Type: new Abstract: Multimodal diffusion transformers (MM-DiTs) have emerged as the prevalent backbone for modern text-to-image generation systems. However, they exhibit cr

model-releasesarxiv-cs-cv
7 Jul 2026
Agents

Fast 3D Foundation Model Initialized Gaussian Splatting

DGX agent

arXiv:2607.03209v1 Announce Type: new Abstract: This paper introduces a fast method for high-quality 3D Gaussian Splatting (3DGS) reconstruction without traditional Structure-from-Motion (SfM). The pr

agentsarxiv-cs-cv
7 Jul 2026
Local Ai

FDR-Occ: Factorized Dense Routing for Full-Spectrum 3D Occupancy Prediction

DGX agent

arXiv:2607.03822v1 Announce Type: new Abstract: Vision-based 3D occupancy prediction fundamentally relies on the 2D-to-3D view transformation. Current paradigms predominantly utilize explicit physical

local-aiarxiv-cs-cv
7 Jul 2026
Model Releases

FedProIn: Mitigating Client Drift for Learnable Prototypes in Federated Medical Imaging

DGX agent

arXiv:2607.04158v1 Announce Type: cross Abstract: Federated learning (FL) is severely hindered by statistical heterogeneity due to variations in scanners, acquisition protocols, and patient population

model-releasesarxiv-cs-cv
7 Jul 2026
Research

Fields of the Planet: Field Boundary Mapping Beyond 10m

DGX agent

arXiv:2607.04449v1 Announce Type: new Abstract: Field-boundary maps support crop monitoring, irrigation planning, and yield estimation, but many smallholder parcels span only a few 10 m Sentinel-2 pix

researcharxiv-cs-cv
7 Jul 2026
Research

Flash-BoN: Instant Drafts for Inference-Time Scaling in Diffusion Models

DGX agent

arXiv:2607.04461v1 Announce Type: new Abstract: Inference-time scaling for text-to-image generation has progressed from simple Best-of-N (BoN) sampling to guided search methods that verify and steer c

researcharxiv-cs-cv
7 Jul 2026
Safety

Flex-Forcing: Towards a Unified Autoregressive and Bidirectional Video Diffusion Model

DGX agent

arXiv:2607.03509v1 Announce Type: new Abstract: Recent progress in large-scale generative models has substantially advanced video generation, yet existing methods remain constrained by a rigid inferen

safetyarxiv-cs-cv
7 Jul 2026
Applications

FlowMark: Mask-Guided Video Watermarking

DGX agent

arXiv:2607.05261v1 Announce Type: new Abstract: We present FlowMark, a video watermarking framework guided by automatically predicted object masks. In contrast to prior region-based approaches that re

applicationsarxiv-cs-cv
7 Jul 2026
Model Releases

Fortifying Fully Convolutional Generative Adversarial Networks for Image Super-Resolution Using Divergence Measures

DGX agent

arXiv:2404.06294v2 Announce Type: replace-cross Abstract: Super-Resolution (SR) is a time-hallowed image processing problem that aims to improve the quality of a Low-Resolution (LR) sample up to the s

model-releasesarxiv-cs-cv
7 Jul 2026
Research

Fourier Splatting: Generalized Fourier encoded primitives for scalable radiance fields

DGX agent

arXiv:2603.19834v3 Announce Type: replace Abstract: Novel view synthesis has recently been revolutionized by 3D Gaussian Splatting (3DGS), which enables real-time rendering through explicit primitive

researcharxiv-cs-cv
7 Jul 2026
Model Releases

Framework and Multi-modal Dataset for Roadwork Zone Detection and Geo-localization

DGX agent

arXiv:2607.04330v1 Announce Type: new Abstract: Autonomous vehicles often rely on high-definition (HD) maps for navigation; however, these maps are not frequently updated and often lack semi-static in

model-releasesarxiv-cs-cv
7 Jul 2026
Safety

FRFDet: Efficient UAV Small Object Detection with Symmetric Sampling and Scalable Fusion

DGX agent

arXiv:2607.04125v1 Announce Type: new Abstract: Small object detection in Unmanned Aerial Vehicle (UAV) imagery remains challenging under adverse conditions, including complex weather, low illuminatio

safetyarxiv-cs-cv
7 Jul 2026
Model Releases

From General Actions to Domain-Specific Monitoring: Prior-Adaptive Transfer for Skeleton-Based Action Recognition

DGX agent

arXiv:2607.03327v1 Announce Type: new Abstract: Skeleton-based action recognition models have recently shown strong performance on large-scale benchmarks with general actions. However, directly transf

model-releasesarxiv-cs-cv
7 Jul 2026
Applications

From Geometric Labels to Semantic Understanding of Indoor Building Components Using Multimodal Large Language Models

DGX agent

arXiv:2607.03661v1 Announce Type: new Abstract: Point cloud-based understanding has become an important enabler for facility operation and maintenance involving indoor building components. However, ex

applicationsarxiv-cs-cv
7 Jul 2026
Safety

From Open Loop to Closed Loop: A Test-Time Iterative Optimization Framework for Reference-Consistent Image Generation

DGX agent

arXiv:2607.04691v1 Announce Type: new Abstract: While controllable image generation has made significant strides by incorporating visual reference conditions, existing methods predominantly operate as

safetyarxiv-cs-cv
7 Jul 2026
Safety

From Region Arrival to Instance-Level Grounding in Vision-and-Language Navigation

DGX agent

arXiv:2607.03792v1 Announce Type: cross Abstract: Vision-and-Language Navigation (VLN) agents may satisfy conventional success criteria while still failing to establish reliable object-level grounding

safetyarxiv-cs-cv
7 Jul 2026
Applications

FSDC-DETR: A Frequency-Spatial Domain Collaborative DETR for Small Object Detection

DGX agent

arXiv:2607.05176v1 Announce Type: new Abstract: Small object detection (SOD) remains a challenging task in real-world applications. Despite recent advances, existing detectors remain limited by rigid

applicationsarxiv-cs-cv
7 Jul 2026
Model Releases

Fully Rotation-Equivariant Spectral-Spatial Learning for Multispectral Object Detection

DGX agent

arXiv:2607.05148v1 Announce Type: new Abstract: Existing multispectral detectors are limited by discrete spectral processing, a scale-dependent shift in the relative reliability of spectral and spatia

model-releasesarxiv-cs-cv
7 Jul 2026
Local Ai

FunPhase: A Periodic Functional Autoencoder for Motion Generation via Phase Manifolds

DGX agent

arXiv:2512.09423v2 Announce Type: replace Abstract: Learning natural body motion remains challenging due to the strong coupling between spatial geometry and temporal dynamics. Embedding motion in phas

local-aiarxiv-cs-cv
7 Jul 2026
Local Ai

G^2TAM: Geometry Grounded Track Anything Model

DGX agent

arXiv:2607.03789v1 Announce Type: new Abstract: Human spatial understanding arises from jointly perceiving geometry and semantics, enabling consistent object identification and localization across vie

local-aiarxiv-cs-cv
7 Jul 2026
Model Releases

G3Splat: Geometrically Consistent Generalizable Gaussian Splatting

DGX agent

arXiv:2512.17547v2 Announce Type: replace Abstract: 3D Gaussians have become a powerful scene representation for real-time splatting and high-quality novel-view synthesis. This has motivated generaliz

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

GALOSH: Blind, Training-Free Denoising of Raw Bayer and sRGB Images by Parallel-Friendly Local Shrinkage

DGX agent

arXiv:2607.03768v1 Announce Type: cross Abstract: Classical training-free denoisers such as BM3D and non-local means owe much of their strength to search: content-dependent block matching whose memory

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

GaussianArt: Unified Modeling of Geometry and Motion for Articulated Objects

DGX agent

arXiv:2508.14891v3 Announce Type: replace Abstract: Reconstructing articulated objects is essential for building digital twins of interactive environments. However, prior methods typically decouple ge

model-releasesarxiv-cs-cv
7 Jul 2026
Research

Geographic Diversity Beats Data Volume for Cross-Domain Generalization in Zero-Label JEPA Driving World Models

DGX agent

arXiv:2607.04500v1 Announce Type: new Abstract: Self-supervised latent world models can assign a surprise score to driving scenarios without any human labels. A natural follow-up question is whether s

researcharxiv-cs-cv
7 Jul 2026
Model Releases

Geometric Observability Index: An Operator-Theoretic Framework for Per-Feature Sensitivity, Weak Observability, and Dynamic Effects in SE(3) Pose Estimation

DGX agent

arXiv:2602.05582v2 Announce Type: replace Abstract: We introduce the Geometric Observability Index (GOI), a per-feature sensitivity measure for pose estimation on SE(3). For a Gauss-Newton curvature m

model-releasesarxiv-cs-cv
7 Jul 2026
Research

Geometric Reciprocity: Unlocking Self-Supervision for Stereoscopic Video Generation

DGX agent

arXiv:2607.05354v1 Announce Type: new Abstract: Monocular-to-stereo conversion synthesizes stereoscopic content from 2D videos for immersive 3D experiences. In modern Depth-Image-Based Rendering (DIBR

researcharxiv-cs-cv
7 Jul 2026
Model Releases

Geometry-aware Depth-guided Representation Learning for Structure-preserving Low-light Image Enhancement

DGX agent

arXiv:2607.05005v1 Announce Type: new Abstract: Low-light degradation reduces image visibility and weakens structural cues that are important for visual representation and scene understanding. Existin

model-releasesarxiv-cs-cv
7 Jul 2026
Research

GeoSAM-Lite: A Lightweight Foundation Model for Onboard Remote Sensing Segmentation

DGX agent

arXiv:2607.03760v1 Announce Type: new Abstract: The deployment of large-scale foundation models like Segment Anything Model (SAM) on resource-constrained Earth observation platforms is hindered by pro

researcharxiv-cs-cv
7 Jul 2026
Applications

GeoWorld: Providing Full-frame Geometry Features to Facilitate 3D Scene Generation

DGX agent

arXiv:2511.23191v2 Announce Type: replace Abstract: Previous works that leverage video models for image-to-3D scene generation often suffer from geometric distortions and blurry content. Using video g

applicationsarxiv-cs-cv
7 Jul 2026
Research

GestaltMML: Enhancing Rare Genetic Disease Diagnosis through Multimodal Machine Learning Combining Facial Images and Clinical Text

DGX agent

arXiv:2312.15320v3 Announce Type: replace-cross Abstract: Individuals with suspected rare genetic disorders often undergo multiple clinical evaluations, imaging studies, laboratory tests, and genetic

researcharxiv-cs-cv
7 Jul 2026
Model Releases

Ghosts Beneath Textures: Texture-Relation Cues for Cross-Paradigm AI-Generated Image Detection

DGX agent

arXiv:2607.03862v1 Announce Type: new Abstract: AI-generated images have proliferated rapidly, motivating extensive research. Most existing AI-generated image detectors are developed and evaluated und

model-releasesarxiv-cs-cv
7 Jul 2026
Applications

GlaBoost: A Multimodal Structured Framework for Glaucoma Risk Stratification

DGX agent

arXiv:2508.03750v2 Announce Type: replace-cross Abstract: Early and accurate glaucoma detection is critical to prevent irreversible vision loss, yet existing AI methods often rely on unimodal inputs a

applicationsarxiv-cs-cv
7 Jul 2026
Research

GlacierCastAI: Predicting Glacier Retreat from Multi-Modal Satellite Imagery and Climate Signals

DGX agent

arXiv:2607.04117v1 Announce Type: cross Abstract: ERA5 seasonal climate variables contain predictive information about future glacier retreat beyond what satellite imagery alone provides, yet existing

researcharxiv-cs-cv
7 Jul 2026
Research

GlaKG: A Biomarker-Centric Fundus Knowledge Graph for Explainable Glaucoma Diagnosis and Risk Assessment

DGX agent

arXiv:2607.04673v1 Announce Type: new Abstract: Glaucoma is a leading cause of irreversible blindness worldwide, yet most automated diagnosis systems rely on opaque deep-learning models that offer lit

researcharxiv-cs-cv
7 Jul 2026
Local Ai

Global Logic and Local Search: Dual-Stream Multimodal In-Context Learning for Verifiable Industrial Anomaly Detection

DGX agent

arXiv:2607.03817v1 Announce Type: new Abstract: Large Multimodal Models (LMMs) show strong few-shot generalization, but industrial anomaly detection remains difficult because defects are small, input

local-aiarxiv-cs-cv
7 Jul 2026
Research

Global Pose Control for Generative View Synthesis in Normalized Object Coordinate Space

DGX agent

arXiv:2607.02712v1 Announce Type: new Abstract: Novel View Synthesis (NVS) enables the generation of unseen views of a scene from a single or multiple images, allowing users to freely explore an objec

researcharxiv-cs-cv
7 Jul 2026
Model Releases

GLOW-FDG: Generalized cancer LesiOn Whole-body segmentation model for ^{18}F-FDG-PET/CT

DGX agent

arXiv:2607.03931v1 Announce Type: cross Abstract: Whole-body fluorodeoxyglucose positron emission tomography combined with computed tomography is widely used in cancer care, but manual lesion delineat

model-releasesarxiv-cs-cv
7 Jul 2026
Tutorials

GMODiff: One-Step Gain Map Refinement with Diffusion Priors for HDR Reconstruction

DGX agent

arXiv:2512.16357v3 Announce Type: replace Abstract: Pre-trained Latent Diffusion Models (LDMs) have recently shown strong perceptual priors for low-level vision tasks, making them a promising directio

tutorialsarxiv-cs-cv
7 Jul 2026
Local Ai

GRCD: Grounded Region Change Detection for Multi-Finding Chest X-Ray Pairs

DGX agent

arXiv:2607.02719v1 Announce Type: new Abstract: Radiologists routinely compare current and prior chest X-rays to track disease progression, producing follow-up reports that describe multiple findings,

local-aiarxiv-cs-cv
7 Jul 2026
Research

Green for Go, Red for No: Visual Grounding via Semantic Segmentation for VLA Navigation Policies

DGX agent

arXiv:2607.05122v1 Announce Type: new Abstract: Vision-language-action (VLA) models enable robot navigation from natural language and visual goals, but remain susceptible to perceptual distractions an

researcharxiv-cs-cv
7 Jul 2026
Tutorials

GrowFields: Compositional 4D Neural Fields for Topology-Changing Plant Growth

DGX agent

arXiv:2607.03330v1 Announce Type: new Abstract: Quantifying plant growth dynamics from sparse longitudinal 3D observations is fundamental for agriculture and plant sciences. Yet, plants pose unique ch

tutorialsarxiv-cs-cv
7 Jul 2026
Model Releases

GuideMe: Multi-Domain Task Guidance and Intervention in Streaming Video

DGX agent

arXiv:2607.02991v1 Announce Type: new Abstract: While multimodal Large Language Models (MLLMs) excel at offline video understanding, an interesting question of how far they are from serving as a real-

model-releasesarxiv-cs-cv
7 Jul 2026
Research

GUSH3R: Everyone Everywhere All at Once as Gaussians

DGX agent

arXiv:2607.05243v1 Announce Type: new Abstract: Reconstructing dynamic human-scene environments from monocular videos is a challenging problem that requires jointly modeling scene geometry, camera mot

researcharxiv-cs-cv
7 Jul 2026
Safety

H-OPD: Confidence Aware Heterogeneous Multi-Teacher Multimodal On-policy Distillation

DGX agent

arXiv:2607.02592v1 Announce Type: new Abstract: On-policy distillation (OPD) has recently emerged as an effective post-training paradigm by providing supervision on student-generated trajectories. How

safetyarxiv-cs-cv
7 Jul 2026
← Previous
1…6263646566…261
Next →