AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,618 results
Model Releases

Enhancing Video Physical Consistency via Role-aware Joint Training and Modality-decoupled Denoising

DGX agent

arXiv:2607.04653v1 Announce Type: new Abstract: While modern video diffusion models excel in visual fidelity, maintaining long-range physical consistency remains a formidable challenge. Conventional p

model-releasesarxiv-cs-cv
7 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Entropy-Coded MS-VQ-VAE with Learned Priors for Ultra-Low Bitrate Video Compression

DGX agent

arXiv:2607.02562v1 Announce Type: new Abstract: Learned video codecs based on continuous latent representations struggle to operate reliably below 0.1 bits per pixel~(bpp): without a differentiable ra

researcharxiv-cs-cv
7 Jul 2026
Model Releases

Erasing Without Collateral Damage: Precise Concept Removal in Diffusion Models

DGX agent

arXiv:2607.05274v1 Announce Type: new Abstract: Training-free concept erasure is an attractive mechanism for controlling text-to-image diffusion models, but precise erasure often comes at the cost of

model-releasesarxiv-cs-cv
7 Jul 2026
Safety

EVA-Client: A Unified Data Collection, Inference, and Deployment Framework for Embodied Policies on Real Robots

DGX agent

arXiv:2607.02646v1 Announce Type: cross Abstract: We present EVA-Client, an open-source framework for deployment, data collection, and evaluation of trained manipulation policies on real robots. Sitti

safetyarxiv-cs-cv
7 Jul 2026
Model Releases

Evaluating Agentic Harness Systems for Autonomous Computational Pathology

DGX agent

arXiv:2607.02598v1 Announce Type: new Abstract: Autonomous computational pathology (ACP) converts high-level pathology analysis goals into executable, traceable and clinically bounded workflows. Reali

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

Evaluating Intellectual Property Guardrails of Generative Image Models: A Technical Report

DGX agent

arXiv:2607.02582v1 Announce Type: new Abstract: Generative image models are capable of producing images that bear a strong resemblance to, or replicate, recognizable intellectual property (IP). In thi

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration

DGX agent

arXiv:2607.04472v1 Announce Type: new Abstract: The rapid proliferation of artificial intelligence-generated content necessitates reliable multimodal forensics. Beyond video-level binary classificatio

model-releasesarxiv-cs-cv
7 Jul 2026
Research

Event Detection in Videos: A Framework for the Development of New Methods

DGX agent

arXiv:2607.04372v1 Announce Type: new Abstract: Event detection tasks in videos, the most important aspect of video surveillance, aim to detect events either at the pixel-level, frame-level, or clip-l

researcharxiv-cs-cv
7 Jul 2026
Model Releases

Explainable Flood Segmentation on Sentinel-1 SAR1 Imagery Using CNN and Transformer Architectures

DGX agent

arXiv:2606.16302v2 Announce Type: replace Abstract: Rapid and accurate flood prediction is essential for disaster response and mitigation planning. Synthetic Aperture Radar (SAR) sensors in satellites

model-releasesarxiv-cs-cv
7 Jul 2026
Applications

Exploring SAM Supervision for Fine-Grained UAV Target Segmentation under Data Scarcity

DGX agent

arXiv:2607.03754v1 Announce Type: new Abstract: Unmanned aerial vehicle (UAV) target segmentation remains challenging due to the small size of objects, appearance variations, cluttered backgrounds, an

applicationsarxiv-cs-cv
7 Jul 2026
Model Releases

ExpoMotion: A Large-Scale Benchmark and A Householder Projection Network for Multi-Exposure Fusion

DGX agent

arXiv:2607.03110v1 Announce Type: new Abstract: Multi-Exposure Fusion (MEF) effectively extends dynamic range, but practical deployment is hindered by motion-induced ghosting and the scarcity of high-

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

FairFlow: Demystifying and Mitigating Stereotype Bias in Text-to-Image Diffusion Transformers

DGX agent

arXiv:2607.03180v1 Announce Type: new Abstract: Multimodal diffusion transformers (MM-DiTs) have emerged as the prevalent backbone for modern text-to-image generation systems. However, they exhibit cr

model-releasesarxiv-cs-cv
7 Jul 2026
Agents

Fast 3D Foundation Model Initialized Gaussian Splatting

DGX agent

arXiv:2607.03209v1 Announce Type: new Abstract: This paper introduces a fast method for high-quality 3D Gaussian Splatting (3DGS) reconstruction without traditional Structure-from-Motion (SfM). The pr

agentsarxiv-cs-cv
7 Jul 2026
Local Ai

FDR-Occ: Factorized Dense Routing for Full-Spectrum 3D Occupancy Prediction

DGX agent

arXiv:2607.03822v1 Announce Type: new Abstract: Vision-based 3D occupancy prediction fundamentally relies on the 2D-to-3D view transformation. Current paradigms predominantly utilize explicit physical

local-aiarxiv-cs-cv
7 Jul 2026
Model Releases

FedProIn: Mitigating Client Drift for Learnable Prototypes in Federated Medical Imaging

DGX agent

arXiv:2607.04158v1 Announce Type: cross Abstract: Federated learning (FL) is severely hindered by statistical heterogeneity due to variations in scanners, acquisition protocols, and patient population

model-releasesarxiv-cs-cv
7 Jul 2026
Research

Fields of the Planet: Field Boundary Mapping Beyond 10m

DGX agent

arXiv:2607.04449v1 Announce Type: new Abstract: Field-boundary maps support crop monitoring, irrigation planning, and yield estimation, but many smallholder parcels span only a few 10 m Sentinel-2 pix

researcharxiv-cs-cv
7 Jul 2026
Research

Flash-BoN: Instant Drafts for Inference-Time Scaling in Diffusion Models

DGX agent

arXiv:2607.04461v1 Announce Type: new Abstract: Inference-time scaling for text-to-image generation has progressed from simple Best-of-N (BoN) sampling to guided search methods that verify and steer c

researcharxiv-cs-cv
7 Jul 2026
Safety

Flex-Forcing: Towards a Unified Autoregressive and Bidirectional Video Diffusion Model

DGX agent

arXiv:2607.03509v1 Announce Type: new Abstract: Recent progress in large-scale generative models has substantially advanced video generation, yet existing methods remain constrained by a rigid inferen

safetyarxiv-cs-cv
7 Jul 2026
Applications

FlowMark: Mask-Guided Video Watermarking

DGX agent

arXiv:2607.05261v1 Announce Type: new Abstract: We present FlowMark, a video watermarking framework guided by automatically predicted object masks. In contrast to prior region-based approaches that re

applicationsarxiv-cs-cv
7 Jul 2026
Model Releases

Fortifying Fully Convolutional Generative Adversarial Networks for Image Super-Resolution Using Divergence Measures

DGX agent

arXiv:2404.06294v2 Announce Type: replace-cross Abstract: Super-Resolution (SR) is a time-hallowed image processing problem that aims to improve the quality of a Low-Resolution (LR) sample up to the s

model-releasesarxiv-cs-cv
7 Jul 2026
Research

Fourier Splatting: Generalized Fourier encoded primitives for scalable radiance fields

DGX agent

arXiv:2603.19834v3 Announce Type: replace Abstract: Novel view synthesis has recently been revolutionized by 3D Gaussian Splatting (3DGS), which enables real-time rendering through explicit primitive

researcharxiv-cs-cv
7 Jul 2026
Model Releases

Framework and Multi-modal Dataset for Roadwork Zone Detection and Geo-localization

DGX agent

arXiv:2607.04330v1 Announce Type: new Abstract: Autonomous vehicles often rely on high-definition (HD) maps for navigation; however, these maps are not frequently updated and often lack semi-static in

model-releasesarxiv-cs-cv
7 Jul 2026
Safety

FRFDet: Efficient UAV Small Object Detection with Symmetric Sampling and Scalable Fusion

DGX agent

arXiv:2607.04125v1 Announce Type: new Abstract: Small object detection in Unmanned Aerial Vehicle (UAV) imagery remains challenging under adverse conditions, including complex weather, low illuminatio

safetyarxiv-cs-cv
7 Jul 2026
Model Releases

From General Actions to Domain-Specific Monitoring: Prior-Adaptive Transfer for Skeleton-Based Action Recognition

DGX agent

arXiv:2607.03327v1 Announce Type: new Abstract: Skeleton-based action recognition models have recently shown strong performance on large-scale benchmarks with general actions. However, directly transf

model-releasesarxiv-cs-cv
7 Jul 2026
Applications

From Geometric Labels to Semantic Understanding of Indoor Building Components Using Multimodal Large Language Models

DGX agent

arXiv:2607.03661v1 Announce Type: new Abstract: Point cloud-based understanding has become an important enabler for facility operation and maintenance involving indoor building components. However, ex

applicationsarxiv-cs-cv
7 Jul 2026
Safety

From Open Loop to Closed Loop: A Test-Time Iterative Optimization Framework for Reference-Consistent Image Generation

DGX agent

arXiv:2607.04691v1 Announce Type: new Abstract: While controllable image generation has made significant strides by incorporating visual reference conditions, existing methods predominantly operate as

safetyarxiv-cs-cv
7 Jul 2026
Safety

From Region Arrival to Instance-Level Grounding in Vision-and-Language Navigation

DGX agent

arXiv:2607.03792v1 Announce Type: cross Abstract: Vision-and-Language Navigation (VLN) agents may satisfy conventional success criteria while still failing to establish reliable object-level grounding

safetyarxiv-cs-cv
7 Jul 2026
Applications

FSDC-DETR: A Frequency-Spatial Domain Collaborative DETR for Small Object Detection

DGX agent

arXiv:2607.05176v1 Announce Type: new Abstract: Small object detection (SOD) remains a challenging task in real-world applications. Despite recent advances, existing detectors remain limited by rigid

applicationsarxiv-cs-cv
7 Jul 2026
Model Releases

Fully Rotation-Equivariant Spectral-Spatial Learning for Multispectral Object Detection

DGX agent

arXiv:2607.05148v1 Announce Type: new Abstract: Existing multispectral detectors are limited by discrete spectral processing, a scale-dependent shift in the relative reliability of spectral and spatia

model-releasesarxiv-cs-cv
7 Jul 2026
Local Ai

FunPhase: A Periodic Functional Autoencoder for Motion Generation via Phase Manifolds

DGX agent

arXiv:2512.09423v2 Announce Type: replace Abstract: Learning natural body motion remains challenging due to the strong coupling between spatial geometry and temporal dynamics. Embedding motion in phas

local-aiarxiv-cs-cv
7 Jul 2026
Local Ai

G^2TAM: Geometry Grounded Track Anything Model

DGX agent

arXiv:2607.03789v1 Announce Type: new Abstract: Human spatial understanding arises from jointly perceiving geometry and semantics, enabling consistent object identification and localization across vie

local-aiarxiv-cs-cv
7 Jul 2026
Model Releases

G3Splat: Geometrically Consistent Generalizable Gaussian Splatting

DGX agent

arXiv:2512.17547v2 Announce Type: replace Abstract: 3D Gaussians have become a powerful scene representation for real-time splatting and high-quality novel-view synthesis. This has motivated generaliz

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

GALOSH: Blind, Training-Free Denoising of Raw Bayer and sRGB Images by Parallel-Friendly Local Shrinkage

DGX agent

arXiv:2607.03768v1 Announce Type: cross Abstract: Classical training-free denoisers such as BM3D and non-local means owe much of their strength to search: content-dependent block matching whose memory

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

GaussianArt: Unified Modeling of Geometry and Motion for Articulated Objects

DGX agent

arXiv:2508.14891v3 Announce Type: replace Abstract: Reconstructing articulated objects is essential for building digital twins of interactive environments. However, prior methods typically decouple ge

model-releasesarxiv-cs-cv
7 Jul 2026
Research

Geographic Diversity Beats Data Volume for Cross-Domain Generalization in Zero-Label JEPA Driving World Models

DGX agent

arXiv:2607.04500v1 Announce Type: new Abstract: Self-supervised latent world models can assign a surprise score to driving scenarios without any human labels. A natural follow-up question is whether s

researcharxiv-cs-cv
7 Jul 2026
Model Releases

Geometric Observability Index: An Operator-Theoretic Framework for Per-Feature Sensitivity, Weak Observability, and Dynamic Effects in SE(3) Pose Estimation

DGX agent

arXiv:2602.05582v2 Announce Type: replace Abstract: We introduce the Geometric Observability Index (GOI), a per-feature sensitivity measure for pose estimation on SE(3). For a Gauss-Newton curvature m

model-releasesarxiv-cs-cv
7 Jul 2026
Research

Geometric Reciprocity: Unlocking Self-Supervision for Stereoscopic Video Generation

DGX agent

arXiv:2607.05354v1 Announce Type: new Abstract: Monocular-to-stereo conversion synthesizes stereoscopic content from 2D videos for immersive 3D experiences. In modern Depth-Image-Based Rendering (DIBR

researcharxiv-cs-cv
7 Jul 2026
Model Releases

Geometry-aware Depth-guided Representation Learning for Structure-preserving Low-light Image Enhancement

DGX agent

arXiv:2607.05005v1 Announce Type: new Abstract: Low-light degradation reduces image visibility and weakens structural cues that are important for visual representation and scene understanding. Existin

model-releasesarxiv-cs-cv
7 Jul 2026
Research

GeoSAM-Lite: A Lightweight Foundation Model for Onboard Remote Sensing Segmentation

DGX agent

arXiv:2607.03760v1 Announce Type: new Abstract: The deployment of large-scale foundation models like Segment Anything Model (SAM) on resource-constrained Earth observation platforms is hindered by pro

researcharxiv-cs-cv
7 Jul 2026
Applications

GeoWorld: Providing Full-frame Geometry Features to Facilitate 3D Scene Generation

DGX agent

arXiv:2511.23191v2 Announce Type: replace Abstract: Previous works that leverage video models for image-to-3D scene generation often suffer from geometric distortions and blurry content. Using video g

applicationsarxiv-cs-cv
7 Jul 2026
Research

GestaltMML: Enhancing Rare Genetic Disease Diagnosis through Multimodal Machine Learning Combining Facial Images and Clinical Text

DGX agent

arXiv:2312.15320v3 Announce Type: replace-cross Abstract: Individuals with suspected rare genetic disorders often undergo multiple clinical evaluations, imaging studies, laboratory tests, and genetic

researcharxiv-cs-cv
7 Jul 2026
Model Releases

Ghosts Beneath Textures: Texture-Relation Cues for Cross-Paradigm AI-Generated Image Detection

DGX agent

arXiv:2607.03862v1 Announce Type: new Abstract: AI-generated images have proliferated rapidly, motivating extensive research. Most existing AI-generated image detectors are developed and evaluated und

model-releasesarxiv-cs-cv
7 Jul 2026
Applications

GlaBoost: A Multimodal Structured Framework for Glaucoma Risk Stratification

DGX agent

arXiv:2508.03750v2 Announce Type: replace-cross Abstract: Early and accurate glaucoma detection is critical to prevent irreversible vision loss, yet existing AI methods often rely on unimodal inputs a

applicationsarxiv-cs-cv
7 Jul 2026
Research

GlacierCastAI: Predicting Glacier Retreat from Multi-Modal Satellite Imagery and Climate Signals

DGX agent

arXiv:2607.04117v1 Announce Type: cross Abstract: ERA5 seasonal climate variables contain predictive information about future glacier retreat beyond what satellite imagery alone provides, yet existing

researcharxiv-cs-cv
7 Jul 2026
Research

GlaKG: A Biomarker-Centric Fundus Knowledge Graph for Explainable Glaucoma Diagnosis and Risk Assessment

DGX agent

arXiv:2607.04673v1 Announce Type: new Abstract: Glaucoma is a leading cause of irreversible blindness worldwide, yet most automated diagnosis systems rely on opaque deep-learning models that offer lit

researcharxiv-cs-cv
7 Jul 2026
Local Ai

Global Logic and Local Search: Dual-Stream Multimodal In-Context Learning for Verifiable Industrial Anomaly Detection

DGX agent

arXiv:2607.03817v1 Announce Type: new Abstract: Large Multimodal Models (LMMs) show strong few-shot generalization, but industrial anomaly detection remains difficult because defects are small, input

local-aiarxiv-cs-cv
7 Jul 2026
Research

Global Pose Control for Generative View Synthesis in Normalized Object Coordinate Space

DGX agent

arXiv:2607.02712v1 Announce Type: new Abstract: Novel View Synthesis (NVS) enables the generation of unseen views of a scene from a single or multiple images, allowing users to freely explore an objec

researcharxiv-cs-cv
7 Jul 2026
Model Releases

GLOW-FDG: Generalized cancer LesiOn Whole-body segmentation model for ^{18}F-FDG-PET/CT

DGX agent

arXiv:2607.03931v1 Announce Type: cross Abstract: Whole-body fluorodeoxyglucose positron emission tomography combined with computed tomography is widely used in cancer care, but manual lesion delineat

model-releasesarxiv-cs-cv
7 Jul 2026
← Previous
1…6465666768…263
Next →