AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,515 results
Research

FAU at ImageCLEF 2026 Task on Multimodal Reasoning Robust Candidate Scoring and Concise Multilingual Visual Answering

DGX agent

arXiv:2608.01664v1 Announce Type: new Abstract: We present our ImageCLEF 2026 Multimodal Reasoning system for the Visual Multiple Choice Question Answering (Visual MCQ) and Visual Open Question Answer

researcharxiv-cs-cv
4 Aug 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Research

FDIR: Harmonizing Fidelity and Human-Machine Preference in Lossy Compression Image Restoration

DGX agent

arXiv:2608.00111v1 Announce Type: cross Abstract: Image restoration quality can be evaluated along three complementary facets: pixel-level fidelity, human perception, and downstream machine preference

researcharxiv-cs-cv
4 Aug 2026
Research

FeDepth: Federated Learning for Depth Estimation under Robot Heterogeneity

DGX agent

arXiv:2608.01129v1 Announce Type: cross Abstract: Although recent robot perception research emphasizes training on data from diverse environments to improve generalization, most existing methods still

researcharxiv-cs-cv
4 Aug 2026
Research

Fermat Active Laplace Learning for Semi-Supervised Hyperspectral Image Classification

DGX agent

arXiv:2608.02483v1 Announce Type: new Abstract: Two active learning algorithms for hyperspectral image (HSI) classification are proposed that combine density-aware Fermat distances with Poisson-reweig

researcharxiv-cs-cv
4 Aug 2026
Research

Few-Shot Concept Prompt Learning for Segmentation Foundation Models via Visual Grounding

DGX agent

arXiv:2608.01663v1 Announce Type: new Abstract: Promptable segmentation foundation models (FMs) such as SAM3 and Medical SAM3 promise few-shot, interactively-specified segmentation for medical imaging

researcharxiv-cs-cv
4 Aug 2026
Safety

FineMoLA: Towards Fine-Grained Motion-Language Alignment from Clip-Level Supervision

DGX agent

arXiv:2608.01392v1 Announce Type: new Abstract: Text-conditioned human motion generation has made rapid progress with the emergence of large-scale motion--language datasets. However, even datasets wit

safetyarxiv-cs-cv
4 Aug 2026
Research

Foveated Probes Recover Localized Binding Information in Vision Foundation Models

DGX agent

arXiv:2608.00726v1 Announce Type: new Abstract: Frozen vision foundation models are commonly evaluated through a single global image embedding, but this interface can conflate missing information with

researcharxiv-cs-cv
4 Aug 2026
Local Ai

FreqAnchorAD: Language-Free Zero-Shot Anomaly Detection via Frequency-Deviation Anchoring

DGX agent

arXiv:2608.00695v1 Announce Type: new Abstract: Zero-shot anomaly detection (ZSAD) aims to detect anomalies and localize defective regions in unseen target domains without target training data. Recent

local-aiarxiv-cs-cv
4 Aug 2026
Research

From Forest to Future Capital: Tracking Land Cover Change in Ibu Kota Nusantara (IKN) from 2021 to 2026 with PlanetScope Imagery

DGX agent

arXiv:2608.01230v1 Announce Type: new Abstract: Indonesia's relocation of its political and administrative capital from Jakarta to Ibu Kota Nusantara (IKN) has been framed around a ``Forest City'' vis

researcharxiv-cs-cv
4 Aug 2026
Tutorials

From Patches to Evidence Balls: Class-Conditioned Evidence Retrieval for Few-Shot Whole Slide Image Classification

DGX agent

arXiv:2608.01104v1 Announce Type: new Abstract: Whole slide image (WSI) classification is an evidence-driven task, where diagnostic cues are often sparse, spatially organized, and class-dependent. Exi

tutorialsarxiv-cs-cv
4 Aug 2026
Research

From Pixels to PCells: A Neurosymbolic Approach to Photonic Component Creation

DGX agent

arXiv:2608.00084v1 Announce Type: new Abstract: We present PixCell, a neurosymbolic system in which multimodal agents convert a visually presented photonic component into a parametric program over a s

researcharxiv-cs-cv
4 Aug 2026
Model Releases

From Representational Complementarity to Dual Systems: Synergizing VLM and Vision-Only Backbones for End-to-End Driving

DGX agent

arXiv:2602.10719v2 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) driving augments end-to-end (E2E) planning with language-enabled visual backbones, yet it remains unclear how vis

model-releasesarxiv-cs-cv
4 Aug 2026
Applications

Fruit-HSNet: A Machine Learning Approach for Hyperspectral Image-Based Fruit Ripeness Prediction

DGX agent

arXiv:2608.01202v1 Announce Type: new Abstract: Fruit ripeness prediction (FRP) is a classification-based agricultural computer vision task that has attracted much attention, thanks to its wide-rangin

applicationsarxiv-cs-cv
4 Aug 2026
Safety

FusionRS: A Large-Scale RGB-Infrared-Style Remote Sensing Dataset for Cross-Modal Vision-Language Learning

DGX agent

arXiv:2606.17020v2 Announce Type: replace Abstract: Remote sensing vision-language models have advanced Earth observation, but available large-scale vision-language resources remain RGB-centered, leav

safetyarxiv-cs-cv
4 Aug 2026
Research

G-Skin: Learning to Bind 3D Gaussians with Generative Visual Priors

DGX agent

arXiv:2608.01726v1 Announce Type: new Abstract: 3D Gaussian Splatting has achieved remarkable success in photorealistic and efficient rendering, leading to a rapid increase in 3D assets represented by

researcharxiv-cs-cv
4 Aug 2026
Applications

GaussianSelector: Lightweight Human-Guided Object Selection in 3D Gaussian Splatting with Graph Optimization

DGX agent

arXiv:2608.01492v1 Announce Type: new Abstract: Selecting a complete 3D object from a reconstructed scene with minimal user effort is essential for practical scene editing and embodied interaction. Ex

applicationsarxiv-cs-cv
4 Aug 2026
Research

Generated Images Are Easier to Forget: A Machine Unlearning Perspective for Synthetic Image Detection

DGX agent

arXiv:2608.00716v1 Announce Type: new Abstract: Robust detection of generated images is critical to counter the misuse of generative models. Existing methods primarily depend on learning from human-an

researcharxiv-cs-cv
4 Aug 2026
Model Releases

Generative AI and Foundation Models in Medical Image

DGX agent

arXiv:2608.01686v1 Announce Type: new Abstract: In recent years, generative AI has attracted significant public attention, and its use has been rapidly expanding across a wide range of domains. From c

model-releasesarxiv-cs-cv
4 Aug 2026
Tutorials

Generative Brownian Bridge Diffusion In Motion Space For Enhanced Myocardial Strain Analysis

DGX agent

arXiv:2608.01677v1 Announce Type: new Abstract: Myocardial strain analysis of cardiac magnetic resonance (CMR) images provides an important tool for evaluating cardiac function. However, current techn

tutorialsarxiv-cs-cv
4 Aug 2026
Research

GenPrior: Unleashing Text-to-Motion Generative Priors for Zero-Shot Skeleton-based Action Recognition

DGX agent

arXiv:2608.02236v1 Announce Type: new Abstract: Zero-shot skeleton-based action recognition (ZSAR) aims to recognize unseen action categories by aligning skeleton features with textual semantics. Howe

researcharxiv-cs-cv
4 Aug 2026
Safety

GenTrack: Physical Alignment for Robot-Native Motion Generation and Zero-Shot Humanoid Tracking

DGX agent

arXiv:2608.01410v1 Announce Type: cross Abstract: General-purpose humanoid trackers can execute diverse references, but their zero-shot coverage depends on large embodied corpora that are costly to ex

safetyarxiv-cs-cv
4 Aug 2026
Model Releases

GeoCore-9B: Towards Geo-Aware Generative Foundation Models in Earth Observation

DGX agent

arXiv:2608.01896v1 Announce Type: new Abstract: Existing generative models for earth observation (EO) predominantly rely on fine-tuning natural image priors, which limits their scalability and introdu

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

GEOID-Flood: A Large-Scale Multi-Modal Benchmark Dataset for Flood Segmentation

DGX agent

arXiv:2608.02315v1 Announce Type: new Abstract: Geospatial foundation models aim to learn representations that transfer across regions and sensors, yet evaluating them on specific tasks requires large

model-releasesarxiv-cs-cv
4 Aug 2026
Local Ai

Geometric-Topological Perception and Motion Prior for Real-Time Satellite Video Object Tracking

DGX agent

arXiv:2603.07564v2 Announce Type: replace Abstract: Satellite video object tracking (SVOT) remains fundamentally challenging due to texture scarcity, arbitrary rotation, aspect ratio changes, and seve

local-aiarxiv-cs-cv
4 Aug 2026
Research

Geometry-guided Emotion Modulation for Controllable and Photorealistic Emotional Talking Face Generation

DGX agent

arXiv:2608.00663v1 Announce Type: new Abstract: Audio-driven emotional talking face generation aims to synthesize realistic videos with expressive facial dynamics. However, existing methods struggle t

researcharxiv-cs-cv
4 Aug 2026
Model Releases

GIFT: Geometry-Invariant Fine-Tuning for Non-Lambertian Monocular Depth Estimation

DGX agent

arXiv:2608.02068v1 Announce Type: new Abstract: Monocular depth foundation models, benefiting from large-scale synthetic training data, have demonstrated strong generalization. However, they often hal

model-releasesarxiv-cs-cv
4 Aug 2026
Research

Gimbal360: Canonicalizing Planar Diffusion for Spherical Panorama Completion

DGX agent

arXiv:2603.23179v2 Announce Type: replace Abstract: Diffusion models provide powerful priors for 2D image completion, but these priors are learned on bounded planar images and do not transfer directly

researcharxiv-cs-cv
4 Aug 2026
Applications

Global-Scale Self-Supervised Spatiotemporal Learning for NDVI Time-Series Reconstruction

DGX agent

arXiv:2608.02322v1 Announce Type: new Abstract: Accurate and efficient reconstruction of cloud-contaminated and noise-corrupted NDVI time series remains a challenge in remote sensing. Deep learning pr

applicationsarxiv-cs-cv
4 Aug 2026
Research

GraRe: Grasp Candidate Re-Ranking for Frozen 6-DoF Grasp Detectors

DGX agent

arXiv:2608.00946v1 Announce Type: cross Abstract: Existing 6-DoF grasp detectors typically rank grasp candidates by detector confidence. However, our analysis on GraspNet-1Billion shows that detector

researcharxiv-cs-cv
4 Aug 2026
Research

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering

DGX agent

arXiv:2608.01660v1 Announce Type: new Abstract: Long-video question answering requires identifying sparse yet critical evidence from videos containing thousands of frames under a constrained visual-to

researcharxiv-cs-cv
4 Aug 2026
Model Releases

Grounding Agentic VLMs with Dedicated Segmentation for Fine-Grained Vehicle Damage Assessment

DGX agent

arXiv:2608.02470v1 Announce Type: new Abstract: Vision-language models (VLMs) are increasingly deployed as reasoning agents in real-world visual assessment pipelines, yet their spatial grounding remai

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

Grounding and Explaining Visual Evidence for AI-Generated Image Detection in Human-Centric Scenes

DGX agent

arXiv:2608.01988v1 Announce Type: new Abstract: Rapid advances in image generation models call for interpretable AI-generated image detection methods that not only determine authenticity but also prov

model-releasesarxiv-cs-cv
4 Aug 2026
Research

GROVE: Growing and Reasoning over Temporally Stratified Memory from Streaming Video Experience

DGX agent

arXiv:2608.02392v1 Announce Type: new Abstract: A wearable assistant should both answer questions about its visual history and recognize when that history is useful to the present situation. Existing

researcharxiv-cs-cv
4 Aug 2026
Agents

GSRAIN: Physically Calibrated High-/Low-Frequency Rainfall Synthesis for 3D Gaussian Driving Scenes

DGX agent

arXiv:2608.02177v1 Announce Type: new Abstract: Existing rainfall simulation methods for autonomous driving remain limited in physical controllability and multi-view consistency. This paper presents G

agentsarxiv-cs-cv
4 Aug 2026
Model Releases

GuideGround: VLM-guided Semantic Understanding and Viewpoint-aware Reasoning for 3D Visual Grounding

DGX agent

arXiv:2608.00518v1 Announce Type: new Abstract: 3D visual grounding aims to localize the target object in a 3D scene from a natural language query, requiring both fine-grained semantic understanding a

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

HarMoE: Multi-Source Chest Radiograph Pretraining with Dataset-Disentangled Experts

DGX agent

arXiv:2608.02252v1 Announce Type: new Abstract: Recent vision-language models for chest X-ray understanding are largely built on image-report alignment and therefore rely heavily on MIMIC-CXR as the d

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

Harnessing Adversarial Distillation to Customise Debiased, Disease-Specific Pathology Foundation Models for Breast Cancer

DGX agent

arXiv:2608.01356v1 Announce Type: new Abstract: Pathology foundation models (PFMs) provide strong tissue representations and have become central to digital pathology. However, deployment in disease-sp

model-releasesarxiv-cs-cv
4 Aug 2026
Safety

Hermite Curves as Trajectory Priors for Vision-Language-Action Models

DGX agent

arXiv:2608.01265v1 Announce Type: cross Abstract: Despite recent progress in Vision-Language-Action (VLA) models for robotic manipulation, the action chunk remains a weakly structured interface. Exist

safetyarxiv-cs-cv
4 Aug 2026
Research

Hi-TOPS: Hierarchical Topology-aware Scoring Prior for 3D Part Decomposition

DGX agent

arXiv:2608.00767v1 Announce Type: cross Abstract: Accurate 3D part decomposition requires separating shapes into structurally meaningful components with precise boundaries while preserving articulatio

researcharxiv-cs-cv
4 Aug 2026
Tutorials

HiResNets: Native Full-HD Video Recognition with Foveal Residual Streams

DGX agent

arXiv:2608.02140v1 Announce Type: new Abstract: Much of the recent progress in image and video recognition has come at the cost of memory: larger models, increased resolution, and longer temporal cont

tutorialsarxiv-cs-cv
4 Aug 2026
Model Releases

HorusEye: Language as Dynamic Attention for Emergency Visual Analysis

DGX agent

arXiv:2606.14741v2 Announce Type: replace Abstract: We introduce HorusEye, Language as Dynamic Attention for Emergency Visual Analysis. Our investigation followed five stages. The first one is benchma

model-releasesarxiv-cs-cv
4 Aug 2026
Agents

Human-like working memory signatures emerge from intrinsically plastic artificial neurons for robust dynamic vision

DGX agent

arXiv:2512.15829v4 Announce Type: replace-cross Abstract: While the unsustainable energy cost of artificial intelligence necessitates physics-driven computing, its performance superiority over full-pr

agentsarxiv-cs-cv
4 Aug 2026
Safety

Hybrid-Domain Posterior Sampling for Inverse Problems via Latent Flow Matching

DGX agent

arXiv:2608.00537v1 Announce Type: new Abstract: Latent Flow Models have revolutionized compressed-space image synthesis, yet their application to high-fidelity inverse problems remains bottlenecked. I

safetyarxiv-cs-cv
4 Aug 2026
Model Releases

HyperGS: Fast and Generalizable Gaussian Video Representation

DGX agent

arXiv:2607.11500v2 Announce Type: replace Abstract: Gaussian Splatting has emerged as an effective representation for video, but existing methods rely on per-video optimization. This leads to slow enc

model-releasesarxiv-cs-cv
4 Aug 2026
Research

IDraw: Artist Verification from Digital Drawing Images

DGX agent

arXiv:2608.01737v1 Announce Type: new Abstract: As digital drawings are increasingly shared online, reliable authorship verification has become important for protecting artists and resolving disputes.

researcharxiv-cs-cv
4 Aug 2026
Model Releases

Image-Space Rule Discovery

DGX agent

arXiv:2608.00490v1 Announce Type: new Abstract: Can image-editing models discover visual rules in image space and complete problem-solving end-to-end? We tackle this question in the spirit of a human

model-releasesarxiv-cs-cv
4 Aug 2026
Tutorials

iMontage: Unified, Versatile, Highly Dynamic Many-to-many Image Generation

DGX agent

arXiv:2511.20635v3 Announce Type: replace Abstract: Pre-trained video models learn powerful priors for generating high-quality, temporally coherent content. While these models excel at temporal cohere

tutorialsarxiv-cs-cv
4 Aug 2026
Applications

Implicit Neural Representations for Multimodal Longitudinal Image Imputation and Interpolation

DGX agent

arXiv:2608.02324v1 Announce Type: new Abstract: Longitudinal multiparametric MRI is central to follow-up imaging in oncology, yet real-world clinical data are characterised by missing sequences, heter

applicationsarxiv-cs-cv
4 Aug 2026
← Previous
1…2122232425…261
Next →