AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlog
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,414 results
Research

Overcoming Data Scarcity and Confidentiality in Hardware Assurance via Synthetic Generation

DGX agent

arXiv:2608.09914v1 Announce Type: cross Abstract: Hardware assurance relies on scanning electron microscopy (SEM) to verify nanoscale structures, but assembling the large, high-quality datasets requir

researcharxiv-cs-cv
11 Aug 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Research

PARAGraph: Pathology-Anatomy-Aware Hierarchical Graph for Diabetic Retinopathy Grading

DGX agent

arXiv:2608.08368v1 Announce Type: new Abstract: Diabetic retinopathy (DR) remains a leading cause of vision loss among working-age adults worldwide, making reliable severity grading clinically importa

researcharxiv-cs-cv
11 Aug 2026
Research

Parcel2Progression: An Anatomy-aware Longitudinal Framework for Alzheimer's Disease Diagnosis

DGX agent

arXiv:2608.08753v1 Announce Type: new Abstract: Alzheimer's disease (AD) progression is a longitudinal process with subtle pathological cues in the early stages. Yet, computational constraints have li

researcharxiv-cs-cv
11 Aug 2026
Research

PatchHead: Learning Spatial Patch Evidence for Generalizable AI-Generated Image Detection

DGX agent

arXiv:2608.09223v1 Announce Type: new Abstract: AI-generated image detectors generalize poorly when their training and test images originate from different generators or datasets. Despite the rich spa

researcharxiv-cs-cv
11 Aug 2026
Research

PE-Mamba: Bidirectional Selective Layer Aggregation for AI-Generated Image Detection

DGX agent

arXiv:2608.07999v1 Announce Type: new Abstract: AI-generated image (AIGI) detection has become increasingly challenging due to the rapid advancement of generative models and the diminishing gap betwee

researcharxiv-cs-cv
11 Aug 2026
Safety

Perception Before Supervision: Self-Contained Visual Distillation from Counterfactual Blind Spots

DGX agent

arXiv:2608.09931v1 Announce Type: new Abstract: Self-improvement for multimodal large language models (MLLMs) is typically driven by reward-based methods that provide only coarse scalar feedback. Dist

safetyarxiv-cs-cv
11 Aug 2026
Research

PhysX-CoT: Structured Physical Reasoning from a Single Image to Simulation-Ready 3D Assets

DGX agent

arXiv:2608.08053v1 Announce Type: cross Abstract: Simulation-ready 3D assets are central to robotics and embodied AI. Generating them from a single image is usually framed as a vision-language model t

researcharxiv-cs-cv
11 Aug 2026
Local Ai

PosBridge: Multi-View Positional Embedding Transplant for Identity-Aware Image Editing

DGX agent

arXiv:2508.17302v2 Announce Type: replace Abstract: Localized subject-driven image editing aims to seamlessly integrate user-specified objects into target scenes. As generative models continue to scal

local-aiarxiv-cs-cv
11 Aug 2026
Research

Predict to Skip: Linear Multistep Feature Forecasting for Efficient Diffusion Transformers

DGX agent

arXiv:2602.18093v2 Announce Type: replace Abstract: Diffusion Transformers (DiT) have emerged as a widely adopted backbone for high-fidelity image and video generation, yet their iterative denoising p

researcharxiv-cs-cv
11 Aug 2026
Research

Predictive Failure Detection in Network Hardware Using Thermal Imaging and Deep Learning with Sensor Fusion

DGX agent

arXiv:2608.07582v1 Announce Type: new Abstract: Unplanned network hardware malfunctions can interrupt services and result in expensive downtime in data centers. A deep learning-based predictive mainte

researcharxiv-cs-cv
11 Aug 2026
Applications

Preserve More Details: Mitigating Content Drift in Real-World Image Super-Resolution

DGX agent

arXiv:2608.09373v1 Announce Type: new Abstract: Real-world image super-resolution (Real-ISR) aims to reconstruct high-quality (HQ) images from low-quality (LQ) inputs subject to diverse real-world deg

applicationsarxiv-cs-cv
11 Aug 2026
Research

PressureMesh: 3D Human Mesh Estimation from Multi-Device Pressure Images

DGX agent

arXiv:2608.09550v1 Announce Type: new Abstract: Human pose monitoring is crucial in fields such as rehabilitation assessment and human-computer interaction. Due to its privacy-preserving nature, press

researcharxiv-cs-cv
11 Aug 2026
Applications

Progressive Learned Image Compression for Machine Perception

DGX agent

arXiv:2512.20070v2 Announce Type: replace Abstract: Recent advances in learned image codecs have extended from human perception toward machine perception However, progressive image compression with fi

applicationsarxiv-cs-cv
11 Aug 2026
Model Releases

RAGMesh with FaME-G2E: Long-Form Text-Driven 3D Face Generation and Editing

DGX agent

arXiv:2608.09186v1 Announce Type: new Abstract: Text-driven 3D face generation and editing remains challenging due to the difficulty of translating long-form descriptions into fine-grained facial geom

model-releasesarxiv-cs-cv
11 Aug 2026
Local Ai

RayLift: Lifting Complementary Ray-Wise Evidence with 3D Geometry Priors for Semantic Scene Completion

DGX agent

arXiv:2608.08476v1 Announce Type: new Abstract: Camera-based 3D semantic scene completion (SSC) provides comprehensive scene understanding for autonomous driving and robotics. However, existing method

local-aiarxiv-cs-cv
11 Aug 2026
Model Releases

Real Data Closes Synthetic-to-Real Gap in Optical Chemical Structure Recognition

DGX agent

arXiv:2608.09100v1 Announce Type: cross Abstract: Millions of chemical structures appear in patents and papers only as drawings, and using that information at scale requires reading the drawings. OCSR

model-releasesarxiv-cs-cv
11 Aug 2026
Hardware

Real-time physics inversion for retrieval of sub-pixel wildfire temperatures from VSWIR imaging spectroscopy

DGX agent

arXiv:2608.07580v1 Announce Type: new Abstract: In this work, we present a wildfire temperature retrieval framework for VSWIR imaging spectroscopy data, employed on data from NASA's Airborne Visible I

hardwarearxiv-cs-cv
11 Aug 2026
Model Releases

RealDenseFace: Real-time Monocular 3D Face Reconstruction from Dense UV-space Priors

DGX agent

arXiv:2608.09238v1 Announce Type: new Abstract: Recent monocular 3D face reconstruction methods achieve high fidelity by fitting a 3D Morphable Model (3DMM) to dense priors predicted by networks, but

model-releasesarxiv-cs-cv
11 Aug 2026
Local Ai

RefineAny3D: Depth Refinement as Semantic Alignment for Monocular 3D Detection

DGX agent

arXiv:2608.09147v1 Announce Type: new Abstract: Monocular 3D object detection spans two regimes: closed-set detectors operating within a fixed category vocabulary, and open-vocabulary detectors that l

local-aiarxiv-cs-cv
11 Aug 2026
Safety

Removing Infrastructure Barriers in Human-Robot Collaboration Through Wireless Reconfigurable Cells

DGX agent

arXiv:2608.09658v1 Announce Type: cross Abstract: Human-Robot Collaboration (HRC) plays a vital role in dynamic, high mix, low volume industrial scenarios such as remanufacturing, which frequently fac

safetyarxiv-cs-cv
11 Aug 2026
Model Releases

RenderMatte: Exact-Alpha Rendering and Group-Relative Alignment for Image Matting

DGX agent

arXiv:2608.08487v1 Announce Type: new Abstract: Image matting is an essential enabling technology for modern visual content production, where foreground extraction determines the realism and editabili

model-releasesarxiv-cs-cv
11 Aug 2026
Research

ResemBrick: Brick Reconstruction from Photographs with Perceptual Fidelity and Buildability

DGX agent

arXiv:2608.09597v1 Announce Type: new Abstract: Producing a hand-buildable, colored brick model of a 3D object from a few casual photographs is a clean testbed for a broader challenge: generating 3D c

researcharxiv-cs-cv
11 Aug 2026
Model Releases

Rethinking 3D Segmentation from Individual LiDAR Scans: Incidence-Aware Sampling on the SIP Benchmark

DGX agent

arXiv:2608.07757v1 Announce Type: new Abstract: 3D scene understanding is increasingly important in construction, yet most methods are developed on curated datasets that do not fully reflect real site

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

Rethinking Attention Locality in Spiking Transformers

DGX agent

arXiv:2608.08541v1 Announce Type: new Abstract: Spiking Transformers provide a promising paradigm for efficient visual processing with spike-driven computation, yet their Softmax-free Spiking Self-Att

model-releasesarxiv-cs-cv
11 Aug 2026
Safety

Retrieval-Augmented Generation-Based Color Restoration for Low-Light Image Enhancement

DGX agent

arXiv:2608.08211v1 Announce Type: cross Abstract: Recent low-light image enhancement (LLIE) methods have driven brightness and structural fidelity close to that of normally-exposed images, yet their o

safetyarxiv-cs-cv
11 Aug 2026
Research

Revisiting the Current Frame: Physical-Trace-Guided Network Output Correction for Video Restoration

DGX agent

arXiv:2608.09342v1 Announce Type: new Abstract: Video restoration methods exploit temporal information to recover information missing from degraded observations. However, reference frames within the s

researcharxiv-cs-cv
11 Aug 2026
Research

Right Answer, Wrong Heat: Explanation-Aware Evaluation and Thermal-Grounded Feedback for MLLMs on Infrared Images

DGX agent

arXiv:2608.09145v1 Announce Type: new Abstract: General-purpose multimodal large language models (MLLMs) are increasingly applied to infrared images, where they are commonly scored by answer accuracy

researcharxiv-cs-cv
11 Aug 2026
Research

RMR-Net: Degradation-Evidence-Guided Road-Image Restoration for Defect Detection

DGX agent

arXiv:2608.08957v1 Announce Type: new Abstract: Vehicle-mounted road cameras are vulnerable to motion blur, defocus, poor illumination, and noise, which can erase thin cracks and pothole boundaries ne

researcharxiv-cs-cv
11 Aug 2026
Safety

RobustDefect-LLM: Explainable and Robustness-Aware Industrial Surface Defect Classification with Decision Support and AI-Assisted Reporting

DGX agent

arXiv:2608.08589v1 Announce Type: new Abstract: This paper presents RobustDefect-LLM, an industrial surface-defect inspection framework integrating deep-learning classification, operator-facing visual

safetyarxiv-cs-cv
11 Aug 2026
Safety

RynnValue: Scaling Robotic Value Foundation Models with Temporal Distance

DGX agent

arXiv:2608.09853v1 Announce Type: cross Abstract: General-purpose reward models are increasingly the bottleneck for scaling robot learning, yet the recipe for learning value-related capabilities from

safetyarxiv-cs-cv
11 Aug 2026
Safety

SC-Diff: Semantically Calibrated Diffusion for Visible-to-Infrared Image Translation

DGX agent

arXiv:2608.08555v1 Announce Type: new Abstract: Visible-to-infrared image translation provides a practical way to expand infrared training data using abundant visible images. Diffusion models are prom

safetyarxiv-cs-cv
11 Aug 2026
Research

SC^{2}-WM: A Self-Correcting World Model with Closed-Loop Feedback for Vision-and-Language Navigation in Continuous Environments

DGX agent

arXiv:2608.07548v1 Announce Type: cross Abstract: Vision-and-Language Navigation in Continuous Environments (VLN-CE) requires agents to make fine-grained navigation decisions under partial observabili

researcharxiv-cs-cv
11 Aug 2026
Research

SCoPE: Training-Free Audio-Visual Event Perception via Sparse Cross-Modal Prior Exchange

DGX agent

arXiv:2608.07923v1 Announce Type: new Abstract: Audio-visual event perception (AVEP) determines which events occur in a video, when they occur, and whether they are audible, visible, or both. Training

researcharxiv-cs-cv
11 Aug 2026
Model Releases

SCTD 3.0: Sonar Common Target Detection in the Wild - A Large-Scale, Multi-Scene Dataset from Real Marine Surveys

DGX agent

arXiv:2608.08106v1 Announce Type: new Abstract: Synthetic Aperture Sonar (SAS) is core for wide-area detection of small underwater targets. However, large-scale, high-quality SAS datasets are scarce,

model-releasesarxiv-cs-cv
11 Aug 2026
Applications

Search over the Visual World: Persistent Visual Memory, Layered Indexes, and Source-Grounded Evidence

DGX agent

arXiv:2608.08075v1 Announce Type: cross Abstract: Most video-retrieval systems assume a bounded corpus and return ranked files or timestamps. Agents operating over cameras, screens, streams, and archi

applicationsarxiv-cs-cv
11 Aug 2026
Research

SegDem: Segmentation helps Demosaicing

DGX agent

arXiv:2608.07916v1 Announce Type: new Abstract: Image demosaicing reconstructs a full-color image from incomplete color measurements produced by a sensor covered with a color filter array (CFA). Most

researcharxiv-cs-cv
11 Aug 2026
Model Releases

Sekai2: From World Exploration to Interactive World Modeling

DGX agent

arXiv:2608.09449v1 Announce Type: new Abstract: Video world models must capture how scenes evolve over time and across viewpoints. Training them for long-horizon generation and camera control therefor

model-releasesarxiv-cs-cv
11 Aug 2026
Research

Semi-Dense Matching Uncertainty Is Not Just Local Confidence

DGX agent

arXiv:2608.08685v1 Announce Type: new Abstract: Reliable semi-dense matching is essential for modern geometric vision systems. Designed under a coarse-to-fine paradigm, it achieves an optimal balance

researcharxiv-cs-cv
11 Aug 2026
Model Releases

SeqLoc: Beyond the Single Frame for Cross-View Geo-Localization in Feature-Sparse Scenes

DGX agent

arXiv:2608.07835v1 Announce Type: new Abstract: Cross-View Geo-Localization (CVGL) with OpenStreetMap (OSM) performs well in structure-rich urban environments but collapses in feature-sparse scenes su

model-releasesarxiv-cs-cv
11 Aug 2026
Applications

SG-WAM: Text-Grounded and Spatial-aware Semantic Guidance for World-Action Models

DGX agent

arXiv:2608.08839v1 Announce Type: cross Abstract: World-Action Models (WAMs) have emerged as a promising paradigm for robotic manipulation. However, most existing WAMs generate future videos and actio

applicationsarxiv-cs-cv
11 Aug 2026
Model Releases

SI-Edit: Toward Sketch-Instruction Guided Local Image Editing with Pixel-Level Precision

DGX agent

arXiv:2608.09097v1 Announce Type: new Abstract: Despite rapid advances in generative models, achieving pixel-level precision in sketch-based image editing remains a persistent challenge, particularly

model-releasesarxiv-cs-cv
11 Aug 2026
Safety

SIP: Site in Pieces- A Dataset of Disaggregated Construction-Phase 3D Scans for Semantic Segmentation and Scene Understanding

DGX agent

arXiv:2512.09062v2 Announce Type: replace Abstract: Accurate 3D scene interpretation in active construction sites is essential for progress monitoring, safety assessment, and digital twin development.

safetyarxiv-cs-cv
11 Aug 2026
Local Ai

SLAP: Selective Local Vision-Language Alignment for Fish Re-Identification via Partial Optimal Transport

DGX agent

arXiv:2608.08840v1 Announce Type: new Abstract: Individual fish re-identification (ReID) is a fine-grained recognition problem in which identity-discriminative cues are often localized to specific bod

local-aiarxiv-cs-cv
11 Aug 2026
Research

Space-Creating versus Dead Possession: An Off-Ball Possession-Quality Index for Broadcast Football

DGX agent

arXiv:2608.09887v1 Announce Type: new Abstract: Ball possession is the most-cited and most-misleading number in football: 60% recycled in one's own half is not 60% spent pinning the opponent back. Exi

researcharxiv-cs-cv
11 Aug 2026
Applications

Sparse Attention to Emotion: Efficient Facial Emotion Recognition via Token Reduction

DGX agent

arXiv:2608.08873v1 Announce Type: new Abstract: Facial Emotion Recognition (FER) is an important task that has significant implications across various fields such as biometrics, health, and human-comp

applicationsarxiv-cs-cv
11 Aug 2026
Research

SplitGaussian: Reconstructing Dynamic Scenes via Visual Geometry Decomposition

DGX agent

arXiv:2508.04224v2 Announce Type: replace Abstract: Reconstructing dynamic 3D scenes from monocular video remains fundamentally challenging due to the need to jointly infer motion, structure, and appe

researcharxiv-cs-cv
11 Aug 2026
Local Ai

SportsGrounder: Proposal-Aided Interleaved Grounding for Dense Sports Video Reasoning

DGX agent

arXiv:2608.07932v1 Announce Type: new Abstract: Sports video analysis is crucial for athletic analytics and broadcasting enhancement. Dense sports video reasoning, however, demands a fine-grained unde

local-aiarxiv-cs-cv
11 Aug 2026
Local Ai

SRE-FER: Regional residual evidence learning for mitigating local evidence dilution in fine-grained facial expression recognition

DGX agent

arXiv:2608.08702v1 Announce Type: new Abstract: Fine-grained facial expression recognition (FER) hinges on capturing subtle muscular cues that distinguish adjacent emotions. Yet capturing these cues p

local-aiarxiv-cs-cv
11 Aug 2026
← Previous
1…56789…259
Next →