AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,515 results
Model Releases

JoyAI-Video-Edit: Real-Time Open-Ended Video Editing with Autoregressive Diffusion

DGX agent

arXiv:2608.03974v1 Announce Type: new Abstract: Real-time video editing requires low-latency causal generation with bounded computational resources while preserving source fidelity and long-term tempo

model-releasesarxiv-cs-cv
5 Aug 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Local Ai

Keep the Needle, Prune the Haystack: Defect-Preserving Token Pruning for Efficient Zero-Shot Anomaly Detection

DGX agent

arXiv:2608.03681v1 Announce Type: new Abstract: Zero-shot visual anomaly detection has achieved remarkable progress, with recent vision-only approaches further improving performance while simplifying

local-aiarxiv-cs-cv
5 Aug 2026
Model Releases

Latent Reward Registers for Diffusion Preference Alignment

DGX agent

arXiv:2608.03929v1 Announce Type: cross Abstract: Aligning diffusion models with human preferences usually relies on a sparse terminal reward evaluated on the final generated samples, presenting a sev

model-releasesarxiv-cs-cv
5 Aug 2026
Model Releases

LDU-Bench: Multimodal LLM Evaluation for Lithography Defect Understanding under Layout-Varying Circuit Backgrounds

DGX agent

arXiv:2608.03078v1 Announce Type: new Abstract: Multimodal large language models have demonstrated strong defect recognition capability in industrial anomaly detection. However, in lithography review,

model-releasesarxiv-cs-cv
5 Aug 2026
Safety

Learning Attribute-aware Representations for Few-shot Scene Text Segmentation

DGX agent

arXiv:2504.11164v2 Announce Type: replace Abstract: Supervised scene text segmentation has achieved notable progress in recent years. However, its development is largely constrained by the scarcity of

safetyarxiv-cs-cv
5 Aug 2026
Research

Learning Biomechanically Plausible Human Motion from Sparse Radar Point Clouds

DGX agent

arXiv:2608.03637v1 Announce Type: new Abstract: Radar-based human pose estimation has focused on improving learning algorithms while representing the body as unconstrained keypoint coordinates. We add

researcharxiv-cs-cv
5 Aug 2026
Safety

Lightweight 3D Object Detection via Mamba-Based Knowledge Distillation

DGX agent

arXiv:2608.03490v1 Announce Type: cross Abstract: 3D object detection using light detection and ranging (LiDAR) sensors requires a balance between accuracy and computational efficiency for onboard per

safetyarxiv-cs-cv
5 Aug 2026
Research

LiteMVS: Efficient Multi-View Stereo with Foundation Distillation and Expert Aggregation

DGX agent

arXiv:2608.03851v1 Announce Type: new Abstract: Real-time 3D perception is crucial for robotics, augmented reality, and embodied intelligence applications. Existing multi-view stereo (MVS) methods pri

researcharxiv-cs-cv
5 Aug 2026
Model Releases

Localize, Don't Beautify: Client-Side Control of Image-Editing APIs for Cosmetic Surgery Previews

DGX agent

arXiv:2608.02841v1 Announce Type: new Abstract: Ask a commercial image editor to preview a cosmetic procedure and it will often change more of the face than the request names: a nose edit can also smo

model-releasesarxiv-cs-cv
5 Aug 2026
Model Releases

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images

DGX agent

arXiv:2608.03322v1 Announce Type: new Abstract: Medical visual grounding connects free-form clinical queries to spatial evidence in medical images and is an important component of interpretable medica

model-releasesarxiv-cs-cv
5 Aug 2026
Model Releases

Low-Dimensional High-Leverage Subspace Optimization: Beyond Full-Parameter Coupled Training for Neural Network Quantization

DGX agent

arXiv:2608.03919v1 Announce Type: new Abstract: Low-bit quantization suffers severe accuracy degradation on compact networks, rooted in the dominant full-parameter coupled training paradigm that ignor

model-releasesarxiv-cs-cv
5 Aug 2026
Model Releases

Material-Segmented Per-Pixel Emissivity Correction for Thermographic Anomaly Detection in Cultural Heritage Digital Twins

DGX agent

arXiv:2608.02964v1 Announce Type: new Abstract: Quantitative longwave thermography of heritage surfaces is limited by the global-constant emissivity assumption in inverse-Planck temperature retrieval;

model-releasesarxiv-cs-cv
5 Aug 2026
Applications

MaterialFusion: High-Quality, Zero-Shot, and Controllable Material Transfer with Diffusion Models

DGX agent

arXiv:2502.06606v3 Announce Type: replace Abstract: Manipulating the material appearance of objects in images is critical for applications like augmented reality, virtual prototyping, and digital cont

applicationsarxiv-cs-cv
5 Aug 2026
Safety

MeSS: City Mesh-Guided Outdoor Scene Generation with Cross-View Consistent Diffusion

DGX agent

arXiv:2508.15169v4 Announce Type: replace Abstract: Mesh models have become increasingly accessible for numerous cities; however, the lack of realistic textures restricts their application in virtual

safetyarxiv-cs-cv
5 Aug 2026
Applications

Micro-Segmentation Anomaly Detection in Zero-Trust Software-Defined Network Fabrics

DGX agent

arXiv:2608.02627v1 Announce Type: cross Abstract: Zero Trust Architecture (ZTA) principles need rigorous network segmentation and ongoing verification to reduce implicit trust and lateral threat propa

applicationsarxiv-cs-cv
5 Aug 2026
Model Releases

MinerU.Chem: A High-Precision System for Optical Chemical Structure and Reaction Recognition

DGX agent

arXiv:2608.03525v1 Announce Type: new Abstract: In organic chemistry papers and patents, molecular structures, reaction schemes, and experimental conditions are often presented as molecular structure

model-releasesarxiv-cs-cv
5 Aug 2026
Model Releases

Modeling Long-Term Memory and Temporal Attention Shifts for Video Salient Object Ranking with a New Benchmark

DGX agent

arXiv:2203.17257v2 Announce Type: replace Abstract: Salient Object Ranking (SOR) aims to estimate the relative saliency order among multiple salient objects. While SOR has been extensively studied in

model-releasesarxiv-cs-cv
5 Aug 2026
Applications

Modeling Scientific Experiment Scenes: Dataset and Model

DGX agent

arXiv:2608.02892v1 Announce Type: new Abstract: Scene Graph Generation (SGG) is fundamental to structured visual understanding, yet existing benchmarks focus mainly on daily life images and overlook s

applicationsarxiv-cs-cv
5 Aug 2026
Research

MoECa: Aligning Feature Reuse with Expert Decomposition in Diffusion Transformers

DGX agent

arXiv:2606.15615v2 Announce Type: replace-cross Abstract: Diffusion Transformers with Mixture-of-Experts (DiT-MoE) improve model capacity under sparse activation, but diffusion inference is still bott

researcharxiv-cs-cv
5 Aug 2026
Model Releases

Morphology-Aware Implicit Super-Resolution Network for Pathological Images

DGX agent

arXiv:2608.03664v1 Announce Type: new Abstract: Accurate diagnosis in Digital Pathology (DP) relies on high-resolution whole-slide images, yet clinical deployment is often limited by hardware costs. S

model-releasesarxiv-cs-cv
5 Aug 2026
Research

MSTAR: Multi-Scale Backbone Architecture Search for Timeseries Classification

DGX agent

arXiv:2402.13822v2 Announce Type: replace Abstract: Most of the previous approaches to Time Series Classification (TSC) highlight the significance of receptive fields and frequencies while overlooking

researcharxiv-cs-cv
5 Aug 2026
Model Releases

MT-Web2Code: Benchmarking Coding Agents on Multi-Turn Regional Reconstruction and Localized Modification

DGX agent

arXiv:2608.03474v1 Announce Type: new Abstract: Recent advances in Large Vision-Language Models (LVLMs) have demonstrated impressive capabilities in web UI generation. However, existing benchmarks pre

model-releasesarxiv-cs-cv
5 Aug 2026
Model Releases

MultiCompose: Multi-Concept Personalized Composition with Per-Subject Attribute Binding

DGX agent

arXiv:2608.03708v1 Announce Type: new Abstract: Text-to-image diffusion models enable personalization of specific visual concepts from a small number of reference images. However, generating a single

model-releasesarxiv-cs-cv
5 Aug 2026
Research

Multimodal Plant Root Phenotyping with Integration of 3D Skeleton Extraction and Language Analysis

DGX agent

arXiv:2608.03109v1 Announce Type: new Abstract: Plant root phenotyping is fundamental to understanding below-ground structures, optimizing crop management, and improving agricultural sustainability. T

researcharxiv-cs-cv
5 Aug 2026
Model Releases

MuRA: Multi-Rank Adaptation for Efficient and Effective Test-Time Vision-Language Generalization

DGX agent

arXiv:2608.03885v1 Announce Type: new Abstract: Vision-language models exhibit remarkable zero-shot capabilities but suffer significant performance degradation under distribution shifts. While test-ti

model-releasesarxiv-cs-cv
5 Aug 2026
Local Ai

NanoMorph-3D: An End-to-End Physics-Driven Unrolling Framework for Nanomaterial Reconstruction

DGX agent

arXiv:2608.03257v1 Announce Type: new Abstract: Precise 3D characterization of nanomaterials is essential for unlocking structure-property relationships. However, standard electron tomography is funda

local-aiarxiv-cs-cv
5 Aug 2026
Research

NCGR: Noise-Conditional Gated Rectification for Camera Extrinsic Perturbations in BEV 3D Object Detection

DGX agent

arXiv:2608.03895v1 Announce Type: new Abstract: Camera-based bird's-eye-view (BEV) 3D detection typically assumes accurate and fixed camera extrinsics. In detectors using spatial cross-attention (SCA)

researcharxiv-cs-cv
5 Aug 2026
Model Releases

NearID: Identity Representation Learning via Near-identity Distractors

DGX agent

arXiv:2604.01973v2 Announce Type: replace Abstract: When evaluating identity-focused tasks such as personalized generation and image editing, existing vision encoders entangle object identity with bac

model-releasesarxiv-cs-cv
5 Aug 2026
Research

Non-Destructive Quantification of Urea Adulteration in Bovine Milk Using Transmittance Multispectral Imaging

DGX agent

arXiv:2608.03113v1 Announce Type: new Abstract: Adulteration of bovine milk using urea remains a major food quality and health concern, motivating the development of rapid and quantitative screening t

researcharxiv-cs-cv
5 Aug 2026
Tutorials

Oh Deer, How Should I Handle This? Seasonal Priors for Selective Wildlife Annotation and Classification

DGX agent

arXiv:2608.02762v1 Announce Type: new Abstract: Fine-grained wildlife classification in aerial imagery is limited not only by model performance, but also by unreliable labels: animals occupy few pixel

tutorialsarxiv-cs-cv
5 Aug 2026
Research

OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models

DGX agent

arXiv:2608.03812v1 Announce Type: new Abstract: Omni-modal large language models (Omni-LLMs) have achieved remarkable performance on audio-visual understanding tasks, but processing long and highly re

researcharxiv-cs-cv
5 Aug 2026
Research

Open-Linguistic Concept Unified Learning for Cross-Site Interpretable Dermatology Image Diagnosis

DGX agent

arXiv:2608.03225v1 Announce Type: new Abstract: Human-interpretable computer-aided diagnosis is crucial for clinical decision making. Concept-based models excel by providing transparent reasoning and

researcharxiv-cs-cv
5 Aug 2026
Safety

Perceptual Anchoring: Prototype-Guided Text Calibration for Training-free Open-Vocabulary Semantic Segmentation

DGX agent

arXiv:2608.03991v1 Announce Type: new Abstract: Training-free open-vocabulary semantic segmentation (OVSS) partitions an image into semantically distinct regions based on arbitrary text descriptions,

safetyarxiv-cs-cv
5 Aug 2026
Research

PixelUp: Zero-Shot Semantic Feature Upsampling for Fine-Grained Vision Tasks

DGX agent

arXiv:2608.02792v1 Announce Type: new Abstract: Self-supervised Vision Foundation Models (VFMs) have become essential backbones for downstream tasks due to their strong and transferable visual represe

researcharxiv-cs-cv
5 Aug 2026
Applications

PLS-Calib: A Partial Least Squares Framework for Event Camera and Odometry Calibration under Ground Motion Constraints

DGX agent

arXiv:2608.03296v1 Announce Type: cross Abstract: Accurate extrinsic rotation calibration between sensors is fundamental to the performance of robotic perception systems. However, most existing calibr

applicationsarxiv-cs-cv
5 Aug 2026
Safety

Poisoning Prompt-Guided Sampling in Video Large Language Models

DGX agent

arXiv:2509.20851v2 Announce Type: replace Abstract: Video Large Language Models (VideoLLMs) are increasingly deployed as automated moderators on user-generated video platforms, where a few unwatched s

safetyarxiv-cs-cv
5 Aug 2026
Research

PolyLayout: Multi-room Manhattan Layout Estimation

DGX agent

arXiv:2608.03323v1 Announce Type: new Abstract: Estimating room layouts from multi-view imagery is a core task for indoor scene understanding. Existing methods are typically limited either by poor gen

researcharxiv-cs-cv
5 Aug 2026
Model Releases

Predictive Enhancement Calibration for Latent Breast MRI Virtual Contrast Enhancement

DGX agent

arXiv:2608.03612v1 Announce Type: cross Abstract: Virtual contrast enhancement (VCE) synthesizes enhanced breast MR images from pre-contrast acquisitions. Modern latent generators offer strong image p

model-releasesarxiv-cs-cv
5 Aug 2026
Tutorials

Progressive Learning of a Diffusion-based Inpainting Model for Separating Overlapped Fingerprints

DGX agent

arXiv:2608.03937v1 Announce Type: new Abstract: Overlapped friction ridge patterns are a recurring problem in latent fingerprints recovered from crime scenes and in live-scan scenarios where residual

tutorialsarxiv-cs-cv
5 Aug 2026
Model Releases

Qwen-3D: A Generalist 3D Vision-Language Model for Spatial Understanding

DGX agent

arXiv:2608.02980v1 Announce Type: new Abstract: Large Multimodal Models (LMMs) have achieved remarkable success on images and short videos, yet scaling them to long videos remains challenging due to f

model-releasesarxiv-cs-cv
5 Aug 2026
Safety

RealWeather: Realistic and Scene-Faithful Weather Translation with Driving World Models

DGX agent

arXiv:2608.02953v1 Announce Type: new Abstract: Realistic weather translation is valuable for developing and evaluating autonomous driving systems, yet collecting paired videos of the same scenes unde

safetyarxiv-cs-cv
5 Aug 2026
Agents

ReCamDriving: LiDAR-Free Camera-Controlled Video Synthesis for Novel Trajectories

DGX agent

arXiv:2512.03621v3 Announce Type: replace Abstract: Synthesizing multi-pass videos is important for autonomous driving. While current repair-based methods often struggle with out-of-distribution artif

agentsarxiv-cs-cv
5 Aug 2026
Research

Recurrent Contrastive Learning for Imbalanced Medical Image Classification

DGX agent

arXiv:2608.03304v1 Announce Type: new Abstract: Medical image classification often suffers from class imbalance due to the inherent disparities in disease incidence. Existing approaches, such as class

researcharxiv-cs-cv
5 Aug 2026
Agents

Residual Flow Matching with Dynamic Cross-Interaction for 3D Multi-Person Motion Prediction

DGX agent

arXiv:2608.03379v1 Announce Type: new Abstract: 3D multi-person motion prediction requires modeling both individual kinematics and inter-person interactions. While Flow Matching is effective for multi

agentsarxiv-cs-cv
5 Aug 2026
Safety

Rethinking Uncertainty Quantification and Entanglement in Image Segmentation

DGX agent

arXiv:2603.18792v2 Announce Type: replace Abstract: Uncertainty quantification (UQ) is crucial in safety-critical applications such as medical image segmentation. Total uncertainty is typically decomp

safetyarxiv-cs-cv
5 Aug 2026
Safety

RIDGE: Re-Noising with Internal Dynamic Guidance for Image Editing

DGX agent

arXiv:2608.03059v1 Announce Type: new Abstract: Inversion-free flow-based image editing avoids latent inversion, but still requires a target-side state at every editing step. The widely used equal-dis

safetyarxiv-cs-cv
5 Aug 2026
Research

S^3-Diff: Structural Semantic Synergy Diffusion Model for High Fidelity Super Resolution of Pathological Images

DGX agent

arXiv:2608.03540v1 Announce Type: new Abstract: Digital pathology relies on high-resolution whole slide images for accurate diagnosis, yet limitations in imaging devices, storage, and transmission oft

researcharxiv-cs-cv
5 Aug 2026
Applications

SAMSEM -- A Generic and Scalable Approach for IC Metal Line Segmentation

DGX agent

arXiv:2603.16548v2 Announce Type: replace-cross Abstract: In light of globalized hardware supply chains, the assurance of hardware components has gained significant interest, particularly in cryptogra

applicationsarxiv-cs-cv
5 Aug 2026
← Previous
1…1718192021…261
Next →