AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlog
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,618 results
Research

Personalizing MLLMs via Reinforced Multimodal Reference Game

DGX agent

arXiv:2606.28845v1 Announce Type: new Abstract: Personalizing Multimodal Large Language Models (MLLMs) aims to recognize users' unique concepts from visual data and provide personalized responses. Alt

researcharxiv-cs-cv
30 Jun 2026
Model Releases
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

PGE-SAM: Prompt-Guided Feature Enhancement for Interactive Segmentation under Degradation

DGX agent

arXiv:2606.30477v1 Announce Type: new Abstract: Segment Anything Model (SAM) has revolutionized promptable image segmentation with strong zero-shot generalization. However, its performance degrades su

model-releasesarxiv-cs-cv
30 Jun 2026
Tutorials

Physics-Grounded Disentangled Flow Modeling for Brain Disease Progression Trajectory

DGX agent

arXiv:2606.28630v1 Announce Type: new Abstract: Forecasting longitudinal brain lesion evolution is critical for disease monitoring and treatment planning. Existing approaches typically learn a direct

tutorialsarxiv-cs-cv
30 Jun 2026
Agents

PLOT: Pseudo-Labeling via Object Tracking for Monocular 3D Object Detection

DGX agent

arXiv:2507.02393v2 Announce Type: replace Abstract: Monocular 3D object detection is crucial for scalable perception across fields like autonomous driving, robotics, and surveillance. However, progres

agentsarxiv-cs-cv
30 Jun 2026
Model Releases

Pointer-CAD v2: Plan-Then-Construct CAD Generation with Dimension-Aware Parametric Precision

DGX agent

arXiv:2606.29301v1 Announce Type: new Abstract: Computer-aided design (CAD) plays a fundamental role in modern manufacturing by providing the high precision required for industrial production. Recent

model-releasesarxiv-cs-cv
30 Jun 2026
Safety

PolarAPP: Beyond Polarization Demosaicking for Polarimetric Applications

DGX agent

arXiv:2603.23071v2 Announce Type: replace Abstract: Polarimetric imaging enables advanced vision applications such as normal estimation and de-reflection by capturing unique surface-material interacti

safetyarxiv-cs-cv
30 Jun 2026
Model Releases

PoseShield: Neural Collision Fields for Human Self-Collision Resolution

DGX agent

arXiv:2606.29686v1 Announce Type: new Abstract: Self-collision remains a persistent challenge in SMPL-based human pose estimation and motion generation. Under extreme articulations or stochastic motio

model-releasesarxiv-cs-cv
30 Jun 2026
Applications

Probing and Leveraging Video Diffusion Transformer Features for Robust Point Tracking

DGX agent

arXiv:2512.20606v2 Announce Type: replace Abstract: Despite achieving strong results on standard benchmarks, current point tracking methods rely on feature backbones that are rarely designed with the

applicationsarxiv-cs-cv
30 Jun 2026
Model Releases

Progressive Self-Supervised Learning with Individualized Community Assignment for Brain Network Analysis

DGX agent

arXiv:2606.29695v1 Announce Type: new Abstract: Brain networks exhibit a modular community structure that varies across individuals and neurological conditions. However, existing self-supervised learn

model-releasesarxiv-cs-cv
30 Jun 2026
Research

Projection-based coupling of infrared thermography and stereocorrelation-based digital image correlation

DGX agent

arXiv:2606.28905v1 Announce Type: new Abstract: Full-field measurement techniques such as digital image correlation and infrared thermography are prevalent in experimental solid mechanics. Digital ima

researcharxiv-cs-cv
30 Jun 2026
Research

PS-MOT: Cultivating Instance Awareness from Point Seeds for Multi-Object Tracking

DGX agent

arXiv:2606.30476v1 Announce Type: new Abstract: We introduce Point-supervised Multi-Object Tracking (PS-MOT) as a cost-effective alternative to traditional bounding box supervision, shifting the focus

researcharxiv-cs-cv
30 Jun 2026
Research

PSP: Harnessing Position and Shape Priors for Cross-Domain Few-Shot Medical Image Segmentation

DGX agent

arXiv:2606.28799v1 Announce Type: new Abstract: Few-Shot Medical Image Segmentation (FSMIS) offers a powerful solution to data scarcity but struggles to generalize across different imaging modalities.

researcharxiv-cs-cv
30 Jun 2026
Model Releases

Qwen-RobotNav Technical Report: A Scalable Navigation Model Designed for an Agentic Navigation System

DGX agent

arXiv:2606.18112v3 Announce Type: replace-cross Abstract: Agentic navigation systems require a base navigation model whose observation strategy can be externally reconfigured at inference time, becaus

model-releasesarxiv-cs-cv
30 Jun 2026
Research

RadarTwin: Scene-Specific mmWave Radar Simulation and Learning for Mobile Indoor Perception

DGX agent

arXiv:2606.28396v1 Announce Type: new Abstract: Millimeter-wave (mmWave) radar perception is limited by data scarcity: models trained on existing radar datasets fail to generalize to new objects, envi

researcharxiv-cs-cv
30 Jun 2026
Hardware

RAGA: Real Time Ray Traced Gaussian Shadow Casting for 3DGS Avatar-Scene Interaction

DGX agent

arXiv:2606.29329v1 Announce Type: new Abstract: We study the problem of physically plausible shadow casting when animating 3D Gaussian Splatting (3DGS) avatars, either individually or in multi-avatar

hardwarearxiv-cs-cv
30 Jun 2026
Local Ai

RainODE: Continuous-Time Precipitation Forecasting with Latent Neural ODEs

DGX agent

arXiv:2606.29855v1 Announce Type: new Abstract: In precipitation forecasting, not only accuracy but also temporal resolution is critical. However, increasing temporal resolution is constrained by obse

local-aiarxiv-cs-cv
30 Jun 2026
Research

RBE-Flow: Recurrent Bayesian Estimation on Feature Manifolds for Cross-Modal Registration

DGX agent

arXiv:2606.30492v1 Announce Type: new Abstract: Cross-modal image registration is essential for multi-sensor perception but remains fundamentally challenging due to severe non-linear radiometric discr

researcharxiv-cs-cv
30 Jun 2026
Applications

Real-time Rendering-based Surgical Instrument Tracking via Evolutionary Optimization

DGX agent

arXiv:2603.11404v3 Announce Type: replace-cross Abstract: Accurate and efficient tracking of surgical instruments is fundamental for Robot-Assisted Minimally Invasive Surgery. Although vision-based ro

applicationsarxiv-cs-cv
30 Jun 2026
Agents

Real-Time Underwater Image Enhancement via Frequency-Guided Dual-Path Attention

DGX agent

arXiv:2606.30314v1 Announce Type: new Abstract: Real-time underwater image enhancement (UIE) is crucial for mobile underwater photography and autonomous robotic systems, where practical deployment typ

agentsarxiv-cs-cv
30 Jun 2026
Research

Rectifying Mask via Entropy for Distractor-Free 3DGS in Ambiguous Scenarios

DGX agent

arXiv:2606.29496v1 Announce Type: new Abstract: We present RefineSplat, a systematic framework that effectively constructs transient masks to identify diverse ambiguous distractors. To do this, we qua

researcharxiv-cs-cv
30 Jun 2026
Model Releases

RefAlign: Representation Alignment for Reference-to-Video Generation

DGX agent

arXiv:2603.25743v2 Announce Type: replace Abstract: Reference-to-video (R2V) generation is a controllable video synthesis paradigm that constrains the generation process using both text prompts and re

model-releasesarxiv-cs-cv
30 Jun 2026
Applications

RefGlass-GS: A UAV-Enabled Fusion Framework for Photorealistic, Semantic and Interactive Digitization of Reflective Glass Facades via Gaussian Splatting

DGX agent

arXiv:2606.28826v1 Announce Type: new Abstract: Existing digitization of buildings with reflective glass facades suffers from geometric reconstruction distortion, unrealistic view-dependent texture re

applicationsarxiv-cs-cv
30 Jun 2026
Model Releases

Reliability-Prioritized Fine-Grained Generation in Multimodal Large

DGX agent

arXiv:2606.29573v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) are increasingly expected to generate fine-grained descriptions of visual content. However, we observe and theo

model-releasesarxiv-cs-cv
30 Jun 2026
Research

RenderFormer++: Scalable and Physically Grounded Feed-Forward Neural Rendering

DGX agent

arXiv:2606.30380v1 Announce Type: cross Abstract: We present RenderFormer++, a scalable and physically grounded feed-forward neural rendering framework for global illumination in mesh scenes. Existing

researcharxiv-cs-cv
30 Jun 2026
Safety

RePer-360: Releasing Perspective Priors for 360^irc Depth Estimation via Self-Modulation

DGX agent

arXiv:2603.05999v2 Announce Type: replace Abstract: Recent depth foundation models trained on perspective imagery achieve strong performance, yet generalize poorly to 360^irc images due to the substan

safetyarxiv-cs-cv
30 Jun 2026
Research

Rethinking Forgery Attacks on Semantic Watermarks in Black-Box Settings: A Geometric Distortion Perspective

DGX agent

arXiv:2606.29807v1 Announce Type: cross Abstract: Recent studies have shown that semantic watermarks, which embed information into the initial noise of latent diffusion models (LDMs), are vulnerable t

researcharxiv-cs-cv
30 Jun 2026
Research

Reweighting Framewise Attention in Video Transformers for Facial Expression Understanding

DGX agent

arXiv:2606.30611v1 Announce Type: new Abstract: Understanding facial expressions in videos requires modeling subtle and localized facial dynamics under unconstrained conditions. Although recent Vision

researcharxiv-cs-cv
30 Jun 2026
Safety

Rigel: Self-Distilled Score Adaptation for Image and Video Captioning Evaluation

DGX agent

arXiv:2606.29997v1 Announce Type: new Abstract: Automatic evaluation of image and video captioning is essential for benchmarking multimodal systems, although standard evaluation metrics show limited a

safetyarxiv-cs-cv
30 Jun 2026
Hardware

Robust and Efficient Monocular 3D Gaussian SLAM for Kilometer-Scale Outdoor Scenes

DGX agent

arXiv:2606.30436v1 Announce Type: new Abstract: Scaling monocular 3D Gaussian Splatting (3DGS) SLAM to kilometer-level outdoor environments poses two tightly coupled challenges: fragile long-term pose

hardwarearxiv-cs-cv
30 Jun 2026
Safety

Robust Trajectory Distillation: Hybrid Reweighting Meets Teacher-Inspired Targets

DGX agent

arXiv:2606.29837v1 Announce Type: new Abstract: Dataset distillation (DD) condenses large corpora into compact, information-rich subsets for efficient training and reuse. However, under noisy supervis

safetyarxiv-cs-cv
30 Jun 2026
Safety

Robust Zero-shot Anomaly Detection under Limited Auxiliary Anomaly Priors

DGX agent

arXiv:2606.29428v1 Announce Type: new Abstract: Zero-shot anomaly detection aims to identify defects in arbitrary novel domains; however, existing models assume that the auxiliary data contains a rich

safetyarxiv-cs-cv
30 Jun 2026
Research

Room Scene Discovery and Grouping in Unstructured Vacation Rental Image Collections

DGX agent

arXiv:2507.00263v2 Announce Type: replace Abstract: The rapid growth of vacation rental (VR) platforms has led to an increasing volume of property images, often uploaded without structured categorizat

researcharxiv-cs-cv
30 Jun 2026
Model Releases

S-Agent: Spatial Tool-Use Elicits Reasoning for Spatial Intelligence

DGX agent

arXiv:2606.20515v2 Announce Type: replace Abstract: Real-world spatial intelligence requires reasoning over a continuous and evolving 3D world, yet existing VLMs and tool-augmented agents largely rema

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

SA-Homo: Scale Adaptive Homography Estimation for Scale Variation Scenarios

DGX agent

arXiv:2606.30408v1 Announce Type: new Abstract: Homography estimation, as one of the fundamental problems in computer vision, remains challenged by scale variation scenarios where image pairs potentia

model-releasesarxiv-cs-cv
30 Jun 2026
Research

SA-VIS: Sparse frame Annotations for training Video Instance Segmentation

DGX agent

arXiv:2606.20140v2 Announce Type: replace Abstract: Recent online video instance segmentation (VIS) methods have achieved impressive results, thus becoming the preferred approach to segment instances

researcharxiv-cs-cv
30 Jun 2026
Research

SAD-GS: Learning Reliable 3D Semantic Gaussian Fields via Dynamic Geo-Semantic Anchoring

DGX agent

arXiv:2606.29376v1 Announce Type: new Abstract: Open-vocabulary 3D semantic Gaussian field learning relies on multi-view 2D supervision, whose semantic targets and spatial assignments are often unreli

researcharxiv-cs-cv
30 Jun 2026
Model Releases

SADL: What to Ignore? A Benchmark for Subject-Aware Distractor Localization

DGX agent

arXiv:2606.30393v1 Announce Type: new Abstract: Photographs frequently contain visual distractors besides foregrounds and backgrounds of the intended subject, competing for attention and weakening com

model-releasesarxiv-cs-cv
30 Jun 2026
Safety

SAFE-DiT: Semantics-Aware Fast-path Execution for High-Resolution Diffusion Transformers

DGX agent

arXiv:2606.29360v1 Announce Type: new Abstract: High-resolution Diffusion Transformer (DiT) inference contains substantial spatial redundancy, but many spatially adaptive implementations encode region

safetyarxiv-cs-cv
30 Jun 2026
Research

Same Concept, Different Directions: Cross-Modal Feature Heterogeneity in Sparse Autoencoders

DGX agent

arXiv:2606.29888v1 Announce Type: cross Abstract: Vision-language models map images and text into a joint embedding space. However, these embeddings often entangle multiple semantic features, which li

researcharxiv-cs-cv
30 Jun 2026
Applications

SATB-VR: Training Few-Step Video Restoration Diffusion Model using SNR-Aware Trajectory Blending

DGX agent

arXiv:2606.28677v1 Announce Type: new Abstract: While diffusion models excel in video restoration, their reliance on extensive iterative steps limits efficiency. Conversely, aggressive single-step dis

applicationsarxiv-cs-cv
30 Jun 2026
Model Releases

SatSplat: Geometrically-Accurate Gaussian Splatting for Satellite Imagery

DGX agent

arXiv:2606.28581v1 Announce Type: new Abstract: High-resolution satellite imagery demands 3D reconstruction methods that deliver both speed and geometric accuracy. Recent adaptations of 3D Gaussian Sp

model-releasesarxiv-cs-cv
30 Jun 2026
Research

ScaleAware-JEPA: Latent Representation for Discovery in Multiscale Physical Fields

DGX agent

arXiv:2606.29723v1 Announce Type: cross Abstract: Continuous physical fields represent a large fraction of data under scientific investigation. Their multiscale structures are central to discovery, ye

researcharxiv-cs-cv
30 Jun 2026
Tutorials

ScaleErasure: Inference-Time Minimal Intervention for Precise Concept Erasure in Next-Scale Autoregressive Image Generation

DGX agent

arXiv:2606.29282v1 Announce Type: new Abstract: Concept erasure aims to prevent image generative models from producing unsafe content while preserving their general generative capability. Meanwhile, n

tutorialsarxiv-cs-cv
30 Jun 2026
Agents

Scene-aware Prediction of Diverse Human Movement Goals

DGX agent

arXiv:2606.29942v1 Announce Type: new Abstract: Anticipation of human behaviours facilitates autonomous systems in proactive planning. Human behaviour could be stochastic due to varying goals. Human g

agentsarxiv-cs-cv
30 Jun 2026
Local Ai

Scenes as Objects, Not Primitives: Instance-Structured 3D Tokenization from Unposed Views

DGX agent

arXiv:2606.29513v1 Announce Type: new Abstract: A 3D scene is understood through its objects, not the primitives that compose them. Yet feed-forward reconstruction methods output dense, unstructured s

local-aiarxiv-cs-cv
30 Jun 2026
Applications

SciFlow: Semantic Cross Interference for Self-Supervised Optical Flow Domain Generalization

DGX agent

arXiv:2606.29004v1 Announce Type: new Abstract: Motions of objects and scenes carry essential intelligence in video understanding, offering rich cues for interpreting dynamic settings and interactions

applicationsarxiv-cs-cv
30 Jun 2026
Model Releases

SciIR: A Large-scale Training Dataset and Benchmark for Scientific Image Reasoning Generation

DGX agent

arXiv:2606.30124v1 Announce Type: new Abstract: While Text-to-Image (T2I) models have shown remarkable success in generating photorealistic visual content, they still struggle with the rigorous semant

model-releasesarxiv-cs-cv
30 Jun 2026
Research

SDGIC: A Semantic Disambiguation-Guided Generative Image Compression Method for Ultra-Low Bitrates

DGX agent

arXiv:2512.06344v2 Announce Type: replace Abstract: Generative image compression has recently shown impressive perceptual quality, but often suffers from semantic inconsistency at ultra-low bitrates (

researcharxiv-cs-cv
30 Jun 2026
← Previous
1…8182838485…263
Next →