AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,618 results
Applications

Multi-modality Image Fusion under Adverse Weather: Mask-Guided Feature Restoration and Interaction

DGX agent

arXiv:2606.26812v1 Announce Type: new Abstract: Multi-modality image fusion (MMIF) enhances scene representation by exploiting complementary cues from different modalities. Adverse weather, however, c

applicationsarxiv-cs-cv
26 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Tutorials

Neural Texture Compression using Hypernetworks

DGX agent

arXiv:2606.26913v1 Announce Type: cross Abstract: Recent work on neural texture compression has demonstrated that it is possible to learn small, per-material texture representations (composed of laten

tutorialsarxiv-cs-cv
26 Jun 2026
Tutorials

Neural Voxel Dynamics: Learning Implicit 3D Physics via Volumetric Feature Advection

DGX agent

arXiv:2606.26410v1 Announce Type: new Abstract: We present a self-supervised framework for learning implicit 3D physical dynamics directly from video-derived supervisory signals. While current generat

tutorialsarxiv-cs-cv
26 Jun 2026
Local Ai

Not All Actions Are Equal: Rethinking Conditioning for Dexterous World Model

DGX agent

arXiv:2606.27325v1 Announce Type: new Abstract: Recent advances in action-conditioned world models show promising progress in modeling complex interactions and forecasting future states under diverse

local-aiarxiv-cs-cv
26 Jun 2026
Hardware

OctoSense: Self-Supervised Learning for Multimodal Robot Perception

DGX agent

arXiv:2606.27317v1 Announce Type: new Abstract: We present OctoSense, an open-source sensor platform with stereo RGB and event cameras, LiDAR, a thermal camera, an inertial measurement unit, RTK-corre

hardwarearxiv-cs-cv
26 Jun 2026
Safety

Ordinal Neural Collapse as a Representation Prior for Visual Navigation

DGX agent

arXiv:2606.26839v1 Announce Type: cross Abstract: Learning robust navigation policies directly from visual observations remains a fundamental challenge in vision-based robotic navigation. In end-to-en

safetyarxiv-cs-cv
26 Jun 2026
Research

PanoImager: Geometry-Guided Novel View Synthesis and Reconstruction from Sparse Panoramic Views

DGX agent

arXiv:2606.27071v1 Announce Type: new Abstract: Panoramic sensing offers wide field-of-view coverage, yet 3D reconstruction from sparse panoramas remains challenging under rotation-dominant, weak-para

researcharxiv-cs-cv
26 Jun 2026
Safety

PathFLIP: Fine-grained Language-Image Pretraining for Versatile Computational Pathology

DGX agent

arXiv:2512.17621v2 Announce Type: replace Abstract: While Vision-Language Models (VLMs) have achieved notable progress in computational pathology (CPath), the gigapixel scale and spatial heterogeneity

safetyarxiv-cs-cv
26 Jun 2026
Safety

Paying More Attention to Visual Tokens in Self-Evolving Large Multimodal Models

DGX agent

arXiv:2606.27373v1 Announce Type: new Abstract: Recently, self-evolving large multimodal models (LMMs) have received attention for improving visual reasoning in a purely unsupervised setting. However,

safetyarxiv-cs-cv
26 Jun 2026
Model Releases

PhyEditBench: A Real-World Multi-Stage Benchmark for Physics-Aware Image Editing

DGX agent

arXiv:2606.26551v1 Announce Type: new Abstract: While instruction-based image editing, enabled by multi-modal generative models, has advanced significantly, existing benchmarks lack a comprehensive ev

model-releasesarxiv-cs-cv
26 Jun 2026
Safety

PhysEditWorld: A Large-Scale Dataset Toward Physics-Editable World Models

DGX agent

arXiv:2606.26694v1 Announce Type: new Abstract: Recent game world models can synthesize visually plausible, action-conditioned rollouts. However, their interaction behaviors often remain limited to ex

safetyarxiv-cs-cv
26 Jun 2026
Applications

PhysiFormer: Learning to Simulate Mechanics in World Space

DGX agent

arXiv:2606.27364v1 Announce Type: new Abstract: We present PhysiFormer, a diffusion transformer for physically-plausible 3D object motion. Unlike video world models that operate in view-dependent pixe

applicationsarxiv-cs-cv
26 Jun 2026
Research

PhysRAG: Enhancing Physics-Awareness in Video Generation via Retrieval-Augmented Generation

DGX agent

arXiv:2606.26916v1 Announce Type: new Abstract: Developing physically aware video generation models remains a significant challenge due to the difficulty in capturing diverse physical phenomena, such

researcharxiv-cs-cv
26 Jun 2026
Model Releases

PortraitGen: Exemplar-Driven GRPO with Dual-Reward Guidance for Photorealistic Portrait Generation

DGX agent

arXiv:2606.26930v1 Announce Type: new Abstract: Reinforcement Learning like Group Relative Policy Optimization (GRPO) has significantly advanced text-to-image post-training. However, current methods o

model-releasesarxiv-cs-cv
26 Jun 2026
Research

Position Rebinding Cache Reuse: Replay-Free Visual Revisiting for Interleaved Multimodal Reasoning

DGX agent

arXiv:2606.26631v1 Announce Type: new Abstract: Interleaved multimodal reasoning improves visual grounding by revisiting visual evidence during multi-step generation, yet existing methods typically re

researcharxiv-cs-cv
26 Jun 2026
Research

Predicting Fruit Quality with a Hybrid Machine Learning and Image Processing Approach

DGX agent

arXiv:2606.26165v1 Announce Type: new Abstract: Fruit spoilage is a significant issue in agriculture, leading to substantial economic losses. Addressing this, our study introduces a hybrid approach co

researcharxiv-cs-cv
26 Jun 2026
Safety

PressMimic: Pressure-Guided Motion Capture and Control for Humanoid Robot Imitation

DGX agent

arXiv:2606.26741v1 Announce Type: cross Abstract: Humanoid motion imitation requires not only accurate perception of human kinematics but also faithful reproduction of physical interactions with the e

safetyarxiv-cs-cv
26 Jun 2026
Agents

PrivacyBench: Privacy Isn't Free in Hybrid Privacy-Preserving Vision Systems

DGX agent

arXiv:2602.18900v2 Announce Type: replace-cross Abstract: Privacy preserving machine learning deployments in sensitive deep learning applications; from medical imaging to autonomous systems; increasin

agentsarxiv-cs-cv
26 Jun 2026
Research

Probabilistic NDVI Forecasting from Sparse Satellite Time Series and Weather Covariates

DGX agent

arXiv:2602.17683v3 Announce Type: replace-cross Abstract: Short-term forecasting of vegetation dynamics is a key enabler for data-driven decision support in precision agriculture. Normalized Differenc

researcharxiv-cs-cv
26 Jun 2026
Safety

Proposal-Conditioned Latent Diffusion for Closed-Loop Traffic Scenario Generation

DGX agent

arXiv:2606.27123v1 Announce Type: cross Abstract: Closed-loop traffic simulation remains challenging because it must generate interactive multi-agent behaviors that are scene-consistent and controllab

safetyarxiv-cs-cv
26 Jun 2026
Hardware

ProtoKV: Streaming Video Understanding under Delayed Query with Summary-State Memory

DGX agent

arXiv:2606.26762v1 Announce Type: new Abstract: Streaming video understanding (SVU) must answer queries that arrive asynchronously while visual tokens stream continuously under strict GPU-memory and q

hardwarearxiv-cs-cv
26 Jun 2026
Local Ai

Pseudo-Text-Conditioned 3D Grounding DINO for Organ Localization in Abdominal CT

DGX agent

arXiv:2606.27084v1 Announce Type: new Abstract: Reliable organ localization in abdominal CT can provide spatial priors for downstream trauma analysis. We propose CT-3GDINO, a lightweight 3D detector t

local-aiarxiv-cs-cv
26 Jun 2026
Model Releases

Qwen-Image-Agent: Bridging the Context Gap in Real-World Image Generation

DGX agent

arXiv:2606.26907v1 Announce Type: new Abstract: While text-to-image (T2I) models have achieved remarkable progress, they struggle with real-world requests that are often underspecified, implicit, or d

model-releasesarxiv-cs-cv
26 Jun 2026
Research

RayPE: Ray-Space Positional Encoding for 3D-Aware Video Generation

DGX agent

arXiv:2606.27345v1 Announce Type: new Abstract: Modern video diffusion transformers position their tokens through RoPE on the (u,v,t) axes -- a description of the camera's sampling grid that says noth

researcharxiv-cs-cv
26 Jun 2026
Research

Rendering Novel Views of MRI Using 3D Gaussian Splatting

DGX agent

arXiv:2606.26236v1 Announce Type: cross Abstract: The objective of this paper is to improve radiological gradings measured on MRIs of spines, by resampling scans so that the new view planes are better

researcharxiv-cs-cv
26 Jun 2026
Agents

Rethinking Training & Inference for Forecasting: Linking Winner-Take-All back to GMMs

DGX agent

arXiv:2606.26424v1 Announce Type: cross Abstract: Trajectory forecasting for autonomous driving has advanced rapidly, yet representative models often produce uninformative posteriors over forecast mod

agentsarxiv-cs-cv
26 Jun 2026
Research

Revealing Mammographic Phenotypes in Deep Learning Breast Cancer Risk Models

DGX agent

arXiv:2606.26431v1 Announce Type: cross Abstract: Mammogram-based deep learning models have improved breast cancer risk prediction, but the learned imaging patterns remain underexplored. Existing inte

researcharxiv-cs-cv
26 Jun 2026
Agents

RIS-Assisted Proactive Handover for Reliable mmWave Wireless Networks

DGX agent

arXiv:2606.26885v1 Announce Type: new Abstract: Millimeter-wave (mmWave) networks are highly susceptible to line-of-sight (LoS) blockages. Vision-aided wireless communications (VAWC) enable proactive

agentsarxiv-cs-cv
26 Jun 2026
Model Releases

Rolling Shutter Relative Pose Estimation Made Practical

DGX agent

arXiv:2606.26863v1 Announce Type: new Abstract: Rolling shutter (RS) cameras equip virtually all consumer devices, yet RS-aware relative pose estimation has remained impractical: the state-of-the-art

model-releasesarxiv-cs-cv
26 Jun 2026
Model Releases

RoPEMover: Depth-Aware Object Relocation via Positional Embeddings

DGX agent

arXiv:2606.27332v1 Announce Type: new Abstract: Moving an object in a single image requires geometry-consistent spatial rearrangement, including handling occlusions, revealing previously unseen region

model-releasesarxiv-cs-cv
26 Jun 2026
Research

SAM2Matting: Generalized Image and Video Matting

DGX agent

arXiv:2606.27339v1 Announce Type: new Abstract: Despite impressive advances in image matting, video matting remains challenging due to the inherent gap between high-level tracking, which requires fram

researcharxiv-cs-cv
26 Jun 2026
Tutorials

SatSplatDiff: Geometry-preserving generative refinement for high-fidelity satellite Gaussian Splatting

DGX agent

arXiv:2606.27223v1 Announce Type: new Abstract: Gaussian Splatting has been recently explored for satellite 3D reconstruction, demonstrating flexibility and efficiency in representing radiometrically

tutorialsarxiv-cs-cv
26 Jun 2026
Research

Sculpting NeRF Geometry: Human-Preference Fine-Tuning of a 3D-Aware Face GAN

DGX agent

arXiv:2606.27305v1 Announce Type: new Abstract: Reinforcement learning from human feedback (RLHF) for 3D generation is now established across a number of works, but most existing pipelines optimise ex

researcharxiv-cs-cv
26 Jun 2026
Model Releases

See & Sniff: Learning Visuo-Olfactory Representations

DGX agent

arXiv:2606.27307v1 Announce Type: new Abstract: While modern multimodal models integrate vision with language, audio, or touch, olfaction remains largely unexplored due to the lack of paired visuo-olf

model-releasesarxiv-cs-cv
26 Jun 2026
Research

Self-Supervised Tree-level Biomass Estimation in Urban Environments From Airborne LiDAR and Optical Observations

DGX agent

arXiv:2606.26194v1 Announce Type: new Abstract: Urban tree biomass remains less spatially explicitly quantified than biomass in managed forests because many estimates rely on inventories or coarse pro

researcharxiv-cs-cv
26 Jun 2026
Applications

SignSparK: Efficient Multilingual Sign Language Production via Sparse Keyframe Learning

DGX agent

arXiv:2603.10446v4 Announce Type: replace Abstract: Sign Language Production (SLP) faces a fundamental trade-off: direct text-to-pose models suffer from regression-to-the-mean effects, while dictionar

applicationsarxiv-cs-cv
26 Jun 2026
Safety

SpatialFlow-GRPO: Where Spatial Credit Drives Image Editing

DGX agent

arXiv:2606.26872v1 Announce Type: new Abstract: Recent online reinforcement learning has substantially improved image editing quality. However, existing Flow-GRPO-style methods usually rely on a singl

safetyarxiv-cs-cv
26 Jun 2026
Research

SubdivAR: Autoregressive Next-Scale Prediction for Neural Mesh Subdivision

DGX agent

arXiv:2606.27088v1 Announce Type: new Abstract: Mesh subdivision is a fundamental operation for converting coarse, editable meshes into high-resolution surfaces, with broad applications in digital ass

researcharxiv-cs-cv
26 Jun 2026
Research

Tailor Made Embeddings for Quantum Machine Learning

DGX agent

arXiv:2606.26312v1 Announce Type: cross Abstract: Autoencoders transformed classical machine learning by solving the curse of dimensionality, enabling principled weight initialization and learning com

researcharxiv-cs-cv
26 Jun 2026
Hardware

TaskNPoint: How to Teach Your Humanoid to Hit a Backhand in Minutes

DGX agent

arXiv:2606.26215v1 Announce Type: cross Abstract: How do we learn to hit a tennis backhand? Not from a thousand hours of tennis tournaments on TV - we work with a coach and practice. We argue this is

hardwarearxiv-cs-cv
26 Jun 2026
Research

TaskTok: Delving into Task Tokens for Task-driven Image Restoration

DGX agent

arXiv:2606.26615v1 Announce Type: new Abstract: While traditional image restoration focuses on perceptual quality, Task-Driven Image Restoration (TDIR) aims to maximize the performance of downstream h

researcharxiv-cs-cv
26 Jun 2026
Model Releases

Temporally Consistent Label Interpolation for Robust Surgical Multi-Task Learning under Challenging Conditions

DGX agent

arXiv:2606.26634v1 Announce Type: new Abstract: Effective multi-task learning for surgical scene understanding is fundamentally hindered by annotation granularity mismatch; temporal workflow tasks suc

model-releasesarxiv-cs-cv
26 Jun 2026
Model Releases

TF-TI2I: Training-Free Text-and-Image-to-Image Generation via Multi-Modal Implicit-Context Learning in Text-to-Image Models

DGX agent

arXiv:2503.15283v2 Announce Type: replace Abstract: Text-and-Image-To-Image (TI2I), an extension of Text-To-Image (T2I), integrates image inputs with textual instructions to enhance image generation.

model-releasesarxiv-cs-cv
26 Jun 2026
Model Releases

TMP: Tree-structured Mixed-policy Pruning for Large-scale Image Generation and Editing

DGX agent

arXiv:2606.27089v1 Announce Type: new Abstract: Modern image generation model rapidly grows their sizes to meet high-fidelity image synthesis. However, they gradually become unaffordable for their eno

model-releasesarxiv-cs-cv
26 Jun 2026
Research

Towards Consistent and Efficient Dataset Distillation via Diffusion-Driven Selection

DGX agent

arXiv:2412.09959v5 Announce Type: replace Abstract: Dataset distillation provides an effective approach to reduce memory and computational costs by optimizing a compact dataset that achieves performan

researcharxiv-cs-cv
26 Jun 2026
Model Releases

Towards Video Anomaly Detection from Event Streams: A Baseline and Benchmark Datasets

DGX agent

arXiv:2603.24991v2 Announce Type: replace Abstract: Event-based vision, characterized by low redundancy, focus on dynamic motion, and inherent privacy-preserving properties, naturally fits the demands

model-releasesarxiv-cs-cv
26 Jun 2026
Research

Tractography-Driven Synthetic Data Generation for Fiber Bundle Segmentation in Tracer Histology

DGX agent

arXiv:2606.26898v1 Announce Type: new Abstract: Diffusion MRI (dMRI) tractography enables non-invasive reconstruction of white-matter pathways, but its accuracy is fundamentally limited by indirect, l

researcharxiv-cs-cv
26 Jun 2026
Model Releases

TraMP-LLaMA: Generative Interpretability with Decoupled Instruction Tuning for Facial Expression Quality Assessment

DGX agent

arXiv:2606.26942v1 Announce Type: new Abstract: Existing facial expression quality assessment (FEQA) methods typically produce only a severity score, without explicitly communicating the observable fa

model-releasesarxiv-cs-cv
26 Jun 2026
← Previous
1…8788899091…263
Next →