AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,618 results
Applications

Dual Quaternion SE(3) Synchronization with Recovery Guarantees

DGX agent

arXiv:2602.00324v2 Announce Type: replace-cross Abstract: Synchronization over the special Euclidean group SE(3) aims to recover absolute poses from noisy pairwise relative transformations and is a co

applicationsarxiv-cs-cv
29 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Dual-Stream Diffusion for World-Model Augmented Vision-Language-Action Model

DGX agent

arXiv:2510.27607v3 Announce Type: replace Abstract: Augmenting vision-language-action models (VLAs) with world models is promising for robotic policy learning but faces challenges in jointly predictin

safetyarxiv-cs-cv
29 May 2026
Research

DVSM: Decoder-only View Synthesis Model Done Right

DGX agent

arXiv:2605.29891v1 Announce Type: new Abstract: Recent Large View Synthesis Models (LVSMs) advocate an encoder-decoder architecture that separates reconstruction and rendering into distinct networks.

researcharxiv-cs-cv
29 May 2026
Hardware

EarlyTom: Early Token Compression Completes Fast Video Understanding

DGX agent

arXiv:2605.30010v1 Announce Type: new Abstract: Video large language models (Video-LLMs) have demonstrated strong capabilities in video understanding tasks. However, their practical deployment is stil

hardwarearxiv-cs-cv
29 May 2026
Model Releases

EarthShift: a benchmark for measuring robustness to real-world distribution shifts in Earth observation

DGX agent

arXiv:2605.29330v1 Announce Type: new Abstract: Current Earth observation benchmarks focus on measuring performance on diverse tasks and applications, typically measuring generalization in-distributio

model-releasesarxiv-cs-cv
29 May 2026
Research

Efficient, Validation-Free Intrinsic Quality Estimation for Large-Scale Face Recognition Datasets

DGX agent

arXiv:2605.29720v1 Announce Type: new Abstract: We propose Intrinsic Quality (IQ), a validation-free metric designed to estimate the inherent potential of face recognition (FR) datasets to produce hig

researcharxiv-cs-cv
29 May 2026
Model Releases

Embodied3DBench: Benchmarking Low-Level Embodied Spatial Intelligence of Vision Language Models

DGX agent

arXiv:2605.29074v1 Announce Type: new Abstract: Are current Vision Language Models (VLMs) ready to comprehend and reason about complex embodied interactions in 3D environments? We introduce Embodied3D

model-releasesarxiv-cs-cv
29 May 2026
Hardware

ESAM++: Efficient Online 3D Perception on the Edge

DGX agent

arXiv:2605.29505v1 Announce Type: new Abstract: Online 3D scene perception in real time is essential for robotics, AR/VR, and autonomous systems, particularly in edge computing scenarios where computa

hardwarearxiv-cs-cv
29 May 2026
Research

Eulerian Gaussian Splatting using Hashed Probability Pyramids

DGX agent

arXiv:2605.29136v1 Announce Type: new Abstract: We introduce a probabilistic splat-based radiance field framework that retains the fast rasterization and test-time efficiency of 3D Gaussian Splatting

researcharxiv-cs-cv
29 May 2026
Applications

Evaluation of Conversational Agents: Understanding Culture, Context and Environment in Emotion Detection

DGX agent

arXiv:2605.30099v1 Announce Type: new Abstract: Valuable decisions and highly prioritized analysis now depend on applications such as facial biometrics, social media photo tagging, and human robots in

applicationsarxiv-cs-cv
29 May 2026
Model Releases

EVL-ECG: Efficient ECG Interpretation With Multi-Aspect Heterogeneous Knowledge Distillation

DGX agent

arXiv:2605.29977v1 Announce Type: new Abstract: High-fidelity ECG interpretation is increasingly reliant on massive foundation models, yet their deployment in clinical edge-care remains hindered by ex

model-releasesarxiv-cs-cv
29 May 2026
Applications

F-RNG: Feed-Forward Relightable Neural Gaussians

DGX agent

arXiv:2605.25975v2 Announce Type: replace-cross Abstract: Capturing relightable 3D assets from real-world objects is a widely researched problem. Several per-scene optimization-based methods, based on

applicationsarxiv-cs-cv
29 May 2026
Model Releases

Fairness Beyond Demographics: Optimizing Performance Across Appearance-Based Hidden Cohorts in Medical Imaging

DGX agent

arXiv:2605.29827v1 Announce Type: new Abstract: Medical image analysis models can exhibit performance disparities across patient subgroups, threatening clinical safety and fairness. Existing methods t

model-releasesarxiv-cs-cv
29 May 2026
Safety

FakeVLM-R1: Internalizing Physical Laws via CoT for Synthetic Image Detection

DGX agent

arXiv:2605.30062v1 Announce Type: new Abstract: The development of generative artificial intelligence technologies has propelled the visual realism of synthetic images to an unprecedented level. Altho

safetyarxiv-cs-cv
29 May 2026
Local Ai

FedSmoothLoRA: Toward Smoother and Faster Convergence in Federated Low-Rank Adaptation

DGX agent

arXiv:2605.29460v1 Announce Type: new Abstract: Federated fine-tuning of foundation models with Low-Rank Adaptation (LoRA) provides an efficient solution for reducing communication and computation cos

local-aiarxiv-cs-cv
29 May 2026
Safety

Fewer Steps, Better Performance: Efficient Cross-Modal Clip Trimming for Video Moment Retrieval Using Language

DGX agent

arXiv:2605.29793v1 Announce Type: new Abstract: Given an untrimmed video and a sentence query, video moment retrieval using language (VMR) aims to locate a target query-relevant moment. Since the untr

safetyarxiv-cs-cv
29 May 2026
Safety

FlowSeg: Dynamic Semantic Guidance for LLM-Conditioned Segmentation

DGX agent

arXiv:2605.29461v1 Announce Type: new Abstract: LLM-conditioned segmentation has recently advanced rapidly by coupling large language models with iterative mask generation frameworks. However, we iden

safetyarxiv-cs-cv
29 May 2026
Research

FreeForm: Reduced-Order Deformable Simulation from Particle-Based Skinning Eigenmodes

DGX agent

arXiv:2605.29318v1 Announce Type: cross Abstract: We present a novel formulation for mesh-free, reduced-order simulation of deformable hyperelastic objects. Existing work in reduced-order elastodynami

researcharxiv-cs-cv
29 May 2026
Safety

From General Vision to Reliable Traversability Estimation: Adapting Vision Foundation Models for Unstructured Outdoor Environments

DGX agent

arXiv:2605.29565v1 Announce Type: new Abstract: Vision-based approaches have become the dominant paradigm for traversability estimation in unstructured outdoor environments, typically adapting vision

safetyarxiv-cs-cv
29 May 2026
Agents

FRUC: Feedforward Dynamic Scene Reconstruction from Uncalibrated Collaborative Driving Views

DGX agent

arXiv:2605.29997v1 Announce Type: new Abstract: We present FRUC, a feed-forward 3D Gaussian splatting framework for dynamic scene reconstruction from uncalibrated collaborative driving views. Existing

agentsarxiv-cs-cv
29 May 2026
Safety

Future Forcing: Future-aware Training-free KV Cache Policy for Autoregressive Video Generation

DGX agent

arXiv:2605.30083v1 Announce Type: new Abstract: Autoregressive (AR) video generation has emerged as a promising paradigm for long-horizon video synthesis, where each frame is generated conditioned on

safetyarxiv-cs-cv
29 May 2026
Applications

Gaga: Group Any Gaussians via 3D-aware Memory Bank

DGX agent

arXiv:2404.07977v4 Announce Type: replace Abstract: We introduce Gaga, a framework that reconstructs and segments open-world 3D scenes by leveraging inconsistent 2D masks predicted by zero-shot class-

applicationsarxiv-cs-cv
29 May 2026
Safety

GAP3D: Generative Alignment of VLM Latents to Patch-Level Embeddings for 3D Generation

DGX agent

arXiv:2605.28995v1 Announce Type: new Abstract: Recent approaches integrating vision-language models (VLMs) as prompt encoders for generative model conditioning typically rely on expensive end-to-end

safetyarxiv-cs-cv
29 May 2026
Safety

GASS: Geometry-Aware Spherical Sampling for Disentangled Diversity Enhancement in Text-to-Image Generation

DGX agent

arXiv:2602.17200v2 Announce Type: replace Abstract: Despite high semantic alignment, modern text-to-image (T2I) generative models still struggle to synthesize diverse images from a given prompt. In th

safetyarxiv-cs-cv
29 May 2026
Agents

GenClaw: Code-Driven Agentic Image Generation

DGX agent

arXiv:2605.30248v1 Announce Type: new Abstract: Image generation models have evolved from text-conditioned pixel synthesis toward multimodal agents endowed with visual comprehension and tool invocatio

agentsarxiv-cs-cv
29 May 2026
Model Releases

GenEraser: Generalizable Video Object Removal via Balanced Text-Mask Guidance and Decoupled Locator-Preserver

DGX agent

arXiv:2605.30045v1 Announce Type: new Abstract: Video object removal frequently struggles to simultaneously eliminate target objects and their associated physical effects (e.g., smoke, reflections, li

model-releasesarxiv-cs-cv
29 May 2026
Applications

GeoMag: Geometric-Aware Video Motion Magnification via State Space Model

DGX agent

arXiv:2605.29762v1 Announce Type: new Abstract: Video Motion Magnification (VMM) reveals imperceptible dynamics but often suffers from structural inconsistencies under complex geometric transformation

applicationsarxiv-cs-cv
29 May 2026
Safety

Geometry-Guided Modeling of Foundation Features Enables Generalizable Object Shape Deformation Learning

DGX agent

arXiv:2605.29661v1 Announce Type: new Abstract: Monocular 3D shape recovery is fundamental to geometric understanding, yet achieving robust generalization across arbitrary viewpoints and unseen object

safetyarxiv-cs-cv
29 May 2026
Tutorials

Geometry Matters: 3D Foundation Priors for Learning Semantic Correspondence

DGX agent

arXiv:2605.30093v1 Announce Type: new Abstract: Foundation features from self-supervised vision models and text-to-image diffusion models have proven effective for semantic correspondence estimation.

tutorialsarxiv-cs-cv
29 May 2026
Applications

GeRaF: Neural Geometry Reconstruction from Radio Frequency Signals

DGX agent

arXiv:2605.29097v1 Announce Type: new Abstract: GeRaF is the first method to use neural implicit learning for near-range 3D geometry reconstruction from radio frequency (RF) signals. Unlike RGB or LiD

applicationsarxiv-cs-cv
29 May 2026
Research

Getting to the Point: Pointing Improves LVLMs at Counting

DGX agent

arXiv:2603.21746v2 Announce Type: replace Abstract: Pointing-based methods decompose complex tasks as sequential grounding and reasoning steps. Given a query, the model first grounds the relevant obje

researcharxiv-cs-cv
29 May 2026
Applications

GMOS: Grounding Moving Object Segmentation in 3D Space and Time

DGX agent

arXiv:2605.30352v1 Announce Type: new Abstract: Moving Object Segmentation (MOS) aims to discover, segment, and track objects that move independently of the camera. Current MOS methods, however, exhib

applicationsarxiv-cs-cv
29 May 2026
Safety

Grounded 3D-Aware Spatial Vision-Language Modeling

DGX agent

arXiv:2605.30307v1 Announce Type: new Abstract: We present GR3D, a spatial vision language model equipped with three complementary grounding capabilities--explicit 2D grounding, implicit 2D grounding,

safetyarxiv-cs-cv
29 May 2026
Safety

Guidance Contrastive Token Credit Assignment for Discrete Policy Optimization

DGX agent

arXiv:2605.29198v1 Announce Type: new Abstract: Group-advantage-based reinforcement learning methods, such as GRPO and DAPO, have demonstrated strong performance across diverse domains, including math

safetyarxiv-cs-cv
29 May 2026
Research

HM-Talker: Hybrid Motion Modeling for High-Fidelity Talking Head Synthesis

DGX agent

arXiv:2508.10566v3 Announce Type: replace Abstract: Audio-driven talking head generation faces a fundamental trade-off between personalization and generalization, limiting its practical application. I

researcharxiv-cs-cv
29 May 2026
Agents

How to Relieve Distribution Shifts in Semantic Segmentation for Off-Road Environments

DGX agent

arXiv:2605.29599v1 Announce Type: cross Abstract: Semantic segmentation is crucial for autonomous navigation in off-road environments, enabling precise classification of surroundings to identify trave

agentsarxiv-cs-cv
29 May 2026
Model Releases

Improving Adversarial Robustness of Attribution via Implicit Regularization

DGX agent

arXiv:2605.29983v1 Announce Type: cross Abstract: The adversarial robustness of attributions is a fundamental requirement for reliable explainability in deep learning, yet existing approaches typicall

model-releasesarxiv-cs-cv
29 May 2026
Safety

Improving CLIP Adaptation by Breaking Tail Alignment for Source-Free Cross-Domain Few-Shot Learning

DGX agent

arXiv:2605.29776v1 Announce Type: new Abstract: Vision-Language Models (VLMs) such as CLIP demonstrate strong zero-shot generalization, but their performance significantly degrades in cross-domain sce

safetyarxiv-cs-cv
29 May 2026
Research

Inspectorch: Efficient rare event exploration in solar observations

DGX agent

arXiv:2602.20316v2 Announce Type: replace-cross Abstract: The Sun is observed in unprecedented detail, enabling studies of its activity on very small spatiotemporal scales. However, the large volume o

researcharxiv-cs-cv
29 May 2026
Research

IP-Adapter Is All You Need: Towards Fine-Tuning-Free Diffusion-Based Talking Face Generation

DGX agent

arXiv:2605.30230v1 Announce Type: new Abstract: With the rapid advancement of diffusion models, talking face generation has made remarkable progress. However, existing diffusion-based methods still re

researcharxiv-cs-cv
29 May 2026
Safety

KGEdit: Ambiguity-Aware Knowledge Graphs for Training-Free Precise Video Generation and Editing

DGX agent

arXiv:2605.29509v1 Announce Type: new Abstract: In recent years, training-free video generation has progressed remarkably. However, when handling complex textual instructions, existing methods still s

safetyarxiv-cs-cv
29 May 2026
Tutorials

Large Depth Completion Model from Sparse Observations

DGX agent

arXiv:2605.30115v1 Announce Type: new Abstract: This work presents the Large Depth Completion Model (LDCM), a simple, effective, and robust framework for single-view metric depth estimation with spars

tutorialsarxiv-cs-cv
29 May 2026
Model Releases

Learning Representations from 3D Gaussian Splats

DGX agent

arXiv:2605.29549v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) is a recent approach for scene rendering. Although primarily designed for view synthesis, its potential for scene understan

model-releasesarxiv-cs-cv
29 May 2026
Research

Lightweight Complementary-Cue Fusion for Robust Video Face Forgery Detection

DGX agent

arXiv:2605.29092v1 Announce Type: new Abstract: Current face video forgery detectors use wide or dual-stream backbones. We show that a single, lightweight fusion of two handcrafted cues can achieve hi

researcharxiv-cs-cv
29 May 2026
Model Releases

LiveSVG: Zero-Shot SVG Animation via Video Generation

DGX agent

arXiv:2605.30174v1 Announce Type: new Abstract: We introduce LiveSVG, a zero-shot approach for generating Scalable Vector Graphics (SVG) animations using video diffusion models. Current SVG animation

model-releasesarxiv-cs-cv
29 May 2026
Safety

Low-Magnification SEM May Suffice: Interpretable Deep Learning for Multi-Scale Fracture-Cause Classification in Zirconia-Toughened Alumina

DGX agent

arXiv:2605.29798v1 Announce Type: new Abstract: Reliable identification of fracture origins in alumina matrix composite hip and knee implants is critical for quality assurance and patient safety, yet

safetyarxiv-cs-cv
29 May 2026
Model Releases

LUMINA: A Multi-Vendor Mammography Benchmark with Energy Harmonization Protocol

DGX agent

arXiv:2603.14644v3 Announce Type: replace-cross Abstract: Publicly available full-field digital mammography (FFDM) datasets remain limited in size, clinical annotations, and vendor diversity, hinderin

model-releasesarxiv-cs-cv
29 May 2026
Research

MARTIAN: A Rendering Framework for Aerial Mars Imagery from HiRISE Orbital Data

DGX agent

arXiv:2605.29647v1 Announce Type: new Abstract: Aerial navigation on Mars requires vision-based pipelines that are robust to the diverse illumination conditions and terrain morphology of the Martian s

researcharxiv-cs-cv
29 May 2026
← Previous
1…135136137138139…263
Next →