AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “synthesis”

GridTimelineEvolution
2,818 results
Safety

SIFT-VTON: Geometric Correspondence Supervision on Cross-Attention for Virtual Try-On

DGX agent

arXiv:2605.01296v1 Announce Type: new Abstract: Diffusion-based virtual try-on methods achieve photorealistic synthesis through cross-attention mechanisms that transfer garment features to target body

safetyarxiv-cs-cv
5 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

SwiftPie: Lightning-fast Subject-driven Image Personalization via One step Diffusion

DGX agent

arXiv:2605.01510v1 Announce Type: new Abstract: Diffusion models have achieved remarkable success in high-quality image synthesis, sparking interest in image-guided generation tasks such as subject-dr

safetyarxiv-cs-cv
5 May 2026
Research

Stepper: Stepwise Immersive Scene Generation with Multiview Panoramas

DGX agent

arXiv:2603.28980v2 Announce Type: replace Abstract: The synthesis of immersive 3D scenes from text is rapidly maturing, driven by novel video generative models and feed-forward 3D reconstruction, with

researcharxiv-cs-cv
4 May 2026
Applications

Sparse-View 3D Gaussian Splatting in the Wild

DGX agent

arXiv:2604.27422v1 Announce Type: new Abstract: We propose a 3D novel sparse-view synthesis framework for unconstrained real-world scenarios that contain distractors. Unlike existing methods that prim

applicationsarxiv-cs-cv
1 May 2026
Safety

The False Resonance: A Critical Examination of Emotion Embedding Similarity for Speech Generation Evaluation

DGX agent

arXiv:2604.26347v1 Announce Type: cross Abstract: Objective metrics for emotional expressiveness are vital for speech generation, particularly in expressive synthesis and voice conversion requiring em

safetyarxiv-cs-cl
30 Apr 2026
Agents

Efficient Multi-Agent System Training with Data Influence-Oriented Tree Search

DGX agent

arXiv:2502.00955v2 Announce Type: replace Abstract: Monte Carlo Tree Search (MCTS) based methods provide promising approaches for generating synthetic data to enhance the self-training of Large Langua

agentsarxiv-cs-cl
27 Apr 2026
Research

PAGaS: Pixel-Aligned 1DoF Gaussian Splatting for Depth Refinement

DGX agent

arXiv:2604.22129v1 Announce Type: new Abstract: Gaussian Splatting (GS) has emerged as an efficient approach for high-quality novel view synthesis. While early GS variants struggled to accurately mode

researcharxiv-cs-cv
27 Apr 2026
Agents

DeVI: Physics-based Dexterous Human-Object Interaction via Synthetic Video Imitation

DGX agent

arXiv:2604.20841v1 Announce Type: new Abstract: Recent advances in video generative models enable the synthesis of realistic human-object interaction videos across a wide range of scenarios and object

agentsarxiv-cs-cv
23 Apr 2026
Tutorials

Render-in-the-Loop: Vector Graphics Generation via Visual Self-Feedback

DGX agent

arXiv:2604.20730v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have shown promising capabilities in generating Scalable Vector Graphics (SVG) via direct code synthesis. Howev

tutorialsarxiv-cs-cv
23 Apr 2026
Research

EgoMotion: Hierarchical Reasoning and Diffusion for Egocentric Vision-Language Motion Generation

DGX agent

arXiv:2604.19105v1 Announce Type: new Abstract: Faithfully modeling human behavior in dynamic environments is a foundational challenge for embodied intelligence. While conditional motion synthesis has

researcharxiv-cs-cv
22 Apr 2026
Model Releases

STReasoner: Empowering LLMs for Spatio-Temporal Reasoning in Time Series via Spatial-Aware Reinforcement Learning

DGX agent

arXiv:2601.03248v2 Announce Type: replace Abstract: Spatio-temporal reasoning in time series involves the explicit synthesis of temporal dynamics, spatial dependencies, and textual context. This capab

model-releasesarxiv-cs-cl
22 Apr 2026
Safety

Reverse Constitutional AI: A Framework for Controllable Toxic Data Generation via Probability-Clamped RLAIF

DGX agent

arXiv:2604.17769v1 Announce Type: new Abstract: Ensuring the safety of large language models (LLMs) requires robust red teaming, yet the systematic synthesis of high-quality toxic data remains under-e

safetyarxiv-cs-cl
21 Apr 2026
Research

Speculative Decoding for Autoregressive Video Generation

DGX agent

arXiv:2604.17397v1 Announce Type: new Abstract: Autoregressive video diffusion is emerging as a promising paradigm for streaming video synthesis, with step distillation serving as the primary means of

researcharxiv-cs-cv
21 Apr 2026
Research

CLIMB: Controllable Longitudinal Brain Image Generation using Mamba-based Latent Diffusion Model and Gaussian-aligned Autoencoder

DGX agent

arXiv:2604.15611v1 Announce Type: cross Abstract: Latent diffusion models have emerged as powerful generative models in medical imaging, enabling the synthesis of high quality brain magnetic resonance

researcharxiv-cs-ai
20 Apr 2026
Applications

Efficient Video Diffusion Models: Advancements and Challenges

DGX agent

arXiv:2604.15911v1 Announce Type: new Abstract: Video diffusion models have rapidly become the dominant paradigm for high-fidelity generative video synthesis, but their practical deployment remains co

applicationsarxiv-cs-cv
20 Apr 2026
Research

From Synchrony to Sequence: Exo-to-Ego Generation via Interpolation

DGX agent

arXiv:2604.13793v1 Announce Type: new Abstract: Exo-to-Ego video generation aims to synthesize a first-person video from a synchronized third-person view and corresponding camera poses. While paired s

researcharxiv-cs-cv
16 Apr 2026
Research

A Survey on 3D Gaussian Splatting Applications: Segmentation, Editing, and Generation

DGX agent

arXiv:2508.09977v4 Announce Type: replace Abstract: In the context of novel view synthesis, 3D Gaussian Splatting (3DGS) has recently emerged as an efficient and competitive counterpart to Neural Radi

researcharxiv-cs-cv
14 Apr 2026
Agents

CountLoop: Training-Free High-Instance Image Generation via Iterative Agent Guidance

DGX agent

arXiv:2508.16644v4 Announce Type: replace Abstract: Diffusion models excel at photorealistic synthesis but struggle with precise object counts, especially in high-density settings. We introduce COUNTL

agentsarxiv-cs-cv
14 Apr 2026
Research

Dark-EvGS: Event Camera as an Eye for Radiance Field in the Dark

DGX agent

arXiv:2507.11931v2 Announce Type: replace Abstract: In low-light environments, conventional cameras often struggle to capture clear multi-view images of objects due to dynamic range limitations and mo

researcharxiv-cs-cv
14 Apr 2026
Research

FlowBind: Efficient Any-to-Any Generation with Bidirectional Flows

DGX agent

arXiv:2512.15420v2 Announce Type: replace Abstract: Any-to-any generation seeks to translate between arbitrary subsets of modalities, enabling flexible cross-modal synthesis. Despite recent success, e

researcharxiv-cs-lg
14 Apr 2026
Research

Progressively Texture-Aware Diffusion for Contrast-Enhanced Sparse-View CT

DGX agent

arXiv:2604.11559v1 Announce Type: new Abstract: Diffusion-based sparse-view CT (SVCT) imaging has achieved remarkable advancements in recent years, thanks to its more stable generative capability. How

researcharxiv-cs-cv
14 Apr 2026
Model Releases

EgoTL: Egocentric Think-Aloud Chains for Long-Horizon Tasks

DGX agent

arXiv:2604.09535v1 Announce Type: new Abstract: Large foundation models have made significant advances in embodied intelligence, enabling synthesis and reasoning over egocentric input for household ta

model-releasesarxiv-cs-cv
13 Apr 2026
Tutorials

BADiff: Bandwidth Adaptive Diffusion Model

DGX agent

arXiv:2510.21366v3 Announce Type: replace Abstract: In this work, we propose a novel framework to enable diffusion models to adapt their generation quality based on real-time network bandwidth constra

tutorialsarxiv-cs-cv
10 Apr 2026
Research

Controller Design for Structured State-space Models via Contraction Theory

DGX agent

arXiv:2604.07069v1 Announce Type: cross Abstract: This paper presents an indirect data-driven output feedback controller synthesis for nonlinear systems, leveraging Structured State-space Models (SSMs

researcharxiv-cs-lg
10 Apr 2026
Research

Neural Harmonic Textures for High-Quality Primitive Based Neural Reconstruction

DGX agent

arXiv:2604.01204v2 Announce Type: replace-cross Abstract: Primitive-based methods such as 3D Gaussian Splatting have recently become the state-of-the-art for novel-view synthesis and related reconstru

researcharxiv-cs-ai
10 Apr 2026
Model Releases

SceneScribe-1M: A Large-Scale Video Dataset with Comprehensive Geometric and Semantic Annotations

DGX agent

arXiv:2604.07990v1 Announce Type: new Abstract: The convergence of 3D geometric perception and video synthesis has created an unprecedented demand for large-scale video data that is rich in both seman

model-releasesarxiv-cs-cv
10 Apr 2026
Applications

CasDeblurGS: Cascaded 2D-to-3D Multi-View Consistency for 3D Gaussian Splatting from Two Blurry Images

DGX agent

arXiv:2608.10345v1 Announce Type: new Abstract: Free-viewpoint 3D scene media is increasingly important for immersive applications, yet practical capture often suffers from severe view sparsity and mo

applicationsarxiv-cs-cv
12 Aug 2026
Tutorials

Gaussian Sculpting: End-to-End Controllable Surface Reconstruction via Field Optimization

DGX agent

arXiv:2608.10602v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) has recently enabled real-time novel view synthesis with impressive quality. However, it struggles to recover accurate surf

tutorialsarxiv-cs-cv
12 Aug 2026
Research

Knowledge-Guided 3D CT Generation: A Conditioning-Centric Taxonomy

DGX agent

arXiv:2608.09992v1 Announce Type: cross Abstract: Controllable generation guided by external knowledge is a key requirement in modern generative deep learning applications, enabling the synthesis of s

researcharxiv-cs-ai
12 Aug 2026
Research

Order Matters: LVLMs as Judges for Temporal Reasoning in Image Sequences

DGX agent

arXiv:2608.10908v1 Announce Type: cross Abstract: As generative multimedia evolves from static image synthesis to complex, interleaved visual narratives, a foundational bottleneck has emerged: the jud

researcharxiv-cs-cl
12 Aug 2026
Research

FlexSplat: Flexible Feed-Forward 3D Gaussian Splatting without Point Cloud Correspondence

DGX agent

arXiv:2608.07937v1 Announce Type: new Abstract: We present FlexSplat, a feed-forward framework for novel view synthesis (NVS) from uncalibrated, object-centric multi-view image collections. A recent l

researcharxiv-cs-cv
11 Aug 2026
Model Releases

MADBench: A Benchmark for Modality-Aware Audio Deepfake Detection

DGX agent

arXiv:2608.09593v1 Announce Type: cross Abstract: Recent advances in speech synthesis and audio generation have made high-fidelity acoustic forgery low-cost and difficult to attribute, enabling a real

model-releasesarxiv-cs-ai
11 Aug 2026
Safety

GOPI: Generation-Oriented 3D Pose Inference for Furniture Insertion from Single-View RGB-D Indoor Scenes

DGX agent

arXiv:2608.06836v1 Announce Type: new Abstract: We study the problem of inserting new furniture into indoor scene images. Under masked single-view 2D image-plane conditioning, however, the physical sc

safetyarxiv-cs-cv
10 Aug 2026
Safety

Identity as Presence: Towards Appearance and Voice Personalized Joint Audio-Video Generation

DGX agent

arXiv:2603.17889v4 Announce Type: replace Abstract: Recent advances in video synthesis have enabled realistic integration of real individuals, driving demand for identity-aware generation. While emerg

safetyarxiv-cs-cv
10 Aug 2026
Applications

Improving Low-Resolution Face Recognition under Limited Data: How Synthetic Data Generation Can Close the Domain Gap

DGX agent

arXiv:2608.06580v1 Announce Type: new Abstract: Face Recognition (FR) systems in surveillance settings often encounter Low Resolution (LR) faces, those whose face region falls below the standard 112 i

applicationsarxiv-cs-cv
10 Aug 2026
Model Releases

MirrorWorld: Taming Video Diffusion Models for Mirror Reflection Generation

DGX agent

arXiv:2608.07463v1 Announce Type: new Abstract: Recent advances in video diffusion models (VDMs) have enabled high-fidelity video synthesis. However, generating mirror reflections remains challenging

model-releasesarxiv-cs-cv
10 Aug 2026
Model Releases

Multi-Agent Forensic Reasoning for Generalizable Deepfake Video Detection

DGX agent

arXiv:2608.06865v1 Announce Type: cross Abstract: The malicious use of generative artificial intelligence to create highly realistic deepfake videos raises serious ethical concerns and poses substanti

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Symbolic Graphics Programming with Large Language Models

DGX agent

arXiv:2509.05208v2 Announce Type: replace Abstract: Large language models (LLMs) excel at program synthesis, yet their ability to produce symbolic graphics programs (SGPs) that render into precise vis

model-releasesarxiv-cs-cv
10 Aug 2026
Safety

A Bridge from Audio to Video: Phoneme-Viseme Alignment Allows Every Face to Speak Multiple Languages

DGX agent

arXiv:2510.06612v2 Announce Type: replace Abstract: Speech-driven talking face synthesis (TFS) focuses on generating lifelike facial animations from speech input. Current TFS models perform well in En

safetyarxiv-cs-cv
7 Aug 2026
Research

Confidence matters: Leveraging Multi-view Geometric Priors for GS-based Reconstruction

DGX agent

arXiv:2608.06117v1 Announce Type: new Abstract: 3D Gaussian splatting (3DGS) has emerged as a widely-used tool for novel view synthesis, offering real-time rendering in a sparse representation. Howeve

researcharxiv-cs-cv
7 Aug 2026
Local Ai

Floating Radiance Networks

DGX agent

arXiv:2608.05920v1 Announce Type: new Abstract: Recent advances in neural scene representations enable photorealistic novel-view synthesis, yet most methods remain tightly coupled to a single renderin

local-aiarxiv-cs-cv
7 Aug 2026
Local Ai

Vorch-Streamer: Extending Human Audio-Visual Generation to Real-Time Long-Form Streaming

DGX agent

arXiv:2608.05663v1 Announce Type: new Abstract: Real-time long-form avatar audio--video generation requires causal, continuous synthesis while maintaining audiovisual synchronization and visual consis

local-aiarxiv-cs-cv
7 Aug 2026
Safety

Flash-VAED: Plug-and-Play VAE Decoders for Efficient Video Generation

DGX agent

arXiv:2602.19161v2 Announce Type: replace Abstract: Latent diffusion models have enabled high-quality video synthesis, yet their inference remains costly and time-consuming. As diffusion transformers

safetyarxiv-cs-cv
6 Aug 2026
Model Releases

NuclearDiffusion: Text-to-Image Foundation Models for Learning Nuclear Energy Concepts

DGX agent

arXiv:2608.04030v1 Announce Type: cross Abstract: Generative artificial intelligence (AI) has transformed text-to-image synthesis, yet its ability to represent specialized engineering domains remains

model-releasesarxiv-cs-ai
6 Aug 2026
Safety

Optimal Constrained sc-LTL Planning in MDPs via Switching Policies

DGX agent

arXiv:2608.05021v1 Announce Type: new Abstract: We study the synthesis of optimal policies for planning problems on Markov decision processes with both objectives and safety constraints specified in c

safetyarxiv-cs-ro
6 Aug 2026
Applications

Visualizing Graph-to-Answer Mechanism Recovery in Materials-Science Hypothesis Generation

DGX agent

arXiv:2608.04170v1 Announce Type: cross Abstract: AI co-scientists can generate fluent materials-science hypotheses, but fluency does not show that an answer preserves a scientifically meaningful mech

applicationsarxiv-cs-ai
6 Aug 2026
Model Releases

3DGSI-Assessor: A Large-Scale Dataset and An LMM-based Method for 3D Gaussian Splatting Image Quality Assessment

DGX agent

arXiv:2608.03279v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) has become a dominant representation for real-time novel view synthesis (NVS), yet its storage footprint makes compression

model-releasesarxiv-cs-cv
5 Aug 2026
Model Releases

UniWorld-Design: From Pixel Generation to Layer-Native Design

DGX agent

arXiv:2608.03971v1 Announce Type: new Abstract: We introduce UniWorld-Design, a framework that redefines image generation from flat pixel synthesis to structured visual composition, with semantic RGBA

model-releasesarxiv-cs-cv
5 Aug 2026
← Previous
1…1213141516…59
Next →