AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “synthesis”

GridTimelineEvolution
2,711 results
Model Releases

SceneScribe-1M: A Large-Scale Video Dataset with Comprehensive Geometric and Semantic Annotations

DGX agent

arXiv:2604.07990v1 Announce Type: new Abstract: The convergence of 3D geometric perception and video synthesis has created an unprecedented demand for large-scale video data that is rich in both seman

model-releasesarxiv-cs-cv
10 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Applications

CasDeblurGS: Cascaded 2D-to-3D Multi-View Consistency for 3D Gaussian Splatting from Two Blurry Images

DGX agent

arXiv:2608.10345v1 Announce Type: new Abstract: Free-viewpoint 3D scene media is increasingly important for immersive applications, yet practical capture often suffers from severe view sparsity and mo

applicationsarxiv-cs-cv
12 Aug 2026
Tutorials

Gaussian Sculpting: End-to-End Controllable Surface Reconstruction via Field Optimization

DGX agent

arXiv:2608.10602v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) has recently enabled real-time novel view synthesis with impressive quality. However, it struggles to recover accurate surf

tutorialsarxiv-cs-cv
12 Aug 2026
Research

Knowledge-Guided 3D CT Generation: A Conditioning-Centric Taxonomy

DGX agent

arXiv:2608.09992v1 Announce Type: cross Abstract: Controllable generation guided by external knowledge is a key requirement in modern generative deep learning applications, enabling the synthesis of s

researcharxiv-cs-ai
12 Aug 2026
Research

Order Matters: LVLMs as Judges for Temporal Reasoning in Image Sequences

DGX agent

arXiv:2608.10908v1 Announce Type: cross Abstract: As generative multimedia evolves from static image synthesis to complex, interleaved visual narratives, a foundational bottleneck has emerged: the jud

researcharxiv-cs-cl
12 Aug 2026
Research

FlexSplat: Flexible Feed-Forward 3D Gaussian Splatting without Point Cloud Correspondence

DGX agent

arXiv:2608.07937v1 Announce Type: new Abstract: We present FlexSplat, a feed-forward framework for novel view synthesis (NVS) from uncalibrated, object-centric multi-view image collections. A recent l

researcharxiv-cs-cv
11 Aug 2026
Model Releases

MADBench: A Benchmark for Modality-Aware Audio Deepfake Detection

DGX agent

arXiv:2608.09593v1 Announce Type: cross Abstract: Recent advances in speech synthesis and audio generation have made high-fidelity acoustic forgery low-cost and difficult to attribute, enabling a real

model-releasesarxiv-cs-ai
11 Aug 2026
Safety

GOPI: Generation-Oriented 3D Pose Inference for Furniture Insertion from Single-View RGB-D Indoor Scenes

DGX agent

arXiv:2608.06836v1 Announce Type: new Abstract: We study the problem of inserting new furniture into indoor scene images. Under masked single-view 2D image-plane conditioning, however, the physical sc

safetyarxiv-cs-cv
10 Aug 2026
Safety

Identity as Presence: Towards Appearance and Voice Personalized Joint Audio-Video Generation

DGX agent

arXiv:2603.17889v4 Announce Type: replace Abstract: Recent advances in video synthesis have enabled realistic integration of real individuals, driving demand for identity-aware generation. While emerg

safetyarxiv-cs-cv
10 Aug 2026
Applications

Improving Low-Resolution Face Recognition under Limited Data: How Synthetic Data Generation Can Close the Domain Gap

DGX agent

arXiv:2608.06580v1 Announce Type: new Abstract: Face Recognition (FR) systems in surveillance settings often encounter Low Resolution (LR) faces, those whose face region falls below the standard 112 i

applicationsarxiv-cs-cv
10 Aug 2026
Model Releases

MirrorWorld: Taming Video Diffusion Models for Mirror Reflection Generation

DGX agent

arXiv:2608.07463v1 Announce Type: new Abstract: Recent advances in video diffusion models (VDMs) have enabled high-fidelity video synthesis. However, generating mirror reflections remains challenging

model-releasesarxiv-cs-cv
10 Aug 2026
Model Releases

Multi-Agent Forensic Reasoning for Generalizable Deepfake Video Detection

DGX agent

arXiv:2608.06865v1 Announce Type: cross Abstract: The malicious use of generative artificial intelligence to create highly realistic deepfake videos raises serious ethical concerns and poses substanti

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Symbolic Graphics Programming with Large Language Models

DGX agent

arXiv:2509.05208v2 Announce Type: replace Abstract: Large language models (LLMs) excel at program synthesis, yet their ability to produce symbolic graphics programs (SGPs) that render into precise vis

model-releasesarxiv-cs-cv
10 Aug 2026
Safety

A Bridge from Audio to Video: Phoneme-Viseme Alignment Allows Every Face to Speak Multiple Languages

DGX agent

arXiv:2510.06612v2 Announce Type: replace Abstract: Speech-driven talking face synthesis (TFS) focuses on generating lifelike facial animations from speech input. Current TFS models perform well in En

safetyarxiv-cs-cv
7 Aug 2026
Research

Confidence matters: Leveraging Multi-view Geometric Priors for GS-based Reconstruction

DGX agent

arXiv:2608.06117v1 Announce Type: new Abstract: 3D Gaussian splatting (3DGS) has emerged as a widely-used tool for novel view synthesis, offering real-time rendering in a sparse representation. Howeve

researcharxiv-cs-cv
7 Aug 2026
Local Ai

Floating Radiance Networks

DGX agent

arXiv:2608.05920v1 Announce Type: new Abstract: Recent advances in neural scene representations enable photorealistic novel-view synthesis, yet most methods remain tightly coupled to a single renderin

local-aiarxiv-cs-cv
7 Aug 2026
Local Ai

Vorch-Streamer: Extending Human Audio-Visual Generation to Real-Time Long-Form Streaming

DGX agent

arXiv:2608.05663v1 Announce Type: new Abstract: Real-time long-form avatar audio--video generation requires causal, continuous synthesis while maintaining audiovisual synchronization and visual consis

local-aiarxiv-cs-cv
7 Aug 2026
Safety

Flash-VAED: Plug-and-Play VAE Decoders for Efficient Video Generation

DGX agent

arXiv:2602.19161v2 Announce Type: replace Abstract: Latent diffusion models have enabled high-quality video synthesis, yet their inference remains costly and time-consuming. As diffusion transformers

safetyarxiv-cs-cv
6 Aug 2026
Model Releases

NuclearDiffusion: Text-to-Image Foundation Models for Learning Nuclear Energy Concepts

DGX agent

arXiv:2608.04030v1 Announce Type: cross Abstract: Generative artificial intelligence (AI) has transformed text-to-image synthesis, yet its ability to represent specialized engineering domains remains

model-releasesarxiv-cs-ai
6 Aug 2026
Safety

Optimal Constrained sc-LTL Planning in MDPs via Switching Policies

DGX agent

arXiv:2608.05021v1 Announce Type: new Abstract: We study the synthesis of optimal policies for planning problems on Markov decision processes with both objectives and safety constraints specified in c

safetyarxiv-cs-ro
6 Aug 2026
Applications

Visualizing Graph-to-Answer Mechanism Recovery in Materials-Science Hypothesis Generation

DGX agent

arXiv:2608.04170v1 Announce Type: cross Abstract: AI co-scientists can generate fluent materials-science hypotheses, but fluency does not show that an answer preserves a scientifically meaningful mech

applicationsarxiv-cs-ai
6 Aug 2026
Model Releases

3DGSI-Assessor: A Large-Scale Dataset and An LMM-based Method for 3D Gaussian Splatting Image Quality Assessment

DGX agent

arXiv:2608.03279v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) has become a dominant representation for real-time novel view synthesis (NVS), yet its storage footprint makes compression

model-releasesarxiv-cs-cv
5 Aug 2026
Model Releases

UniWorld-Design: From Pixel Generation to Layer-Native Design

DGX agent

arXiv:2608.03971v1 Announce Type: new Abstract: We introduce UniWorld-Design, a framework that redefines image generation from flat pixel synthesis to structured visual composition, with semantic RGBA

model-releasesarxiv-cs-cv
5 Aug 2026
Research

D^2-4DGS: Dual-Depth Guided Sparse-Camera 4D Gaussian Splatting

DGX agent

arXiv:2608.01588v1 Announce Type: new Abstract: Dynamic 4D Gaussian Splatting has emerged as an efficient representation for dynamic novel view synthesis through explicit scene modeling and real-time

researcharxiv-cs-cv
4 Aug 2026
Research

DeGS: A Scalable 3DGS Architecture via Decoupled Workload Parsing and Reorganization

DGX agent

arXiv:2608.02099v1 Announce Type: cross Abstract: 3D Gaussian Splatting (3DGS) has emerged as a leading technique for real-time novel view synthesis, yet existing 3DGS accelerators suffer from poor ar

researcharxiv-cs-cv
4 Aug 2026
Research

Exploring More to Solve More: Boosting Diversity in Text Diffusion Models via Entropy-Based Guidance

DGX agent

arXiv:2608.00024v1 Announce Type: new Abstract: Although diffusion models have revolutionized continuous domains like image synthesis through high quality generations and controllable guidance mechani

researcharxiv-cs-cl
4 Aug 2026
Safety

Hybrid-Domain Posterior Sampling for Inverse Problems via Latent Flow Matching

DGX agent

arXiv:2608.00537v1 Announce Type: new Abstract: Latent Flow Models have revolutionized compressed-space image synthesis, yet their application to high-fidelity inverse problems remains bottlenecked. I

safetyarxiv-cs-cv
4 Aug 2026
Model Releases

LongChart VQA: A Comprehensive Benchmark for MLLMs with Complex Multi-Chart Reasoning

DGX agent

arXiv:2608.01328v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) are rapidly evolving with expanded context windows and stronger reasoning capabilities, enabling multi-chart un

model-releasesarxiv-cs-cl
4 Aug 2026
Tutorials

DualDiT: A Conditional Dual-Output Diffusion Transformer for Joint OCT Image and Segmentation Mask Generation

DGX agent

arXiv:2607.29337v1 Announce Type: cross Abstract: Background and Objective: Generating realistic medical images with anatomically accurate segmentation masks helps address the shortage of annotated da

tutorialsarxiv-cs-ai
3 Aug 2026
Research

Enabling Low-Latency Machine learning on Radiation-Hard FPGAs with hls4ml

DGX agent

arXiv:2602.15751v2 Announce Type: replace-cross Abstract: This paper presents an end-to-end demonstration of a viable, ultra-fast, radiation-hard machine learning (ML) application on FPGAs, which coul

researcharxiv-cs-lg
3 Aug 2026
Hardware

Split and Drive: Dual-Axis Disentanglement for Real-Time Gaussian Head Avatars

DGX agent

arXiv:2607.28032v1 Announce Type: new Abstract: Creating photorealistic animatable head avatars from a single image remains a fundamental challenge in digital human synthesis. While recent 3D Gaussian

hardwarearxiv-cs-cv
31 Jul 2026
Safety

ContractHIL-HLS: Contract-Aligned Multi-Agent Workflow with Hardware-in-the-Loop Feedback for HLS Design

DGX agent

arXiv:2607.25283v1 Announce Type: new Abstract: This paper presents ContractHIL-HLS, a contract-aligned multi-agent workflow for practical high-level synthesis (HLS) engineering. The workflow makes th

safetyarxiv-cs-ai
29 Jul 2026
Research

Diff-ID: Identity Consistent Facial Image Generation and Morphing via Diffusion Models

DGX agent

arXiv:2607.25078v1 Announce Type: new Abstract: Generative diffusion models have revolutionized facial image synthesis, yet robust identity preservation in high resolution outputs remains a critical c

researcharxiv-cs-cv
29 Jul 2026
Research

TIGA: Trajectory-Injected Generative Attack against Black-box AIGC Detectors

DGX agent

arXiv:2607.25894v1 Announce Type: new Abstract: Recent diffusion models have achieved remarkable realism in facial image synthesis, posing growing challenges to artificial intelligence-generated conte

researcharxiv-cs-cv
29 Jul 2026
Safety

OmniVAE: An Audio-Video VAE with Cross-Modal Alignment for Joint Generation

DGX agent

arXiv:2607.23855v1 Announce Type: cross Abstract: Recent generative models are moving beyond silent video or standalone audio synthesis toward the joint generation of synchronized audio and video. Des

safetyarxiv-cs-cv
28 Jul 2026
Research

TopoFE: topology-aware LLM-guided Automated Feature Engineering

DGX agent

arXiv:2607.23286v1 Announce Type: new Abstract: Automatic feature engineering (AutoFE) for tabular learning can be naturally formulated as a program synthesis problem, where the objective is to discov

researcharxiv-cs-ai
28 Jul 2026
Safety

Deformable Triangle Splatting: Flexible Primitives for Real-Time Radiance Field Rendering

DGX agent

arXiv:2607.22446v1 Announce Type: new Abstract: Recent radiance field methods represent scenes with 2D primitives that offer surface alignment and efficient rasterization, from Gaussian disks to trian

safetyarxiv-cs-cv
27 Jul 2026
Research

Hash-QNeRF: Multiresolution Hash Encoding for Quantum Neural Radiance Fields

DGX agent

arXiv:2607.21675v1 Announce Type: cross Abstract: Neural Radiance Fields (NeRF) have revolutionized novel view synthesis, yet their classical implementations remain computationally intensive for high-

researcharxiv-cs-cv
27 Jul 2026
Research

InnoText: A Unified Model for Visual Text Generation and Editing

DGX agent

arXiv:2607.22101v1 Announce Type: new Abstract: Diffusion models have recently achieved remarkable success in high-fidelity image synthesis, yet their application to visual text generation and editing

researcharxiv-cs-cv
27 Jul 2026
Model Releases

GLAN-QnA-KR: A Seedless Taxonomy-Driven Korean Instruction Corpus

DGX agent

arXiv:2607.20443v1 Announce Type: new Abstract: We release GLAN-QnA-KR, a 303,581-row openly redistributable Korean instruction-QA corpus produced via the seedless taxonomy-driven GLAN synthesis pipel

model-releasesarxiv-cs-cl
24 Jul 2026
Model Releases

Is Deep Research Reliable? Misleading Knowledge Induces False Conclusions

DGX agent

arXiv:2607.20891v1 Announce Type: new Abstract: Deep Research agents extend LLM-based assistants into long-horizon workflows involving planning, retrieval, evidence synthesis, and report generation, y

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

SciExplore: Evaluating Autonomous Agents from Scientific Navigation to Information Integration

DGX agent

arXiv:2607.20926v1 Announce Type: new Abstract: Scientific research involves complex information-seeking and reasoning workflows across heterogeneous sources. However, existing benchmarks primarily em

model-releasesarxiv-cs-ai
24 Jul 2026
Research

SubSplat: High-Resolution Pixel-aligned 3DGS via Sub-pixel Gaussian Reparameterization

DGX agent

arXiv:2607.20813v1 Announce Type: new Abstract: Pixel-aligned Gaussian splatting enables efficient and generalizable novel-view synthesis. However, high-resolution rendering faces a critical trade-off

researcharxiv-cs-cv
24 Jul 2026
Model Releases

WhereEdit: Mask-aware Local Latent Editing for One-Step Image Editing

DGX agent

arXiv:2607.20883v1 Announce Type: new Abstract: Recent one-step text-to-image (T2I) models enable efficient image synthesis and provide new opportunities for real-time image editing. However, existing

model-releasesarxiv-cs-cv
24 Jul 2026
Research

HeadCast: Casting Attention Heads for Efficient Autoregressive Video Generation

DGX agent

arXiv:2607.20125v1 Announce Type: cross Abstract: Autoregressive (AR) video diffusion models have become a promising paradigm for long and streaming video synthesis, but the continuously growing Key-V

researcharxiv-cs-lg
23 Jul 2026
Research

Pixel-Space Diffusion Transformers

DGX agent

arXiv:2607.17585v2 Announce Type: replace Abstract: Latent diffusion models (LDMs) enable efficient high-resolution image synthesis by denoising in a VAE-compressed latent space. However, fixed visual

researcharxiv-cs-cv
23 Jul 2026
Research

Surprise Forcing: What to Remember, When to Skip in Long Video Generation

DGX agent

arXiv:2607.18436v1 Announce Type: new Abstract: Streaming autoregressive diffusion makes minute-scale video synthesis practical, but its bounded context and fixed denoising schedule allocate resources

researcharxiv-cs-cv
23 Jul 2026
Research

Timeripple: Accelerating vDiTs by Understanding the Spatio-Temporal Correlations in Latent Space

DGX agent

arXiv:2511.12035v2 Announce Type: replace-cross Abstract: The recent surge in video generation has shown the growing demand for high-quality video synthesis using large vision models. Existing video g

researcharxiv-cs-cv
23 Jul 2026
← Previous
1…1213141516…57
Next →