AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “synthesis”

GridTimelineEvolution
2,818 results
10 Aug 2026

Multi-Agent Forensic Reasoning for Generalizable Deepfake Video Detection

Model ReleasesDGX agent

arXiv:2608.06865v1 Announce Type: cross Abstract: The malicious use of generative artificial intelligence to create highly realistic deepfake videos raises serious ethical concerns and poses substanti

Symbolic Graphics Programming with Large Language Models

Model ReleasesDGX agent

arXiv:2509.05208v2 Announce Type: replace Abstract: Large language models (LLMs) excel at program synthesis, yet their ability to produce symbolic graphics programs (SGPs) that render into precise vis

7 Aug 2026

A Bridge from Audio to Video: Phoneme-Viseme Alignment Allows Every Face to Speak Multiple Languages


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety
DGX agent

arXiv:2510.06612v2 Announce Type: replace Abstract: Speech-driven talking face synthesis (TFS) focuses on generating lifelike facial animations from speech input. Current TFS models perform well in En

Confidence matters: Leveraging Multi-view Geometric Priors for GS-based Reconstruction

ResearchDGX agent

arXiv:2608.06117v1 Announce Type: new Abstract: 3D Gaussian splatting (3DGS) has emerged as a widely-used tool for novel view synthesis, offering real-time rendering in a sparse representation. Howeve

Floating Radiance Networks

Local AiDGX agent

arXiv:2608.05920v1 Announce Type: new Abstract: Recent advances in neural scene representations enable photorealistic novel-view synthesis, yet most methods remain tightly coupled to a single renderin

Vorch-Streamer: Extending Human Audio-Visual Generation to Real-Time Long-Form Streaming

Local AiDGX agent

arXiv:2608.05663v1 Announce Type: new Abstract: Real-time long-form avatar audio--video generation requires causal, continuous synthesis while maintaining audiovisual synchronization and visual consis

6 Aug 2026

Flash-VAED: Plug-and-Play VAE Decoders for Efficient Video Generation

SafetyDGX agent

arXiv:2602.19161v2 Announce Type: replace Abstract: Latent diffusion models have enabled high-quality video synthesis, yet their inference remains costly and time-consuming. As diffusion transformers

NuclearDiffusion: Text-to-Image Foundation Models for Learning Nuclear Energy Concepts

Model ReleasesDGX agent

arXiv:2608.04030v1 Announce Type: cross Abstract: Generative artificial intelligence (AI) has transformed text-to-image synthesis, yet its ability to represent specialized engineering domains remains

Optimal Constrained sc-LTL Planning in MDPs via Switching Policies

SafetyDGX agent

arXiv:2608.05021v1 Announce Type: new Abstract: We study the synthesis of optimal policies for planning problems on Markov decision processes with both objectives and safety constraints specified in c

Visualizing Graph-to-Answer Mechanism Recovery in Materials-Science Hypothesis Generation

ApplicationsDGX agent

arXiv:2608.04170v1 Announce Type: cross Abstract: AI co-scientists can generate fluent materials-science hypotheses, but fluency does not show that an answer preserves a scientifically meaningful mech

5 Aug 2026

3DGSI-Assessor: A Large-Scale Dataset and An LMM-based Method for 3D Gaussian Splatting Image Quality Assessment

Model ReleasesDGX agent

arXiv:2608.03279v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) has become a dominant representation for real-time novel view synthesis (NVS), yet its storage footprint makes compression

UniWorld-Design: From Pixel Generation to Layer-Native Design

Model ReleasesDGX agent

arXiv:2608.03971v1 Announce Type: new Abstract: We introduce UniWorld-Design, a framework that redefines image generation from flat pixel synthesis to structured visual composition, with semantic RGBA

4 Aug 2026

D^2-4DGS: Dual-Depth Guided Sparse-Camera 4D Gaussian Splatting

ResearchDGX agent

arXiv:2608.01588v1 Announce Type: new Abstract: Dynamic 4D Gaussian Splatting has emerged as an efficient representation for dynamic novel view synthesis through explicit scene modeling and real-time

DeGS: A Scalable 3DGS Architecture via Decoupled Workload Parsing and Reorganization

ResearchDGX agent

arXiv:2608.02099v1 Announce Type: cross Abstract: 3D Gaussian Splatting (3DGS) has emerged as a leading technique for real-time novel view synthesis, yet existing 3DGS accelerators suffer from poor ar

Exploring More to Solve More: Boosting Diversity in Text Diffusion Models via Entropy-Based Guidance

ResearchDGX agent

arXiv:2608.00024v1 Announce Type: new Abstract: Although diffusion models have revolutionized continuous domains like image synthesis through high quality generations and controllable guidance mechani

Hybrid-Domain Posterior Sampling for Inverse Problems via Latent Flow Matching

SafetyDGX agent

arXiv:2608.00537v1 Announce Type: new Abstract: Latent Flow Models have revolutionized compressed-space image synthesis, yet their application to high-fidelity inverse problems remains bottlenecked. I

LongChart VQA: A Comprehensive Benchmark for MLLMs with Complex Multi-Chart Reasoning

Model ReleasesDGX agent

arXiv:2608.01328v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) are rapidly evolving with expanded context windows and stronger reasoning capabilities, enabling multi-chart un

3 Aug 2026

DualDiT: A Conditional Dual-Output Diffusion Transformer for Joint OCT Image and Segmentation Mask Generation

TutorialsDGX agent

arXiv:2607.29337v1 Announce Type: cross Abstract: Background and Objective: Generating realistic medical images with anatomically accurate segmentation masks helps address the shortage of annotated da

Enabling Low-Latency Machine learning on Radiation-Hard FPGAs with hls4ml

ResearchDGX agent

arXiv:2602.15751v2 Announce Type: replace-cross Abstract: This paper presents an end-to-end demonstration of a viable, ultra-fast, radiation-hard machine learning (ML) application on FPGAs, which coul

31 Jul 2026

Split and Drive: Dual-Axis Disentanglement for Real-Time Gaussian Head Avatars

HardwareDGX agent

arXiv:2607.28032v1 Announce Type: new Abstract: Creating photorealistic animatable head avatars from a single image remains a fundamental challenge in digital human synthesis. While recent 3D Gaussian

29 Jul 2026

ContractHIL-HLS: Contract-Aligned Multi-Agent Workflow with Hardware-in-the-Loop Feedback for HLS Design

SafetyDGX agent

arXiv:2607.25283v1 Announce Type: new Abstract: This paper presents ContractHIL-HLS, a contract-aligned multi-agent workflow for practical high-level synthesis (HLS) engineering. The workflow makes th

Diff-ID: Identity Consistent Facial Image Generation and Morphing via Diffusion Models

ResearchDGX agent

arXiv:2607.25078v1 Announce Type: new Abstract: Generative diffusion models have revolutionized facial image synthesis, yet robust identity preservation in high resolution outputs remains a critical c

TIGA: Trajectory-Injected Generative Attack against Black-box AIGC Detectors

ResearchDGX agent

arXiv:2607.25894v1 Announce Type: new Abstract: Recent diffusion models have achieved remarkable realism in facial image synthesis, posing growing challenges to artificial intelligence-generated conte

28 Jul 2026

OmniVAE: An Audio-Video VAE with Cross-Modal Alignment for Joint Generation

SafetyDGX agent

arXiv:2607.23855v1 Announce Type: cross Abstract: Recent generative models are moving beyond silent video or standalone audio synthesis toward the joint generation of synchronized audio and video. Des

TopoFE: topology-aware LLM-guided Automated Feature Engineering

ResearchDGX agent

arXiv:2607.23286v1 Announce Type: new Abstract: Automatic feature engineering (AutoFE) for tabular learning can be naturally formulated as a program synthesis problem, where the objective is to discov

27 Jul 2026

Deformable Triangle Splatting: Flexible Primitives for Real-Time Radiance Field Rendering

SafetyDGX agent

arXiv:2607.22446v1 Announce Type: new Abstract: Recent radiance field methods represent scenes with 2D primitives that offer surface alignment and efficient rasterization, from Gaussian disks to trian

Hash-QNeRF: Multiresolution Hash Encoding for Quantum Neural Radiance Fields

ResearchDGX agent

arXiv:2607.21675v1 Announce Type: cross Abstract: Neural Radiance Fields (NeRF) have revolutionized novel view synthesis, yet their classical implementations remain computationally intensive for high-

InnoText: A Unified Model for Visual Text Generation and Editing

ResearchDGX agent

arXiv:2607.22101v1 Announce Type: new Abstract: Diffusion models have recently achieved remarkable success in high-fidelity image synthesis, yet their application to visual text generation and editing

24 Jul 2026

GLAN-QnA-KR: A Seedless Taxonomy-Driven Korean Instruction Corpus

Model ReleasesDGX agent

arXiv:2607.20443v1 Announce Type: new Abstract: We release GLAN-QnA-KR, a 303,581-row openly redistributable Korean instruction-QA corpus produced via the seedless taxonomy-driven GLAN synthesis pipel

Is Deep Research Reliable? Misleading Knowledge Induces False Conclusions

Model ReleasesDGX agent

arXiv:2607.20891v1 Announce Type: new Abstract: Deep Research agents extend LLM-based assistants into long-horizon workflows involving planning, retrieval, evidence synthesis, and report generation, y

SciExplore: Evaluating Autonomous Agents from Scientific Navigation to Information Integration

Model ReleasesDGX agent

arXiv:2607.20926v1 Announce Type: new Abstract: Scientific research involves complex information-seeking and reasoning workflows across heterogeneous sources. However, existing benchmarks primarily em

SubSplat: High-Resolution Pixel-aligned 3DGS via Sub-pixel Gaussian Reparameterization

ResearchDGX agent

arXiv:2607.20813v1 Announce Type: new Abstract: Pixel-aligned Gaussian splatting enables efficient and generalizable novel-view synthesis. However, high-resolution rendering faces a critical trade-off

WhereEdit: Mask-aware Local Latent Editing for One-Step Image Editing

Model ReleasesDGX agent

arXiv:2607.20883v1 Announce Type: new Abstract: Recent one-step text-to-image (T2I) models enable efficient image synthesis and provide new opportunities for real-time image editing. However, existing

23 Jul 2026

HeadCast: Casting Attention Heads for Efficient Autoregressive Video Generation

ResearchDGX agent

arXiv:2607.20125v1 Announce Type: cross Abstract: Autoregressive (AR) video diffusion models have become a promising paradigm for long and streaming video synthesis, but the continuously growing Key-V

Pixel-Space Diffusion Transformers

ResearchDGX agent

arXiv:2607.17585v2 Announce Type: replace Abstract: Latent diffusion models (LDMs) enable efficient high-resolution image synthesis by denoising in a VAE-compressed latent space. However, fixed visual

Surprise Forcing: What to Remember, When to Skip in Long Video Generation

ResearchDGX agent

arXiv:2607.18436v1 Announce Type: new Abstract: Streaming autoregressive diffusion makes minute-scale video synthesis practical, but its bounded context and fixed denoising schedule allocate resources

Timeripple: Accelerating vDiTs by Understanding the Spatio-Temporal Correlations in Latent Space

ResearchDGX agent

arXiv:2511.12035v2 Announce Type: replace-cross Abstract: The recent surge in video generation has shown the growing demand for high-quality video synthesis using large vision models. Existing video g

16 Jul 2026

NanoGS: Training-Free Gaussian Splat Simplification

HardwareDGX agent

arXiv:2603.16103v2 Announce Type: replace Abstract: 3D Gaussian Splat (3DGS) enables high-fidelity, real-time novel view synthesis by representing scenes with large sets of anisotropic primitives, but

PersGuard: Preventing Malicious Personalization in Text-to-Image Diffusion Models via Model Backdoors

ResearchDGX agent

arXiv:2502.16167v2 Announce Type: replace-cross Abstract: Diffusion models (DMs) have advanced text-to-image (T2I) synthesis, yet their personalization capabilities raise serious privacy and copyright

15 Jul 2026

ExtraGS: Enhancing Endoscopic View Extrapolation via Diffusion-Guided 3D Gaussian Splatting

SafetyDGX agent

arXiv:2607.12785v1 Announce Type: new Abstract: Robot-assisted minimally invasive surgery (MIS) critically depends on reliable endoscopic perception for navigation and safety. However, conventional en

Hallo4D: Multi-Modal Hallucination Mitigation for Consistent Spatio-Temporal Generation

SafetyDGX agent

arXiv:2607.12752v1 Announce Type: cross Abstract: While recent advances in 3D generation have enabled impressive visual synthesis, existing methods often rely on 2D diffusion supervision without expli

Wiki Lint Report — 2026-07-15

SynthesesDGX agent

Automated lint: 26 errors, 6728 warnings, 3 info

SymbOmni: Evolving Agentic Omni Models via Symbolic Concept Learning

AgentsDGX agent

arXiv:2607.12042v1 Announce Type: new Abstract: Visual generation is increasingly ubiquitous in diverse domains, from text-to-image/video synthesis to multimodal interactive creation. Yet prevailing m

10 Jul 2026

DR-Arena: an Automated Evaluation Framework for Deep Research Agents

SafetyDGX agent

arXiv:2601.10504v2 Announce Type: replace Abstract: As Large Language Models (LLMs) increasingly operate as Deep Research (DR) Agents capable of autonomous investigation and information synthesis, rel

9 Jul 2026

Safe Reinforcement Learning using Ideas from Model Predictive Control

SafetyDGX agent

arXiv:2607.07252v1 Announce Type: new Abstract: Reinforcement learning (RL) enables the synthesis of control policies directly from data, making it highly appealing for complex cyber-physical systems

8 Jul 2026

FirstResearch: Auditable Question Formation for LLM Scientific Discovery Agents

Model ReleasesDGX agent

arXiv:2607.05682v1 Announce Type: new Abstract: LLM systems for scientific discovery increasingly assist with ideation, literature synthesis, experiment planning, and report generation, but the first

Integrating knowledge graphs and multilingual scholarly corpora for domain-adaptive LLMs in SSH

ApplicationsDGX agent

arXiv:2607.05956v1 Announce Type: new Abstract: The integration of Large Language Models (LLMs) into scientific research workflows, particularly for bibliographic discovery and literature synthesis, r

Kling 3.0: https://links.comfy.org/4d2AMN9

Local AiDGX agent

Kling 3.0 is a video generation model or update shared by the ComfyUI project, likely featuring improvements to video synthesis capabilities or integration features. The announcement was made via Comf

Rendering-Aware Bayesian 3D Gaussian Splatting with Native Uncertainty and Adaptive Complexity Control

ResearchDGX agent

arXiv:2607.05522v1 Announce Type: cross Abstract: 3D Gaussian splatting (3DGS) is a strong representation for real-time novel-view synthesis, but its standard training pipeline relies on point estimat

SSA-3DGS: Unsupervised Removal of Screen-Space Artifacts for 3D Gaussian Splatting

ApplicationsDGX agent

arXiv:2607.05598v1 Announce Type: cross Abstract: Novel View Synthesis (NVS) methods, such as 3D Gaussian Splatting (3DGS), rely heavily on the assumption of clean, multi-view consistent, posed input

7 Jul 2026

Aura: Consistent Multi-Subject Video Generation via VLM-Grounded Semantic Alignment

SafetyDGX agent

arXiv:2607.04311v1 Announce Type: new Abstract: Subject-driven and multi-element video generation are central to controllable video synthesis, but existing methods still struggle to preserve identity

BiSLW: Bi-Spectral Latent Watermarking for Generative Diffusion Models

SafetyDGX agent

arXiv:2607.02643v1 Announce Type: new Abstract: Diffusion-based generative models have transformed visual content synthesis, yet they remain vulnerable to unauthorized usage and lack reliable attribut

Diagnosing Aerial-View Object Detectors with Foundational Image Generative Models

TutorialsDGX agent

arXiv:2607.02718v1 Announce Type: cross Abstract: Recent advances in large-scale image generative models enable photorealistic scene synthesis with controllable attributes. Beyond data augmentation, t

EmoteGPT: 3D Human Facial Expressions from Natural Language Descriptions

Model ReleasesDGX agent

arXiv:2607.02674v1 Announce Type: new Abstract: Precise control of 3D facial expressions from text is crucial for virtual avatars, animation, and human-computer interaction, yet existing text-to-3D me

Fourier Splatting: Generalized Fourier encoded primitives for scalable radiance fields

ResearchDGX agent

arXiv:2603.19834v3 Announce Type: replace Abstract: Novel view synthesis has recently been revolutionized by 3D Gaussian Splatting (3DGS), which enables real-time rendering through explicit primitive

InSpace: Structure-Aware 3D Indoor Scene Generation from a Single 360{eg} Image

ResearchDGX agent

arXiv:2607.03990v1 Announce Type: new Abstract: Recent advances in single image-to-3D generation have enabled high-quality asset synthesis, yet extending these capabilities to indoor scene generation

KARMA: Knowledge graph-based Automated Reasoning Materialization and Alignment

SafetyDGX agent

arXiv:2607.03166v1 Announce Type: cross Abstract: Template-based contrastive synthesis is scalable, but its candidates often differ only in a few entity-slots while sequence-level optimization spreads

LAW & ORDER: Adaptive Spatial Weighting for Medical Diffusion and Segmentation

ResearchDGX agent

arXiv:2603.04795v2 Announce Type: replace-cross Abstract: Medical image analysis depends on accurate segmentation and controllable synthesis, but both tasks face severe spatial imbalance: lesions occu

Learning on the Manifold: Unlocking Standard Diffusion Transformers with Representation Encoders

ResearchDGX agent

arXiv:2602.10099v2 Announce Type: replace-cross Abstract: Leveraging representation encoders for generative modeling offers a path for efficient, high-fidelity synthesis. However, standard diffusion t

MACRO: Training-free Multi-plane Attention for Closeup Render Optimization

ApplicationsDGX agent

arXiv:2607.03875v1 Announce Type: new Abstract: Close-up rendering, zooming into a scene well beyond any training camera, is important for virtual production and interactive 3D content, yet remains an

← Previous
1…1011121314…47
Next →