AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “synthesis”

GridTimelineEvolution
2,844 results
Model Releases

ControlFoley: Unified and Controllable Video-to-Audio Generation with Cross-Modal Conflict Handling

DGX agent

arXiv:2604.15086v1 Announce Type: cross Abstract: Recent advances in video-to-audio (V2A) generation enable high-quality audio synthesis from visual content, yet achieving robust and fine-grained cont

model-releasesarxiv-cs-cv
17 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Local Ai

Deepfake Detection Generalization with Diffusion Noise

DGX agent

arXiv:2604.14570v1 Announce Type: new Abstract: Deepfake detectors face growing challenges in generalization as new image synthesis techniques emerge. In particular, deepfakes generated by diffusion m

local-aiarxiv-cs-cv
17 Apr 2026
Research

Edge-preserving noise for diffusion models

DGX agent

arXiv:2410.01540v4 Announce Type: replace Abstract: Classical diffusion models typically rely on isotropic Gaussian noise, treating all regions uniformly and overlooking structural information importa

researcharxiv-cs-cv
17 Apr 2026
Tutorials

How to Fine-Tune a Reasoning Model? A Teacher-Student Cooperation Framework to Synthesize Student-Consistent SFT Data

DGX agent

arXiv:2604.14164v1 Announce Type: new Abstract: A widely adopted strategy for model enhancement is to use synthetic data generated by a stronger model for supervised fine-tuning (SFT). However, for em

tutorialsarxiv-cs-cl
17 Apr 2026
Research

Improved Multiscale Structural Mapping with Supervertex Vision Transformer for the Detection of Alzheimer's Disease Neurodegeneration

DGX agent

arXiv:2604.14837v1 Announce Type: new Abstract: Alzheimer's disease (AD) confirmation often relies on positron emission tomography (PET) or cerebrospinal fluid (CSF) analysis, which are costly and inv

researcharxiv-cs-cv
17 Apr 2026
Safety

Modeling LLM Unlearning as an Asymmetric Two-Task Learning Problem

DGX agent

arXiv:2604.14808v1 Announce Type: new Abstract: Machine unlearning for large language models (LLMs) aims to remove targeted knowledge while preserving general capability. In this paper, we recast LLM

safetyarxiv-cs-cl
17 Apr 2026
Safety

NG-GS: NeRF-Guided 3D Gaussian Splatting Segmentation

DGX agent

arXiv:2604.14706v1 Announce Type: new Abstract: Recent advances in 3D Gaussian Splatting (3DGS) have enabled highly efficient and photorealistic novel view synthesis. However, segmenting objects accur

safetyarxiv-cs-cv
17 Apr 2026
Research

PixelDiT: Pixel Diffusion Transformers for Image Generation

DGX agent

arXiv:2511.20645v2 Announce Type: replace Abstract: Latent-space modeling has been the standard for Diffusion Transformers (DiTs). However, it relies on a two-stage pipeline where the pretrained autoe

researcharxiv-cs-cv
17 Apr 2026
Safety

Irregularly Sampled Time Series Interpolation for Binary Evolution Simulations Using Dynamic Time Warping

DGX agent

arXiv:2604.13604v1 Announce Type: cross Abstract: Binary stellar evolution simulations are computationally expensive. Stellar population synthesis relies on these detailed evolution models at a fundam

safetyarxiv-cs-lg
16 Apr 2026
Applications

Mathematical Reasoning Enhanced LLM for Formula Derivation: A Case Study on Fiber NLI Modellin

DGX agent

arXiv:2604.13062v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) have demonstrated strong capabilities in code generation and text synthesis, yet their potential for sym

applicationsarxiv-cs-cl
16 Apr 2026
Applications

MSGS: Multispectral 3D Gaussian Splatting

DGX agent

arXiv:2604.13340v1 Announce Type: new Abstract: We present a multispectral extension to 3D Gaussian Splatting (3DGS) for wavelength-aware view synthesis. Each Gaussian is augmented with spectral radia

applicationsarxiv-cs-cv
16 Apr 2026
Research

Lyra 2.0: Explorable Generative 3D Worlds

DGX agent

arXiv:2604.13036v1 Announce Type: new Abstract: Recent advances in video generation enable a new paradigm for 3D scene creation: generating camera-controlled videos that simulate scene walkthroughs, t

researcharxiv-cs-cv
15 Apr 2026
Research

Prompt Evolution for Generative AI: A Classifier-Guided Approach

DGX agent

arXiv:2305.16347v2 Announce Type: replace-cross Abstract: Synthesis of digital artifacts conditioned on user prompts has become an important paradigm facilitating an explosion of use cases with genera

researcharxiv-cs-ai
15 Apr 2026
Research

Scaling Exposes the Trigger: Input-Level Backdoor Detection in Text-to-Image Diffusion Models via Cross-Attention Scaling

DGX agent

arXiv:2604.12446v1 Announce Type: cross Abstract: Text-to-image (T2I) diffusion models have achieved remarkable success in image synthesis, but their reliance on large-scale data and open ecosystems i

researcharxiv-cs-cv
15 Apr 2026
Model Releases

Self-Adversarial One Step Generation via Condition Shifting

DGX agent

arXiv:2604.12322v1 Announce Type: new Abstract: The push for efficient text to image synthesis has moved the field toward one step sampling, yet existing methods still face a three way tradeoff among

model-releasesarxiv-cs-cv
15 Apr 2026
Applications

FreeScale: Scaling 3D Scenes via Certainty-Aware Free-View Generation

DGX agent

arXiv:2604.10512v1 Announce Type: new Abstract: The development of generalizable Novel View Synthesis (NVS) models is critically limited by the scarcity of large-scale training data featuring diverse

applicationsarxiv-cs-cv
14 Apr 2026
Model Releases

Immunizing 3D Gaussian Generative Models Against Unauthorized Fine-Tuning via Attribute-Space Traps

DGX agent

arXiv:2604.09688v1 Announce Type: new Abstract: Recent large-scale generative models enable high-quality 3D synthesis. However, the public accessibility of pre-trained weights introduces a critical vu

model-releasesarxiv-cs-cv
14 Apr 2026
Safety

Large Language Model as An Operator: An Experience-Driven Solution for Distribution Network Voltage Control

DGX agent

arXiv:2507.14800v2 Announce Type: replace-cross Abstract: With the advanced reasoning, contextual understanding, and information synthesis capabilities of large language models (LLMs), a novel paradig

safetyarxiv-cs-ai
14 Apr 2026
Tutorials

Learning from Contrasts: Synthesizing Reasoning Paths from Diverse Search Trajectories

DGX agent

arXiv:2604.11365v1 Announce Type: new Abstract: Monte Carlo Tree Search (MCTS) has been widely used for automated reasoning data exploration, but current supervision extraction methods remain ineffici

tutorialsarxiv-cs-ai
14 Apr 2026
Research

PointSplat: Efficient Geometry-Driven Pruning and Transformer Refinement for 3D Gaussian Splatting

DGX agent

arXiv:2604.09903v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) has recently unlocked real-time, high-fidelity novel view synthesis by representing scenes using explicit 3D primitives. Ho

researcharxiv-cs-cv
14 Apr 2026
Research

RECIPER: A Dual-View Retrieval Pipeline for Procedure-Oriented Materials Question Answering

DGX agent

arXiv:2604.11229v1 Announce Type: cross Abstract: Retrieving procedure-oriented evidence from materials science papers is difficult because key synthesis details are often scattered across long, conte

researcharxiv-cs-ai
14 Apr 2026
Research

Sat2Sound: A Unified Framework for Zero-Shot Soundscape Mapping

DGX agent

arXiv:2505.13777v2 Announce Type: replace-cross Abstract: We present Sat2Sound, a unified multimodal framework for geospatial soundscape understanding, designed to predict and map the distribution of

researcharxiv-cs-ai
14 Apr 2026
Research

SyncFix: Fixing 3D Reconstructions via Multi-View Synchronization

DGX agent

arXiv:2604.11797v1 Announce Type: new Abstract: We present SyncFix, a framework that enforces cross-view consistency during the diffusion-based refinement of reconstructed scenes. SyncFix formulates r

researcharxiv-cs-cv
14 Apr 2026
Model Releases

Training-Free Object-Background Compositional T2I via Dynamic Spatial Guidance and Multi-Path Pruning

DGX agent

arXiv:2604.09850v1 Announce Type: new Abstract: Existing text-to-image diffusion models, while excelling at subject synthesis, exhibit a persistent foreground bias that treats the background as a pass

model-releasesarxiv-cs-cv
14 Apr 2026
Research

Unfolding 3D Gaussian Splatting via Iterative Gaussian Synopsis

DGX agent

arXiv:2604.11685v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) has become a state-of-the-art framework for real-time, high-fidelity novel view synthesis. However, its substantial storage

researcharxiv-cs-cv
14 Apr 2026
Safety

Adversarial Concept Distillation for One-Step Diffusion Personalization

DGX agent

arXiv:2510.20512v2 Announce Type: replace Abstract: Recent progress in accelerating text-to-image diffusion models enables high-fidelity synthesis within a single denoising step. However, customizing

safetyarxiv-cs-cv
13 Apr 2026
Model Releases

CAD 100K: A Comprehensive Multi-Task Dataset for Car Related Visual Anomaly Detection

DGX agent

arXiv:2604.09023v1 Announce Type: new Abstract: Multi-task visual anomaly detection is critical for car-related manufacturing quality assessment. However, existing methods remain task-specific, hinder

model-releasesarxiv-cs-cv
13 Apr 2026
Research

DDSP-QbE++: Improving Speech Quality for Speech Anonymisation for Atypical Speech

DGX agent

arXiv:2604.09246v1 Announce Type: cross Abstract: Differentiable Digital Signal Processing (DDSP) pipelines for voice conversion rely on subtractive synthesis, where a periodic excitation signal is sh

researcharxiv-cs-ai
13 Apr 2026
Model Releases

Hierarchical SVG Tokenization: Learning Compact Visual Programs for Scalable Vector Graphics Modeling

DGX agent

arXiv:2604.05072v2 Announce Type: replace Abstract: Recent large language models have shifted SVG generation from differentiable rendering optimization to autoregressive program synthesis. However, ex

model-releasesarxiv-cs-lg
13 Apr 2026
Research

PRADA: Probability-Ratio-Based Attribution and Detection of Autoregressive-Generated Images

DGX agent

arXiv:2511.20068v2 Announce Type: replace Abstract: Autoregressive (AR) image generation has recently emerged as a powerful paradigm for image synthesis. Leveraging the generation principle of large l

researcharxiv-cs-cv
13 Apr 2026
Tutorials

Cross-Modal Emotion Transfer for Emotion Editing in Talking Face Video

DGX agent

arXiv:2604.07786v1 Announce Type: new Abstract: Talking face generation has gained significant attention as a core application of generative models. To enhance the expressiveness and realism of synthe

tutorialsarxiv-cs-cv
10 Apr 2026
Safety

Harnessing Hyperbolic Geometry for Harmful Prompt Detection and Sanitization

DGX agent

arXiv:2604.06285v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have become essential for tasks such as image synthesis, captioning, and retrieval by aligning textual and visual inform

safetyarxiv-cs-ai
10 Apr 2026
Research

SurfelSplat: Learning Efficient and Generalizable Gaussian Surfel Representations for Sparse-View Surface Reconstruction

DGX agent

arXiv:2604.08370v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) has demonstrated impressive performance in 3D scene reconstruction. Beyond novel view synthesis, it shows great potential f

researcharxiv-cs-cv
10 Apr 2026
Safety

When Numbers Speak: Aligning Textual Numerals and Visual Instances in Text-to-Video Diffusion Models

DGX agent

arXiv:2604.08546v1 Announce Type: new Abstract: Text-to-video diffusion models have enabled open-ended video synthesis, but often struggle with generating the correct number of objects specified in a

safetyarxiv-cs-cv
10 Apr 2026
Agents

3D Scene Generation: A Survey

DGX agent

arXiv:2505.05474v2 Announce Type: replace Abstract: 3D scene generation seeks to synthesize spatially structured, semantically meaningful, and photorealistic environments for applications such as imme

agentsarxiv-cs-cv
13 Aug 2026
Agents

LLMs in Process Diagram Engineering: From Optimal PFDs to Validated P&IDs

DGX agent

arXiv:2608.11220v1 Announce Type: new Abstract: Nowadays, the creation of a process flow diagram (PFD) and its subsequent transformation into a piping and instrumentation diagram (P&ID) is predominant

agentsarxiv-cs-ai
13 Aug 2026
Tutorials

Self-Supervised Weighted Image Guided Quantitative MRI Super-Resolution

DGX agent

arXiv:2512.17612v2 Announce Type: replace Abstract: Object: To present and evaluate Self-supervised Weighted Image Guided quantitative MRI Super-Resolution (SWIG qMRI SR), a physics-informed framework

tutorialsarxiv-cs-cv
13 Aug 2026
Research

TGRHuman: Text-Guided Realistic 3D Human Generation via Diffusion Renderer

DGX agent

arXiv:2608.12175v1 Announce Type: new Abstract: Realistic 3D human generation plays a crucial role in many graphics applications. However, current methods still struggle to generate high-quality human

researcharxiv-cs-cv
13 Aug 2026
Research

Beyond Pixels: From Video Priors to 4D Worlds

DGX agent

arXiv:2608.10744v1 Announce Type: new Abstract: 4D generation synthesizes dynamic 3D scenes from conditions such as text or images. Existing methods either reconstruct generated RGB videos with a sepa

researcharxiv-cs-cv
12 Aug 2026
Model Releases

Cost-Efficient Estimation of General Abilities Across Benchmarks

DGX agent

arXiv:2604.01418v2 Announce Type: replace Abstract: Thousands of diverse benchmarks have been developed to measure the quality of large language models (LLMs). Yet prior work has demonstrated that LLM

model-releasesarxiv-cs-cl
12 Aug 2026
Local Ai

Easy3D-Labels: Supervising Semantic Occupancy Estimation with 3D Pseudo-Labels for Automotive Perception

DGX agent

arXiv:2509.26087v5 Announce Type: replace Abstract: In perception for automated vehicles, safety is critical not only for the driver but also for other agents in the scene, particularly vulnerable roa

local-aiarxiv-cs-cv
12 Aug 2026
Research

Emergent Neural Network Mechanisms for Generalization to Objects in Novel Orientations

DGX agent

arXiv:2109.13445v3 Announce Type: replace-cross Abstract: The capability of Deep Neural Networks (DNNs) to recognize objects in orientations outside the distribution of the training data is not well u

researcharxiv-cs-ai
12 Aug 2026
Local Ai

ENCORE: Efficient Noise Context-Aware Representation for Low-Dose CT Denoising

DGX agent

arXiv:2608.10343v1 Announce Type: new Abstract: While deep learning-based denoising has become widely adopted in low-dose CT, conventional models use generic architectures designed for natural images,

local-aiarxiv-cs-cv
12 Aug 2026
Research

Longitudinal 3D Foundation Modeling for Neoadjuvant Breast Cancer Response Prediction from Serial DCE-MRI

DGX agent

arXiv:2608.09991v1 Announce Type: cross Abstract: Pathologic complete response (pCR) is an important endpoint in neoadjuvant chemotherapy (NAC) for breast cancer, and predicting pCR from imaging durin

researcharxiv-cs-cv
12 Aug 2026
Tutorials

MRIComp4Flow: Compression of 3D Brain MRI for Training Multi-Modal Generative Models

DGX agent

arXiv:2608.10291v1 Announce Type: cross Abstract: Large-scale multi-modal MRI datasets impose substantial storage and I/O costs, limiting the training of 3D generative models on commodity infrastructu

tutorialsarxiv-cs-ai
12 Aug 2026
Agents

When Agent Automation Becomes Profitable: Quantifying and Insuring Autonomous AI Risk through Trace-Economic Underwriting

DGX agent

arXiv:2606.16465v2 Announce Type: replace Abstract: AI agents can now take irreversible actions in operational systems, but agent-caused losses are still not clearly assigned, priced, or transferred.

agentsarxiv-cs-ai
12 Aug 2026
Research

Alpha as an Efficiency Signal: Visibility-Routed RGBA Image-to-Video Generation

DGX agent

arXiv:2608.09355v1 Announce Type: new Abstract: RGBA videos combine RGB appearance with an alpha channel, enabling animated assets to be applied across arbitrary backgrounds, which are heavily used in

researcharxiv-cs-cv
11 Aug 2026
Research

AMD:Anatomical Motion Diffusion with Interpretable Motion Decomposition and Fusion

DGX agent

arXiv:2312.12763v3 Announce Type: replace Abstract: Generating realistic human motion sequences from text descriptions is a challenging task that requires capturing the rich expressiveness of both nat

researcharxiv-cs-cv
11 Aug 2026
← Previous
1…1920212223…60
Next →