AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “synthesis”

GridTimelineEvolution
2,818 results
5 May 2026

Application Research of a Deep Learning Model Integrating CycleGAN and YOLO in PCB Infrared Defect Detection

ResearchDGX agent

arXiv:2601.00237v2 Announce Type: replace Abstract: This paper addresses the critical bottleneck of infrared (IR) data scarcity in Printed Circuit Board (PCB) defect detection by proposing a cross-mod

ARGUS: Policy-Adaptive Ad Governance via Evolving Reinforcement with Adversarial Umpiring

SafetyDGX agent

arXiv:2605.02200v1 Announce Type: new Abstract: Online advertising governance faces significant challenges due to the non-stationary nature of regulatory policies, where emerging mandates (e.g., restr

Balalaika: Data-Centric, Prosody-Aware Annotation Pipeline for Russian Speech

ResearchDGX agent

arXiv:2507.13563v2 Announce Type: replace Abstract: We introduce Balalaika, an open-source, data-centric pipeline for processing audio and producing prosody-aware annotations. It combines semantic VAD


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Behavior-Grounded Lane Representation Learning for Multi-Task Traffic Digital Twins

SafetyDGX agent

arXiv:2605.01901v1 Announce Type: new Abstract: Traffic digital twins are powerful tools for advanced traffic management, and most systems are built on static geometric representations. However, these

Compositional Multi-hop Factual Error Correction via Decomposition-and-Injection

ApplicationsDGX agent

arXiv:2605.02277v1 Announce Type: new Abstract: Factual Error Correction (FEC) aims to revise inaccurate text into statements that are factually consistent with external evidence. Although recent meth

DissolveStereo: Coarse Depth Injection for Zero-Shot Stereo Video Generation

ResearchDGX agent

arXiv:2411.14295v3 Announce Type: replace Abstract: Generating high-quality stereo videos requires consistent depth perception and temporal coherence across frames. Despite advances in image and video

Embody4D: A Generalist 4D World Model for Embodied AI

ResearchDGX agent

arXiv:2605.01799v1 Announce Type: new Abstract: World models have made significant progress in modeling dynamic environments; however, most embodied world models are still restricted to 2D representat

Gen-Searcher: Reinforcing Agentic Search for Image Generation

Model ReleasesDGX agent

arXiv:2603.28767v2 Announce Type: replace Abstract: Recent image generation models have shown strong capabilities in generating high-fidelity and photorealistic images. However, they are fundamentally

How Label Imbalance Shapes Geometry: A General Spectral Analysis of Multi-Label Neural Collapse

ResearchDGX agent

arXiv:2605.01897v1 Announce Type: new Abstract: This work investigates the phenomenon of Neural Collapse (NC) in multi-label classification, extending its conceptual framework from multi-class learnin

HumanSplatHMR: Closing the Loop Between Human Mesh Recovery and Gaussian Splatting Avatar

SafetyDGX agent

arXiv:2605.02784v1 Announce Type: new Abstract: Accurately recovering human pose and appearance from video is an essential component of scene reconstruction, with applications to motion capture, motio

MedScribe: Clinically Grounded CT Reporting through Agentic Workflows

AgentsDGX agent

arXiv:2605.01779v1 Announce Type: new Abstract: Vision-language models (VLMs) have shown potential for automated radiology report generation, yet existing approaches rely on global embedding compressi

oMeBench: Towards Robust Benchmarking of LLMs in Organic Mechanism Elucidation and Reasoning

Model ReleasesDGX agent

arXiv:2510.07731v3 Announce Type: replace-cross Abstract: Organic reaction mechanisms are the stepwise elementary reactions by which reactants form intermediates and products, and are fundamental to u

OphMAE: Bridging Volumetric and Planar Imaging with a Foundation Model for Adaptive Ophthalmological Diagnosis

Model ReleasesDGX agent

arXiv:2605.02714v1 Announce Type: new Abstract: The advent of foundation models has heralded a new era in medical artificial intelligence (AI), enabling the extraction of generalizable representations

Refracting Reality: Generating Images with Realistic Transparent Objects

ResearchDGX agent

arXiv:2511.17340v3 Announce Type: replace Abstract: Generative image models can produce convincingly real images, with plausible shapes, textures, layouts and lighting. However, one domain in which th

Robo3R: Enhancing Robotic Manipulation with Accurate Feed-Forward 3D Reconstruction

SafetyDGX agent

arXiv:2602.10101v2 Announce Type: replace Abstract: 3D spatial perception is fundamental to generalizable robotic manipulation, yet obtaining reliable, high-quality 3D geometry remains challenging. De

Robust Parameter Learning for Uncertain MDPs

Model ReleasesDGX agent

arXiv:2605.01339v1 Announce Type: new Abstract: Learning-based approaches to verifying unknown Markov decision processes (MDPs) often employ uncertain MDPs. These models use, for example, confidence i

S^3-R1: Learning to Retrieve and Answer Step-by-Step with Synthetic Data

AgentsDGX agent

arXiv:2605.01248v1 Announce Type: new Abstract: Reinforcement learning (RL) post-training has enabled newer capabilities in models, such as agentic tool-use for search. However, these models struggle

Self-Supervised Spatial And Zero-Shot Angular Super-Resolution by Spatial-Angular Implicit Representation For Rotating-View SNR-Efficient Diffusion MRI

ResearchDGX agent

arXiv:2605.02575v1 Announce Type: new Abstract: Rotating-view thick-slice acquisition is highly SNR-efficient for mesoscale diffusion MRI (dMRI) but requires numerous rotating views to satisfy Nyquist

Sparse Representation Learning for Vessels

ResearchDGX agent

arXiv:2605.01382v1 Announce Type: new Abstract: Analyzing human vasculature and vessel-like, tubular structures, such as airways, is crucial for disease diagnosis and treatment. Current methods often

StyleShield: Exposing the Fragility of AIGC Detectors through Continuous Controllable Style Transfer

Model ReleasesDGX agent

arXiv:2605.00924v1 Announce Type: new Abstract: AI-generated content (AIGC) detectors are increasingly deployed in high-stakes settings such as academic integrity screening, yet their reliability rest

SwiftChannel: Algorithm-Hardware Co-Design for Deep Learning-Based 5G Channel Estimation

Model ReleasesDGX agent

arXiv:2605.01931v1 Announce Type: cross Abstract: Channel estimation is crucial in 5G communication networks for optimizing transmission parameters and ensuring reliable, high-speed communication. How

Synergistic Perception and Generative Recomposition: A Multi-Agent Orchestration for Expert-Level Building Inspection

AgentsDGX agent

arXiv:2603.20143v2 Announce Type: replace Abstract: Building facade defect inspection is fundamental to structural health monitoring and sustainable urban maintenance, yet it remains a formidable chal

The Banach-Butterfly Invariant: Influence-Adaptive Walsh Geometry for Ternary Polynomial Threshold Functions

ResearchDGX agent

arXiv:2605.01637v1 Announce Type: new Abstract: We introduce the Banach-Butterfly Invariant (BBT), an influence-adaptive Banach geometry on the Walsh-Hadamard butterfly factorization. For a Boolean fu

To Do or Not to Do: Ensuring the Safety of Visuomotor Policies Learned from Demonstrations

SafetyDGX agent

arXiv:2605.01201v1 Announce Type: new Abstract: Task success has historically been the primary measure of policy performance in imitation learning (IL) research. This characteristics strictly limits t

Vanast: Virtual Try-On with Human Image Animation via Synthetic Triplet Supervision

ResearchDGX agent

arXiv:2604.04934v2 Announce Type: replace Abstract: We present Vanast, a unified framework that generates garment-transferred human animation videos directly from a single human image, garment images,

XekRung Technical Report

TutorialsDGX agent

arXiv:2605.00072v1 Announce Type: cross Abstract: We present XekRung, a frontier large language model for cybersecurity, designed to provide comprehensive security capabilities. To achieve this, we de

4 May 2026

2D-SuGaR: Surface-Aware Gaussian Splatting for Geometrically Accurate Mesh Reconstruction

ResearchDGX agent

arXiv:2605.00569v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) has emerged as a powerful technique for generating photorealistic renderings of a scene in real-time. However, the volumetr

Certifiable Factor Graph Optimization

ResearchDGX agent

arXiv:2603.01267v2 Announce Type: replace-cross Abstract: We show that the factor graph and certifiable estimation paradigms, which have thus far been treated as essentially independent in the literat

CollaFuse: Collaborative Diffusion Models

ResearchDGX agent

arXiv:2406.14429v3 Announce Type: replace-cross Abstract: In the landscape of generative artificial intelligence, diffusion-based models have emerged as a promising method for generating synthetic ima

Generating Statistical Charts with Validation-Driven LLM Workflows

ResearchDGX agent

arXiv:2605.00800v1 Announce Type: new Abstract: Generating diverse, readable statistical charts from tabular data remains challenging for LLMs, as many failures become apparent after rendering and are

Learning from the Unseen: Generative Data Augmentation for Geometric-Semantic Accident Anticipation

Model ReleasesDGX agent

arXiv:2605.00051v1 Announce Type: new Abstract: Anticipating traffic accidents is a critical yet unresolved problem for autonomous driving, hindered by the inherent complexity of modeling interactions

Learning Multimodal Energy-Based Model with Multimodal Variational Auto-Encoder via MCMC Revision

ResearchDGX agent

arXiv:2605.00644v1 Announce Type: new Abstract: Energy-based models (EBMs) are a flexible class of deep generative models and are well-suited to capture complex dependencies in multimodal data. Howeve

MMAudio-LABEL: Audio Event Labeling via Audio Generation for Silent Video

ApplicationsDGX agent

arXiv:2605.00495v1 Announce Type: cross Abstract: Recent advances in multimodal generation have enabled high-quality audio generation from silent videos. Practical applications, such as sound producti

Physically Native World Models: A Hamiltonian Perspective on Generative World Modeling

AgentsDGX agent

arXiv:2605.00412v1 Announce Type: cross Abstract: World models have recently re-emerged as a central paradigm for embodied intelligence, robotics, autonomous driving, and model-based reinforcement lea

Reinforcement Learning with LLM-Guided Action Spaces for Synthesizable Lead Optimization

SafetyDGX agent

arXiv:2604.07669v2 Announce Type: replace Abstract: Lead optimization in drug discovery requires improving therapeutic properties while ensuring that molecular modifications correspond to feasible syn

SCAN: Structured Capability Assessment and Navigation for LLMs

ResearchDGX agent

arXiv:2505.06698v4 Announce Type: replace Abstract: Evaluating Large Language Models (LLMs) has become increasingly important, with automatic evaluation benchmarks gaining prominence as alternatives t

TimesNet-Gen: Deep Learning-based Site Specific Strong Motion Generation

Local AiDGX agent

arXiv:2512.04694v3 Announce Type: replace Abstract: Effective earthquake risk reduction relies on accurate site-specific evaluations, which require models capable of representing the influence of loca

UniVidX: A Unified Multimodal Framework for Versatile Video Generation via Diffusion Priors

SafetyDGX agent

arXiv:2605.00658v1 Announce Type: new Abstract: Recent progress has shown that video diffusion models (VDMs) can be repurposed for diverse multimodal graphics tasks. However, existing methods often tr

1 May 2026

A Pattern Language for Resilient Visual Agents

ApplicationsDGX agent

arXiv:2604.28001v1 Announce Type: new Abstract: Integrating multimodal foundation models into enterprise ecosystems presents a fundamental software architecture challenge. Architects must balance comp

AesRM: Improving Video Aesthetics with Expert-Level Feedback

Model ReleasesDGX agent

arXiv:2604.28078v1 Announce Type: new Abstract: Despite rapid advances in photorealistic video generation, real-world applications such as filmmaking require video aesthetics, e.g., harmonious colors

Aligning Perception, Reasoning, Modeling and Interaction: A Survey on Physical AI

ApplicationsDGX agent

arXiv:2510.04978v5 Announce Type: replace Abstract: The rapid advancement of embodied intelligence and world models has intensified efforts to integrate physical laws into AI systems, yet physical per

Generative Human Geometry Distribution

ResearchDGX agent

arXiv:2503.01448v5 Announce Type: replace Abstract: Realistic human geometry generation is an important yet challenging task, requiring both the preservation of fine clothing details and the accurate

Omni-Attribute: Open-vocabulary Attribute Encoder for Visual Concept Personalization

TutorialsDGX agent

arXiv:2512.10955v2 Announce Type: replace Abstract: Visual concept personalization aims to transfer only specific image attributes, such as identity, expression, lighting, and style, into unseen conte

OptimusKG: Unifying biomedical knowledge in a modern multimodal graph

AgentsDGX agent

arXiv:2604.27269v1 Announce Type: new Abstract: Biomedical knowledge graphs (KGs) are widely used in the life sciences, yet many are derived from unstructured documents and therefore lack schema-level

Progressive Multi-Agent Reasoning for Biological Perturbation Prediction

Model ReleasesDGX agent

arXiv:2602.07408v2 Announce Type: replace Abstract: Predicting gene regulation responses to biological perturbations requires reasoning about underlying biological causalities. While large language mo

Simple Self-Conditioning Adaptation for Masked Diffusion Models

ResearchDGX agent

arXiv:2604.26985v1 Announce Type: cross Abstract: Masked diffusion models (MDMs) generate discrete sequences by iterative denoising under an absorbing masking process. In standard masked diffusion, if

SQuadGen: Generating Simple Quad Layouts via Chart Distance Fields

ResearchDGX agent

arXiv:2604.27329v1 Announce Type: cross Abstract: 3D shapes from scanning, reconstruction, or AI-generated content often lack simple quad mesh layouts -- critical for efficient editing and modeling. E

Visual Generation in the New Era: An Evolution from Atomic Mapping to Agentic World Modeling

Model ReleasesDGX agent

arXiv:2604.28185v1 Announce Type: new Abstract: Recent visual generation models have made major progress in photorealism, typography, instruction following, and interactive editing, yet they still str

30 Apr 2026

AdaMem: Adaptive User-Centric Memory for Long-Horizon Dialogue Agents

Model ReleasesDGX agent

arXiv:2603.16496v2 Announce Type: replace Abstract: Large language model (LLM) agents increasingly rely on external memory to support long-horizon interaction, personalized assistance, and multi-step

Deterministic Legal Agents: A Canonical Primitive API for Auditable Reasoning over Temporal Knowledge Graphs

Model ReleasesDGX agent

arXiv:2510.06002v3 Announce Type: replace Abstract: In high-stakes legal domains, retrieval must preserve not only semantic relevance, but also the hierarchy, temporality, and causal provenance of leg

EmoTransCap: Dataset and Pipeline for Emotion Transition-Aware Speech Captioning in Discourses

AgentsDGX agent

arXiv:2604.26417v1 Announce Type: new Abstract: Emotion perception and adaptive expression are fundamental capabilities in human-agent interaction. While recent advances in speech emotion captioning (

FLARE: Fully Integration of Vision-Language Representations for Deep Cross-Modal Understanding

SafetyDGX agent

arXiv:2504.09925v3 Announce Type: replace Abstract: We introduce FLARE, a family of vision language models (VLMs) with a fully vision-language alignment and integration paradigm. Unlike existing appro

Hybrid Diffusion for Simultaneous Symbolic and Continuous Planning

ResearchDGX agent

arXiv:2509.21983v2 Announce Type: replace-cross Abstract: Constructing robots to accomplish long-horizon tasks is a long-standing challenge within artificial intelligence. Approaches using generative

Inferix: A Block-Diffusion based Next-Generation Inference Engine for World Simulation

Model ReleasesDGX agent

arXiv:2511.20714v2 Announce Type: replace-cross Abstract: World models serve as core simulators for fields such as agentic AI, embodied AI, and gaming, capable of generating long, physically realistic

Learning to Rewrite Tool Descriptions for Reliable LLM-Agent Tool Use

AgentsDGX agent

arXiv:2602.20426v2 Announce Type: replace Abstract: While most efforts to improve LLM-based tool-using agents focus on the agent itself - through larger models, better prompting, or fine-tuning - agen

Planar Gaussian Splatting with Bilinear Spatial Transformer for Wireless Radiance Field Reconstruction

TutorialsDGX agent

arXiv:2604.25945v1 Announce Type: cross Abstract: Wireless radiance field (WRF) reconstruction aims to learn a continuous, queryable representation of radio frequency characteristics over 3D space and

Probe-then-Plan: Environment-Aware Planning for Industrial E-commerce Search

SafetyDGX agent

arXiv:2603.15262v2 Announce Type: replace Abstract: Modern e-commerce search is evolving to resolve complex user intents. While Large Language Models (LLMs) offer strong reasoning, existing LLM-based

SciMDR: Advancing Scientific Multimodal Document Reasoning

Model ReleasesDGX agent

arXiv:2603.12249v2 Announce Type: replace-cross Abstract: Constructing scientific multimodal document reasoning datasets for foundation model training involves an inherent trade-off among scale, faith

SynSur: An end-to-end generative pipeline for synthetic industrial surface defect generation and detection

ResearchDGX agent

arXiv:2604.26633v1 Announce Type: cross Abstract: The bottleneck in learning-based industrial defect detection is often limited not by model capacity, but by the scarcity of labeled defect data: defec

The devil is in the details: Enhancing Video Virtual Try-On via Keyframe-Driven Details Injection

ResearchDGX agent

arXiv:2512.20340v3 Announce Type: replace Abstract: Although diffusion transformer (DiT)-based video virtual try-on (VVT) has made significant progress in synthesizing realistic videos, existing metho

← Previous
1…3839404142…47
Next →