AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “synthesis”

GridTimelineEvolution
2,818 results
28 May 2026

Sketch2Motion: Text-driven 2D Sketch to 3D Animation via Diffusion-guided Skeleton Optimization

TutorialsDGX agent

arXiv:2605.28394v1 Announce Type: new Abstract: Animation of 2D hand-drawn sketches provides an effective medium for visual communication. However, these sketches pose challenges, particularly in hand

SmartDirector: Keyframe-Conditioned Cinematic Video Generation with Narrative Pacing Control

ResearchDGX agent

arXiv:2605.27891v1 Announce Type: cross Abstract: The narrative quality of a video fundamentally determines its perceptual value. Although existing video generation methods can produce visually appeal

Smoothed Score Queries and the Complexity of Sampling

ResearchDGX agent

arXiv:2605.27769v1 Announce Type: cross Abstract: We study the query complexity of sampling from high-dimensional Gaussian distributions using gradient information. In the standard oracle model, exact


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
27 May 2026

ADRD-Bench: A Preliminary LLM Benchmark for Alzheimer's Disease and Related Dementias

Model ReleasesDGX agent

arXiv:2602.11460v2 Announce Type: replace Abstract: Large language models (LLMs) have shown great potential for healthcare applications. However, existing evaluation benchmarks provide minimal coverag

AI Agent for Reverse-Engineering Legacy Finite-Difference Code and Translating to Devito

AgentsDGX agent

arXiv:2601.18381v2 Announce Type: replace Abstract: To facilitate the transformation of legacy finite difference implementations into the Devito environment, this study develops an integrated AI agent

Aligning Few-Step Generative Models by Amortizing Sample-based Variational Inference

SafetyDGX agent

arXiv:2605.26552v1 Announce Type: cross Abstract: Aligning a few-step generative model is challenging, since existing alignment frameworks typically rely on restrictive assumptions: a tractable likeli

Alignment Tuning for Large Language Models: A Data-Centric Lens on Alignment Data Pipelines

SafetyDGX agent

arXiv:2605.26442v1 Announce Type: cross Abstract: Much of the alignment tuning literature is organized around optimization objectives, while the construction of alignment data is often treated implici

DGLD: Domain-Gated Latent Diffusion for the Discovery of Novel Energetic Materials

Model ReleasesDGX agent

arXiv:2605.26540v1 Announce Type: cross Abstract: Energetic-materials performance gains translate directly into reduced propellant mass, smaller warheads, and more efficient civilian gas-generators, y

DSA-Tokenizer: Disentangled Semantic-Acoustic Tokenization via Flow Matching-based Hierarchical Fusion

ResearchDGX agent

arXiv:2601.09239v3 Announce Type: replace-cross Abstract: Speech tokenizers are a key building block of fully discrete Speech LLMs. Existing tokenizers either prioritize semantic encoding, fuse semant

DuoGesture: Neuro-Inspired and Biomechanically Informed Dual-Stream Co-Speech Gesture Generation

SafetyDGX agent

arXiv:2605.26236v1 Announce Type: new Abstract: Co-speech gesture generation requires both semantic expressivity and biomechanically plausible rhythmic motion. Existing holistic gesture models mix lex

E^3C: Video Generation with 3D Environmental Memory and Ego-Exo Human Pose Control

ResearchDGX agent

arXiv:2605.26316v1 Announce Type: cross Abstract: Controllable and physically grounded egocentric video generation is essential for embodied agents to reason about how their own and others' actions ma

Eroding Trust in Real Speech: A Large-Scale Study of Human Audio Deepfake Perception

ResearchDGX agent

arXiv:2605.26136v1 Announce Type: cross Abstract: Audio deepfakes have improved rapidly recently, yet their effect on human trust in real speech remains unstudied. We present the largest listening stu

FAB-Bench: A Framework for Adaptive RAG Benchmarking in Semiconductor Manufacturing

Model ReleasesDGX agent

arXiv:2605.26476v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) has become critical for knowledge-intensive applications, yet evaluating its performance in vertical domains remain

HyperSim: A Holistic Sim-To-Real Framework For Robust Robotic Manipulation

SafetyDGX agent

arXiv:2605.26638v1 Announce Type: new Abstract: Scaling data volume and diversity is critical for generalizing embodied intelligence. While synthetic data generation offers a scalable alternative to e

Morphling: Fast, Fused, and Flexible GNN Training at Scale

HardwareDGX agent

arXiv:2512.01678v5 Announce Type: replace Abstract: Graph Neural Networks (GNNs) present a fundamental hardware challenge by fusing irregular, memory-bound graph traversals with regular, compute-inten

MRT: Masked Region Transformer for Layered Image Generation and Editing at Scale

Model ReleasesDGX agent

arXiv:2605.27235v1 Announce Type: new Abstract: Layered image generation and editing is a fundamental capability that enables layer-wise reuse, editing, and composition of generated visual content, an

MULTISEISMO: A Multimodal Seismic Dataset and Model for Cross-Modal Seismic Understanding

Model ReleasesDGX agent

arXiv:2605.26320v1 Announce Type: cross Abstract: The application of generalist multimodal models (GMMs) to specialized scientific domains remains limited due to the scarcity of comprehensive domain-s

New skill from K-Dense: LiteParse in Scientific Agent Skills — built for fast, local research paper ingestion. Your AI co-scientist can now:…

Local AiDGX agent

New skill from K-Dense: LiteParse in Scientific Agent Skills — built for fast, local research paper ingestion. Your AI co-scientist can now: * Parse PDFs and supplementary files on your machine (no do

PashtoTTS-Bench: automated screening for low-resource non-Latin-script text-to-speech

Model ReleasesDGX agent

arXiv:2605.26978v1 Announce Type: new Abstract: Text-to-speech (TTS) evaluation for low-resource non-Latin-script languages can fail when it relies on a single ASR round-trip word error rate (WER). A

Personalized Generative Models for Contextual Debiasing

TutorialsDGX agent

arXiv:2605.26353v1 Announce Type: cross Abstract: Different visual patterns appear with different frequencies in the world: e.g., beach balls appear on sand more often than they do on a road. These st

Pusa V1.0: Unlocking Temporal Control in Pretrained Video Diffusion Models via Vectorized Timestep Adaptation

ResearchDGX agent

arXiv:2507.16116v2 Announce Type: replace Abstract: The rapid advancement of video diffusion models has been hindered by fundamental limitations in temporal modeling, particularly the rigid synchroniz

RadarSim: Simulating Single-Chip Radar via Multimodal Neural Fields

ResearchDGX agent

arXiv:2605.26328v1 Announce Type: new Abstract: Radars are an ideal complement to cameras: both are inexpensive, solid-state sensors, with cameras offering fine angular resolution, while radars provid

Slide Deck Q&A Quality Assurance App: A Multi-Stage Pipeline for Pedagogical Question Generation

ResearchDGX agent

arXiv:2605.26428v1 Announce Type: new Abstract: Generating high-quality, pedagogically useful questions from lecture slide decks is difficult because important instructional content is distributed acr

SOLE-R1: Video-Language Reasoning as the Sole Reward for On-Robot Reinforcement Learning

Model ReleasesDGX agent

arXiv:2603.28730v2 Announce Type: replace-cross Abstract: Vision-language models (VLMs) have shown impressive capabilities across diverse tasks, motivating efforts to leverage these models to supervis

Underwater360: Reconstructing Underwater Scenes from Panoramic Images with Omnidirectional Gaussian Splatting

Model ReleasesDGX agent

arXiv:2605.26447v1 Announce Type: new Abstract: Underwater scene reconstruction is essential for immersive exploration of aquatic environments, yet remains challenging due to complex participating-med

VERA-V: Variational Inference Framework for Jailbreaking Vision-Language Models

SafetyDGX agent

arXiv:2510.17759v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) extend large language models with visual reasoning, but their multimodal design also introduces new, underexplor

VesselSim: learning 3D blood vessel segmentation without expert annotations

ApplicationsDGX agent

arXiv:2605.26277v1 Announce Type: cross Abstract: Blood vessel segmentation is a core task in medical image analysis for the care of vascular diseases and surgical planning, yet the challenges of prov

26 May 2026

AnyScene: Towards Highly Controllable Driving Scene Generation at Anywhere and Beyond

SafetyDGX agent

arXiv:2605.26113v1 Announce Type: new Abstract: Generating high-fidelity and controllable synthetic data is critical for advancing end-to-end autonomous driving, particularly for addressing the long t

Beyond Inference-Only Deployment: Comparing Weight-Based Consolidation Against Cascading Compaction

Model ReleasesDGX agent

arXiv:2605.24657v1 Announce Type: new Abstract: Major LLM platforms deploy models in an inference-only configuration: the model serves requests but never updates per-user weights. Users must repeatedl

BODHI: Precise OS Kernel Specification Inference

Model ReleasesDGX agent

arXiv:2605.23931v1 Announce Type: new Abstract: The formal verification of operating system kernels requires precise specifications that capture the intended behavior of system calls. Writing these sp

Cross-Domain Generalization Limits of Vision Foundation Models in Facial Deepfake Detection

Model ReleasesDGX agent

arXiv:2605.24965v1 Announce Type: cross Abstract: The rapid evolution of generative models has enabled the creation of hyper-realistic facial deepfakes, exposing a critical vulnerability in modern dig

CUA-Gym: Scaling Verifiable Training Environments and Tasks for Computer-Use Agents

Model ReleasesDGX agent

arXiv:2605.25624v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has driven breakthroughs in domains such as math, tool-use, and software engineering, yet its exte

D3S2: Diffusion-Guided Dataset Distillation for Semantic Segmentation

SafetyDGX agent

arXiv:2605.25022v1 Announce Type: cross Abstract: Dataset distillation (DD) aims to compress large-scale datasets into compact synthetic sets while preserving training efficacy. However, existing stud

Designing Singing Syllabi with Virtual Avatars: AI-Assisted Syllabus Reauthoring

TutorialsDGX agent

arXiv:2508.11872v3 Announce Type: replace-cross Abstract: Traditional syllabi often function as static reference documents rather than engaging introductions to a course. In practical teaching, we obs

Energy Shields for Fairness

SafetyDGX agent

arXiv:2605.24926v1 Announce Type: new Abstract: Runtime fairness is not a one-time constraint but a dynamic property evaluated over a sequence of decisions. To ensure fairness at runtime, it is necess

Geo-Expert: Towards Expert-Level Geological Reasoning via Parameter-Efficient Fine-Tuning

Model ReleasesDGX agent

arXiv:2605.24844v1 Announce Type: new Abstract: While general-purpose Large Language Models (LLMs) applied to Geology often hallucinate when reasoning about subsurface structures and deep-time evoluti

IterInject: Indirect Prompt Injection Against LLM Agents via Feedback-Guided Iterative Optimization

Model ReleasesDGX agent

arXiv:2605.24659v1 Announce Type: new Abstract: LLM-based agents are increasingly deployed for complex tasks requiring planning, tool use, and interaction with external services. Their reliance on unt

KAME: Tandem Architecture for Enhancing Knowledge in Real-Time Speech-to-Speech Conversational AI

Model ReleasesDGX agent

arXiv:2510.02327v2 Announce Type: replace-cross Abstract: Real-time speech-to-speech (S2S) models excel at generating natural, low-latency conversational responses but often lack deep knowledge and se

MDIA: A Multi-Agent Diagnostic Intelligence Pipeline on HealthBench Professional

Model ReleasesDGX agent

arXiv:2605.24699v1 Announce Type: new Abstract: Most reported gains on agentic-LLM clinical benchmarks are often attributed to prompt engineering, yet our results suggest that larger improvements can

MIND: Multi-Scale Intent Diffusion for Text-Driven Physics-Based Humanoid Control

Model ReleasesDGX agent

arXiv:2605.26006v1 Announce Type: cross Abstract: Enabling physics-based humanoids to execute diverse behaviors from high-level textual commands remains a significant challenge. Existing methods typic

Multi-Persona Debate System for Automated Scientific Hypothesis Generation

AgentsDGX agent

arXiv:2605.23917v1 Announce Type: new Abstract: Modern scientific discovery is bottlenecked not by data scarcity, but by the inability to synthesize fragmented knowledge into actionable hypotheses. Th

Multi-view Consistent 3D Gaussian Head Avatars 'without' Multi-view Generation

SafetyDGX agent

arXiv:2605.25220v1 Announce Type: cross Abstract: High-fidelity 3D Gaussian head avatar generation is critical for applications such as AR/VR, telepresence, and digital humans. Existing methods depend

Multimodal Alignment and Preference Optimization for Zero-Shot Conditional RNA Generation

SafetyDGX agent

arXiv:2605.23961v1 Announce Type: cross Abstract: The design of RNA molecules that interact with specific proteins is a critical challenge in experimental and computational biology. Despite recent pro

OrpQuant: Geometric Orthogonal Residual Projection for Multiplier-Free Power-of-Two Transformer Quantization

Model ReleasesDGX agent

arXiv:2605.26092v1 Announce Type: cross Abstract: The deployment of Large Language Models (LLMs) and Vision Transformers (ViTs) on edge devices is significantly constrained by memory limitations and t

Parameter-Efficient VLMs for Gastrointestinal Endoscopy: Medical Image Generation and Clinical Visual Question Answering

Model ReleasesDGX agent

arXiv:2605.24792v1 Announce Type: cross Abstract: The major limitations of gastrointestinal (GI) endoscopy AI systems arise from a shortage of annotated data, strict privacy policies, and significant

PRIMA: Operational Patterns for Resilient Multi-Agent Research with Verifiable Identity and Convergent Feedback

AgentsDGX agent

arXiv:2605.24775v1 Announce Type: new Abstract: Operating LLMs as coordinated multi-agent research systems over multi-hour runs surfaces failure modes that single-shot evaluation cannot: upstream prov

RCTs & Human Uplift Studies: Methodological Challenges and Practical Solutions for Frontier AI Evaluation

ApplicationsDGX agent

arXiv:2603.11001v2 Announce Type: replace-cross Abstract: Human uplift studies, or studies that measure the effects of AI access on human performance via randomized controlled trials (RCT) or similar

Re-defining Humor Data Objects for AI Humor Research

ResearchDGX agent

arXiv:2605.25171v1 Announce Type: new Abstract: In most existing AI humor research, humor was treated as either 'present' or 'not present.' We explore the concept of humor as a social interaction with

SA-Kura: An Energy-Efficient Systolic Array Accelerator for Locally-Coupled Kuramoto Drift in Diffusion Sampling

HardwareDGX agent

arXiv:2605.24016v1 Announce Type: cross Abstract: Diffusion inference remains costly for edge deployment, yet existing accelerators focus almost exclusively on score networks because standard drift is

The Time is Here for Just-in-Time Systems: Challenges and Opportunities

Model ReleasesDGX agent

arXiv:2605.24096v1 Announce Type: cross Abstract: Core systems like key-value stores have historically taken years to build, and are designed to be general so as to amortize cost across deployments, p

25 May 2026

Agentic-VLA: Efficient Online Adaptation for Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2605.22896v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have emerged as a promising paradigm for robotic manipulation by leveraging pre-trained vision-language representa

Diffusion and Flow Matching Models for Tabular Data: A Survey

SafetyDGX agent

arXiv:2502.17119v2 Announce Type: replace-cross Abstract: Deep generative models have made rapid progress in image, text, audio, and video generation, and are increasingly being applied to structured

Efficient One-Step Diffusion Restoration Model with Compact Token Compression and Linear Attention

Model ReleasesDGX agent

arXiv:2605.23451v1 Announce Type: new Abstract: Real-world image super-resolution aims to recover high-quality images from complex and unknown real-world degradations. However, existing generative Rea

Knowledge Distillation for Low-Resource Open-source Text-to-SQL Model

ApplicationsDGX agent

arXiv:2605.22843v1 Announce Type: new Abstract: Text-to-SQL converts natural language questions into executable SQL queries, enabling non-technical users to access relational databases for analytics a

LangFlash: Feed-forward 3D Language Gaussian Splatting from Sparse Unposed Images

ResearchDGX agent

arXiv:2605.23287v1 Announce Type: new Abstract: We present LangFlash, a feed-forward framework for 3D Language Gaussian Splatting that reconstructs 3D scenes parameterized by Gaussian primitives enric

R^3L: Reflect-then-Retry Reinforcement Learning with Language-Guided Exploration, Pivotal Credit, and Positive Amplification

Model ReleasesDGX agent

arXiv:2601.03715v2 Announce Type: replace-cross Abstract: Reinforcement learning drives recent advances in LLM reasoning and agentic capabilities, yet current approaches struggle with both exploration

RiGS: Rigid-aware 4D Gaussian Splatting from a Single Monocular Video

TutorialsDGX agent

arXiv:2605.23672v1 Announce Type: new Abstract: Reconstructing dynamic 3D scenes from monocular videos is a fundamental yet highly challenging task, as real-world motions often involve both long-term

SciAtlas: A Large-Scale Knowledge Graph for Automated Scientific Research

Model ReleasesDGX agent

arXiv:2605.22878v1 Announce Type: new Abstract: The exponential growth of global academic output has confronted researchers and AI agents with an unprecedented ``information explosion,'' where fragmen

VINS-120K: Ultra High-Resolution Image Editing with A Large-Scale Dataset

Model ReleasesDGX agent

arXiv:2605.23518v1 Announce Type: new Abstract: Directly editing ultra-high-resolution (UHR) images is valuable but underexplored, primarily due to the lack of high-quality data and the challenge in m

23 May 2026

Beyond One-Size-Fits-All: Adaptive Subgraph Denoising for Zero-Shot Graph Learning with Large Language Models

SafetyDGX agent

arXiv:2603.02938v2 Announce Type: replace Abstract: Graph-based tasks in the zero-shot setting remain a significant challenge due to data scarcity and the inability of traditional Graph Neural Network

← Previous
1…3334353637…47
Next →