AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “synthesis”

GridTimelineEvolution
2,818 results
23 May 2026

Live Music Diffusion Models: Efficient Fine-Tuning and Post-Training of Interactive Diffusion Music Generators

SafetyDGX agent

arXiv:2605.22717v1 Announce Type: cross Abstract: Interactive streaming music generation promises the use of generative models for live performance and co-creation that is impossible with offline mode

TimeGuard: Channel-wise Pool Training for Backdoor Defense in Time Series Forecasting

ResearchDGX agent

arXiv:2605.22365v1 Announce Type: cross Abstract: Time Series Forecasting (TSF) plays a critical role across many domains, yet it is vulnerable to backdoor attacks. However, backdoor defenses tailored

22 May 2026

ACC: Compiling Agent Trajectories for Long-Context Training


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents
DGX agent

arXiv:2605.21850v1 Announce Type: new Abstract: Recent development of agents has renewed demand for long-context reasoning capacity of LLMs. However, training LLMs for this capacity requires costly lo

AgroTools: A Benchmark for Tool-Augmented Multimodal Agents in Agriculture

Model ReleasesDGX agent

arXiv:2605.22366v1 Announce Type: new Abstract: Agricultural decision-making increasingly requires multimodal systems that can transform visual observations into reliable, executable actions. However,

Cell Phantom Video Generation in Elliptical Fourier Descriptor Domain

ResearchDGX agent

arXiv:2605.22563v1 Announce Type: new Abstract: Training Deep Neural Networks for tracking individual cells in biomedical videos requires a large amount of annotated data. The annotation of videos for

EasyVFX: Frequency-Driven Decoupling for Resource-Efficient VFX Generation

HardwareDGX agent

arXiv:2605.22051v1 Announce Type: new Abstract: Generating high-fidelity visual effects (VFX) typically demands massive datasets and prohibitive computational power due to the intricate coupling of sp

Evaluating Commercial AI Chatbots as News Intermediaries

Model ReleasesDGX agent

arXiv:2605.22785v1 Announce Type: new Abstract: AI chatbots are rapidly shaping how people encounter the news, yet no prior study has systematically measured how accurately these systems, with their p

EventGait: Towards Robust Gait Recognition with Event Streams

Model ReleasesDGX agent

arXiv:2605.22139v1 Announce Type: new Abstract: Gait recognition enables non-intrusive, privacy-preserving identification but suffers in uncontrolled environments due to illumination and motion sensit

Guided Trajectory Optimization with Sparse Scaling for Test-Time Diffusion

ResearchDGX agent

arXiv:2605.21907v1 Announce Type: new Abstract: The efficient Test-Time Scaling (TTS) paradigm offers a promising perspective for enhancing the generation performance of diffusion models. However, cur

OSCToM: RL-Guided Adversarial Generation for High-Order Theory of Mind

Model ReleasesDGX agent

arXiv:2605.20423v1 Announce Type: new Abstract: Large Language Models (LLMs) perform well on many language tasks, but their Theory of Mind (ToM) reasoning is still uneven in complex social settings. E

Physiology and Anatomy Aware Inverse Inference of Myocardial Infarction for Cardiac Digital Twin

Model ReleasesDGX agent

arXiv:2605.22044v1 Announce Type: new Abstract: Accurate localization of myocardial infarction is essential for risk stratification. While LGE-MRI remains the gold standard, it is resource-intensive.

Quantifying Full-Body Immersion

ResearchDGX agent

arXiv:2605.22521v1 Announce Type: new Abstract: Humanity is at the forefront of yet another digital revolution, where the lines between real and virtual worlds are dissolving, reshaping how we perceiv

SEGA: Spectral-Energy Guided Attention for Resolution Extrapolation in Diffusion Transformers

ResearchDGX agent

arXiv:2605.22668v1 Announce Type: new Abstract: Diffusion transformers (DiTs) have emerged as a dominant architecture for text-to-image generation, yet their performance drops when generating at resol

SpaceDG: Benchmarking Spatial Intelligence under Visual Degradation

Model ReleasesDGX agent

arXiv:2605.22536v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have made rapid progress in spatial intelligence, yet existing spatial reasoning benchmarks largely assume pr

Video as Natural Augmentation: Towards Unified AI-Generated Image and Video Detection

ResearchDGX agent

arXiv:2605.21977v1 Announce Type: new Abstract: AI-generated content (AIGC) is rapidly improving, creating an urgent need for detectors that generalize across data sources, deployment pipelines, and v

Video-o3: Native Interleaved Clue Seeking for Long Video Multi-Hop Reasoning

ResearchDGX agent

arXiv:2601.23224v2 Announce Type: replace Abstract: Existing multimodal large language models for long-video understanding predominantly rely on uniform sampling and single-turn inference, limiting th

21 May 2026

CHOIR: Contact-aware 4D Hand-Object Interaction Reconstruction

ResearchDGX agent

arXiv:2605.20992v1 Announce Type: new Abstract: We ask whether everyday open-world monocular videos can be turned into reusable 4D interaction primitives: articulated hand motion, object shape with 6D

Depth Completion in Unseen Field Robotics Environments Using Extremely Sparse Depth Measurements

HardwareDGX agent

arXiv:2602.03209v2 Announce Type: replace Abstract: Autonomous field robots operating in unstructured environments require robust perception to ensure safe and reliable operations. Recent advances in

Domain-Adaptable Reinforcement Learning for Code Generation with Dense Rewards

SafetyDGX agent

arXiv:2605.21180v1 Announce Type: new Abstract: Large language models show strong potential for automated code generation, but lack guarantees for correctness, quality, safety, and domain-specific con

Enhancing Speech Large Language Models through Reinforced Behavior Alignment

SafetyDGX agent

arXiv:2509.03526v2 Announce Type: replace Abstract: The recent advancements of Large Language Models (LLMs) have spurred considerable research interest in extending their linguistic capabilities beyon

Generation of Heterogeneous PET Images from Uniform Organ Activity Maps Using a Pretrained Domain-Adapted Diffusion Model

ResearchDGX agent

arXiv:2605.20267v1 Announce Type: new Abstract: Synthetic PET images are valuable for quantitative imaging workflow development, scalable virtual imaging trials, and deep learning model training, but

HiRes: Inspectable Precedent Memory for Reaction Condition Recommendation

ResearchDGX agent

arXiv:2605.21420v1 Announce Type: new Abstract: Reaction condition recommendation sits immediately after retrosynthetic disconnection selection, and in practice, chemists require both accurate predict

Learning First Integrals via Backward-Generated Data and Guided Reinforcement Learning

ResearchDGX agent

arXiv:2605.21160v1 Announce Type: new Abstract: The discovery of first integrals is of fundamental scientific importance for understanding conservation laws in dynamical systems. However, existing sym

OpenSeisML: Open Large-Scale Real Seismic and well-log Dataset for Generative AI

ResearchDGX agent

arXiv:2605.20539v1 Announce Type: new Abstract: The advent of machine learning (ML) and computer vision has significantly accelerated seismic inversion workflows by reducing the computational cost of

PlanningBench: Generating Scalable and Verifiable Planning Data for Evaluating and Training Large Language Models

Model ReleasesDGX agent

arXiv:2605.20873v1 Announce Type: cross Abstract: Planning is a fundamental capability for large language models (LLMs) because such complex tasks require models to coordinate goals, constraints, reso

Q-SYNTH: Hybrid Quantum-Classical Adversarial Augmentation for Imbalanced Fraud Detection

ResearchDGX agent

arXiv:2605.21164v1 Announce Type: new Abstract: Credit card fraud detection is fundamentally challenged by extreme class imbalance, where fraudulent transactions are rare yet operationally critical. T

RoPeSLR: 3D RoPE-driven Sparse-LowRank Attention for Efficient Diffusion Transformers

ResearchDGX agent

arXiv:2605.20659v1 Announce Type: new Abstract: Diffusion Transformers (DiTs) have revolutionized high-fidelity video generation, yet their O(L^2) attention complexity poses a formidable bottleneck fo

TelePhysics: Physics-Grounded Multi-Object Scene Generation from a Single Image with Real-Time Interaction

SafetyDGX agent

arXiv:2605.20290v1 Announce Type: cross Abstract: Recent generative video models achieve impressive visual quality but remain constrained by limited physical consistency and controllability. Existing

TextSculptor: Training and Benchmarking Scene Text Editing

Model ReleasesDGX agent

arXiv:2605.21090v1 Announce Type: new Abstract: Recent advances in Multimodal Large Language Models (MLLMs) and diffusion-based generative models have substantially improved prompt-driven image editin

Uni-Edit: Intelligent Editing Is A General Task For Unified Model Tuning

ResearchDGX agent

arXiv:2605.21487v1 Announce Type: new Abstract: Currently, enhancing Unified Multimodal Models (UMMs) with image understanding, generation, and editing capabilities mainly relies on mixed multi-task t

20 May 2026

3D aperture-engineered diffractive neural networks for super-resolution electromagnetic wave computing

ResearchDGX agent

arXiv:2603.00995v2 Announce Type: replace-cross Abstract: The rapid progress in 6G communication and high-bandwidth radar has driven an unprecedented surge in the spatial density of signal sources, re

Agentic GraphRAG: Navigating Unstructured Financial Data with Collaborative AI

AgentsDGX agent

arXiv:2605.18770v1 Announce Type: cross Abstract: We present a collaborative agentic GraphRAG framework for expert analysis of commercial registry data. Public registries are often formally accessible

Are Rationales Necessary and Sufficient? Tuning LLMs for Explainable Misinformation Detection

ResearchDGX agent

arXiv:2605.19285v1 Announce Type: cross Abstract: The rapid spread of misinformation on social media platforms has become a formidable challenge. To mitigate its proliferation, Misinformation Detectio

Character-Centered Dialogue Generation from Scene-Level Prompts

TutorialsDGX agent

arXiv:2505.16819v4 Announce Type: replace Abstract: Recent advances in scene-based video generation enable coherent visual narratives from structured prompts, yet a key aspect of storytelling -- chara

CriterAlign: Criterion-Centric Rationale Alignment for Code Preference Judging

SafetyDGX agent

arXiv:2605.19665v1 Announce Type: cross Abstract: Pairwise human preference prediction is central to evaluating code-generation systems, where quality often depends on task-specific trade-offs beyond

Feed-Forward Gaussian Splatting from Sparse Aerial Views

ApplicationsDGX agent

arXiv:2605.19949v1 Announce Type: new Abstract: Reconstructing large-scale urban scenes from sparse aerial views is a crucial yet challenging task. Due to biased top-down and shallow-oblique camera po

GenAI-FDIA: Physics-Informed Generative Models for False Data Injection Attacks

ResearchDGX agent

arXiv:2605.18873v1 Announce Type: cross Abstract: Training and evaluating false data injection attack (FDIA) detectors for power systems is constrained by data scarcity. Operational grid measurements

GRLoc: Geometric Representation Regression for Visual Localization

Local AiDGX agent

arXiv:2511.13864v2 Announce Type: replace Abstract: Absolute Pose Regression (APR) has emerged as a compelling paradigm for visual localization. However, APR models typically operate as black boxes, d

k-Inductive Neural Barrier Certificates for Unknown Nonlinear Dynamics

SafetyDGX agent

arXiv:2605.20108v1 Announce Type: cross Abstract: While conventional (k=1) discrete-time barrier certificate conditions impose strict safety constraints by requiring the function to be non-increasing

Neural Configuration-Space Barriers for Manipulation Planning and Control

SafetyDGX agent

arXiv:2503.04929v3 Announce Type: replace-cross Abstract: Planning and control for high-dimensional robot manipulators in cluttered dynamic environments require computational efficiency and robust saf

Receptogenesis in a Vascularized Robotic Embodiment

ResearchDGX agent

arXiv:2603.09473v2 Announce Type: replace Abstract: Equipping robotic systems with the capacity to generate extit{ex novo} hardware during operation extends control of physical adaptability. Unlike mo

Search Self-play: Pushing the Frontier of Agent Capability without Supervision

Model ReleasesDGX agent

arXiv:2510.18821v3 Announce Type: replace Abstract: Reinforcement learning with verifiable rewards (RLVR) has become the mainstream technique for training LLM agents. However, RLVR highly depends on w

See MiniMax Speech 2.8 Turbo in voice finder and try the voices directly: https://voicefinder.together.ai/minimax--speech-2.8-turbo Learn mo…

TutorialsDGX agent

MiniMax Speech 2.8 Turbo is a text-to-speech model available through Together AI's voice finder tool at voicefinder.together.ai, allowing users to preview and test different voice options directly. Th

Time to REFLECT: Can We Trust LLM Judges for Evidence-based Research Agents?

Model ReleasesDGX agent

arXiv:2605.19196v1 Announce Type: new Abstract: Deep research agents increasingly automate complex information-seeking tasks, producing evidence-grounded reports via multi-step reasoning, tool use, an

Towards LLM-Assisted Architecture Recovery for Real-World ROS~2 Systems: An Agent-Based Multi-Level Approach to Hierarchical Structural Architecture Reconstruction

AgentsDGX agent

arXiv:2605.20055v1 Announce Type: cross Abstract: Explicit software architecture models are essential artifacts for communicating, analyzing, and evolving complex software-intensive systems. In ROS~2-

What Makes Synthetic Data Effective in Image Segmentation

Model ReleasesDGX agent

arXiv:2605.19289v1 Announce Type: new Abstract: Driven by rapid advances in large-scale generative models, synthetic data has emerged as a promising solution for visual understanding. While modern dif

19 May 2026

A-ProS: Towards Reliable Autonomous Programming Through Multi-Model Feedback

Model ReleasesDGX agent

arXiv:2605.18073v1 Announce Type: cross Abstract: Large Language Models (LLMs) demonstrate strong potential for automated code generation, yet their ability to iteratively refine solutions using execu

A Retrieval-Augmented Generation Approach to Extracting Algorithmic Logic from Neural Networks

ResearchDGX agent

arXiv:2512.04329v2 Announce Type: replace Abstract: Reusing existing neural-network components is central to research efficiency, yet discovering, extracting, and validating such modules across thousa

Agentic AI Governance and Lifecycle Management in Healthcare

Model ReleasesDGX agent

arXiv:2601.15630v2 Announce Type: replace Abstract: Healthcare organizations are beginning to embed agentic AI into routine workflows, including clinical documentation support and early-warning monito

AgentSteerTTS: A Multi-Agent Closed-Loop Framework for Composite-Instruction Text-to-Speech

Model ReleasesDGX agent

arXiv:2605.17583v1 Announce Type: new Abstract: While existing text-to-speech (TTS) models exhibit high expressiveness, fine-grained control over composite instructions remains challenging due to the

AtlasVid: Efficient Ultra-High-Resolution Long Video Generation via Decoupled Global-Local Modeling

Local AiDGX agent

arXiv:2605.16649v1 Announce Type: new Abstract: Recent diffusion-based video generators have achieved remarkable visual fidelity and prompt controllability, yet scaling them to ultra-high-resolution (

AutoVecCoder: Teaching LLMs to Generate Explicitly Vectorized Code

ResearchDGX agent

arXiv:2605.17978v1 Announce Type: new Abstract: Vectorization via Single Instruction, Multiple Data (SIMD) architectures is a cornerstone of high-performance computing. To fully exploit hardware poten

Beyond Patches: Global-aware Autoregressive Model for Multimodal Few-Shot Font Generation

ResearchDGX agent

arXiv:2601.01593v2 Announce Type: replace Abstract: Manual font design is an intricate process that transforms a stylistic visual concept into a coherent glyph set. This challenge persists in automate

Bridging the Gap: Converting Read Text to Conversational Dialogue

ResearchDGX agent

arXiv:2605.18001v1 Announce Type: new Abstract: In recent advancements within speech processing, converting read speech to conversational speech has gained significant attention. The primary challenge

CheckSupport: A Local LLM-Powered Tool for Automated Manuscript Submission Checklist Selection and Completion

Local AiDGX agent

arXiv:2605.16377v1 Announce Type: cross Abstract: Transparent and standardized reporting is essential for reproducible scientific research, yet adherence to reporting guidelines remains inconsistent b

DecoRec: Decomposed 3D Scene Reconstruction from Single-View Images via Object-Level Diffusion

ResearchDGX agent

arXiv:2605.16807v1 Announce Type: new Abstract: In this paper, we introduce extit{DecoRec}, a novel system designed to elevate single-view 2D images to a decomposed 3D scene mesh. Current methods for

Diffusion-Based sRGB Real Noise Generation via Prompt-Driven Noise Representation Learning

Model ReleasesDGX agent

arXiv:2603.04870v2 Announce Type: replace Abstract: Denoising in the sRGB image space is challenging due to large noise variability. Although end-to-end methods perform well, their effectiveness in re

ECG-WM: A Physiology-Informed ECG World Model for Clinical Intervention Simulation

SafetyDGX agent

arXiv:2605.17580v1 Announce Type: new Abstract: Electrocardiogram (ECG)-based models have achieved strong performance in diagnostic tasks, yet they remain limited in modeling how cardiac dynamics evol

Efficient 3D Content Reconstruction and Generation

HardwareDGX agent

arXiv:2605.18052v1 Announce Type: new Abstract: Automatic 3D content creation seeks to replace labor-intensive modeling and scanning pipelines with systems that can synthesize or recover 3D assets dir

Experiment-as-Code Labs: A Declarative Stack for AI-Driven Scientific Discovery

SafetyDGX agent

arXiv:2605.04375v2 Announce Type: replace-cross Abstract: To unleash the full potential of AI for Science, we must untether the agents from a purely digital environment. The agent's ability to control

← Previous
1…3435363738…47
Next →