AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “synthesis”

GridTimelineEvolution
2,818 results
1 Jul 2026

Towards Flexible, Natural, Efficient Interaction for Conversational Talking Face Generation

ResearchDGX agent

arXiv:2606.31088v1 Announce Type: new Abstract: Conversational talking face generation has recently attracted increasing attention, aiming to synthesize interactive talking videos where characters spe

UniCoder: Unified Visual-to-Code Generation via Symbolic Rewards and Reference-Guided Code Optimization

Model ReleasesDGX agent

arXiv:2606.31732v1 Announce Type: new Abstract: Visual-to-Code generation, which transforms scientific plots, vector graphics, and webpages into executable scripts, demands a level of pixel-precise al

VS3R: Robust Full-frame Video Stabilization via Deep 3D Reconstruction

ResearchDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2603.05851v2 Announce Type: replace Abstract: Video stabilization aims to mitigate camera shake but faces a fundamental trade-off between geometric robustness and full-frame consistency. While 2

30 Jun 2026

Agents-K1: Towards Agent-native Knowledge Orchestration

AgentsDGX agent

arXiv:2606.13669v2 Announce Type: replace Abstract: Current LLM-based research agents have advanced through agent orchestration, yet largely overlook scientific knowledge orchestration. Existing works

CLQT: A Closed-Loop, Cost-Aware, Strategy-Consistent Benchmark for Diagnostic Evaluation of LLM Portfolio-Management Agents

Model ReleasesDGX agent

arXiv:2606.29771v1 Announce Type: new Abstract: LLM agents are increasingly cast as autonomous portfolio managers, and benchmarks have moved from financial question-answering to sequential trading. Ye

Conversational Query Engine for Mixed-Modality Heterogeneous Enterprise Data Sources

Model ReleasesDGX agent

arXiv:2606.28370v1 Announce Type: cross Abstract: Enterprise business intelligence queries span structured warehouses and unstructured document repositories -- modalities with fundamentally different

CORTEX: High-Quality Cross-Domain Organization of Web-Scale Corpora through Ontological Corpus Graph

Model ReleasesDGX agent

arXiv:2606.30175v1 Announce Type: new Abstract: The continuous evolution of large language models drives escalating demands on data scale and quality, and as different training stages impose increasin

Data Provenance for Image Auto-Regressive Generation

ResearchDGX agent

arXiv:2606.28386v1 Announce Type: cross Abstract: Image autoregressive models (IARs) have recently demonstrated remarkable capabilities in visual content generation, achieving photorealistic quality a

DCGrasp: Distance-aware Controllable Grasp Generation

ResearchDGX agent

arXiv:2606.29924v1 Announce Type: new Abstract: Generating 3D hand-object interactions is essential for applications in robotics, XR, and synthetic data generation, where flexible controllability and

DEEPMED Search: An Open-Source Agentic Platform for Medical Deep Research with Introspective Verification

AgentsDGX agent

arXiv:2606.29746v1 Announce Type: new Abstract: Navigating the deluge of heterogeneous medical data, from academic literature (PubMed) to clinical guidelines (Web) and private knowledge bases, remains

DialogPII: A multilingual dataset of synthetic dialog transcripts to detect personal information

Model ReleasesDGX agent

arXiv:2606.30312v1 Announce Type: new Abstract: Conversational data collected in domains such as healthcare or social sciences is a valuable resource for research and automated analysis. However, resp

Distribution Matching Variational AutoEncoder

SafetyDGX agent

arXiv:2512.07778v2 Announce Type: replace Abstract: Most visual generative models compress images into a latent space before applying diffusion or autoregressive modelling. Yet, existing approaches su

Ground Truths in Suicide Research: The Current State of AI-Based Suicide Detection in Social Media

ResearchDGX agent

arXiv:2606.28334v1 Announce Type: cross Abstract: Recent advances in artificial intelligence (AI) and social media data have led to growing optimism about the ability to detect suicide risk at scale.

InsertAnywhere: Geometrically Grounded and Optics-Aware Video Object Insertion

SafetyDGX agent

arXiv:2512.17504v2 Announce Type: replace-cross Abstract: Recent advances in diffusion models have enabled impressive video editing capabilities, yet production-grade Video Object Insertion (VOI) rema

Learning Efficient 4D Gaussian Representations from Monocular Videos with Flow Splatting

ResearchDGX agent

arXiv:2606.29976v1 Announce Type: new Abstract: Reconstructing dynamic 3D scenes from monocular videos is challenging due to scene complexity and temporal dynamics. With the advancement of 3D Gaussian

MeEvo: Metacognitive Evolution Combined with Natural Evolution for Automatic Heuristic Design

ResearchDGX agent

arXiv:2606.14202v4 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have advanced Automatic Heuristic Design (AHD) by enabling heuristic generation through reasoning and code synthe

meta-pipe: An LLM-agent pipeline for end-to-end automated systematic review and meta-analysis

Model ReleasesDGX agent

arXiv:2606.28363v1 Announce Type: cross Abstract: Objective: To describe the architecture and design rationale of meta-pipe, an open-source large language model (LLM)-agent pipeline that integrates th

Occlusion-Robust Multi-Object Decoupling for Physics-Based Interaction

ApplicationsDGX agent

arXiv:2606.29303v1 Announce Type: new Abstract: We propose a mask-free method for lossless multi-object 3D reconstruction from sparse and occluded real-world views, enabling physically plausible inter

OLIVE: View-Augmented Latent Prediction with Waveform Reconstruction for Speech SSL

ResearchDGX agent

arXiv:2606.30356v1 Announce Type: new Abstract: We propose Online Latent prediction with Invariant Views and rEconstruction (OLIVE), a self-supervised speech representation learning framework that joi

Physically-Constrained Harmonic Separation for Robust Heart and Respiratory Rate Estimation from Wrist Photoplethysmography

ResearchDGX agent

arXiv:2606.30156v1 Announce Type: cross Abstract: Wrist-worn photoplethysmography (PPG) enables continuous monitoring of cardiopulmonary physiology, but reliable heart rate (HR) and respiratory rate (

PlantExpertVQA: A Visual Question Answering Dataset for Benchmarking Vision-Language Models in Plant Science

Model ReleasesDGX agent

arXiv:2508.17117v3 Announce Type: replace-cross Abstract: Existing plant-disease datasets target classification and detection, leaving vision-language models unable to support interactive, reasoning-b

Pondering the Way: Spatial-perceiving World Action Model for Embodied Navigation

SafetyDGX agent

arXiv:2606.29908v1 Announce Type: cross Abstract: Existing world model-based planners for visual navigation typically follow a verification-centric paradigm, decoupling goal intent from trajectory syn

PoseShield: Neural Collision Fields for Human Self-Collision Resolution

Model ReleasesDGX agent

arXiv:2606.29686v1 Announce Type: new Abstract: Self-collision remains a persistent challenge in SMPL-based human pose estimation and motion generation. Under extreme articulations or stochastic motio

Rectifying Mask via Entropy for Distractor-Free 3DGS in Ambiguous Scenarios

ResearchDGX agent

arXiv:2606.29496v1 Announce Type: new Abstract: We present RefineSplat, a systematic framework that effectively constructs transient masks to identify diverse ambiguous distractors. To do this, we qua

RefGlass-GS: A UAV-Enabled Fusion Framework for Photorealistic, Semantic and Interactive Digitization of Reflective Glass Facades via Gaussian Splatting

ApplicationsDGX agent

arXiv:2606.28826v1 Announce Type: new Abstract: Existing digitization of buildings with reflective glass facades suffers from geometric reconstruction distortion, unrealistic view-dependent texture re

RoamFlow: Reinforcement-Aligned One-Step Action MeanFlow Policy for Image-Goal Navigation

SafetyDGX agent

arXiv:2606.29934v1 Announce Type: new Abstract: Image-goal navigation is a key challenge in embodied robotics, where an agent must reach a target specified solely by a goal image. While existing reinf

Scenes as Objects, Not Primitives: Instance-Structured 3D Tokenization from Unposed Views

Local AiDGX agent

arXiv:2606.29513v1 Announce Type: new Abstract: A 3D scene is understood through its objects, not the primitives that compose them. Yet feed-forward reconstruction methods output dense, unstructured s

Shell-Supervised Gaussian Splatting for Urban Real-to-Sim Reconstruction

AgentsDGX agent

arXiv:2606.30014v1 Announce Type: new Abstract: Real-to-sim reconstruction for embodied AI requires geometry that is useful for collision reasoning, navigation, and agent-environment interaction, not

SICAGE: Speaker-Independent Culture-Aware Gesture Generation using TED4C-L Dataset

AgentsDGX agent

arXiv:2606.30001v1 Announce Type: new Abstract: Recent co-speech gesture generation methods often overlook cultural differences, limiting their effectiveness in human-agent interaction. Moreover, cult

UltraImageGen: Efficient Ultra-High-Resolution Image Generation with Hierarchical Local Attention

Local AiDGX agent

arXiv:2510.16325v3 Announce Type: replace Abstract: Ultra-high-resolution text-to-image generation is increasingly vital for applications requiring fine-grained textures and global structural fidelity

ViPSim: Collaborating Visual and Parameter Spaces for Consistent Long-Horizon Embodied World Models

Model ReleasesDGX agent

arXiv:2606.28804v1 Announce Type: new Abstract: Embodied World Models (EWMs) have emerged as a scalable and risk-free paradigm for advancing embodied intelligence, enabling the safety-critical evaluat

Walking in the Implicit: Interactive World Exploration via Neural Scene Representation

Local AiDGX agent

arXiv:2606.30045v1 Announce Type: new Abstract: Interactive video generation systems for camera-controlled world exploration roll out growing sequences of latent video frames, entangling state transit

29 Jun 2026

A Survey of Automated Presentation Coaching: Systems, Methods, and Open Challenges

ResearchDGX agent

arXiv:2606.27380v1 Announce Type: new Abstract: Automated coaching for oral presentations sits at the intersection of computer-assisted pronunciation training (CAPT), prosody modeling, and speech synt

Beyond MoCap: Scaling Motion Tokenizers with Synthetic Human Motion for Generative Modeling

ResearchDGX agent

arXiv:2606.27547v1 Announce Type: new Abstract: Human motion generation models are fundamentally constrained by the limited diversity of motion capture datasets, which predominantly contain common, re

From Detection to Action: Using LLM Agents for Fault-Tolerant Control

Model ReleasesDGX agent

arXiv:2606.28011v1 Announce Type: cross Abstract: We propose an agentic Large Language Model (LLM) framework for active Fault-Tolerant Control (FTC) that transforms fault detection outputs into constr

GeoFace: Consistent Multi-View Face Generation with Geometry-Constrained Diffusion

SafetyDGX agent

arXiv:2606.27659v1 Announce Type: new Abstract: We present GeoFace, a geometry-constrained multi-view diffusion framework for consistent face generation from a single input. % While recent multi-view

Learning Stochastic Bridges for Video Object Removal via Video-to-Video Translation

ResearchDGX agent

arXiv:2601.12066v4 Announce Type: replace Abstract: Existing video object removal methods predominantly rely on diffusion models following a noise-to-data paradigm, where generation starts from uninfo

PAC-Bayesian Certificates for Quadratic Closed-Loop Control

ResearchDGX agent

arXiv:2606.28281v1 Announce Type: cross Abstract: PAC-Bayesian bounds provide finite-sample guarantees for data-dependent randomized predictors, but applying them to learning-based control is difficul

Perceptual 3D Simulation With Physical World Modeling

Local AiDGX agent

arXiv:2606.27575v1 Announce Type: new Abstract: Predicting how a scene will evolve after a desired 3D transformation from images is a central goal in vision, graphics, and robotics. Yet unlike ideal s

26 Jun 2026

Axon: A Synthesizing Superoptimizer for Tensor Programs

ResearchDGX agent

arXiv:2606.26344v1 Announce Type: cross Abstract: Writing high performance kernels for AI accelerators requires deep expertise in tiling, instruction selection, data layout, and operator fusion placin

Bayesian Optimization for General Reaction Conditions

Model ReleasesDGX agent

arXiv:2502.18966v2 Announce Type: replace Abstract: General chemical reaction conditions that achieve consistently high performance across multiple substrates are important for practical applications

Charting the Growth of Social-Physical HRI (spHRI): A Systematic Review Pipeline Augmented by Small Language Models

ResearchDGX agent

arXiv:2606.26382v1 Announce Type: cross Abstract: Social-physical human-robot interaction (spHRI) has grown rapidly across robotics, human-computer interaction, human-robot interaction, and haptics. Y

Confidence-Aware Tool Orchestration for Robust Video Understanding

Model ReleasesDGX agent

arXiv:2606.26904v1 Announce Type: cross Abstract: Video reasoning language models implicitly assume that every input frame is equally reliable. This leads to what we term the Blind Trust Problem: unde

Context Recycling for Long-Horizon LLM Inference

Model ReleasesDGX agent

arXiv:2606.26105v1 Announce Type: cross Abstract: Large language models (LLMs) exhibit strong capabilities in short-context reasoning but degrade in performance over long conversational horizons due t

CORTEX: A Structured Reasoning Benchmark for Trustworthy 3D Chest CT MLLMs

Model ReleasesDGX agent

arXiv:2606.27264v1 Announce Type: new Abstract: Reasoning in multimodal large language models (MLLMs) has shown strong promise in medical imaging. However, this reasoning is usually free-form text jud

CyberChainBench: Can AI Agents Secure Smart Contracts Against Real-World On-Chain Vulnerabilities?

Model ReleasesDGX agent

arXiv:2606.26216v1 Announce Type: cross Abstract: We present CyberChainBench, a benchmark for evaluating LLM-based agents on smart contract security across three complementary tasks: vulnerability det

HyperDFlash: MHC-Aligned Block Speculative Decoding with Gated Residual Reduction

Model ReleasesDGX agent

arXiv:2606.26744v1 Announce Type: cross Abstract: We present HyperDFlash, a block-parallel speculative decoding framework tailored to the novel multi-hyper-connection (MHC) architecture proposed by De

LiveEdit: Towards Real-Time Diffusion-Based Streaming Video Editing

Model ReleasesDGX agent

arXiv:2606.26740v1 Announce Type: new Abstract: Streaming video editing has made rapid progress, yet practical deployment is still limited by two core issues: maintaining stable backgrounds and non-ed

Pianist Transformer: Towards Expressive Piano Performance Rendering via Scalable Self-Supervised Pre-Training

ApplicationsDGX agent

arXiv:2512.02652v2 Announce Type: replace-cross Abstract: Existing methods for expressive music performance rendering, a conditional generation task that aims to generate a human-like performance from

SciFig: Towards Automating Editable Figure Generation for Scientific Papers

Model ReleasesDGX agent

arXiv:2601.04390v2 Announce Type: replace Abstract: High-quality methodology figures are central to scientific communication, yet they remain difficult and time-consuming to create. Such figures must

SubdivAR: Autoregressive Next-Scale Prediction for Neural Mesh Subdivision

ResearchDGX agent

arXiv:2606.27088v1 Announce Type: new Abstract: Mesh subdivision is a fundamental operation for converting coarse, editable meshes into high-resolution surfaces, with broad applications in digital ass

Target-Aware Bandit Allocation for Scalable Surrogate Optimization in Chemical Space

ApplicationsDGX agent

arXiv:2606.26657v1 Announce Type: new Abstract: Identifying high-utility candidates from massive discrete spaces under expensive evaluations is a recurring challenge across the sciences, with structur

25 Jun 2026

ConSolv: Solvent-Conditional Machine Learning Implicit Solvent Potential

ResearchDGX agent

arXiv:2606.24983v1 Announce Type: cross Abstract: Implicit solvent machine learning potentials (MLPs) offer a powerful route to bridging the gap between accuracy and efficiency in molecular simulation

Enhancing Brain MRI Anomaly Detection and Reasoning with ROI Rethink and Synthetic Data

Model ReleasesDGX agent

arXiv:2606.25894v1 Announce Type: new Abstract: Medical vision-language models typically generate diagnoses through single-pass inference without indicating which image regions support their conclusio

Follow Your Track: Precise Skeleton Animation Controlled by 3D Trajectories

SafetyDGX agent

arXiv:2606.25344v1 Announce Type: new Abstract: 4D generation aims to animate 3D objects with realistic motion, holding great promise for applications. Existing methods typically decouple 3D asset gen

HEART: Coordination of Heterogeneous Expert Agents for Physically Grounded Robotic Task Planning

ApplicationsDGX agent

arXiv:2606.25404v1 Announce Type: new Abstract: Large Language Models (LLMs) can reason over complex instructions but often fail to satisfy the physical and spatial constraints required for robotic ta

MIMFlow: Integrating Masked Image Modeling with Normalizing Flows for End-to-End Image Generation

ResearchDGX agent

arXiv:2606.26016v1 Announce Type: new Abstract: Normalizing Flows (NFs) are powerful generative models capable of exact density estimation and sampling. However, their strict invertibility often force

OPERA: Aligning Open-Ended Reasoning via Objective Perplexity-based Reinforcement Learning

SafetyDGX agent

arXiv:2606.25757v1 Announce Type: new Abstract: Reinforcement Learning (RL) has enabled LLMs to excel in objective reasoning tasks such as mathematics and code generation. However, applying RL to open

ReaDy-Go: Real-to-Sim Dynamic 3D Gaussian Splatting Simulation for Environment-Specific Visual Navigation with Moving Obstacles

ApplicationsDGX agent

arXiv:2602.11575v3 Announce Type: replace-cross Abstract: Visual navigation models often struggle in real-world dynamic environments due to limited robustness to the sim-to-real gap and the difficulty

SSMNBench: Diagnosing Image-based Cross-View Human-Object Understanding via Single-View Sufficiency and Multi-View Necessity

Model ReleasesDGX agent

arXiv:2606.25634v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have shown remarkable progress in single-image perception, yet their ability to reason about complex cross-view

← Previous
1…2829303132…47
Next →