AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “synthesis”

GridTimelineEvolution
2,818 results
15 Apr 2026

StructDiff: A Structure-Preserving and Spatially Controllable Diffusion Model for Single-Image Generation

TutorialsDGX agent

arXiv:2604.12575v1 Announce Type: new Abstract: This paper introduces StructDiff, a generative framework based on a single-scale diffusion model for single-image generation. Single-image generation ai

SynthPix: A lightspeed PIV image generator

Model ReleasesDGX agent

arXiv:2512.09664v2 Announce Type: replace-cross Abstract: We describe SynthPix, a synthetic image generator for Particle Image Velocimetry (PIV) with a focus on performance and parallelism on accelera

Towards Long-horizon Agentic Multimodal Search

Model ReleasesDGX agent

arXiv:2604.12890v1 Announce Type: cross Abstract: Multimodal deep search agents have shown great potential in solving complex tasks by iteratively collecting textual and visual evidence. However, mana


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Towards Realistic and Consistent Orbital Video Generation via 3D Foundation Priors

ResearchDGX agent

arXiv:2604.12309v1 Announce Type: new Abstract: We present a novel method for generating geometrically realistic and consistent orbital videos from a single image of an object. Existing video generati

WebFactory: Automated Compression of Foundational Language Intelligence into Grounded Web Agents

AgentsDGX agent

arXiv:2603.05044v2 Announce Type: replace Abstract: Current paradigms for training GUI agents are fundamentally limited by a reliance on either unsafe, non-reproducible live web interactions or costly

When Robert Burns meets LTX and Ace Step xl

Local AiDGX agent

This r/StableDiffusion post likely showcases a creative AI-generated video or audio experiment combining the poetry or likeness of Scottish poet Robert Burns with two AI generation tools: LTX (Lightri

XRZero-G0: Pushing the Frontier of Dexterous Robotic Manipulation with Interfaces, Quality and Ratios

SafetyDGX agent

arXiv:2604.13001v1 Announce Type: new Abstract: The acquisition of high-quality, action-aligned demonstration data remains a fundamental bottleneck in scaling foundation models for dexterous robot man

14 Apr 2026

A collaborative agent with two lightweight synergistic models for autonomous crystal materials research

AgentsDGX agent

arXiv:2604.11540v1 Announce Type: new Abstract: Current large language models require hundreds of billions of parameters yet struggle with domain-specific reasoning and tool coordination in materials

A Temporally Augmented Graph Attention Network for Affordance Classification

ResearchDGX agent

arXiv:2604.10149v1 Announce Type: cross Abstract: Graph attention networks (GATs) provide one of the best frameworks for learning node representations in relational data; but, existing variants such a

Adaptive Multi-Expert Reasoning via Difficulty-Aware Routing and Uncertainty-Guided Aggregation

ResearchDGX agent

arXiv:2604.10335v1 Announce Type: new Abstract: Large language models (LLMs) demonstrate strong performance in math reasoning benchmarks, but their performance varies inconsistently across problems wi

Adoption and Effectiveness of AI-Based Anomaly Detection for Cross Provider Health Data Exchange

TutorialsDGX agent

arXiv:2604.09630v1 Announce Type: cross Abstract: This study investigates the adoption and effectiveness of AI-based anomaly detection in cross-provider electronic health record (EHR) environments. It

Affostruction: 3D Affordance Grounding with Generative Reconstruction

ResearchDGX agent

arXiv:2601.09211v2 Announce Type: replace Abstract: This paper addresses the problem of affordance grounding from RGBD images of an object, which aims to localize surface regions corresponding to a te

Agentic AI in Engineering and Manufacturing: Industry Perspectives on Utility, Adoption, Challenges, and Opportunities

AgentsDGX agent

arXiv:2604.09633v1 Announce Type: cross Abstract: This work examines how AI, especially agentic systems, is being adopted in engineering and manufacturing workflows, what value it provides today, and

Agentic Video Generation: From Text to Executable Event Graphs via Tool-Constrained LLM Planning

SafetyDGX agent

arXiv:2604.10383v1 Announce Type: new Abstract: Existing multi-agent video generation systems use LLM agents to orchestrate neural video generators, producing visually impressive but semantically unre

ANCHOR: Branch-Point Data Generation for GUI Agents

AgentsDGX agent

arXiv:2602.07153v2 Announce Type: replace Abstract: End-to-end GUI agents for real desktop environments require large amounts of high-quality interaction data, yet collecting human demonstrations is e

Audio-Omni: Extending Multi-modal Understanding to Versatile Audio Generation and Editing

Model ReleasesDGX agent

arXiv:2604.10708v1 Announce Type: cross Abstract: Recent progress in multimodal models has spurred rapid advances in audio understanding, generation, and editing. However, these capabilities are typic

Automatic Uncertainty-Aware Synthetic Data Bootstrapping for Historical Map Segmentation

ApplicationsDGX agent

arXiv:2511.15875v2 Announce Type: replace Abstract: The automated analysis of historical documents, particularly maps, has drastically benefited from advances in deep learning and its success across v

Beyond RAG for Cyber Threat Intelligence: A Systematic Evaluation of Graph-Based and Agentic Retrieval

AgentsDGX agent

arXiv:2604.11419v1 Announce Type: new Abstract: Cyber threat intelligence (CTI) analysts must answer complex questions over large collections of narrative security reports. Retrieval-augmented generat

ComSim: Building Scalable Real-World Robot Data Generation via Compositional Simulation

SafetyDGX agent

arXiv:2604.11386v1 Announce Type: cross Abstract: Recent advancements in foundational models, such as large language models and world models, have greatly enhanced the capabilities of robotics, enabli

COSMIK-MPPI: Scaling Constrained Model Predictive Control to Collision Avoidance in Close-Proximity Dynamic Human Environments

SafetyDGX agent

arXiv:2604.10358v1 Announce Type: new Abstract: Ensuring safe physical interaction between torque-controlled manipulators and humans is essential for deploying robots in everyday environments. Model P

CraftGraffiti: Exploring Human Identity with Custom Graffiti Art via Facial-Preserving Diffusion Models

ApplicationsDGX agent

arXiv:2508.20640v2 Announce Type: replace Abstract: Preserving facial identity under extreme stylistic transformation remains a major challenge in generative art. In graffiti, a high-contrast, abstrac

CylinderDepth: Cylindrical Spatial Attention for Multi-View Consistent Self-Supervised Surround Depth Estimation

ResearchDGX agent

arXiv:2511.16428v3 Announce Type: replace Abstract: Self-supervised surround-view depth estimation enables dense, low-cost 3D perception with a 360{eg} field of view from multiple minimally overlappin

Deep-Reporter: Deep Research for Grounded Multimodal Long-Form Generation

Local AiDGX agent

arXiv:2604.10741v1 Announce Type: cross Abstract: Recent agentic search frameworks enable deep research via iterative planning and retrieval, reducing hallucinations and enhancing factual grounding. H

DIB-OD: Preserving the Invariant Core for Robust Heterogeneous Graph Adaptation via Decoupled Information Bottleneck and Online Distillation

ResearchDGX agent

arXiv:2604.10882v1 Announce Type: cross Abstract: Graph Neural Network pretraining is pivotal for leveraging unlabeled graph data. However, generalizing across heterogeneous domains remains a major ch

Distributionally Robust PAC-Bayesian Control

SafetyDGX agent

arXiv:2604.10588v1 Announce Type: new Abstract: We present a distributionally robust PAC-Bayesian framework for certifying the performance of learning-based finite-horizon controllers. While existing

Doc-PP: Document Policy Preservation Benchmark for Large Vision-Language Models

Model ReleasesDGX agent

arXiv:2601.03926v2 Announce Type: replace Abstract: The deployment of Large Vision-Language Models (LVLMs) for real-world document question answering is often constrained by dynamic, user-defined poli

EE-MCP: Self-Evolving MCP-GUI Agents via Automated Environment Generation and Experience Learning

SafetyDGX agent

arXiv:2604.09815v1 Announce Type: new Abstract: Computer-use agents that combine GUI interaction with structured API calls via the Model Context Protocol (MCP) show promise for automating software tas

Evaluating Memory Capability in Continuous Lifelog Scenario

Model ReleasesDGX agent

arXiv:2604.11182v1 Announce Type: new Abstract: Nowadays, wearable devices can continuously lifelog ambient conversations, creating substantial opportunities for memory systems. However, existing benc

EvoDiagram: Agentic Editable Diagram Creation via Design Expertise Evolution

Model ReleasesDGX agent

arXiv:2604.09568v1 Announce Type: cross Abstract: High-fidelity diagram creation requires the complex orchestration of semantic topology, visual styling, and spatial layout, posing a significant chall

Exploring Knowledge Purification in Multi-Teacher Knowledge Distillation for LLMs

ResearchDGX agent

arXiv:2602.01064v2 Announce Type: replace Abstract: Knowledge distillation has emerged as a pivotal technique for transferring knowledge from stronger large language models (LLMs) to smaller, more eff

FF3R: Feedforward Feature 3D Reconstruction from Unconstrained views

ResearchDGX agent

arXiv:2604.09862v1 Announce Type: new Abstract: Recent advances in vision foundation models have revolutionized geometry reconstruction and semantic understanding. Yet, most of the existing approaches

GeoArena: Evaluating Open-World Geographic Reasoning in Large Vision-Language Models

Model ReleasesDGX agent

arXiv:2509.04334v4 Announce Type: replace Abstract: Geographic reasoning is a fundamental cognitive capability that requires models to infer plausible locations by synthesizing visual evidence with sp

GIANTS: Generative Insight Anticipation from Scientific Literature

Model ReleasesDGX agent

arXiv:2604.09793v1 Announce Type: cross Abstract: Scientific breakthroughs often emerge from synthesizing prior ideas into novel contributions. While language models (LMs) show promise in scientific d

GroupRank: A Groupwise Paradigm for Effective and Efficient Passage Reranking with LLMs

Model ReleasesDGX agent

arXiv:2511.11653v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have emerged as powerful tools for passage reranking in information retrieval, leveraging their superior reasonin

HO-Flow: Generalizable Hand-Object Interaction Generation with Latent Flow Matching

TutorialsDGX agent

arXiv:2604.10836v1 Announce Type: new Abstract: Generating realistic 3D hand-object interactions (HOI) is a fundamental challenge in computer vision and robotics, requiring both temporal coherence and

HOG-Layout: Hierarchical 3D Scene Generation, Optimization and Editing via Vision-Language Models

ResearchDGX agent

arXiv:2604.10772v1 Announce Type: new Abstract: 3D layout generation and editing play a crucial role in Embodied AI and immersive VR interaction. However, manual creation requires tedious labor, while

HumanVBench: Probing Human-Centric Video Understanding in MLLMs with Automatically Synthesized Benchmarks

Model ReleasesDGX agent

arXiv:2412.17574v3 Announce Type: replace-cross Abstract: Evaluating the nuanced human-centric video understanding capabilities of Multimodal Large Language Models (MLLMs) remains a great challenge, a

Intra-finger Variability of Diffusion-based Latent Fingerprint Generation

ResearchDGX agent

arXiv:2604.10040v1 Announce Type: new Abstract: The primary goal of this work is to systematically evaluate the intra-finger variability of synthetic fingerprints (particularly latent prints) generate

Knowing What to Stress: A Discourse-Conditioned Text-to-Speech Benchmark

Model ReleasesDGX agent

arXiv:2604.10580v1 Announce Type: new Abstract: Spoken meaning often depends not only on what is said, but also on which word is emphasized. The same sentence can convey correction, contrast, or clari

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation

ResearchDGX agent

arXiv:2509.20128v2 Announce Type: replace-cross Abstract: Audio-driven facial animation has made significant progress in multimedia applications, with diffusion models showing strong potential for tal

Learning Long-term Motion Embeddings for Efficient Kinematics Generation

TutorialsDGX agent

arXiv:2604.11737v1 Announce Type: new Abstract: Understanding and predicting motion is a fundamental component of visual intelligence. Although modern video models exhibit strong comprehension of scen

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment

SafetyDGX agent

arXiv:2604.10677v1 Announce Type: cross Abstract: Scaling up robot learning is hindered by the scarcity of robotic demonstrations, whereas human videos offer a vast, untapped source of interaction dat

LLMs for Text-Based Exploration and Navigation Under Partial Observability

Model ReleasesDGX agent

arXiv:2604.09604v1 Announce Type: new Abstract: Exploration and goal-directed navigation in unknown layouts are central to inspection, logistics, and search-and-rescue. We ask whether large language m

MorphoFlow: Sparse-Supervised Generative Shape Modeling with Adaptive Latent Relevance

TutorialsDGX agent

arXiv:2604.11636v1 Announce Type: new Abstract: Statistical shape modeling (SSM) is central to population level analysis of anatomical variability, yet most existing approaches rely on densely annotat

OccuBench: Evaluating AI Agents on Real-World Professional Tasks via Language World Models

Model ReleasesDGX agent

arXiv:2604.10866v1 Announce Type: new Abstract: AI agents are expected to perform professional work across hundreds of occupational domains (from emergency department triage to nuclear reactor safety

Online Learning-Enhanced High Order Adaptive Safety Control

SafetyDGX agent

arXiv:2511.19651v2 Announce Type: replace Abstract: Control barrier functions (CBFs) are an effective model-based tool to formally certify the safety of a system. With the growing complexity of modern

Pioneer Agent: Continual Improvement of Small Language Models in Production

Model ReleasesDGX agent

arXiv:2604.09791v1 Announce Type: new Abstract: Small language models are attractive for production deployment due to their low cost, fast inference, and ease of specialization. However, adapting them

Prompt Relay: Inference-Time Temporal Control for Multi-Event Video Generation

SafetyDGX agent

arXiv:2604.10030v1 Announce Type: new Abstract: Video diffusion models have achieved remarkable progress in generating high-quality videos. However, these models struggle to represent the temporal suc

Representations Before Pixels: Semantics-Guided Hierarchical Video Prediction

AgentsDGX agent

arXiv:2604.11707v1 Announce Type: new Abstract: Accurate future video prediction requires both high visual fidelity and consistent scene semantics, particularly in complex dynamic environments such as

RiTeK: A Dataset for Large Language Models Complex Reasoning over Textual Knowledge Graphs in Medicine

Model ReleasesDGX agent

arXiv:2410.13987v3 Announce Type: replace Abstract: Answering complex real-world questions in the medical domain often requires accurate retrieval from medical Textual Knowledge Graphs (medical TKGs),

Ro-SLM: Onboard Small Language Models for Robot Task Planning and Operation Code Generation

TutorialsDGX agent

arXiv:2604.10929v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) provide robots with contextual reasoning abilities to comprehend human instructions. Yet, current LLM-en

RoboStereo: Dual-Tower 4D Embodied World Models for Unified Policy Optimization

SafetyDGX agent

arXiv:2603.12639v2 Announce Type: replace Abstract: Scalable Embodied AI faces fundamental constraints due to prohibitive costs and safety risks of real-world interaction. While Embodied World Models

Robust Real-Time Coordination of CAVs: A Distributed Optimization Framework under Uncertainty

SafetyDGX agent

arXiv:2508.21322v2 Announce Type: replace Abstract: Achieving both safety guarantees and real-time performance in cooperative vehicle coordination remains a fundamental challenge, particularly in dyna

Seeing Through Deception: Uncovering Misleading Creator Intent in Multimodal News with Vision-Language Models

Model ReleasesDGX agent

arXiv:2505.15489v4 Announce Type: replace-cross Abstract: The impact of multimodal misinformation arises not only from factual inaccuracies but also from the misleading narratives that creators delibe

SIGMA: An Efficient Heterophilous Graph Neural Network with Fast Global Aggregation

ResearchDGX agent

arXiv:2305.09958v5 Announce Type: replace Abstract: Graph neural networks (GNNs) realize great success in graph learning but suffer from performance loss when meeting heterophily, i.e. neighboring nod

Simply retrieving a reasoning trace looks a lot like human reasoning, until it's time to navigate uncharted territory. If you memorized all …

ResearchDGX agent

Simply retrieving a reasoning trace looks a lot like human reasoning, until it's time to navigate uncharted territory. If you memorized all reasoning traces of humans from 10,000 BC, you could automat

StaMo: Unsupervised Learning of Generalizable Robot Motion from Compact State Representation

SafetyDGX agent

arXiv:2510.05057v2 Announce Type: replace-cross Abstract: A fundamental challenge in embodied intelligence is developing expressive and compact state representations for efficient world modeling and d

Switch-JustDance: Benchmarking Whole Body Motion Tracking Controllers Using a Commercial Console Game

Model ReleasesDGX agent

arXiv:2511.17925v3 Announce Type: replace-cross Abstract: Recent advances in whole-body robot control have enabled humanoid and legged robots to perform increasingly agile and coordinated motions. How

The Phantom of PCIe: Constraining Generative Artificial Intelligences for Practical Peripherals Trace Synthesizing

ResearchDGX agent

arXiv:2411.06376v3 Announce Type: replace-cross Abstract: Peripheral Component Interconnect Express (PCIe) is the de facto interconnect standard for high-speed peripherals and CPUs. The development of

UDAPose: Unsupervised Domain Adaptation for Low-Light Human Pose Estimation

ResearchDGX agent

arXiv:2604.10485v1 Announce Type: cross Abstract: Low-visibility scenarios, such as low-light conditions, pose significant challenges to human pose estimation due to the scarcity of annotated low-ligh

← Previous
1…44454647
Next →