AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “synthesis”

GridTimelineEvolution
2,818 results
17 Apr 2026

From Reactive to Proactive: Assessing the Proactivity of Voice Agents via ProVoice-Bench

ResearchDGX agent

arXiv:2604.15037v1 Announce Type: cross Abstract: Recent advancements in LLM agents are gradually shifting from reactive, text-based paradigms toward proactive, multimodal interaction. However, existi

Generative Data Augmentation for Skeleton Action Recognition

ResearchDGX agent

arXiv:2604.14933v1 Announce Type: new Abstract: Skeleton-based human action recognition is a powerful approach for understanding human behaviour from pose data, but collecting large-scale, diverse, an

Generative Modeling of Complex-Valued Brain MRI Data

ResearchDGX agent

arXiv:2604.14800v1 Announce Type: cross Abstract: Objective. Standard Magnetic Resonance Imaging (MRI) reconstruction pipelines discard phase information captured during acquisition, despite evidence


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

GlobalSplat: Efficient Feed-Forward 3D Gaussian Splatting via Global Scene Tokens

Local AiDGX agent

arXiv:2604.15284v1 Announce Type: new Abstract: The efficient spatial allocation of primitives serves as the foundation of 3D Gaussian Splatting, as it directly dictates the synergy between representa

Hybrid Latents -- Geometry-Appearance-Aware Surfel Splatting

SafetyDGX agent

arXiv:2604.14928v1 Announce Type: new Abstract: We introduce a hybrid Gaussian-hash-grid radiance representation for reconstructing 2D Gaussian scene models from multi-view images. Similar to NeST spl

Interpretable and Explainable Surrogate Modeling for Simulations: A State-of-the-Art Survey and Perspectives on Explainable AI for Decision-Making

AgentsDGX agent

arXiv:2604.14240v1 Announce Type: cross Abstract: The simulation of complex systems increasingly relies on sophisticated but fundamentally opaque computational black-box simulators. Surrogate models p

LongAct: Harnessing Intrinsic Activation Patterns for Long-Context Reinforcement Learning

Model ReleasesDGX agent

arXiv:2604.14922v1 Announce Type: cross Abstract: Reinforcement Learning (RL) has emerged as a critical driver for enhancing the reasoning capabilities of Large Language Models (LLMs). While recent ad

MADE: A Living Benchmark for Multi-Label Text Classification with Uncertainty Quantification of Medical Device Adverse Events

Model ReleasesDGX agent

arXiv:2604.15203v1 Announce Type: new Abstract: Machine learning in high-stakes domains such as healthcare requires not only strong predictive performance but also reliable uncertainty quantification

MARS: Sound Generation via Multi-Channel Autoregression on Spectrograms

ResearchDGX agent

arXiv:2509.26007v2 Announce Type: replace-cross Abstract: Research on audio generation has progressively developed along both waveform-based and spectrogram-based directions, giving rise to diverse st

Prompt-to-Gesture: Measuring the Capabilities of Image-to-Video Deictic Gesture Generation

ResearchDGX agent

arXiv:2604.14953v1 Announce Type: new Abstract: Gesture recognition research, unlike NLP, continues to face acute data scarcity, with progress constrained by the need for costly human recordings or im

Reference-Free Sampling-Based Model Predictive Control

HardwareDGX agent

arXiv:2511.19204v3 Announce Type: replace Abstract: We present a sampling-based model predictive control (MPC) framework that enables emergent locomotion without relying on handcrafted gait patterns o

Separation is Optimal for LQR under Intermittent Feedback

SafetyDGX agent

arXiv:2603.27833v2 Announce Type: cross Abstract: In this work, we first prove that the separation principle holds for communication-constrained LQR problems under i.i.d. zero-mean disturbances with a

TableNet A Large-Scale Table Dataset with LLM-Powered Autonomous

AgentsDGX agent

arXiv:2604.13041v1 Announce Type: cross Abstract: Table Structure Recognition (TSR) requires the logical reasoning ability of large language models (LLMs) to handle complex table layouts, but current

The PICCO Framework for Large Language Model Prompting: A Taxonomy and Reference Architecture for Prompt Structure

SafetyDGX agent

arXiv:2604.14197v1 Announce Type: new Abstract: Large language model (LLM) performance depends heavily on prompt design, yet prompt construction is often described and applied inconsistently. Our purp

Towards a Multi-Embodied Grasping Agent

AgentsDGX agent

arXiv:2510.27420v3 Announce Type: replace Abstract: Multi-embodiment grasping focuses on developing approaches that exhibit generalist behavior across diverse gripper designs. Existing methods often l

Towards Scalable Lightweight GUI Agents via Multi-role Orchestration

SafetyDGX agent

arXiv:2604.13488v1 Announce Type: new Abstract: Autonomous Graphical User Interface (GUI) agents powered by Multimodal Large Language Models (MLLMs) enable digital automation on end-user devices. Whil

Trajectory Planning for a Multi-UAV Rigid-Payload Cascaded Transportation System Based on Enhanced Tube-RRT*

SafetyDGX agent

arXiv:2604.15074v1 Announce Type: new Abstract: This paper presents a two-stage trajectory planning framework for a multi-UAV rigid-payload cascaded transportation system, aiming to address planning c

Unraveling the Mechanism of Drug Binding to SARS-CoV-2 RNA Pseudoknot with Thermodynamics-Driven Machine Learning

TutorialsDGX agent

arXiv:2604.14906v1 Announce Type: cross Abstract: The SARS-CoV-2 RNA pseudoknot is a promising target for antiviral intervention, as it regulates the efficiency of -1 programmed ribosomal frameshiftin

16 Apr 2026

3DRealHead: Few-Shot Detailed Head Avatar

ResearchDGX agent

arXiv:2604.13171v1 Announce Type: new Abstract: The human face is central to communication. For immersive applications, the digital presence of a person should mirror the physical reality, capturing t

Action Images: End-to-End Policy Learning via Multiview Video Generation

SafetyDGX agent

arXiv:2604.06168v2 Announce Type: replace Abstract: World action models (WAMs) have emerged as a promising direction for robot policy learning, as they can leverage powerful video backbones to model t

Add any kind of audio(Voices,SFX, ambience) to existing video?

Local AiDGX agent

This r/StableDiffusion thread likely discusses methods and AI tools for adding AI-generated or sourced audio — including voice, sound effects, and ambient sound — to existing videos created with Stabl

ASTER: Latent Pseudo-Anomaly Generation for Unsupervised Time-Series Anomaly Detection

Model ReleasesDGX agent

arXiv:2604.13924v1 Announce Type: cross Abstract: Time-series anomaly detection (TSAD) is critical in domains such as industrial monitoring, healthcare, and cybersecurity, but it remains challenging d

ASTRA: Enhancing Multi-Subject Generation with Retrieval-Augmented Pose Guidance and Disentangled Position Embedding

Model ReleasesDGX agent

arXiv:2604.13938v1 Announce Type: new Abstract: Subject-driven image generation has shown great success in creating personalized content, but its capabilities are largely confined to single subjects i

Asymmetric-Loss-Guided Hybrid CNN-BiLSTM-Attention Model for Industrial RUL Prediction with Interpretable Failure Heatmaps

SafetyDGX agent

arXiv:2604.13459v1 Announce Type: new Abstract: Turbofan engine degradation under sustained operational stress necessitates robust prognostic systems capable of accurately estimating the Remaining Use

Beyond Conservative Automated Driving in Multi-Agent Scenarios via Coupled Model Predictive Control and Deep Reinforcement Learning

SafetyDGX agent

arXiv:2604.13891v1 Announce Type: new Abstract: Automated driving at unsignalized intersections is challenging due to complex multi-vehicle interactions and the need to balance safety and efficiency.

Boundary Sampling to Learn Predictive Safety Filters via Pontryagin's Maximum Principle

SafetyDGX agent

arXiv:2604.13325v1 Announce Type: new Abstract: Safety filters provide a practical approach for enforcing safety constraints in autonomous systems. While learning-based tools scale to high-dimensional

ChartNet: A Million-Scale, High-Quality Multimodal Dataset for Robust Chart Understanding

SafetyDGX agent

arXiv:2603.27064v2 Announce Type: replace-cross Abstract: Understanding charts requires models to jointly reason over geometric visual patterns, structured numerical data, and natural language -- a ca

Computational framework for multistep metabolic pathway design

ResearchDGX agent

arXiv:2604.13471v1 Announce Type: new Abstract: In silico tools are important for generating novel hypotheses and exploring alternatives in de novo metabolic pathway design. However, while many comput

Correct Chains, Wrong Answers: Dissociating Reasoning from Output in LLM Logic

Model ReleasesDGX agent

arXiv:2604.13065v1 Announce Type: new Abstract: LLMs can execute every step of chain-of-thought reasoning correctly and still produce wrong final answers. We introduce the Novel Operator Test, a bench

Cross-Layer Co-Optimized LSTM Accelerator for Real-Time Gait Analysis

ApplicationsDGX agent

arXiv:2604.13543v1 Announce Type: cross Abstract: Long Short-Term Memory (LSTM) neural networks have penetrated healthcare applications where real-time requirements and edge computing capabilities are

DiT as Real-Time Rerenderer: Streaming Video Stylization with Autoregressive Diffusion Transformer

ResearchDGX agent

arXiv:2604.13509v1 Announce Type: new Abstract: Recent advances in video generation models has significantly accelerated video generation and related downstream tasks. Among these, video stylization h

EmbodiedClaw: Conversational Workflow Execution for Embodied AI Development

Model ReleasesDGX agent

arXiv:2604.13800v1 Announce Type: new Abstract: Embodied AI research is increasingly moving beyond single-task, single-environment policy learning toward multi-task, multi-scene, and multi-model setti

ExpSeek: Self-Triggered Experience Seeking for Web Agents

Model ReleasesDGX agent

arXiv:2601.08605v2 Announce Type: replace Abstract: Experience intervention in web agents emerges as a promising technical paradigm, enhancing agent interaction capabilities by providing valuable insi

From Pixels to Nucleotides: End-to-End Token-Based Video Compression for DNA Storage

ResearchDGX agent

arXiv:2604.13667v1 Announce Type: new Abstract: DNA-based storage has emerged as a promising approach to the global data crisis, offering molecular-scale density and millennial-scale stability at low

Frozen Forecasting: A Unified Evaluation

ResearchDGX agent

arXiv:2507.13942v2 Announce Type: replace Abstract: Forecasting future events is a fundamental capability for general-purpose systems that plan or act across different levels of abstraction. Yet, eval

Giving Voice to the Constitution: Low-Resource Text-to-Speech for Quechua and Spanish Using a Bilingual Legal Corpus

ApplicationsDGX agent

arXiv:2604.13288v1 Announce Type: new Abstract: We present a unified pipeline for synthesizing high-quality Quechua and Spanish speech for the Peruvian Constitution using three state-of-the-art text-t

Hydra: Unifying Document Retrieval and Generation in a Single Vision-Language Model

HardwareDGX agent

arXiv:2603.28554v2 Announce Type: replace Abstract: Visual document understanding typically requires separate retrieval and generation models, doubling memory and system complexity. We present Hydra,

MixAtlas: Uncertainty-aware Data Mixture Optimization for Multimodal LLM Midtraining

ResearchDGX agent

This paper was accepted at the Workshop on Navigating and Addressing Data Problems for Foundation Models (NADPFM) at ICLR 2026. Principled domain reweighting can substantially improve sample efficienc

MM-Doc-R1: Training Agents for Long Document Visual Question Answering through Multi-turn Reinforcement Learning

Model ReleasesDGX agent

arXiv:2604.13579v1 Announce Type: new Abstract: Conventional Retrieval-Augmented Generation (RAG) systems often struggle with complex multi-hop queries over long documents due to their single-pass ret

OneHOI: Unifying Human-Object Interaction Generation and Editing

ResearchDGX agent

arXiv:2604.14062v1 Announce Type: new Abstract: Human-Object Interaction (HOI) modelling captures how humans act upon and relate to objects, typically expressed as triplets. Existing approaches split

PostureObjectstitch: Anomaly Image Generation Considering Assembly Relationships in Industrial Scenarios

ApplicationsDGX agent

arXiv:2604.13863v1 Announce Type: new Abstract: Image generation technology can synthesize condition-specific images to supplement real-world industrial anomaly data and enhance anomaly detection mode

SemiFA: An Agentic Multi-Modal Framework for Autonomous Semiconductor Failure Analysis Report Generation

HardwareDGX agent

arXiv:2604.13236v1 Announce Type: new Abstract: Semiconductor failure analysis (FA) requires engineers to examine inspection images, correlate equipment telemetry, consult historical defect records, a

SSD-GS: Scattering and Shadow Decomposition for Relightable 3D Gaussian Splatting

ResearchDGX agent

arXiv:2604.13333v1 Announce Type: new Abstract: We present SSD-GS, a physically-based relighting framework built upon 3D Gaussian Splatting (3DGS) that achieves high-quality reconstruction and photore

There's a broadly held misconception in AI that methods that scale well are simple methods -- even, that simple methods usually scale. This …

ResearchDGX agent

There's a broadly held misconception in AI that methods that scale well are simple methods -- even, that simple methods usually scale. This is completely wrong. Pretty much none of the truly simple me

Towards Unconstrained Human-Object Interaction

ResearchDGX agent

arXiv:2604.14069v1 Announce Type: new Abstract: Human-Object Interaction (HOI) detection is a longstanding computer vision problem concerned with predicting the interaction between humans and objects.

15 Apr 2026

Algorithmic Analysis of Dense Associative Memory: Finite-Size Guarantees and Adversarial Robustness

ResearchDGX agent

arXiv:2604.12811v1 Announce Type: cross Abstract: Dense Associative Memory (DAM) generalizes Hopfield networks through higher-order interactions and achieves storage capacity that scales as O(N^{n-1})

ARGen: Affect-Reinforced Generative Augmentation towards Vision-based Dynamic Emotion Perception

SafetyDGX agent

arXiv:2604.12255v1 Announce Type: cross Abstract: Dynamic facial expression recognition in the wild remains challenging due to data scarcity and long-tail distributions, which hinder models from effec

ArtifactWorld: Scaling 3D Gaussian Splatting Artifact Restoration via Video Generation Models

ApplicationsDGX agent

arXiv:2604.12251v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) delivers high-fidelity real-time rendering but suffers from geometric and photometric degradations under sparse-view constr

Artificial Intelligence for Modeling and Simulation of Mixed Automated and Human Traffic

AgentsDGX agent

arXiv:2604.12857v1 Announce Type: new Abstract: Autonomous vehicles (AVs) are now operating on public roads, which makes their testing and validation more critical than ever. Simulation offers a safe

Beyond Factual Grounding: The Case for Opinion-Aware Retrieval-Augmented Generation

SafetyDGX agent

arXiv:2604.12138v1 Announce Type: new Abstract: RAG systems have transformed how LLMs access external knowledge, but we find that current implementations exhibit a bias toward factual, objective conte

Causal Fingerprints of AI Generative Models

ResearchDGX agent

arXiv:2509.15406v2 Announce Type: replace Abstract: AI generative models leave implicit traces in their generated images, which are commonly referred to as model fingerprints and are exploited for sou

DAV-GSWT: Diffusion-Active-View Sampling for Data-Efficient Gaussian Splatting Wang Tiles

ResearchDGX agent

arXiv:2602.15355v3 Announce Type: replace Abstract: The emergence of 3D Gaussian Splatting has fundamentally redefined the capabilities of photorealistic neural rendering by enabling high-throughput s

Deep QP Safety Filter: Model-free Learning for Reachability-based Safety Filter

SafetyDGX agent

arXiv:2601.21297v2 Announce Type: replace Abstract: We introduce Deep QP Safety Filter, a fully data-driven safety layer for black-box dynamical systems. Our method learns a Quadratic-Program (QP) saf

Enhancing Agentic Textual Graph Retrieval with Synthetic Stepwise Supervision

SafetyDGX agent

arXiv:2510.03323v2 Announce Type: replace Abstract: Integrating textual graphs into Large Language Models (LLMs) is promising for complex graph-based QA. However, a key bottleneck is retrieving inform

MedGS: Gaussian Splatting for Multi-Modal 3D Medical Imaging

ApplicationsDGX agent

arXiv:2509.16806v4 Announce Type: replace Abstract: Endoluminal endoscopic procedures are essential for diagnosing colorectal cancer and other severe conditions in the digestive tract, urogenital syst

NaviRAG: Towards Active Knowledge Navigation for Retrieval-Augmented Generation

AgentsDGX agent

arXiv:2604.12766v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) typically relies on a flat retrieval paradigm that maps queries directly to static, isolated text segments. This ap

On Efficient Variants of Segment Anything Model: A Survey

ApplicationsDGX agent

arXiv:2410.04960v5 Announce Type: replace Abstract: The Segment Anything Model (SAM) is a foundational model for image segmentation tasks, known for its strong generalization across diverse applicatio

PianoFlow: Music-Aware Streaming Piano Motion Generation with Bimanual Coordination

ResearchDGX agent

arXiv:2604.12856v1 Announce Type: new Abstract: Audio-driven bimanual piano motion generation requires precise modeling of complex musical structures and dynamic cross-hand coordination. However, exis

QuarkMedSearch: A Long-Horizon Deep Search Agent for Exploring Medical Intelligence

Model ReleasesDGX agent

arXiv:2604.12867v1 Announce Type: new Abstract: As agentic foundation models continue to evolve, how to further improve their performance in vertical domains has become an important challenge. To this

SIRI-Bench: Challenging VLMs' Spatial Intelligence through Complex Reasoning Tasks

Model ReleasesDGX agent

arXiv:2506.14512v4 Announce Type: replace Abstract: Large Language Models (LLMs) have undergone rapid progress, largely attributed to reinforcement learning on complex reasoning tasks. In contrast, wh

← Previous
1…4344454647
Next →