AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “synthesis”

GridTimelineEvolution
2,818 results
28 Apr 2026

Sliding Mode Control for Safe Trajectory Tracking with Moving Obstacles Avoidance: Experimental Validation on Planar Robots

SafetyDGX agent

arXiv:2604.24518v1 Announce Type: cross Abstract: This paper presents a unified control framework for robust trajectory tracking and moving obstacle avoidance applicable to a broad class of mobile rob

Spatiotemporal Degradation-Aware 3D Gaussian Splatting for Realistic Underwater Scene Reconstruction

Model ReleasesDGX agent

arXiv:2604.23551v1 Announce Type: new Abstract: Reconstructing realistic underwater scenes from underwater video remains a meaningful yet challenging task in the multimedia domain. The inherent spatio

Talker-T2AV: Joint Talking Audio-Video Generation with Autoregressive Diffusion Modeling

ResearchDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.23586v1 Announce Type: cross Abstract: Joint audio-video generation models have shown that unified generation yields stronger cross-modal coherence than cascaded approaches. However, existi

Talking Slide Avatars: Open-Source Multimodal Communication Approach for Teaching

ApplicationsDGX agent

arXiv:2604.23703v1 Announce Type: cross Abstract: Slide-based teaching is widely used in higher education, yet in online, hybrid, and asynchronous contexts, slides often lose the instructor presence,

Towards High-Fidelity CAD Generation via LLM-Driven Program Generation and Text-Based B-Rep Primitive Grounding

ApplicationsDGX agent

arXiv:2603.11831v2 Announce Type: replace Abstract: The field of Computer-Aided Design (CAD) generation has made significant progress in recent years. Existing methods typically fall into two separate

V-GRPO: Online Reinforcement Learning for Denoising Generative Models Is Easier than You Think

SafetyDGX agent

arXiv:2604.23380v1 Announce Type: cross Abstract: Aligning denoising generative models with human preferences or verifiable rewards remains a key challenge. While policy-gradient online reinforcement

27 Apr 2026

An Artifact-based Agent Framework for Adaptive and Reproducible Medical Image Processing

Model ReleasesDGX agent

arXiv:2604.21936v1 Announce Type: new Abstract: Medical imaging research is increasingly shifting from controlled benchmark evaluation toward real-world clinical deployment. In such settings, applying

An Interdisciplinary and Cross-Task Review on Missing Data Imputation

ApplicationsDGX agent

arXiv:2511.01196v3 Announce Type: replace-cross Abstract: Missing data is a fundamental challenge in data science, significantly hindering analysis and decision-making across a wide range of disciplin

Eidolon: A Post-Quantum Signature Scheme Based on k-Colorability in the Age of Graph Neural Networks

ResearchDGX agent

arXiv:2602.02689v2 Announce Type: replace-cross Abstract: We propose Eidolon, a post-quantum signature scheme grounded on the NP-complete k-colorability problem. Our construction generalizes the Goldr

Holo360D: A Large-Scale Real-World Dataset with Continuous Trajectories for Advancing Panoramic 3D Reconstruction and Beyond

Model ReleasesDGX agent

arXiv:2604.22482v1 Announce Type: new Abstract: While feed-forward 3D reconstruction models have advanced rapidly, they still exhibit degraded performance on panoramas due to spherical distortions. Mo

Learning Reactive Human Motion Generation from Paired Interaction Data Using Transformer-Based Models

AgentsDGX agent

arXiv:2604.22164v1 Announce Type: new Abstract: Recent advances in deep learning have enabled the generation of videos from textual descriptions as well as the prediction of future sequences from inpu

OccDirector: Language-Guided Behavior and Interaction Generation in 4D Occupancy Space

Model ReleasesDGX agent

arXiv:2604.22240v1 Announce Type: new Abstract: Generative world models increasingly rely on 4D occupancy for realistic autonomous driving simulation. However, existing generation frameworks depend on

RedVLA: Physical Red Teaming for Vision-Language-Action Models

SafetyDGX agent

arXiv:2604.22591v1 Announce Type: new Abstract: The real-world deployment of Vision-Language-Action (VLA) models remains limited by the risk of unpredictable and irreversible physical harm. However, w

RouteLMT: Learned Sample Routing for Hybrid LLM Translation Deployment

ResearchDGX agent

arXiv:2604.22520v1 Announce Type: new Abstract: Large Language Models (LLMs) have achieved remarkable performance in Machine Translation (MT), but deploying them at scale remains prohibitively expensi

SHAPE: Unifying Safety, Helpfulness and Pedagogy for Educational LLMs

Model ReleasesDGX agent

arXiv:2604.22134v1 Announce Type: new Abstract: Large Language Models (LLMs) have been widely explored in educational scenarios. We identify a critical vulnerability in current educational LLMs, pedag

SOLAR-RL: Semi-Online Long-horizon Assignment Reinforcement Learning

AgentsDGX agent

arXiv:2604.22558v1 Announce Type: cross Abstract: As Multimodal Large Language Models (MLLMs) mature, GUI agents are evolving from static interactions to complex navigation. While Reinforcement Learni

Superminds Test: Actively Evaluating Collective Intelligence of Agent Society via Probing Agents

AgentsDGX agent

arXiv:2604.22452v1 Announce Type: new Abstract: Collective intelligence refers to the ability of a group to achieve outcomes beyond what any individual member can accomplish alone. As large language m

TTS-PRISM: A Perceptual Reasoning and Interpretable Speech Model for Fine-Grained Diagnosis

SafetyDGX agent

arXiv:2604.22225v1 Announce Type: new Abstract: While generative text-to-speech (TTS) models approach human-level quality, monolithic metrics fail to diagnose fine-grained acoustic artifacts or explai

UniSonate: A Unified Model for Speech, Music, and Sound Effect Generation with Text Instructions

ResearchDGX agent

arXiv:2604.22209v1 Announce Type: cross Abstract: Generative audio modeling has largely been fragmented into specialized tasks, text-to-speech (TTS), text-to-music (TTM), and text-to-audio (TTA), each

24 Apr 2026

Agentic AI for Personalized Physiotherapy: A Multi-Agent Framework for Generative Video Training and Real-Time Pose Correction

AgentsDGX agent

arXiv:2604.21154v1 Announce Type: new Abstract: At-home physiotherapy compliance remains critically low due to a lack of personalized supervision and dynamic feedback. Existing digital health solution

An Overlay Multicast Routing Method Based on Network Situational Awareness and Hierarchical Multi-Agent Reinforcement Learning

AgentsDGX agent

arXiv:2602.13211v2 Announce Type: replace-cross Abstract: Compared with IP multicast, Overlay Multicast (OM) offers better compatibility and flexible deployment in heterogeneous, cross-domain networks

Architectures for Robust Self-Organizing Energy Systems under Information and Control Constraints

AgentsDGX agent

arXiv:2604.21529v1 Announce Type: cross Abstract: Applying the concept of controlled self-organization in agent-based Cyber-Physical Energy Systems (CPES) is a promising approach to ensure system robu

AttDiff-GAN: A Hybrid Diffusion-GAN Framework for Facial Attribute Editing

SafetyDGX agent

arXiv:2604.21289v1 Announce Type: new Abstract: Facial attribute editing aims to modify target attributes while preserving attribute-irrelevant content and overall image fidelity. Existing GAN-based m

BiTDiff: Fine-Grained 3D Conducting Motion Generation via BiMamba-Transformer Diffusion

SafetyDGX agent

arXiv:2604.04395v2 Announce Type: replace Abstract: 3D conducting motion generation aims to synthesize fine-grained conductor motions from music, with broad potential in music education, virtual perfo

C-SHAP for time series: An approach to high-level temporal explanations

ApplicationsDGX agent

arXiv:2504.11159v2 Announce Type: replace Abstract: In high-stakes domains, such as healthcare and industry, the explainability of AI-based decision-making has become crucial. Without insight into mod

EngramaBench: Evaluating Long-Term Conversational Memory with Structured Graph Retrieval

Model ReleasesDGX agent

arXiv:2604.21229v1 Announce Type: cross Abstract: Large language model assistants are increasingly expected to retain and reason over information accumulated across many sessions. We introduce Engrama

Enhancing Online Recruitment with Category-Aware MoE and LLM-based Data Augmentation

ResearchDGX agent

arXiv:2604.21264v1 Announce Type: new Abstract: Person-Job Fit (PJF) is a critical component for online recruitment. Existing approaches face several challenges, particularly in handling low-quality j

Exploring the Role of Synthetic Data Augmentation in Controllable Human-Centric Video Generation

ApplicationsDGX agent

arXiv:2604.21291v1 Announce Type: cross Abstract: Controllable human video generation aims to produce realistic videos of humans with explicitly guided motions and appearances,serving as a foundation

Full-Body Dynamic Safety for Robot Manipulators: 3D Poisson Safety Functions for CBF-Based Safety Filters

SafetyDGX agent

arXiv:2604.21189v1 Announce Type: new Abstract: Collision avoidance for robotic manipulators requires enforcing full-body safety constraints in high-dimensional configuration spaces. Control Barrier F

GraphLeap: Decoupling Graph Construction and Convolution for Vision GNN Acceleration on FPGA

HardwareDGX agent

arXiv:2604.21290v1 Announce Type: new Abstract: Vision Graph Neural Networks (ViGs) represent an image as a graph of patch tokens, enabling adaptive, feature-driven neighborhoods. Unlike CNNs with fix

Hierarchical Policy Optimization for Simultaneous Translation of Unbounded Speech

SafetyDGX agent

arXiv:2604.21045v1 Announce Type: new Abstract: Simultaneous speech translation (SST) generates translations while receiving partial speech input. Recent advances show that large language models (LLMs

HypEHR: Hyperbolic Modeling of Electronic Health Records for Efficient Question Answering

ApplicationsDGX agent

arXiv:2604.21027v1 Announce Type: new Abstract: Electronic health record (EHR) question answering is often handled by LLM-based pipelines that are costly to deploy and do not explicitly leverage the h

Job Skill Extraction via LLM-Centric Multi-Module Framework

ResearchDGX agent

arXiv:2604.21525v1 Announce Type: new Abstract: Span-level skill extraction from job advertisements underpins candidate-job matching and labor-market analytics, yet generative large language models (L

Learning Long-Term Motion Embeddings for Efficient Kinematics Generation

ResearchDGX agent

Understanding and predicting motion is a fundamental component of visual intelligence. Although modern video models exhibit strong comprehension of scene dynamics, exploring multiple possible futures

PercHead: Perceptual Head Model for Single-Image 3D Head Reconstruction & Editing

ResearchDGX agent

arXiv:2511.02777v2 Announce Type: replace Abstract: We present PercHead, a model for single-image 3D head reconstruction and disentangled 3D editing - two tasks that are inherently challenging due to

Process Supervision via Verbal Critique Improves Reasoning in Large Language Models

Model ReleasesDGX agent

arXiv:2604.21611v1 Announce Type: cross Abstract: Inference-time scaling for LLM reasoning has focused on three axes: chain depth, sample breadth, and learned step-scorers (PRMs). We introduce a fourt

RELOOP: Recursive Retrieval with Multi-Hop Reasoner and Planners for Heterogeneous QA

SafetyDGX agent

arXiv:2510.20505v4 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) remains brittle on multi-step questions and heterogeneous evidence sources, trading accuracy against late

Reshoot-Anything: A Self-Supervised Model for In-the-Wild Video Reshooting

TutorialsDGX agent

arXiv:2604.21776v1 Announce Type: new Abstract: Precise camera control for reshooting dynamic videos is bottlenecked by the severe scarcity of paired multi-view data for non-rigid scenes. We overcome

The CriticalSet problem: Identifying Critical Contributors in Bipartite Dependency Networks

ResearchDGX agent

arXiv:2604.21537v1 Announce Type: new Abstract: Identifying critical nodes in complex networks is a fundamental task in graph mining. Yet, methods addressing an all-or-nothing coverage mechanics in a

The Last Harness You'll Ever Build

AgentsDGX agent

arXiv:2604.21003v1 Announce Type: new Abstract: AI agents are increasingly deployed on complex, domain-specific workflows -- navigating enterprise web applications that require dozens of clicks and fo

Unlocking the Power of Critical Factors for 3D Visual Geometry Estimation

Local AiDGX agent

arXiv:2604.21713v1 Announce Type: new Abstract: Feed-forward visual geometry estimation has recently made rapid progress. However, an important gap remains: multi-frame models usually produce better c

23 Apr 2026

Amodal SAM: A Unified Amodal Segmentation Framework with Generalization

ApplicationsDGX agent

arXiv:2604.20748v1 Announce Type: new Abstract: Amodal segmentation is a challenging task that aims to predict the complete geometric shape of objects, including their occluded regions. Although exist

ChipCraftBrain: Validation-First RTL Generation via Multi-Agent Orchestration

SafetyDGX agent

arXiv:2604.19856v1 Announce Type: cross Abstract: Large Language Models (LLMs) show promise for generating Register-Transfer Level (RTL) code from natural language specifications, but single-shot gene

Computing the Reachability Value of Posterior-Deterministic POMDPs

ResearchDGX agent

arXiv:2602.07473v2 Announce Type: replace Abstract: Partially observable Markov decision processes (POMDPs) are a fundamental model for sequential decision-making under uncertainty. However, many veri

Efficient Reinforcement Learning using Linear Koopman Dynamics for Nonlinear Robotic Systems

SafetyDGX agent

arXiv:2604.19980v1 Announce Type: new Abstract: This paper presents a model-based reinforcement learning (RL) framework for optimal closed-loop control of nonlinear robotic systems. The proposed appro

FlexServe: A Fast and Secure LLM Serving System for Mobile Devices with Flexible Resource Isolation

AgentsDGX agent

arXiv:2603.09046v2 Announce Type: replace-cross Abstract: Device-side Large Language Models (LLMs) have witnessed explosive growth, offering higher privacy and availability compared to cloud-side LLMs

From Diffusion to Flow: Efficient Motion Generation in MotionGPT3

ResearchDGX agent

arXiv:2603.26747v2 Announce Type: replace Abstract: Recent text-driven motion generation methods span both discrete token-based approaches and continuous-latent formulations. MotionGPT3 exemplifies th

GeoRelight: Learning Joint Geometrical Relighting and Reconstruction with Flexible Multi-Modal Diffusion Transformers

ResearchDGX agent

arXiv:2604.20715v1 Announce Type: new Abstract: Relighting a person from a single photo is an attractive but ill-posed task, as a 2D image ambiguously entangles 3D geometry, intrinsic appearance, and

HumanScore: Benchmarking Human Motions in Generated Videos

ResearchDGX agent

arXiv:2604.20157v1 Announce Type: new Abstract: Recent advances in model architectures, compute, and data scale have driven rapid progress in video generation, producing increasingly realistic content

LaplacianFormer:Rethinking Linear Attention with Laplacian Kernel

HardwareDGX agent

arXiv:2604.20368v1 Announce Type: cross Abstract: The quadratic complexity of softmax attention presents a major obstacle for scaling Transformers to high-resolution vision tasks. Existing linear atte

Large language models perceive cities through a culturally uneven baseline

SafetyDGX agent

arXiv:2604.20048v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used to describe, evaluate and interpret places, yet it remains unclear whether they do so from a cultural

LEXIS: LatEnt ProXimal Interaction Signatures for 3D HOI from an Image

ResearchDGX agent

arXiv:2604.20800v1 Announce Type: new Abstract: Reconstructing 3D Human-Object Interaction from an RGB image is essential for perceptive systems. Yet, this remains challenging as it requires capturing

LLM StructCore: Schema-Guided Reasoning Condensation and Deterministic Compilation

ResearchDGX agent

arXiv:2604.20560v1 Announce Type: new Abstract: Automatically filling Case Report Forms (CRFs) from clinical notes is challenging due to noisy language, strict output contracts, and the high cost of f

Lucky High Dynamic Range Smartphone Imaging

ResearchDGX agent

arXiv:2604.19976v1 Announce Type: new Abstract: While the human eye can perceive an impressive twenty stops of dynamic range, smartphone camera sensors remain limited to about twelve stops despite dec

Measuring Creativity in the Age of Generative AI: Distinguishing Human and AI-Generated Creative Performance in Hiring and Talent Systems

ResearchDGX agent

arXiv:2604.19799v1 Announce Type: cross Abstract: Generative AI is rapidly transforming how organizations create value and evaluate talent. While large language models enhance baseline output quality,

🔹 One Prompt → 100-page PDF report +cited dataset + 30-page executive PPT + 20 financial charts

Model ReleasesDGX agent

Moonshot's Kimi AI demonstrated capabilities to generate comprehensive business reports from a single prompt, including a 100-page PDF with citations, a 30-page executive PowerPoint presentation, and

Online Survival Analysis: A Bandit Approach under Cox PH Model

ResearchDGX agent

arXiv:2604.20296v1 Announce Type: cross Abstract: Survival analysis is a widely used statistical framework for modeling time-to-event data under censoring. Classical methods, such as the Cox proportio

Open-source agent for long-horizon deep research https://github.com/TIGER-AI-Lab/OpenResearcher

AgentsDGX agent

OpenResearcher is an open-source AI agent designed to conduct long-horizon deep research tasks, enabling autonomous investigation and analysis across extended research workflows. The project is mainta

ReasonRank: Empowering Passage Ranking with Strong Reasoning Ability

Model ReleasesDGX agent

arXiv:2508.07050v3 Announce Type: replace-cross Abstract: Large Language Model (LLM) based listwise ranking has shown superior performance in many passage ranking tasks. With the development of Large

SolidCoder: Bridging the Mental-Reality Gap in LLM Code Generation through Concrete Execution

ResearchDGX agent

arXiv:2604.19825v1 Announce Type: cross Abstract: State-of-the-art code generation frameworks rely on mental simulation, where LLMs internally trace execution to verify correctness. We expose a fundam

← Previous
1…4041424344…47
Next →