AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “synthesis”

GridTimelineEvolution
2,711 results
Research

HOG-Layout: Hierarchical 3D Scene Generation, Optimization and Editing via Vision-Language Models

DGX agent

arXiv:2604.10772v1 Announce Type: new Abstract: 3D layout generation and editing play a crucial role in Embodied AI and immersive VR interaction. However, manual creation requires tedious labor, while

researcharxiv-cs-cv
14 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

HumanVBench: Probing Human-Centric Video Understanding in MLLMs with Automatically Synthesized Benchmarks

DGX agent

arXiv:2412.17574v3 Announce Type: replace-cross Abstract: Evaluating the nuanced human-centric video understanding capabilities of Multimodal Large Language Models (MLLMs) remains a great challenge, a

model-releasesarxiv-cs-ai
14 Apr 2026
Research

Intra-finger Variability of Diffusion-based Latent Fingerprint Generation

DGX agent

arXiv:2604.10040v1 Announce Type: new Abstract: The primary goal of this work is to systematically evaluate the intra-finger variability of synthetic fingerprints (particularly latent prints) generate

researcharxiv-cs-cv
14 Apr 2026
Model Releases

Knowing What to Stress: A Discourse-Conditioned Text-to-Speech Benchmark

DGX agent

arXiv:2604.10580v1 Announce Type: new Abstract: Spoken meaning often depends not only on what is said, but also on which word is emphasized. The same sentence can convey correction, contrast, or clari

model-releasesarxiv-cs-cl
14 Apr 2026
Research

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation

DGX agent

arXiv:2509.20128v2 Announce Type: replace-cross Abstract: Audio-driven facial animation has made significant progress in multimedia applications, with diffusion models showing strong potential for tal

researcharxiv-cs-ai
14 Apr 2026
Tutorials

Learning Long-term Motion Embeddings for Efficient Kinematics Generation

DGX agent

arXiv:2604.11737v1 Announce Type: new Abstract: Understanding and predicting motion is a fundamental component of visual intelligence. Although modern video models exhibit strong comprehension of scen

tutorialsarxiv-cs-cv
14 Apr 2026
Safety

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment

DGX agent

arXiv:2604.10677v1 Announce Type: cross Abstract: Scaling up robot learning is hindered by the scarcity of robotic demonstrations, whereas human videos offer a vast, untapped source of interaction dat

safetyarxiv-cs-cv
14 Apr 2026
Model Releases

LLMs for Text-Based Exploration and Navigation Under Partial Observability

DGX agent

arXiv:2604.09604v1 Announce Type: new Abstract: Exploration and goal-directed navigation in unknown layouts are central to inspection, logistics, and search-and-rescue. We ask whether large language m

model-releasesarxiv-cs-ai
14 Apr 2026
Tutorials

MorphoFlow: Sparse-Supervised Generative Shape Modeling with Adaptive Latent Relevance

DGX agent

arXiv:2604.11636v1 Announce Type: new Abstract: Statistical shape modeling (SSM) is central to population level analysis of anatomical variability, yet most existing approaches rely on densely annotat

tutorialsarxiv-cs-cv
14 Apr 2026
Model Releases

OccuBench: Evaluating AI Agents on Real-World Professional Tasks via Language World Models

DGX agent

arXiv:2604.10866v1 Announce Type: new Abstract: AI agents are expected to perform professional work across hundreds of occupational domains (from emergency department triage to nuclear reactor safety

model-releasesarxiv-cs-cl
14 Apr 2026
Safety

Online Learning-Enhanced High Order Adaptive Safety Control

DGX agent

arXiv:2511.19651v2 Announce Type: replace Abstract: Control barrier functions (CBFs) are an effective model-based tool to formally certify the safety of a system. With the growing complexity of modern

safetyarxiv-cs-ro
14 Apr 2026
Model Releases

Pioneer Agent: Continual Improvement of Small Language Models in Production

DGX agent

arXiv:2604.09791v1 Announce Type: new Abstract: Small language models are attractive for production deployment due to their low cost, fast inference, and ease of specialization. However, adapting them

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

Prompt Relay: Inference-Time Temporal Control for Multi-Event Video Generation

DGX agent

arXiv:2604.10030v1 Announce Type: new Abstract: Video diffusion models have achieved remarkable progress in generating high-quality videos. However, these models struggle to represent the temporal suc

safetyarxiv-cs-cv
14 Apr 2026
Agents

Representations Before Pixels: Semantics-Guided Hierarchical Video Prediction

DGX agent

arXiv:2604.11707v1 Announce Type: new Abstract: Accurate future video prediction requires both high visual fidelity and consistent scene semantics, particularly in complex dynamic environments such as

agentsarxiv-cs-cv
14 Apr 2026
Model Releases

RiTeK: A Dataset for Large Language Models Complex Reasoning over Textual Knowledge Graphs in Medicine

DGX agent

arXiv:2410.13987v3 Announce Type: replace Abstract: Answering complex real-world questions in the medical domain often requires accurate retrieval from medical Textual Knowledge Graphs (medical TKGs),

model-releasesarxiv-cs-cl
14 Apr 2026
Tutorials

Ro-SLM: Onboard Small Language Models for Robot Task Planning and Operation Code Generation

DGX agent

arXiv:2604.10929v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) provide robots with contextual reasoning abilities to comprehend human instructions. Yet, current LLM-en

tutorialsarxiv-cs-ro
14 Apr 2026
Safety

RoboStereo: Dual-Tower 4D Embodied World Models for Unified Policy Optimization

DGX agent

arXiv:2603.12639v2 Announce Type: replace Abstract: Scalable Embodied AI faces fundamental constraints due to prohibitive costs and safety risks of real-world interaction. While Embodied World Models

safetyarxiv-cs-cv
14 Apr 2026
Safety

Robust Real-Time Coordination of CAVs: A Distributed Optimization Framework under Uncertainty

DGX agent

arXiv:2508.21322v2 Announce Type: replace Abstract: Achieving both safety guarantees and real-time performance in cooperative vehicle coordination remains a fundamental challenge, particularly in dyna

safetyarxiv-cs-ro
14 Apr 2026
Model Releases

Seeing Through Deception: Uncovering Misleading Creator Intent in Multimodal News with Vision-Language Models

DGX agent

arXiv:2505.15489v4 Announce Type: replace-cross Abstract: The impact of multimodal misinformation arises not only from factual inaccuracies but also from the misleading narratives that creators delibe

model-releasesarxiv-cs-cl
14 Apr 2026
Research

SIGMA: An Efficient Heterophilous Graph Neural Network with Fast Global Aggregation

DGX agent

arXiv:2305.09958v5 Announce Type: replace Abstract: Graph neural networks (GNNs) realize great success in graph learning but suffer from performance loss when meeting heterophily, i.e. neighboring nod

researcharxiv-cs-lg
14 Apr 2026
Safety

StaMo: Unsupervised Learning of Generalizable Robot Motion from Compact State Representation

DGX agent

arXiv:2510.05057v2 Announce Type: replace-cross Abstract: A fundamental challenge in embodied intelligence is developing expressive and compact state representations for efficient world modeling and d

safetyarxiv-cs-cv
14 Apr 2026
Model Releases

Switch-JustDance: Benchmarking Whole Body Motion Tracking Controllers Using a Commercial Console Game

DGX agent

arXiv:2511.17925v3 Announce Type: replace-cross Abstract: Recent advances in whole-body robot control have enabled humanoid and legged robots to perform increasingly agile and coordinated motions. How

model-releasesarxiv-cs-cv
14 Apr 2026
Research

The Phantom of PCIe: Constraining Generative Artificial Intelligences for Practical Peripherals Trace Synthesizing

DGX agent

arXiv:2411.06376v3 Announce Type: replace-cross Abstract: Peripheral Component Interconnect Express (PCIe) is the de facto interconnect standard for high-speed peripherals and CPUs. The development of

researcharxiv-cs-ai
14 Apr 2026
Research

UDAPose: Unsupervised Domain Adaptation for Low-Light Human Pose Estimation

DGX agent

arXiv:2604.10485v1 Announce Type: cross Abstract: Low-visibility scenarios, such as low-light conditions, pose significant challenges to human pose estimation due to the scarcity of annotated low-ligh

researcharxiv-cs-ai
14 Apr 2026
Safety

ViserDex: Visual Sim-to-Real for Robust Dexterous In-hand Reorientation

DGX agent

arXiv:2604.11138v1 Announce Type: cross Abstract: In-hand object reorientation requires precise estimation of the object pose to handle complex task dynamics. While RGB sensing offers rich semantic cu

safetyarxiv-cs-cv
14 Apr 2026
Safety

VLMaterial: Vision-Language Model-Based Camera-Radar Fusion for Physics-Grounded Material Identification

DGX agent

arXiv:2604.11671v1 Announce Type: cross Abstract: Accurate material recognition is a fundamental capability for intelligent perception systems to interact safely and effectively with the physical worl

safetyarxiv-cs-ro
14 Apr 2026
Safety

WM-DAgger: Enabling Efficient Data Aggregation for Imitation Learning with World Models

DGX agent

arXiv:2604.11351v1 Announce Type: new Abstract: Imitation learning is a powerful paradigm for training robotic policies, yet its performance is limited by compounding errors: minor policy inaccuracies

safetyarxiv-cs-ro
14 Apr 2026
Local Ai

Active Learning for Generalizable Detonation Performance Prediction of Energetic Materials

DGX agent

arXiv:2604.08744v1 Announce Type: cross Abstract: The discovery of new energetic materials is critical for advancing technologies from defense to private industry. However, experimental approaches rem

local-aiarxiv-cs-lg
13 Apr 2026
Agents

ActivityEditor: Learning to Synthesize Physically Valid Human Mobility

DGX agent

arXiv:2604.05529v2 Announce Type: replace Abstract: Human mobility modeling is indispensable for diverse urban applications. However, existing data-driven methods often suffer from data scarcity, limi

agentsarxiv-cs-ai
13 Apr 2026
Safety

AVA-VLA: Improving Vision-Language-Action models with Active Visual Attention

DGX agent

arXiv:2511.18960v3 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models have shown remarkable progress in embodied tasks recently, but most methods process visual observations in

safetyarxiv-cs-cv
13 Apr 2026
Safety

ChipSeek: Optimizing Verilog Generation via EDA-Integrated Reinforcement Learning

DGX agent

arXiv:2507.04736v2 Announce Type: replace Abstract: Large Language Models have emerged as powerful tools for automating Register-Transfer Level (RTL) code generation, yet they face critical limitation

safetyarxiv-cs-ai
13 Apr 2026
Tutorials

CT-1: Vision-Language-Camera Models Transfer Spatial Reasoning Knowledge to Camera-Controllable Video Generation

DGX agent

arXiv:2604.09201v1 Announce Type: new Abstract: Camera-controllable video generation aims to synthesize videos with flexible and physically plausible camera movements. However, existing methods either

tutorialsarxiv-cs-cv
13 Apr 2026
Research

FDIF: Formula-Driven supervised Learning with Implicit Functions for 3D Medical Image Segmentation

DGX agent

arXiv:2603.23199v2 Announce Type: replace Abstract: Deep learning-based 3D medical image segmentation methods relies on large-scale labeled datasets, yet acquiring such data is difficult due to privac

researcharxiv-cs-cv
13 Apr 2026
Model Releases

From Reasoning to Agentic: Credit Assignment in Reinforcement Learning for Large Language Models

DGX agent

arXiv:2604.09459v1 Announce Type: new Abstract: Reinforcement learning (RL) for large language models (LLMs) increasingly relies on sparse, outcome-level rewards -- yet determining which actions withi

model-releasesarxiv-cs-cl
13 Apr 2026
Safety

GAN-Enhanced Deep Reinforcement Learning for Semantic-Aware Resource Allocation in 6G Network Slicing

DGX agent

arXiv:2604.08576v1 Announce Type: cross Abstract: Sixth-generation (6G) wireless networks must support heterogeneous services: enhanced Mobile Broadband (eMBB) requiring 1 Tbps data rates, massive Mac

safetyarxiv-cs-ai
13 Apr 2026
Safety

Generative Simulation for Policy Learning in Physical Human-Robot Interaction

DGX agent

arXiv:2604.08664v1 Announce Type: new Abstract: Developing autonomous physical human-robot interaction (pHRI) systems is limited by the scarcity of large-scale training data to learn robust robot beha

safetyarxiv-cs-ro
13 Apr 2026
Research

Large-Scale Universal Defect Generation: Foundation Models and Datasets

DGX agent

arXiv:2604.08915v1 Announce Type: cross Abstract: Existing defect/anomaly generation methods often rely on few-shot learning, which overfits to specific defect categories due to the lack of large-scal

researcharxiv-cs-ai
13 Apr 2026
Safety

Multimodal Anomaly Detection for Human-Robot Interaction

DGX agent

arXiv:2604.09326v1 Announce Type: cross Abstract: Ensuring safety and reliability in human-robot interaction (HRI) requires the timely detection of unexpected events that could lead to system failures

safetyarxiv-cs-cv
13 Apr 2026
Safety

Neural Distribution Prior for LiDAR Out-of-Distribution Detection

DGX agent

arXiv:2604.09232v1 Announce Type: cross Abstract: LiDAR-based perception is critical for autonomous driving due to its robustness to poor lighting and visibility conditions. Yet, current models operat

safetyarxiv-cs-ai
13 Apr 2026
Model Releases

R2G: A Multi-View Circuit Graph Benchmark Suite from RTL to GDSII

DGX agent

arXiv:2604.08810v1 Announce Type: new Abstract: Graph neural networks (GNNs) are increasingly applied to physical design tasks such as congestion prediction and wirelength estimation, yet progress is

model-releasesarxiv-cs-cv
13 Apr 2026
Model Releases

SAGE: A Service Agent Graph-guided Evaluation Benchmark

DGX agent

arXiv:2604.09285v1 Announce Type: new Abstract: The development of Large Language Models (LLMs) has catalyzed automation in customer service, yet benchmarking their performance remains challenging. Ex

model-releasesarxiv-cs-ai
13 Apr 2026
Safety

Scene-Agnostic Object-Centric Representation Learning for 3D Gaussian Splatting

DGX agent

arXiv:2604.09045v1 Announce Type: new Abstract: Recent works on 3D scene understanding leverage 2D masks from visual foundation models (VFMs) to supervise radiance fields, enabling instance-level 3D s

safetyarxiv-cs-cv
13 Apr 2026
Agents

SkillForge: Forging Domain-Specific, Self-Evolving Agent Skills in Cloud Technical Support

DGX agent

arXiv:2604.08618v1 Announce Type: cross Abstract: Deploying LLM-powered agents in enterprise scenarios such as cloud technical support demands high-quality, domain-specific skills. However, existing s

agentsarxiv-cs-ai
13 Apr 2026
Research

Structured Exploration and Exploitation of Label Functions for Automated Data Annotation

DGX agent

arXiv:2604.08578v1 Announce Type: cross Abstract: High-quality labeled data is critical for training reliable machine learning and deep learning models, yet manual annotation remains costly and error-

researcharxiv-cs-ai
13 Apr 2026
Applications

SynDocDis: A Metadata-Driven Framework for Generating Synthetic Physician Discussions Using Large Language Models

DGX agent

arXiv:2604.08555v1 Announce Type: new Abstract: Physician-physician discussions of patient cases represent a rich source of clinical knowledge and reasoning that could feed AI agents to enrich and eve

applicationsarxiv-cs-cl
13 Apr 2026
Safety

Traj2Action: A Co-Denoising Framework for Trajectory-Guided Human-to-Robot Skill Transfer

DGX agent

arXiv:2510.00491v3 Announce Type: replace-cross Abstract: Learning diverse manipulation skills for real-world robots is severely bottlenecked by the reliance on costly and hard-to-scale teleoperated d

safetyarxiv-cs-ai
13 Apr 2026
Research

WAND: Windowed Attention and Knowledge Distillation for Efficient Autoregressive Text-to-Speech Models

DGX agent

arXiv:2604.08558v1 Announce Type: cross Abstract: Recent decoder-only autoregressive text-to-speech (AR-TTS) models produce high-fidelity speech, but their memory and compute costs scale quadratically

researcharxiv-cs-ai
13 Apr 2026
Research

Zero-Shot Generative De-identification: Inversion-Free Flow for Privacy-Preserving Skin Image Analysis

DGX agent

arXiv:2602.00821v2 Announce Type: replace Abstract: The secure analysis of dermatological images in clinical environments is fundamentally restricted by the critical trade-off between patient privacy

researcharxiv-cs-cv
13 Apr 2026
← Previous
1…54555657
Next →