AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “synthesis”

GridTimelineEvolution
2,818 results
15 Jul 2026

PluRel: Synthetic Data unlocks Scaling Laws for Relational Foundation Models

TutorialsDGX agent

arXiv:2602.04029v2 Announce Type: replace-cross Abstract: Relational Foundation Models (RFMs) facilitate data-driven decision-making by learning from complex multi-table databases. However, the divers

Steering Diffusion Models via Class-Contrastive Influence for Few-Shot Medical Classification

ResearchDGX agent

arXiv:2607.12464v1 Announce Type: new Abstract: When labeled data are scarce, off-the-shelf diffusion models can augment training sets for few-shot medical image classification, but not all generated

tencent/Hy-Embodied-RxBrain-1.0 · Hugging Face

Model ReleasesDGX agent

Introduction RxBrain (Hy-Embodied-RxBrain-1.0) is a unified multimodal foundation model for embodied cognition — a single model that couples language reasoning with visual imagination to deliver three


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
10 Jul 2026

ARDY: Autoregressive Diffusion with Hybrid Representation for Interactive Human Motion Generation

Model ReleasesDGX agent

arXiv:2607.08741v1 Announce Type: cross Abstract: Generating realistic 3D human motions in real-time within interactive applications is key for animation, simulation, and humanoid robotics. While rece

Cross-Modal Generative Framework for Signal Translation from Fetal-Maternal Electrocardiograms to Fetal Doppler Waveforms

ResearchDGX agent

arXiv:2607.08073v1 Announce Type: new Abstract: Fetal electrocardiogram (fECG) and Doppler ultrasound provide complementary views of fetal cardiovascular function: fECG captures electrical activity wh

From Solvers to Research: Large Language Model-Driven Formal Mathematics at the Research Frontier

ResearchDGX agent

arXiv:2607.07779v1 Announce Type: cross Abstract: Recent developments in AI for Mathematics (AI4Math), especially Large Language Model (LLM)-driven theorem provers, has achieved remarkable success in

FunHOI: Annotation-Free 3D Hand-Object Interaction Generation via Functional Text Guidance

ResearchDGX agent

arXiv:2502.20805v3 Announce Type: replace-cross Abstract: Hand-object interaction(HOI) is the fundamental link between human and environment, yet its dexterous and complex pose significantly challenge

How Do I Know What to Say Next? Barenholtz's Autogenerative Theory as an Enrichment of Harrisean Integrationism

ResearchDGX agent

arXiv:2607.07891v1 Announce Type: cross Abstract: Roy Harris's Integrationist linguistics offers a compelling critique of the referentialist tradition embedded deep at the heart of computational appro

HumanForge: A Human-Centric Deepfake Video Benchmark with Multi-Agent Forgery Rationales

Model ReleasesDGX agent

arXiv:2607.08705v1 Announce Type: new Abstract: Rapid advancements in video diffusion models and temporal editing tools have enabled the generation of highly realistic human-centric videos, posing unp

LSRM: High-Fidelity Object-Centric Reconstruction via Scaled Context Windows

Model ReleasesDGX agent

arXiv:2604.05182v2 Announce Type: replace-cross Abstract: We introduce the Large Sparse Reconstruction Model to study how scaling transformer context windows affects feed-forward 3D reconstruction. Al

On the Design of Mixture-of-Experts for Dynamic Gaussian Splatting

ApplicationsDGX agent

arXiv:2607.08250v1 Announce Type: new Abstract: Dynamic scene reconstruction remains challenging due to the heterogeneous and spatially varying nature of real-world motion. Although recent 3D Gaussian

Reaction-network reasoning with frontier models for experimentally confirmed catalyst-selectivity hypotheses

ApplicationsDGX agent

arXiv:2607.08003v1 Announce Type: cross Abstract: Catalysts are essential for sustainable chemical manufacturing, yet discovering novel architectures remains a bottleneck dominated by trial-and-error

Selective Left-Shift: Turning Test-Time Compute and Difficulty-based Curation into Training Data for Low-Resource Code Generation

ResearchDGX agent

arXiv:2607.07748v1 Announce Type: new Abstract: Large Language Models achieve strong code generation for high resource languages like Python and Java but suffer sharp performance drops on Low-Resource

SPHINX: First Explain, Then Explore

SafetyDGX agent

arXiv:2606.17482v2 Announce Type: replace Abstract: Generating adversarial driving scenarios is critical for evaluating and improving autonomous vehicle decision-making systems in simulation. Recent a

Tool-Making and Self-Evolving LLM Agents in Low-Latency Systems

AgentsDGX agent

arXiv:2607.08010v1 Announce Type: new Abstract: Production LLM agents often waste latency and reliability by regenerating code for the same procedural steps on every request. We replace this inference

9 Jul 2026

Behavior Foundations for Quadruped Robots: ABot-C0 Technical Report

SafetyDGX agent

arXiv:2607.07370v1 Announce Type: cross Abstract: In embodied intelligence systems, the motion controller serves as the critical bridge between semantic reasoning and physical execution. Humanoid cont

CompDiff: Hierarchical Compositional Diffusion for Fair and Zero-Shot Intersectional Medical Image Generation

Model ReleasesDGX agent

arXiv:2603.16551v2 Announce Type: replace-cross Abstract: Generative models are increasingly used to augment medical imaging datasets for fairer AI, yet a key assumption often goes unexamined: that ge

Explain Before You Answer: A Survey on Compositional Visual Reasoning

Model ReleasesDGX agent

arXiv:2508.17298v3 Announce Type: replace-cross Abstract: Compositional visual reasoning has emerged as a key research frontier in multimodal AI, aiming to endow machines with the human-like ability t

GeoGS-SLAM: Geometry-Only Gaussian Splatting for Dense Monocular SLAM

ApplicationsDGX agent

arXiv:2607.07452v1 Announce Type: new Abstract: Dense visual SLAM is a fundamental problem in robotics. Recent advances in 3DGS have demonstrated its potential for dense SLAM. Existing 3DGS frameworks

R^3: Advertisement Compliance Rectification via Group-Relative Experience Extractor and Curriculum Reinforcement

SafetyDGX agent

arXiv:2607.07318v1 Announce Type: new Abstract: Rigorous content moderation is crucial for online advertising but leads to millions of daily rejections. This scale renders manual rectification infeasi

SoccerNet 2026 Challenges Results

ApplicationsDGX agent

arXiv:2607.07320v1 Announce Type: new Abstract: The SoccerNet 2026 Challenges constitute the sixth annual edition of the SoccerNet open benchmarking effort, dedicated to advancing computer vision rese

Statistical inverse learning and ell^1-regularization

Model ReleasesDGX agent

arXiv:2607.07468v1 Announce Type: cross Abstract: We study the recovery of sparse functions from finite, noisy, and indirect observations in the framework of statistical inverse learning. The unknown

Successor-Generator Planning with LLM-generated Heuristics

ResearchDGX agent

arXiv:2501.18784v5 Announce Type: replace Abstract: Heuristics are a central component of deterministic planning, particularly in domain-independent settings where general applicability is prioritized

The Blind Curator: How a Biased Judge Silently Disables Skill Retirement in Self-Evolving Agents

SafetyDGX agent

arXiv:2607.07436v1 Announce Type: new Abstract: A self-evolving agent retires its bad skills by watching them fail, so what happens when the judge cannot see the failures? Skill retirement is the stru

Tree-of-Thoughts Reasoning for Text-to-Image In-Context Learning

Model ReleasesDGX agent

arXiv:2607.07117v1 Announce Type: cross Abstract: In text-to-image in-context learning (T2I-ICL), a model has to infer a latent compositional pattern from fewshot demonstrations for generating a query

8 Jul 2026

DBNN: Neural Spike Classification Using a Deep Binarized Neural Network

HardwareDGX agent

arXiv:2607.05590v1 Announce Type: cross Abstract: Implantable brain-computer interfaces require on-node spike sorting to reduce telemetry bandwidth and power while maintaining reliable neural decoding

Demonstrating TOFFEE: A Learned System for Synthesizing Data Agent Trajectories at Scale

AgentsDGX agent

arXiv:2607.06233v1 Announce Type: new Abstract: LLM-powered data agents are playing an increasingly important role in data-driven decision making. However, existing data agents struggle to generalize

DeSeG: Decoupling Semantic Intent and Geometric Constraints for Physically Plausible Human-Scene Interaction

SafetyDGX agent

arXiv:2607.05787v1 Announce Type: new Abstract: Synthesizing physically plausible human-scene interactions (HSI) remains a critical challenge in computer vision and the development of human avatars. A

DreamPartGen: Semantically Grounded Part-Level 3D Generation via Collaborative Latent Denoising

SafetyDGX agent

arXiv:2603.19216v2 Announce Type: replace-cross Abstract: Understanding and generating 3D objects as compositions of meaningful parts is fundamental to human perception and reasoning. However, most te

EvalLoop: A Methodology for Evaluation-Driven Iterative Improvement of Business AI Systems

AgentsDGX agent

arXiv:2607.05638v1 Announce Type: cross Abstract: Teams deploying large language models in business contexts need evaluation systems, yet most treat evaluation as static model selection: run benchmark

Foundation Models for Automatic CAD Generation

Model ReleasesDGX agent

arXiv:2607.05573v1 Announce Type: new Abstract: Recent advances in Large Language Models (LLMs) and Vision-Language Models (VLMs) enable the automatic generation of parametric 3D designs from natural-

From RGB Generation to Dense Field Readout: Pixel-Space Dense Prediction with Text-to-Image Models

ResearchDGX agent

arXiv:2607.06553v1 Announce Type: new Abstract: Large-scale text-to-image models are attractive backbones for dense prediction because RGB generation pretraining learns rich semantic, structural, and

High-Resolution Artwork Outpainting with Global Blueprint Guidance and Layout Control

Local AiDGX agent

arXiv:2607.06162v1 Announce Type: new Abstract: Image outpainting extends an image beyond its original borders, requiring seamless style integration and globally coherent scene completion. Building on

Image2Sim: Scaling Embodied Navigation via Generative Neural Simulator

ApplicationsDGX agent

arXiv:2607.05765v1 Announce Type: new Abstract: Embodied navigation aims to build agents that interpret multimodal goals, reason in 3D space, and reach target destinations reliably in the real world.

KernelEvolve: Scaling Agentic Kernel Coding for Heterogeneous AI Accelerators at Meta

SafetyDGX agent

arXiv:2512.23236v4 Announce Type: replace-cross Abstract: Making deep learning recommendation model (DLRM) training and inference fast and efficient is important. However, this presents three key syst

MorphGS: Morphology-Adaptive Articulated 3D Motion Transfer from Videos

ApplicationsDGX agent

arXiv:2601.02716v3 Announce Type: replace Abstract: Transferring articulated motion from monocular videos to rigged 3D characters is challenging due to pose ambiguity in 2D observations and morphologi

PIPBench: A Profile-Inclusive Framework for Personalized Image Generation Evaluation

Model ReleasesDGX agent

arXiv:2607.06440v1 Announce Type: new Abstract: Recent text-to-image models such as DALLE-3 excel at following diverse prompts yet remain blind to individual aesthetic preferences. We study personaliz

RayRoPE: Projective Ray Positional Encoding for Multi-view Attention

ResearchDGX agent

arXiv:2601.15275v3 Announce Type: replace Abstract: We study positional encodings for multi-view transformers that process tokens from a set of posed input images, and seek a mechanism that encodes pa

Recovering Cloud Microstructures with Cascaded Diffusion Inversion

SafetyDGX agent

arXiv:2607.05637v1 Announce Type: new Abstract: High-resolution satellite imagery is critical for observing fine-scale cloud structures that inform weather modification strategies like cloud seeding f

SparseCtrl-HOI: Sparse Temporal Control for Human-Object Interaction Video Generation

ResearchDGX agent

arXiv:2607.05994v1 Announce Type: new Abstract: Human-Object Interaction (HOI) video generation aims to synthesize realistic videos of humans manipulating diverse objects, serving as a promising avenu

Volumetric Directional Diffusion: Anchoring Uncertainty Quantification in Anatomical Consensus for Ambiguous Medical Image Segmentation

SafetyDGX agent

arXiv:2603.04024v2 Announce Type: replace-cross Abstract: Ambiguous 3D medical image segmentation often involves boundaries where different expert delineations are non-identical yet clinically plausib

WordVoice: Explicit and Decoupled Multi-Dimensional Word-Level Control for LLM-Based TTS

SafetyDGX agent

arXiv:2607.06461v1 Announce Type: cross Abstract: While recent Large Language Model (LLM)-based Text-to-Speech (TTS) systems have achieved remarkable naturalness, they predominantly rely on implicit e

7 Jul 2026

ASSEMCAD: Production-Ready CAD Assembly Generation from Natural Language

ApplicationsDGX agent

arXiv:2607.05123v1 Announce Type: new Abstract: Recent advances in large language models and programmatic CAD have significantly improved Text-to-CAD generation for individual parts. However, producti

AudioX-Turbo: A Unified Framework for Efficient Anything-to-Audio Generation

Model ReleasesDGX agent

arXiv:2606.12555v2 Announce Type: replace-cross Abstract: Audio and music generation based on flexible multimodal control signals is a widely applicable topic, with the following key challenges: 1) a

Awakening Diffusion Transformers: Eliciting Stronger Generation and Understanding via Massive Activation Modulation

ResearchDGX agent

arXiv:2607.02968v1 Announce Type: new Abstract: Massive Activations (MAs) have been widely observed in Transformer-based models, yet their structure and functional roles in Diffusion Transformers (DiT

City-Level 3D Surface Reconstruction with Viewpoint Orientation Partitioning and Scene Completion

ResearchDGX agent

arXiv:2607.03771v1 Announce Type: new Abstract: Multi-view 3D surface reconstruction is a longstanding challenge in computer vision. Although recent large-scale reconstruction methods based on 3D Gaus

Deep Learning for Semen Analysis in Male Infertility: Computer Vision, Multimodal Fusion, and Clinical Translation

ApplicationsDGX agent

arXiv:2607.05311v1 Announce Type: new Abstract: Male infertility contributes substantially to the global infertility burden, and sperm analysis remains central to diagnosis, treatment planning, and as

DELTA-TTS: Adapting Autoregressive Model into Diffusion Language Model for Text-to-Speech

Local AiDGX agent

arXiv:2607.04140v1 Announce Type: cross Abstract: Autoregressive (AR) text-to-speech (TTS) models generate discrete speech tokens sequentially, which makes inference slow and can degrade robustness by

ELiTeFormer: An Efficient Transformer for FPGAs

Model ReleasesDGX agent

arXiv:2607.03652v1 Announce Type: cross Abstract: Transformer blocks are prevalent in large language model (LLM) but present deployment challenges due to their challenging computational and memory dem

Enhancing Facial Expression Recognition in Head-Mounted Displays with Synthetic Data

SafetyDGX agent

arXiv:2607.04490v1 Announce Type: new Abstract: Facial expression recognition (FER) is crucial for social interaction in mixed reality environments that employ head-mounted displays (HMD). However, co

Few-Shot Demonstration-Driven Task Coordination and Trajectory Execution for Multi-Robot Systems

SafetyDGX agent

arXiv:2510.15686v2 Announce Type: replace Abstract: Learning coordinated behaviors for multi-robot systems from only a few demonstrations is difficult because temporal task dependencies and spatial tr

Flex-Forcing: Towards a Unified Autoregressive and Bidirectional Video Diffusion Model

SafetyDGX agent

arXiv:2607.03509v1 Announce Type: new Abstract: Recent progress in large-scale generative models has substantially advanced video generation, yet existing methods remain constrained by a rigid inferen

FORGE: Research-Trajectory Hijacking Attacks on Deep Research Agents

AgentsDGX agent

arXiv:2607.04718v1 Announce Type: new Abstract: Deep research agents decompose open-ended queries into subtasks, retrieve web evidence over multiple rounds, and synthesize long-form reports. This work

GenShin: Guiding Rational Liposome Design by Ranking Liposomal Protein Corona through a Docking-Pose-Free GNN

Model ReleasesDGX agent

arXiv:2504.13853v2 Announce Type: replace-cross Abstract: Rational design of lipid nanoparticles (LNPs) for tissue-specific delivery critically depends on predicting the composition of the protein cor

InverseCrafter: Efficient Video ReCapture as a Latent Domain Inverse Problem

ResearchDGX agent

arXiv:2512.05672v2 Announce Type: replace-cross Abstract: Recent approaches in controllable novel view video generation often rely on fine-tuning pre-trained Video Diffusion Models (VDMs). This domina

iVISION-2DCD: A Long-Term Change Detection Dataset for Large-Scale Outdoor Construction Monitoring

Model ReleasesDGX agent

arXiv:2607.03553v1 Announce Type: new Abstract: Automation in construction is essential for reducing costs and human errors in large-scale projects. We approach the construction progress monitoring fr

Measuring the Robustness of Audio Deepfake Detection under Real-World Corruption

ApplicationsDGX agent

arXiv:2503.17577v2 Announce Type: replace-cross Abstract: Deepfakes have emerged as a widespread and rapidly escalating concern in generative AI, spanning images, audio, and videos. Among these, audio

Medical Heuristic Learning: An LLM-Driven Framework for Interpretable and Auditable Clinical Decision Rules

ResearchDGX agent

arXiv:2606.16337v3 Announce Type: replace Abstract: Predictive modeling for clinical decision support requires not only strong predictive performance but also transparent decision logic. Although deep

ProLaViT: Learning Progressive Latent Visual Thoughts in Structured Latent Space

ResearchDGX agent

arXiv:2607.02907v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable progress but still struggle with complex visual reasoning tasks requiring multi-step

Sparse-View Surface Reconstruction using Gaussian Splatting through High-Confidence Depth Propagation with Normal Priors

ResearchDGX agent

arXiv:2607.03765v1 Announce Type: new Abstract: 3D reconstruction from sparse views is a challenging task in 3D computer vision. Recent studies on 3D Gaussian Splatting (3DGS) have achieved remarkable

← Previous
1…2627282930…47
Next →