AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “latent-space”

GridTimelineEvolution
293 results
15 May 2026

Unsupervised learning of acquisition variability in structural connectomes via hybrid latent space modeling

Model ReleasesDGX agent

arXiv:2605.13933v1 Announce Type: cross Abstract: Acquisition differences across sites, scanners, and protocols in dMRI introduce variability that complicates structural connectome analysis. This moti

14 May 2026

ArcVQ-VAE: A Spherical Vector Quantization Framework with ArcCosine Additive Margin

ResearchDGX agent

arXiv:2605.13517v1 Announce Type: cross Abstract: Vector Quantized Variational Autoencoder (VQ-VAE) has become a fundamental framework for learning discrete representations in image modeling. However,

Make-It-Poseable: Feed-forward Latent Posing Model for 3D Characters


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
ResearchDGX agent

arXiv:2512.16767v2 Announce Type: replace Abstract: Posing 3D characters is a fundamental task in computer graphics. However, existing paradigms, ranging from traditional auto-rigging to recent pose-c

13 May 2026

G^2TR: Generation-Guided Visual Token Reduction for Separate-Encoder Unified Multimodal Models

ResearchDGX agent

arXiv:2605.12309v1 Announce Type: new Abstract: The development of separate-encoder Unified multimodal models (UMMs) comes with a rapidly growing inference cost due to dense visual token processing. I

Martingale-Consistent Self-Supervised Learning

ResearchDGX agent

arXiv:2605.11846v1 Announce Type: new Abstract: Self-supervised learning (SSL) is often deployed under changing information, such as shorter histories, missing features, or partially observed images.

12 May 2026

Clin-JEPA: A Multi-Phase Co-Training Framework for Joint-Embedding Predictive Pretraining on EHR Patient Trajectories

Model ReleasesDGX agent

arXiv:2605.10840v1 Announce Type: cross Abstract: We present Clin-JEPA, a multi-phase co-training framework for joint-embedding predictive (JEPA) pretraining on EHR patient trajectories. JEPA architec

CoLA-Flow Policy: Temporally Coherent Imitation Learning via Continuous Latent Action Flow Matching for Robotic Manipulation

SafetyDGX agent

arXiv:2601.23087v3 Announce Type: replace Abstract: Learning long-horizon robotic manipulation requires jointly achieving expressive behavior modeling, real-time inference, and stable execution, which

Deep Dreams Are Made of This: Visualizing Monosemantic Features in Diffusion Models

ResearchDGX agent

arXiv:2605.08218v1 Announce Type: cross Abstract: This paper proposes latent visualization by optimization (LVO), a mechanistic interpretability technique that extends feature visualization by optimiz

From Syntax to Semantics: Unveiling the Emergence of Chirality in SMILES Translation Models

TutorialsDGX agent

arXiv:2605.09949v1 Announce Type: new Abstract: Understanding how chemical language models (CLMs) learn chemical meaning from molecular string representations, rather than only surface-level string pa

Why Do DiT Editors Drift? Plug-and-Play Low Frequency Alignment in VAE Latent Space

SafetyDGX agent

arXiv:2605.08250v1 Announce Type: cross Abstract: Recent advances in diffusion transformers (DiTs) have enabled promising single-turn image editing capabilities. However, multi-turn editing often lead

11 May 2026

ProteinJEPA: Latent prediction complements protein language models

ResearchDGX agent

arXiv:2605.07554v1 Announce Type: cross Abstract: Protein language models are trained primarily with masked language modeling (MLM), which predicts amino-acid identities at masked positions. We ask wh

7 May 2026

Geometry-Aware Neural Optimizer for Shape Optimization and Inversion

ResearchDGX agent

arXiv:2605.04474v1 Announce Type: new Abstract: Geometry is central to PDE-governed systems, motivating shape optimization and inversion. Classical pipelines conduct costly forward simulation with geo

6 May 2026

Pretrained Model Representations as Acquisition Signals for Active Learning of MLIPs

ResearchDGX agent

arXiv:2605.03964v1 Announce Type: new Abstract: Training machine learning interatomic potentials (MLIPs) for reactive chemistry is often bottlenecked by the high cost of quantum chemical labels and th

5 May 2026

Latent Space Probing for Adult Content Detection in Video Generative Models

ResearchDGX agent

arXiv:2605.00874v1 Announce Type: new Abstract: The rapid proliferation of AI-powered video generation systems has introduced significant challenges in content moderation, particularly with respect to

Toward a Scientific Discovery Engine for Weather and Climate Data: A Visual Analytics Workbench for Embedding-Based Exploration

SafetyDGX agent

arXiv:2605.00972v1 Announce Type: cross Abstract: Earth system science is producing increasingly large, high-dimensional datasets from physics based Earth system models to AI-based weather and climate

4 May 2026

Latent Generative Modeling of Random Fields from Limited Training Data

TutorialsDGX agent

arXiv:2505.13007v2 Announce Type: replace Abstract: The ability to accurately model random fields plays a critical role in science and engineering for problems involving uncertain, spatially-varying q

1 May 2026

Data-Efficient Indentation Size Effect Correction in Steels Using Machine Learning and Physics-Guided Augmentation

ResearchDGX agent

arXiv:2604.27775v1 Announce Type: cross Abstract: Shallow nanoindentation enables mechanical characterization of thin films, individual phases and other volume-constrained materials, but measured hard

30 Apr 2026

Exploring the Potential of Probabilistic Transformer for Time Series Modeling: A Report on the ST-PT Framework

ResearchDGX agent

arXiv:2604.26762v1 Announce Type: cross Abstract: The Probabilistic Transformer (PT) establishes that the Transformer's self-attention plus its feed-forward block is mathematically equivalent to Mean-

Latent Autoencoder Ensemble Kalman Filter for Nonlinear Data assimilation

ResearchDGX agent

arXiv:2603.06752v2 Announce Type: replace Abstract: The ensemble Kalman filter (EnKF) is widely used for data assimilation in high-dimensional systems, but its performance often deteriorates for stron

29 Apr 2026

Categorical Optimization with Bayesian Anchored Latent Trust Regions for Structural Design under High-Dimensional Uncertainty

ResearchDGX agent

arXiv:2604.25241v1 Announce Type: new Abstract: Categorical structural optimization under aleatoric uncertainty is challenging because each design variable must be selected from a finite catalog of ad

Recursive Multi-Agent Systems

AgentsDGX agent

arXiv:2604.25917v1 Announce Type: cross Abstract: Recursive or looped language models have recently emerged as a new scaling axis by iteratively refining the same model computation over latent states

28 Apr 2026

DeepSignature: Digitally Signed, Content-Encoding Watermarks for Robust and Transparent Image Authentication

Local AiDGX agent

arXiv:2604.23016v1 Announce Type: cross Abstract: AI-powered generative models have significantly expanded the possibilities for editing, manipulating, and creating high-quality images. Particularly,

Information bottleneck for learning the phase space of dynamics from high-dimensional experimental data

TutorialsDGX agent

arXiv:2604.24662v1 Announce Type: cross Abstract: Identifying the dynamical state variables of a system from high-dimensional observations is a central problem across physical sciences. The challenge

LatentStealth: Unnoticeable and Efficient Adversarial Attacks on Expressive Human Pose and Shape Estimation

ApplicationsDGX agent

arXiv:2505.12009v2 Announce Type: replace Abstract: Expressive human pose and shape estimation (EHPS) plays a central role in digital human generation, particularly in live-streaming applications. How

Meta is about to release a pixel space model (Tuna-2)

Local AiDGX agent

Tuna-2 is a unified multimodal model that performs visual understanding and generation directly based on pixel embeddings, employing simple patch embedding layers to encode visual input without a VAE

MUSIC: Learning Muscle-Driven Dexterous Hand Control

ResearchDGX agent

arXiv:2604.23886v1 Announce Type: cross Abstract: We present a data-driven approach for physics-based, muscle-driven dexterous control that enables musculoskeletal hands to perform precise piano playi

Tuna-2: Pixel Embeddings Beat Vision Encoders for Multimodal Understanding and Generation

ResearchDGX agent

arXiv:2604.24763v1 Announce Type: new Abstract: Unified multimodal models typically rely on pretrained vision encoders and use separate visual representations for understanding and generation, creatin

27 Apr 2026

FeudalNav: A Simple Framework for Visual Navigation

ResearchDGX agent

arXiv:2602.06974v2 Announce Type: replace-cross Abstract: Visual navigation for robotics is inspired by the human ability to navigate environments using visual cues and memory, eliminating the need fo

24 Apr 2026

Flow Matching for Conditional MRI-CT and CBCT-CT Image Synthesis

Model ReleasesDGX agent

arXiv:2510.04823v2 Announce Type: replace Abstract: Generating synthetic CT (sCT) from MRI or CBCT plays a crucial role in enabling MRI-only and CBCT-based adaptive radiotherapy, improving treatment p

Frequency-Forcing: From Scaling-as-Time to Soft Frequency Guidance

ResearchDGX agent

arXiv:2604.20902v1 Announce Type: cross Abstract: While standard flow-matching models transport noise to data uniformly, incorporating an explicit generation order - specifically, establishing coarse,

JEPAMatch: Geometric Representation Shaping for Semi-Supervised Learning

TutorialsDGX agent

arXiv:2604.21046v1 Announce Type: new Abstract: Semi-supervised learning has emerged as a powerful paradigm for leveraging large amounts of unlabeled data to improve the performance of machine learnin

MISTY: High-Throughput Motion Planning via Mixer-based Single-step Drifting

Model ReleasesDGX agent

arXiv:2604.21489v1 Announce Type: cross Abstract: Multi-modal trajectory generation is essential for safe autonomous driving, yet existing diffusion-based planners suffer from high inference latency d

23 Apr 2026

Combo-Gait: Unified Transformer Framework for Multi-Modal Gait Recognition and Attribute Analysis

TutorialsDGX agent

arXiv:2510.10417v2 Announce Type: replace-cross Abstract: Gait recognition is an important biometric for human identification at a distance, particularly under low-resolution or unconstrained environm

22 Apr 2026

OLLM: Options-based Large Language Models

Model ReleasesDGX agent

arXiv:2604.19087v1 Announce Type: new Abstract: We introduce Options LLM (OLLM), a simple, general method that replaces the single next-token prediction of standard LLMs with a extit{set of learned op

21 Apr 2026

Does AI See like Art Historians? Interpreting How Vision Language Models Recognize Artistic Style

ResearchDGX agent

arXiv:2603.11024v2 Announce Type: replace Abstract: VLMs have become increasingly proficient at a range of computer vision tasks, such as visual question answering and object detection. This includes

20 Apr 2026

Hierarchical Active Inference using Successor Representations

TutorialsDGX agent

arXiv:2604.15679v1 Announce Type: cross Abstract: Active inference, a neurally-inspired model for inferring actions based on the free energy principle (FEP), has been proposed as a unifying framework

Limits of Lamarckian Evolution Under Pressure of Morphological Novelty

TutorialsDGX agent

arXiv:2604.15854v1 Announce Type: new Abstract: Lamarckian inheritance has been shown to be a powerful accelerator in systems where the joint evolution of robot morphologies and controllers is enhance

Similarity-Based Bike Station Expansion via Hybrid Denoising Autoencoders

ResearchDGX agent

arXiv:2604.15783v1 Announce Type: new Abstract: Urban bike-sharing systems require strategic station expansion to meet growing demand. Traditional allocation approaches rely on explicit demand modelli

17 Apr 2026

Compressing Sequences in the Latent Embedding Space: K-Token Merging for Large Language Models

ResearchDGX agent

arXiv:2604.15153v1 Announce Type: new Abstract: Large Language Models (LLMs) incur significant computational and memory costs when processing long prompts, as full self-attention scales quadratically

Differentiable Object Pose Connectivity Metrics for Regrasp Sequence Optimization

TutorialsDGX agent

arXiv:2604.14733v1 Announce Type: new Abstract: Regrasp planning is often required when one pick-and-place cannot transfer an object from an initial pose to a goal pose while maintaining grasp feasibi

How robots learn: A brief, contemporary history

TutorialsDGX agent

Roboticists used to dream big but build small. They’d hope to match or exceed the extraordinary complexity of the human body, and then they’d spend their career refining robotic arms for auto plants.

World-Value-Action Model: Implicit Planning for Vision-Language-Action Systems

TutorialsDGX agent

arXiv:2604.14732v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have emerged as a promising paradigm for building embodied agents that ground perception and language into action.

16 Apr 2026

ASTER: Latent Pseudo-Anomaly Generation for Unsupervised Time-Series Anomaly Detection

Model ReleasesDGX agent

arXiv:2604.13924v1 Announce Type: cross Abstract: Time-series anomaly detection (TSAD) is critical in domains such as industrial monitoring, healthcare, and cybersecurity, but it remains challenging d

Language steering in latent space to mitigate unintended code-switching

Model ReleasesDGX agent

arXiv:2510.13849v3 Announce Type: replace Abstract: Multilingual Large Language Models (LLMs) often exhibit hallucinations such as unintended code-switching, reducing reliability in downstream tasks.

15 Apr 2026

Learning Versatile Humanoid Manipulation with Touch Dreaming

SafetyDGX agent

arXiv:2604.13015v1 Announce Type: new Abstract: Humanoid robots promise general-purpose assistance, yet real-world humanoid loco-manipulation remains challenging because it requires whole-body stabili

Social Learning Strategies for Evolved Virtual Soft Robots

TutorialsDGX agent

arXiv:2604.12482v1 Announce Type: cross Abstract: Optimizing the body and brain of a robot is a coupled challenge: the morphology determines what control strategies are effective, while the control pa

14 Apr 2026

Continuous Adversarial Flow Models

TutorialsDGX agent

arXiv:2604.11521v1 Announce Type: cross Abstract: We propose continuous adversarial flow models, a type of continuous-time flow model trained with an adversarial objective. Unlike flow matching, which

Differentiable free energy surface: a variational approach to directly observing rare events using generative deep-learning models

SafetyDGX agent

arXiv:2604.09769v1 Announce Type: cross Abstract: Rare events are central to the evolution of complex many-body systems, characterized as key transitional configurations on the free energy surface (FE

Hide-and-Seek Attribution: Weakly Supervised Segmentation of Vertebral Metastases in CT

ResearchDGX agent

arXiv:2512.06849v2 Announce Type: replace Abstract: Accurate segmentation of vertebral metastasis in CT is clinically important yet difficult to scale, as voxel-level annotations are scarce and both l

Tipiano: Cascaded Piano Hand Motion Synthesis via Fingertip Priors

TutorialsDGX agent

arXiv:2604.09692v1 Announce Type: new Abstract: Synthesizing realistic piano hand motions requires both precision and naturalness. Physics-based methods achieve precision but produce stiff motions; da

13 Apr 2026

Envisioning the Future, One Step at a Time

Model ReleasesDGX agent

arXiv:2604.09527v1 Announce Type: cross Abstract: Accurately anticipating how complex, diverse scenes will evolve requires models that represent uncertainty, simulate along extended interaction chains

FlashLips: 100-FPS Mask-Free Latent Lip-Sync using Reconstruction Instead of Diffusion or GANs

HardwareDGX agent

arXiv:2512.20033v2 Announce Type: replace Abstract: We present FlashLips, a two-stage, mask-free lip-sync system that decouples lips control from rendering and achieves real-time performance, with our

10 Apr 2026

EchoTorrent: Towards Swift, Sustained, and Streaming Multi-Modal Video Generation

SafetyDGX agent

arXiv:2602.13669v4 Announce Type: replace Abstract: Recent multi-modal video generation models have achieved high visual quality, but their prohibitive latency and limited temporal stability hinder re

← Previous
1…345
Next →