AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “latent-space”

GridTimelineEvolution
293 results
5 May 2026

Visual Latents Know More Than They Say: Unsilencing Latent Reasoning in MLLMs

Model ReleasesDGX agent

arXiv:2605.02735v1 Announce Type: new Abstract: Continuous latent-space reasoning offers a compact alternative to textual chain-of-thought for multimodal models, enabling high-dimensional visual evide

17 Apr 2026

Calibrate-Then-Delegate: Safety Monitoring with Risk and Budget Guarantees via Model Cascades

SafetyDGX agent

arXiv:2604.14251v1 Announce Type: new Abstract: Monitoring LLM safety at scale requires balancing cost and accuracy: a cheap latent-space probe can screen every input, but hard cases should be escalat

PixelDiT: Pixel Diffusion Transformers for Image Generation


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research
DGX agent

arXiv:2511.20645v2 Announce Type: replace Abstract: Latent-space modeling has been the standard for Diffusion Transformers (DiTs). However, it relies on a two-stage pipeline where the pretrained autoe

11 Aug 2026

MoNo: Multiscale Optimal Transport Neural Operator for Solving PDEs on General Geometries

ResearchDGX agent

arXiv:2608.09764v1 Announce Type: cross Abstract: Transformer-based neural operators have achieved substantial progress in solving Partial Differential Equations (PDEs) by projecting spatial observati

Biologically Informed Representation Learning for Robust Cross-Center Generalization of MALDI-TOF Mass Spectrometry

Model ReleasesDGX agent

arXiv:2608.08182v1 Announce Type: cross Abstract: Machine learning models for MALDI-TOF mass spectrometry have shown considerable promise for clinical microbiology tasks such as microbial identificati

6 Aug 2026

AI-based single-shot structured-light depth reconstruction for real-time laparoscopic surgical guidance

HardwareDGX agent

arXiv:2608.05109v1 Announce Type: cross Abstract: Significance. Accurate intraoperative depth perception is important for autonomous and semi-autonomous robotic laparoscopic surgery. Conventional frin

9 Jul 2026

Latent Policy Steering through One-Step Flow Policies

SafetyDGX agent

arXiv:2603.05296v2 Announce Type: replace-cross Abstract: Offline reinforcement learning (RL) allows robots to learn from offline datasets without risky exploration. Yet, offline RL's performance ofte

7 Jul 2026

Interpretable Human-Label-Free Deep Learning for Real-Bogus Classification with Uncertainty Quantification

Model ReleasesDGX agent

arXiv:2607.05393v1 Announce Type: cross Abstract: Time-domain surveys generate many transient candidates, making Real-Bogus classification a critical step in automated discovery pipelines. Reliable la

Physics-Informed Domain-Invariant Feature Learning with Autoencoder-Driven Gaussian Clustering for Robust Non-line-of-Sight Scenarios

Local AiDGX agent

arXiv:2607.02537v1 Announce Type: cross Abstract: Jamming and spoofing pose significant threats to wireless and satellite navigation by disrupting radio-frequency (RF) signals and compromising availab

PixWorld: Unifying 3D Scene Generation and Reconstruction in Pixel Space

ResearchDGX agent

arXiv:2607.05373v1 Announce Type: new Abstract: 3D reconstruction and generation are commonly tackled by separate paradigms: pixel-based regression for reconstruction, and latent diffusion for generat

9 Jun 2026

Latent Spatial Memory for Video World Models

ResearchDGX agent

arXiv:2606.09828v1 Announce Type: new Abstract: Video world models that maintain 3D spatial consistency across generated frames typically rely on explicit point cloud memory constructed in RGB space.

Beyond Linear Activation Steering: Invertible Latent Transformations for Controlling LLM Behavior

SafetyDGX agent

arXiv:2606.08454v1 Announce Type: new Abstract: Activation steering provides a lightweight inference-time mechanism for controlling large language models (LLMs) by modifying their internal activation

Graph Mamba Operator: A Latent Simulator for Interacting Particle Systems

ResearchDGX agent

arXiv:2606.09432v1 Announce Type: new Abstract: Modeling interacting dynamical systems requires capturing spatial interactions alongside long-range temporal dependencies. Graph neural networks (GNNs)

How Well Do Latent World Models Understand Partially Observable Safety Constraints?

SafetyDGX agent

arXiv:2510.06492v2 Announce Type: replace Abstract: Latent world models are a promising approach for learning state representations and dynamics directly from high-dimensional observations, enabling r

22 May 2026

LatentOmni: Rethinking Omni-Modal Understanding via Unified Audio-Visual Latent Reasoning

ResearchDGX agent

arXiv:2605.22012v1 Announce Type: new Abstract: Joint audio-visual reasoning is essential for omnimodal understanding, yet current multimodal large language models (MLLMs) still struggle when reasonin

14 May 2026

REALISTA: Realistic Latent Adversarial Attacks that Elicit LLM Hallucinations

ResearchDGX agent

arXiv:2605.12813v1 Announce Type: cross Abstract: Large language models (LLMs) achieve strong performance across many tasks but remain vulnerable to hallucinations, motivating the need for realistic a

21 Apr 2026

UniCSG: Unified High-Fidelity Content-Constrained Style-Driven Generation via Staged Semantic and Frequency Disentanglement

SafetyDGX agent

arXiv:2604.17850v1 Announce Type: new Abstract: Style transfer must match a target style while preserving content semantics. DiT-based diffusion models often suffer from content-style entanglement, le

12 Aug 2026

FoR-SALE: Frame of Reference-guided Spatial Adjustment in LLM-based Diffusion Editing

SafetyDGX agent

arXiv:2509.23452v2 Announce Type: replace-cross Abstract: Current text-to-image generation models, even state-of-the-art models, exhibit a significant performance gap when spatial expressions are desc

MarkNull: Model-Agnostic Watermark Removal in AI-Generated Images via On-Manifold Latent Manipulation

SafetyDGX agent

arXiv:2608.10166v1 Announce Type: cross Abstract: Digital watermarking has emerged as a critical technique for provenance and copyright attribution in AI-generated imagery, yet its robustness against

RLMOpt: Adaptive Prompt Optimization via Recursive Language Models

Model ReleasesDGX agent

arXiv:2608.10471v1 Announce Type: new Abstract: Prompt optimizers automate the search for prompts that improve language-model performance, but existing methods rely on a predefined optimization proced

Sheaf-Based Federated Representation Learning

Local AiDGX agent

arXiv:2608.10016v1 Announce Type: cross Abstract: Heterogeneous federated systems require agents to learn and exchange informative representations despite differences in data distributions, sensing mo

SynBoost: A Synergistic Framework for Fast Sampling of Diffusion Models

ResearchDGX agent

arXiv:2506.13058v2 Announce Type: replace-cross Abstract: Diffusion probabilistic models (DPMs) have demonstrated remarkable success in visual generation. However, their iterative sampling mechanism r

10 Aug 2026

Fluid-DiT: Graph-Free Diffusion Transformers for Fluid Flow Simulations Learning

ResearchDGX agent

arXiv:2608.07161v1 Announce Type: cross Abstract: Simulating complex fluid flows requires capturing full equilibrium distributions rather than just mean trajectories, yet high-fidelity solvers remain

LoRAScan: Detecting Backdoor Prompts in Low-Rank Adapters for Large Language Models via Down-Projection Activation Spikes

TutorialsDGX agent

arXiv:2608.06795v1 Announce Type: cross Abstract: Low-rank adaptation (LoRA) enables efficient specialization and distribution of large language models through compact adapters. However, untrusted ada

Representation-driven Endoscopic Visual Embedding Alignment for Latent Generation

SafetyDGX agent

arXiv:2608.07176v1 Announce Type: cross Abstract: Developing foundation generative models for endoscopy is limited by the gap between natural and clinical images and the computational cost of training

7 Aug 2026

Hierarchical Latent Prediction for Language Models

ResearchDGX agent

arXiv:2608.05806v1 Announce Type: cross Abstract: While standard Next-Token Prediction (NTP) lays the foundation of language model pre- training, its teacher-forced training paradigm may not be optima

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training

SafetyDGX agent

arXiv:2608.06125v1 Announce Type: new Abstract: Latent reward models can supervise visual diffusion models without decoding intermediate states into pixel space. This makes alignment with human prefer

5 Aug 2026

LiLa-WAM: Lightweight Latent Reasoning World-Action Model for Robotic Manipulation

HardwareDGX agent

arXiv:2608.03701v1 Announce Type: cross Abstract: World-action modeling has emerged as a promising paradigm for robotic control, as it empowers models to go beyond reacting to observations and anticip

SpreadMark: Robust Image Watermarking via Spread-Spectrum Embedding

ResearchDGX agent

arXiv:2608.03165v1 Announce Type: cross Abstract: Invisible image watermarks are increasingly used for deepfake detection and provenance tracking, where they must survive not only incidental distortio

4 Aug 2026

Unleashing the Power of Text: Text-Guided Flow Matching for Image Fusion under Complex Degradations

TutorialsDGX agent

arXiv:2608.00530v1 Announce Type: new Abstract: Infrared-visible image fusion under realistic degradation scenarios is a challenging task, as degradations not only cause a loss of reliable modality-sp

3 Aug 2026

DiffAttack: Evasion Attacks Against Face Recognition via Latent Diffusion Models

Model ReleasesDGX agent

arXiv:2607.28936v1 Announce Type: cross Abstract: Facial biometric identification relies on the distinctiveness of user attributes within a high-dimensional embedding space. However, the decision boun

WaiT for the Signal: Simple Frequency-Aware Flow-Matching

Local AiDGX agent

arXiv:2607.28760v1 Announce Type: cross Abstract: As image generation models scale to ever higher resolutions, global coherence, local detail, and texture fidelity become critical axes for generation

31 Jul 2026

Borrowed Strength: Best-of-N Search over a Code EncodingBreaks Self-Check Jailbreak Defenses

SafetyDGX agent

arXiv:2607.26639v1 Announce Type: cross Abstract: A self-check defense asks the target model to assess a request before answering it; SAGE, the strongest published instance, reports an average 99% def

30 Jul 2026

From Interface to Inference: Eliciting Any-Order Inference from Any-Order Models

Local AiDGX agent

arXiv:2607.26504v1 Announce Type: new Abstract: Many discrete reasoning tasks, such as code generation, are inherently non-causal: programmers move between high-level structure and local details, a pr

29 Jul 2026

Understand Kimi K3 from first principles: a recommended order for anyone trying to understand this beast

Local AiDGX agent

Everyone is talking about Kimi K3, but if you jump straight into the technical report, you’ll quickly realize it’s standing on years of research -- just like any breakthrough is! If you want to unders

28 Jul 2026

Latent Confounded Causal Discovery via Lie Bracket Geometry

ResearchDGX agent

arXiv:2606.19610v2 Announce Type: replace-cross Abstract: We study causal discovery from observational and interventional regimes when latent variables may affect the measured system. Our first algori

27 Jul 2026

On the Identifiability of Controlled World Models

SafetyDGX agent

arXiv:2607.22430v1 Announce Type: new Abstract: Learning world models that infer environment dynamics from high-dimensional observations and predict outcomes under candidate actions is central to plan

23 Jul 2026

Closing the Lab-to-Store Gap: A Data-Efficient Post-Training and Experience-Driven Learning VLA Framework for Retail Humanoids

Model ReleasesDGX agent

arXiv:2607.20345v1 Announce Type: cross Abstract: Closing the gap between benchmark performance and reliable real-world operation remains a central challenge for Vision-Language-Action (VLA) humanoid

Latent Riemannian Flow Matching for Geometry-Grounded 3D Foundation Models

ResearchDGX agent

arXiv:2607.19120v1 Announce Type: new Abstract: Geometric foundation models, such as the Visual Geometry Grounded Transformer (VGGT), provide strong 3D priors from unposed images. However, such models

15 Jul 2026

BattVAE-GP: Generative Modeling of Long-Horizon Battery Degradation with Uncertainty Quantification

ResearchDGX agent

arXiv:2607.11943v1 Announce Type: cross Abstract: Long-horizon physics-based simulations of battery degradation provide mechanistic insight but remain computationally expensive, limiting their use for

LatentFlow: A General Framework for Conditioning Stochastic Processes

ResearchDGX agent

arXiv:2607.12922v1 Announce Type: cross Abstract: Stochastic-process models are, as a rule, far easier to simulate than to condition. Non-linear observations, non-Gaussian likelihoods, black-box infor

10 Jul 2026

Latent Memory Palace: Reasoning for Control as Autoregressive Variational Inference

SafetyDGX agent

arXiv:2607.08724v1 Announce Type: new Abstract: Human decision-making is highly flexible -- some actions are taken immediately; others require longer deliberation. Language models have exhibited a sim

Omni-Sleep: A Sleep Foundation Model via Hierarchical Contrastive Learning of CNS--ANS Dynamic

ResearchDGX agent

arXiv:2607.07720v1 Announce Type: cross Abstract: Sleep physiology arises from the coordinated dynamics of the central nervous system (CNS) and autonomic nervous system (ANS), as reflected by multimod

8 Jul 2026

Physics-Informed Neural Embeddings of PDE Solution Families

ResearchDGX agent

arXiv:2607.06348v1 Announce Type: new Abstract: We introduce a physics-informed framework for learning finite-dimensional embeddings of solution families of partial differential equations. The method

3 Jul 2026

VLA-Corrector: Lightweight Detect-and-Correct Inference for Adaptive Action Horizon

Local AiDGX agent

arXiv:2607.01804v1 Announce Type: new Abstract: Vision-Language-Action (VLA) foundation models have recently achieved strong progress in embodied intelligence. To reduce policy-call frequency while pr

2 Jul 2026

Multimodal Continuous Reasoning via Asymmetric Mutual Variational Learning

Model ReleasesDGX agent

arXiv:2607.00461v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) are often constrained by a language-space bottleneck, forcing complex visual reasoning into discrete tokens whi

1 Jul 2026

Cross-Space Distillation: Teaching One-Step Students with Modern Diffusion Teachers

SafetyDGX agent

arXiv:2606.32020v1 Announce Type: new Abstract: Modern one-step diffusion models achieve impressive quality through distribution-based timestep distillation. Yet, they rely on a critical assumption: T

30 Jun 2026

Few-Shot Domain Incremental Learning via Continual Vision-Language Consolidation

Model ReleasesDGX agent

arXiv:2606.30190v1 Announce Type: cross Abstract: Existing domain-incremental learning (DIL) strategies call for massive amounts of data to adapt to new domains and suffer from the overfitting problem

Granular-ball computing: an efficient, robust, and interpretable adaptive multi-granularity representation and computation method

ResearchDGX agent

arXiv:2304.11171v5 Announce Type: replace-cross Abstract: To overcome the limitations of point-based inputs, overly fine computation and limited adaptability in existing artificial intelligence method

TacGen: Touch Is a Necessary Dimension of Physical-World Representation -- Addressing Tactile Data Scarcity with Scalable Vision-to-Touch Alignment and Generation

SafetyDGX agent

arXiv:2606.29173v1 Announce Type: new Abstract: Touch resolves the physical-property ambiguity left by vision: exploratory contact recovers shape, texture, compliance, and material, and visuo-haptic o

29 Jun 2026

A Multi-Attribute Latent Space for Visual Analysis of Watches

Model ReleasesDGX agent

arXiv:2606.27897v1 Announce Type: new Abstract: We present a design rationale, embedding model, and interactive visual-analysis system for exploring large wristwatch collections through heterogeneous

From Tokens to States: LLMs as a Special Case of World Models and the Continuous Path Beyond

ResearchDGX agent

arXiv:2606.28127v1 Announce Type: cross Abstract: The AI community has framed the relationship between large language models (LLMs) and world models as a dichotomy: LLMs predict tokens; world models s

25 Jun 2026

Latent Block-Diffusion Temporal Point Processes: A Semi-Autoregressive Framework for Asynchronous Event Sequence Generation

Model ReleasesDGX agent

arXiv:2606.24982v1 Announce Type: new Abstract: Modeling and sampling from the underlying distribution of asynchronous event sequences are crucial in various real-world applications, including social

23 Jun 2026

A Stitch in Time Saves Nine: Preserving Policy Compatibility Under Perception Updates in End-to-End Autonomous Driving

SafetyDGX agent

arXiv:2606.21509v1 Announce Type: new Abstract: End-to-end autonomous driving systems tightly couple perception and decision-making through latent representations. Consequently, updates to perception

AdaReP:Adaptive Re-Planning under Model Mismatch for Neural World-Model Predictive Control

Local AiDGX agent

arXiv:2606.23079v1 Announce Type: new Abstract: Neural world models coupled with model predictive control (MPC) replan at every environment step to bound accumulated prediction error, but this incurs

Understanding Latent Flow Models for Tabular Data Synthesis: Targets, Paths, and Sampling

ResearchDGX agent

arXiv:2606.20878v1 Announce Type: new Abstract: Synthetic tabular data enables microdata sharing in regulated domains, yet deploying continuous-time generative models requires balancing analytical uti

11 Jun 2026

Continual Learning with Support Boundary Experience Blending

ResearchDGX agent

arXiv:2507.23534v3 Announce Type: replace-cross Abstract: Continual learning (CL) seeks to mitigate catastrophic forgetting when models are trained with sequential tasks. A common approach, experience

The Latent Color Subspace: Emergent Order in High-Dimensional Chaos

ResearchDGX agent

arXiv:2603.12261v2 Announce Type: replace-cross Abstract: Text-to-image generation models have advanced rapidly, yet achieving fine-grained control over generated images remains difficult, largely due

10 Jun 2026

One Token per Multimodal Evidence: Latent Memory for Resource-Constrained QA

ResearchDGX agent

arXiv:2606.10572v1 Announce Type: new Abstract: External memory effectively grounds large language models (LLMs) and vision-language models (VLMs)-based question answering (QA) in relevant multimodal

8 Jun 2026

Consistency-Preserving Diverse Video Generation

ResearchDGX agent

arXiv:2602.15287v2 Announce Type: replace Abstract: Text-to-video generation is expensive, so only a few samples are typically produced per prompt. In this low-sample regime, maximizing the value of e

← Previous
12345
Next →