TechniqueRLHF / Alignment8 recent entries4 Aug 2026Unleashing the Power of Text: Text-Guided Flow Matching for Image Fusion under Complex DegradationsarXiv:2608.00530v1 Announce Type: new Abstract: Infrared-visible image fusion under realistic degradation scenarios is a challenging task, as degradations not only cause a loss of reliable modality-sp→5 Aug 2026LiLa-WAM: Lightweight Latent Reasoning World-Action Model for Robotic ManipulationarXiv:2608.03701v1 Announce Type: cross Abstract: World-action modeling has emerged as a promising paradigm for robotic control, as it empowers models to go beyond reacting to observations and anticip
TechniqueRAG5 recent entries13 May 2026[AINews] The End of FinetuningThis article likely discusses how advances in large language models, prompt engineering, and in-context learning are making traditional finetuning less necessary for many applications. It probably exp→19 May 2026CLAP: Contrastive Latent-space Prompt Optimization for End-to-end Autonomous DrivingarXiv:2605.17284v1 Announce Type: cross Abstract: End-to-end autonomous driving systems powered by Vision-Language-Action (VLA) models achieve strong performance on common driving scenarios, yet remai→1 Jun 2026Latent Geometric Chords for Query-Efficient Decision-Based Adversarial AttacksarXiv:2605.31219v1 Announce Type: new Abstract: While decision-based black-box adversarial attacks present a severe security threat, current methodologies suffer from fundamental limitations. Pixel-wi→1 Jun 2026ElasticMem: Latent Memory as a Learnable Resource for LLM AgentsarXiv:2605.30690v1 Announce Type: new Abstract: Long-term memory is essential for LLM agents to reason coherently across extended interactions, personalize responses, and reuse past experience. Howeve→10 Jun 2026One Token per Multimodal Evidence: Latent Memory for Resource-Constrained QAarXiv:2606.10572v1 Announce Type: new Abstract: External memory effectively grounds large language models (LLMs) and vision-language models (VLMs)-based question answering (QA) in relevant multimodal
TechniqueAgents8 recent entries8 Jul 2026Why AI Infrastructure must evolve for Agent Experience — Akshat Bubna, Modal CTOThis article discusses how AI infrastructure needs to adapt to support autonomous agents effectively, with insights from Modal's CTO on the technical and architectural requirements for agent deploymen→15 Jul 2026Mobility-Aware Cache Framework for Scalable LLM-Based Human Mobility SimulationarXiv:2602.16727v2 Announce Type: replace Abstract: Simulating large-scale human mobility is fundamental to understanding population movement patterns and supporting real-world geospatial applications→29 Jul 2026Understand Kimi K3 from first principles: a recommended order for anyone trying to understand this beastEveryone is talking about Kimi K3, but if you jump straight into the technical report, you’ll quickly realize it’s standing on years of research -- just like any breakthrough is! If you want to unders→31 Jul 2026Borrowed Strength: Best-of-N Search over a Code EncodingBreaks Self-Check Jailbreak DefensesarXiv:2607.26639v1 Announce Type: cross Abstract: A self-check defense asks the target model to assess a request before answering it; SAGE, the strongest published instance, reports an average 99% def→4 Aug 2026Unpacking ChatGPT Work: the Agent for a Billion UsersChatGPT Work was launched by OpenAI on July 9, 2026 as an agent‑oriented knowledge‑work platform that combines chat, Codex tools and cloud agents across fourteen model configurations. Within three wee→12 Aug 2026Sheaf-Based Federated Representation LearningarXiv:2608.10016v1 Announce Type: cross Abstract: Heterogeneous federated systems require agents to learn and exchange informative representations despite differences in data distributions, sensing mo→12 Aug 2026RLMOpt: Adaptive Prompt Optimization via Recursive Language ModelsarXiv:2608.10471v1 Announce Type: new Abstract: Prompt optimizers automate the search for prompts that improve language-model performance, but existing methods rely on a predefined optimization proced→12 Aug 2026Post-Hoc Sparse Coding of Latent Communication Between Vision-Language Model AgentsarXiv:2608.10198v1 Announce Type: new Abstract: Latent-space communication allows heterogeneous vision-language model agents to exchange continuous representations without serializing visual and reaso
TechniqueFine-tuning8 recent entries26 May 2026Lngram: N-gram Conditional Memory in Latent SpacearXiv:2605.24869v1 Announce Type: new Abstract: Sequence modeling requires both compositional reasoning and local static knowledge retrieval, yet standard Transformers handle both through dense comput→4 Jun 2026TamperBench: Systematically Stress-Testing LLM Safety Under Fine-Tuning and TamperingarXiv:2602.06911v2 Announce Type: replace-cross Abstract: As increasingly capable open-weight large language models (LLMs) are deployed, improving their tamper resistance against unsafe modifications,→23 Jun 2026A Stitch in Time Saves Nine: Preserving Policy Compatibility Under Perception Updates in End-to-End Autonomous DrivingarXiv:2606.21509v1 Announce Type: new Abstract: End-to-end autonomous driving systems tightly couple perception and decision-making through latent representations. Consequently, updates to perception →30 Jun 2026Few-Shot Domain Incremental Learning via Continual Vision-Language ConsolidationarXiv:2606.30190v1 Announce Type: cross Abstract: Existing domain-incremental learning (DIL) strategies call for massive amounts of data to adapt to new domains and suffer from the overfitting problem→23 Jul 2026Closing the Lab-to-Store Gap: A Data-Efficient Post-Training and Experience-Driven Learning VLA Framework for Retail HumanoidsarXiv:2607.20345v1 Announce Type: cross Abstract: Closing the gap between benchmark performance and reliable real-world operation remains a central challenge for Vision-Language-Action (VLA) humanoid →28 Jul 2026Latent-LoRA: Compact Latent-Space Adapters with Gradient-Free Routing for Continual LearningarXiv:2607.23837v1 Announce Type: cross Abstract: Large language models generalize well to individual tasks but lack an inherent mechanism for learning them sequentially, leading to catastrophic forge→10 Aug 2026LoRAScan: Detecting Backdoor Prompts in Low-Rank Adapters for Large Language Models via Down-Projection Activation SpikesarXiv:2608.06795v1 Announce Type: cross Abstract: Low-rank adaptation (LoRA) enables efficient specialization and distribution of large language models through compact adapters. However, untrusted ada→10 Aug 2026CellWorld: From Gene-Level Reconstruction to Latent Cell Prediction in Spatial Transcriptomics Foundation ModelsarXiv:2608.06659v1 Announce Type: new Abstract: This paper shows that latent-space predictive pretraining can provide a scalable route to foundation models for spatial transcriptomics. Existing spatia
TechniqueMultimodal8 recent entries29 Jul 2026Understand Kimi K3 from first principles: a recommended order for anyone trying to understand this beastEveryone is talking about Kimi K3, but if you jump straight into the technical report, you’ll quickly realize it’s standing on years of research -- just like any breakthrough is! If you want to unders→30 Jul 2026RLMM-Flow: A Flow-based Mobile Manipulation Framework with Latent-Space Reinforcement LearningarXiv:2607.26460v1 Announce Type: new Abstract: Mobile manipulation requires generating whole-body action chunks that jointly satisfy goal reaching, collision avoidance, base kinematic constraints, ma→4 Aug 2026Unleashing the Power of Text: Text-Guided Flow Matching for Image Fusion under Complex DegradationsarXiv:2608.00530v1 Announce Type: new Abstract: Infrared-visible image fusion under realistic degradation scenarios is a challenging task, as degradations not only cause a loss of reliable modality-sp→10 Aug 2026Representation-driven Endoscopic Visual Embedding Alignment for Latent GenerationarXiv:2608.07176v1 Announce Type: cross Abstract: Developing foundation generative models for endoscopy is limited by the gap between natural and clinical images and the computational cost of training→10 Aug 2026LoRAScan: Detecting Backdoor Prompts in Low-Rank Adapters for Large Language Models via Down-Projection Activation SpikesarXiv:2608.06795v1 Announce Type: cross Abstract: Low-rank adaptation (LoRA) enables efficient specialization and distribution of large language models through compact adapters. However, untrusted ada→12 Aug 2026Sheaf-Based Federated Representation LearningarXiv:2608.10016v1 Announce Type: cross Abstract: Heterogeneous federated systems require agents to learn and exchange informative representations despite differences in data distributions, sensing mo→12 Aug 2026Post-Hoc Sparse Coding of Latent Communication Between Vision-Language Model AgentsarXiv:2608.10198v1 Announce Type: new Abstract: Latent-space communication allows heterogeneous vision-language model agents to exchange continuous representations without serializing visual and reaso→12 Aug 2026FoR-SALE: Frame of Reference-guided Spatial Adjustment in LLM-based Diffusion EditingarXiv:2509.23452v2 Announce Type: replace-cross Abstract: Current text-to-image generation models, even state-of-the-art models, exhibit a significant performance gap when spatial expressions are desc
TechniqueSafety8 recent entries27 Jul 2026On the Identifiability of Controlled World ModelsarXiv:2607.22430v1 Announce Type: new Abstract: Learning world models that infer environment dynamics from high-dimensional observations and predict outcomes under candidate actions is central to plan→30 Jul 2026RLMM-Flow: A Flow-based Mobile Manipulation Framework with Latent-Space Reinforcement LearningarXiv:2607.26460v1 Announce Type: new Abstract: Mobile manipulation requires generating whole-body action chunks that jointly satisfy goal reaching, collision avoidance, base kinematic constraints, ma→31 Jul 2026Borrowed Strength: Best-of-N Search over a Code EncodingBreaks Self-Check Jailbreak DefensesarXiv:2607.26639v1 Announce Type: cross Abstract: A self-check defense asks the target model to assess a request before answering it; SAGE, the strongest published instance, reports an average 99% def→7 Aug 2026Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-TrainingarXiv:2608.06125v1 Announce Type: new Abstract: Latent reward models can supervise visual diffusion models without decoding intermediate states into pixel space. This makes alignment with human prefer→10 Aug 2026Representation-driven Endoscopic Visual Embedding Alignment for Latent GenerationarXiv:2608.07176v1 Announce Type: cross Abstract: Developing foundation generative models for endoscopy is limited by the gap between natural and clinical images and the computational cost of training→11 Aug 2026Do All LLMs Know When They're Being Harmful? A Reproducibility Study of Latent-Space Safety Probes Across Model FamiliesarXiv:2608.08029v1 Announce Type: cross Abstract: Khatri et al. (2026) [DOI: 10.1109/DSN-W70714.2026.00027] show that lightweight MLP probes on final-layer activations of a single 8B model (LLaMA-3.1-→12 Aug 2026MarkNull: Model-Agnostic Watermark Removal in AI-Generated Images via On-Manifold Latent ManipulationarXiv:2608.10166v1 Announce Type: cross Abstract: Digital watermarking has emerged as a critical technique for provenance and copyright attribution in AI-generated imagery, yet its robustness against →12 Aug 2026FoR-SALE: Frame of Reference-guided Spatial Adjustment in LLM-based Diffusion EditingarXiv:2509.23452v2 Announce Type: replace-cross Abstract: Current text-to-image generation models, even state-of-the-art models, exhibit a significant performance gap when spatial expressions are desc