TechniqueRLHF / Alignment8 recent entries12 Aug 2026MIDAS: Mutual Information Disentanglement with Uncertainty-Aware Fusion for Incomplete Multimodal Sentiment AnalysisarXiv:2608.09986v1 Announce Type: new Abstract: Most existing multimodal sentiment analysis approaches assume access to complete multimodal inputs. However, real-world applications frequently encounte→12 Aug 2026Measuring Semantic Abstractness of SAE Features via NonlocalityarXiv:2608.10537v1 Announce Type: new Abstract: Sparse autoencoders (SAEs) have helped uncover mechanistic explanations for LLM behaviours such as reasoning, jailbreaking etc., via understanding the c
TechniqueRAG8 recent entries12 Aug 2026When should you start post-training your own models? @FireworksAI_HQ CEO @lqiao’s answer: after product-market fit. Not because it's hard...…When should you start post-training your own models? @FireworksAI_HQ CEO @lqiao’s answer: after product-market fit. Not because it's hard... but because only after PMF is the data coming off your prod→12 Aug 2026SkillLens: Visual Skill Cards for Retrieval-Augmented GUI Action Prediction and On-Policy DistillationarXiv:2608.10775v1 Announce Type: new Abstract: Computer-using agents can perceive rich software interfaces, yet their decisions often lack visual procedural memory: they may recognize individual cont→12 Aug 2026Self-Knowledge Retrieval Augmented Generation Framework for Patent MatchingarXiv:2608.11030v1 Announce Type: cross Abstract: Patent retrieval and matching based on large language models (LLMs) play a vital role in intellectual property protection. However, due to the complex→12 Aug 2026Retrieval-Augmented Vision Foundation Models for Robust Leukemia Cell Classification across Multiple Microscopy DatasetsarXiv:2608.10657v1 Announce Type: cross Abstract: Leukemia cell image classification is challenged by real-world domain shifts from acquisition, staining, illumination, and site protocols, causing sin→12 Aug 2026REATS: LLM Reasoning-based Ensemble Learning for Adaptive Time Series ForecastingarXiv:2608.10149v1 Announce Type: new Abstract: Due to the diversity of real-world time series, no single forecasting model consistently dominates across all samples. Ensemble learning addresses this →12 Aug 2026Graphical Models of False Information and Fact Checking EcosystemsarXiv:2208.11582v2 Announce Type: replace-cross Abstract: The wide spread of false information online, including misinformation and disinformation, has become a major problem for our highly digitised →12 Aug 2026Covert Visual Prompt Injection against Commercial Multimodal Large Language ModelsarXiv:2603.29418v2 Announce Type: replace-cross Abstract: Although multimodal large language models (MLLMs) are increasingly deployed in real-world applications, their instruction-following behavior l→12 Aug 2026ConVAWG: A Retrieval-Grounded Framework for Controlled Synthetic Dialogue Generation in Violence Against Women and GirlsarXiv:2608.11200v1 Announce Type: cross Abstract: Synthetic dialogue generation offers a way to study conversational dynamics in sensitive domains where real data are difficult to access, release, or
TechniqueAgents8 recent entries12 Aug 2026Hand-Written PTX Tensor-Core GEMM Kernels: A Multi-Precision Study on NVIDIA L4arXiv:2608.10103v1 Announce Type: cross Abstract: High-performance Tensor Core kernels rely on a low-level PTX pipeline built from asynchronous data movement with cp.async, warp-level matrix loads wit→12 Aug 2026Coordinating the Unknown Lipschitz Constant in Multiplayer BanditsarXiv:2608.10526v1 Announce Type: cross Abstract: Motivated by decentralized applications, we study cooperative multi-agent bandits in continuous (Lipschitz) action spaces when the Lipschitz constant →12 Aug 2026CohereLabs/North-Micro-Vision-Instruct · Hugging FaceNorth Micro Vision Instruct is a 2.4B-parameter open-weight vision-language model with native-resolution image support, released under the Apache 2.0 license. It is designed as a compact foundation fo→12 Aug 2026Closed-Loop LLM Co-Pilots for Digital AgriculturearXiv:2608.09949v1 Announce Type: new Abstract: This study evaluates the application of Large Language Models (LLMs) in complex biological systems, evolving from data analysis to autonomous, AI-guided→12 Aug 2026Chain of Spatial Thoughts: Modality-Agnostic Spatial Grounding for Vision Language ModelsarXiv:2608.10278v1 Announce Type: new Abstract: Spatial understanding is fundamental to embodied intelligence, underpinning applications such as robotic manipulation, embodied navigation, and autonomo→12 Aug 2026Bandwidth-Efficient Multi-Agent Communication through Information Bottleneck and Vector QuantizationarXiv:2602.02035v2 Announce Type: replace-cross Abstract: Multi-agent reinforcement learning systems deployed in real-world robotics applications face severe communication constraints that significant→12 Aug 2026Agentic Instruction Data Selection: Let DataMaster Interpret Your IntentarXiv:2608.10579v1 Announce Type: new Abstract: Although existing instruction data selection methods have introduced various metrics, the inherent complexity of real-world datasets makes it impractica→12 Aug 2026Agentic AI infrastructure shifts enterprise focus from model choice to platform controlAs agentic AI infrastructure moves from experimentation into production, enterprises are confronting a more complex question than which model to use: how to control the cost, data exposure and infrast
TechniqueFine-tuning8 recent entries12 Aug 2026Pretrained Optimization Model for Zero-Shot Black Box OptimizationarXiv:2405.03728v3 Announce Type: replace-cross Abstract: Zero-shot optimization involves optimizing a target task that was not seen during training, aiming to provide the optimal solution without or →12 Aug 2026LLM Agents Factory: Retrieval of Domain-Specific LLM AgentsarXiv:2608.09934v1 Announce Type: cross Abstract: Large language model (LLM) agents improve task performance by decomposing problems into role-specialized behaviors. However, their practical deploymen→12 Aug 2026***Lights, inference, action…*** I’m so happy to share that @sequoia has led the Seed in @previewio. Developers are flying in magical AI-nat…***Lights, inference, action…*** I’m so happy to share that @sequoia has led the Seed in @previewio. Developers are flying in magical AI-native editors. But creative tooling is still stuck in the pre-→12 Aug 2026INSIDE the Student's Mind: Jointly Modeling Latent Reasoning and Action in LLM Student SimulatorsarXiv:2608.10492v1 Announce Type: new Abstract: Large Language Model (LLM)-based simulators often reproduce observable actions but fail to capture the underlying reasoning behind them. In education, w→12 Aug 2026CohereLabs/North-Micro-Vision-Instruct · Hugging FaceNorth Micro Vision Instruct is a 2.4B-parameter open-weight vision-language model with native-resolution image support, released under the Apache 2.0 license. It is designed as a compact foundation fo→12 Aug 2026CHORUS: Complementary Experts for High-Coverage Testbench Stimulus GenerationarXiv:2608.10090v1 Announce Type: new Abstract: Large language models (LLMs) have advanced code generation, where executable feedback provides a more reliable learning signal than textual imitation al→12 Aug 2026A Systematic Sample Size Analysis of ML-Based Path Loss Prediction for LPWANarXiv:2608.11083v1 Announce Type: cross Abstract: Low Power Wide Area Networks like LoRa are increasingly deployed for smart city applications, requiring accurate path loss prediction for effective ne→12 Aug 2026A Cost-Efficient Routing Pipeline for Multilingual Short-Text Classification Using Small Language ModelsarXiv:2608.10939v1 Announce Type: cross Abstract: Multilingual short-text classification supports operational systems such as content moderation, customer support routing, and intent recognition, yet
TechniqueMultimodal8 recent entries12 Aug 2026MIDAS: Mutual Information Disentanglement with Uncertainty-Aware Fusion for Incomplete Multimodal Sentiment AnalysisarXiv:2608.09986v1 Announce Type: new Abstract: Most existing multimodal sentiment analysis approaches assume access to complete multimodal inputs. However, real-world applications frequently encounte→12 Aug 2026LiquidAI/LFM2.5-VL-3B · Hugging FaceLFM2.5-VL-3B is a multimodal variant of LFM2.5, a family of hybrid models designed for on-device deployment. It builds on LFM2-VL-3B with further mid- and post-training. LFM2.5-VL-3B can process both →12 Aug 2026Dynamic Context Adapters: Efficiently Infusing History into Vision-and-Language ModelsarXiv:2608.10525v1 Announce Type: cross Abstract: Historical context integration presents a fundamental challenge for Vision-Language Models (VLMs) in sequential decision-making tasks. Current VLMs pr→12 Aug 2026DIMOS: Disentangling Instance-level Moving Object SegmentationarXiv:2606.12826v2 Announce Type: replace-cross Abstract: Moving instance segmentation (MIS) attracts increasing attention due to its broad applications in traffic surveillance, autonomous driving, an→12 Aug 2026Covert Visual Prompt Injection against Commercial Multimodal Large Language ModelsarXiv:2603.29418v2 Announce Type: replace-cross Abstract: Although multimodal large language models (MLLMs) are increasingly deployed in real-world applications, their instruction-following behavior l→12 Aug 2026CohereLabs/North-Micro-Vision-Instruct · Hugging FaceNorth Micro Vision Instruct is a 2.4B-parameter open-weight vision-language model with native-resolution image support, released under the Apache 2.0 license. It is designed as a compact foundation fo→12 Aug 2026Chain of Spatial Thoughts: Modality-Agnostic Spatial Grounding for Vision Language ModelsarXiv:2608.10278v1 Announce Type: new Abstract: Spatial understanding is fundamental to embodied intelligence, underpinning applications such as robotic manipulation, embodied navigation, and autonomo→12 Aug 2026Capturing Uncertainty in Human Motion for Representation Learning in SoccerarXiv:2608.11203v1 Announce Type: new Abstract: This paper presents a self-supervised representation learning framework for understanding 3D skeleton-based human motion in soccer, using future motion
TechniqueSafety8 recent entries12 Aug 2026Rethinking LLM Verification: Evidence Structure, Uncertainty, and Selective RefinementarXiv:2608.10725v1 Announce Type: new Abstract: Large language models (LLMs) often rely on shortcuts rather than systematic reasoning, raising safety concerns in medical applications. Allowing models →12 Aug 2026Physics-Informed Machine Learning in Prognostics and Health Management: A Systematic Literature ReviewarXiv:2608.10047v1 Announce Type: cross Abstract: In modern industry, keeping complex systems reliable, safe, and efficient hinges on Prognostics and Health Management (PHM). Machine Learning (ML) has→12 Aug 2026MIDAS: Mutual Information Disentanglement with Uncertainty-Aware Fusion for Incomplete Multimodal Sentiment AnalysisarXiv:2608.09986v1 Announce Type: new Abstract: Most existing multimodal sentiment analysis approaches assume access to complete multimodal inputs. However, real-world applications frequently encounte→12 Aug 2026INSIDE the Student's Mind: Jointly Modeling Latent Reasoning and Action in LLM Student SimulatorsarXiv:2608.10492v1 Announce Type: new Abstract: Large Language Model (LLM)-based simulators often reproduce observable actions but fail to capture the underlying reasoning behind them. In education, w→12 Aug 2026Expert-Guided g-computation with Large Language Models for Estimating Causal Effects on Timings: Applications to Hospital Quality ImprovementarXiv:2608.10339v1 Announce Type: cross Abstract: Hospital quality improvement (QI) programs routinely face multiple candidate interventions to optimize hospital flow, but existing methods struggle to→12 Aug 2026DIMOS: Disentangling Instance-level Moving Object SegmentationarXiv:2606.12826v2 Announce Type: replace-cross Abstract: Moving instance segmentation (MIS) attracts increasing attention due to its broad applications in traffic surveillance, autonomous driving, an→12 Aug 2026Automated Data Enrichment using Confidence-Aware Fine-Grained Debate among Open-Source LLMs for Mental Health and Online SafetyarXiv:2512.06227v3 Announce Type: replace Abstract: Real-world indicators play an important role in many Natural Language Processing (NLP) applications, such as life events for mental health analysis →12 Aug 2026A Convolutional Layer Activation Dimensionality Reduction for Out-of-Distribution and Adversarial Attack Detection MethodsarXiv:2608.10203v1 Announce Type: new Abstract: Despite the success of convolutional neural networks in image classification tasks and their general application in multi-modal models, their susceptibi