Safety
When Missing Becomes Structure: Intent-Preserving Policy Completion from Financial KOL Discourse
arXiv:2604.14333v1 Announce Type: new Abstract: Key Opinion Leader (KOL) discourse on social media is widely consumed as investment guidance, yet turning it into executable trading strategies without
arXiv:2604.14333v1 Announce Type: new Abstract: Key Opinion Leader (KOL) discourse on social media is widely consumed as investment guidance, yet turning it into executable trading strategies without injecting assumptions about unspecified execution decisions remains an open problem. We observe that the gaps in KOL statements are not random deficiencies but a structured separation: KOLs express directional intent (what to buy or sell and why) while leaving execution decisions (when, how much, how long) systematically unspecified. Building on this observation, we propose an intent-preserving policy completion framework that treats KOL discourse as a partial trading policy and uses offline reinforcement learning to complete the missing execution decisions around the KOL-expressed intent. Experiments on multimodal KOL discourse from YouTube and X (2022-2025) show that KICL achieves the best return and Sharpe ratio on both platforms while maintaining zero unsupported entries and zero directional reversals, and ablations confirm that the full framework yields an 18.9% return improvement over the KOL-aligned baseline.
Related
- Policy-Aware Design of Large-Scale Factorial Experiments
- Optimistic Policy Learning under Pessimistic Adversaries with Regret and Violation Guarantees
- PROXIMA: A Reliability Scoring Framework for Proxy Metrics in Online Controlled Experiments
- Beyond Importance Sampling: Rejection-Gated Policy Optimization
Source: arXiv cs.LG | 2026-04-17