Cost of Structural Learning Under Censored Feedback: A Threshold-Bandit Approach
arXiv:2605.27076v1 Announce Type: cross Abstract: In many multi-agent applications, tasks yield rewards only when executed by a coalition meeting an unknown size threshold; otherwise, feedback is full