UniClawBench: A Universal Benchmark for Proactive Agents on Real-World Tasks
arXiv:2607.08768v1 Announce Type: new Abstract: The rapid development of large language models and multimodal large language models has accelerated the emergence of proactive agents capable of operati