SafetyALFRED: Evaluating Safety-Conscious Planning of Multimodal Large Language Models
arXiv:2604.19638v1 Announce Type: new Abstract: Multimodal Large Language Models are increasingly adopted as autonomous agents in interactive environments, yet their ability to proactively address saf