CAREBench: A Child-Safety Risk Benchmark for Language Models
DGX agentarXiv:2606.29685v1 Announce Type: new Abstract: How can we evaluate whether frontier AI systems recognize child-safety risks before they escalate into explicit harm? Existing child safety evaluations