Benchmarking AI for low-resource contexts: Thinking beyond leaderboards
DGX agentarXiv:2605.28508v1 Announce Type: new Abstract: Existing AI evaluation practices often fail to capture how systems actually perform in low-resource environments, where operational constraints shape us