Safety
Sorry, @peterwildeford, but this is wrong. Please don’t play along. The measurement “wall” you mention is hit ONLY if you don’t insist on re…
Sorry, @peterwildeford, but this is wrong. Please don’t play along. The measurement “wall” you mention is hit ONLY if you don’t insist on reliability. If you demanded 95% accuracy on the task, the sys
Sorry, @peterwildeford, but this is wrong. Please don’t play along. The measurement “wall” you mention is hit ONLY if you don’t insist on reliability. If you demanded 95% accuracy on the task, the systems wouldn’t be close to the measurement wall. The measurement problem you allude to is an artifact of artificially lowered expectations. Deep learning is hitting a wall (the wall being our ability to measure AI capabilities)
Source: Gary Marcus (X) | 2026-05-09