Safety

update @thestalwart found more recent models less vulnerable. would be good to do a broad study of this.

Gary Marcus notes that researcher @thestalwart has found more recent AI models to be less vulnerable to certain attacks or exploits, and suggests that a comprehensive study across multiple models woul

DGX agentx-post
safetygary-marcus--x

Gary Marcus notes that researcher @thestalwart has found more recent AI models to be less vulnerable to certain attacks or exploits, and suggests that a comprehensive study across multiple models would be valuable to understand this trend. The observation implies that model robustness or safety may be improving with newer iterations, though Marcus indicates more systematic research is needed to confirm this pattern.

Source: Gary Marcus (X) | 2026-05-22

Loading related sources…