From Training to Deployment: Post-Hoc Causal Feature Identification via Sensitivity Ratios
DGX agentarXiv:2607.25546v1 Announce Type: new Abstract: Given a model that is already trained, which features does it rely on causally versus spuriously? Existing methods require access to the training proced