From Preimage Search To Source-Grounded Feature Inversion
arXiv:2607.12526v1 Announce Type: new Abstract: Interpreting a neural network requires understanding what its internal features extract from a particular input. Feature inversion seeks to express a se