On the Generalization of Steering Vectors for Chain-of-Thought Faithfulness
DGX agentarXiv:2607.29062v1 Announce Type: new Abstract: Model capabilities have improved in large part due to scaling chain of thought. This has been a promising development for AI safety--where models verbal