One of the fastest ways to lose trust in a self-hosted LLM: prompt injection compliance [P]
This r/MachineLearning post discusses how self-hosted LLMs that comply with prompt injection attempts — effectively following malicious or overriding instructions embedded in user input — represent a