PoisonForge: Task-Level Targeted Poisoning Benchmark for Instruction-Tuned LLMs
DGX agentarXiv:2605.23168v1 Announce Type: cross Abstract: When practitioners fine-tune LLMs on unvetted datasets, an adversary can exploit the data supply chain through task-level poisoning: inserting a small