SLAP: Stratified Loss-based Pruning for On-Policy Data-Efficient Instruction Tuning
DGX agentarXiv:2605.23969v1 Announce Type: new Abstract: Instruction tuning has optimized the specialized capabilities of large language models (LLMs), but it often requires extensive datasets and prolonged tr