GRZO: Group-Relative Zeroth-Order Optimization for Large Language Model Fine-Tuning
DGX agentarXiv:2606.02857v1 Announce Type: cross Abstract: Zeroth-order (ZO) optimization is a memory-efficient alternative to backpropagation for fine-tuning large language models, but its deployment is limit