VALUEFLOW: Toward Pluralistic and Steerable Value-based Alignment in Large Language Models
arXiv:2602.03160v2 Announce Type: replace Abstract: Aligning Large Language Models (LLMs) with the diverse spectrum of human values remains a central challenge: preference-based methods often fail to