On the Limits of LLM Adaptability: Impact of Model-Internalized Priors on Annotation Task Performance
DGX agentarXiv:2606.00467v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used for zero-shot annotation and LLM-as-a-judge tasks, yet their reliability hinges on how model-intern