which-values-do-llms-confuse-a-schwartz-based-recognition-study-0857e78f·1 events·first seen Aliases: Which Values Do LLMs Confuse? A Schwartz-Based Recognition Study
A new arXiv paper evaluates 21 instruction-tuned LLMs on their ability to identify which of Schwartz's ten basic human values is expressed in 1,000 Russian situational texts. Models achieve pooled Acc@1 of 0.683 and Acc@3 of 0.892, but exhibit systematic directional confusions — e.g., Universalism→Benevolence and Tradition→Conformity — that are consistent across checkpoints. The findings challenge the validity of value-alignment evaluations that assume models can reliably recognize values in context, and propose a richer evaluation protocol combining exact accuracy, ranked recovery, and directed error analysis.