#ICML2026 If you are interested in value alignment, mechanistic interpretability, or pluralistic values, come see our poster @ JUL/8 (Wed)!
Time: 10:30
Location: Hall A, #3410
Title: Dual Mechanisms of Value Expression: Intrinsic vs. Prompted Values in Large Language Models
📢 New paper @ ACL 2025 Main!
TL;DR: What humans or LLMs think a text reflects ≠ how actual value-holders respond.
🔍 We propose a psychometrically grounded way to evaluate LLM values.
📄 https://t.co/o3cCI3sCLu
#ACL2025NLP#ValueAlignment#Values#LLM#Benchmark