Deep Value Benchmark: Measuring Whether Models Generalize Deep Values or Shallow Preferences
NeurIPSSpotlight2025
TL;DR
We introduce the Deep Value Benchmark (DVB), an evaluation framework that directly tests whether large language models (LLMs) learn fundamental human values or merely surface-level preferences…
Opening excerpt from the authors’ abstract. source
Read the paper
Topics
large language model language model evaluation benchmark llm
← All NeurIPS 2025 Spotlight papers · Browse the whole archive