Deep Value Benchmark: Measuring Whether Models Generalize Deep Values or Shallow Preferences

NeurIPSSpotlight2025

Authors
Joshua Ashkinaze, Hua Shen, Sai Avula, Eric Gilbert, Ceren Budak
Venue
NeurIPS 2025
Track
Spotlight

TL;DR

We introduce the Deep Value Benchmark (DVB), an evaluation framework that directly tests whether large language models (LLMs) learn fundamental human values or merely surface-level preferences…

Opening excerpt from the authors’ abstract. source

Read the paper

Topics

large language model language model evaluation benchmark llm

← All NeurIPS 2025 Spotlight papers · Browse the whole archive