1 paper
Joshua Ashkinaze, Hua Shen, Saipranav Avula +2
We introduce the Deep Value Benchmark (DVB), an evaluation framework that directly tests whether large language models (LLMs) learn fundamental human values or merely surface-level…