Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Information Discernment in Large Language Models
Joshua Ashkinaze, Laura Kurek, Alina Faisal +4
LLMs are increasingly used with external knowledge sources like the internet. Do they weigh information appropriately -- updating more for reliable sources (source discernment) and…
cs.AI2026
Deep Value Benchmark: Measuring Whether Models Generalize Deep Values or Shallow Preferences
Joshua Ashkinaze, Hua Shen, Saipranav Avula +2
We introduce the Deep Value Benchmark (DVB), an evaluation framework that directly tests whether large language models (LLMs) learn fundamental human values or merely surface-level…