1 citations · 1 across the 3 of their papers we have counts for
3 papers
STRIVE: Probing Reasoning Limits in Graded Plausibility Generation and Evaluation
Bhiman Kumar Baghel, Anna Chrabaszcz, Tessa Warren +3
Event knowledge concerns who does what to whom. Psycholinguists use event-plausibility judgments to examine how this knowledge supports human language processing. To isolate plausi…
MindTrellis: Co-Creating Knowledge Structures with AI through Interactive Visual Exploration
Xiang Li, Cara Li, Emily Kuang +2
Knowledge workers face increasing challenges in synthesizing information from multiple documents into structured conceptual understanding. This process is inherently iterative: use…
Every Answer Matters: Evaluating Commonsense with Probabilistic Measures
Qi Cheng, Michael Boratko, Pranay Kumar Yelugam +4
Large language models have demonstrated impressive performance on commonsense tasks; however, these tasks are often posed as multiple-choice questions, allowing models to exploit s…