1 citations · 1 across the 11 of their papers we have counts for
6 papers · 1 filter
Scoring Rules! Statistical and Strategic Alignment for Text Evaluation Metrics
Shengwei Xu, Yuxuan Lu, Yifan Wu +2
Reference-based text evaluation metrics, which are widely used to assess natural language generation systems, score a candidate response by comparing it with a reference response.…
ComplLLM: Fine-tuning LLMs to Discover Complementary Signals for Decision-making
Ziyang Guo, Yifan Wu, Jason Hartline +2
Multi-agent decision pipelines can outperform single agent workflows when complementarity holds, i.e., different agents bring unique information to the table to inform a final deci…
Aligned Textual Scoring Rules
Yuxuan Lu, Yifan Wu, Jason Hartline +1
Scoring rules elicit probabilistic predictions from a strategic agent by scoring the prediction against a ground truth state. A scoring rule is proper if, from the agent's perspect…
Explaining and Improving Information Complementarities in Multi-Agent Decision-making
Ziyang Guo, Yifan Wu, Jason Hartline +1
Multiple agents are increasingly combined to make decisions with the expectation of achieving complementary performance, where the decisions they make together outperform those mad…
ElicitationGPT: Text Elicitation Mechanisms via Language Models
Yifan Wu, Jason Hartline
Scoring rules evaluate probabilistic forecasts of an unknown state against the realized state and are a fundamental building block in the incentivized elicitation of information. T…
A Decision Theoretic Framework for Measuring AI Reliance
Ziyang Guo, Yifan Wu, Jason Hartline +1
Humans frequently make decisions with the aid of artificially intelligent (AI) systems. A common pattern is for the AI to recommend an action to the human who retains control over…