6 papers
PMIScore: An Unsupervised Approach to Quantify Dialogue Engagement
Yongkang Guo, Zhihuan Huang, Yuqing Kong
High dialogue engagement is a crucial indicator of an effective conversation. A reliable measure of engagement could help benchmark large language models, enhance the effectiveness…
Jailbreaking LLMs via Calibration
Yuxuan Lu, Yongkang Guo, Yuqing Kong
Safety alignment in Large Language Models (LLMs) often creates a systematic discrepancy between a model's aligned output and the underlying pre-aligned data distribution. We propos…
Mitigating the Participation Bias by Balancing Extreme Ratings
Yongkang Guo, Yuqing Kong, Jialiang Liu
Rating aggregation plays a crucial role in various fields, such as product recommendations, hotel rankings, and teaching evaluations. However, traditional averaging methods can be…
Robust Decision Aggregation with Adversarial Experts
Yongkang Guo, Yuqing Kong
We consider a robust aggregation problem in the presence of both truthful and adversarial experts. The truthful experts will report their private signals truthfully, while the adve…
How Gold to Make the Golden Snitch: Designing the "Game Changer" in Esports
Zhihuan Huang, Yuxuan Lu, Yongkang Guo +1
Many battling games utilize a special item (e.g. Roshan in Defense of the Ancients 2 (DOTA 2), Baron Nashor in League of Legends (LOL), Golden Snitch in Quidditch) as a potential `…
Algorithmic Robust Forecast Aggregation
Yongkang Guo, Jason D. Hartline, Zhihuan Huang +3
Forecast aggregation combines the predictions of multiple forecasters to improve accuracy. However, the lack of knowledge about forecasters' information structure hinders optimal a…