2 papers
cs.CL2025
PokerBench: Training Large Language Models to become Professional Poker Players
Richard Zhuang, Akshat Gupta, Richard Yang +3
We introduce PokerBench - a benchmark for evaluating the poker-playing abilities of large language models (LLMs). As LLMs excel in traditional NLP tasks, their application to compl…
cs.DB2024
GeoLife+: Large-Scale Simulated Trajectory Datasets Calibrated to the GeoLife Dataset
Hossein Amiri, Richard Yang, Andreas Zufle
Analyzing individual human trajectory data helps our understanding of human mobility and finds many commercial and academic applications. There are two main approaches to accessing…