5 papers
PERRY: Policy Evaluation with Confidence Intervals using Auxiliary Data
Aishwarya Mandyam, Jason Meng, Ge Gao +4
Off-policy evaluation (OPE) methods estimate the value of a new reinforcement learning (RL) policy prior to deployment. Recent advances have shown that leveraging auxiliary dataset…
Improving Hybrid Human-AI Tutoring by Differentiating Human Tutor Roles Based on Student Needs
Ashish Gurung, Ge Gao, Jordan Gutterman +6
Hybrid human-AI tutoring, where technology and humans jointly facilitate student learning, can be more beneficial than AI-only tutoring. However, preliminary evidence suggests that…
GIANTS: Generative Insight Anticipation from Scientific Literature
Joy He-Yueya, Anikait Singh, Ge Gao +5
Scientific breakthroughs often emerge from synthesizing prior ideas into novel contributions. While language models (LMs) show promise in scientific discovery, their ability to per…
Predicting Long Term Sequential Policy Value Using Softer Surrogates
Hyunji Nam, Allen Nie, Ge Gao +2
Off-policy policy evaluation (OPE) estimates the outcome of a new policy using historical data collected from a different policy. However, existing OPE methods cannot handle cases…
Predicting Long-Term Student Outcomes from Short-Term EdTech Log Data
Ge Gao, Amelia Leon, Andrea Jetten +4
Educational stakeholders are often particularly interested in sparse, delayed student outcomes, like end-of-year statewide exams. The rare occurrence of such assessments makes it h…