4 papers
ReplicatorBench: Benchmarking LLM Agents for Replicability in Social and Behavioral Sciences
Bang Nguyen, Dominik Soós, Qian Ma +8
The literature has witnessed an emerging interest in AI agents for automated assessment of scientific papers. Existing benchmarks focus primarily on the computational aspect of thi…
Re-defining Humor Data Objects for AI Humor Research
Anna Arnett, Bang Nguyen, Meng Jiang
In most existing AI humor research, humor was treated as either "present" or "not present." We explore the concept of humor as a social interaction with context and explanations. D…
QG-SMS: Enhancing Test Item Analysis via Student Modeling and Simulation
Bang Nguyen, Tingting Du, Mengxia Yu +2
While the Question Generation (QG) task has been increasingly adopted in educational assessments, its evaluation remains limited by approaches that lack a clear connection to the e…
Context Selection and Rewriting for Video-based Educational Question Generation
Mengxia Yu, Bang Nguyen, Olivia Zino +1
Educational question generation (EQG) is a crucial component of intelligent educational systems, significantly aiding self-assessment, active learning, and personalized education.…