activity
20242026
collaborators

5 papers

cs.AI2026

ReplicatorBench: Benchmarking LLM Agents for Replicability in Social and Behavioral Sciences

Bang Nguyen, Dominik Soós, Qian Ma +8

The literature has witnessed an emerging interest in AI agents for automated assessment of scientific papers. Existing benchmarks focus primarily on the computational aspect of thi…

cs.CL2026

Re-defining Humor Data Objects for AI Humor Research

Anna Arnett, Bang Nguyen, Meng Jiang

In most existing AI humor research, humor was treated as either "present" or "not present." We explore the concept of humor as a social interaction with context and explanations. D…

cs.CL2025

QG-SMS: Enhancing Test Item Analysis via Student Modeling and Simulation

Bang Nguyen, Tingting Du, Mengxia Yu +2

While the Question Generation (QG) task has been increasingly adopted in educational assessments, its evaluation remains limited by approaches that lack a clear connection to the e…

cs.CL2025

Context Selection and Rewriting for Video-based Educational Question Generation

Mengxia Yu, Bang Nguyen, Olivia Zino +1

Educational question generation (EQG) is a crucial component of intelligent educational systems, significantly aiding self-assessment, active learning, and personalized education.…

cs.CL2024

Reference-based Metrics Disprove Themselves in Question Generation

Bang Nguyen, Mengxia Yu, Yun Huang +1

Reference-based metrics such as BLEU and BERTScore are widely used to evaluate question generation (QG). In this study, on QG benchmarks such as SQuAD and HotpotQA, we find that us…