2 papers
cs.CL2026
GRADE: Generalizable Reasoning-Aware Dialogue Evaluation for AI Tutors
Parth Bhalerao, Jeromy Chang, David Chou +1
Evaluating AI tutor responses requires more than factual correctness: tutors must identify mistakes, locate errors, provide guidance, and offer actionable next steps. We present GR…
cs.CV2026
When Cultures Move: Measuring and Improving Multicultural Text-to-Video Generation
Shuowei Li, Yuming Zhao, Parth Bhalerao +1
Text-to-video (T2V) generation has rapidly progressed in visual fidelity, yet its ability to faithfully represent multiple cultures within a single prompt remains underexplored. We…