From the 1 of 7 linked papers with an AI index.
7 papers
ChartGenEval: Corruption-Tested Multi-Dimensional Feedback for Rhythm-Game Chart Generation
Jhen-Ke Lin
The paper introduces ChartGenEval, an evaluation framework for rhythm-game chart generation that uses timing maps and controlled corruption tests to provide multi-dimensional feedb…
The NTNU System at the S&I Challenge 2025 SLA Open Track
Hong-Yun Lin, Tien-Hong Lo, Yu-Hsuan Fang +4
A recent line of research on spoken language assessment (SLA) employs neural models such as BERT and wav2vec 2.0 (W2V) to evaluate speaking proficiency across linguistic and acoust…
BUILD-AND-FIND: An Effort-Aware Protocol for Evaluating Agent-Managed Codebases
Jhen-Ke Lin
Most coding-agent benchmarks ask whether generated code behaves correctly. That remains essential, but repository-level engineering is increasingly agent-managed: one agent writes…
Session-Level Spoken Language Assessment with a Multimodal Foundation Model via Multi-Target Learning
Hong-Yun Lin, Jhen-Ke Lin, Chung-Chun Wang +2
Spoken Language Assessment (SLA) estimates a learner's oral proficiency from spontaneous speech. The growing population of L2 English speakers has intensified the demand for reliab…
A Novel Data Augmentation Approach for Automatic Speaking Assessment on Opinion Expressions
Chung-Chun Wang, Jhen-Ke Lin, Hao-Chien Lu +2
Automated speaking assessment (ASA) on opinion expressions is often hampered by the scarcity of labeled recordings, which restricts prompt diversity and undermines scoring reliabil…
Acoustically Precise Hesitation Tagging Is Essential for End-to-End Verbatim Transcription Systems
Jhen-Ke Lin, Hao-Chien Lu, Chung-Chun Wang +2
Verbatim transcription for automatic speaking assessment demands accurate capture of disfluencies, crucial for downstream tasks like error analysis and feedback. However, many ASR…