From the 3 of 26 linked papers with an AI index.
1 citations · 1 across the 15 of their papers we have counts for
4 papers · 1 filter
TACT: Taxonomy-Aligned Post-Training for Pedagogically Adaptive English Tutoring
Dongjie Yang, Siyan Lin, Leixian Shen +3
Large language models (LLMs) are increasingly used to provide conversational practice for English-as-a-second-language (ESL) learners. Effective ESL tutoring, however, requires mor…
Fewer Clarifications, Better Code: Benchmarking Cross-Session Personalized Ambiguity Adaptation in Coding Assistants
Zijian Xu, Wenshuo Zhang, Zisen Qin +4
The paper defines personalized ambiguity adaptation for coding assistants, introduces the CAPA benchmark to evaluate how well models use a user's past resolved sessions to handle r…
Are LLMs Ready for Scientific Discovery? A Capability-Oriented Benchmark for AI Scientists
Chuhan Shi, Xiaoquan Ren, Sicheng Song +3
The paper presents SDABench, a capability-oriented benchmark that evaluates large language models on scientific data analysis tasks across biology, chemistry, environment, geograph…
TeachArena: Are Language Agents Ready for Realistic Teaching Work?
Zixin Chen, Peng Liu, Rui Sheng +6
Language agents are increasingly deployed in professional workflows, yet tutoring remains a high-stakes capability that existing evaluations only partially capture. Effective tutor…