3 papers
stat.ML2026
MultiwayPAM: Multiway Partitioning Around Medoids for LLM-as-a-Judge Score Analysis
Chihiro Watanabe, Jingyu Sun
LLM-as-a-Judge is a flexible framework for text evaluation, which allows us to obtain scores for the quality of a given text from various perspectives by changing the prompt templa…
stat.ML2026
Empirical Cumulative Distribution Function Clustering for LLM-based Agent System Analysis
Chihiro Watanabe, Jingyu Sun
Large language models (LLMs) are increasingly used as agents to solve complex tasks such as question answering (QA), scientific debate, and software development. A standard evaluat…
eess.AS2024
GE2E-AC: Generalized End-to-End Loss Training for Accent Classification
Chihiro Watanabe, Hirokazu Kameoka
Accent classification or AC is a task to predict the accent type of an input utterance, and it can be used as a preliminary step toward accented speech recognition and accent conve…