works on

From the 1 of 5 linked papers with an AI index.

activity
20242026
collaborators

5 papers

cs.SD2026

Auditing Protocol-Level Shortcuts in Large Audio Language Model Judges for Speech Evaluation

Joonyong Park, David M. Chan, Yuki Saito +1

The paper investigates how large audio-language models used as automatic judges for speech evaluation can exploit protocol-level shortcuts—relying on provided labels or reference d…

cs.SD2026

Probing Token Spaces under Generator Shift in AI-Generated Music Detection

Joonyong Park, Jungwoo Kim, Junyoung Koh +1

AI-generated music detectors can appear robust on standard benchmark splits, yet their deployments require transfer to generator sources absent during training. We study this probl…

cs.SD2026

AnimeScore: A Preference-Based Dataset and Framework for Evaluating Anime-Like Speech Style

Joonyong Park, Jerry Li

Evaluating 'anime-like' voices currently relies on costly subjective judgments, yet no standardized objective metric exists. A key challenge is that anime-likeness, unlike naturaln…

cs.CL2025

Analysing the Language of Neural Audio Codecs

Joonyong Park, Shinnosuke Takamichi, David M. Chan +3

This study presents a comparative analysis of the statistical and linguistic properties of neural audio codecs (NACs). We investigate discrete speech tokens produced by various NAC…

cs.CV2024

Analyzing The Language of Visual Tokens

David M. Chan, Rodolfo Corona, Joonyong Park +3

With the introduction of transformer-based models for vision and language tasks, such as LLaVA and Chameleon, there has been renewed interest in the discrete tokenized representati…