5 papers
TalkPlay-Tools: Conversational Music Recommendation with LLM Tool Calling
Seungheon Doh, Keunwoo Choi, Juhan Nam
While the recent developments in large language models (LLMs) have successfully enabled generative recommenders with natural language interactions, their recommendation behavior is…
TalkPlayData 2: An Agentic Synthetic Data Pipeline for Multimodal Conversational Music Recommendation
Keunwoo Choi, Seungheon Doh, Juhan Nam
We present TalkPlayData 2, a synthetic dataset for multimodal conversational music recommendation generated by an agentic data pipeline. In the proposed pipeline, multiple large la…
TALKPLAY: Multimodal Music Recommendation with Large Language Models
Seungheon Doh, Keunwoo Choi, Juhan Nam
We present TALKPLAY, a novel multimodal music recommendation system that reformulates recommendation as a token generation problem using large language models (LLMs). By leveraging…
KAD: No More FAD! An Effective and Efficient Evaluation Metric for Audio Generation
Yoonjin Chung, Pilsun Eu, Junwon Lee +3
Although being widely adopted for evaluating generated audio signals, the Fréchet Audio Distance (FAD) suffers from significant limitations, including reliance on Gaussian assumpt…
Music Discovery Dialogue Generation Using Human Intent Analysis and Large Language Models
SeungHeon Doh, Keunwoo Choi, Daeyong Kwon +2
A conversational music retrieval system can help users discover music that matches their preferences through dialogue. To achieve this, a conversational music retrieval system shou…