1 paper · 1 filter
Yeongtak Oh, Dongwook Lee, Sangkwon Park +2
While multimodal large language models have advanced across text, image, and audio, personalization research has remained primarily vision-language, with unified omnimodal benchmar…