4 papers
Rethinking Perplexity: Revealing the Impact of Input Length on Perplexity Evaluation in LLMs
Letian Cheng, Junyan Wang, Yan Gao +3
Perplexity is a widely adopted metric for assessing the predictive quality of large language models (LLMs) and often serves as a reference metric for downstream evaluations. Howeve…
Decoding Ambiguous Emotions with Test-Time Scaling in Audio-Language Models
Hong Jia, Weibin Li, Jingyao Wu +6
Emotion recognition from human speech is a critical enabler for socially aware conversational AI. However, while most prior work frames emotion recognition as a categorical classif…
FlowerTune: A Cross-Domain Benchmark for Federated Fine-Tuning of Large Language Models
Yan Gao, Massimo Roberto Scamarcia, Javier Fernandez-Marques +18
Large Language Models (LLMs) have achieved state-of-the-art results across diverse domains, yet their development remains reliant on vast amounts of publicly available data, raisin…
Scaling Auditory Cognition via Test-Time Compute in Audio Language Models
Ting Dang, Yan Gao, Hong Jia
Large language models (LLMs) have shown exceptional versatility in natural language processing, prompting recent efforts to extend their multimodal capabilities to speech processin…