From the 1 of 13 linked papers with an AI index.
13 papers
TreeProbe : A Tibetan Medicine Benchmark for Cultural Bias in LLMs
Jin Zhang, Linyu Li, Weili Jiang +7
Large language models are increasingly viewed as a potential means of mitigating global health inequities, yet their outputs often reflect dominant high-resource medical traditions…
Knowledge before Reasoning: EC-Reason-Bench, a Training-Free Diagnostic Benchmark for LLM Enzyme Classification
Linyu Li, Zhi Jin, Yichi Zhang +6
The paper introduces EC-Reason-Bench, a training-free diagnostic benchmark that evaluates why general large language models struggle with detailed enzyme classification and how per…
SkillSight: Calibrating Generic Content Bias for Skill Retrieval
Jinying Xiao, Bin Li, Bin Ji +9
As large language model agents gain access to increasingly large skill libraries, retrieving the right skill becomes critical to reliable capability selection and execution. Existi…
TFD: A Comprehensive Structured Tibetan Foundation Dataset for Low-Resource Language Processing and Large-Scale Modeling
Cheng Huang, Fan Gao, Nyima Tashi +5
Large Language Models (LLMs) have achieved remarkable success in high-resource languages, yet progress in Tibetan remains severely constrained. While recent efforts have begun to a…
FMSD-TTS: Few-shot Multi-Speaker Multi-Dialect Text-to-Speech Synthesis for Ã-Tsang, Amdo and Kham Speech Dataset Generation
Yutong Liu, Ziyue Zhang, Ban Ma-bao +7
Tibetan is a low-resource language with minimal parallel speech corpora spanning its three major dialects-Ã-Tsang, Amdo, and Kham-limiting progress in speech modeling. To address…
POTSA: A Cross-Lingual Speech Alignment Framework for Speech-to-Text Translation
Xuanchen Li, Chenrui Cui, Tianrui Wang +9
Speech Large Language Models have achieved breakthroughs in multilingual speech-to-text translation. However, existing approaches often overlook semantic commonalities across sourc…