1 citations · 1 across the 4 of their papers we have counts for
4 papers
Bhaasha, Bhasa, Zaban: A Survey for Low-Resourced Languages in South Asia -- Current Stage and Challenges
Sampoorna Poria, Xiaolei Huang
Rapid developments of large language models have revolutionized many NLP tasks for English data. Unfortunately, the models and their evaluations for low-resource languages are bein…
Attributes as Textual Genes: Leveraging LLMs as Genetic Algorithm Simulators for Conditional Synthetic Data Generation
Guangzeng Han, Weisi Liu, Xiaolei Huang
Large Language Models (LLMs) excel at generating synthetic data, but ensuring its quality and diversity remains challenging. We propose Genetic Prompt, a novel framework that combi…
Examining and Adapting Time for Multilingual Classification via Mixture of Temporal Experts
Weisi Liu, Guangzeng Han, Xiaolei Huang
Time is implicitly embedded in classification process: classifiers are usually built on existing data while to be applied on future data whose distributions (e.g., label and token)…
Examining Imbalance Effects on Performance and Demographic Fairness of Clinical Language Models
Precious Jones, Weisi Liu, I-Chan Huang +1
Data imbalance is a fundamental challenge in applying language models to biomedical applications, particularly in ICD code prediction tasks where label and demographic distribution…