9 papers · 1 filter
Below the Noise Floor: Bimodal Seed Collapse and Distinct Failure Modes in Small-Model Knowledge Distillation
Dipto Sumit, Sakib Ul Haque, Farig Sadeque
Function routing -- selecting the correct API call from a fixed catalog given a natural-language request -- is a deployment problem where small students are attractive but knowledg…
BengaliMCQ: Automatic Generation and Answer Prediction of Academic Multiple-Choice Questions in a Low-Resource Language
Abu Tarabin Surzo, A. K. M. Nihalul Kabir, Sm Azmain Faysal +3
Traditional retrieval-augmented generation (RAG) frameworks process documents without attending to their hierarchical structure, leading to poor performance, especially in low-reso…
Characterizing Human-Likeness in AI Generated Poetry: A Zero-shot Classification Study
A. N. Biswas, T. Tabassum, A. A. Shohid +4
With the advancement of AI technologies, Generative AI (GenAI) and human written text have become nearly indistinguishable. Additionally, the global standardization of AI chatbots…
When Does Knowledge Distillation Hurt? Reliability-Aware Distillation for Low-Resource Language Summarization
Dipto Sumit, Ankan Kumar Roy Srizon, Sadia Khair Rodela +4
Knowledge distillation (KD) is a standard approach for compressing sequence-to-sequence models, but its per-sample effects are rarely examined. On the BanSum Bangla summarization b…
Contaminated Collaboration: Measuring Gender Bias Transfer in LLM-Assisted Student Writing
Ariyan Hossain, Kazi Kamruzzaman Rabbi, Farig Sadeque +1
Gender bias in LLMs has been studied extensively in model outputs, with biased prompts shown to amplify stereotyped generations. Whether such bias propagates into text produced by…
Exploring the Limits of Pruning: Task-Specific Neurons, Model Collapse, and Recovery in Task-Specific Large Language Models
M. K. Khalidi Siam, Md. Tausif-Ul-Islam, Md. Reshad Romim Khan +5
Neuron pruning is widely used to reduce the computational cost and parameter footprint of large language models, yet it remains unclear whether neurons in task-specific models cont…