3 papers
cs.CL2026
Index SLM Technical Report
Tianjiao Li, Lusheng Zhang, Shien He +8
We present Index-1.9B, a series of open small language models developed at Bilibili. The series comprises four models: Index-1.9B-Base, a foundation model with 1.9 billion non-embe…
cs.CL2025
SABER: Switchable and Balanced Training for Efficient LLM Reasoning
Kai Zhao, Yanjun Zhao, Jiaming Song +4
Large language models (LLMs) empowered by chain-of-thought reasoning have achieved impressive accuracy on complex tasks but suffer from excessive inference costs and latency when a…
cs.CL2024
LLMCL-GEC: Advancing Grammatical Error Correction with LLM-Driven Curriculum Learning
Tao Fang, Derek F. Wong, Lusheng Zhang +5
While large-scale language models (LLMs) have demonstrated remarkable capabilities in specific natural language processing (NLP) tasks, they may still lack proficiency compared to…