2 papers
cs.CL2026
Less Is More: Reducing Token Counts Without Compromising Performance
Gyeongje Cho, Yeonkyoung So, Sangmin Lee +1
Tokenization directly affects the inference efficiency of large language models, since fragmented tokenization increases sequence length and generation cost. Although longer, multi…
cs.CL2026
Choices Speak Louder than Questions
Gyeongje Cho, Yeonkyoung So, Jaejin Lee
Recent findings raise concerns about whether the evaluation of Multiple-Choice Question Answering (MCQA) accurately reflects the comprehension abilities of large language models. T…