From the 1 of 11 linked papers with an AI index.
4 papers · 1 filter
The Missing Piece in Pre-trained Model Evaluation: Reward-Guided Decoding Unlocks Task-Oriented Behavior Without Parameter Updates
Shaobo Wang, Guo Chen, Ziyue Wang +5
With the rapid progress of large language models (LLMs), reliably evaluating the capabilities of pre-trained LLMs has become increasingly important. The challenge is that base pre-…
OPUS: Towards Efficient and Principled Data Selection in Large Language Model Pre-training in Every Iteration
Shaobo Wang, Xuan Ouyang, Tianyi Xu +9
As high-quality public text approaches exhaustion, a phenomenon known as the Data Wall, pre-training is shifting from more tokens to better tokens. However, existing methods either…
LIRAG: A Lightweight Rerank Reasoning Strategy Framework for Retrieval-Augmented Generation
Guo Chen, Junjie Huang, Huaijin Xie +2
Retrieval-Augmented Generation (RAG) effectively enhances Large Language Models (LLMs) by incorporating retrieved external knowledge into the generation process. Reasoning models i…
M2G-Eval: Enhancing and Evaluating Multi-granularity Multilingual Code Generation
Fanglin Xu, Wei Zhang, Jian Yang +5
The rapid advancement of code large language models (LLMs) has sparked significant research interest in systematically evaluating their code generation capabilities, yet existing b…