2 papers
cs.CL2025
MiniCPM4: Ultra-Efficient LLMs on End Devices
MiniCPM Team, Chaojun Xiao, Yuxuan Li +80
This paper introduces MiniCPM4, a highly efficient large language model (LLM) designed explicitly for end-side devices. We achieve this efficiency through systematic innovation in…
cs.LG2023
Dynamic Ensemble of Low-fidelity Experts: Mitigating NAS "Cold-Start"
Junbo Zhao, Xuefei Ning, Enshu Liu +7
Predictor-based Neural Architecture Search (NAS) employs an architecture performance predictor to improve the sample efficiency. However, predictor-based NAS suffers from the sever…