2 papers
cs.AI2025
Video-VoT-R1: An efficient video inference model integrating image packing and AoE architecture
Cheng Li, Jiexiong Liu, Yixuan Chen +1
In the field of video-language pretraining, existing models face numerous challenges in terms of inference efficiency and multimodal data processing. This paper proposes a KunLunBa…
cs.CL2025
KunlunBaize: LLM with Multi-Scale Convolution and Multi-Token Prediction Under TransformerX Framework
Cheng Li, Jiexiong Liu, Yixuan Chen +2
Large language models have demonstrated remarkable performance across various tasks, yet they face challenges such as low computational efficiency, gradient vanishing, and difficul…