3 papers
cs.AI2026
OPOD: On-Policy Omni Distillation
Tong Zhao, Yuyang Hu, Reed Li +5
Omni-modal models provide a unified interface for text, images, and audio. However, improving these abilities together remains difficult, as post-training on pooled multimodal data…
cs.LG2025
A Text-guided Protein Design Framework
Shengchao Liu, Yanjing Li, Zhuoxinran Li +10
Current AI-assisted protein design mainly utilizes protein sequential and structural information. Meanwhile, there exists tremendous knowledge curated by humans in the text format…
cs.CL2024
YuLan-Mini: An Open Data-efficient Language Model
Yiwen Hu, Huatong Song, Jia Deng +8
Effective pre-training of large language models (LLMs) has been challenging due to the immense resource demands and the complexity of the technical processes involved. This paper p…