activity
20242026
collaborators
Showing 2024Show all

13 papers · 1 filter

cs.LG2024

Free Process Rewards without Process Labels

Lifan Yuan, Wendi Li, Huayu Chen +6

Different from its counterpart outcome reward models (ORMs), which evaluate the entire responses, a process reward model (PRM) scores a reasoning trajectory step by step, providing…

cs.AI2024

Automating Exploratory Proteomics Research via Language Models

Ning Ding, Shang Qu, Linhai Xie +13

With the development of artificial intelligence, its contribution to science is evolving from simulating a complex problem to automating entire research processes and producing nov…

cs.CL2024

Scalable Efficient Training of Large Language Models with Low-dimensional Projected Attention

Xingtai Lv, Ning Ding, Kaiyan Zhang +3

Improving the effectiveness and efficiency of large language models (LLMs) simultaneously is a critical yet challenging research goal. In this paper, we find that low-rank pre-trai…

cs.CL2024

UltraMedical: Building Specialized Generalists in Biomedicine

Kaiyan Zhang, Sihang Zeng, Ermo Hua +11

Large Language Models (LLMs) have demonstrated remarkable capabilities across various domains and are moving towards more specialized areas. Recent advanced proprietary models such…

cs.CL2024

Fast and Slow Generating: An Empirical Study on Large and Small Language Models Collaborative Decoding

Kaiyan Zhang, Jianyu Wang, Ning Ding +4

Large Language Models (LLMs) exhibit impressive capabilities across various applications but encounter substantial challenges such as high inference latency, considerable training…

cs.CL2024

Tool Learning with Foundation Models

Yujia Qin, Shengding Hu, Yankai Lin +38

Humans possess an extraordinary ability to create and utilize tools, allowing them to overcome physical limitations and explore new frontiers. With the advent of foundation models,…