agent evaluation 1asynchronous runtime 1benchmark 1benchmarking 1computer-use agents 1cross-platform evaluation 1large language models 1reward modeling 1software infrastructure 1vision-language models 1
From the 2 of 25 linked papers with an AI index.
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Intern-S1-Pro: Scientific Multimodal Foundation Model at Trillion Scale
Yicheng Zou, Dongsheng Zhu, Lin Zhu +174
We introduce Intern-S1-Pro, the first one-trillion-parameter scientific multimodal foundation model. Scaling to this unprecedented size, the model delivers a comprehensive enhancem…
cs.LG2025
RELIEF: Reinforcement Learning Empowered Graph Feature Prompt Tuning
Jiapeng Zhu, Zichen Ding, Jianxiang Yu +3
The advent of the "pre-train, prompt" paradigm has recently extended its generalization ability and data efficiency to graph representation learning, following its achievements in…