137 citations · 239 across the 24 of their papers we have counts for
5 papers · 1 filter
From Trajectories to Prefixes: Reusing Teacher Trajectories via Replayed Prefixes and Online Continuation
Yihan Wang, Zhong Guan, Haoran Sun +3
Small language models are attractive backbones for interactive agents, but direct distillation from strong teacher trajectories often turns rich multi-turn behavior into one-shot i…
Missing Old Logits in Asynchronous Agentic RL: Semantic Mismatch and Repair Methods for Off-Policy Correction
Zhong Guan, Yongjian Guo, Haoran Sun +5
Asynchronous reinforcement learning improves rollout throughput for large language model agents by decoupling sample generation from policy optimization, but it also introduces a c…
KMF: Knowledge-Aware Multi-Faceted Representation Learning for Zero-Shot Node Classification
Likang Wu, Junji Jiang, Hongke Zhao +4
Recently, Zero-Shot Node Classification (ZNC) has been an emerging and crucial task in graph data analysis. This task aims to predict nodes from unseen classes which are unobserved…
Estimating Fund-Raising Performance for Start-up Projects from a Market Graph Perspective
Likang Wu, Zhi Li, Hongke Zhao +2
In the online innovation market, the fund-raising performance of the start-up project is a concerning issue for creators, investors and platforms. Unfortunately, existing studies a…
Estimating Early Fundraising Performance of Innovations via Graph-based Market Environment Model
Likang Wu, Zhi Li, Hongke Zhao +3
Well begun is half done. In the crowdfunding market, the early fundraising performance of the project is a concerned issue for both creators and platforms. However, estimating the…