3 citations · 3 across the 6 of their papers we have counts for
4 papers · 1 filter
AutoScientist-Quant: Self-Evolving Coding Agents for Automatic Research in Quantitative Investment
Zongqian Li, Yaoyiran Li, Yaohui Guo +3
Large language model agents can discover alphas, yet current methods have three weaknesses. The search cannot adapt during the run, automation usually ends at alpha generation whil…
WebRider: Persona-Conditioned Intent Controllers for Live-Web Assistance
Zhi Li, Tao Zhou, Yeqing Li +2
Delegating a web task involves more than asking a question; it requires transferring a policy: what to verify, how to handle uncertainty, which preferences matter, and when to stop…
How Fast Should a Model Commit to Supervision? Training Reasoning Models on the Tsallis Loss Continuum
Chu-Cheng Lin, Eugene Ie
SFT-then-RLVR is widely used for post-training reasoning models, but why this specific ordering, and why RLVR-only stalls at cold start, have lacked a unifying theoretical account.…
ConvApparel: A Benchmark Dataset and Validation Framework for User Simulators in Conversational Recommenders
Ofer Meshi, Krisztian Balog, Sally Goldman +5
The promise of LLM-based user simulators to improve conversational AI is hindered by a critical "realism gap," leading to systems that are optimized for simulated interactions, but…