3 papers
cs.IR2026
LLM Retrieval for Stable and Predictable Ad Recommendations
Vinodh Kumar Sunkara, Satheeshkumar Karuppusamy, Hangjun Xu +13
Traditional ads recommendation systems have primarily focused on optimizing for prediction accuracy of click or conversion events using canonical metrics such as recall or normaliz…
cs.AI2026
When AI reviews science: Can we trust the referee?
Jialiang Wang, Yuchen Liu, Hang Xu +7
The volume of scientific submissions continues to climb, outpacing the capacity of qualified human referees and stretching editorial timelines. At the same time, modern large langu…
cs.CL2025
Beyond path selection: Better LLMs for Scientific Information Extraction with MimicSFT and Relevance and Rule-induced(R)GRPO
Ran Li, Shimin Di, Yuchen Liu +3
Previous study suggest that powerful Large Language Models (LLMs) trained with Reinforcement Learning with Verifiable Rewards (RLVR) only refines reasoning path without improving t…