3 papers
cs.IR2025
Optimizing Generative Ranking Relevance via Reinforcement Learning in Xiaohongshu Search
Ziyang Zeng, Heming Jing, Jindong Chen +11
Ranking relevance is a fundamental task in search engines, aiming to identify the items most relevant to a given user query. Traditional relevance models typically produce scalar s…
cs.AI2025
BIRD-INTERACT: Re-imagining Text-to-SQL Evaluation for Large Language Models via Lens of Dynamic Interactions
Nan Huo, Xiaohan Xu, Jinyang Li +21
Large language models (LLMs) have demonstrated remarkable performance on single-turn text-to-SQL tasks, but real-world database applications predominantly require multi-turn intera…
cs.CL2025
Uncovering the Fragility of Trustworthy LLMs through Chinese Textual Ambiguity
Xinwei Wu, Haojie Li, Hongyu Liu +4
In this work, we study a critical research problem regarding the trustworthiness of large language models (LLMs): how LLMs behave when encountering ambiguous narrative text, with a…