1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.AI2025
PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization
Zouying Cao, Runze Wang, Yifei Yang +4
Large Language Model (LLM) agents have demonstrated impressive capabilities in handling complex interactive problems. Existing LLM agents mainly generate natural language plans to…
cs.CL2024★ 1 cited
CO3: Low-resource Contrastive Co-training for Generative Conversational Query Rewrite
Yifei Yuan, Chen Shi, Runze Wang +5
Generative query rewrite generates reconstructed query rewrites using the conversation history while rely heavily on gold rewrite pairs that are expensive to obtain. Recently, few-…