distribution shift 1LLM agents 1multi-turn interaction 1offline training 1on-policy distillation 1prefix replay 1
From the 1 of 33 linked papers with an AI index.
1 citations · 1 across the 17 of their papers we have counts for
Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
DocReward: A Document Reward Model for Structuring and Stylizing
Junpeng Liu, Yuzhong Zhao, Bowen Cao +17
Recent agentic workflows automate professional document generation but focus narrowly on textual quality, overlooking structural and stylistic professionalism, which is equally cri…
cs.CV2025
Model as a Game: On Numerical and Spatial Consistency for Generative Games
Jingye Chen, Yuzhong Zhao, Yupan Huang +5
Recent advances in generative models have significantly impacted game generation. However, despite producing high-quality graphics and adequately receiving player input, existing m…