Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Pairwise Preference Reward and Group-Based Diversity Enhancement for Superior Open-Ended Generation
Guining Cao, Jiaxin Peng, Chu Zeng +3
Current reinforcement learning(RL) methods are broadly applicable and powerful in verifiable settings where scalar rewards can be provided. However, in open-ended generation tasks,…
cs.AI2025
TableReasoner: Advancing Table Reasoning Framework with Large Language Models
Sishi Xiong, Dakai Wang, Yu Zhao +8
The paper presents our system developed for table question answering (TQA). TQA tasks face challenges due to the characteristics of real-world tabular data, such as large size, inc…