4 papers
Pairwise Preference Reward and Group-Based Diversity Enhancement for Superior Open-Ended Generation
Guining Cao, Jiaxin Peng, Chu Zeng +3
Current reinforcement learning(RL) methods are broadly applicable and powerful in verifiable settings where scalar rewards can be provided. However, in open-ended generation tasks,…
T2R-bench: A Benchmark for Generating Article-Level Reports from Real World Industrial Tables
Jie Zhang, Changzai Pan, Kaiwen Wei +12
Extensive research has been conducted to explore the capabilities of large language models (LLMs) in table reasoning. However, the essential task of transforming tables information…
Technical Report of TeleChat2, TeleChat2.5 and T1
Zihan Wang, Xinzhang Liu, Yitong Yao +35
We introduce the latest series of TeleChat models: \textbf{TeleChat2}, \textbf{TeleChat2.5}, and \textbf{T1}, offering a significant upgrade over their predecessor, TeleChat. Despi…
Acceleration noise due to Space Magnetic Field for Heliocentric Gravitational Wave Detector
Peng Jia-Hui, Zhang Ji-Xiang, Hong W +4
The space-borne gravitational wave observatory is to detect low-frequency gravitational wave signals in the range of 0.1 mHz to 100 mHz. The inertial sensors of space gravitational…