4 papers
Solving Formal Math Problems by Decomposition and Iterative Reflection
Yichi Zhou, Jianqiu Zhao, Yongxin Zhang +14
General-purpose Large Language Models (LLMs) have achieved remarkable success in intelligence, performing comparably to human experts on complex reasoning tasks such as coding and…
FinTeam: A Multi-Agent Collaborative Intelligence System for Comprehensive Financial Scenarios
Yingqian Wu, Qiushi Wang, Zefei Long +7
Financial report generation tasks range from macro- to micro-economics analysis, also requiring extensive data analysis. Existing LLM models are usually fine-tuned on simple QA tas…
Multi-agent KTO: Reinforcing Strategic Interactions of Large Language Model in Language Game
Rong Ye, Yongxin Zhang, Yikai Zhang +3
Achieving Artificial General Intelligence (AGI) requires AI agents that can not only make stratigic decisions but also engage in flexible and meaningful communication. Inspired by…
AgentSense: Benchmarking Social Intelligence of Language Agents through Interactive Scenarios
Xinyi Mou, Jingcong Liang, Jiayu Lin +8
Large language models (LLMs) are increasingly leveraged to empower autonomous agents to simulate human beings in various fields of behavioral research. However, evaluating their ca…