1 paper
Pengzuo Wu, Yuhang Yang, Guangcheng Zhu +10
With the rapid advancement of Large Language Models (LLMs), there is an increasing need for challenging benchmarks to evaluate their capabilities in handling complex tabular data.…