1 paper
Nikhil Abhyankar, Purvi Chaurasia, Sanchit Kabra +3
Existing tabular reasoning benchmarks mostly test models on small, uniform tables, underrepresenting the complexity of real-world data and giving an incomplete view of Large Langua…