3 papers
cs.CL2026
TrustDABench: Benchmarking Reliability and Robustness of LLMs for Structured Data Analysis
Boshen Shi, Yize Liu, Chen Zhao +4
LLMs are increasingly used to analyze spreadsheets, CSV files, and other structured data, but producing a correct-looking answer is not the same as producing a trustworthy analysis…
cs.AI2025
JT-DA: Enhancing Data Analysis with Tool-Integrated Table Reasoning Large Language Models
Ce Chi, Xing Wang, Zhendong Wang +9
In this work, we present JT-DA-8B (JiuTian Data Analyst 8B), a specialized large language model designed for complex table reasoning tasks across diverse real-world scenarios. To a…
cs.CL2025
TReB: A Comprehensive Benchmark for Evaluating Table Reasoning Capabilities of Large Language Models
Ce Li, Xiaofan Liu, Zhiyan Song +10
The majority of data in businesses and industries is stored in tables, databases, and data warehouses. Reasoning with table-structured data poses significant challenges for large l…