2 papers
cs.CV2026
ChartSync: A Benchmark for Visuo-Logical Cascading Chart Editing
Jiakang Yu, Yixuan Chai, Tianci Wang +7
Generative image editing models struggle with structured statistical charts when data modifications require geometric synchronization. We formalize this task as Visuo-Logical Casca…
cs.CL2026
TaxPraBen: A Scalable Benchmark for Structured Evaluation of LLMs in Chinese Real-World Tax Practice
Gang Hu, Yating Chen, Haiyan Ding +5
While Large Language Models (LLMs) excel in various general domains, they exhibit notable gaps in the highly specialized, knowledge-intensive, and legally regulated Chinese tax dom…