2 papers
cs.CL2026
DocScope: Benchmarking Verifiable Reasoning for Trustworthy Long-Document Understanding
Xiang Feng, Jiawei Zhou, Zhangfeng Huang +6
Evaluating whether Multimodal Large Language Models can produce trustworthy, verifiable reasoning over long, visually rich documents requires evaluation beyond end-to-end answer ac…
cs.LG2025
The Lossy Horizon: Error-Bounded Predictive Coding for Lossy Text Compression (Episode I)
Nnamdi Aghanya, Jun Li, Kewei Wang
Large Language Models (LLMs) can achieve near-optimal lossless compression by acting as powerful probability models. We investigate their use in the lossy domain, where reconstruct…