From the 1 of 3 linked papers with an AI index.
3 papers
cs.SE2026
Inference Economics of Enterprise Coding Agents: A Case Study of Cloud vs. On-Premise LLMs
Sheng-Wei Peng, Yi-Hsun Lin, Yi-Pei Lee
The paper compares the economic and quality trade‑offs of using an API‑based large language model versus an on‑premise quantized model for autonomous coding agents in a real‑world…
cs.AI2025
Multi-TW: Benchmarking Multimodal Models on Traditional Chinese Question Answering in Taiwan
Jui-Ming Yao, Bing-Cheng Xie, Sheng-Wei Peng +5
Multimodal Large Language Models (MLLMs) process visual, acoustic, and textual inputs, addressing the limitations of single-modality LLMs. However, existing benchmarks often overlo…
cs.CL2025
Token Constraint Decoding Improves Robustness on Question Answering for Large Language Models
Jui-Ming Yao, Hao-Yuan Chen, Zi-Xian Tang +4
Large Language Models (LLMs) have demonstrated impressive performance on multiple-choice question answering (MCQA) benchmarks, yet they remain highly vulnerable to minor input pert…