2 papers
cs.CL2026
OI-Bench: An Option Injection Benchmark for Evaluating LLM Susceptibility to Directive Interference
Yow-Fu Liou, Yu-Chien Tang, Yu-Hsiang Liu +1
Benchmarking large language models (LLMs) is critical for understanding their capabilities, limitations, and robustness. In addition to interface artifacts, prior studies have show…
cs.CL2025
Stay Hungry, Stay Foolish: On the Extended Reading Articles Generation with LLMs
Yow-Fu Liou, Yu-Chien Tang, An-Zi Yen
The process of creating educational materials is both time-consuming and demanding for educators. This research explores the potential of Large Language Models (LLMs) to streamline…