10 papers
VDAR-Router: Adaptive LLMs Routing via Verbalized Query Difficulty Analysis Retrieval
Yu-Chien Tang, Jun-Chen Hung, Wen-Chih Peng +1
Large language models are increasingly used in practical systems, making efficient model selection important for reducing deployment cost. LLM routing has emerged as a practical so…
OI-Bench: An Option Injection Benchmark for Evaluating LLM Susceptibility to Directive Interference
Yow-Fu Liou, Yu-Chien Tang, Yu-Hsiang Liu +1
Benchmarking large language models (LLMs) is critical for understanding their capabilities, limitations, and robustness. In addition to interface artifacts, prior studies have show…
VISTA: A Controllable Platform for Generating and Auditing Egocentric Assistance Scenarios
Yu-Hsiang Liu, Yu-Chien Tang, An-Zi Yen
Evaluating whether AI agents can proactively assist humans in daily activities, ranging from routine household tasks to urgent safety-critical situations, requires diverse visual d…
Tree-of-Text: A Tree-based Prompting Framework for Table-to-Text Generation in the Sports Domain
Shang-Hsuan Chiang, Tsan-Tsung Yang, An-Zi Yen +1
Generating sports game reports from structured tables is a complex table-to-text task that demands both precise data interpretation and fluent narrative generation. Traditional mod…
ConceptKT: A Benchmark for Concept-Level Deficiency Prediction in Knowledge Tracing
Yu-Chen Kang, Yu-Chien Tang, An-Zi Yen
Knowledge Tracing (KT) is a critical technique for modeling student knowledge to support personalized learning. However, most KT systems focus on binary correctness prediction and…
DaMO: A Data-Efficient Multimodal Orchestrator for Temporal Reasoning with Video LLMs
Bo-Cheng Chiu, Jen-Jee Chen, Yu-Chee Tseng +2
Large Language Models (LLMs) have recently been extended to the video domain, enabling sophisticated video-language understanding. However, existing Video LLMs often exhibit limita…