collaborators

6 papers

cs.CV2025

The Jumping Reasoning Curve? Tracking the Evolution of Reasoning Performance in GPT-[n] and o-[n] Models on Multimodal Puzzles

Vernon Y. H. Toh, Yew Ken Chia, Deepanway Ghosal +1

The releases of OpenAI's o-[n] series, such as o1, o3, and o4-mini, mark a significant paradigm shift in Large Language Models towards advanced reasoning capabilities. Notably, mod…

cs.CL2025

PromptDistill: Query-based Selective Token Retention in Intermediate Layers for Efficient Large Language Model Inference

Weisheng Jin, Maojia Song, Tej Deep Pala +4

As large language models (LLMs) tackle increasingly complex tasks and longer documents, their computational and memory costs during inference become a major bottleneck. To address…

cs.CL2024

M-Longdoc: A Benchmark For Multimodal Super-Long Document Understanding And A Retrieval-Aware Tuning Framework

Yew Ken Chia, Liying Cheng, Hou Pong Chan +5

The ability to understand and answer questions over documents can be useful in many business and practical applications. However, documents often contain lengthy and diverse multim…

cs.CL2024

Domain-Expanded ASTE: Rethinking Generalization in Aspect Sentiment Triplet Extraction

Yew Ken Chia, Hui Chen, Wei Han +4

Aspect Sentiment Triplet Extraction (ASTE) is a challenging task in sentiment analysis, aiming to provide fine-grained insights into human sentiments. However, existing benchmarks…

cs.CL2024

Reasoning Paths Optimization: Learning to Reason and Explore From Diverse Paths

Yew Ken Chia, Guizhen Chen, Weiwen Xu +3

Advanced models such as OpenAI o1 exhibit impressive problem-solving capabilities through step-by-step reasoning. However, they may still falter on more complex problems, making er…

cs.CL2024

Auto-Arena: Automating LLM Evaluations with Agent Peer Battles and Committee Discussions

Ruochen Zhao, Wenxuan Zhang, Yew Ken Chia +3

As LLMs continuously evolve, there is an urgent need for a reliable evaluation method that delivers trustworthy results promptly. Currently, static benchmarks suffer from inflexibi…