4 papers
Learning to Evaluate: Cost-Effective Model Evaluation on Unlabeled Data with Meta-Learning
Trinh Pham, Viet Huynh, Hongzhi Yin +2
The rapid advancement of machine learning has led to an unprecedented expansion of model ecosystems, making it increasingly difficult to assess the reliability of newly released mo…
PRISM: A Multi-Dimensional Benchmark for Evaluating LLM Peer Reviewers
Ngoc Phan Phuoc Loc, Toan Huynh La Viet, Thanh Tran Khanh +8
The rapid growth in submissions to machine learning venues has strained the scientific peer-review system and intensified interest in LLM-based automated peer reviewers. However, h…
Comparing Without Saying: A Dataset and Benchmark for Implicit Comparative Opinion Mining from Same-User Reviews
Thanh-Lam T. Nguyen, Ngoc-Quang Le, Quoc-Trung Phu +4
Existing studies on comparative opinion mining have mainly focused on explicit comparative expressions, which are uncommon in real-world reviews. This leaves implicit comparisons -…
Knowledge Graph Enrichment and Reasoning for Nobel Laureates
Thanh-Lam T. Nguyen, Ngoc-Quang Le, Thu-Trang Pham +1
This project aims to construct and analyze a comprehensive knowledge graph of Nobel Prize and Laureates by enriching existing datasets with biographical information extracted from…