2 papers
cs.CL2024
A Step Towards Mixture of Grader: Statistical Analysis of Existing Automatic Evaluation Metrics
Yun Joon Soh, Jishen Zhao
The explosion of open-sourced models and Question-Answering (QA) datasets emphasizes the importance of automated QA evaluation. We studied the statistics of the existing evaluation…
cs.CL2024
You Only Use Reactive Attention Slice For Long Context Retrieval
Yun Joon Soh, Hanxian Huang, Yuandong Tian +1
Supporting longer context for Large Language Models (LLM) is a promising direction to advance LLMs. As training a model for a longer context window is computationally expensive, ma…