4 papers
OmniTable: A Unified Wide-Table System for Petabyte-Scale LLM Data Curation and Exploration
Yuzhuo Fu, Xiangchun Wang, Chao Huang +16
Data curation is a critical bottleneck in industrial-grade LLM development, where petabyte-scale unstructured corpora are scattered across hundreds of physical tables, feature engi…
Revisiting a Pain in the Neck: A Semantic Reasoning Benchmark for Language Models
Yang Liu, Hongming Li, Melissa Xiaohui Qin +2
We present SemanticQA, an evaluation suite designed to assess language models (LMs) in semantic phrase processing tasks. The benchmark consolidates existing multiword expression (M…
Entropy-Based Data Selection for Language Models
Hongming Li, Yang Liu, Chao Huang
Modern language models (LMs) increasingly require two critical resources: computational resources and data resources. Data selection techniques can effectively reduce the amount of…
Revisiting a Pain in the Neck: Semantic Phrase Processing Benchmark for Language Models
Yang Liu, Melissa Xiaohui Qin, Hongming Li +1
We introduce LexBench, a comprehensive evaluation suite enabled to test language models (LMs) on ten semantic phrase processing tasks. Unlike prior studies, it is the first work to…