20 citations · 20 across the 2 of their papers we have counts for
2 papers
cs.CL2025
ExpertLongBench: Benchmarking Language Models on Expert-Level Long-Form Generation Tasks with Structured Checklists
Jie Ruan, Inderjeet Nair, Shuyang Cao +14
This paper introduces ExpertLongBench, an expert-level benchmark containing 11 tasks from 9 domains that reflect realistic expert workflows and applications. Beyond question answer…
cs.CL2022★ 20 cited
Multi-LexSum: Real-World Summaries of Civil Rights Lawsuits at Multiple Granularities
Zejiang Shen, Kyle Lo, Lauren Yu +3
With the advent of large language models, methods for abstractive summarization have made great strides, creating potential for use in applications to aid knowledge workers process…