1 citations · 1 across the 1 of their papers we have counts for
2 papers
cs.CL2026★ 1 cited
DEER: A Benchmark for Evaluating Deep Research Agents on Expert Report Generation
Janghoon Han, Heegyu Kim, Changho Lee +6
Recent advances in large language models have enabled deep research systems that generate expert-level reports through multi-step reasoning and evidence-based synthesis. However, e…
cs.CL2025
LRAGE: Legal Retrieval Augmented Generation Evaluation Tool
Minhu Park, Hongseok Oh, Eunkyung Choi +1
Recently, building retrieval-augmented generation (RAG) systems to enhance the capability of large language models (LLMs) has become a common practice. Especially in the legal doma…