2 papers
cs.LG2026
Data Scout: Targeted Web Crawling for Domain-Specific Pretraining Corpora
Chirag Garg, Eelaaf Zahid, Farhan Ahmed +3
The dominant approach to building domain-specific pretraining corpora is to filter large web archives such as CommonCrawl. This works well for popular domains but breaks down for s…
cs.CL2026
NC-Bench: An LLM Benchmark for Evaluating Conversational Competence
Robert J. Moore, Sungeun An, Farhan Ahmed +1
The Natural Conversation Benchmark (NC-Bench) introduces a new approach to evaluating the general conversational competence of large language models (LLMs). Unlike prior benchmarks…