Showing 2025Show all
2 papers · 1 filter
cs.CL2025
DeepWideSearch: Benchmarking Depth and Width in Agentic Information Seeking
Tian Lan, Bin Zhu, Qianghuai Jia +6
Current search agents fundamentally lack the ability to simultaneously perform \textit{deep} reasoning over multi-hop retrieval and \textit{wide}-scale information collection-a cri…
cs.CL2025
TransBench: Benchmarking Machine Translation for Industrial-Scale Applications
Haijun Li, Tianqi Shi, Zifu Shang +13
Machine translation (MT) has become indispensable for cross-border communication in globalized industries like e-commerce, finance, and legal services, with recent advancements in…