3 papers
cs.CL2026
AppTek Call-Center Dialogues: A Multi-Accent Long-Form Benchmark for English ASR
Eugen Beck, Sarah Beranek, Uma Moothiringote +4
Evaluating English ASR systems for conversational AI applications remains difficult, as many publicly available corpora are either pre-segmented into short segments, consist of rea…
cs.CL2024
Vietnamese AI Generated Text Detection
Quang-Dan Tran, Van-Quan Nguyen, Quang-Huy Pham +2
In recent years, Large Language Models (LLMs) have become integrated into our daily lives, serving as invaluable assistants in completing tasks. Widely embraced by users, the abuse…
cs.LG2024
ClimateGPT: Towards AI Synthesizing Interdisciplinary Research on Climate Change
David Thulke, Yingbo Gao, Petrus Pelser +23
This paper introduces ClimateGPT, a model family of domain-specific large language models that synthesize interdisciplinary research on climate change. We trained two 7B models fro…