5 papers
Dynamic Role Assignment for Multi-Agent Debate
Miao Zhang, Junsik Kim, Siyuan Xiang +2
Multi-agent large language model (LLM) and vision-language model (VLM) debate systems employ specialized roles for complex problem-solving, yet model specializations are not levera…
Fact-Checking with Large Language Models via Probabilistic Certainty and Consistency
Haoran Wang, Maryam Khalid, Qiong Wu +2
Large language models (LLMs) are increasingly used in applications requiring factual accuracy, yet their outputs often contain hallucinated responses. While fact-checking can mitig…
Efficient Toxic Content Detection by Bootstrapping and Distilling Large Language Models
Jiang Zhang, Qiong Wu, Yiming Xu +3
Toxic content detection is crucial for online services to remove inappropriate content that violates community standards. To automate the detection process, prior works have propos…
Human Transcription Quality Improvement
Jian Gao, Hanbo Sun, Cheng Cao +1
High quality transcription data is crucial for training automatic speech recognition (ASR) systems. However, the existing industry-level data collection pipelines are expensive to…
HTEC: Human Transcription Error Correction
Hanbo Sun, Jian Gao, Xiaomin Wu +3
High-quality human transcription is essential for training and improving Automatic Speech Recognition (ASR) models. Recent study~\cite{libricrowd} has found that every 1% worse tra…