2 papers
cs.CL2022
PerPaDa: A Persian Paraphrase Dataset based on Implicit Crowdsourcing Data Collection
Salar Mohtaj, Fatemeh Tavakkoli, Habibollah Asghari
In this paper we introduce PerPaDa, a Persian paraphrase dataset that is collected from users' input in a plagiarism detection system. As an implicit crowdsourcing experience, we h…
cs.CL2021
Hamtajoo: A Persian Plagiarism Checker for Academic Manuscripts
Vahid Zarrabi, Salar Mohtaj, Habibollah Asghari
In recent years, due to the high availability of electronic documents through the Web, the plagiarism has become a serious challenge, especially among scholars. Various plagiarism…