most citedBanglaBait: Semi-Supervised Adversarial Approach for Clickbait Detection on Bangla Clickbait Dataset

5 citations · 8 across the 5 of their papers we have counts for

collaborators

5 papers

cs.CL20235 cited

BanglaBait: Semi-Supervised Adversarial Approach for Clickbait Detection on Bangla Clickbait Dataset

Md. Motahar Mahtab, Monirul Haque, Mehedi Hasan +1

Intentionally luring readers to click on a particular content by exploiting their curiosity defines a title as clickbait. Although several studies focused on detecting clickbait ti…

cs.HC2023

Bornil: An open-source sign language data crowdsourcing platform for AI enabled dialect-agnostic communication

Shahriar Elahi Dhruvo, Mohammad Akhlaqur Rahman, Manash Kumar Mandal +14

The absence of annotated sign language datasets has hindered the development of sign language recognition and translation technologies. In this paper, we introduce Bornil; a crowds…

cs.CV20232 cited

bbOCR: An Open-source Multi-domain OCR Pipeline for Bengali Documents

Imam Mohammad Zulkarnain, Shayekh Bin Islam, Md. Zami Al Zunaed Farabe +9

Despite the existence of numerous Optical Character Recognition (OCR) tools, the lack of comprehensive open-source systems hampers the progress of document digitization in various…

eess.AS2023

OOD-Speech: A Large Bengali Speech Recognition Dataset for Out-of-Distribution Benchmarking

Fazle Rabbi Rakib, Souhardya Saha Dip, Samiul Alam +11

We present OOD-Speech, the first out-of-distribution (OOD) benchmarking dataset for Bengali automatic speech recognition (ASR). Being one of the most spoken languages globally, Ben…

cs.CV20231 cited

BaDLAD: A Large Multi-Domain Bengali Document Layout Analysis Dataset

Md. Istiak Hossain Shihab, Md. Rakibul Hasan, Mahfuzur Rahman Emon +14

While strides have been made in deep learning based Bengali Optical Character Recognition (OCR) in the past decade, the absence of large Document Layout Analysis (DLA) datasets has…