activity
20222024
most citedApplying wav2vec2 for Speech Recognition on Bengali Common Voices Dataset

5 citations · 10 across the 2 of their papers we have counts for

collaborators

5 papers

cs.CL2024

Too Late to Train, Too Early To Use? A Study on Necessity and Viability of Low-Resource Bengali LLMs

Tamzeed Mahfuz, Satak Kumar Dey, Ruwad Naswan +3

Each new generation of English-oriented Large Language Models (LLMs) exhibits enhanced cross-lingual transfer capabilities and significantly outperforms older LLMs on low-resource…

cs.CV2024

IllusionVQA: A Challenging Optical Illusion Dataset for Vision Language Models

Haz Sameen Shahgir, Khondker Salman Sayeed, Abhik Bhattacharjee +3

The advent of Vision Language Models (VLM) has allowed researchers to investigate the visual understanding of a neural network using natural language. Beyond object classification…

eess.IV2023

Leveraging Complementary Attention maps in vision transformers for OCT image analysis

Haz Sameen Shahgir, Tanjeem Azwad Zaman, Khondker Salman Sayeed +3

Optical Coherence Tomography (OCT) scan yields all possible cross-section images of a retina for detecting biomarkers linked to optical defects. Due to the high volume of data gene…

cs.CL20235 cited

Bangla Grammatical Error Detection Using T5 Transformer Model

H. A. Z. Sameen Shahgir, Khondker Salman Sayeed

This paper presents a method for detecting grammatical errors in Bangla using a Text-to-Text Transfer Transformer (T5) Language Model, using the small variant of BanglaT5, fine-tun…

eess.AS20225 cited

Applying wav2vec2 for Speech Recognition on Bengali Common Voices Dataset

H. A. Z. Sameen Shahgir, Khondker Salman Sayeed, Tanjeem Azwad Zaman

Speech is inherently continuous, where discrete words, phonemes and other units are not clearly segmented, and so speech recognition has been an active research problem for decades…