3 papers
cs.AI2026
OSCToM: RL-Guided Adversarial Generation for High-Order Theory of Mind
Sharmin Sultana Srishty, Kazi Mahathir Rahman, Malaika Parizat Sakkhi +2
Large Language Models (LLMs) perform well on many language tasks, but their Theory of Mind (ToM) reasoning is still uneven in complex social settings. Existing benchmarks, includin…
cs.CV2025
Speak2Sign3D: A Multi-modal Pipeline for English Speech to American Sign Language Animation
Kazi Mahathir Rahman, Naveed Imtiaz Nafis, Md. Farhan Sadik +2
Helping deaf and hard-of-hearing people communicate more easily is the main goal of Automatic Sign Language Translation. Although most past research has focused on turning sign lan…
cs.CV2025
TextDiffuser-RL: Efficient and Robust Text Layout Optimization for High-Fidelity Text-to-Image Synthesis
Kazi Mahathir Rahman, Showrin Rahman, Sharmin Sultana Srishty
Text-embedded image generation plays a critical role in industries such as graphic design, advertising, and digital content creation. Text-to-Image generation methods leveraging di…