← Search

Md Fahim

4 accepted papers

2025

BANMIME : Misogyny Detection with Metaphor Explanation on Bangla Memes

EMNLP 2025

Detecting misogyny in multimodal content remains a notable challenge, particularly in culturally conservative and low-resource contexts like Bangladesh. While existing research has explored hate speech and general meme classification, the nuanced identification of misogyny in Bangla memes, rich in m

2025

BanTH: A Multi-label Hate Speech Detection Dataset for Transliterated Bangla

NAACL 2025findings

The proliferation of transliterated texts in digital spaces has emphasized the need for detecting and classifying hate speech in languages beyond English, particularly in low-resource languages. As online discourse can perpetuate discrimination based on target groups, e.g. gender, religion, and orig…

Cited by 0SourcePDFScholar
2025

DM-Codec: Distilling Multimodal Representations for Speech Tokenization

EMNLP 2025

Recent advancements in speech-language models have yielded significant improvements in speech tokenization and synthesis. However, effectively mapping the complex, multidimensional attributes of speech into discrete tokens remains challenging. This process demands acoustic, semantic, and contextual

2024

BanglaTLit: A Benchmark Dataset for Back-Transliteration of Romanized Bangla

EMNLP 2024finding

Low-resource languages like Bangla are severely limited by the lack of datasets. Romanized Bangla texts are ubiquitous on the internet, offering a rich source of data for Bangla NLP tasks and extending the available data sources. However, due to the informal nature of romanized text, they often lack…