Collected Bangla Comments Dataset
收藏资源简介:
The dataset comprises over 70,000 Bangla-language comments.Approximately 50,000 were sourced from a Kaggle dataset focused on social media discourse, while 20,000 were gathered from publicly available Bangla text corpora, including Facebook, Twitter and forums. These comments reflect diverse communication styles and topics. Data were filtered for relevance and quality, and duplicates were removed. No specialized instruments were used.Data were sourced from publicly available platforms relevant to Bangladesh, including Facebook comments and Google-hosted datasets. While user content originated in Bangladesh, no geographic coordinates were recorded due to anonymization. The dataset is stored at the Department of CSE, International Islamic University Chittagong (IIUC).



