EmptyMind at BLP-2023 Task 1: A Transformer-based Hierarchical-BERT Model for Bangla Violence-Inciting Text Detection

Udoy Das, Karnis Fatema, Md Ayon Mia, Mahshar Yahan, Md Sajidul Mowla, Md Fayez Ullah, Arpita Sarker, Hasan Murad · 2023

The availability of the internet has made it easier for people to share information via social media.People with ill intent can use this widespread availability of the internet to share violent content easily.A significant portion of social media users prefer using their regional language which makes it quite difficult to detect violence-inciting text.The objective of our research work is to detect Bangla violence-inciting text from social media content.A shared task on Bangla violenceinciting text detection has been organized by the First Bangla Language Processing Workshop (BLP) co-located with EMNLP, where the organizer has provided a dataset named VITD with three categories: nonviolence, passive violence, and direct violence text.To accomplish this task, we have implemented three machine learning models (RF, SVM, XG-Boost), two deep learning models (LSTM, BiLSTM), and two transformer-based models (BanglaBERT, Hierarchical-BERT).We have conducted a comparative study among different models by training and evaluating each model on the VITD dataset.We have found that Hierarchical-BERT has provided the best result with an F1 score of 0.73797 on the test set and ranked 9 th position among all participants in the shared task 1 of the BLP Workshop co-located with EMNLP 2023.

Read the paper · More papers on PaperTik