Explainable AI Method for Cyber bullying Detection
Varsha Pawar, Deepa V. Jose, Ashwini Patil · 2022 IEEE 2nd International Conference on Mobile Networks and Wireless Communications (ICMNWC) · 2022
People of all ages and genders are using social media platforms to engage themselves in all sorts of activities. People create profiles on online social networks in order to communicate with one another in this virtual environment. Hundreds or thousands of friends and followers are split across many profiles. Along with the virtual communication in this social media life, cyber-crimes also creep in many distinguished forms to grab user’s information and emotionally degrade them with harassment and arrogant behavior. A set of machine learning methods are proposed and used to detect such a bullying behavior. Along with the detection of such an act, the model should also provide the logical reasoning of the evidence extracted. The explain ability of the models classification will give us a view of the way towards portraying a suspect as a bullier. This paper illustrates a machine learning model that works on a twitter data set to suggest the tweets as category bullying or non-bullying. LIME a tool to predict the interpretability of the model is used to depict the performance of model and provides explainability.