Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:Bangla Text Dataset and Exploratory Analysis for Online Harassment Detection

Feb 04, 2021

Md Faisal Ahmed, Zalish Mahmud, Zarin Tasnim Biash, Ahmed Ann Noor Ryen, Arman Hossain, Faisal Bin Ashraf

Figure 1 for Bangla Text Dataset and Exploratory Analysis for Online Harassment Detection

Figure 2 for Bangla Text Dataset and Exploratory Analysis for Online Harassment Detection

Figure 3 for Bangla Text Dataset and Exploratory Analysis for Online Harassment Detection

Figure 4 for Bangla Text Dataset and Exploratory Analysis for Online Harassment Detection

Share this with someone who'll enjoy it:

Abstract:Being the seventh most spoken language in the world, the use of the Bangla language online has increased in recent times. Hence, it has become very important to analyze Bangla text data to maintain a safe and harassment-free online place. The data that has been made accessible in this article has been gathered and marked from the comments of people in public posts by celebrities, government officials, athletes on Facebook. The total amount of collected comments is 44001. The dataset is compiled with the aim of developing the ability of machines to differentiate whether a comment is a bully expression or not with the help of Natural Language Processing and to what extent it is improper if it is an inappropriate comment. The comments are labeled with different categories of harassment. Exploratory analysis from different perspectives is also included in this paper to have a detailed overview. Due to the scarcity of data collection of categorized Bengali language comments, this dataset can have a significant role for research in detecting bully words, identifying inappropriate comments, detecting different categories of Bengali bullies, etc. The dataset is publicly available at https://data.mendeley.com/datasets/9xjx8twk8p.

* 3 pages, 5 tables, 6 figures

View paper on

Share this with someone who'll enjoy it:

Title:Bangla Text Dataset and Exploratory Analysis for Online Harassment Detection

Paper and Code