Towards More Robust Methods of Cyberbullying Detection Öffentlichkeit

Ziems, Caleb (Spring 2020)

Permanent URL: https://etd.library.emory.edu/concern/etds/zs25x960d?locale=de
Published

Abstract

Cyberbullying is a pervasive problem in online communities. To identify cyberbullying cases in large-scale social networks, content moderators depend on machine learning classifiers for automatic cyberbullying detection. However, existing models remain unfit for real-world applications, largely due to a shortage of publicly available training data and a lack of standard criteria for assigning ground truth labels. In this study, we address the need for reliable data using an original annotation framework. Inspired by social sciences research into bullying behavior, we characterize the nuanced problem of cyberbullying using five explicit factors to represent its social and linguistic aspects. We model this behavior using social network and language-based features, which improves classifier performance. Lastly, we develop a method for inferring the target of aggression in the message thread, and we evaluate this approach on hand-labeled data. These results demonstrate the importance of representing and modeling cyberbullying as a social phenomenon.

Introduction.............................1

Background..............................4

Data.......................................10

Feature Engineering................16

Model Evaluation....................27

Inferring the Target User.........32

Conclusion.............................36

About this Honors Thesis

Rights statement

Permission granted by the author to include this thesis or dissertation in this repository. All rights reserved by the author. Please contact the author for information regarding the reproduction and use of this thesis or dissertation.

School	Emory College
Department	Computer Science
Degree	B.S.
Submission	Honors Thesis
Language	English
Research Field	Computer Science
Stichwort	machine learning dataset detection features harassment cyberbullying social definition context twitter
Committee Chair / Thesis Advisor	Ymir Vigfusson, Emory University
Committee Members	Marjorie Pak, Emory University Phillip Wolff, Emory University Eugene Agichtein, Emory University

Zuletzt geändert

Primary PDF

Thumbnail	Title	Date Uploaded	Actions
	Towards More Robust Methods of Cyberbullying Detection ()	2020-04-27 21:40:26 -0400	Download

Towards More Robust Methods of Cyberbullying Detection Öffentlichkeit

Ziems, Caleb (Spring 2020)

Abstract

Table of Contents

About this Honors Thesis

Primary PDF

Supplemental Files