Ralf Krestel

You are here: Home > Publications > Workshop Papers > ALW 18

ALW 18

Challenges for Toxic Comment Classification: An In-Depth Error Analysis

Abstract

Toxic comment classification has become an active research field with many recently proposed approaches. However, while these approaches address some of the task’s challenges others still remain unsolved and directions for further research are needed. To this end, we compare different approaches on a new, large comment dataset and propose an ensemble that outperforms all individual models. Further, we validate our findings on a second dataset. The results of the ensemble enable us to perform an extensive error analysis, which reveals open challenges for state-of- the-art methods and directions towards pending future research. These challenges include missing paradigmatic context and inconsistent dataset labels.

Full Paper

ALW18.pdf

Workshop Homepage

ALW 2018

BibTex Entry

@inproceedings{krestel-alw18, author = {van Aken, Betty and Risch, Julian and Krestel, Ralf and Löser, Alexander}, booktitle = {Proceedings of the 2nd Workshop on Abusive Language Online (co-located with EMNLP)}, month = {October 31st}, title = {Challenges for Toxic Comment Classification: An In-Depth Error Analysis}, year = {2018} }

« prev| top| next »

News

Watch our new MOOC in German about hate and fake in the Internet ("Trolle, Hass und Fake-News: Wie können wir das Internet retten?") on openHPI (link).

New Publication

Our work on Measuring and Comparing Dimensionality Reduction Algorithms for Robust Visualisation of Dynamic Text Collections will be presented at CHIIR 2021.

New Photos

I added some photos from my trip to Hildesheim.