login

Overview of the GermEval 2018 Shared Task on the Identification of Offensive Language

Publication Server of the Institute for German Language (Institute for German Language)Published 13 February 2019Open access
Michael Wiegand, Melanie Siegel, Josef Ruppenhofer
Citations237
View PDF

TL;DR

This pilot edition of the GermEval Shared Task on the Identification of Offensive Language deals with the classification of German tweets from Twitter and describes the process of extracting the raw-data for the data collection and the annotation schema.

Abstract

We present the pilot edition of the GermEval Shared Task on the Identification of Offensive Language. This shared task deals with the classification of German tweets from Twitter. It comprises two tasks, a coarse-grained binary classification task and a fine-grained multi-class classification task. The shared task had 20 participants submitting 51 runs for the coarse-grained task and 25 runs for the fine-grained task. Since this is a pilot task, we describe the process of extracting the raw-data for the data collection and the annotation schema. We evaluate the results of the systems submitted to the shared task. The shared task homepage can be found at https://projects.cai. fbi.h-da.de/iggsa/

Keywords

Computer ScienceSocial Sciences