login

Using N-gram and Word Network Features for Native Language Identification

Published 1 June 2013
Shibamouli Lahiri, Rada Mihalcea
Citations8

TL;DR

Experiments show that word networks have competitive performance against the baseline feature set in the Native Language Identification Shared Task, which is a promising result.

Abstract

We report on the performance of two different feature sets in the Native Language Identification Shared Task (Tetreault et al., 2013). Our feature sets were inspired by existing literature on native language identification and word networks. Exper-iments show that word networks have competitive performance against the baseline feature set, which is a promising result. We also present a discussion of feature analysis based on information gain, and an overview on the performance of different word net-work features in the Native Language Identification task.

Keywords

Computer Science