login

Multilingual Language Processing From Bytes

Published 1 January 2016Open access
Dan Gillick, Cliff Brunk, Oriol Vinyals, Amarnag Subramanya
Citations220
View PDF

TL;DR

An LSTM-based model that reads text as bytes and outputs span annotations of the form [start, length, label] where start positions, lengths, and labels are separate entries in the authors' vocabulary is described.

Abstract

Dan Gillick, Cliff Brunk, Oriol Vinyals, Amarnag Subramanya. Proceedings of the 2016 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies. 2016.

Keywords

Computer Science