Attending to characters in neural sequence labeling models

Rei, M; Crichton, GKO; Pyysalo, S

Attending to characters in neural sequence labeling models

Published version

Peer-reviewed

Repository URI

https://www.repository.cam.ac.uk/handle/1810/274252

Repository DOI

https://doi.org/10.17863/CAM.21366

Files

Published version (262.88 KB)

Type

Conference Object

Authors

Rei, M

Crichton, GKO

Pyysalo, S

Abstract

Sequence labeling architectures use word embeddings for capturing similarity, but suffer when handling previously unseen or rare words. We investigate character-level extensions to such models and propose a novel architecture for combining alternative word representations. By using an attention mechanism, the model is able to dynamically decide how much information to use from a word- or character-level component. We evaluated different architectures on a range of sequence labeling datasets, and character-level extensions were found to improve performance on every benchmark. In addition, the proposed attention-based architecture delivered the best results even with a smaller number of trainable parameters.

Journal Title

COLING 2016 - 26th International Conference on Computational Linguistics, Proceedings of COLING 2016: Technical Papers

Conference Name

The International Conference on Computational Linguistics (COLING)

Publisher DOI

https://doi.org/10.17863/CAM.21366

Rights

Attribution 4.0 International

Sponsorship

Cambridge Assessment (unknown)

Collections

Scholarly Works - Computer Science and Technology