Switch-LSTMs for Multi-Criteria Chinese Word Segmentation

Authors

  • Jingjing Gong Fudan University
  • Xinchi Chen Fudan University
  • Tao Gui Fudan University
  • Xipeng Qiu Fudan University

DOI:

https://doi.org/10.1609/aaai.v33i01.33016457

Abstract

Multi-criteria Chinese word segmentation is a promising but challenging task, which exploits several different segmentation criteria and mines their common underlying knowledge. In this paper, we propose a flexible multi-criteria learning for Chinese word segmentation. Usually, a segmentation criterion could be decomposed into multiple sub-criteria, which are shareable with other segmentation criteria. The process of word segmentation is a routing among these sub-criteria. From this perspective, we present Switch-LSTMs to segment words, which consist of several long short-term memory neural networks (LSTM), and a switcher to automatically switch the routing among these LSTMs. With these auto-switched LSTMs, our model provides a more flexible solution for multi-criteria CWS, which is also easy to transfer the learned knowledge to new criteria. Experiments show that our model obtains significant improvements on eight corpora with heterogeneous segmentation criteria, compared to the previous method and single-criterion learning.

Downloads

Published

2019-07-17

How to Cite

Gong, J., Chen, X., Gui, T., & Qiu, X. (2019). Switch-LSTMs for Multi-Criteria Chinese Word Segmentation. Proceedings of the AAAI Conference on Artificial Intelligence, 33(01), 6457-6464. https://doi.org/10.1609/aaai.v33i01.33016457

Issue

Section

AAAI Technical Track: Natural Language Processing