Natural Language Processing Service Based on Stroke-Level Convolutional Networks for Chinese Text Classification

Hang Zhuang, Chao Wang, Changlong Li, Qingfeng Wang, Xuehai Zhou

Research output: Chapter in Book/Report/Conference proceedingConference contributionpeer-review

27 Scopus citations

Abstract

With the development of deep learning and artificial intelligence, more and more research apply neural networks to natural language processing tasks. However, while the majority of these research take English corpus as the dataset, few studies have been done using Chinese corpus. Meanwhile, Existing Chinese processing algorithms typically regard Chinese word or Chinese character as the basic unit but ignore the deeper information into the Chinese character. In Chinese linguistic, strokes are the basic unit of Chinese character who are similar to letters of the English word. Inspired by the recent success of deep learning at character-level, we delve deeper to Chinese stroke level for Chinese language processing and developed it into service for Chinese text classification. In this paper, we dig the basic feature of the strokes considering the similar Chinese character components and propose a new method to leverage Chinese stroke for learning the continuous representation of Chinese character and develop it into a service for Chinese text classification. We develop a dedicated neural architecture based on the convolutional neural network to effectively learn character embedding and apply it to Chinese word similarity judgment and Chinese text classification. Both experiments results show that the stroke level method is effective for Chinese language processing.

Original languageEnglish
Title of host publicationProceedings - 2017 IEEE 24th International Conference on Web Services, ICWS 2017
EditorsShiping Chen, Ilkay Altintas
PublisherInstitute of Electrical and Electronics Engineers Inc.
Pages404-411
Number of pages8
ISBN (Electronic)9781538607527
DOIs
StatePublished - 7 Sep 2017
Externally publishedYes
Event24th IEEE International Conference on Web Services, ICWS 2017 - Honolulu, United States
Duration: 25 Jun 201730 Jun 2017

Publication series

NameProceedings - 2017 IEEE 24th International Conference on Web Services, ICWS 2017

Conference

Conference24th IEEE International Conference on Web Services, ICWS 2017
Country/TerritoryUnited States
CityHonolulu
Period25/06/1730/06/17

Fingerprint

Dive into the research topics of 'Natural Language Processing Service Based on Stroke-Level Convolutional Networks for Chinese Text Classification'. Together they form a unique fingerprint.

Cite this