跳到主要导航 跳到搜索 跳到主要内容

Natural Language Processing Service Based on Stroke-Level Convolutional Networks for Chinese Text Classification

  • Hang Zhuang
  • , Chao Wang
  • , Changlong Li
  • , Qingfeng Wang
  • , Xuehai Zhou

科研成果: 书/报告/会议事项章节会议稿件同行评审

摘要

With the development of deep learning and artificial intelligence, more and more research apply neural networks to natural language processing tasks. However, while the majority of these research take English corpus as the dataset, few studies have been done using Chinese corpus. Meanwhile, Existing Chinese processing algorithms typically regard Chinese word or Chinese character as the basic unit but ignore the deeper information into the Chinese character. In Chinese linguistic, strokes are the basic unit of Chinese character who are similar to letters of the English word. Inspired by the recent success of deep learning at character-level, we delve deeper to Chinese stroke level for Chinese language processing and developed it into service for Chinese text classification. In this paper, we dig the basic feature of the strokes considering the similar Chinese character components and propose a new method to leverage Chinese stroke for learning the continuous representation of Chinese character and develop it into a service for Chinese text classification. We develop a dedicated neural architecture based on the convolutional neural network to effectively learn character embedding and apply it to Chinese word similarity judgment and Chinese text classification. Both experiments results show that the stroke level method is effective for Chinese language processing.

源语言英语
主期刊名Proceedings - 2017 IEEE 24th International Conference on Web Services, ICWS 2017
编辑Shiping Chen, Ilkay Altintas
出版商Institute of Electrical and Electronics Engineers Inc.
404-411
页数8
ISBN(电子版)9781538607527
DOI
出版状态已出版 - 7 9月 2017
已对外发布
活动24th IEEE International Conference on Web Services, ICWS 2017 - Honolulu, 美国
期限: 25 6月 201730 6月 2017

出版系列

姓名Proceedings - 2017 IEEE 24th International Conference on Web Services, ICWS 2017

会议

会议24th IEEE International Conference on Web Services, ICWS 2017
国家/地区美国
Honolulu
时期25/06/1730/06/17

指纹

探究 'Natural Language Processing Service Based on Stroke-Level Convolutional Networks for Chinese Text Classification' 的科研主题。它们共同构成独一无二的指纹。

引用此