https://scholars.lib.ntu.edu.tw/handle/123456789/413125
標題: | Sentence rephrasing for parsing sentences with OOV words | 作者: | Huang H.-H. Chen H.-Y. Yu C.-S. Chen H.-H. Lee P.-C. Chen C.-H. |
關鍵字: | Dependency parsing;Named entity;Sentence rephrasing | 公開日期: | 2014 | 起(迄)頁: | 2859-2862 | 來源出版物: | 9th International Conference on Language Resources and Evaluation | 摘要: | This paper addresses the problems of out-of-vocabulary (OOV) words, named entities in particular, in dependency parsing. The OOV words, whose word forms are unknown to the learning-based parser, in a sentence may decrease the parsing performance. To deal with this problem, we propose a sentence rephrasing approach to replace each OOV word in a sentence with a popular word of the same named entity type in the training set, so that the knowledge of the word forms can be used for parsing. The highest-frequency-based rephrasing strategy and the information-retrieval-based rephrasing strategy are explored to select the word to replace, and the Chinese Treebank 6.0 (CTB6) corpus is adopted to evaluate the feasibility of the proposed sentence rephrasing strategies. Experimental results show that rephrasing some specific types of OOV words such as Corporation, Organization, and Competition increases the parsing performances. This methodology can be applied to domain adaptation to deal with OOV problems. |
URI: | https://scholars.lib.ntu.edu.tw/handle/123456789/413125 | ISBN: | 9782951740884 |
顯示於: | 資訊工程學系 |
在 IR 系統中的文件,除了特別指名其著作權條款之外,均受到著作權保護,並且保留所有的權利。