Chinese-Khmer Named Entity Equivalents Excavation Based on Feature Similarity in Comparable Corpus
XU Lu
YAN Xin
XIA Qing
ZHOU Feng
MO Yuanyuan
Abstract:Named entity translation equivalent has been playing a significant role in the processing of cross-language informa?tion. However limited by the corpora resource,few in-depth studies have been made on the extraction of the bilingual Chi?nese-Khmer named entity equivalents. Starting from the comparable corpus text,according to the type of entity characteristics and comparable corpus characteristics,the paper selects transliteration feature,translation feature,context feature of the bilingual Chi?nese-Khmer named entity equivalents and length feature. So a method based on multi-feature fusion is proposed to calculate the sim?ilarity to excavate the bilingual Chinese-Khmer named entity equivalents. The experiment shows this method has a good perfor?mance when the bilingual Chinese-Khmer named entity equivalents are acquired through the computation of feature similarity,turn?ing out that the method proposed in this paper is able to give better effect compared with the method using only a single feature.
Keywords:named entity equivalentsChinese-Khmer bilingualmulti-feature fusioncomparable corpustransliteration model
Publication Date:2017-01-01
Online Publishing Date:2025-08-15(First online date of this platform, not the publication date of the document)
Pages:5( 882-885,910 )
