当前位置: 首页 > 文章 > Heuristics based semantic annotation of biodiversity documents in Chinese 数据与情报科学学报(英文) 2013,6 (2)
Position: Home > Articles > Heuristics based semantic annotation of biodiversity documents in Chinese Journal of Data and Information Science 2013,6 (2)

Heuristics based semantic annotation of biodiversity documents in Chinese

作  者:
Yufeng Yufeng;Duan;Zhenzhen;Wendy Ju;Hong Li;Cu
关键词:
documents;approach;based;heuristics;biodiversity;leadin
摘  要:
Purpose : To design an efficient high-performance algorithm for semantic annotation of biodiversity documents in Chinese. Design/methodology/approach : Data set consists of 1,000 randomly selected documents from Flora of China. Comparative evaluation of the proposed approach with the Na ve Bayes algorithm have been developed before for the same purpose. Findings : Experimental results show that the heuristics based algorithm outperformed the Na ve Bayes algorithm. The use of leading words helped improving the annotation performance while prioritizing rule application based on their weights had no significant impact on algorithm performance. Research limitations : The ICTCLAS was used to identify word boundaries off-shelf without optimatization for biodiversity domain. This may have not made the best use of the tool. Practical implications & Originality/value : The performance of heuristics based approach, enhanced by leading words analysis, reached an F value of 0.9216, which is sufficiently accurate for practical use.

相似文章

计量
文章访问数: 15
HTML全文浏览量: 0
PDF下载量: 0

所属期刊

推荐期刊