首页 | 本学科首页   官方微博 | 高级检索  
     检索      

基于向量空间模型的古汉语词义自动消歧研究
引用本文:常娥,张长秀,侯汉清,惠富平.基于向量空间模型的古汉语词义自动消歧研究[J].图书情报工作,2013,57(2):114-118.
作者姓名:常娥  张长秀  侯汉清  惠富平
作者单位:1. 东南大学图书馆; 2. 南京农业大学信息科技学院; 3. 南京农业大学人文学院
基金项目:国家社会科学基金项目“古籍整理与开发智能化技术研究”(项目编号:08ATQ002);高等学校博士学科点专项科研基金资助课题“古农书资料自动编纂及注释系统的设计与构建”(项目编号:20090097110033)研究成果之一
摘    要:借鉴现代汉语词义消歧的研究成果,提出一种改进的向量空间模型词义消歧方法,即在古汉语义项词语知识库的支持下,将待消歧多义词上下文与多义词的义项映射到向量空间模型中,完成语义消歧任务。以中国农业古籍全文数据库为统计语料,对10个典型古汉语多义词,共29个义项、1 836条待消歧上下文进行义项标注的实验,消歧平均正确率达到79.5%。

关 键 词:向量空间模型  词义消歧  古汉语  
收稿时间:2012-08-15

Automatic Word Sense Disambiguation of Ancient Chinese Based on Vector Space Model
Chang E,Zhang Changxiu,Hou Hanqing,Hui Fuping.Automatic Word Sense Disambiguation of Ancient Chinese Based on Vector Space Model[J].Library and Information Service,2013,57(2):114-118.
Authors:Chang E  Zhang Changxiu  Hou Hanqing  Hui Fuping
Institution:1. Southeast University Library, Nanjing 210096; 2. School of Information Science and Technology, Nanjing Agricultural University, Nanjing 210095; 3. School of Humanities and Social Sciences, Nanjing Agricultural University, Nanjing 210095
Abstract:How to annotate the meaning of words is an important research work on collation of Chinese ancient books. The manual interpretation is time-consuming and laborious. According to the word sense disambiguation of modern Chinese, an improved unsupervised disambiguation method of ancient Chinese is proposed based on the vector space model. In order to disambiguate the word sense, the knowledge repository of ancient Chinese polysemous words is build, and the contexts and the meanings of the polysemous words are mapped into the vector space model. This paper takes the full-text database of Chinese agricultural ancient books for statistics corpus, and conducts the experiment using 10 typical polysemous words of ancient Chinese which include 29 senses and 1836 contexts. The result shows that the average disambiguation accuracy achieves 79.5%.
Keywords:vector space model  semantic disambiguation  ancient Chinese  
本文献已被 CNKI 万方数据 等数据库收录!
点击此处可从《图书情报工作》浏览原始摘要信息
点击此处可从《图书情报工作》下载免费的PDF全文
设为首页 | 免责声明 | 关于勤云 | 加入收藏

Copyright©北京勤云科技发展有限公司  京ICP备09084417号