基于加权词汇衔接的文档级机器翻译自动评价 Document-Level Automatic Machine Translation Evaluation Based on Weighted Lexical Cohesion期刊界 All Journals 搜尽天下杂志传播学术成果专业期刊搜索期刊信息化学术搜索

基于加权词汇衔接的文档级机器翻译自动评价

引用本文：	贡正仙,李良友. 基于加权词汇衔接的文档级机器翻译自动评价[J]. 北京大学学报(自然科学版), 2014, 50(1): 173

作者姓名：	贡正仙李良友

作者单位：	苏州大学计算机科学与技术学院, 苏州 215006;

基金项目：	863计划(2012AA011102);国家自然科学基金(61305088)资助

摘要：	在文档词汇衔接评价LC方法的基础上, 提出基于权重的LC, 即WLC, 该方法通过在文档词图上运行PageRank算法获得词汇权重。根据词性信息使得PageRank算法偏向特定的词汇, 并提出PWLC方法。实验表明, 在文档级别上, 所提出的两种方法与人工评价的相关度都优于LC; 融合两种方法后, BLEU和TER在文档级别上的评价性能有显著提高。
关键词：	词汇衔接文档级评价机器翻译自动评价 PageRank
收稿时间：	2013-06-18
Document-Level Automatic Machine Translation Evaluation Based on Weighted Lexical Cohesion

GONG Zhengxian,LI Liangyou. Document-Level Automatic Machine Translation Evaluation Based on Weighted Lexical Cohesion[J]. Acta Scientiarum Naturalium Universitatis Pekinensis, 2014, 50(1): 173

Authors:	GONG Zhengxian LI Liangyou

Affiliation:	School of Computer Science and Technology, Soochow University, Suzhou 215006;

Abstract:	Based on LC method, weighted LC (WLC) method is proposed, which assigns weights for words by PageRank algorithm running on word graph of documents. Furthermore, a new method named PWLC is also proposed, which biases PageRank algorithm to words with specific POS tags. The experiment results show that WLC and PWLC have higher Spearman correlation than LC at document-level evaluation. Combined with others metrics, such as BLEU and TER, the proposed metrics both show better performance of evaluation at document level.

Keywords:	lexical cohesion document-level evaluation machine translation automatic evaluation PageRank
本文献已被 CNKI 万方数据等数据库收录！
	点击此处可从《北京大学学报(自然科学版)》浏览原始摘要信息
	点击此处可从《北京大学学报(自然科学版)》下载全文

设为首页 | 免责声明 | 关于勤云 | 加入收藏