会议专题

Web-Based Technical Term Translation Pairs Mining for Patent Document Translation

This paper proposes a simple but powerful approach for obtaining technical term translation pairs in patent domain from Web automatically. First, several technical terms are used as seed queries and submitted to search engineering. Secondly, an extraction algorithm is proposed to extract some key word translation pairs from the returned web pages. Finally, a multi-feature based evaluation method is proposed to pick up those translation pairs that are true technical term translation pairs in patent domain. With this method, we obtain about 8,890,000 key word translation pairs which can be used to translate the technical terms in patent documents. And experimental results show that the precision of these translation pairs are more than 99%, and the coverage of these translation pairs for the technical terms in patent documents are more than 84%.

Term translation patent document translation web-based key word extraction key word selection machine translation

Feiliang REN Jingbo ZHU Huizhen WANG

Northeastern University Shenyang, China

国际会议

The 6th International Conference on Natural Language Processing and Knowledge Engineering(第六届IEEE自然语言处理与知识工程国际会议 NLP-KE 2010)

北京

英文

1-8

2010-08-21(万方平台首次上网日期,不代表论文的发表时间)