Web-Based Technical Term Translation Pairs Mining for Patent Document Translation
This paper proposes a simple but powerful approach for obtaining technical term translation pairs in patent domain from Web automatically. First, several technical terms are used as seed queries and submitted to search engineering. Secondly, an extraction algorithm is proposed to extract some key word translation pairs from the returned web pages. Finally, a multi-feature based evaluation method is proposed to pick up those translation pairs that are true technical term translation pairs in patent domain. With this method, we obtain about 8,890,000 key word translation pairs which can be used to translate the technical terms in patent documents. And experimental results show that the precision of these translation pairs are more than 99%, and the coverage of these translation pairs for the technical terms in patent documents are more than 84%.
Term translation patent document translation web-based key word extraction key word selection machine translation
Feiliang REN Jingbo ZHU Huizhen WANG
Northeastern University Shenyang, China
国际会议
北京
英文
1-8
2010-08-21(万方平台首次上网日期,不代表论文的发表时间)