Distant BI-Gram model, collocation, and their applications in post-processing for Chinese character recognition

Proceedings of 2002 International Conference on Machine Learning and Cybernetics(2002)

引用 0|浏览10
暂无评分
摘要
In this paper, we present a Distant BI-Gram model, which extended the regular BI-Gram model by considering the distance information and weight parameters, in order to describe the long-distance restrictions among the Chinese sentence. The extraction of the statistical information and weight parameters of this language model is discussed. Based on this work, the word combination strength and spread are employed to extract the recurrent word combinations, i.e. collocations. The Distant BI-Gram Model and collocation are applied to a statistic-based post-processing system for improving the recognition performance of Chinese character. The experimental results show that by employ these two language models, the post-processing system achieves a higher improvement performance.
更多
查看译文
关键词
Distant BI-Gram, collocation, post-processing, character, recognition
AI 理解论文
溯源树
样例
生成溯源树,研究论文发展脉络
Chat Paper
正在生成论文摘要