全部 标题 作者
关键词 摘要

OALib Journal期刊
ISSN: 2333-9721
费用:99美元

查看量下载量

相关文章

更多...
-  2015 

版面相似中文表单的分类方法研究
A Study on Classification of Forms with Similar Layout

Keywords: 表单分类,距离度量,权重计算,表单分类,距离度量,权重计算
form classification
,distance metric,weight calculation,form classification,distance metric,weight calculation

Full-Text   Cite this paper   Add to My Lib

Abstract:

摘要 针对具有相似版面的中文表单, 提出一种简单有效的基于距离度量的表单分类方法, 该方法对表单的用户填写信息、布局信息和位置偏移分别进行距离度量, 并通过3种权重有效地降低用户填写信息的随机性、版面相似表单的布局一致性和位置抖动性对表单分类的影响。实验表明, 所提方法在多个中文表单图像库上的分类准确率达到90%以上, 比目前最新的表单分类方法有明显提高。
Abstract The authors propose a simple but effective distance based method to identify forms with similar layouts by measuring the user filled-in data, preprinted data and dithering data. The proposed method utilizes three kinds of weight components to mitigate the impact of randomness of user filled-in data, consistency of similar layouts and position dithering respectively. Experimental results show that the proposed method can achieve more than 90% classification accuracy on a series of data sets, which is significantly better than the results of the state-of-the-art method.

Full-Text

Contact Us

service@oalib.com

QQ:3279437679

WhatsApp +8615387084133