全部 标题 作者
关键词 摘要

OALib Journal期刊
ISSN: 2333-9721
费用:99美元

查看量下载量

相关文章

更多...

FEATURES EXTRACTION ALGORITHM FROM SGML FOR CLASSIFICATION

Keywords: Preprocessing , text categorization , algorithm , .net

Full-Text   Cite this paper   Add to My Lib

Abstract:

The basic phases in text categorization include preprocessing features, extracting relevant features against the features in a database, and finally categorizing a set of documents into predefined categories. Most of the researches in text categorization are focusing more on the development of algorithms and computer techniques. An algorithm for pre-processing features is seem to be like a "black-box" and ignored by them. Thus, it is significant and worthwhile to develop an algorithm for preprocessing features and finally can be used by other beginners before going in depth in the field of text categorization. This research proposes an algorithm for preprocessing features with capability of Microsoft .NET framework technology. The actual implementation shows that, this algorithm can extract interested features from the standard corpus of collection and upload into a relational database.

Full-Text

Contact Us

service@oalib.com

QQ:3279437679

WhatsApp +8615387084133