作成者 |
|
|
|
本文言語 |
|
出版者 |
|
|
発行日 |
|
収録物名 |
|
巻 |
|
出版タイプ |
|
アクセス権 |
|
関連DOI |
|
|
関連URI |
|
|
関連情報 |
|
|
概要 |
Single–character recognition of mathematical symbols poses challenges from its twodimensional pattern, the variety of similar symbols that must be recognized distinctly, the imbalance and paucity of t...raining data available, and the impossibility of final verification through spell check. We investigate the use of support vector machines to improve the classification of InftyReader, a free system for the OCR of mathematical documents. First, we compare the performance of SVM kernels and feature definitions on pairs of letters that InftyReader usually confuses. Second, we describe a successful approach to multi–class classification with SVM, utilizing the ranking of alternatives within InftyReader’s confusion clusters. The inclusion of our technique in InftyReader reduces its misrecognition rate by 41%.続きを見る
|