TY - JOUR
T1 - Mathematical symbol recognition with support vector machines
AU - Malon, Christopher
AU - Uchida, Seiichi
AU - Suzuki, Masakazu
N1 - Funding Information:
This work was supported by the Kyushu University 21st Century COE Program, “Development of Mathematics with High Functionality”.
PY - 2008/7/1
Y1 - 2008/7/1
N2 - Single-character recognition of mathematical symbols poses challenges from its two-dimensional pattern, the variety of similar symbols that must be recognized distinctly, the imbalance and paucity of training data available, and the impossibility of final verification through spell check. We investigate the use of support vector machines to improve the classification of InftyReader, a free system for the OCR of mathematical documents. First, we compare the performance of SVM kernels and feature definitions on pairs of letters that InftyReader usually confuses. Second, we describe a successful approach to multi-class classification with SVM, utilizing the ranking of alternatives within InftyReader's confusion clusters. The inclusion of our technique in InftyReader reduces its misrecognition rate by 41%.
AB - Single-character recognition of mathematical symbols poses challenges from its two-dimensional pattern, the variety of similar symbols that must be recognized distinctly, the imbalance and paucity of training data available, and the impossibility of final verification through spell check. We investigate the use of support vector machines to improve the classification of InftyReader, a free system for the OCR of mathematical documents. First, we compare the performance of SVM kernels and feature definitions on pairs of letters that InftyReader usually confuses. Second, we describe a successful approach to multi-class classification with SVM, utilizing the ranking of alternatives within InftyReader's confusion clusters. The inclusion of our technique in InftyReader reduces its misrecognition rate by 41%.
UR - http://www.scopus.com/inward/record.url?scp=43249087434&partnerID=8YFLogxK
UR - http://www.scopus.com/inward/citedby.url?scp=43249087434&partnerID=8YFLogxK
U2 - 10.1016/j.patrec.2008.02.005
DO - 10.1016/j.patrec.2008.02.005
M3 - Article
AN - SCOPUS:43249087434
SN - 0167-8655
VL - 29
SP - 1326
EP - 1332
JO - Pattern Recognition Letters
JF - Pattern Recognition Letters
IS - 9
ER -