Using Off-Line Features and Synthetic Data for On-Line Handwritten Math Symbol Recognition
Using Off-Line Features and Synthetic Data for On-Line Handwritten Math Symbol Recognition
复制标题
使用离线特征和合成数据进行在线手写数学符号识别
DOI:
10.1109/icfhr.2014.61
复制
发表时间:
2014
期刊:
影响因子:
--
通讯作者:
R. Zanibbi
中科院分区:
文献类型:
--
作者:
Kenny Davila;S. Ludi;R. Zanibbi
We present an approach for on-line recognition of handwritten math symbols using adaptations of off-line features and synthetic data generation. We compare the performance of our approach using four different classification methods: AdaBoost. M1 with C4.5 decision trees, Random Forests and Support-Vector Machines with linear and Gaussian kernels. Despite the fact that timing information can be extracted from on-line data, our feature set is based on shape description for greater tolerance to variations of the drawing process. Our main datasets come from the Competition on Recognition of Online Handwritten Mathematical Expressions (CROHME) 2012 and 2013. Class representation bias in CROHME datasets is mitigated by generating samples for underrepresented classes using an elastic distortion model. Our results show that generation of synthetic data for underrepresented classes might lead to improvements of the average per-class accuracy. We also tested our system using the Math Brush dataset achieving a top-1 accuracy of 89.87% which is comparable with the best results of other recently published approaches on the same dataset.