Exploring deep learning approaches for Urdu text classification in product manufacturing
Exploring deep learning approaches for Urdu text classification in product manufacturing
复制标题
DOI:
10.1080/17517575.2020.1755455
复制
发表时间:
2020-05-06
影响因子:
4.4
通讯作者:
Fayyaz, Muhammad
中科院分区:
文献类型:
--
作者:
Akhter, Muhammad Pervez;Jiangbin, Zheng;Fayyaz, Muhammad
From last decade, machine learning (ML) techniques have been used for Urdu text processing. Due to lack of language resources, potential of deep learning (DL) models have not been exploited yet for Urdu text document classification. A text document has more noise, redundant information, and large vocabulary than short text like tweets. This study is the systematic comparison of four well-known DL models. We also compare DL models with four ML models. We also explore the various text preprocessing techniques. Experimental results show that CNN outperforms the others. Further, single-layer architecture of LSTM and BiLSTM performs better than multiple-layers architecture.