Temporal Dynamic Matrix Factorization for Missing Data Prediction in Large Scale Coevolving Time Series
Temporal Dynamic Matrix Factorization for Missing Data Prediction in Large Scale Coevolving Time Series
复制标题
用于大规模协同演化时间序列中缺失数据预测的时间动态矩阵分解
DOI:
10.1109/access.2016.2606242
复制
发表时间:
2016-09
期刊:
影响因子:
3.9
通讯作者:
Yufeng Chen
中科院分区:
文献类型:
--
作者:
Weiwei Shi;Yongxin Zhu;Philip S. Yu;Tian Huang;Chang Wang;Yishu Mao;Yufeng Chen
Data missing in collections of time series occurs frequently in practical applications and turns out to be a major menace to precise data analysis. However, most of the existing methods either might be infeasible or could be inefficient to predict the missing values in large-scale coevolving time series. Also, the evolving of time series needs to be handled properly to adapt to the temporal characteristic. Furthermore, more massive volume of data is generated in many areas than ever before. In this paper, we have taken up the challenge of missing data prediction in coevolving time series by employing temporal dynamic matrix factorization techniques. First, our approaches are optimally designed to largely utilize both the interior patterns of each time series and the information of time series across multiple sources to build an initial model. Based on the idea, we have imposed hybrid regularization terms to constrain the objective functions of matrix factorization. Then, temporal dynamic matrix factorization is proposed to effectively update the initial already trained models. In the process of dynamic matrix factorization, batch updating and fine-tuning strategies are also employed to build an effective and efficient model. Extensive experiments on real-world data sets and synthetic data set demonstrate that the proposed approaches can effectively improve the performance of missing data prediction. Even when the missing ratio reaches as high as 90%, our proposed methods still show low prediction errors. Dynamic performance demonstrates that the methods can obtain satisfactory effectiveness and efficiency. Furthermore, we have also demonstrated how to take advantage of the high processing power of Apache Spark to perform missing data prediction in large-scale coevolving time series.
登录
查看更多内容
DOI:
10.1007/s00034-015-0047-z
发表时间:
2015-04
期刊:
Circuits, Systems, and Signal Processing
影响因子:
--
作者:
Junyou Shi;Long Chen;Wei-Wei Cui-Wei
通讯作者:
Junyou Shi;Long Chen;Wei-Wei Cui-Wei
DOI:
10.1109/jstars.2015.2424683
发表时间:
2015-10-01
影响因子:
5.5
作者:
Rathore, Muhammad Mazhar Ullah;Paul, Anand;Ji, Wen
通讯作者:
Ji, Wen
DOI:
10.1007/bf02898786
发表时间:
1896-09
期刊:
American Potato Journal
影响因子:
--
作者:
K. Fernow
通讯作者:
K. Fernow
DOI:
10.1016/j.trc.2013.05.008
发表时间:
2013-09-01
影响因子:
8.3
作者:
Li, Li;Li, Yuebiao;Li, Zhiheng
通讯作者:
Li, Zhiheng
影响因子:
8.4
作者:
Fonollosa, Jordi;Sheik, Sadique;Marco, Santiago
通讯作者:
Marco, Santiago