Structured Latent Factor Analysis for Large-scale Data: Identifiability, Estimability, and Their Implications
Structured Latent Factor Analysis for Large-scale Data: Identifiability, Estimability, and Their Implications
复制标题
DOI:
10.1080/01621459.2019.1635485
复制
发表时间:
2019-07-20
影响因子:
3.7
通讯作者:
Zhang, Siliang
中科院分区:
文献类型:
--
作者:
Chen, Yunxiao;Li, Xiaoou;Zhang, Siliang
Abstract–Latent factor models are widely used to measure unobserved latent traits in social and behavioral sciences, including psychology, education, and marketing. When used in a confirmatory manner, design information is incorporated as zero constraints on corresponding parameters, yielding structured (confirmatory) latent factor models. In this article, we study how such design information affects the identifiability and the estimation of a structured latent factor model. Insights are gained through both asymptotic and nonasymptotic analyses. Our asymptotic results are established under a regime where both the number of manifest variables and the sample size diverge, motivated by applications to large-scale data. Under this regime, we define the structural identifiability of the latent factors and establish necessary and sufficient conditions that ensure structural identifiability. In addition, we propose an estimator which is shown to be consistent and rate optimal when structural identifiability holds. Finally, a nonasymptotic error bound is derived for this estimator, through which the effect of design information is further quantified. Our results shed lights on the design of large-scale measurement in education and psychology and have important implications on measurement validity and reliability.