聚类数据线性混合效应模型的选择

Selection of linear mixed‐effects models for clustered data

Scandinavian Journal of Statistics · 2022
被引 0
ABS 3

中文导读

针对聚类结构线性混合效应模型,提出条件广义信息准则(CGIC)选择随机效应,在聚类数固定或随样本量增大时均能实现渐近损失效率,适用于经济学等领域的聚类数据分析。

Abstract

Abstract We consider model selection for linear mixed‐effects models with clustered structure, where conditional Kullback–Leibler (CKL) loss is applied to measure the efficiency of the selection. We estimate the CKL loss by substituting the empirical best linear unbiased predictors (EBLUPs) into random effects with model parameters estimated by maximum likelihood. Although the BLUP approach is commonly used in predicting random effects and future observations, selecting random effects to achieve asymptotic loss efficiency concerning CKL loss is challenging and has not been well studied. In this paper, we propose addressing this difficulty using a conditional generalized information criterion (CGIC) with two tuning parameters. We further consider a challenging but practically relevant situation where the number, , of clusters does not go to infinity with the sample size. Hence the random‐effects variances are not consistently estimable. We show that via a novel decomposition of the CKL risk, the CGIC achieves consistency and asymptotic loss efficiency, whether is fixed or increases to infinity with the sample size. We also conduct numerical experiments to illustrate the theoretical findings.

计量经济学统计学应用数学模型选择