Next Article in Journal
Synthetic Hydrograph Estimation for Ungauged Basins: Exploring the Role of Statistical Distributions
Previous Article in Journal
Expansions for the Conditional Density and Distribution of a Standard Estimate
 
 
Font Type:
Arial Georgia Verdana
Font Size:
Aa Aa Aa
Line Spacing:
Column Width:
Background:
Article

Model-Free Feature Screening Based on Data Aggregation for Ultra-High-Dimensional Longitudinal Data

1
School of Mathematics, Southwest Jiaotong University, Chengdu 611756, China
2
Office of Medical Information and Data, Medical Support Center, The General Hospital of Western Theater Command, PLA, Chengdu 610083, China
3
Department of Information, Medical Support Center, The General Hospital of Western Theater Command, PLA, Chengdu 610083, China
*
Author to whom correspondence should be addressed.
Stats 2025, 8(4), 99; https://doi.org/10.3390/stats8040099
Submission received: 3 September 2025 / Revised: 6 October 2025 / Accepted: 14 October 2025 / Published: 16 October 2025

Abstract

Ultra-high dimensional longitudinal data feature screening procedures are widely studied, but most require model assumptions. The screening performance of these methods may not be excellent if we specify an incorrect model. To resolve the above problem, a new model-free method is introduced where feature screening is performed by sample splitting and data aggregation. Distance correlation is used to measure the association at each time point separately, while longitudinal correlation is modeled by a specific cumulative distribution function to achieve efficiency. In addition, we extend this new method to handle situations where the predictors are correlated. Both methods possess excellent asymptotic properties and are capable of handling longitudinal data with unequal numbers of repeated measurements and unequal intervals between repeated measurement time points. Compared to other model-free methods, the two new methods are relatively insensitive to within-subject correlation, and they can help reduce the computational burden when applied to longitudinal data. Finally, we use some simulated and empirical examples to show that both new methods have better screening performance.
Keywords: model-free; longitudinal data; ultra-high-dimensional; feature screening; data aggregation model-free; longitudinal data; ultra-high-dimensional; feature screening; data aggregation

Share and Cite

MDPI and ACS Style

Chen, J.; Yang, X.; Dai, J.; Li, Y. Model-Free Feature Screening Based on Data Aggregation for Ultra-High-Dimensional Longitudinal Data. Stats 2025, 8, 99. https://doi.org/10.3390/stats8040099

AMA Style

Chen J, Yang X, Dai J, Li Y. Model-Free Feature Screening Based on Data Aggregation for Ultra-High-Dimensional Longitudinal Data. Stats. 2025; 8(4):99. https://doi.org/10.3390/stats8040099

Chicago/Turabian Style

Chen, Junfeng, Xiaoguang Yang, Jing Dai, and Yunming Li. 2025. "Model-Free Feature Screening Based on Data Aggregation for Ultra-High-Dimensional Longitudinal Data" Stats 8, no. 4: 99. https://doi.org/10.3390/stats8040099

APA Style

Chen, J., Yang, X., Dai, J., & Li, Y. (2025). Model-Free Feature Screening Based on Data Aggregation for Ultra-High-Dimensional Longitudinal Data. Stats, 8(4), 99. https://doi.org/10.3390/stats8040099

Article Metrics

Back to TopTop