Towards the interpretation of time-varying regularization parameters in streaming penalized regression models

التفاصيل البيبلوغرافية
العنوان: Towards the interpretation of time-varying regularization parameters in streaming penalized regression models
المؤلفون: Zboňáková, Lenka, Monti, Ricardo Pio, Härdle, Wolfgang Karl
المصدر: Pattern Recognition Letters, Volume 125, 1 July 2019, Pages 542 to 548
سنة النشر: 2020
المجموعة: Statistics
مصطلحات موضوعية: Statistics - Methodology
الوصف: High-dimensional, streaming datasets are ubiquitous in modern applications. Examples range from finance and e-commerce to the study of biomedical and neuroimaging data. As a result, many novel algorithms have been proposed to address challenges posed by such datasets. In this work, we focus on the use of $\ell_1$ regularized linear models in the context of (possibly non-stationary) streaming data Recently, it has been noted that the choice of the regularization parameter is fundamental in such models and several methods have been proposed which iteratively tune such a parameter in a~time-varying manner; thereby allowing the underlying sparsity of estimated models to vary. Moreover, in many applications, inference on the regularization parameter may itself be of interest, as such a parameter is related to the underlying \textit{sparsity} of the model. However, in this work, we highlight and provide extensive empirical evidence regarding how various (often unrelated) statistical properties in the data can lead to changes in the regularization parameter. In particular, through various synthetic experiments, we demonstrate that changes in the regularization parameter may be driven by changes in the true underlying sparsity, signal-to-noise ratio or even model misspecification. The purpose of this letter is, therefore, to highlight and catalog various statistical properties which induce changes in the associated regularization parameter. We conclude by presenting two applications: one relating to financial data and another to neuroimaging data, where the aforementioned discussion is relevant.
نوع الوثيقة: Working Paper
DOI: 10.1016/j.patrec.2019.06.021
URL الوصول: http://arxiv.org/abs/2009.12113
رقم الأكسشن: edsarx.2009.12113
قاعدة البيانات: arXiv
الوصف
DOI:10.1016/j.patrec.2019.06.021