Gradient-induced Model-free Variable Selection with Composite Quantile Regression

Statistica Sinica ◽

10.5705/ss.202016.0222 ◽

2018 ◽

Author(s):

Shaogao Lv ◽

Xin He ◽

Junhui Wang

Keyword(s):

Variable Selection ◽

Quantile Regression ◽

Free Variable ◽

Composite Quantile Regression ◽

Download Full-text

Robust communication-efficient distributed composite quantile regression and variable selection for massive data

Computational Statistics & Data Analysis ◽

10.1016/j.csda.2021.107262 ◽

2021 ◽

pp. 107262

Author(s):

Kangning Wang ◽

Shaomin Li ◽

Benle Zhang

Keyword(s):

Variable Selection ◽

Quantile Regression ◽

Massive Data ◽

Composite Quantile Regression ◽

Download Full-text

A model-free variable selection method for reducing the number of redundant variables

10.1080/02331888.2018.1515949 ◽

2018 ◽

Vol 52 (6) ◽

pp. 1212-1248

Author(s):

Anchao Song ◽

Tiefeng Ma ◽

Shaogao Lv ◽

Changsheng Lin

Keyword(s):

Variable Selection ◽

Free Variable ◽

Selection Method ◽

Variable Selection Method ◽

Download Full-text

On dual model-free variable selection with two groups of variables

Journal of Multivariate Analysis ◽

10.1016/j.jmva.2018.06.003 ◽

2018 ◽

Vol 167 ◽

pp. 366-377

Author(s):

Ahmad Alothman ◽

Yuexiao Dong ◽

Andreas Artemiou

Keyword(s):

Variable Selection ◽

Free Variable ◽

Download Full-text

Knockoff boosted tree for model-free variable selection

Bioinformatics ◽

10.1093/bioinformatics/btaa770 ◽

2020 ◽

Author(s):

Tao Jiang ◽

Yuanyuan Li ◽

Alison A Motsinger-Reif

Keyword(s):

Variable Selection ◽

Principal Component ◽

Free Variable ◽

Supplementary Information ◽

Type I ◽

Test Statistics ◽

Linear Regression Models ◽

Tree Models ◽

Abstract Motivation The recently proposed knockoff filter is a general framework for controlling the false discovery rate (FDR) when performing variable selection. This powerful new approach generates a ‘knockoff’ of each variable tested for exact FDR control. Imitation variables that mimic the correlation structure found within the original variables serve as negative controls for statistical inference. Current applications of knockoff methods use linear regression models and conduct variable selection only for variables existing in model functions. Here, we extend the use of knockoffs for machine learning with boosted trees, which are successful and widely used in problems where no prior knowledge of model function is required. However, currently available importance scores in tree models are insufficient for variable selection with FDR control. Results We propose a novel strategy for conducting variable selection without prior model topology knowledge using the knockoff method with boosted tree models. We extend the current knockoff method to model-free variable selection through the use of tree-based models. Additionally, we propose and evaluate two new sampling methods for generating knockoffs, namely the sparse covariance and principal component knockoff methods. We test and compare these methods with the original knockoff method regarding their ability to control type I errors and power. In simulation tests, we compare the properties and performance of importance test statistics of tree models. The results include different combinations of knockoffs and importance test statistics. We consider scenarios that include main-effect, interaction, exponential and second-order models while assuming the true model structures are unknown. We apply our algorithm for tumor purity estimation and tumor classification using Cancer Genome Atlas (TCGA) gene expression data. Our results show improved discrimination between difficult-to-discriminate cancer types. Availability and implementation The proposed algorithm is included in the KOBT package, which is available at https://cran.r-project.org/web/packages/KOBT/index.html. Supplementary information Supplementary data are available at Bioinformatics online.

Download Full-text

Estimation and variable selection in single-index composite quantile regression

Communications in Statistics - Simulation and Computation ◽

10.1080/03610918.2016.1222424 ◽

2017 ◽

Vol 46 (9) ◽

pp. 7022-7039 ◽

Author(s):

Huilan Liu ◽

Hu Yang

Keyword(s):

Variable Selection ◽

Quantile Regression ◽

Composite Quantile Regression ◽

Download Full-text

Model-free variable selection

Journal of the Royal Statistical Society Series B (Statistical Methodology) ◽

10.1111/j.1467-9868.2005.00502.x ◽

2005 ◽

Vol 67 (2) ◽

pp. 285-299 ◽

Author(s):

Lexin Li ◽

R. Dennis Cook ◽

Christopher J. Nachtsheim

Keyword(s):

Variable Selection ◽

Free Variable ◽

Download Full-text

Multiple Loci Mapping via Model-free Variable Selection

10.1111/j.1541-0420.2011.01650.x ◽

2011 ◽

Vol 68 (1) ◽

pp. 12-22

Author(s):

Wei Sun ◽

Lexin Li

Keyword(s):

Variable Selection ◽

Free Variable ◽

Download Full-text

Composite quantile regression and variable selection in single-index coefficient model

Journal of Statistical Planning and Inference ◽

10.1016/j.jspi.2016.04.003 ◽

2016 ◽

Vol 176 ◽

pp. 1-21 ◽

Author(s):

Riquan Zhang ◽

Yazhao Lv ◽

Weihua Zhao ◽

Jicai Liu

Keyword(s):

Variable Selection ◽

Quantile Regression ◽

Composite Quantile Regression ◽

Download Full-text

Variable selection via composite quantile regression with dependent errors

Statistica Neerlandica ◽

10.1111/stan.12035 ◽

2014 ◽

Vol 69 (1) ◽

pp. 1-20 ◽

Author(s):

Yanlin Tang ◽

Xinyuan Song ◽

Zhongyi Zhu

Keyword(s):

Variable Selection ◽

Quantile Regression ◽

Composite Quantile Regression

Download Full-text

Copula based composite quantile regression for longitudinal data and variable selection

Scientia Sinica Mathematica ◽

10.1360/n012018-00298 ◽

2020 ◽

Vol 50 (8) ◽

pp. 1097

Author(s):

Lin Lu ◽

Wang Kangning ◽

Li Shaomin

Keyword(s):

Variable Selection ◽

Longitudinal Data ◽

Quantile Regression ◽

Composite Quantile Regression

Download Full-text