STRUCTURED, SPARSE REGRESSION WITH APPLICATION TO HIV DRUG RESISTANCE

Ann Appl Stat. 2011 Jun 1;5(2A):628-644. doi: 10.1214/10-AOAS428.

Abstract

We introduce a new version of forward stepwise regression. Our modification finds solutions to regression problems where the selected predictors appear in a structured pattern, with respect to a predefined distance measure over the candidate predictors. Our method is motivated by the problem of predicting HIV-1 drug resistance from protein sequences. We find that our method improves the interpretability of drug resistance while producing comparable predictive accuracy to standard methods. We also demonstrate our method in a simulation study and present some theoretical results and connection.