Date of Original Version
Annals of Applied Statistics
© Institute of Mathematical Statistics, 2011 The version of record is available online at http://dx.doi.org/10.1214/10-AOAS428
Abstract or Description
We introduce a new version of forward stepwise regression. Our modification finds solutions to regression problems where the selected predictors appear in a structured pattern, with respect to a predefined distance measure over the candidate predictors. Our method is motivated by the problem of predicting HIV-1 drug resistance from protein sequences. We find that our method improves the interpretability of drug resistance while producing comparable predictive accuracy to standard methods. We also demonstrate our method in a simulation study and present some theoretical results and connections.
Annals of Applied Statistics, 5, 2A, 628-644.