Sparse Penalized Forward Selection for Support Vector Classification

Name: Sparse Penalized Forward Selection for Support Vector Classification
Creator: Taylor & Francis
Published: 2020-09-04 18:06:37
License: 暂无描述

DataCite Commons2020-09-04 更新2024-07-25 收录

下载链接：

https://tandf.figshare.com/articles/dataset/Sparse_Penalized_Forward_Selection_for_Support_Vector_Classification/1378868/1

下载链接

链接失效反馈

官方服务：

资源简介：

We propose a new binary classification and variable selection technique especially designed for high-dimensional predictors. Among many predictors, typically, only a small fraction of them have significant impact on prediction. In such a situation, more interpretable models with better prediction accuracy can be obtained by variable selection along with classification. By adding an ℓ<sub>1</sub>-type penalty to the loss function, common classification methods such as logistic regression or support vector machines (SVM) can perform variable selection. Existing penalized SVM methods all attempt to jointly solve all the parameters involved in the penalization problem altogether. When data dimension is very high, the joint optimization problem is very complex and involves a lot of memory allocation. In this article, we propose a new penalized forward search technique that can reduce high-dimensional optimization problems to one-dimensional optimization by iterating the selection steps. The new algorithm can be regarded as a forward selection version of the penalized SVM and its variants. The advantage of optimizing in one dimension is that the location of the optimum solution can be obtained with intelligent search by exploiting convexity and a piecewise linear or quadratic structure of the criterion function. In each step, the predictor that is most able to predict the outcome is chosen in the model. The search is then repeatedly used in an iterative fashion until convergence occurs. Comparison of our new classification rule with ℓ<sub>1</sub>-SVM and other common methods show very promising performance, in that the proposed method leads to much leaner models without compromising misclassification rates, particularly for high-dimensional predictors.

提供机构：

Taylor & Francis

创建时间：

2015-04-14

5,000+

优质数据集

54 个

任务类型

进入经典数据集