Convolution smoothing and non-convex regularization for support vector machine in high dimensions

Kangning Wang; Junning Yang; Kemal Polat; Adi Alhudhaif; Xiaofei Sun

doi:10.1016/j.asoc.2024.111433

Convolution smoothing and non-convex regularization for support vector machine in high dimensions

Kangning Wang, Junning Yang, Kemal Polat, Adi Alhudhaif, Xiaofei Sun

Computer Sciences

Research output: Contribution to journal › Article › peer-review

5 Scopus citations

Abstract

The support vector machine (SVM) is a well-known statistical learning tool for binary classification. One serious drawback of SVM is that it can be adversely affected by redundant variables, and research has shown that variable selection is crucial and necessary for achieving good classification accuracy. Hence some SVM variable selection studies have been devoted, and they have an unified “empirical hinge loss plus sparse penalty” formulation. However, a noteworthy issue is the computational complexity of existing methods is high especially for large-scale problems, due to the non-smoothness of the hinge loss. To solve this issue, we first propose a convolution smoothing approach, which turns the non-smooth hinge loss into a smooth surrogate one, and they are asymptotically equivalent. Moreover, we construct computationally more efficient SVM variable selection procedure by implementing non-convex penalized convolution smooth hinge loss. In theory, we prove that the resulting variable selection possesses the oracle property when the number of predictors is diverging. Numerical experiments also confirm the good performance of the new method.

Original language	English
Article number	111433
Journal	Applied Soft Computing
Volume	155
DOIs	https://doi.org/10.1016/j.asoc.2024.111433
State	Published - Apr 2024

Keywords

Convolution-type smoothing
High dimensionality
Penalized learning
Support vector machine

Access to Document

10.1016/j.asoc.2024.111433

Cite this

@article{8c59edb5ff2a47bf9b8a2c14aac68f29,

title = "Convolution smoothing and non-convex regularization for support vector machine in high dimensions",

abstract = "The support vector machine (SVM) is a well-known statistical learning tool for binary classification. One serious drawback of SVM is that it can be adversely affected by redundant variables, and research has shown that variable selection is crucial and necessary for achieving good classification accuracy. Hence some SVM variable selection studies have been devoted, and they have an unified “empirical hinge loss plus sparse penalty” formulation. However, a noteworthy issue is the computational complexity of existing methods is high especially for large-scale problems, due to the non-smoothness of the hinge loss. To solve this issue, we first propose a convolution smoothing approach, which turns the non-smooth hinge loss into a smooth surrogate one, and they are asymptotically equivalent. Moreover, we construct computationally more efficient SVM variable selection procedure by implementing non-convex penalized convolution smooth hinge loss. In theory, we prove that the resulting variable selection possesses the oracle property when the number of predictors is diverging. Numerical experiments also confirm the good performance of the new method.",

keywords = "Convolution-type smoothing, High dimensionality, Penalized learning, Support vector machine",

author = "Kangning Wang and Junning Yang and Kemal Polat and Adi Alhudhaif and Xiaofei Sun",

note = "Publisher Copyright: {\textcopyright} 2024 Elsevier B.V.",

year = "2024",

month = apr,

doi = "10.1016/j.asoc.2024.111433",

language = "English",

volume = "155",

journal = "Applied Soft Computing",

issn = "1568-4946",

publisher = "Elsevier B.V.",

}

TY - JOUR

T1 - Convolution smoothing and non-convex regularization for support vector machine in high dimensions

AU - Wang, Kangning

AU - Yang, Junning

AU - Polat, Kemal

AU - Alhudhaif, Adi

AU - Sun, Xiaofei

PY - 2024/4

Y1 - 2024/4

N2 - The support vector machine (SVM) is a well-known statistical learning tool for binary classification. One serious drawback of SVM is that it can be adversely affected by redundant variables, and research has shown that variable selection is crucial and necessary for achieving good classification accuracy. Hence some SVM variable selection studies have been devoted, and they have an unified “empirical hinge loss plus sparse penalty” formulation. However, a noteworthy issue is the computational complexity of existing methods is high especially for large-scale problems, due to the non-smoothness of the hinge loss. To solve this issue, we first propose a convolution smoothing approach, which turns the non-smooth hinge loss into a smooth surrogate one, and they are asymptotically equivalent. Moreover, we construct computationally more efficient SVM variable selection procedure by implementing non-convex penalized convolution smooth hinge loss. In theory, we prove that the resulting variable selection possesses the oracle property when the number of predictors is diverging. Numerical experiments also confirm the good performance of the new method.

AB - The support vector machine (SVM) is a well-known statistical learning tool for binary classification. One serious drawback of SVM is that it can be adversely affected by redundant variables, and research has shown that variable selection is crucial and necessary for achieving good classification accuracy. Hence some SVM variable selection studies have been devoted, and they have an unified “empirical hinge loss plus sparse penalty” formulation. However, a noteworthy issue is the computational complexity of existing methods is high especially for large-scale problems, due to the non-smoothness of the hinge loss. To solve this issue, we first propose a convolution smoothing approach, which turns the non-smooth hinge loss into a smooth surrogate one, and they are asymptotically equivalent. Moreover, we construct computationally more efficient SVM variable selection procedure by implementing non-convex penalized convolution smooth hinge loss. In theory, we prove that the resulting variable selection possesses the oracle property when the number of predictors is diverging. Numerical experiments also confirm the good performance of the new method.

KW - Convolution-type smoothing

KW - High dimensionality

KW - Penalized learning

KW - Support vector machine

UR - http://www.scopus.com/inward/record.url?scp=85186633514&partnerID=8YFLogxK

U2 - 10.1016/j.asoc.2024.111433

DO - 10.1016/j.asoc.2024.111433

M3 - Article

AN - SCOPUS:85186633514

SN - 1568-4946

VL - 155

JO - Applied Soft Computing

JF - Applied Soft Computing

M1 - 111433

ER -

Convolution smoothing and non-convex regularization for support vector machine in high dimensions

Abstract

Keywords

Access to Document

Other files and links

Fingerprint

Cite this