arrow
返回

Efficient Feature Selection via l2,0-norm Constrained Sparse Regression

delete2019-05-01
delete82
PRE
AI
T
Tianji Pang
聂
聂飞平 (Feiping Nie)
J
Junwei Han *
李学龙 封面图
李学龙 (Xuelong Li)
DOI:10.1109/TKDE.2018.2847685delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Sparse regression based feature selection method has been extensively investigated these years. However, because it has a non-convex constraint, i.e., l(2,0)-norm constraint, this problem is very hard to solve. In this paper, unlike most of the other methods which only solve its slack version by introducing sparsity regularization into objective function forcibly, a novel framework is proposed by us to solve the original l(2,0)-norm constrained sparse regression based feature selection problem. We transform our objective function into Linear Discriminant Analysis (LDA) by using a new label coding method, thus enabling our model to calculate the ratio of inter-class scatter to intra-class scatter of features which is the most widely used feature discrimination evaluation metric. According to that ratio, features can be selected by a simple sorting method. The projection gradient descent algorithm is introduced to further improve the performance of our algorithm by using the solution obtained before as its initial solution. This ensures the stability of this iterative algorithm. We prove that the proposed method can get the global optimal solution of this non-convex problem when all features are statistically independent. For the general case where features are statistically dependent, extensive experiments on six small sample size datasets and one large-scale dataset show that our algorithm has comparable or better classification capability comparing with other eight state-of-the-art feature selection methods by the SVM classifier. We also show that our algorithm can obtain a low loss value, which means the solution of our algorithm can get very close to this NP-hard problem's real solution. What is more, because we solve the original l(2,0)-norm constrained problem, we avoid the heavy work of tuning the regularization parameter because its meaning is explicit in our method, i.e., the number of selected features. At last, we evaluate the stability of our algorithm from two perspectives, i.e., the objective function values and the selected features, by experiments. From both perspectives, our algorithm shows satisfactory stability performance.
Keyword:
Feature selection
LDA
sparse regression
embedding
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

IEEE Transactions on Knowledge and Data Engineering 封面图
IEEE Transactions on Knowledge and Data Engineering
IF:
10.4
论文数:
6.8K
被引数:
3.2W

机构

N
Northwestern Polytechnical University
学者数:
4.6W
论文数: 3.7W
被引数: 5.3W
引用论文

引用论文

err分享
err收藏
The Effect of Phosphorus Exposure on Diesel Oxidation Catalysts—Part I: Activity Measurements, Elementary and Surface Analyses
err2015-08-25
err0
PREAI
errMarja Kärkkäinen; Tanja Kolli; Mari Honkanen; Olli Heikkinen; Mika Huuhtanen; Kauko Kallinen; Toivo Lepistö; Jouko Lahtinen; Minnamari Vippola; Riitta L. Keiski
err分享
err收藏
The electroneutrality approximation in electrochemistry电化学中的电中性近似
err2011-02-22
err0
PREAI
errEdmund J. F. Dickinson; Juan G. Limon-Petersen; Richard G. Compton
err分享
err收藏
err分享
err收藏
学者 查看更多内容