arrow
Return

Weakly supervised multi-label feature selection based on shared subspace

delete2024-11-11
delete0
PRE
AI
S
Shi, Rongyi
A
Anhui Tan
S
Suwei Shi
J
Jin Wang
S
Shenming Gu
W
Wei-Zhi Wu *
DOI:10.1007/s13042-024-02426-7delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Multi-label feature selection (MLFS) improves classification accuracy and alleviates the curse of dimensionality by retaining relevant features and eliminating redundant and irrelevant features. However, the incompleteness of the label space not only makes the models difficult to learn the latent structure of the label space, but also leads to the unreliability of correlation between the labels and features. Therefore, it is essential and difficult how to discover the credible correlation information between the features and labels, so as to enable the effective selection of the key feature subset by the algorithms learning from multi-label data with missing labels. To more effectively excavate the implicit shared information within the feature matrix and the label matrix, we propose a novel MLFS method named WSMF which combines the feature matrix and the label matrix together to identify vital feature subsets in the absence of a large portion of labeled data. First, we utilize joint matrix factorization to uncover a low-dimensional shared mode between the feature matrix and the label matrix, thereby diminishing the effect of incomplete label information. Second, we employ non-negative matrix factorization (NMF) to enhance the interpretability of the succeeding feature selection process. Additionally, we employ the structural consistency assumption to retrieve the absent labels and incorporate l2,1\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$l_{2,1}$$\end{document}-norm to control the feature redundancy and irrelevant features. In the end, we conduct experiments on 14 datasets to distinctly explain the effectiveness of WSMF against other established algorithms.
Keywords:
Feature selection
Missing labels
Multi-label learning
Non-negative matrix factorization

Journal

International Journal of Machine Learning and Cybernetics cover
International Journal of Machine Learning and Cybernetics
IF:
2.7
Papers:
3.1K
Citations:
5.6K

Organization

Z
Zhejiang Ocean University
Scholars:
4.9K
Papers: 3.1K
Citations: 8.1K
H
huaqiao university
Scholars:
1.1W
Papers: 7.1K
Citations: 131
X
xiamen university
Scholars:
5.8W
Papers: 3.8W
Citations: 67
researcher View more organizations