• Home
  • Search
  • Exploiting Multilabel Information for Noise-Resilient Feature Selection
  • Cite Icon10
  • https://doi.org/10.1145/3158675Copy DOI Icon

Exploiting Multilabel Information for Noise-Resilient Feature Selection

Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

In a conventional supervised learning paradigm, each data instance is associated with one single class label. Multilabel learning differs in the way that data instances may belong to multiple concepts simultaneously, which naturally appear in a variety of high impact domains, ranging from bioinformatics and information retrieval to multimedia analysis. It targets leveraging the multiple label information of data instances to build a predictive learning model that can classify unlabeled instances into one or multiple predefined target classes. In multilabel learning, even though each instance is associated with a rich set of class labels, the label information could be noisy and incomplete as the labeling process is both time consuming and labor expensive, leading to potential missing annotations or even erroneous annotations. The existence of noisy and missing labels could negatively affect the performance of underlying learning algorithms. More often than not, multilabeled data often has noisy, irrelevant, and redundant features of high dimensionality. The existence of these uninformative features may also deteriorate the predictive power of the learning model due to the curse of dimensionality. Feature selection, as an effective dimensionality reduction technique, has shown to be powerful in preparing high-dimensional data for numerous data mining and machine-learning tasks. However, a vast majority of existing multilabel feature selection algorithms either boil down to solving multiple single-labeled feature selection problems or directly make use of the imperfect labels to guide the selection of representative features. As a result, they may not be able to obtain discriminative features shared across multiple labels. In this article, to bridge the gap between a rich source of multilabel information and its blemish in practical usage, we propose a novel noise-resilient multilabel informed feature selection framework (MIFS) by exploiting the correlations among different labels. In particular, to reduce the negative effects of imperfect label information in obtaining label correlations, we decompose the multilabel information of data instances into a low-dimensional space and then employ the reduced label representation to guide the feature selection phase via a joint sparse regression framework. Empirical studies on both synthetic and real-world datasets demonstrate the effectiveness and efficiency of the proposed MIFS framework.

Similar Papers
  • Research Article
  • Citations39

Feature selection for multilabel classification with missing labels via multi-scale fusion fuzzy uncertainty measures

  • May 11, 2024
  • Pattern Recognition
  • Tengyu Yin +6
  • Conference Article
  • Citations77

Robust Unsupervised Feature Selection on Networked Data

  • Jun 30, 2016
  • Jundong Li +3
  • PDF
  • Research Article
  • Citations6

Multi-Label Learning via Feature and Label Space Dimension Reduction

  • Jan 01, 2020
  • IEEE Access
  • Jun Huang +4
  • Research Article
  • Citations48

Joint label-specific features and label correlation for multi-label learning with missing label

  • Jul 08, 2020
  • Applied Intelligence
  • Ziwei Cheng +1
  • Book Chapter
  • Citations4

Towards a Feature Selection for Multi-label Text Classification in Big Data

  • Jan 01, 2020
  • Houda Amazal +2
  • Research Article
  • Citations95

Robust and Discriminative Labeling for Multi-Label Active Learning Based on Maximum Correntropy Criterion.

  • Jan 10, 2017
  • IEEE Transactions on Image Processing
  • Bo Du +4
  • Research Article
  • Citations55

A bipartite matching-based feature selection for multi-label learning

  • Aug 11, 2020
  • International Journal of Machine Learning and Cybernetics
  • Amin Hashemi +2
  • Research Article
  • Citations53

Semi-supervised multi-label feature selection with adaptive structure learning and manifold learning

  • Jan 07, 2021
  • Knowledge-Based Systems
  • Sitao Lv +3
  • Conference Article
  • Citations23

Multi-label Learning with Highly Incomplete Data via Collaborative Embedding

  • Jul 19, 2018
  • Yufei Han +3
  • Conference Article
  • Citations88

Multi-label Feature Selection via Global Relevance and Redundancy Optimization

  • Jul 01, 2020
  • Jia Zhang +5
  • Research Article
  • Citations67

Feature selection based on label distribution and fuzzy mutual information

  • Jun 07, 2021
  • Information Sciences
  • Chuanzhen Xiong +3
  • PDF
  • Research Article
  • Citations10

Multi-label feature selection method based on dynamic weight

  • Jan 30, 2022
  • Soft Computing
  • Ping Zhang +4
  • Research Article

Zero-Shot Feature Selection via Transferring Supervised Knowledge

  • Apr 01, 2021
  • International Journal of Data Warehousing and Mining
  • Zheng Wang +4
  • Conference Article
  • Citations1

ML-LRC: Low-rank-constraint-based Multi-label Learning with Label Noise

  • May 05, 2020
  • Xiaoying Wang +3
  • Book Chapter

LIA: A Label-Independent Algorithm for Feature Selection for Supervised Learning

  • Jan 01, 2019
  • Gail Gilboa-Freedman +2
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.