Start Over

Multi-label Learning with Highly Incomplete Data via Collaborative Embedding

Authors :: Yun Shen
Xiangliang Zhang
Yufei Han
Guolei Sun
Source :: KDD
Publication Year :: 2018
Publisher :: ACM, 2018.
Abstract: Tremendous efforts have been dedicated to improving the effectiveness of multi-label learning with incomplete label assignments. Most of the current techniques assume that the input features of data instances are complete. Nevertheless, the co-occurrence of highly incomplete features and weak label assignments is a challenging and widely perceived issue in real-world multi-label learning applications due to a number of practical reasons including incomplete data collection, moderate labels from annotators, etc. Existing multi-label learning algorithms are not directly applicable when the observed features are highly incomplete. In this work, we attack this problem by proposing a weakly supervised multi-label learning approach, based on the idea of collaborative embedding. This approach provides a flexible framework to conduct efficient multi-label classification at both transductive and inductive mode by coupling the process of reconstructing missing features and weak label assignments in a joint optimisation framework. It is designed to collaboratively recover feature and label information, and extract the predictive association between the feature profile and the multi-label tag of the same data instance. Substantial experiments on public benchmark datasets and real security event data validate that our proposed method can provide distinctively more accurate transductive and inductive classification than other state-of-the-art algorithms.

Subjects :: Transduction (machine learning)
Computer science
business.industry
Process (engineering)
Association (object-oriented programming)
Multi label learning
02 engineering and technology
Machine learning
computer.software_genre
ComputingMethodologies_PATTERNRECOGNITION
020204 information systems
0202 electrical engineering, electronic engineering, information engineering
Feature (machine learning)
Benchmark (computing)
Embedding
020201 artificial intelligence & image processing
Artificial intelligence
business
computer

Details

Database :: OpenAIRE
Journal :: Proceedings of the 24th ACM SIGKDD International Conference on Knowledge Discovery & Data Mining
Accession number :: edsair.doi...........33790bf79c5d04888d03474da66f9b0c
Full Text :: https://doi.org/10.1145/3219819.3220038