Covering assisted intuitionistic fuzzy bi-selection technique for data reduction and its applications

Rajat Saini; Anoop Kumar Tiwari; Abhigyan Nath; Phool Singh; S. P. Maurya; Mohd Asif Shah

doi:10.1038/s41598-024-62099-8

Scientific Reports (Jun 2024)

Covering assisted intuitionistic fuzzy bi-selection technique for data reduction and its applications

Rajat Saini,
Anoop Kumar Tiwari,
Abhigyan Nath,
Phool Singh,
S. P. Maurya,
Mohd Asif Shah

Affiliations

Rajat Saini: Department of Mathematics, School of Basic Sciences, Central University of Haryana
Anoop Kumar Tiwari: Department of Computer Science and Information Technology, Central University of Haryana
Abhigyan Nath: Department of Biochemistry, Pt. Jawahar Lal Nehru Memorial Medical College
Phool Singh: Department of Mathematics (SoET), Central University of Haryana
S. P. Maurya: Department of Geophysics, Institute of Science, Banaras Hindu University
Mohd Asif Shah: Department of Economics, Kebri Dehar University

DOI: https://doi.org/10.1038/s41598-024-62099-8
Journal volume & issue: Vol. 14, no. 1
pp. 1 – 23

Abstract

Read online

Abstract The dimension and size of data is growing rapidly with the extensive applications of computer science and lab based engineering in daily life. Due to availability of vagueness, later uncertainty, redundancy, irrelevancy, and noise, which imposes concerns in building effective learning models. Fuzzy rough set and its extensions have been applied to deal with these issues by various data reduction approaches. However, construction of a model that can cope with all these issues simultaneously is always a challenging task. None of the studies till date has addressed all these issues simultaneously. This paper investigates a method based on the notions of intuitionistic fuzzy (IF) and rough sets to avoid these obstacles simultaneously by putting forward an interesting data reduction technique. To accomplish this task, firstly, a novel IF similarity relation is addressed. Secondly, we establish an IF rough set model on the basis of this similarity relation. Thirdly, an IF granular structure is presented by using the established similarity relation and the lower approximation. Next, the mathematical theorems are used to validate the proposed notions. Then, the importance-degree of the IF granules is employed for redundant size elimination. Further, significance-degree-preserved dimensionality reduction is discussed. Hence, simultaneous instance and feature selection for large volume of high-dimensional datasets can be performed to eliminate redundancy and irrelevancy in both dimension and size, where vagueness and later uncertainty are handled with rough and IF sets respectively, whilst noise is tackled with IF granular structure. Thereafter, a comprehensive experiment is carried out over the benchmark datasets to demonstrate the effectiveness of simultaneous feature and data point selection methods. Finally, our proposed methodology aided framework is discussed to enhance the regression performance for IC50 of Antiviral Peptides.

Published in Scientific Reports

ISSN: 2045-2322 (Online)
Publisher: Nature Portfolio
Country of publisher: United Kingdom
LCC subjects: Medicine; Science
Website: https://www.nature.com/srep/

About the journal

Abstract

Keywords