RECENT DEVELOPMENTS IN SOFT COMPUTING BASED TECHNIQUES FOR FEATURE SELECTION AND DISEASE CLASSIFICATION
DOI:
https://doi.org/10.55766/sujst-2023-02-e01872Keywords:
Clinical Symptom, Covid-19 Prediction, Disease Classification, Feature Selection, Gene Expression, Soft ComputingAbstract
Computational prediction of diseases is vital in medical research that contributes to computer-aided diagnostics and helps doctors and medical practitioners in critical decision-making for various diseases such as bacterial and viral kinds of disease, including COVID-19 of the current pandemic situation. Feature selection techniques function as a preprocessing phase for classification and prediction algorithms. For disease prediction, these features may be the patient’s clinical profiles or genomic features such as gene expression profiles from microarray and read counts from RNA-Seq. The performance of a classifier depends primarily on the selected features. In addition, genomic features are too large in numbers, resulting in the curse of dimensionality problem. In the last few years, several feature selection algorithms have been developed to overcome the existing problems to get rid of eliminating chronic diseases, such as various cancers, Zika virus, Ebola virus, and the COVID-19 pandemic. In this review article, we systematically associate soft computing-based approaches for feature selection and disease prediction by applying three data types: patients’ clinical profiles, microarray gene expression profiles, and RNA-Seq sample profiles. According to related work, when the discussion took place, the percentage of medical data types highlighted through pictorial representation and the respective ratio of percentages mentioned were 52%, 27%, 9% and 12% for clinical symptoms, gene expression, MRI-Image and other data types such as signal or text-based utilized, respectively. We also highlight the significant challenges and future directions in this research domain.
References
Abbass, H.A. (2002). An evolutionary artificial neural networks approach for breast cancer diagnosis. Artificial intelligence in Medicine., 25(3):265-281. https://doi.org/10.1016/S0933-3657(02)00028-3
Abdel-Aal, R.E. (2005). GMDH-based feature ranking and selection for improved classification of medical data. Journal of Biomedical Informatics., 38(6):456-468. https://doi.org/10.1016/j.jbi.2005.03.003
Abdi, M.J., Hosseini, S.M., and Rezghi, M. (2012). A novel weighted support vector machine based on particle swarm optimization for gene selection and tumor classification. Computational and mathematical methods in medicine, 2012. https://doi.org/10.1155/2012/320698
Abeel, T., Helleputte, T., Van de Peer, Y., Dupont, P., and Saeys, Y. (2009). Robust biomarker identification for cancer diagnosis with ensemble feature selection methods. Bioinformatics., 26(3):392-398. https://doi.org/10.1093/bioinformatics/btp630
Ahmad, W.M.T.W., Ab Ghani, N.L., and Drus, S.M. (2018). Data mining techniques for disease risk prediction model: A systematic literature review. In International Conference of Reliable Information and Communication Technology. Springer, Cham. p. 40-46. https://doi.org/10.1007/978-3-319-99007-1_4
Ang, J.C., Mirzal, A., Haron, H., and Hamed, H.N.A. (2016). Supervised, unsupervised, and semi-supervised feature selection: a review on gene selection. IEEE/ACM transactions on computational biology and bioinformatics., 13(5):971-989. https://doi.org/10.1109/TCBB.2015.2478454
Avci, E. (2009). A new intelligent diagnosis system for the heart valve diseases by using genetic-SVM classifier. Expert Systems with Applications., 36(7):10,618-10,626. https://doi.org/10.1016/j.eswa.2009.02.053
Azar, A.T., and Hassanien, A.E. (2015). Dimensionality reduction of medical big data using neural-fuzzy classifier. Soft computing., 19(4):1,115-1,127. https://doi.org/10.1007/s00500-014-1327-4
Azar, A.T., El-Said, S.A., Balas, V.E., and Olariu, T. (2013). Linguistic hedges fuzzy feature selection for differential diagnosis of Erythemato-Squamous diseases. In Soft computing applications (pp. 487-500). Springer, Berlin, Heidelberg. https://doi.org/10.1007/978-3-642-33941-7_43
Bellazzi, R., and Zupan, B. (2008). Predictive data mining in clinical medicine: current issues and guidelines. International journal of medical informatics., 77(2):81-97. https://doi.org/10.1016/j.ijmedinf.2006.11.006
Beloufa, F., and Chikh, M.A. (2013). Design of fuzzy classifier for diabetes disease using Modified Artificial Bee Colony algorithm. Computer methods and programs in biomedicine., 112(1):92-103. https://doi.org/10.1016/j.cmpb.2013.07.009
Bolón-Canedo, V., Sánchez-Marono, N., Alonso-Betanzos, A., Benítez, J.M., and Herrera, F. (2014). A review of microarray datasets and applied feature selection methods. Information Sciences., 282:111-135. https://doi.org/10.1016/j.ins.2014.05.042
Bolón-Canedo, V., Sánchez-Maroño, N., and Alonso-Betanzos, A. (2015). Recent advances and emerging challenges of feature selection in the context of big data. Knowledge-Based Systems., 86:33-45. https://doi.org/10.1016/j.knosys.2015.05.014
Buturovic, L.J. (2006). PCP-Pattern Classification Program, version 2.2 User’s Guide.
Çalişir, D., and Dogantekin, E. (2011). A new intelligent hepatitis diagnosis system: PCA-LSSVM. Expert Systems with Applications., 38(8):10,705-10,708. https://doi.org/10.1016/j.eswa.2011.01.014
Chaves, R., Ramírez, J., Górriz, J.M., Puntonet, C.G., and Alzheimer’s Disease Neuroimaging Initiative. (2012). Association rule-based feature selection method for Alzheimer’s disease diagnosis. Expert Systems with Applications., 39(14):11,766-11,774. https://doi.org/10.1016/j.eswa.2012.04.075
Chen, H.L., Yang, B., Wang, G., Wang, S.J., Liu, J., and Liu, D. Y. (2012). Support vector machine based diagnostic system for breast cancer using swarm intelligence. Journal of medical systems., 36(4):2,505-2,519. https://doi.org/10.1007/s10916-011-9723-0
Chiang, J.H., and Ho, S.H. (2008). A combination of rough-based feature selection and RBF neural network for classification using gene expression data. IEEE transactions on nanobioscience., 7(1):91-99. https://doi.org/10.1109/TNB.2008.2000142
Cho, S.B., and Won, H.H. (2003). Machine learning in DNA microarray analysis for cancer classification. In Proceedings of the First Asia-Pacific Bioinformatics Conference on Bioinformatics., 19:189-198
Das, P., Roychudhury, S., and Tripathy, S. (2018). sigFeature: Significant feature selection using SVM-RFE and t-statistic. https://doi.org/10.24870/cjb.2017-a22
De Falco, I. (2013). Differential Evolution for automatic rule extraction from medical databases. Applied soft computing., 13(2):1,265-1,283. https://doi.org/10.1016/j.asoc.2012.10.022
Determan Jr, C.E., and Determan Jr, M.C.E. (2017). Package ‘OmicsMarkeR’.
Dey, C., Bose, R., Ghosh, K.K., Malakar, S., and Sarkar, R. (2021). LAGOA: Learning automata-based grasshopper optimisation algorithm for feature selection in disease datasets. Journal of Ambient Intelligence and Humanized Computing, p. 1-20. https://doi.org/10.1007/s12652-021-03155-3
Divya, R., and Kumari, R.S.S. (2021). Genetic algorithm with logistic regression feature selection for Alzheimer’s disease classification. Neural Computing and Applications, p. 1-10. https://doi.org/10.1007/s00521-020-05596-x
El-Kenawy, E.S.M., Ibrahim, A., Mirjalili, S., Eid, M.M., and Hussein, S.E. (2020). Novel feature selection and voting classifier algorithms for COVID-19 classification in CT images. IEEE Access, 8, 179317-179335. https://doi.org/10.1109/ACCESS.2020.3028012
Fazlic, L.B., Avdagic, K., and Omanovic, S. (2015). GA-ANFIS Expert System Prototype for Prediction of Dermatological Diseases. In MIE (pp. 622-626).
Florescu, D., and Kossmann, D. (2009). Rethinking cost and performance of database systems. ACM Sigmod Record., 38(1):43-48. https://doi.org/10.1145/1558334.1558339
García-Nieto, J., Alba, E., Jourdan, L., and Talbi, E. (2009). Sensitivity and specificity based multiobjective approach for feature selection: Application to cancer diagnosis. Information Processing Letters., 109(16):887-896. https://doi.org/10.1016/j.ipl.2009.03.029
Geeitha, S., and Thangamani, M. (2018). Incorporating EBO-HSIC with SVM for Gene Selection Associated with Cervical Cancer Classification. Journal of medical systems., 42(11):225. https://doi.org/10.1007/s10916-018-1092-5
Glez-Peña, D., Álvarez, R., Díaz, F., and Fdez-Riverola, F. (2009). DFP: A Bioconductor package for fuzzy profile identification and gene reduction of microarray data. BMC Bioinformatics., 10(1):1-8. https://doi.org/10.1186/1471-2105-10-37
Gould, J., Getz, G., Monti, S., Reich, M., and Mesirov, J.P. (2006). Comparative gene marker selection suite. Bioinformatics., 22(15):1,924-1,925. https://doi.org/10.1093/bioinformatics/btl196
Gupta, D., Julka, A., Jain, S., Aggarwal, T., Khanna, A., Arunkumar, N., and de Albuquerque, V.H.C. (2018). Optimised Cuttlefish Algorithm for Diagnosis of Parkinson’s Disease. Cognitive Systems Research. https://doi.org/10.1016/j.cogsys.2018.06.006
Guyon, I., and Elisseeff, A. (2003). An introduction to variable and feature selection. Journal of machine learning research., 3(Mar), 1,157-1,182.
Hall, M., Frank, E., Holmes, G., Pfahringer, B., Reutemann, P., and Witten, I.H. (2009). The WEKA data mining software: an update. ACM SIGKDD explorations newsletter., 11(1):10-18. https://doi.org/10.1145/1656274.1656278
Hancer, E., Xue, B., Karaboga, D., and Zhang, M. (2015). A binary ABC algorithm based on advanced similarity scheme for feature selection. Applied Soft Computing., 36:334-348. https://doi.org/10.1016/j.asoc.2015.07.023
Hariharan, M., Polat, K., and Sindhu, R. (2014). A new hybrid intelligent system for accurate detection of Parkinson’s disease. Computer methods and programs in biomedicine., 113(3):904-913. https://doi.org/10.1016/j.cmpb.2014.01.004
Hira, Z.M., and Gillies, D.F. (2015). A review of feature selection and feature extraction methods applied on microarray data. Advances in bioinformatics, 2015. https://doi.org/10.1155/2015/198363
Hong, F., Breitling, R., McEntee, C.W., Wittner, B.S., Nemhauser, J.L., and Chory, J. (2006). RankProd: a bioconductor package for detecting differentially expressed genes in meta-analysis. Bioinformatics., 22(22):2,825-2,827. https://doi.org/10.1093/bioinformatics/btl476
Hong, J.H., and Cho, S.B. (2004). Lymphoma cancer classification using genetic programming with SNR features. In European Conference on Genetic Programming. Springer, Berlin, Heidelberg. P. 78-88. https://doi.org/10.1007/978-3-540-24650-3_8
Hsu, H.H., Hsieh, C.W., and Lu, M.D. (2011). Hybrid feature selection by combining filters and wrappers. Expert Systems with Applications., 38(7):8,144-8,150. https://doi.org/10.1016/j.eswa.2010.12.156
Huang, C., Huang, X., Fang, Y., Xu, J., Qu, Y., Zhai, P., ... and Li, J. (2020). Sample imbalance disease classification
model based on association rule feature selection. Pattern Recognition Letters., 133:280-286. https://doi.org/10.1016/j.patrec.2020.03.016
Illán, I.A., Górriz, J.M., López, M.M., Ramírez, J., Salas-Gonzalez, D., Segovia, F., and Puntonet, C.G. (2011). Computer-aided diagnosis of Alzheimer’s disease using component-based SVM. Applied Soft Computing., 11(2):2,376-2,382. https://doi.org/10.1016/j.asoc.2010.08.019
Inbarani, H.H., Azar, A.T., and Jothi, G. (2014). Supervised hybrid feature selection based on PSO and rough sets for medical diagnosis. Computer methods and programs in biomedicine., 113(1):175-185. https://doi.org/10.1016/j.cmpb.2013.10.007
Inbarani, H.H., Bagyamathi, M., and Azar, A.T. (2015). A novel hybrid feature selection method based on rough set and improved harmony search. Neural Computing and Applications., 26(8):1,859-1,880. https://doi.org/10.1007/s00521-015-1840-0
Inza, I., Sierra, B., Blanco, R., and Larrañaga, P. (2002). Gene selection by sequential search wrapper approaches in microarray cancer class prediction. Journal of Intelligent and Fuzzy Systems., 12(1):25-33.
Iqbal, N. and Kumar, P. (2019). I-NFG: An integrated neuro-fuzzy-genetic based soft computing techniques for feature selection and disease prediction using gene expression. Journal of Applied Computing., 4(1):1-8.
Iqbal, N., and Islam, M. (2019). Machine learning for dengue outbreak prediction: A performance evaluation of different prominent classifiers. Informatica., 43(3). https://doi.org/10.31449/inf.v43i3.1548
Iqbal, N., and Kumar, P. (2020). A framework for the RNA-Seq based classification and prediction of disease. In ICDSMLA 2019 (pp. 74-81). Springer, Singapore. https://doi.org/10.1007/978-981-15-1420-3_8
Iqbal, N., and Kumar, P. (2021, November). Coronavirus Disease Predictor: An RNA-Seq based pipeline for dimension reduction and prediction of COVID-19. In Journal of Physics: Conference Series., 2,089(1):012025. IOP Publishing. https://doi.org/10.1088/1742-6596/2089/1/012025
Iqbal, N., and Kumar, P. (2022). Integrated COVID-19 predictor: Differential expression analysis to reveal potential biomarkers and prediction of coronavirus using RNA-Seq profile data. Computers in Biology and Medicine., 147:105684. https://doi.org/10.1016/j.compbiomed.2022.105684
Jabeen, A., Ahmad, N. and Raza, K. (2018). Machine Learning-based State-of-the-art Methods for the Classification of RNA-Seq Data. In: Dey N., Ashour A., Borra S. (eds) Classification in BioApps. Lecture Notes in Computational Vision and Biomechanics, Springer., 26:133-172. https://doi.org/10.1007/978-3-319-65981-7_6
Kabir, M.M., Shahjahan, M., and Murase, K. (2012). A new hybrid ant colony optimization algorithm for feature selection. Expert Systems with Applications., 39(3):3,747-3,763. https://doi.org/10.1016/j.eswa.2011.09.073
Karabatak, M., and Ince, M.C. (2009). A new feature selection method based on association rules for diagnosis of erythemato-squamous diseases. Expert Systems with Applications., 36(10):12,500-12,505. https://doi.org/10.1016/j.eswa.2009.04.073
Kaya, Y., and Uyar, M. (2013). A hybrid decision support system based on rough set and extreme learning machine for diagnosis of hepatitis disease. Applied Soft Computing., 13(8):3,429-3,438. https://doi.org/10.1016/j.asoc.2013.03.008
Khan, F.N., Qazi, S., Tanveer, K., and Raza, K. (2017). A review on the antagonist Ebola: A prophylactic approach. Biomedicine and Pharmacotherapy., 96:1,513-1,526. https://doi.org/10.1016/j.biopha.2017.11.103
Khan, M.M., Mendes, A., and Chalup, S.K. (2018). Evolutionary Wavelet Neural Network ensembles for breast cancer and Parkinson’s disease prediction. PloS one., 13(2):e0192192. https://doi.org/10.1371/journal.pone.0192192
Kim, J., Lee, J., and Lee, Y. (2015). Data-mining-based coronary heart disease risk prediction model using fuzzy logic and decision tree. Healthcare informatics research., 21(3):167-174. https://doi.org/10.4258/hir.2015.21.3.167
Krishnapuram, B., Carin, L., and Hartemink, A. (2004). 1 Gene expression analysis: Joint feature selection and classifier design. Kernel Methods in Computational Biology., p. 299-317.
Lahsasna, A., Ainon, R.N., Zainuddin, R., and Bulgiba, A. (2012). Design of a fuzzy-based decision support system for coronary heart disease diagnosis. Journal of medical systems., 36(5):3,293-3,306. https://doi.org/10.1007/s10916-012-9821-7
Li, J., and Liu, H. (2017). Challenges of feature selection for big data analytics. IEEE Intelligent Systems., 32(2):9-15. https://doi.org/10.1109/MIS.2017.38
Li, L., Umbach, D.M., Terry, P., and Taylor, J.A. (2004). Application of the GA/KNN method to SELDI proteomics data. Bioinformatics., 20(10):1,638-1,640. https://doi.org/10.1093/bioinformatics/bth098
Lin, J.H., and Haug, P.J. (2008). Exploiting missing clinical data in Bayesian network modeling for predicting medical problems. Journal of biomedical informatics., 41(1):1-14. https://doi.org/10.1016/j.jbi.2007.06.001
Lin, S.W., and Chen, S.C. (2009). PSOLDA: A particle swarm optimization approach for enhancing classification accuracy rate of linear discriminant analysis. Applied Soft Computing., 9(3):1,008-1,015. https://doi.org/10.1016/j.asoc.2009.01.001
Liu, B., Cui, Q., Jiang, T., and Ma, S. (2004a). A combinational feature selection and ensemble neural network method for classification of gene expression data. BMC bioinformatics., 5(1):1-12.
Liu, H., and Yu, L. (2005). Toward integrating feature selection algorithms for classification and clustering. IEEE Transactions on knowledge and data engineering., 17(4):491-502. https://doi.org/10.1109/TKDE.2005.66
Liu, H., Motoda, H., and Yu, L. (2004b). A selective sampling approach to active feature selection. Artificial Intelligence., 159(1-2):49-74. https://doi.org/10.1016/j.artint.2004.05.009
Liu, J., Chen, C., Liu, Z., Jermsittiparsert, K., and Ghadimi, N. (2020). An IGDT-based risk-involved optimal bidding strategy for hydrogen storage-based intelligent parking lot of electric vehicles. Journal of Energy Storage., 27:101057. https://doi.org/10.1016/j.est.2019.101057
Liu, L., and Liu, J. (2020). Reconstructing gene regulatory networks via memetic algorithm and LASSO based on recurrent neural networks. Soft Computing., 24(6):4,205-4,221. https://doi.org/10.1007/s00500-019-04185-y
Liu, X., and Fu, H. (2014). PSO-based support vector machine with Cuckoo search technique for clinical disease diagnoses. The Scientific World Journal, 2014. https://doi.org/10.1155/2014/548483
Liu, Y., and Zheng, Y.F. (2006). FS_SFS: A novel feature selection method for support vector machines. Pattern recognition., 39(7):1,333-1,345. https://doi.org/10.1016/j.patcog.2005.10.006
Luo, P., Tian, L.P., Ruan, J., and Wu, F. (2017). Disease gene prediction by integrating PPI networks, clinical RNA-Seq data and OMIM data. IEEE/ACM Transactions on Computational Biology and Bioinformatics.
Mehrpooya, M., Ghadimi, N., Marefati, M., and Ghorbanian, S.A. (2021). Numerical investigation of a new combined energy system includes parabolic dish solar collector, Stirling engine and thermoelectric device. International Journal of Energy Research. https://doi.org/10.1002/er.6891
Miller, J.F. (2019). Cartesian genetic programming: its status and future. Genetic Programming and Evolvable Machines, 1-40. https://doi.org/10.1007/s10710-019-09360-6
Molina, L.C., Belanche, L., and Nebot, À. (2002). Feature selection algorithms: A survey and experimental evaluation. In 2002 IEEE International Conference on Data Mining, 2002. P. 306-313. https://doi.org/10.1109/ICDM.2002.1183917
Nahato, K.B., Harichandran, K.N., and Arputharaj, K. (2015). Knowledge mining from clinical datasets using rough sets and backpropagation neural network. Computational and mathematical methods in medicine, 2015. https://doi.org/10.1155/2015/460189
Nilashi, M., Bin Ibrahim, O., Mardani, A., Ahani, A., and Jusoh, A. (2016). A soft computing approach for diabetes disease classification. Health Informatics Journal, 1460458216675500. https://doi.org/10.1177/1460458216675500
Ozcift, A. (2012). SVM feature selection based rotation forest ensemble classifiers to improve computer-aided diagnosis of Parkinson disease. Journal of medical systems., 36(4):2,141-2,147. https://doi.org/10.1007/s10916-011-9678-1
Özçift, A., and Gülten, A. (2013). Genetic algorithm wrapped Bayesian network feature selection applied to differential diagnosis of erythematous-squamous diseases. Digital Signal Processing., 23(1):230-237. https://doi.org/10.1016/j.dsp.2012.07.008
Pahikkala, T., Okser, S., Airola, A., Salakoski, T., and Aittokallio, T. (2012). Wrapper-based selection of genetic features in genome-wide association studies through fast matrix operations. Algorithms for Molecular Biology., 7(1):11. https://doi.org/10.1186/1748-7188-7-11
Pavlenko, T. (2003). On feature selection, curse-of-dimensionality and error probability in discriminant analysis. Journal of Statistical Planning and Inference., 115(2):565-584. https://doi.org/10.1016/S0378-3758(02)00166-0
Peng, Y., Wu, Z., and Jiang, J. (2010). A novel feature selection approach for biomedical data classification. Journal of Biomedical Informatics., 43(1):15-23. https://doi.org/10.1016/j.jbi.2009.07.008
Plant, C., Teipel, S.J., Oswald, A., Böhm, C., Meindl, T., Mourao-Miranda, J., and Ewers, M. (2010). Automated detection of brain atrophy patterns based on MRI for the prediction of Alzheimer’s disease. Neuroimage., 50(1):162-174. https://doi.org/10.1016/j.neuroimage.2009.11.046
Polat, K., and Güneş, S. (2007). A hybrid approach to medical decision support systems: Combining feature selection, fuzzy weighted pre-processing and AIRS. Computer methods and programs in biomedicine., 88(2):164-174. https://doi.org/10.1016/j.cmpb.2007.07.013
Pollard, K.S., Dudoit, S., and van der Laan, M.J. (2005). Multiple testing procedures: the multtest package and applications to genomics. In Bioinformatics and computational biology solutions using R and bioconductor. Springer, New York, NY. P. 249-271. https://doi.org/10.1007/0-387-29362-0_15
Rawat, K., and Burse, K. (2013). A soft computing genetic-neuro fuzzy approach for data mining and its application to medical diagnosis. Int J Eng Adv Technol., 3:409-411.
Raymer, M.L., Punch, W.F., Goodman, E.D., Kuhn, L.A., and Jain, A.K. (2000). Dimensionality reduction using genetic algorithms. IEEE Transactions on evolutionary computation., 4(2):164-171. https://doi.org/10.1109/4235.850656
Raza, K. (2017). Formal concept analysis for knowledge discovery from biological data. International journal of data mining and bioinformatics., 18(4):281-300. https://doi.org/10.1504/IJDMB.2017.088138
Raza, K. (2019). Improving the Prediction Accuracy of Heart Disease with Ensemble Learning and Majority Voting Rule. In Advances in Ubiquitous Sensing Applications for Healthcare, U-Healthcare Monitoring Systems: Design and Applications, Academic Press, Elsevier, 179-196.138 https://doi.org/10.1016/B978-0-12-815370-3.00008-6
Raza, K. and Hasan, A.N. (2015). A Comprehensive Evaluation of Machine Learning Techniques for Cancer Class Prediction Based on Microarray Data. International Journal of Bioinformatics Research and Applications, Inderscience., 11(5):397-416. https://doi.org/10.1504/IJBRA.2015.071940
Raza, K., and Parveen, R. (2013). Soft computing approach for modeling genetic regulatory networks. Advances in computing and information technology, P. 1-11. https://doi.org/10.1007/978-3-642-31600-5_1
Roffo, G. (2016). Feature selection library (MATLAB toolbox). arXiv preprint arXiv:1607.01327.
Romalt, A.A., and Kumar, R.M.S. (2020). An analysis on feature selection methods, clustering and classification used in heart disease prediction-A machine learning approach. Journal of Critical Reviews., 7(6):138-142. https://doi.org/10.31838/jcr.07.06.27
Saeys, Y., Inza, I., and Larrañaga, P. (2007). A review of feature selection techniques in bioinformatics. Bioinformatics., 23(19):2,507-2,517. https://doi.org/10.1093/bioinformatics/btm344
Seera, M., and Lim, C.P. (2014). A hybrid intelligent system for medical data classification. Expert Systems with Applications., 41(5):2,239-2,249. https://doi.org/10.1016/j.eswa.2013.09.022
Senliol, B., Gulgezen, G., Yu, L., and Cataltepe, Z. (2008). Fast Correlation Based Filter (FCBF) with a different search strategy. In 2008 23rd international symposium on computer and information sciences. IEEE. P.1-4. https://doi.org/10.1109/ISCIS.2008.4717949
Shaban, W.M., Rabie, A.H., Saleh, A.I., and Abo-Elsoud, M.A. (2020). A new COVID-19 Patients Detection Strategy (CPDS) based on hybrid feature selection and enhanced KNN classifier. Knowledge-Based Systems, 205, 106270. https://doi.org/10.1016/j.knosys.2020.106270
Shankar, K., Lakshmanaprabu, S.K., Gupta, D., Maseleno, A., and De Albuquerque, V.H.C. (2020). Optimal feature-based multi-kernel SVM approach for thyroid disease classification. The journal of supercomputing., 76(2):1128-1143. https://doi.org/10.1007/s11227-018-2469-4
Sivagaminathan, R.K., and Ramakrishnan, S. (2007). A hybrid approach for feature subset selection using neural networks and ant colony optimization. Expert systems with applications., 33(1):49-60. https://doi.org/10.1016/j.eswa.2006.04.010
Smyth, G.K., Ritchie, M., Thorne, N., and Wettenhall, J. (2005). LIMMA: linear models for microarray data. In Bioinformatics and Computational Biology Solutions Using R and Bioconductor. Statistics for Biology and Health. https://doi.org/10.1007/0-387-29362-0_23
Sudha, M. (2017). Evolutionary and Neural Computing Based Decision Support System for Disease Diagnosis from Clinical Data Sets in Medical Practice. Journal of medical systems., 41(11):178. https://doi.org/10.1007/s10916-017-0823-3
Suk, H.I., Lee, S.W., Shen, D., and Alzheimer’s Disease Neuroimaging Initiative. (2016). Deep sparse multi-task learning for feature selection in Alzheimer’s disease diagnosis. Brain Structure and Function., 221(5):2,569-2,587. https://doi.org/10.1007/s00429-015-1059-y
Sun, L., Xu, J.C., Wang, W., and Yin, Y. (2016). Locally linear embedding and neighborhood rough set-based gene selection for gene expression data classification. Genetics and molecular research: GMR, 15(3). https://doi.org/10.4238/gmr.15038990
Thangavel, K., and Pethalakshmi, A. (2009). Dimensionality reduction based on rough set theory: A review. Applied Soft Computing., 9(1):1-12. https://doi.org/10.1016/j.asoc.2008.05.006
Thévenot, E.A. (2016). ropls: PCA, PLS (-DA) and OPLS (-DA) for multivariate analysis and feature selection of omics data.
Too, J., and Mirjalili, S. (2021). A hyper learning binary dragonfly algorithm for feature selection: A COVID-19 case study. Knowledge-Based Systems, 212, 106553. https://doi.org/10.1016/j.knosys.2020.106553
Trevino, V., and Falciani, F. (2006). GALGO: an R package for multivariate variable selection using genetic algorithms. Bioinformatics., 22(9):1,154-1,156. https://doi.org/10.1093/bioinformatics/btl074
Tusher, V.G., Tibshirani, R., and Chu, G. (2001). Significance analysis of microarrays applied to the ionizing radiation response. Proceedings of the National Academy of Sciences., 98(9):5,116-5,121. https://doi.org/10.1073/pnas.091062498
Vergara, J.R., and Estévez, P.A. (2014). A review of feature selection methods based on mutual information. Neural computing and applications., 24(1):175-186. https://doi.org/10.1007/s00521-013-1368-0
Wang, F., Xu, J., and Li, L. (2014, October). A novel rough set reduct algorithm to feature selection based on artificial fish swarm algorithm. In International Conference in Swarm Intelligence. Springer, Cham. P. 24-33. https://doi.org/10.1007/978-3-319-11897-0_4
Xing, E.P., and Karp, R.M. (2001). CLIFF: clustering of high-dimensional microarray data via iterative feature filtering using normalized cuts. Bioinformatics., 17(1):S306-S315. https://doi.org/10.1093/bioinformatics/17.suppl_1.S306
Xue, B., Zhang, M., Browne, W.N., and Yao, X. (2016). A survey on evolutionary computation approaches to feature selection. IEEE Transactions on Evolutionary Computation., 20(4):606-626. https://doi.org/10.1109/TEVC.2015.2504420
Ye, H., Jin, G., Fei, W., and Ghadimi, N. (2020). High step-up interleaved dc/dc converter with high efficiency. Energy sources, Part A: recovery, utilization, and environmental effects, P. 1-20. https://doi.org/10.1080/15567036.2020.1716111
Yin, L., Ge, Y., Xiao, K., Wang, X., and Quan, X. (2013). Feature selection for high-dimensional imbalanced data. Neurocomputing., 105:3-11. https://doi.org/10.1016/j.neucom.2012.04.039
Zeng, N., Qiu, H., Wang, Z., Liu, W., Zhang, H., and Li, Y. (2018). A new switching-delayed-PSO-based optimized SVM algorithm for diagnosis of Alzheimer’s disease. Neurocomputing. https://doi.org/10.1016/j.neucom.2018.09.001
Zhang, L., Chen, J., Gao, C., Liu, C., and Xu, K. (2018). An efficient model for auxiliary diagnosis of hepatocellular carcinoma based on gene expression programming. Medical and biological engineering and computing, P.1-9. https://doi.org/10.1007/s11517-018-1811-6
Zhao, M., Fu, C., Ji, L., Tang, K., and Zhou, M. (2011). Feature selection and parameter optimization for support vector machines: A new approach based on genetic algorithm with feature chromosomes. Expert Systems with Applications., 38(5):5,197-5,204. https://doi.org/10.1016/j.eswa.2010.10.041
Zomaya, A.Y. (2013). Stability of feature selection algorithms and ensemble feature selection methods in bioinformatics. Biological Knowledge Discovery Handbook: Preprocessing, Mining and Postprocessing of Biological Data, 23, 333. https://doi.org/10.1002/9781118617151.ch14








