Mostrar el registro sencillo de la publicación
Semi-supervised learning for MALDI–TOF mass spectrometry data classification: an application in the salmon industry
dc.contributor.author | González, Camila | |
dc.contributor.author | Astudillo, César A. | |
dc.contributor.author | López-Cortés, Xaviera A. | |
dc.contributor.author | Maldonado, Sebastián | |
dc.date.accessioned | 2024-08-06T20:12:41Z | |
dc.date.available | 2024-08-06T20:12:41Z | |
dc.date.issued | 2023 | |
dc.identifier.uri | http://repositorio.ucm.cl/handle/ucm/5553 | |
dc.description.abstract | MALDI–TOF mass spectrometry (Matrix-Assisted Laser Desorption-Ionization (MALDI) and a Time-of-Flight detector (TOF) is a promising strategy for identifying patterns in data, establishing a relevant methodology for rapid and accurate microorganisms identification. However, this type of data is challenging to analyze due to its high complexity, and sometimes it is impossible to make a correct labeling. To address this problem, advanced data analysis techniques such as machine learning methods can be applied. In this work, we propose a novel approach using the semi-supervised paradigm for classifying MALDI–TOF mass spectrometry data. In addition, our study considers the use of labeled and unlabeled data to alleviate the issue of data labeling. Specifically, mass spectrometry data of healthy and infected salmon with the Piscirickettsia salmonis pathogen was analyzed. Our proposed algorithm based on self-training showed superior performance compared to traditional ML methods (NB, RF, SVM). Even considering a small percentage of labeled instances (25%), semi-supervised learning attains equilibrated performance across all metrics. Experimental results showed that self-training with a random forest classifier reached an accuracy of 0.9, sensitivity of 0.75, and specificity of 1. Furthermore, the feature selection allowed the identification of 15 potential biomarkers that define healthy and infected salmon profiles accurately. From a more general perspective, these results demonstrate the potential of the proposed semi-supervised learning methodology for classifying MALDI–TOF mass spectrometry data. | es_CL |
dc.language.iso | en | es_CL |
dc.rights | Atribución-NoComercial-SinDerivadas 3.0 Chile | * |
dc.rights.uri | http://creativecommons.org/licenses/by-nc-nd/3.0/cl/ | * |
dc.source | Neural Computing and Applications, 35(13), 9381-9391 | es_CL |
dc.subject | Semi-supervised learning | es_CL |
dc.subject | Mass spectrometry | es_CL |
dc.subject | Aquaculture | es_CL |
dc.title | Semi-supervised learning for MALDI–TOF mass spectrometry data classification: an application in the salmon industry | es_CL |
dc.type | Article | es_CL |
dc.ucm.facultad | Facultad de Ciencias de la Ingeniería | es_CL |
dc.ucm.indexacion | Scopus | es_CL |
dc.ucm.indexacion | Isi | es_CL |
dc.ucm.uri | https://springerlink.ucm.elogim.com/article/10.1007/s00521-023-08333-2 | es_CL |
dc.ucm.doi | doi.org/10.1007/s00521-023-08333-2 | es_CL |
Ficheros en la publicación
Ficheros | Tamaño | Formato | Ver |
---|---|---|---|
No hay ficheros asociados a esta publicación. |