» Articles » PMID: 38213739

Unveiling Protein Corona Composition: Predicting with Resampling Embedding and Machine Learning

Overview
Journal Regen Biomater
Date 2024 Jan 12
PMID 38213739
Authors
Affiliations
Soon will be listed here.
Abstract

Biomaterials with surface nanostructures effectively enhance protein secretion and stimulate tissue regeneration. When nanoparticles (NPs) enter the living system, they quickly interact with proteins in the body fluid, forming the protein corona (PC). The accurate prediction of the PC composition is critical for analyzing the osteoinductivity of biomaterials and guiding the reverse design of NPs. However, achieving accurate predictions remains a significant challenge. Although several machine learning (ML) models like Random Forest (RF) have been used for PC prediction, they often fail to consider the extreme values in the abundance region of PC absorption and struggle to improve accuracy due to the imbalanced data distribution. In this study, resampling embedding was introduced to resolve the issue of imbalanced distribution in PC data. Various ML models were evaluated, and RF model was finally used for prediction, and good correlation coefficient () and root-mean-square deviation (RMSE) values were obtained. Our ablation experiments demonstrated that the proposed method achieved an of 0.68, indicating an improvement of approximately 10%, and an RMSE of 0.90, representing a reduction of approximately 10%. Furthermore, through the verification of label-free quantification of four NPs: hydroxyapatite (HA), titanium dioxide (TiO), silicon dioxide (SiO) and silver (Ag), and we achieved a prediction performance with an value >0.70 using Random Oversampling. Additionally, the feature analysis revealed that the composition of the PC is most significantly influenced by the incubation plasma concentration, PDI and surface modification.

References
1.
Some D . Light-scattering-based analysis of biomolecular interactions. Biophys Rev. 2013; 5(2):147-158. PMC: 3641300. DOI: 10.1007/s12551-013-0107-1. View

2.
Pareek V, Bhargava A, Bhanot V, Gupta R, Jain N, Panwar J . Formation and Characterization of Protein Corona Around Nanoparticles: A Review. J Nanosci Nanotechnol. 2018; 18(10):6653-6670. DOI: 10.1166/jnn.2018.15766. View

3.
Duan Y, Coreas R, Liu Y, Bitounis D, Zhang Z, Parviz D . Prediction of protein corona on nanomaterials by machine learning using novel descriptors. NanoImpact. 2020; 17. PMC: 7043407. DOI: 10.1016/j.impact.2020.100207. View

4.
Poulsen K, Payne C . Concentration and composition of the protein corona as a function of incubation time and serum concentration: an automated approach to the protein corona. Anal Bioanal Chem. 2022; 414(24):7265-7275. DOI: 10.1007/s00216-022-04278-y. View

5.
Corbo C, Molinaro R, Tabatabaei M, Farokhzad O, Mahmoudi M . Personalized protein corona on nanoparticles and its clinical implications. Biomater Sci. 2017; 5(3):378-387. PMC: 5592724. DOI: 10.1039/c6bm00921b. View