Asset Details
MbrlCatalogueTitleDetail
Do you wish to reserve the book?
A two-step ensemble learning for predicting protein hot spot residues from whole protein sequence
in
Amino acid sequence
/ Computer applications
/ Datasets
/ Ensemble learning
/ Model testing
/ Protein interaction
/ Proteins
/ Residues
/ Test sets
2022
Hey, we have placed the reservation for you!
By the way, why not check out events that you can attend while you pick your title.
You are currently in the queue to collect this book. You will be notified once it is your turn to collect the book.
Oops! Something went wrong.
Looks like we were not able to place the reservation. Kindly try again later.
Are you sure you want to remove the book from the shelf?
Oops! Something went wrong.
While trying to remove the title from your shelf something went wrong :( Kindly try again later!
Do you wish to request the book?
A two-step ensemble learning for predicting protein hot spot residues from whole protein sequence
in
Amino acid sequence
/ Computer applications
/ Datasets
/ Ensemble learning
/ Model testing
/ Protein interaction
/ Proteins
/ Residues
/ Test sets
2022
Please be aware that the book you have requested cannot be checked out. If you would like to checkout this book, you can reserve another copy
We have requested the book for you!
Your request is successful and it will be processed during the Library working hours. Please check the status of your request in My Requests.
Oops! Something went wrong.
Looks like we were not able to place your request. Kindly try again later.
A two-step ensemble learning for predicting protein hot spot residues from whole protein sequence
Journal Article
A two-step ensemble learning for predicting protein hot spot residues from whole protein sequence
2022
Request Book From Autostore
and Choose the Collection Method
Overview
Protein hot spot residues are functional sites in protein–protein interactions. Biological experimental methods are traditionally used to identify hot spot residues, which is laborious and time-consuming. Thus a variety of computational methods were widely used in recent years. Despite the success of computational methods in hot spot identification, most of them are impractical in reality because they can recognize hot spot residues only from known protein–protein interface residues. Therefore, identifying hot spots from whole protein sequence is a meaningful and interesting issue. However, it will bring extreme imbalance between positive and negative samples. Hot spot residues only account for about 1–2% of whole protein sequences. To address the issue, this paper proposes a two-step ensemble model for identifying hot spot residues from extremely unbalanced data set. The model is composed of 134 classifiers constructed by base KNN and SVM. Compared to the previous methods, our model yields good performance with an F1 score of 0.593 on the BID test set. Furthermore, to validate the robustness of our model, it was tested on other three independent test sets and also achieved good predictions. More importantly, the performance of our model tested on unbalanced data set is comparable with other methods tested on balanced hot spot data set.
Publisher
Springer Nature B.V
Subject
This website uses cookies to ensure you get the best experience on our website.