Article ; Online: A Novel Feature Selection Method for Uncertain Features: An Application to the Prediction of Pro-/Anti-Longevity Genes.
IEEE/ACM transactions on computational biology and bioinformatics
2021 Volume 18, Issue 6, Page(s) 2230–2238
Abstract: Understanding the ageing process is a very challenging problem for biologists. To help in this task, there has been a growing use of classification methods (from machine learning) to learn models that predict whether a gene influences the process of ... ...
Abstract | Understanding the ageing process is a very challenging problem for biologists. To help in this task, there has been a growing use of classification methods (from machine learning) to learn models that predict whether a gene influences the process of ageing or promotes longevity. One type of predictive feature often used for learning such classification models is Protein-Protein Interaction (PPI) features. One important property of PPI features is their uncertainty, i.e., a given feature (PPI annotation) is often associated with a confidence score, which is usually ignored by conventional classification methods. Hence, we propose the Lazy Feature Selection for Uncertain Features (LFSUF) method, which is tailored for coping with the uncertainty in PPI confidence scores. In addition, following the lazy learning paradigm, LFSUF selects features for each instance to be classified, making the feature selection process more flexible. We show that our LFSUF method achieves better predictive accuracy when compared to other feature selection methods that either do not explicitly take PPI confidence scores into account or deal with uncertainty globally rather than using a per-instance approach. Also, we interpret the results of the classification process using the features selected by LFSUF, showing that the number of selected features is significantly reduced, assisting the interpretability of the results. The datasets used in the experiments and the program code of the LFSUF method are freely available on the web at http://github.com/pablonsilva/FSforUncertainFeatureSpaces. |
---|---|
MeSH term(s) | Aging/genetics ; Algorithms ; Animals ; Computational Biology/methods ; Drosophila melanogaster/genetics ; Genome, Human/genetics ; Humans ; Machine Learning ; Mice ; Protein Interaction Maps/genetics ; Uncertainty ; Yeasts/genetics |
Language | English |
Publishing date | 2021-12-08 |
Publishing country | United States |
Document type | Journal Article ; Research Support, Non-U.S. Gov't |
ISSN | 1557-9964 |
ISSN (online) | 1557-9964 |
DOI | 10.1109/TCBB.2020.2988450 |
Database | MEDical Literature Analysis and Retrieval System OnLINE |
More links
Kategorien
Order via subito
This service is chargeable due to the Delivery terms set by subito. Orders including an article and supplementary material will be classified as separate orders. In these cases, fees will be demanded for each order.
Inter-library loan at ZB MED
Your chosen title can be delivered directly to ZB MED Cologne location if you are registered as a user at ZB MED Cologne.