4.6 Article

MRMD2.0: A Python Tool for Machine Learning with Feature Ranking and Reduction

Journal

CURRENT BIOINFORMATICS
Volume 15, Issue 10, Pages 1213-1221

Publisher

BENTHAM SCIENCE PUBL LTD
DOI: 10.2174/1574893615999200503030350

Keywords

Feature ranking; bioinformatics; machine learning; python; feature selection; dimension reduction

Funding

  1. National Key R&D Program of China [2018YFC0910405]
  2. Natural Science Foundation of China [61771331, 61922020]

Ask authors/readers for more resources

Aims: The study aims to find a way to reduce the dimensionality of the dataset. Background: Dimensionality reduction is the key issue of the machine learning process. It does not only improve the prediction performance but also could recommend the intrinsic features and help to explore the biological expression of the machine learning black box. Objective: A variety of feature selection algorithms are used to select data features to achieve dimensionality reduction. Methods: First, MRMD2.0 integrated 7 different popular feature ranking algorithms with PageRank strategy. Second, optimized dimensionality was detected with forward adding strategy. Result: We have achieved good results in our experiments. Conclusion: Several works have been tested with MRMD2.0. It showed well performance. Otherwise, it also can draw the performance curves according to the feature dimensionality. If users want to sacrifice accuracy for fewer features, they can select the dimensionality from the performance curves. Other: We developed friendly python tools together with the web server. The users could upload their csv, arff or libsvm format files. Then the webserver would help to rank features and find the optimized dimensionality.

Authors

I am an author on this paper
Click your name to claim this paper and add it to your profile.

Reviews

Primary Rating

4.6
Not enough ratings

Secondary Ratings

Novelty
-
Significance
-
Scientific rigor
-
Rate this paper

Recommended

No Data Available
No Data Available