.

DriveML: Self-Drive Machine Learning Projects

Abstract

Implementing some of the pillars of an automated machine learning pipeline such as (i) Automated data preparation, (ii) Feature engineering, (iii) Model building in classification context that includes techniques such as (a) Regularised regression [1], (b) Logistic regression [2], (c) Random Forest [3], (d) Decision tree [4] and (e) Extreme Gradient Boosting (xgboost) [5], and finally, (iv) Model explanation (using lift chart and partial dependency plots). Accomplishes the above tasks by running the function instead of writing lengthy R codes. Also provides some additional features such as generating missing at random (MAR) variables and automated exploratory data analysis. Moreover, function exports the model results with the required plots in an HTML vignette report format that follows the best practices of the industry and the academia.

The various functionalities of DriveML

Resources:

DriveML paper is now available in ArXiv

R CRAN

Help File

GitHub Page

Views: 586

Tags: dsc_ml, dsc_tagged

Comment

You need to be a member of Data Science Central to add comments!

Join Data Science Central

Comment by Capri Granville on February 25, 2021 at 10:58pm

thank goodness there's no linear regression

© 2021   TechTarget, Inc.   Powered by

Badges  |  Report an Issue  |  Privacy Policy  |  Terms of Service