Data pre-processing for Machine Learning in Python

Data pre-processing for Machine Learning in Python
item image
 Buy Now
Facebook Twitter Pinterest

Price: 69.99$

In this course, we are going to focus on pre-processing techniques for machine learning. Pre-processing is the set of manipulations that transform a raw dataset to make it used by a machine learning model. It is necessary for making our data suitable for some machine learning models, to reduce the dimensionality, to better identify the relevant data, and to increase model performance. It’s the most important part of a machine learning pipeline and it’s strongly able to affect the success of a project. In fact, if we don’t feed a machine learning model with the correctly shaped data, it won’t work at all. Sometimes, aspiring Data Scientists start studying neural networks and other complex models and forget to study how to manipulate a dataset in order to make it used by their algorithms. So, they fail in creating good models and only at the end they realize that good pre-processing would make them save a lot of time and increase the performance of their algorithms. So, handling pre-processing techniques is a very important skill. That’s why I have created an entire course that focuses only on data pre-processing. With this course, you are going to learn: Data cleaning Encoding of the categorical variables Transformation of the numerical features Scikit-learn Pipeline and Column Transformer objects Scaling of the numerical features Principal Component Analysis Filter-based feature selection Oversampling using SMOTEAll the examples will be given using Python programming language and its powerful scikit-learn library. The environment that will be used is Jupyter, which is a standard in the data science industry. All the sections of this course end with some practical exercises and the Jupyter notebooks are all downloadable.

Leave a Reply