0% found this document useful (0 votes)
3 views5 pages

Supervised Machine Learning

Supervised machine learning involves training models on labeled data to make predictions and improve accuracy over time. It can be applied to classification and regression problems, with common algorithms including Linear Regression, Decision Trees, and Support Vector Machines. The process involves collecting data, splitting it into training and testing sets, training the model, validating its performance, and deploying it for predictions on new data.

Uploaded by

anupamabhan
Copyright
© All Rights Reserved
We take content rights seriously. If you suspect this is your content, claim it here.
Available Formats
Download as DOCX, PDF, TXT or read online on Scribd
0% found this document useful (0 votes)
3 views5 pages

Supervised Machine Learning

Supervised machine learning involves training models on labeled data to make predictions and improve accuracy over time. It can be applied to classification and regression problems, with common algorithms including Linear Regression, Decision Trees, and Support Vector Machines. The process involves collecting data, splitting it into training and testing sets, training the model, validating its performance, and deploying it for predictions on new data.

Uploaded by

anupamabhan
Copyright
© All Rights Reserved
We take content rights seriously. If you suspect this is your content, claim it here.
Available Formats
Download as DOCX, PDF, TXT or read online on Scribd

Supervised Machine Learning

Last Updated : 12 Sep, 2025

Supervised learning is a type of machine learning where a model learns


from labelled data—meaning every input has a corresponding correct
output. The model makes predictions and compares them with the true
outputs, adjusting itself to reduce errors and improve accuracy over time.
The goal is to make accurate predictions on new, unseen data. For
example, a model trained on images of handwritten digits can recognise
new digits it has never seen before.

Supervised Machine Learning

Types of Supervised Learning in Machine Learning

Now, Supervised learning can be applied to two main types of problems:


 Classification: Where the output is a categorical variable (e.g.,
spam vs. non-spam emails, yes vs. no).

 Regression: Where the output is a continuous variable (e.g.,


predicting house prices, stock prices).

Types
of Supervised Learning

While training the model, data is usually split in the ratio of 80:20 i.e. 80%
as training data and the rest as testing data. In training data, we feed
input as well as output for 80% of data. The model learns from training
data only. We use different supervised learning algorithms (which we will
discuss in detail in the next section) to build our model. Let's first
understand the classification and regression data through the table below:

Sample
Both the above figures have labelled data set as follows:

Figure A: It is a dataset of a shopping store that is useful in predicting


whether a customer will purchase a particular product under consideration
or not based on his/her gender, age and salary.

 Input: Gender, Age, Salary

 Output: Purchased i.e. 0 or 1; 1 means yes the customer will


purchase and 0 means that the customer won't purchase it.

Figure B: It is a Meteorological dataset that serves the purpose of


predicting wind speed based on different parameters.

 Input: Dew Point, Temperature, Pressure, Relative Humidity, Wind


Direction

 Output: Wind Speed

Working of Supervised Machine Learning

The working of supervised machine learning follows these key steps:

1. Collect Labeled Data

 Gather a dataset where each input has a known correct output


(label).

 Example: Images of handwritten digits with their actual numbers as


labels.

2. Split the Dataset

 Divide the data into training data (about 80%) and testing data
(about 20%).

 The model will learn from the training data and be evaluated on the
testing data.

3. Train the Model

 Feed the training data (inputs and their labels) to a suitable


supervised learning algorithm (like Decision Trees, SVM or Linear
Regression).

 The model tries to find patterns that map inputs to correct outputs.

4. Validate and Test the Model

 Evaluate the model using testing data it has never seen before.

 The model predicts outputs and these predictions are compared


with the actual labels to calculate accuracy or error.
5. Deploy and Predict on New Data

 Once the model performs well, it can be used to predict outputs for
completely new, unseen data.

Supervised Machine Learning Algorithms

Supervised learning can be further divided into several different types,


each with its own unique characteristics and applications. Here are some
of the most common types of supervised learning algorithms:

 Linear Regression: Linear regression is a type of supervised


learning regression algorithm that is used to predict a continuous
output value. It is one of the simplest and most widely used
algorithms in supervised learning.

 Logistic Regression: Logistic regression is a type of supervised


learning classification algorithm that is used to predict a binary
output variable.

 Decision Trees : Decision tree is a tree-like structure that is used to


model decisions and their possible consequences. Each internal
node in the tree represents a decision, while each leaf node
represents a possible outcome.

 Random Forests: Random forests again are made up of multiple


decision trees that work together to make predictions. Each tree in
the forest is trained on a different subset of the input features and
data. The final prediction is made by aggregating the predictions of
all the trees in the forest.

 Support Vector Machine(SVM): The SVM algorithm creates a


hyperplane to segregate n-dimensional space into classes and
identify the correct category of new data points. The extreme cases
that help create the hyperplane are called support vectors, hence
the name Support Vector Machine.

 K-Nearest Neighbors: KNN works by finding k training examples


closest to a given input and then predicts the class or value based
on the majority class or average value of these neighbors. The
performance of KNN can be influenced by the choice of k and the
distance metric used to measure proximity.

 Gradient Boosting: Gradient Boosting combines weak learners,


like decision trees, to create a strong model. It iteratively builds new
models that correct errors made by previous ones.

 Naive Bayes Algorithm: The Naive Bayes algorithm is a


supervised machine learning algorithm based on applying Bayes'
Theorem with the “naive” assumption that features are independent
of each other given the class label.

You might also like