Experiment No.
Title: Write steps for installing the HADOOP in windows 10
The Apache Hadoop software library is a framework that allows for the distributed
processing of large data sets across clusters of computers using simple programming
models. It is designed to scale up from single servers to thousands of machines, each
offering local computation and storage. Rather than rely on hardware to deliver high-
availability, the library itself is designed to detect and handle failures at the application
layer, so delivering a highly-available service on top of a cluster of computers, each of
which may be prone to failures.
Install Java
– Java JDK Link to download
[Link]
– extract and install Java in C:\Java
– open cmd and type -> javac -version
Download Hadoop
– [Link]
– extract to C:\Hadoop
1. Set the path JAVA_HOME Environment variable
2. Set the path HADOOP_HOME Environment variable
Configurations: -
Edit file C:/Hadoop-3.3.0/etc/hadoop/[Link],
paste the xml code in folder and save
<configuration>
<property>
<name>[Link]</name>
<value>hdfs://localhost:9000</value>
</property>
</configuration>
======================================================
Rename “[Link]” to “[Link]” and edit this file C:/Hadoop-
3.3.0/etc/hadoop/[Link], paste xml code and save this file.
<configuration>
<property>
<name>[Link]</name>
<value>yarn</value>
</property>
</configuration>
======================================================
Create folder “data” under “C:\Hadoop-3.3.0”
Create folder “datanode” under “C:\Hadoop-3.3.0\data”
Create folder “namenode” under “C:\Hadoop-3.3.0\data”
======================================================
Edit file C:\Hadoop-3.3.0/etc/hadoop/[Link],
paste xml code and save this file.
<configuration>
<property>
<name>[Link]</name>
<value>1</value>
</property>
<property>
<name>[Link]</name>
<value>/hadoop-3.3.0/data/namenode</value>
</property>
<property>
<name>[Link]</name>
<value>/hadoop-3.3.0/data/datanode</value>
</property>
</configuration>
======================================================
Edit file C:/Hadoop-3.3.0/etc/hadoop/[Link],
paste xml code and save this file.
<configuration>
<property>
<name>[Link]-services</name>
<value>mapreduce_shuffle</value>
</property>
<property>
<name>[Link]</name>
<value>[Link]</value>
</property>
</configuration>
======================================================
Edit file C:/Hadoop-3.3.0/etc/hadoop/[Link]
by closing the command line
“JAVA_HOME=%JAVA_HOME%” instead of set “JAVA_HOME=C:\Java”
======================================================
Testing: -
– Open cmd and change directory to C:\Hadoop-3.3.0\sbin
– type [Link]
– Start namenode and datanode with this command
– type [Link]
– Start yarn through this command
– type [Link]
Make sure these apps are running
– Hadoop Namenode
– Hadoop datanode
– YARN Resource Manager
– YARN Node Manager
Open: [Link]
======================================================
Hadoop installed Successfully…………
======================================================
Applications of Hadoop
Hadoop is developed by Doug Cutting and Michale J. It is managed by apache software
foundation and licensed under the Apache license 2.0 Hadoop. It is beneficial for the big
business because it is based on cheap servers, requiring less cost to store the data and process
the data. Hadoop helps make a better business decision by providing a history of data and
various company records. So by using this technology company can improve its business.
Hadoop does lots of processing over collected data from the company to deduce the result
which can help to make a future decision.