0% found this document useful (0 votes)
7 views4 pages

Data Analytics Week1 Assignment.

Copyright
© All Rights Reserved
We take content rights seriously. If you suspect this is your content, claim it here.
Available Formats
Download as DOCX, PDF, TXT or read online on Scribd
0% found this document useful (0 votes)
7 views4 pages

Data Analytics Week1 Assignment.

Copyright
© All Rights Reserved
We take content rights seriously. If you suspect this is your content, claim it here.
Available Formats
Download as DOCX, PDF, TXT or read online on Scribd

Data Analytics Week 1 Assignment.

Part1.
1. Define Data Analytics and its importance in decision-making.
Data Analytics is the process of turning raw, scattered data into meaningful insights. It
answers questions like what happened, why it happened, and what might happen next. Its
importance? It transforms decisions from guesswork into evidence-based actions, helping
organizations cut risks, spot opportunities, and make effective choices.

2. Differentiate between Descriptive, Diagnostic, Predictive, and Prescriptive


Analytics.
• Descriptive: Asks “What happened?” *summarizes past events. (e.g. Monthly sales
Reports.)
• Diagnostic: Asks “Why did it happen?” * digs deeper to find reasons. (e.g. sales ,dropped
due to low Inventory)
• Predictive: Tells “What will happen?” * uses patterns to forecast the future. (e.g.
Forecasting next Month’s sales.)
• Prescriptive: Gives “What should we do?” * offers recommendations for action. (e.g. stock
up before high demand hits.)
Together, they move from history → investigation → forecast → action plan.

3. Explain the Data Analytics Life Cycle (stages & activities).


The life cycle is like a journey:
• Ask: Define the problem clearly. Frame the right questions.
• Prepare/Process: Gather, collect, clean, and organize the data.
• Analyze: Apply techniques to explore patterns and extract insights.
• Share: Present findings with clear visuals and stories.
• Act: Turn insights into real actions and monitor results and refine.

In short: Ask the right questions, prepare carefully, analyze smartly, share clearly, and act
without bias.

4. Describe the role of a Data Analyst in driving insight and decision-making.


A Data Analyst is like a bridge between data and decisions. They collect and clean data,
analyze it, and then transform numbers into stories that organizations can act on. By
presenting insights through reports and visualizations, they guide stakeholders toward
smarter choices, helping businesses solve problems, improve performance, and plan for the
future.
5. Discuss the difference between Structured and Unstructured Data.
• Structured Data: Neat and organized, fits into rows and columns. Example: sales records
or customer databases.
• Unstructured Data: Messy and without a clear format. Example: social media posts, videos,
or emails.

Think of structured data as a well-arranged bookshelf, while unstructured data is a pile of


scattered books waiting to be sorted.

6. Explain the concept of Data Visualization and its significance in Data


Analytics.
Data visualization is the art of turning raw numbers into charts, graphs, or dashboards that
people can instantly understand. Its significance lies in clarity, stakeholders can quickly
grasp insights, spot patterns, and act faster. In short, it makes data not just useful, but also
engaging and easy to interpret.

7. Describe the difference between Correlation and Causation in Data Analysis.


• Correlation: Two variables move together, but one doesn’t cause the other.
• Causation: One variable directly causes the change in the other.

In simple terms, correlation shows a relationship, causation shows a cause-and-effect link.

8. Explain the difference between Data Mining, Machine Learning, and Artificial
Intelligence.
• Data Mining: Finding hidden patterns in big piles of data, like treasure hunting.
• Machine Learning: Teaching computers to learn from data and make predictions, like
training a pet with examples.
• Artificial Intelligence (AI): Making computers act smart, almost like humans, learning,
deciding, and solving problems on their own.

Think of it this way: Data mining finds clues, machine learning learns from clues, and AI
uses clues to act smart.

9. How can Data Analytics be applied in Finance, Healthcare, and Retail?


• Finance: Helps spot fraud, predict risks, and guide smart investments.
• Healthcare: Tracks patient data, improves diagnosis, and predicts diseases before they
spread.
• Retail: Knows what customers like, boosts sales with recommendations, and manages
stock better.

In short: finance gets safer, healthcare gets smarter, and retail gets closer to the customer.
10. What is the main purpose of a Data Warehouse and how does it differ from
a Data Lake?
A Data Warehouse is like a well-organized library. Every book (data) is cleaned, labeled, and
placed on the right shelf. The purpose is simple: when you need answers, like sales numbers
or performance reports, you can find them quickly because everything is neatly arranged.

A Data Lake is different. It’s like a huge storage hall where all kinds of information are kept
in their raw form, books, loose papers, photos, videos, even recordings, without sorting first.
You keep everything together and decide later how you want to use it.

In short: Warehouse = organized, structured, decision-ready;

Lake = raw, unstructured, full of different data types waiting to be explored.

11. Describe the ETL process in data management. What do the letters E, T, and
L stand for?
The ETL process is like preparing a meal before serving it, but here the “ingredients” are
data.
• E – Extract: This is like picking ingredients from the market. In data terms, it means
collecting raw data from different sources such as files, databases, or apps.
• T – Transform: Just as you wash, cut, and cook food to make it ready to eat, this step
means cleaning, sorting, and converting the raw data into the right format for use.
• L – Load: Finally, just like serving the cooked meal on a plate, this step means storing the
prepared data neatly into a data warehouse or database, ready for analysis.

In short: ETL = collect the data, prepare it properly, then store it where it can be easily used.

12. What is a Data Pipeline and how does it differ from an ETL Pipeline?
A Data Pipeline is like a series of connected pipes moving water from one place to another—
but instead of water, it moves data. It automatically collects data from different sources and
delivers it safely to a target location, like a database or warehouse.

An ETL Pipeline is a type of data pipeline with a specific recipe: Extract, Transform, Load.
This means it doesn’t just move the data. It also cleans and reshapes it before storing it.

In short: All ETL pipelines are examples of data pipelines, but not all data pipelines are
ETL.

13. Summarize the roles and responsibilities of a Data Scientist and a Data
Engineer
• Data Scientist: Like a problem-solver, they use statistics and machine learning to build
models, make predictions, and uncover deep insights. They work on the full process, from
collecting and cleaning data to building algorithms and guiding decisions.
• Data Engineer: Like a builder, they design and maintain the systems that hold and move
data. They create data pipelines, warehouses, and databases so analysts and scientists have
reliable, high-quality data to work with.

In short: Data Engineers build the roads; Data Scientists drive on them to reach insights.

14. How can Data Bias affect analytics outcomes and what steps can analysts
take to mitigate bias?
Bias in data is like looking through a cracked mirror, it gives you a distorted view. If the data
isn’t representative or carries old stereotypes, the results can be unfair or misleading,
leading to wrong decisions.

To reduce bias, analysts should:


• Use diverse and balanced datasets.
• Check for hidden assumptions in the data.
• Apply fairness-aware methods.

In short: Bad data in = bad results out. Balanced data in = fairer, clearer insights.

15. What are the ethical considerations data analysts must keep in mind when
working with data?
Analysts must treat data with care and responsibility, like handling someone else’s personal
diary. Key considerations include:
• Privacy & Consent: Collect and use data only with proper permission.
• Anonymity: Protect identities by removing personal details.
• Ownership: Respect that data belongs to the people or organizations who provided it.
• Fairness: Avoid bias that could lead to discrimination.
• Transparency: Report results honestly, with context and limitations.
• Retention: Don’t keep data longer than needed.

In short: Be respectful, be fair, be honest, and protect people’s privacy.

You might also like