0% found this document useful (0 votes)
2 views10 pages

Statistical Analysis and Observations

The document contains a series of questions and answers related to statistical observations, cumulative frequency polygons, and data analysis. It discusses the differences between raw and grouped data, the calculation of mean, median, and quartiles, as well as the interpretation of results. The document emphasizes the importance of understanding data distribution and the implications of data grouping on statistical measures.
Copyright
© All Rights Reserved
We take content rights seriously. If you suspect this is your content, claim it here.
Available Formats
Download as PDF, TXT or read online on Scribd
0% found this document useful (0 votes)
2 views10 pages

Statistical Analysis and Observations

The document contains a series of questions and answers related to statistical observations, cumulative frequency polygons, and data analysis. It discusses the differences between raw and grouped data, the calculation of mean, median, and quartiles, as well as the interpretation of results. The document emphasizes the importance of understanding data distribution and the implications of data grouping on statistical measures.
Copyright
© All Rights Reserved
We take content rights seriously. If you suspect this is your content, claim it here.
Available Formats
Download as PDF, TXT or read online on Scribd

STTN122

Hoofstuk 4 / Chapter 4
Vraag 1 / Question 1
1.1.
1.2. waarneming / observation

1.3.
1.4. Vir / For waarneming / observation

Vir / For waarneming / observation

Vraag 2 / Question 2
Vir tak 1: / For branch 1:
2.1.
2.2.
2.3. en / and
Vir tak 2: / For branch 2:
2.1.
2.2.
2.3. en / and

Vraag 3 / Question 3
3.1.
3.2.
3.3.
3.4.

Vraag 4 / Question 4
4.1.
4.2.

Cumulative Frequency Polygon


40

35

30

25

20

15

10

0
10 20 30 40 50 60 70 80
Hoogte

4.3. en / and
4.4. Omdat die gemiddeld en mediaan byna gelyk is ( ) en die eerste kwartiel, mediaan en
derde kwartiel bykans ewe ver uitmekaar is, lei ons af dat die verdeling simmetries is.
Since the mean and the median are almost equal ( ) and the first quartile, median and
third quartile are almost equally far away from each other, we conclude that the distribution is
symmetric.

Vraag 5 / Question 5
5.1.
5.2.
Cumulative Frequency Polygon
80

70

60

50

40

30

20

10

0
80 90 100 110 120 130 140
IK

5.3. Modale interval: / Modal interval: . Gebruik histogram om die modus te benader
(soos in handboek verduidelik). / Use histogram to estimate the mode (as explained in the
textbook).
5.4. en / and

Vraag 6 / Question 6
a.
b.
Die antwoorde verskil omdat ons in (b) nie soos in (a) van die oorspronklike datawaardes gebruik
gemaak het nie, maar van die klasmiddelwaardes en ooreenstemmende frekwensies van die
gegroepeerde datastel. Ons het dus in werklikheid aangeneem dat die klasmiddelwaarde van elke klas
as beraming vir die waarnemings se waardes binne daardie klas gebruik kan word.
The answers differ since in (a) we make use of all the original data values (raw data) but in (b) we make
use of the class midpoints and corresponding frequencies. In reality we thus made the assumption that
the class midpoint of each class can be used as estimator for the observations’ values within that class.

Vraag 7 / Question 7
Let wel: daar is 24 waarnemings in die datastel!
Note: there are 24 observations in the dataset!
7.1.
7.2.
Class boundaries Frequency ( ) Relative frequency ( )

Total:

7.3.
Class boundaries Cumulative frequency ( ) Relative cumulative frequency
( )

Cumulative Frequency Polygon


25

20

15

10

0
0 20 40 60 80 100 120 140
Duration

7.4. oproepe / calls


7.5. Vanuit rou data: / From raw data:

Vanuit gegroepeerde data: / From grouped data:

Die twee antwoorde verskil omdat daar inligting verlore gaan wanneer data gegroepeer word.
The two answers differ because information is lost when data are grouped.
7.6. Vanuit rou data: / From raw data:

Vanuit gegroepeerde data: / From grouped data:

Vraag 8 / Question 8
8.1. en / and .
Mans eis op gemiddeld die meeste.
On average, men claim the most.
8.2. en / and .
Mans het die hoogste mediaan-eisbedrag.
Men have the highest median claim amount.
8.3. en / and
en / and .

Vraag 9 / Question 9
Werk eers terug na frekwensietabel om die vrae te beantwoord: / First work back to frequency table to
answer the questions:
Class boundaries Frequency ( ) Cumulative frequency ( )

9.1. modale interval / modal interval


9.2. Benodig klasmiddelwaardes: / Requires class midpoints:

Blaai om vir tabel / Turn page for table


Class boundaries Class midpoint ( ) Frequency ( )

Total:

Nou / Now

9.3.
Cumulative Frequency Polygon
50
40
30
F

20
10
0

20 30 40 50 60

Class Upper Bounds

Mediaan se posisie / Median’s position waarneming / observation

mediaan / median

9.4. Deur weer grafiek in (9.3) te gebruik: / By again using graph in (9.3):
-posisie / position waarneming / observation
Vraag 10 / Question 10
10.1. Die formule waarmee die rekenkundige gemiddeld uitgewerk is, is . Ons

benodig dus die frekwensiekolom ( ) en kolom met klasmiddelwaardes ( ) om die


vergelyking op te los.
The formula with which the arithmetic mean was calculated is . Hence we

need the frequency column ( ) and column with class midpoints ( ) to solve the equation.
Marks (%) Class midpoint Frequency ( ) Cumulative
( ) frequency ( )

Total:
Los nou die volgende vergelyking op: / Now solve the following equation:

10.2. Die eerste kwartiel ( ) is die punt wat 75% van die studente oorskry, dus moet grafies
bepaal word deur die kumulatiewe frekwensieveelhoek te gebruik: / The first quartile ( ) is the
point which 75% of the students exceeded, hence must be determined graphically by using
the cumulative frequency polygon:

Blaai om vir grafiek / Turn over for graph


Cumulative Frequency Polygon

100
80
60
F

40
20
0

0 20 40 60 80 100

Mark (%)

Die eerste kwartiel word grafies vasgestel as . / The first quartile is graphically determined as
.

Vraag 11 / Question 11
11. Net soos in Vraag 10.1 benodig ons die res van die tabel as volg: / Just as in Question 10.1 we
need the rest of the table as follows:
Class interval Frequency ( ) Cumulative Class midpoint
frequency ( ) ( )

Los nou die volgende vergelyking op: / Now solve the following equation:
Vraag 12 / Question 12
12. Ja, aangesien die mediaan slegs van posisie afhang en nie van waardes nie. Gestel dat die
aanvanklike vyf waarnemings georden is van klein na groot, dan is die mediaan in die derde
posisie:
Yes, since the median depends only on position and not on values. Suppose that the initial five
observations are ordered from smallest to largest, then the median is in the third position:

Indien daar ʼn sesde persoon opklim moet die waarnemings weer van klein na groot gerangskik
word met hierdie nuwe waarneming in ag genome. Let op dat die -persoon slegs in die
vierde, vyfde of sesde posisie kan val aangesien die derde persoon reeds is. Die
mediaan is nou die gemiddeld van die derde en vierde waarnemings.
If a sixth person enters the elevator, then the observations need to be arranged from smallest
to largest again with this sixth observation taken into account. Note that the person can
only fall into the fourth, fifth or sixth position since the third person is already . The
median is now the mean of the third and fourth observations.
Indien die -persoon in die vierde posisie val is dit onmoontlik dat die mediaan konstant

kan bly op , aangesien die derde persoon weeg, en dus . Maar

as die -persoon in die vyfde of sesde posisie is, dan is dit moontlik vir die mediaan om

konstant te bly op MITS die vierde persoon ook weeg, sodat .

If the person falls in the fourth position then it is impossible that the median stays

constant on since the third person weighs , and hence . But if the

person is in the fifth or sixth position, then it is possible for the median to stay constant

on BUT ONLY IF the fourth person also weighs , so that .

You might also like