/*! This file is auto-generated */ .wp-block-button__link{color:#fff;background-color:#32373c;border-radius:9999px;box-shadow:none;text-decoration:none;padding:calc(.667em + 2px) calc(1.333em + 2px);font-size:1.125em}.wp-block-file__button{background:#32373c;color:#fff;text-decoration:none} Q45 Why Divide by n 鈭 1? Let a pop... [FREE SOLUTION] | 91影视

91影视

Why Divide by n 鈭 1? Let a population consist of the values 9 cigarettes, 10 cigarettes, and 20 cigarettes smoked in a day (based on data from the California Health Interview Survey). Assume that samples of two values are randomly selected with replacement from this population. (That is, a selected value is replaced before the second selection is made.)

a. Find the variance2 of the population {9 cigarettes, 10 cigarettes, 20 cigarettes}.

b. After listing the nine different possible samples of two values selected with replacement, find the sample variance s2 (which includes division by n - 1) for each of them; then find the mean of the nine sample variances s2.

c. For each of the nine different possible samples of two values selected with replacement, find the variance by treating each sample as if it is a population (using the formula for population variance, which includes division by n); then find the mean of those nine population variances.

d. Which approach results in values that are better estimates of2 part (b) or part (c)? Why? When computing variances of samples, should you use division by n or n - 1?

e. The preceding parts show that s2 is an unbiased estimator of 2. Is s an unbiased estimator of ? Explain

Short Answer

Expert verified

(a) Population variance 2 is equal to 24.7.

(b)s12= 0.5,s22= 60.5,s32= 50.0,s42= 0.0,s52= 0.0,s62= 0.0,s72= 0.5,s82= 60.5, ands92= 50.0. The mean of the 9-sample variances is 24.7.

(c) 12= 0.25,22= 30.25,32= 25.0,42= 0.0,52= 0.0,62= 0.0,72= 0.25,82= 30.25, and92= 25.0. The mean of the 9-population variances is 12.3.

(d) The method in part (b) results in a better estimate as multiple samples are used to compute the mean of the sample variances. Thus, the value becomes equal to the population variance. Moreover, using n鈥1 gives a precise estimate.

(e) No, s is not an unbiased estimator of as the mean of the sample standard deviations is not equal to the population standard deviation.

Step by step solution

01

Given information

A population of three values (number of cigarettes) is given.

Out of these, nine samples are selected with replacement.

02

Population variance and sample variance

Population variance2 is calculated by dividing the sum of the squared differences of the population observations (from the mean) by the count of observations.

Mathematically,

2=i=1nxi-2n

Here, n is the total number of observations.

Sample variances2 is calculated by dividing the sum of the squared differences of the sample observations from the mean by n鈥1.

Mathematically,

s2=i=1nxi-x2n-1

03

Compute the population variance

(a)

To compute the value of the population variance, find the population mean as shown below.

=9+10+203=13.0

The population variance is computed as follows:

2=i=1nxi-2n=9-13.02+10-13.02+20-13.023=24.7

The population variance is 24.7.

04

Describe the feasible samples of size two from the collection

(b)

The nine different samples selected with replacement are shown below:

Sample 1

Sample 2

Sample 3

9

9

10

10

20

20

Sample 4

Sample 5

Sample 6

9

10

20

9

10

20

Sample 7

Sample 8

Sample 9

10

20

20

9

9

10

The mean of each sample is computed using the formula x=xn .

The mean for each sample is stated in the brackets in the following table.

Sample 1

Sample 2

Sample 3

9

9

10

10

20

20

x1=9.5

x2=14.5

x3=15

Sample 4

Sample 5

Sample 6

9

10

20

9

10

20

x4=9

x5=10

x6=20

Sample 7

Sample 8

Sample 9

10

20

20

9

9

10

x7=9.5

x8=14.5

x9=15

The sample variances are computed as shown below.

s12=9-9.52+10-9.522-1=0.5s22=9-14.52+20-14.522-1=60.5

s32=10-152+20-1522-1=50.0s42=9-92+9-922-1=0.0

s52=10-102+10-1022-1=0.0s62=20-202+20-2022-1=0.0

s72=10-9.52+9-9.522-1=0.5s82=20-14.52+9-14.522-1=60.5

s92=20-152+10-1522-1=50.0

The mean of the nine sample variances is

s2=i=19si29=24.7

Thus, the mean of the sample variances is 24.7.

05

Describe the variances for each sample  using the population variance

(c)

The variance of samples is computed using the formula for population variance, as shown below.

Considering the above nine samples as populations, you can compute the population variances as shown below.

12=9-9.52+10-9.522=0.2522=9-14.52+20-14.522=30.25

32=10-152+20-1522=25.042=9-92+9-922=0.0

52=10-102+10-1022=0.062=20-202+20-2022=0.0

72=10-9.52+9-9.522=0.2582=20-14.52+9-14.522=30.25

92=20-152+10-1522=25.0

The mean of the nine population variances is

2=i=19i29=12.3.

Thus, the mean of the population variances is 12.3.

06

Compare the results of parts (b) and (c)

(d)

Part (b) gives a better estimate. By usingn鈥1for sample variance, the value gives a precise estimate of the population variance.

Here, a repeated number of samples tends tocenterthe value of the resultant variance close to the population variance. In the case of sample variance,division by n鈥1 is performed rather than by n. If divided by n, the value of the sample variance underestimates the value of population variance.

07

Explain if the sample standard deviation is an unbiased estimator of the population standard deviation

(e)

An unbiased estimate is a measure for sample values that have a mean equivalent or are close to the population value of the measure.

The standard deviations for the nine samples are calculated below:

s1=s12=0.7s2=s22=7.8

s3=s32=7.1s4=s42=0.0

s5=s52=0.0s6=s62=0.0

s7=s72=0.7s8=s82=7.8

s9=s92=7.1

The mean of these nine sample standard deviations is

s=i=19si9=3.5.

Therefore, the mean of the sample standard deviations is 3.5.

The population standard deviation is

=2=5.0.

Thus, the value of the population standard deviation is 5.0.

Here, the mean of the sample standard deviations is not equal to the population standard deviation.

Therefore, the sample standard deviation s is not an unbiased estimator of the population standard deviation .

Unlock Step-by-Step Solutions & Ace Your Exams!

  • Full Textbook Solutions

    Get detailed explanations and key concepts

  • Unlimited Al creation

    Al flashcards, explanations, exams and more...

  • Ads-free access

    To over 500 millions flashcards

  • Money-back guarantee

    We refund you if you fail your exam.

Over 30 million students worldwide already upgrade their learning with 91影视!

One App. One Place for Learning.

All the tools & learning materials you need for study success - in one app.

Get started for free

Most popular questions from this chapter

Symbols Identify the symbols used for each of the following: (a) sample standard deviation;(b) population standard deviation;(c) sample variance;(d) population variance.

Critical Thinking. For Exercises 5鈥20, watch out for these little buggers. Each of these exercises involves some feature that is somewhat tricky. Find the (a) mean, (b) median, (c) mode, (d) midrange, and then answer the given question

Sales of LP Vinyl Record Albums Listed below are annual U.S. sales of vinyl record albums (millions of units). The numbers of albums sold are listed in chronological order, and the last entry represents the most recent year. Do the measures of center give us any information about a changing trend over time?

0.3 0.6 0.8 1.1 1.1 1.4 1.4 1.5 1.2 1.3 1.4 1.2 0.9 0.9 1 1.9 2.5 2.8 3.9 4.6 6.1

Range Rule of Thumb for Interpreting s The 20 brain volumes (cm3 ) from Data Set 8 鈥淚Q and Brain Size鈥 in Appendix B have a mean of 1126.0 cm3 and a standard deviation of 124.9 cm3. Use the range rule of thumb to identify the limits separating values that are significantly low or significantly high. For such data, would a brain volume of 1440 cm3 be significantly high?

In Exercises 21鈥24, find the mean and median for each of the two samples, then compare the two sets of results.

Blood Pressure A sample of blood pressure measurements is taken from Data Set 1 鈥淏ody Data鈥 in Appendix B, and those values (mm Hg) are listed below. The values are matched so that 10 subjects each have systolic and diastolic measurements. (Systolic is a measure of the

force of blood being pushed through arteries, but diastolic is a measure of blood pressure when the heart is at rest between beats.) Are the measures of center the best statistics to use with these data? What else might be better?

Systolic: 118 128 158 96 156 122 116 136 126 120

Diastolic: 80 76 74 52 90 88 58 64 72 82

In Exercises 21鈥24, find the mean and median for each of the two samples, then compare the two sets of results.

Bank Queues Waiting times (in seconds) of customers at the Madison Savings Bank are recorded with two configurations: single customer line; individual customer lines. Carefully examine the data to determine whether there is a difference between the two data sets that is not apparent from a comparison of the measures of center. If so, what is it?

Single Line 390 396 402 408 426 438 444 462 462 462

Individual Lines 252 324 348 372 402 462 462 510 558 600

See all solutions

Recommended explanations on Math Textbooks

View all explanations

What do you think about this solution?

We value your feedback to improve our textbook solutions.

Study anywhere. Anytime. Across all devices.