/*! This file is auto-generated */ .wp-block-button__link{color:#fff;background-color:#32373c;border-radius:9999px;box-shadow:none;text-decoration:none;padding:calc(.667em + 2px) calc(1.333em + 2px);font-size:1.125em}.wp-block-file__button{background:#32373c;color:#fff;text-decoration:none} Q25 E The accompanying data consists o... [FREE SOLUTION] | 91影视

91影视

The accompanying data consists of prices (\$) for one sample of California cabernet sauvignon wines that received ratings of 93 or higher in the May 2013 issue of Wine Spectator and another sample of California cabernets that received ratings of 89 or lower in the same issue.

\(\begin{array}{*{20}{c}}{ \ge 93:}&{100}&{100}&{60}&{135}&{195}&{195}&{}\\{}&{125}&{135}&{95}&{42}&{75}&{72}&{}\\{ \le 89:}&{80}&{75}&{75}&{85}&{75}&{35}&{85}\\{}&{65}&{45}&{100}&{28}&{38}&{50}&{28}\end{array}\)

Assume that these are both random samples of prices from the population of all wines recently reviewed that received ratings of at least 93 and at most 89 , respectively.

a. Investigate the plausibility of assuming that both sampled populations are normal.

b. Construct a comparative boxplot. What does it suggest about the difference in true average prices?

c. Calculate a confidence interval at the\(95\% \)confidence level to estimate the difference between\({\mu _1}\), the mean price in the higher rating population, and\({\mu _2}\), the mean price in the lower rating population. Is the interval consistent with the statement "Price rarely equates to quality" made by a columnist in the cited issue of the magazine?

Short Answer

Expert verified

(a) Plausible

(b) A large difference .

(c) \((16.1180,81.9534)\)

The interval is not consistent with the statement.

Step by step solution

01

a)Step 1: Determine the normal probability plot

Given:

\(\begin{array}{l} \ge 93:100,100,60,135,195,195,125,135,95,42,75,72\\ \le 89:80,75,75,85,75,35,85,65,45,100,28,38,50,28\end{array}\)

If we want to perform a two-sample\(t\)test, then we require that both sampling distributions of the sample mean are approximately normal.

FIRST DATA SET

We will create a normal probability plot.

The data values are on the horizontal axis and the standardized normal scores are on the vertical axis.

If the data contains\(n\)data values, then the standardized normal scores are the z-scores in the normal probability table of the appendix corresponding to an area of\(\frac{{j - 0.5}}{n}\)(or the closest area) with\(j \in \{ 1,2,3, \ldots ,n\} \).

The smallest standardized score corresponds with the smallest data value, the second smallest standardized score corresponds with the second smallest data value, and so on.

02

b)Step 2: Determine the normal probability plot

\(\begin{array}{l}{{\bar x}_1} = 13.4\\{{\bar x}_2} = 9.7\\{n_1} = 65\\{n_2} = 50\\{\sigma _{{{\bar x}_1}}} = 2.05 \Rightarrow {s_1} = {\sigma _{{{\bar x}_1}}}\sqrt n = 2.05\sqrt {65} \approx 16.5276\\{\sigma _{{{\bar x}_2}}} = 1.76 \Rightarrow {s_2} = {\sigma _{{{\bar x}_2}}}\sqrt n = 1.76\sqrt {50} \approx 12.4451\end{array}\)

Let us assume: \(\alpha = 0.05\)

Given claim: exceeds

The claim is either the null hypothesis or the alternative hypothesis. The null hypothesis and the alternative hypothesis state the opposite of each other. The null hypothesis needs to contain the value mentioned in the claim.

\(\begin{array}{l}{H_0}:{\mu _1} = {\mu _2}\\{H_a}:{\mu _1} > {\mu _2}\end{array}\)

SECOND DATA SET

We will create a normal probability plot.

The data values are on the horizontal axis and the standardized normal scores are on the vertical axis.

If the data contains \(n\) data values, then the standardized normal scores are the z-scores in the normal probability table of the appendix corresponding to an area of \(\frac{{j - 0.5}}{n}\) (or the closest area) with\(j \in \{ 1,2,3, \ldots ,n\} \).

The smallest standardized score corresponds with the smallest data value, the second smallest standardized score corresponds with the second smallest data value, and so on.

If the pattern in the normal probability plot is roughly linear and does not contain strong curvature, then the population distribution is approximately normal.

Both probability plots do not contain strong curvature and are roughly linear, thus both population distributions are approximately normal.

Since the population distributions are approximately normal, the sampling distribution of the sample mean(s) \(\bar x\) are also approximately normal. and thus it is appropriate to use the two-sample\(t\) test.

03

B)Step 3: Fild the quartile for first data set

Given:

\(\begin{array}{l} \ge 93:100,100,60,135,195,195,125,135,95,42,75,72\\ \le 89:80,75,75,85,75,35,85,65,45,100,28,38,50,28\end{array}\)

Sort the data values from smallest to largest:

\(\begin{array}{l} \ge 93:42,60,72,75,95,100,100,125,135,135,195,195\\ \le 89:28,28,35,38,45,50,65,75,75,75,80,85,85,100\end{array}\)

FIRST DATA SET

The minimum is \(42.\)

Since the number of data values is even, the median is the average of the two middle values of the sorted data set:

\(M = {Q_2} = \frac{{100 + 100}}{2} = 100\)

The first quartile is the median of the data values below the median (or at \(25\% \) of the data):

\({Q_1} = \frac{{72 + 75}}{2} = 73.5\)

The third quartile is the median of the data values above the median (or at \(75\% \) of the data):

\({Q_3} = \frac{{135 + 135}}{2} = 135\)

The maximum is \(195.\)

04

Find the quartile for second data set

SECOND DATA SET

The minimum is\(28.\)

Since the number of data values is even, the median is the average of the two middle values of the sorted data set:

\(M = {Q_2} = \frac{{65 + 75}}{2} = 70\)

The first quartile is the median of the data values below the median (or at\(25\% \)of the data):

\({Q_1} = 38\)

The third quartile is the median of the data values above the median (or at\(75\% \)of the data):

\({Q_3} = 80\)

The maximum is \(100\) .

05

Mapping the graph

The whiskers of the boxplot are at the minimum and maximum value. The box starts at the first quartile, ends at the third quartile and has a vertical line at the median.

The first quartile is at \(25\% \) of the sorted data list, the median at \(50\% \) and the third quartile at\(75\% \).

There appears to be a large difference between the true average prices, because the vertical lines corresponding to the median in the box of th boxplots lie are not roughly at the same location (on the horizontal axis)

06

c)Step 6: Determine the standard deviation

Given:

\(\begin{array}{l} \ge 93:100,100,60,135,195,195,125,135,95,42,75,72\\ \le 89:80,75,75,85,75,35,85,65,45,100,28,38,50,28\end{array}\)

The mean is the sum of all values divided by the number of values:

\(\begin{array}{l}{{\bar x}_1} = \frac{{100 + 100 + 60 + \ldots + 42 + 75 + 72}}{{12}} = 110.75\\{{\bar x}_2} = \frac{{80 + 75 + 75 + \ldots + 38 + 50 + 28}}{{14}} \approx 61.7143\end{array}\)

The variance is the sum of squared deviations from the mean divided by\(n - 1\). The standard deviation is the square root of the variance: \(\begin{array}{l}{s_1} = \sqrt {\frac{{{{(100 - 110.75)}^2} + \ldots . + {{(72 - 110.75)}^2}}}{{12 - 1}}} \approx 48.7445\\{s_2} = \sqrt {\frac{{{{(80 - 61.7143)}^2} + \ldots . + {{(28 - 61.7143)}^2}}}{{14 - 1}}} \approx 23.8438\end{array}\)

07

Find the endpoint of the confidence interval

Given:

\(c = 95\% = 0.95\)

Determine the degrees of freedom (rounded down to the nearest integer):

\(\Delta = \frac{{{{\left( {\frac{{s_1^2}}{{{n_1}}} + \frac{{s_2^2}}{{{n_2}}}} \right)}^2}}}{{\frac{{{{\left( {s_1^2/{n_1}} \right)}^2}}}{{{n_1} - 1}} + \frac{{{{\left( {s_2^2/{n_2}} \right)}^2}}}{{{n_2} - 1}}}} = \frac{{{{\left( {\frac{{{{48.7445}^2}}}{{12}} + \frac{{{{23.8438}^2}}}{{14}}} \right)}^2}}}{{\frac{{{{\left( {{{48.7445}^2}/12} \right)}^2}}}{{12 - 1}} + \frac{{{{\left( {{{23.8438}^2}/14} \right)}^2}}}{{14 - 1}}}} \approx 15\)

Determine the t-value by looking in the row starting with degrees of freedom \(df = 15\) and in the column with \(1 - c/2 = 0.025\) in the Student's t distribution table in the appendix:

\({t_{\alpha /2}} = 2.131\)

The margin of error is then:

\(E = {t_{\alpha /2}} \cdot \sqrt {\frac{{s_1^2}}{{{n_1}}} + \frac{{s_2^2}}{{{n_2}}}} = 2.131 \cdot \sqrt {\frac{{{{48.7445}^2}}}{{12}} + \frac{{{{23.8438}^2}}}{{14}}} \approx 32.9177\)

The endpoints of the confidence interval for \({\mu _1} - {\mu _2}\) are: \(\begin{array}{l}\left( {{{\bar x}_1} - {{\bar x}_2}} \right) - E = (110.75 - 61.7143) - 32.9177 = 49.0357 - 32.9177 = 16.1180\\\left( {{{\bar x}_1} - {{\bar x}_2}} \right) + E = (110.75 - 61.7143) + 32.9177 = 49.0357 + 32.9177 = 81.9534\end{array}\)

Unlock Step-by-Step Solutions & Ace Your Exams!

  • Full Textbook Solutions

    Get detailed explanations and key concepts

  • Unlimited Al creation

    Al flashcards, explanations, exams and more...

  • Ads-free access

    To over 500 millions flashcards

  • Money-back guarantee

    We refund you if you fail your exam.

Over 30 million students worldwide already upgrade their learning with 91影视!

One App. One Place for Learning.

All the tools & learning materials you need for study success - in one app.

Get started for free

Most popular questions from this chapter

Suppose \({\mu _1}\) and \({\mu _2}\) are true mean stopping distances at \(50mph\) for cars of a certain type equipped with two different types of braking systems. Use the two-sample t test at significance level

t test at significance level \(.01\) to test \({H_0}:{\mu _1} - {\mu _2} = - 10\) versus \({H_a}:{\mu _1} - {\mu _2} < - 10\) for the following data: \(m = 6,\;\;\;\bar x = 115.7,{s_1} = 5.03,n = 6,\bar y = 129.3,\;\)and \({s_2} = 5.38.\)

Quantitative noninvasive techniques are needed for routinely assessing symptoms of peripheral neuropathies, such as carpal tunnel syndrome (CTS). The article "A Gap Detection Tactility Test for Sensory Deficits Associated with Carpal Tunnel Syndrome" (Ergonomics, \(1995: 2588 - 2601\)) reported on a test that involved sensing a tiny gap in an otherwise smooth surface by probing with a finger; this functionally resembles many work-related tactile activities, such as detecting scratches or surface defects. When finger probing was not allowed, the sample average gap detection threshold for\(m = 8\)normal subjects was\(1.71\;mm\), and the sample standard deviation was\(.53\); for\(n = 10\)CTS subjects, the sample mean and sample standard deviation were\(2.53\)and\(.87\), respectively. Does this data suggest that the true average gap detection threshold for CTS subjects exceeds that for normal subjects? State and test the relevant hypotheses using a significance level of\(.01\).

Reliance on solid biomass fuel for cooking and heating exposes many children from developing countries to high levels of indoor air pollution. The article 鈥淒omestic Fuels, Indoor Air Pollution, and Children鈥檚 Health鈥 (Annals of the N.Y. Academy of Sciences, \(2008:209 - 217\)) presented information on various pulmonary characteristics in samples of children whose households in India used either biomass fuel or liquefied petroleum gas (\(LPG\)). For the \(755\) children in biomass households, the sample mean peak expiratory flow (a person鈥檚 maximum speed of expiration) was \(3.30L/s\), and the sample standard deviation was \(1.20\). For the \(750\) children whose households used liquefied petroleum gas, the sample mean \(PEF\) was \(4.25\) and the sample standard deviation was \(1.75\).

a. Calculate a confidence interval at the \(95\% \) confidence level for the population mean \(PEF\) for children in biomass households and then do likewise for children in \(LPG\) households. What is the simultaneous confidence level for the two intervals?

b. Carry out a test of hypotheses at significance level \(.01\) to decide whether true average \(PEF\) is lower for children in biomass households than it is for children in \(LPG\) households (the cited article included a P-value for this test)

c. \(FE{V_1}\), the forced expiratory volume in \(1\) second, is another measure of pulmonary function. The cited article reported that for the biomass households the sample mean FEV1 was \(2.3L/s\) and the sample standard deviation was \(.5L/s\). If this information is used to compute a \(95\% \) \(CI\) for population mean \(FE{V_1}\), would the simultaneous confidence level for this interval and the first interval calculated in (a) be the same as the simultaneous confidence level determined there? Explain

The level of lead in the blood was determined for a sample of \(152\) male hazardous-waste workers ages \(20 - 30\) and also for a sample of \(86\) female workers, resulting in a mean \(6\) standard error of \(5.5 \pm 0.3\)for the men and \(3.8 \pm 0.2\) for the women (鈥淭emporal Changes in Blood Lead Levels of Hazardous Waste Workers in New Jersey, 1984鈥1987,鈥 Environ. Monitoring and Assessment, 1993: 99鈥107). Calculate an estimate of the difference between true average blood lead levels for male and female workers in a way that provides information about reliability and precision.

Is there any systematic tendency for part-time college faculty to hold their students to different standards than do full-time faculty? The article 鈥淎re There Instructional Differences Between Full-Time and Part-Time Faculty?鈥 (College Teaching, \(2009: 23 - 26)\) reported that for a sample of \(125 \) courses taught by fulltime faculty, the mean course \(GPA\) was \(2.7186\) and the standard deviation was \(.63342\), whereas for a sample of \(88\) courses taught by part-timers, the mean and standard deviation were \(2.8639\) and \(.49241,\) respectively. Does it appear that true average course \(GPA\) for part-time faculty differs from that for faculty teaching full-time? Test the appropriate hypotheses at significance level \( .01\).

See all solutions

Recommended explanations on Math Textbooks

View all explanations

What do you think about this solution?

We value your feedback to improve our textbook solutions.

Study anywhere. Anytime. Across all devices.