/*! This file is auto-generated */ .wp-block-button__link{color:#fff;background-color:#32373c;border-radius:9999px;box-shadow:none;text-decoration:none;padding:calc(.667em + 2px) calc(1.333em + 2px);font-size:1.125em}.wp-block-file__button{background:#32373c;color:#fff;text-decoration:none} Problem 131 Flying Home for the Holidays, On... [FREE SOLUTION] | 91Ó°ÊÓ

91Ó°ÊÓ

Flying Home for the Holidays, On Time In Exercise 4.115 on page \(302,\) we compared the average difference between actual and scheduled arrival times for December flights on two major airlines: Delta and United. Suppose now that we are only interested in the proportion of flights arriving more than 30 minutes after the scheduled time. Of the 1,000 Delta flights, 67 arrived more than 30 minutes late, and of the 1,000 United flights, 160 arrived more than 30 minutes late. We are testing to see if this provides evidence to conclude that the proportion of flights that are over 30 minutes late is different between flying United or Delta. (a) State the null and alternative hypothesis. (b) What statistic will be recorded for each of the simulated samples to create the randomization distribution? What is the value of that statistic for the observed sample? (c) Use StatKey or other technology to create a randomization distribution. Estimate the p-value for the observed statistic found in part (b). (d) At a significance level of \(\alpha=0.01\), what is the conclusion of the test? Interpret in context. (e) Now assume we had only collected samples of size \(75,\) but got essentially the same proportions (5/75 late flights for Delta and \(12 / 75\) late flights for United). Repeating steps (b) through (d) on these smaller samples, do you come to the same conclusion?

Short Answer

Expert verified
The answer depends on the estimated p-value. If the p-value is less than 0.01, we reject the null hypothesis and conclude that there is a significant difference in late arrival proportions between Delta and United. Otherwise, we do not reject the null hypothesis. The conclusion might change when repeating the process with samples of size 75.

Step by step solution

01

State the hypothesis

The null hypothesis (H0) is that the proportion of flights over 30 minutes late is the same for United and Delta. The alternative hypothesis (H1) is that the proportion is different for the two airlines.
02

Calculate the observed statistic

The statistic in this case is the difference in proportions. For Delta, the proportion is \( \frac{67}{1000} = 0.067 \). For United, the proportion is \( \frac{160}{1000} = 0.16 \). Hence, the observed difference is \( 0.16 - 0.067 = 0.093 \). This statistic will be calculated for each simulated sample.
03

Create a randomization distribution and estimate the p-value

By simulating a large number of samples, one can count the number of times the difference in proportions exceeds the observed statistic of 0.093. Divide this count by the total number of simulated samples to get the p-value. Note that the p-value is estimated using statistical software or tools like StatKey.
04

Interpret the results in regards to the alpha level

If the p-value is less than alpha (\( \alpha = 0.01 \)), reject the null hypothesis and conclude there is evidence of a difference in proportions. Otherwise, do not reject the null hypothesis, meaning the observed difference could be by random chance alone.
05

Repeat steps for samples of size 75

Assume that samples of size 75 were collected and proportions were \( \frac{5}{75} \) for Delta and \( \frac{12}{75} \) for United. Perform steps (b) to (d) again. If the conclusion changes because of the different sample size, it might imply that the sample size affects the power of the test.

Unlock Step-by-Step Solutions & Ace Your Exams!

  • Full Textbook Solutions

    Get detailed explanations and key concepts

  • Unlimited Al creation

    Al flashcards, explanations, exams and more...

  • Ads-free access

    To over 500 millions flashcards

  • Money-back guarantee

    We refund you if you fail your exam.

Over 30 million students worldwide already upgrade their learning with 91Ó°ÊÓ!

Key Concepts

These are the key concepts you need to understand to accurately answer the question.

Null and Alternative Hypothesis
Understanding the null and alternative hypotheses is crucial in hypothesis testing. These hypotheses are mutually exclusive statements about a population parameter, often involving population proportions or means.

In our exercise, the null hypothesis (\(H_0\)) posits that there is no difference in the proportion of flights arriving more than 30 minutes late between Delta and United Airlines. Mathematically, this can be expressed as \(P_{Delta} = P_{United}\). Conversely, the alternative hypothesis (\(H_1\)) suggests that there is indeed a difference, and it's formulated as \(P_{Delta} eq P_{United}\).

The goal of hypothesis testing is to determine which hypothesis the data supports, using a pre-determined significance level to judge whether the observed differences are statistically significant, or if they could have occurred by chance.
Proportions Comparison
When comparing proportions, such as the exercise's percentage of late flights for two airlines, we calculate the difference between the two sample proportions. This difference is pivotal in the hypothesis testing process. For the given data, the proportions of late flights for Delta and United are \(0.067\) and \(0.16\), respectively.

The observed statistic, the difference in these two proportions (\(0.16 - 0.067 = 0.093\)), serves as the basis for comparing what we would expect if the null hypothesis were true to what we actually observe. A significant difference implies that the airlines do not perform equally relative to the specified criterion, prompting further investigation into the statistical significance of this observed difference.
Randomization Distribution
A randomization distribution is a powerful tool in statistics that enables us to understand what kind of sample statistics we might observe if the null hypothesis were true. To create this, we simulate taking many samples from a combined population that assumes the null hypothesis is true, then calculate the statistic of interest for each one.

In our case, the statistic is the difference in late flight proportions. We use computer simulations to produce a distribution of this difference assuming no real difference between the airlines. This randomization distribution forms the backdrop against which we compare our observed statistic, allowing us to make inferences about the likelihood that our observed difference occurred under the null hypothesis.
P-Value Estimation
The p-value is essential in hypothesis testing; it quantifies the probability of observing a statistic as extreme as, or more extreme than, the one calculated from our sample data, assuming that the null hypothesis is true. To estimate the p-value, statistical software or tools like StatKey are typically employed.

Upon creating a randomization distribution, we estimate the p-value by determining the proportion of simulated differences that are as large as or larger than our observed difference of \(0.093\). A small p-value indicates that such an extreme result is unlikely under the null hypothesis, suggesting that the observed difference may be attributable to an actual effect rather than random variation.
Statistical Significance
The concept of statistical significance pertains to the decisiveness of a hypothesis test result. It determines whether the evidence in the sample data is strong enough to reject the null hypothesis. This is done by comparing the p-value to a pre-set alpha level (\(\alpha\)), which is the threshold for statistical significance.

In our exercise, the alpha level is set at \(0.01\), indicating a 1% risk of concluding that there is a difference in late flight proportions when there is none. If the p-value is smaller than this alpha level, we reject the null hypothesis, implying that there is statistically significant evidence of a difference. Conversely, if the p-value is higher, we do not reject the null hypothesis, suggesting that the observed difference could be just a result of sampling variability.

One App. One Place for Learning.

All the tools & learning materials you need for study success - in one app.

Get started for free

Most popular questions from this chapter

Influencing Voters Exercise 4.39 on page 272 describes a possible study to see if there is evidence that a recorded phone call is more effective than a mailed flyer in getting voters to support a certain candidate. The study assumes a significance level of \(\alpha=0.05\) (a) What is the conclusion in the context of thisstudy if the p-value for the test is \(0.027 ?\) (b) In the conclusion in part (a), which type of error are we possibly making: Type I or Type II? Describe what that type of error means in this situation. (c) What is the conclusion if the p-value for the test is \(0.18 ?\)

Interpreting a P-value In each case, indicate whether the statement is a proper interpretation of what a p-value measures. (a) The probability the null hypothesis \(H_{0}\) is true. (b) The probability that the alternative hypothesis \(H_{a}\) is true. (c) The probability of seeing data as extreme as the sample, when the null hypothesis \(H_{0}\) is true. (d) The probability of making a Type I error if the null hypothesis \(H_{0}\) is true. (e) The probability of making a Type II error if the alternative hypothesis \(H_{a}\) is true.

In Exercises 4.40 to 4.44 , null and alternative hypotheses for a test are given. Give the notation \((\bar{x},\) for example) for a sample statistic we might record for each simulated sample to create the randomization distribution. \(H_{0}: p=0.5\) vs \(H_{a}: p \neq 0.5\)

Could owning a cat as a child be related to mental illness later in life? Toxoplasmosis is a disease transmitted primarily through contact with cat feces, and has recently been linked with schizophrenia and other mental illnesses. Also, people infected with Toxoplasmosis tend to like cats more and are 2.5 times more likely to get in a car accident, due to delayed reaction times. The CDC estimates that about \(22.5 \%\) of Americans are infected with Toxoplasmosis (most have no symptoms), and this prevalence can be as high as \(95 \%\) in other parts of the world. A study \(^{37}\) randomly selected 262 people registered with the National Alliance for the Mentally Ill (NAMI), almost all of whom had schizophrenia, and for each person selected, chose two people from families without mental illness who were the same age, sex, and socioeconomic status as the person selected from NAMI. Each participant was asked whether or not they owned a cat as a child. The results showed that 136 of the 262 people in the mentally ill group had owned a cat, while 220 of the 522 people in the not mentally ill group had owned a cat. (a) This is known as a case-control study, where cases are selected as people with a specific disease or trait, and controls are chosen to be people without the disease or trait being studied. Both cases and controls are then asked about some variable from their past being studied as a potential risk factor. This is particularly useful for studying rare diseases (such as schizophrenia), because the design ensures a sufficient sample size of people with the disease. Can casecontrol studies such as this be used to infer a causal relationship between the hypothesized risk factor (e.g., cat ownership) and the disease (e.g., schizophrenia)? Why or why not? (b) In case-control studies, controls are usually chosen to be similar to the cases. For example, in this study each control was chosen to be the same age, sex, and socioeconomic status as the corresponding case. Why choose controls who are similar to the cases? (c) For this study, calculate the relevant difference in proportions; proportion of cases (those with schizophrenia) who owned a cat as a child minus proportion of controls (no mental illness) who owned a cat as a child. (d) For testing the hypothesis that the proportion of cat owners is higher in the schizophrenic group than the control group, use technology to generate a randomization distribution and calculate the p-value. (e) Do you think this provides evidence that there is an association between owning a cat as a child and developing schizophrenia? \(^{38}\) Why or why not?

It is well established that exercise is beneficial for our bodies. Recent studies appear to indicate that exercise can also do wonders for our brains, or, at least, the brains of mice. In a randomized experiment, one group of mice was given access to a running wheel while a second group of mice was kept sedentary. According to an article describing the study, "The brains of mice and rats that were allowed to run on wheels pulsed with vigorous, newly born neurons, and those animals then breezed through mazes and other tests of rodent IQ"9 compared to the sedentary mice. Studies are examining the reasons for these beneficial effects of exercise on rodent (and perhaps human) intelligence. High levels of BMP (bonemorphogenetic protein) in the brain seem to make stem cells less active, which makes the brain slower and less nimble. Exercise seems to reduce the level of BMP in the brain. Additionally, exercise increases a brain protein called noggin, which improves the brain's ability. Indeed, large doses of noggin turned mice into "little mouse geniuses," according to Dr. Kessler, one of the lead authors of the study. While research is ongoing in determining how strong the effects are, all evidence points to the fact that exercise is good for the brain. Several tests involving these studies are described. In each case, define the relevant parameters and state the null and alternative hypotheses. (a) Testing to see if there is evidence that mice allowed to exercise have lower levels of BMP in the brain on average than sedentary mice. (b) Testing to see if there is evidence that mice allowed to exercise have higher levels of noggin in the brain on average than sedentary mice. (c) Testing to see if there is evidence of a negative correlation between the level of BMP and the level of noggin in the brains of mice.

See all solutions

Recommended explanations on Math Textbooks

View all explanations

What do you think about this solution?

We value your feedback to improve our textbook solutions.

Study anywhere. Anytime. Across all devices.