/*! This file is auto-generated */ .wp-block-button__link{color:#fff;background-color:#32373c;border-radius:9999px;box-shadow:none;text-decoration:none;padding:calc(.667em + 2px) calc(1.333em + 2px);font-size:1.125em}.wp-block-file__button{background:#32373c;color:#fff;text-decoration:none} Q56 E Sometimes experiments involving ... [FREE SOLUTION] | 魅影直播

魅影直播

Sometimes experiments involving success or failure responses are run in a paired or before/after manner. Suppose that before a major policy speech by a political candidate, n individuals are selected and asked whether \((S)\)or not (F) they favor the candidate. Then after the speech the same n people are asked the same question. The responses can be entered in a table as follows:

Before

After

S

F

S

\({{\bf{X}}_{\bf{1}}}\)

\({{\bf{X}}_{\bf{2}}}\)

F

\({{\bf{X}}_{\bf{3}}}\)

\({{\bf{X}}_{\bf{4}}}\)

Where\({{\bf{x}}_{\bf{1}}}{\bf{ + }}{{\bf{x}}_{\bf{2}}}{\bf{ + }}{{\bf{x}}_{\bf{3}}}{\bf{ + }}{{\bf{x}}_{\bf{4}}}{\bf{ = n}}\). Let\({{\bf{p}}_{\bf{1}}}{\bf{,}}{{\bf{p}}_{\bf{2}}}{\bf{,}}{{\bf{p}}_{\bf{3}}}\), and \({p_4}\)denote the four cell probabilities, so that \({p_1} = P(S\) before and S after), and so on. We wish to test the hypothesis that the true proportion of supporters (S) after the speech has not increased against the alternative that it has increased.

a. State the two hypotheses of interest in terms of\({p_1},{p_2}\),\({p_3}\), and \({p_4}\).

b. Construct an estimator for the after/before difference in success probabilities

c. When n is large, it can be shown that the rv \(\left( {{{\bf{X}}_{\bf{i}}}{\bf{ - }}{{\bf{X}}_{\bf{j}}}} \right){\bf{/n}}\) has approximately a normal distribution with variance given by\(\left[ {{{\bf{p}}_{\bf{i}}}{\bf{ + }}{{\bf{p}}_{\bf{j}}}{\bf{ - }}{{\left( {{{\bf{p}}_{\bf{i}}}{\bf{ - }}{{\bf{p}}_{\bf{j}}}} \right)}^{\bf{2}}}} \right]{\bf{/n}}\). Use this to construct a test statistic with approximately a standard normal distribution when \({H_0}\)is true (the result is called McNemar's test).

d. If\({{\bf{x}}_{\bf{1}}}{\bf{ = 350,}}\;\;\;{{\bf{x}}_{\bf{2}}}{\bf{ = 150,}}\;\;\;{{\bf{x}}_{\bf{3}}}{\bf{ = 200}}\), and\({x_4} = 300\), what do you conclude?

Short Answer

Expert verified

a. The hypotheses of interest are \({H_0}:{p_3} = {p_2}\)versus \({H_a}:{p_3} > {p_2};\)

b. \(\frac{{{X_3} - {X_2}}}{n};\)

\(c.Z = \frac{{{X_3} - {X_2}}}{{\sqrt {{X_3} + {X_2}} }};\)

d. Reject null hypothesis.

Step by step solution

01

Step 1: State the two hypotheses of interest in terms of\({{\bf{p}}_{\bf{1}}}{\bf{,}}{{\bf{p}}_{\bf{2}}}\), \({{\bf{p}}_{\bf{3}}}\), and \({{\bf{p}}_{\bf{4}}}\).

(a):

The probability that after the speech people favor the major is the sum of \({p_1}\)and \({p_3}\)- both are the corresponding probabilities to cells 1 and 3. The probability that before the speech people do not favor the major is the sum of \({p_1}\)and \({p_2}\)both are the corresponding probabilities to cells 1 and 2. This is because the first column represents success for "after" and the first row represents the row for "before" success.

From\({p_1}\)+\({p_3}\)=\({p_1}\)+\({p_2}\), which would be the null hypothesis,

The hypotheses of interest are \({H_0}:{p_3} = {p_2}\) versus \({H_a}:{p_3} > {p_2}\)

02

Step 2:Construct an estimator for the after/before difference in success probabilities

(b):

Every cell in the table represents number of people for corresponding described event, this indicates that using that data you could estimate the probabilities. There are total of n individuals, and\({x_1} + {x_2} + {x_3} + {x_4} = n\). From this, the estimate of \({p_1} + {p_3} - \left( {{p_1} + {p_2}} \right)\) is simply

\(\frac{{{x_1} + {x_3} - \left( {{x_1} + {x_2}} \right)}}{n} = \frac{{{x_3} - {x_2}}}{n}\)

W hich is obviously the after/before difference in success probabilities estimate. \(\sigma = \sqrt {\frac{{{p_3} + {p_2}}}{n}} ,\)However, the estimator requires random variables and it is\(\frac{{{X_3} - {X_2}}}{n}\)

03

Step 3:to construct a test statistic with approximately a standard normal distribution

(c):

Assume that \({H_0}\)is true, thus\({p_3} = {p_2}\). Based on the given information, the variance of the given estimator in (b) is

\({\mathop{\rm Var}\nolimits} \frac{{{X_3} - {X_2}}}{n} = \frac{{{p_2} + {p_3} - {{\left( {{p_2} - {p_3}} \right)}^2}}}{n}\)

which is, for \({p_3} = {p_2}\)

\({\mathop{\rm Var}\nolimits} \frac{{{X_3} - {X_2}}}{n} = \frac{{{p_3} + {p_2}}}{n}.\)

The standard deviation is which needs to be estimated with\(\hat \sigma = \sqrt {\frac{{{{\hat p}_3} + {{\hat p}_2}}}{n}} \)

Since the expected value is obviously zero, and random variable

\(\frac{{{X_3} - {X_2}}}{n}\)has approximately normal distribution with parameter mean 0 and standard deviation\(\sigma \), you can construct the Z statistic as

\(\begin{array}{c}Z = \frac{{\frac{{{X_3} - {X_2}}}{n}}}{{\sqrt {\frac{{{{\hat p}_3} + {{\hat p}_2}}}{n}} }}\\ = \frac{{{X_3} - {X_2}}}{{\sqrt {{X_3} + {X_2}} }}\end{array}\)

which has standard normal distribution.

04

Step 4:To conclude for the final proof

(d):

the hypotheses of interest are \({H_0}:{p_3} = {p_2}\) versus\({H_a}:{p_3} > {p_2}\), and the value of the Z test statistic is

\(\begin{array}{c}z = \frac{{{X_3} - {X_2}}}{{\sqrt {{X_3} + {X_2}} }}\\ = \frac{{200 - 150}}{{\sqrt {200 + 150} }}\\ = 2.68\end{array}\)

The test is one sided, upper, thus the P value is the area under the standard normal curve to the right of z. Therefore, the P value is

\(\begin{array}{c}P = P(Z > 2.68)\\ = 1 - P(Z \le 2.68)\\ = 1 - \Phi (2.68)\\ = 0.0037\end{array}\)

where the \(\Phi ( \cdot )\)value can be obtained from the table in the appendix. Depending on significance level, the conclusion might change. At significance level 0.01 or 0.05 which are usually used, because

\(P = 0.0037 < 0.01 < 0.05\)we reject the null hypothesis

However, you might not reject in for other significance levels.

Unlock Step-by-Step Solutions & Ace Your Exams!

  • Full Textbook Solutions

    Get detailed explanations and key concepts

  • Unlimited Al creation

    Al flashcards, explanations, exams and more...

  • Ads-free access

    To over 500 millions flashcards

  • Money-back guarantee

    We refund you if you fail your exam.

Over 30 million students worldwide already upgrade their learning with 魅影直播!

One App. One Place for Learning.

All the tools & learning materials you need for study success - in one app.

Get started for free

Most popular questions from this chapter

Many freeways have service (or logo) signs that give information on attractions, camping, lodging, food, and gas services prior to off-ramps. These signs typically do not provide information on distances. The article "Evaluation of Adding Distance Information to Freeway-Specific Service (Logo) Signs" \((J\). of Transp. Engr., 2011: 782-788) reported that in one investigation, six sites along Virginia interstate highways where service signs are posted were selected. For each site, crash data was obtained for a three-year period before distance information was added to the service signs and for a one-year period afterward. The number of crashes per year before and after the sign changes were as follows:

\(\begin{array}{*{20}{l}}{Before:\;\;15\;26\;66\;115\;62\;64}\\{After:\;\;\;\;\;16\;24\;42\;80\;\;78\;73}\end{array}\)

a. The cited article included the statement "A paired\(t\)test was performed to determine whether there was any change in the mean number of crashes before and after the addition of distance information on the signs." Carry out such a test. (Note: The relevant normal probability plot shows a substantial linear pattern.)

b. If a seventh site were to be randomly selected among locations bearing service signs, between what values would you predict the difference in number of crashes to lie?

Refer to Example 9.7. Does the data suggest that the standard deviation of the strength distribution for fused specimens is smaller than that for not-fused specimens? Carry out a test at significance level .01.

Shoveling is not exactly a high-tech activity, but it will continue to be a required task even in our information age. The article "A Shovel with a Perforated Blade Reduces Energy Expenditure Required for Digging Wet Clay" (Human Factors, 2010: 492-502) reported on an experiment in which 13 workers were each provided with both a conventional shovel and a shovel whose blade was perforated with small holes. The authors of the cited article provided the following data on stable energy expenditure ((kcal/kg(subject)//b(clay)):

Worker : 1 2 3 4

Conventional : .0011 .0014 .0018 .0022

Perforated : .0011 .0010 .0019 .0013

Worker: 5 6 7

Conventional : .0010 .0016 .0028

Perforated : .0011 .0017 .0024

Worker 8 9 10

Conventional : .0020 .0015 .0014

Perforated : .0020 .0013 .0013

Worker: 11 12 13

Conventional : .0023 .0017 .0020

Perforated : .0017 .0015 .0013

a. Calculate a confidence interval at the 95 % confidence level for the true average difference between energy expenditure for the conventional shovel and the perforated shovel (the relevant normal probability plot shows a reasonably linear pattern). Based on this interval, does it appear that the shovels differ with respect to true average energy expenditure? Explain.

b. Carry out a test of hypotheses at significance level .05 to see if true average energy expenditure using the conventional shovel exceeds that using the perforated shovel.

An experiment to compare the tension bond strength of polymer latex modified mortar (Portland cement mortar to which polymer latex emulsions have been added during mixing) to that of unmodified mortar resulted in \(\bar x = 18.12kgf/c{m^2}\)for the modified mortar \((m = 40)\) and \(\bar y = 16.87kgf/c{m^2}\) for the unmodified mortar \((n = 32)\). Let \({\mu _1}\) and \({\mu _2}\) be the true average tension bond strengths for the modified and unmodified mortars, respectively. Assume that the bond strength distributions are both normal.

a. Assuming that \({\sigma _1} = 1.6\) and \({\sigma _2} = 1.4\), test \({H_0}:{\mu _1} - {\mu _2} = 0\) versus \({H_a}:{\mu _1} - {\mu _2} > 0\) at level . \(01.\)

b. Compute the probability of a type II error for the test of part (a) when \({\mu _1} - {\mu _2} = 1\).

c. Suppose the investigator decided to use a level \(.05\) test and wished \(\beta = .10\) when \({\mu _1} - {\mu _2} = 1. \). If \(m = 40\), what value of n is necessary?

d. How would the analysis and conclusion of part (a) change if \({\sigma _1}\) and \({\sigma _2}\) were unknown but \({s_1} = 1.6\)and \({s_7} = 1.4?\)

Reliance on solid biomass fuel for cooking and heating exposes many children from developing countries to high levels of indoor air pollution. The article 鈥淒omestic Fuels, Indoor Air Pollution, and Children鈥檚 Health鈥 (Annals of the N.Y. Academy of Sciences, \(2008:209 - 217\)) presented information on various pulmonary characteristics in samples of children whose households in India used either biomass fuel or liquefied petroleum gas (\(LPG\)). For the \(755\) children in biomass households, the sample mean peak expiratory flow (a person鈥檚 maximum speed of expiration) was \(3.30L/s\), and the sample standard deviation was \(1.20\). For the \(750\) children whose households used liquefied petroleum gas, the sample mean \(PEF\) was \(4.25\) and the sample standard deviation was \(1.75\).

a. Calculate a confidence interval at the \(95\% \) confidence level for the population mean \(PEF\) for children in biomass households and then do likewise for children in \(LPG\) households. What is the simultaneous confidence level for the two intervals?

b. Carry out a test of hypotheses at significance level \(.01\) to decide whether true average \(PEF\) is lower for children in biomass households than it is for children in \(LPG\) households (the cited article included a P-value for this test)

c. \(FE{V_1}\), the forced expiratory volume in \(1\) second, is another measure of pulmonary function. The cited article reported that for the biomass households the sample mean FEV1 was \(2.3L/s\) and the sample standard deviation was \(.5L/s\). If this information is used to compute a \(95\% \) \(CI\) for population mean \(FE{V_1}\), would the simultaneous confidence level for this interval and the first interval calculated in (a) be the same as the simultaneous confidence level determined there? Explain

See all solutions

Recommended explanations on Math Textbooks

View all explanations

What do you think about this solution?

We value your feedback to improve our textbook solutions.

Study anywhere. Anytime. Across all devices.