Home / AP® Exam / AP® Statistics / AP Statistics 1.12 Potential Problems with Sampling- Exam Style Questions – FRQs

AP Statistics 1.12 Potential Problems with Sampling- Exam Style Questions - FRQs - New Syllabus

Question

Aphids are tiny insects that feed on plants such as cabbage plants. A farmer wants to reduce the number of aphids in a cabbage field. A river is located 100 meters south of the cabbage field. The farmer divides the field into 25 regions of equal size, as shown in the diagram. Each region has approximately the same number of cabbage plants.
The farmer would like to estimate the proportion of cabbage plants in the field that are affected by aphids and believes that the extent of aphid damage is greater for the regions in the cabbage field closer to the river. To obtain the estimate, the farmer is considering three sampling methods.
  • Sampling method I: Select region 3, which is closest to the farmer’s house and farthest from the river. Examine every cabbage plant in the region for aphid damage.
  • Sampling method II: Randomly select one row (A, B, C, D, or E). For every region in the selected row, examine every cabbage plant for aphid damage.
  • Sampling method III: Randomly select one region from each of rows A, B, C, D, and E. For each selected region, examine every cabbage plant for aphid damage.
A. Explain whether sampling method I is an appropriate sampling method for the farmer to use to estimate the proportion of cabbage plants in the field that are damaged by aphids.
B. Using sampling method II, the farmer randomly selected row E and examined every cabbage plant in row E. If the farmer’s belief is correct, determine whether the selection of row E is likely to provide an overestimate or an underestimate of the proportion of cabbage plants in the field that are damaged by aphids. Justify your answer.
C. Using the information provided in the diagram of the cabbage field, describe how to implement sampling method III, which requires a random selection of one region from each of rows A, B, C, D, and E.

Most-appropriate topic codes (AP Statistics):

• Topic \(1.11\) — Random Sampling (Parts \( \mathrm{A} \), \( \mathrm{B} \), \( \mathrm{C} \))
• Topic \(1.12\) — Potential Problems with Sampling (Parts \( \mathrm{A} \), \( \mathrm{B} \))
▶️ Answer/Explanation

A.
• Sampling method I is not an appropriate sampling method because it uses a convenience sample that lacks randomization.
• Since region 3 is located farthest from the river where aphid damage is believed to be lowest, this region is not representative of the whole field and will likely lead to an underestimate of the true population proportion.

B.
• The selection of row E is likely to provide an overestimate of the true proportion of damaged cabbage plants.
• Row E is the row positioned closest to the river, meaning every single plant checked in this sample belongs to the high-risk zone where the farmer expects aphid damage to be at its peak concentration.

C.
• Label the 5 individual regions within row A with unique identifiers from 1 to 5.
• Use a random number generator or draw numbered slips from a hat to select one single region from row A.
• Repeat this exact independent drawing process for row B (regions 6 to 10), row C (regions 11 to 15), row D (regions 16 to 20), and row E (regions 21 to 25) to complete a stratified sample containing exactly 5 distinct regions.

Question

Researchers will conduct a year-long investigation of walking and cholesterol levels in adults. They will select a random sample of \(100\) adults from the target population to participate as subjects in the study.
(a) One aspect of the study is to record the number of miles each subject walks per day. The researchers are deciding whether to have subjects wear an activity tracker to record the data or to have subjects keep a daily journal of the miles they walk each day. Describe what bias could be introduced by keeping the daily journal instead of wearing the activity tracker.
During the course of the study, the subjects will have their cholesterol levels measured each month by a doctor. The researchers will perform a significance test at the end of the study to determine whether the average cholesterol level for subjects who walk fewer miles each day is greater than for those who walk more miles each day.
(b) Selecting a random sample creates a reasonable representative sample of the target population. Explain the benefit of using a representative sample from the population.
(c) Suppose the researchers conduct the test and find a statistically significant result. Would it be valid to claim that increased walking causes a decrease in average cholesterol levels for adults in the target population? Explain your reasoning.

Most-appropriate topic codes (AP Statistics):

• Topic \(1.11\) — Random Sampling (Parts \( \mathrm{b} \), \( \mathrm{c} \))
• Topic \(1.12\) — Potential Problems with Sampling (Part \( \mathrm{a} \))
▶️ Answer/Explanation

(a)
Keeping a daily journal introduces response bias because subjects are self-reporting. They might forget to log short walks or underestimate their distances, which would cause our sample’s average daily miles to be systematically biased too low compared to what they actually walked.

(b)
A representative sample is essential because it allows us to generalize our findings. By selecting randomly, our \(100\) adults look like the larger target population, meaning any inference we make about the relationship between walking and cholesterol isn’t skewed by an unrepresentative group.

(c)
No, we can’t claim causation here because this is an observational study, not an experiment. We didn’t randomly assign subjects to walk specific distances. Because of this lack of assignment, there could be confounding variables at play—like a person’s diet or general health consciousness—that simultaneously affect both how much they walk and their cholesterol levels.

Question

Emma is moving to a large city and is investigating typical monthly rental prices of available one-bedroom apartments. She obtained a random sample of rental prices for \(50\) one-bedroom apartments taken from a Web site where people voluntarily list available apartments.
(a) Describe the population for which it is appropriate for Emma to generalize the results from her sample.
The distribution of the \(50\) rental prices of the available apartments is shown in the following histogram.
(b) Emma wants to estimate the typical rental price of a one-bedroom apartment in the city. Based on the distribution shown, what is a disadvantage of using the mean rather than the median as an estimate of the typical rental price?
(c) Instead of using the sample median as the point estimate for the population median, Emma wants to use an interval estimate. However, computing an interval estimate requires knowing the sampling distribution of the sample median for samples of size \(50\). Emma has one point, her sample median, in that sampling distribution.
Using information about rental prices that are available on the Web site, describe how someone could develop a theoretical sampling distribution of the sample median for samples of size \(50\).
Because Emma does not have the resources to develop the theoretical sampling distribution, she estimates the sampling distribution of the sample median using a process called bootstrapping. In the bootstrapping process, a computer program performs the following steps.
• Take a random sample, with replacement, of size \(50\) from the original sample.
• Calculate and record the median of the sample.
• Repeat the process to obtain a total of \(15,000\) medians.
Emma ran the bootstrap process, and the following frequency table is the bootstrap distribution showing her results of generating \(15,000\) medians.
The bootstrap distribution provides an approximation of the sampling distribution of the sample median. A confidence interval for the median can be constructed using a percentage of the values in the middle of the bootstrap distribution.
(d) Use the frequency table to find the following.
i. Value of the \(5\text{th}\) percentile:
ii. Value of the \(95\text{th}\) percentile:
(e) Find the percentage of bootstrap medians in the table that are equal to or between the values found in part (d).
(f) Use your values from parts (d) and (e) to construct and interpret a confidence interval for the median rental price.

Most-appropriate topic codes (AP Statistics):

• Topic \(1.11\) — Random Sampling (Part \( \mathrm{a} \))
• Topic \(1.12\) — Potential Problems with Sampling (Part \( \mathrm{a} \))
• Topic \(1.7\) — Summary Statistics for One Quantitative Variable (Parts \( \mathrm{b} \), \( \mathrm{d} \), \( \mathrm{e} \))
• Topic \(2.12\) — Central Limit Theorem and Sampling Distributions for Sample Means (Parts \( \mathrm{c} \), \( \mathrm{f} \))
▶️ Answer/Explanation

(a)
Because a random sample was used, it is appropriate to generalize the results to the population of all rental prices for one-bedroom apartments in the city that are listed on this particular website at the time the sample was taken.

(b)
The histogram indicates that the distribution of rental prices is strongly skewed to the right.
Because of this right skewness, the sample mean will be pulled substantially higher by a few very large rental prices.
Consequently, the sample mean would overestimate the typical rental price, whereas the sample median provides a more accurate representation of the center.

(c)
To develop a theoretical sampling distribution, Emma would need to obtain every possible sample of size \(50\) from the population of apartments listed on the website.
She would then compute the median rental price for each of these possible samples.
The entire collection of all these computed sample medians forms the theoretical sampling distribution for the sample median.

(d)
i. The \(5\text{th}\) percentile corresponds to the \((0.05)(15,000) = 750\text{th}\) value in the ordered data.
By keeping a running cumulative total of the frequencies down the columns, the \(750\text{th}\) value first falls in the row for the median of \(\$2,500\).
ii. The \(95\text{th}\) percentile corresponds to the \((0.95)(15,000) = 14,250\text{th}\) value in the ordered data.
Continuing to add the frequencies, the \(14,250\text{th}\) value falls in the row for the median of \(\$2,950\).

(e)
We need to sum the frequencies of all bootstrap medians from \(\$2,500\) to \(\$2,950\) inclusive.
\(\text{Sum} = 1,899 + 2 + \dots + 10 + 700 = 14,404\)
\(\text{Percentage} = \left(\dfrac{14,404}{15,000}\right) \times 100\% \approx 96.03\%\)

(f)
Based on our previous parts, an approximate \(96\%\) confidence interval for the median is \((\$2,500, \$2,950)\).
Interpretation: We are approximately \(96\%\) confident that the true median rental price of all one-bedroom apartments listed on this website for this city is between \(\$2,500\) and \(\$2,950\).

Question

An environmental science teacher at a high school with a large population of students wanted to estimate the proportion of students at the school who regularly recycle plastic bottles. The teacher selected a random sample of students at the school to survey. Each selected student went into the teacher’s office, one at a time, and was asked to respond yes or no to the following question.
Do you regularly recycle plastic bottles?
Based on the responses, a $95$ percent confidence interval for the proportion of all students at the school who would respond yes to the question was calculated as $(0.584, 0.816)$.
(a) How many students were in the sample selected by the environmental science teacher?
(b) Given the method used by the environmental science teacher to collect the responses, explain how bias might have been introduced and describe how the bias might affect the point estimate of the proportion of all students at the school who would respond yes to the question.
(c) The statistics teacher at the high school was concerned about the potential bias in the survey. To obtain a potentially less biased estimate of the proportion, the statistics teacher used an alternate method for collecting student responses. A random sample of $300$ students was selected, and each student was given the following instructions on how to respond to the question.
In private, flip a fair coin.
If heads, you must respond no, regardless of whether you regularly recycle. If tails, please truthfully respond yes or no.
i. What is the expected number of students from the sample of $300$ who would be required to respond no because the coin flip resulted in heads?
ii. The results of the sample showed that $213$ of the $300$ selected students responded no. Based on the results of the sample, give a point estimate for the proportion of all students at the high school who would respond yes to the question.

Most-appropriate topic codes (AP Statistics):

• Topic \(1.12\) — Potential Problems with Sampling (Part \( \mathrm{b} \))
• Topic \(2.6\) — Conditional Probability (Part \( \mathrm{c}\text{-}\mathrm{ii} \))
• Topic \(2.8\) — Introduction to Random Variables and Probability Distributions (Part \( \mathrm{c}\text{-}\mathrm{i} \))
• Topic \(3.3\) — Constructing a Confidence Interval for a Population Proportion (Part \( \mathrm{a} \))
▶️ Answer/Explanation

(a)
The sample selected by the environmental science teacher contained $60$ students.
Detailed Solution:
First, find the point estimate $\hat{p}$ which is the midpoint of the confidence interval: $\hat{p} = \frac{0.584 + 0.816}{2} = 0.70$.
Next, determine the margin of error ($ME$) by calculating the distance from the midpoint to an endpoint: $ME = 0.816 – 0.70 = 0.116$.
Using the margin of error formula $ME = z^* \sqrt{\frac{\hat{p}(1-\hat{p})}{n}}$ for a $95\%$ confidence level ($z^* = 1.96$), we set up the equation $0.116 = 1.96 \sqrt{\frac{0.70(1-0.70)}{n}}$.
Solving for $n$ gives us $\sqrt{n} = \frac{1.96 \sqrt{0.21}}{0.116} \approx 7.74$, which squares to $n \approx 59.9$, revealing that the teacher’s sample size is exactly $60$ students.

(b)
Bias might have been introduced because students were asked directly by their environmental science teacher, which likely creates response bias.
Detailed Solution:
Because the survey is conducted face-to-face by a teacher who is expected to care about the environment, students may feel strong social pressure to give the “desirable” answer.
This phenomenon is known as response bias, where respondents do not answer truthfully in order to avoid judgment or please the interviewer.
As a result, more students will claim they recycle than actually do, artificially inflating the number of “yes” responses and causing the point estimate to be higher than the true population proportion.

(c)(i)
The expected number of students required to respond “no” due to the coin flip is $150$.
Detailed Solution:
Since the students are flipping a fair coin, the theoretical probability of getting heads is exactly $0.5$.
With a total random sample of $n = 300$ students, the expected number of heads is calculated as $n \times p = 300 \times 0.5$.
Therefore, we can expect exactly half the students, or $150$, to be forced to respond “no” based on the coin flip instructions.

(c)(ii)
The point estimate for the proportion of all students at the high school who would respond “yes” is $0.58$.
Detailed Solution:
Out of the $300$ total students, $213$ responded “no”, and we expect $150$ of these “no” responses to come from the students who flipped heads.
This means the remaining $213 – 150 = 63$ “no” responses came from the $150$ students who flipped tails and answered truthfully about not recycling.
Since $150$ students flipped tails and $63$ of them truthfully said “no”, the remaining $150 – 63 = 87$ students must have truthfully answered “yes”.
Thus, the point estimate for the proportion of students who actually recycle is $\frac{87}{150} = 0.58$.

Question

Alzheimer’s disease results in a loss of cognitive ability beyond what is expected with typical aging. A local newspaper published an article with the following headline.
Study Finds Strong Association Between Smoking and Alzheimer’s
The article reported that a study tracked the medical histories of 21,123 men and women for 23 years. The article stated that, for those who smoked at least two packs of cigarettes a day, the risk of developing Alzheimer’s disease was 2.57 times the risk for those who did not smoke.
(a) Identify the explanatory and response variables in the study.
Explanatory variable:
Response variable:
(b) Is the study described in the article an observational study or an experiment? Explain.
(c) Exercise status (regular weekly exercise versus no regular weekly exercise) was mentioned in the article as a possible confounding variable. Explain how exercise status could be a confounding variable in the study.

Most-appropriate topic codes (AP Statistics):

• Topic \(1.10\) — The Investigative Question Revisited and Data Collection (Parts \( \mathrm{a} \), \( \mathrm{b} \))
• Topic \(1.11\) — Random Smapling (Part \( \mathrm{b} \))
• Topic \(1.12\) — Potential Problems with Sampling (Part \( \mathrm{c} \))
▶️ Answer/Explanation

(a)

Explanatory variable: The person’s degree of cigarette smoking — specifically, whether or not the individual smoked at least two packs of cigarettes per day (smoker of at least two packs per day versus non-smoker at that level).
Response variable: Whether or not the person developed Alzheimer’s disease during the course of the 23-year study.

(b)

This is an observational study, not an experiment. In an experiment, researchers would have to actively assign the treatment — in this case, the level of cigarette smoking — to the participants. Instead, the researchers simply tracked the existing medical histories of 21,123 men and women over 23 years. The smoking status of each person was passively observed and recorded, not controlled or manipulated by the researchers. Since no treatment was imposed, this is an observational study.

(c)

A confounding variable is one that is related to the explanatory variable and also independently influences the response variable, making it difficult to determine whether the explanatory variable alone is responsible for the observed association.
Exercise status could be a confounding variable here for two reasons working together:
First, people who exercise regularly tend to be more health-conscious overall, and as a result are less likely to smoke heavily. So exercise status is related to the explanatory variable — smoking status. Heavy smokers are, on average, less likely to exercise regularly than non-smokers.
Second, regular exercise may independently reduce the risk of developing Alzheimer’s disease. So exercise status is also related to the response variable — development of Alzheimer’s disease.
Because of both of these relationships, the observed association between heavy smoking and higher rates of Alzheimer’s could be at least partly explained by the fact that heavy smokers tend to exercise less — and it is the lack of exercise, not the smoking itself, that contributes to the increased risk. This makes it impossible to determine from this study alone whether the association between smoking and Alzheimer’s reflects a true causal relationship or is merely the result of exercise status being linked to both.

Question

As part of its twenty-fifth reunion celebration, the class of 1988 (students who graduated in 1988) at a state university held a reception on campus. In an informal survey, the director of alumni development asked 50 of the attendees about their incomes. The director computed the mean income of the 50 attendees to be $189,952. In a news release, the director announced, “The members of our class of 1988 enjoyed resounding success. Last year’s mean income of its members was $189,952!”
(a) What would be a statistical advantage of using the median of the reported incomes, rather than the mean, as the estimate of the typical income?
(b) The director felt the members who attended the reception may be different from the class as a whole. A more detailed survey of the class was planned to find a better estimate of the income as well as other facts about the alumni. The staff developed two methods based on the available funds to carry out the survey.
Method 1: Send out an e-mail to all 6,826 members of the class asking them to complete an online form. The staff estimates that at least 600 members will respond.
Method 2: Select a simple random sample of members of the class and contact the selected members directly by phone. Follow up to ensure that all responses are obtained. Because method 2 will require more time than method 1, the staff estimates that only 100 members of the class could be contacted using method 2.
Which of the two methods would you select for estimating the average yearly income of all 6,826 members of the class of 1988? Explain your reasoning by comparing the two methods and the effect of each method on the estimate.

Most-appropriate topic codes (AP Statistics):

• Topic \(1.7\) — Summary Statistics for One Quantitative Variable (Part \( \mathrm{a} \))
• Topic \(1.12\) — Potential Biases in Sampling Methods (Part \( \mathrm{b} \))
▶️ Answer/Explanation

(a)

The median is a better measure of typical income because it is resistant to skewness and outliers, while the mean is not. Income distributions tend to be right-skewed — a small number of very high earners can pull the mean far above what most people actually earn, making the mean an inflated and misleading estimate of the typical income. The median, by contrast, simply reflects the middle value and is not distorted by a few extremely large incomes.

(b)

Method 2 is the better choice.

Method 1 relies on voluntary response — members choose whether or not to reply to the e-mail. This introduces voluntary response bias: alumni with higher incomes are likely more motivated to respond (to show their success), while alumni with lower incomes may be less likely to reply. As a result, the sample from Method 1 would not be representative of the entire class, and the estimated mean income would be inflated — higher than the true mean income of all 6,826 members.

Method 2, despite its smaller sample size of 100, uses a simple random sample with guaranteed follow-up to ensure all selected members respond. Random selection makes the sample much more representative of the full class, producing an approximately unbiased estimate of the true average yearly income. A smaller but unbiased sample is far preferable to a larger but systematically biased one.

Question

An administrator at a large university wants to conduct a survey to estimate the proportion of students who are satisfied with the appearance of the university buildings and grounds. The administrator is considering three methods of obtaining a sample of \(500\) students from the \(70{,}000\) students at the university.
(a) Because of financial constraints, the first method the administrator is considering consists of taking a convenience sample to keep the expenses low. A very large number of students will attend the first football game of the season, and the first \(500\) students who enter the football stadium could be used as a sample. Why might such a sampling method be biased in producing an estimate of the proportion of students who are satisfied with the appearance of the buildings and grounds?
(b) Because of the large number of students at the university, the second method the administrator is considering consists of using a computer with a random number generator to select a simple random sample of \(500\) students from a list of \(70{,}000\) student names. Describe how to implement such a method.
(c) Because stratification can often provide a more precise estimate than a simple random sample, the third method the administrator is considering consists of selecting a stratified random sample of \(500\) students. The university has two campuses with male and female students at each campus. Under what circumstance(s) would stratification by campus provide a more precise estimate of the proportion of students who are satisfied with the appearance of the university buildings and grounds than stratification by gender?

Most-appropriate topic codes (AP Statistics):

• Topic \(1.12\) — Potential Problems with Sampling (Part \( \mathrm{a} \))
• Topic \(1.11\) — Random Sampling (Part \( \mathrm{b} \))
• Topic \(1.11\) — Random Sampling (Part \( \mathrm{c} \) — stratified random sampling)
▶️ Answer/Explanation

(a)

The first \(500\) students who enter the football stadium are not likely to be representative of all \(70{,}000\) students at the university. Students who attend football games tend to have stronger school pride and a more positive connection to the university overall, which likely makes them more satisfied with the appearance of the buildings and grounds than the general student population. This means the sample would systematically overestimate the proportion of students who are satisfied — producing a positively biased estimate.

(b)

Step 1: Obtain a complete list of all \(70{,}000\) students at the university and assign each student a unique identification number from \(1\) to \(70{,}000\).
Step 2: Use a computer random number generator to generate \(500\) distinct random integers between \(1\) and \(70{,}000\), ignoring any repeated numbers that appear.
Step 3: Select the students whose assigned ID numbers match the \(500\) randomly generated numbers — these students form the simple random sample.

(c)

Stratification by campus would provide a more precise estimate than stratification by gender when the variability in students’ opinions about the appearance of the university buildings and grounds is greater between the two campuses than between the two genders. In other words, if the two campuses differ noticeably from each other in appearance (for example, one campus is newer or more attractive than the other), then students on different campuses will tend to have systematically different satisfaction levels — making campus a more effective stratification variable than gender.

Question

Before beginning a unit on frog anatomy, a seventh-grade biology teacher gives each of the 24 students in the class a pretest to assess their knowledge of frog anatomy. The teacher wants to compare the effectiveness of an instructional program in which students physically dissect frogs with the effectiveness of a different program in which students use computer software that only simulates the dissection of a frog. After completing one of the two programs, students will be given a posttest to assess their knowledge of frog anatomy. The teacher will then analyze the changes in the test scores (score on posttest minus score on pretest).
(a) Describe a method for assigning the 24 students to two groups of equal size that allows for a statistically valid comparison of the two instructional programs.
(b) Suppose the teacher decided to allow the students in the class to select which instructional program on frog anatomy (physical dissection or computer simulation) they prefer to take, and 11 students choose actual dissection and 13 students choose computer simulation. How might that self-selection process jeopardize a statistically valid comparison of the changes in the test scores (score on posttest minus score on pretest) for the two instructional programs? Provide a specific example to support your answer.

Most-appropriate topic codes (AP Statistics):

• Topic \(1.13\) — Experimental Design (Part \(\mathrm{a}\))
• Topic \(1.13\) — Experimental Design (Part \(\mathrm{b}\))
• Topic \(1.12\) — Potential Problems with Sampling (Part \(\mathrm{b}\))

▶️ Answer/Explanation

(a) — Completely Randomized Design
Assign each of the 24 students a unique two-digit number from \(01\) to \(24\).
Use a random number table or a random number generator to produce a sequence of two-digit numbers from \(01\) to \(24\), ignoring repeats and any numbers outside that range.
The first 12 distinct numbers that appear correspond to the students assigned to the physical dissection program; the remaining 12 students are assigned to the computer simulation program.
This ensures both groups are of equal size (\(n = 12\) each) and that assignment is governed entirely by chance, removing any systematic differences between the groups before the study begins.

Alternative: Randomized Block Design
Rank all 24 students from lowest to highest pretest score.
Form 12 blocks of 2 students each: Block 1 contains the two students with the lowest pretest scores, Block 2 the next two, and so on, with Block 12 containing the two students with the highest pretest scores.
Within each block, randomly assign one student to the physical dissection program and the other to the computer simulation program — for example, by flipping a fair coin or using a random number generator to decide which student in each pair gets which treatment.
This design controls for prior knowledge of frog anatomy (as measured by the pretest), making the comparison between treatments more precise.

(b)
When students self-select into groups, the two groups may differ systematically in ways that affect posttest performance — completely apart from which instructional method they received.
For example, suppose students who already know a great deal about frog anatomy tend to be the ones who choose the physical dissection program, because they are enthusiastic about frogs and eager to work with them hands-on. Since these students enter with higher prior knowledge, there is less room for them to improve between the pretest and the posttest, so their score changes (posttest \(-\) pretest) will tend to be smaller.
Meanwhile, the students who choose the computer simulation tend to know less about frog anatomy to begin with, so they have more room to improve, and their score changes will tend to be larger.
If the computer simulation group shows a larger average improvement, we cannot tell whether the simulation program itself is more effective or whether the difference simply reflects the fact that lower-knowledge students had more room to grow. The self-selection has introduced a confounding variable (prior knowledge of frog anatomy) that makes it impossible to isolate the effect of the instructional method alone.

Question

A local school board plans to conduct a survey of parents’ opinions about year-round schooling in elementary schools. The school board obtains a list of all families in the district with at least one child in an elementary school and sends the survey to a random sample of 500 of the families. The survey question is provided below.
A proposal has been submitted that would require students in elementary schools to attend school on a year-round basis. Do you support this proposal? (Yes or No)
The school board received responses from 98 of the families, with 76 of the responses indicating support for year-round schools. Based on this outcome, the local school board concludes that most of the families with at least one child in elementary school prefer year-round schooling.
(a) What is a possible consequence of nonresponse bias for interpreting the results of this survey?
(b) Someone advised the local school board to take an additional random sample of 500 families and to use the combined results to make their decision. Would this be a suitable solution to the issue raised in part (a)? Explain.
(c) Suggest a different follow-up step from the one suggested in part (b) that the local school board could take to address the issue raised in part (a).

Most-appropriate topic codes (AP Statistics):

• Topic \(1.11\) — Random Sampling (Parts \(\mathrm{a}\), \(\mathrm{b}\), \(\mathrm{c}\))
• Topic \(1.12\) — Potential Problems with Sampling (Parts \(\mathrm{a}\), \(\mathrm{b}\), \(\mathrm{c}\))
▶️ Answer/Explanation

(a)

Only 98 out of 500 families responded — that is a response rate of just \(\dfrac{98}{500} = 19.6\%\), meaning \(80.4\%\) did not reply at all.
For the survey results to be unbiased, we would need the 402 non-responding families to have similar opinions to those who did respond. But that is unlikely to be true — families who feel strongly in favour of year-round schooling may be much more motivated to respond, while those who oppose or feel indifferent may simply ignore the survey.
As a result, the observed proportion \(\dfrac{76}{98} \approx 77.6\%\) in favour likely overestimates the true proportion of all families who support the proposal. The school board’s conclusion that “most families prefer year-round schooling” may therefore be misleading.

(b)

No, taking an additional random sample of 500 families and combining the results would not be a suitable solution.
• The core problem is not the size of the sample — it is the low response rate in the original survey.
• The original 402 non-responses represent a biased portion of the data that will still be present in the combined sample, regardless of how the second sample turns out.
• If the second survey is conducted in the same way (mailed questionnaire with no follow-up), it is very likely to suffer from the same nonresponse bias, leaving the combined results just as unreliable.
Simply increasing the number of people surveyed does not fix bias — it only gives you a larger biased sample.

(c)

The school board should directly contact the 402 families who did not respond to the original survey, using a different mode of communication such as telephone calls or in-person visits, to obtain their opinions.
• This targets the exact group whose missing responses are causing the bias.
• Combining these newly obtained responses with the original 98 would give a much more complete and representative picture of opinion across all families in the district.
Alternatively, the board could design a new survey from scratch using in-person interviews or telephone calls to ensure a much higher response rate from the outset, reducing the opportunity for nonresponse bias to arise.

Question

The goal of a nutritional study was to compare the caloric intake of adolescents living in rural areas of the United States with the caloric intake of adolescents living in urban areas of the United States. A random sample of ninth-grade students from one high school in a rural area was selected. Another random sample of ninth graders from one high school in an urban area was also selected. Each student in each sample kept records of all the food he or she consumed in one day.
The back-to-back stemplot below displays the number of calories of food consumed per kilogram of body weight for each student on that day.
(a) Write a few sentences comparing the distribution of the daily caloric intake of ninth-grade students in the rural high school with the distribution of the daily caloric intake of ninth-grade students in the urban high school.
(b) Is it reasonable to generalize the findings of this study to all rural and urban ninth-grade students in the United States? Explain.
(c) Researchers who want to conduct a similar study are debating which of the following two plans to use.
Plan I: Have each student in the study record all the food he or she consumed in one day. Then researchers would compute the number of calories of food consumed per kilogram of body weight for each student for that day.
Plan II: Have each student in the study record all the food he or she consumed over the same 7-day period. Then researchers would compute the average daily number of calories of food consumed per kilogram of body weight for each student during that 7-day period.
Assuming that the students keep accurate records, which plan, I or II, would better meet the goal of the study? Justify your answer.

Most-appropriate topic codes (AP Statistics):

• Topic 1.9 — Comparisons of the Distributions for One Quantitative Variable (Part \(\mathrm{a}\))
• Topic 1.11 — Random Sampling (Part \(\mathrm{b}\))
• Topic 1.12 — Potential Problems with Sampling (Part \(\mathrm{b}\))
• Topic 1.10 — The Investigative Question Revisited and Data Collection (Part \(\mathrm{c}\))
▶️ Answer/Explanation

(a)
Reading the stemplot, the rural distribution is centered higher and is more spread out than the urban distribution.

For the rural students:

Mean \(\approx 40.45\) cal/kg
Median \(\approx 41\) cal/kg
Range \(= 19\)
SD \(\approx 6.04\)
IQR \(\approx 10\)

For the urban students:

Mean \(\approx 32.6\) cal/kg
Median \(\approx 32\) cal/kg
Range \(= 16\)
SD \(\approx 4.67\)
IQR \(\approx 7\)

So both the typical value and the spread are larger for the rural group. In terms of shape, the rural data look fairly symmetric and spread evenly between about 32 and 51 cal/kg, while the urban data appear skewed toward the larger values.

\( \boxed{\text{Rural: higher center and more spread; Urban: lower center, less spread, right-skewed}} \)

(b)
No. Each sample came from just one rural school and one urban school, so these two specific schools may not represent the much larger and more diverse population of all rural and urban ninth graders across the country. Because the schools themselves were not randomly chosen from all such schools, the results can’t be safely extended beyond these two schools.

\( \boxed{\text{No — only one school of each type was sampled, so results cannot be generalized nationally}} \)

(c)
Plan II is the better choice.

Both plans already adjust for body size by dividing calories by body weight, so that part is the same. The real issue is that a single day’s eating can be unusually high or low depending on what happened that day — a birthday party, a sick day, a weekend versus a school day, and so on. By recording food over a full 7-day period and averaging, Plan II smooths out this day-to-day variability and gives a more stable, precise picture of each student’s typical caloric intake.

\( \boxed{\text{Plan II — averaging over 7 days reduces day-to-day variability and gives a more precise estimate}} \)

Question

A survey will be conducted to examine the educational level of adult heads of households in the United States. Each respondent in the survey will be placed into one of the following two categories:
• Does not have a high school diploma
• Has a high school diploma
The survey will be conducted using a telephone interview. Random-digit dialing will be used to select the sample.
(a) For this survey, state one potential source of bias and describe how it might affect the estimate of the proportion of adult heads of households in the United States who do not have a high school diploma.
(b) A pilot survey indicated that about 22 percent of the population of adult heads of households do not have a high school diploma. Using this information, how many respondents should be obtained if the goal of the survey is to estimate the proportion of the population who do not have a high school diploma to within \(0.03\) with \(95\) percent confidence? Justify your answer.
(c) Since education is largely the responsibility of each state, the agency wants to be sure that estimates are available for each state as well as for the nation. Identify a sampling method that will achieve this additional goal and briefly describe a way to select the survey sample using this method.

Most-appropriate topic codes (AP Statistics):

• Topic 1.12 — Potential Problems with Sampling (Part \(\mathrm{a}\))
• Topic 3.3 — Constructing a Confidence Interval for a Population Proportion (Part \(\mathrm{b}\))
• Topic 1.11 — Random Sampling (Part \(\mathrm{c}\))
▶️ Answer/Explanation

(a)
One issue is that random-digit dialing only reaches people who have a telephone. People without a high school diploma tend to have lower-paying jobs, so they may be less likely to be able to afford phone service. As a result, this group could be underrepresented in the sample.
\( \boxed{\text{Households without phones are missed, and these are more likely to lack a diploma, so the estimate may be too low}} \)

(b)
The margin of error for a proportion is
\( ME=z^*\sqrt{\dfrac{p(1-p)}{n}} \)
We want \(ME=0.03\) with \(95\%\) confidence, so \(z^*=1.96\), and we use the pilot estimate \(p=0.22\):
\( 0.03=1.96\sqrt{\dfrac{0.22(0.78)}{n}} \)
Solving for \(n\), first isolate the square root:
\( \sqrt{\dfrac{0.22(0.78)}{n}}=\dfrac{0.03}{1.96} \)
Square both sides:
\( \dfrac{0.22(0.78)}{n}=\left(\dfrac{0.03}{1.96}\right)^2 \)
Solve for \(n\):
\( n=\dfrac{0.22(0.78)}{\left(\dfrac{0.03}{1.96}\right)^2} \)
\( n=\left(\dfrac{1.96}{0.03}\right)^2(0.22)(0.78) \)
\( n\approx 732.47 \)
Since \(n\) must be a whole number and we need at least this many respondents, we round up:
\( \boxed{n=733\text{ respondents}} \)

(c)
A good approach here is stratified random sampling, using each state as a stratum. Within every state, a random sample of adult heads of households would be selected and surveyed, with the sample size in each state chosen based on the precision needed for that state. Once all the state-level samples are collected, the results can be combined to produce an overall national estimate.
\( \boxed{\text{Stratified random sampling — treat each state as a stratum and take a random sample within each state}} \)

Question

At an archaeological site that was an ancient swamp, the bones from 20 brontosaur skeletons have been unearthed. The bones do not show any sign of disease or malformation. It is thought that these animals wandered into a deep area of the swamp and became trapped in the swamp bottom. The 20 left femur bones (thigh bones) were located and 4 of these left femurs are to be randomly selected without replacement for DNA testing to determine gender.
(a) Let \(X\) be the number out of the 4 selected left femurs that are from males. Based on how these bones were sampled, explain why the probability distribution of \(X\) is not binomial.
(b) Suppose that the group of 20 brontosaurs whose remains were found in the swamp had been made up of 10 males and 10 females. What is the probability that all 4 in the sample to be tested are male?
(c) The DNA testing revealed that all 4 femurs tested were from males. Based on this result and your answer from part (b), do you think that males and females were equally represented in the group of 20 brontosaurs stuck in the swamp? Explain.
(d) Is it reasonable to generalize your conclusion in part (c) pertaining to the group of 20 brontosaurs to the population of all brontosaurs? Explain why or why not.

Most-appropriate topic codes (AP Statistics):

• Topic 2.10 — The Binomial Distribution (Part a)
• Topic 2.7 — Independent Events and Unions of Events (Part a,b)
• Topic 2.4 — Introduction to Probability (Part b)
• Topic 2.3 — Estimating Probabilities Using Simulation (Part c)
Topic 1.12 — Potential Problems with Sampling (Part d)
▶️ Answer/Explanation

(a)
The probability distribution of \(X\) is not binomial because the bones are selected without replacement from a finite population of only 20 femurs. For a binomial distribution to apply, each trial must be independent — that is, the probability of success (selecting a male femur) must remain constant from one draw to the next. However, when sampling without replacement, the composition of the remaining pool changes with each selection, so the probability of drawing a male femur on each successive draw depends on what was drawn before it. Since the trials are not independent and the probability of success is not fixed, the distribution of \(X\) is hypergeometric, not binomial.
\(\boxed{X \text{ is not binomial because sampling is without replacement, making trials dependent}}\)

(b)
With 10 males and 10 females among the 20 brontosaurs, compute the probability that all 4 selected femurs are male using the multiplication rule for dependent events (without replacement):
\(P(\text{1st is male}) = \dfrac{10}{20}\)
\(P(\text{2nd is male} \mid \text{1st is male}) = \dfrac{9}{19}\)
\(P(\text{3rd is male} \mid \text{first two are male}) = \dfrac{8}{18}\)
\(P(\text{4th is male} \mid \text{first three are male}) = \dfrac{7}{17}\)
Therefore:
\(P(\text{all 4 are male}) = \dfrac{10}{20} \times \dfrac{9}{19} \times \dfrac{8}{18} \times \dfrac{7}{17}\)
\(= \dfrac{10 \times 9 \times 8 \times 7}{20 \times 19 \times 18 \times 17} = \dfrac{5040}{116280} \approx 0.0433\)
This can also be expressed using combinations:
\(P(\text{all 4 are male}) = \dfrac{\dbinom{10}{4}}{\dbinom{20}{4}} = \dfrac{210}{4845} \approx 0.0433\)
\(\boxed{P(\text{all 4 male}) \approx 0.0433}\)

(c)
No, it does not seem likely that males and females were equally represented in the group of 20 brontosaurs. From part (b), if the group had exactly 10 males and 10 females, the probability of randomly selecting 4 males in a row is only about \(4.33\%\). Since this probability is quite small (less than 5%), observing all 4 selected femurs being male is an unusual result under the assumption of equal representation. It is therefore more reasonable to think that males outnumbered females in this particular group of brontosaurs trapped in the swamp, though equal representation is possible — just unlikely given the data.
\(\boxed{\text{Equal representation is unlikely; evidence suggests more males than females in the group}}\)

(d)
No, it is not reasonable to generalize the conclusion from part (c) to the entire population of brontosaurs. The 20 brontosaurs found at the site do not constitute a random sample from the population of all brontosaurs — they represent only those individuals that happened to wander into that particular swamp and become trapped. This is a highly specific and non-random group. It is plausible that behavioral differences between male and female brontosaurs (for example, males may have been more likely to venture into deep swamp areas while foraging) could explain why males are overrepresented in this particular site. Such a non-representative sample cannot be used to draw conclusions about the broader population of all brontosaurs.
\(\boxed{\text{Cannot generalize; the 20 brontosaurs are not a random sample of all brontosaurs}}\)

Question

At a certain university, students who live in the dormitories eat at a common dining hall. Recently, some students have been complaining about the quality of the food served there. The dining hall manager decided to do a survey to estimate the proportion of students living in the dormitories who think that the quality of the food should be improved. One evening, the manager asked the first 100 students entering the dining hall to answer the following question.

Many students believe that the food served in the dining hall needs improvement. Do you think that the quality of food served here needs improvement, even though that would increase the cost of the meal plan?

_____ Yes_____ No_____ No opinion
(a) In this setting, explain how bias may have been introduced based on the way this convenience sample was selected and suggest how the sample could have been selected differently to avoid that bias.
(b) In this setting, explain how bias may have been introduced based on the way the question was worded and suggest how it could have been worded differently to avoid that bias.

Most-appropriate topic codes (AP Statistics):

• Topic 1.11 — Random Sampling (Part a)
• Topic 1.12 — Potential Problems with Sampling (Parts a, b)
▶️ Answer/Explanation

(a)

Since the manager used a convenience sample — the first 100 students entering the cafeteria — bias may have been introduced because students who arrive at the dining hall early may have opinions about food quality that differ systematically from other dormitory residents who come later or not at all. For example, students who are very hungry or who have strong feelings about the food may be more likely to arrive early, making this group unrepresentative of all dormitory students.
To avoid this bias, the manager should have selected a random sample of 100 dormitory residents — for instance, using a simple random sample from a list of all dormitory residents, a stratified random sample by dormitory building, or a systematic random sample with a random starting point. Any of these approaches gives every dormitory resident a known, non-zero chance of being selected, which eliminates the selection bias introduced by the convenience sample.

(b)

The question as worded contains two sources of wording bias. First, the opening statement — “Many students believe that the food served in the dining hall needs improvement” — is leading because it tells respondents what other students think, which may pressure them to agree and respond “Yes” even if they do not truly feel that way.
Second, the phrase “even though that would increase the cost of the meal plan” introduces a second bias in the opposite direction, making students less likely to say “Yes” because they are reminded of a financial consequence. These two biases may push responses in opposite directions, making the results unreliable in either direction.
A better, more neutral wording would simply ask: “Do you think that the quality of food served in the dining hall needs improvement?” This removes the leading statement and the cost reminder, allowing students to respond based only on their true opinion about food quality.

Question

In order to monitor the populations of birds of a particular species on two islands, the following procedure was implemented.
Researchers captured an initial sample of 200 birds of the species on Island A; they attached leg bands to each of the birds, and then released the birds. Similarly, a sample of 250 birds of the same species on Island B was captured, banded, and released. Sufficient time was allowed for the birds to return to their normal routine and location.
Subsequent samples of birds of the species of interest were then taken from each island. The number of birds captured and the number of birds with leg bands were recorded. The results are summarized in the following table.

Assume that both the initial sample and the subsequent samples that were taken on each island can be regarded as random samples from the population of birds of this species.
(a) Do the data from the subsequent samples indicate that there is a difference in proportions of the banded birds on these two islands? Give statistical evidence to support your answer.
(b) Researchers can estimate the total number of birds of this species on an island by using information on the number of birds in the initial sample and the proportion of banded birds in the subsequent sample. Use this information to estimate the total number of birds of this species on Island A. Show your work.
(c) The analyses in parts (a) and (b) assume that the samples of birds captured in both the initial and subsequent samples can be regarded as random samples of the population of birds of this species that live on the respective islands. This is a common assumption made by wildlife researchers. Describe two concerns that should be addressed before making this assumption.

Most-appropriate topic codes (AP Statistics):

• Topic 3.12 — Setting Up a Test for the Difference Between Two Population Proportions (Part a)
• Topic 3.13 — Carrying Out a Test for the Difference Between Two Population Proportions (Part a)
• Topic 3.1 — Estimators (Part b)
• Topic 1.11 — Random Sampling (Part c)
• Topic 1.12 — Potential Problems with Sampling (Part c)
▶️ Answer/Explanation

(a)

Step 1: State hypotheses.
Let \(p_A\) = true proportion of banded birds on Island A, and \(p_B\) = true proportion of banded birds on Island B.
\(H_0: p_A – p_B = 0 \qquad H_a: p_A – p_B \neq 0\)
Step 2: Identify the test and check assumptions.
We use a two-sample \(z\)-test for a difference in proportions. The test statistic is:
\(z = \dfrac{\hat{p}_A – \hat{p}_B}{\sqrt{\hat{p}(1-\hat{p})\left(\dfrac{1}{n_1} + \dfrac{1}{n_2}\right)}}\)
The problem states the samples are random. Since the two islands are separate, the samples are independent. We check the large sample condition using the pooled estimate:
\(\hat{p} = \dfrac{n_A\hat{p}_A + n_B\hat{p}_B}{n_A + n_B} = \dfrac{12 + 35}{180 + 220} = \dfrac{47}{400} = 0.1175\)
Expected counts: \(n_A\hat{p} = 21.15,\quad n_A(1-\hat{p}) = 158.85,\quad n_B\hat{p} = 25.85,\quad n_B(1-\hat{p}) = 194.15\)
All expected counts are well above 5, so the large sample condition is satisfied.
Step 3: Compute the test statistic and p-value.
\(\hat{p}_A = \dfrac{12}{180} = 0.067 \qquad \hat{p}_B = \dfrac{35}{220} = 0.159\)
\(z = \dfrac{0.067 – 0.159}{\sqrt{\dfrac{(0.1175)(0.8825)}{180} + \dfrac{(0.1175)(0.8825)}{220}}} = \dfrac{-0.092}{\sqrt{0.00105}} = \dfrac{-0.092}{0.032} = -2.875\)
\(\text{p-value} = 2 \times P(Z < -2.875) \approx 0.00429\)
Step 4: State conclusion in context.
Since the p-value of \(0.00429\) is less than \(\alpha = 0.05\), we reject the null hypothesis. There is convincing statistical evidence that the proportions of banded birds on the two islands are different — Island B has a notably higher proportion of banded birds than Island A.

(b)

We use the capture-recapture logic: the proportion of banded birds in the subsequent sample estimates the proportion of banded birds in the whole population.
For Island A, the number of birds banded in the initial sample is \(n_I = 200\), and the proportion of banded birds observed in the subsequent sample is:
\(\hat{p}_S = \dfrac{12}{180} \approx 0.06667\)
Setting this equal to the fraction of banded birds in the population:
\(\hat{p}_S \approx \dfrac{n_I}{\text{population size}}\)
Solving for the estimated population size:
\(\text{Estimated population size} = \dfrac{n_I}{\hat{p}_S} = \dfrac{200}{12/180} = \dfrac{200 \times 180}{12} = \dfrac{36{,}000}{12} = \boxed{3{,}000 \text{ birds}}\)

(c)

Two concerns that should be addressed before assuming the captures can be treated as random samples are:
Concern 1 — Differential catchability: Some birds may be more likely to be captured than others — for example, slower, older, or less wary birds might be caught at a higher rate than the general population. If the same birds that were easy to capture in the initial sample are also more likely to appear in the subsequent sample, then banded birds would be overrepresented in the subsequent sample, leading us to underestimate the true population size.
Concern 2 — Behavioural change after banding: Birds that were captured and banded in the initial sample may become more trap-shy (avoiding capture in the future) or, conversely, may be more conspicuous to predators due to the bands, altering their survival or behaviour. If banded birds are less likely to be recaptured, we would overestimate the population size. In either case, if banding changes the birds’ behaviour or survival, the subsequent sample can no longer be treated as a true random sample of the population.

Question

Because of concerns about employee stress, a large company is conducting a study to compare two programs (tai chi or yoga) that may help employees reduce their stress levels. Tai chi is a \(1{,}200\)-year-old practice, originating in China, that consists of slow, fluid movements. Yoga is a practice, originating in India, that consists of breathing exercises and movements designed to stretch and relax muscles. The company has assembled a group of volunteer employees to participate in the study during the first half of their lunch hour each day for a \(10\)-week period. Each volunteer will be assigned at random to one of the two programs. Volunteers will have their stress levels measured just before beginning the program and \(10\) weeks later at the completion of it.
(a) A group of volunteers who work together ask to be assigned to the same program so that they can participate in that program together. Give an example of a problem that might arise if this is permitted. Explain to this volunteer group why random assignment to the two programs will address this problem.
(b) Someone proposes that a control group be included in the design as well. The stress level would be measured for each volunteer assigned to the control group at the start of the study and again \(10\) weeks later. What additional information, if any, would this provide about the effectiveness of the two programs?
(c) Is it reasonable to generalize the findings of this study to all employees of this company? Explain.

Most-appropriate topic codes (AP Statistics):

• Topic 1.13 — Experimental Design (Parts a, b)
• Topic 1.11 — Random Sampling (Part c)
• Topic 1.12 — Potential Problems with Sampling (Part c)
▶️ Answer/Explanation

(a)
If volunteers who work together are all placed in the same program, there’s a risk that something specific to their workplace situation gets mixed up with the effect of the program itself. For example, suppose this group’s department recently had a deadline pushed back, which on its own would lower everyone’s stress level in that department regardless of which program they’re doing. If this entire group ends up in, say, the tai chi group, then the drop in stress they experience could mistakenly be credited to tai chi, when really it was caused by the lighter workload.

Random assignment fixes this issue. By randomly assigning volunteers to the two programs instead of letting groups choose, we spread people from this department across both the tai chi and yoga groups. This way, any unusual circumstance affecting that department’s stress levels — like the deadline change — gets “evened out” between the two treatment groups rather than being concentrated in just one. Randomization helps make sure the two groups are comparable at the start, so that any difference we see at the end can be attributed to the program rather than to some other confounding factor.

(b)
Yes, a control group would add useful information. Without one, the company could only compare tai chi to yoga directly — they could say which of the two programs led to a bigger drop in stress, but they couldn’t say whether either program actually caused a reduction in stress at all.

Here’s the issue: stress levels might naturally go down over a \(10\)-week period for reasons that have nothing to do with either program — for instance, if the overall work environment becomes less hectic during that time, everyone’s stress might drop a little just from that. A control group, which doesn’t participate in either program but still has its stress measured at the start and end of the \(10\) weeks, gives a baseline for what “no treatment” looks like under those same conditions.

By comparing each treatment group’s change in stress to the control group’s change, the company can tell how much of the reduction is actually attributable to tai chi or yoga specifically, rather than just background changes that would have happened anyway.

(c)
No, it is not reasonable to generalize these findings to all employees of the company. The participants in this study were volunteers, not a random sample of employees. People who choose to volunteer for a stress-reduction study might already be different from the typical employee — for example, they might be more motivated to manage their stress, more open to trying tai chi or yoga, or have more flexible schedules that let them give up part of their lunch hour.

Because the group wasn’t randomly selected from the entire employee population, there’s no guarantee that what works (or doesn’t work) for these volunteers would apply the same way to employees who didn’t volunteer. So while the random assignment within the study supports drawing cause-and-effect conclusions about tai chi versus yoga for people like these volunteers, it doesn’t justify extending those conclusions to the company as a whole.

Scroll to Top