AP Statistics Inference for Quantitative Data: Slopes — Worked Answer Explanations

Unit 9 · 12 questions explained

Below is a complete answer key for our AP Statistics Inference for Quantitative Data: Slopes practice questions. For each question you'll find the correct choice, a full written explanation of how to get there, and — for every wrong answer — a short note on exactly why it's tempting and where it goes wrong. Reading these straight through is one of the fastest ways to find the gaps in a unit before exam day.

Prefer to test yourself first? Take the timed Inference for Quantitative Data: Slopes practice test and come back here to review, or head back to the Inference for Quantitative Data: Slopes unit overview.

In-content ad
  1. Question 1 · Easy

    A regression output for predicting salary from years of experience shows: slope , . What is the statistic for testing ?

    • A
      Why not A: Dividing by instead of by .
    • B
      Correct
    • C
      Why not C: Squaring before reporting: , or some other transformation.
    • D
      Why not D: Reporting itself as the test statistic, ignoring .
    Explanation

    . The test statistic measures how many standard errors is from the hypothesized value of 0.

    Key takeaway

    $t$ statistic for slope: $t = b/SE_b$ (under $H_0: \beta = 0$).

  2. Question 2 · Easy

    Computer output for a regression of plant height on water amount shows -value = 0.03 for the test vs. . At , what is the correct interpretation?

    • A
      Fail to reject ; there is no linear relationship between water and height.
      Why not A: , so we reject .
    • B
      Reject ; there is convincing evidence of a linear relationship between water amount and plant height.Correct
    • C
      Reject ; water amount causes plant height to increase.
      Why not C: Statistical significance of a regression slope does not establish causation from an observational study.
    • D
      Reject ; the regression model explains 97% of variation in height.
      Why not D: The -value is not related to ; refers to the slope test, not the proportion of explained variation.
    Explanation

    Since , we reject . There is convincing evidence that the true slope is not zero — that is, there is a statistically significant linear relationship between water amount and plant height in the population.

    Key takeaway

    Reject $H_0: \beta = 0$ when $p$-value $< \alpha$; this provides evidence of a linear relationship in the population.

  3. Question 3 · Easy

    A regression of test score on hours of study is based on observations. What are the degrees of freedom for the -test of the slope?

    • A
      25
      Why not A: Using instead of ; one degree of freedom is lost for each estimated parameter (intercept and slope).
    • B
      24
      Why not B: Using ; simple linear regression estimates two parameters, so .
    • C
      23Correct
    • D
      1
      Why not D: for the slope in the ANOVA table, but for the -test of slope, .
    Explanation

    For inference on the slope in simple linear regression, . Two degrees of freedom are lost because two parameters are estimated: the intercept and the slope . So .

    Key takeaway

    For a $t$-test on slope in simple linear regression: $df = n - 2$.

  4. Question 4 · Easy

    Regression output for predicting fuel efficiency (, mpg) from vehicle weight (, pounds) gives , , . A 95% confidence interval for the true slope is reported as . What does this interval mean?

    • A
      We are 95% confident that each additional pound reduces fuel efficiency by between 0.004 and 0.008 mpg.Correct
    • B
      There is a 95% probability the slope is negative.
      Why not B: The probability interpretation of a frequentist CI is incorrect; the slope is fixed (though unknown).
    • C
      95% of vehicles lose between 0.004 and 0.008 mpg per pound.
      Why not C: The confidence interval estimates the population parameter (true slope), not individual vehicle characteristics.
    • D
      We are 95% confident the regression line has a negative -intercept.
      Why not D: The CI is for the slope, not the intercept.
    Explanation

    The confidence interval is for the true slope . We are 95% confident that in the population, each additional pound of vehicle weight is associated with a decrease in fuel efficiency of between 0.004 and 0.008 mpg.

    Key takeaway

    A CI for the slope: we are C% confident the true rate of change in $y$ per unit of $x$ is within this range.

  5. Question 5 · Easy

    A researcher tests vs. with , . Using a -table, the one-sided -value is between 0.025 and 0.05. At , what is the conclusion?

    • A
      Fail to reject ; both bounds (0.025 and 0.05) need to be below to reject.
      Why not A: Since the entire range (0.025 to 0.05) is below or at , we can reject .
    • B
      Reject ; the -value is less than .Correct
    • C
      Inconclusive; the range straddles .
      Why not C: The range (0.025 to 0.05) is entirely ≤ ; we can reject at this level.
    • D
      Fail to reject ; the two-sided -value would be between 0.05 and 0.10, which exceeds .
      Why not D: The test is one-sided (); do not double the one-sided -value.
    Explanation

    The one-sided -value is between 0.025 and 0.05, so it is . We reject and conclude there is statistically significant evidence that the slope is positive.

    Key takeaway

    For a one-sided test, use the one-sided $p$-value; do not double it when $H_a$ specifies a direction.

  6. Question 6 · Easy

    Regression output for predicting blood pressure () from sodium intake () gives , , , for a 95% CI. Compute the 95% confidence interval for the slope.

    • A
      Correct
    • B
      Why not B: Using instead of .
    • C
      Why not C: Using instead of the correct for 95% with .
    • D
      Why not D: Using or the full instead of the reported .
    Explanation

    95% CI for : . With more precise arithmetic: , giving matches the closest choice.

    Key takeaway

    Confidence interval for slope: $b \pm t^* \cdot SE_b$, where $t^*$ uses $df = n - 2$.

  7. Question 7 · Easy

    A study fits a regression of reading score on age for 30 children. The ANOVA table for regression shows , . A researcher notes that the statistic for the slope also gives . What is the relationship between these two tests?

    • A
      They test different things; tests model fit and tests the slope.
      Why not A: In simple linear regression, both the test and the test for the slope are equivalent tests of the same null hypothesis .
    • B
      In simple linear regression, ; both tests are equivalent tests of .Correct
    • C
      The test requires more assumptions than the test.
      Why not C: Both tests in the regression context require the same conditions (linearity, independence, normality, equal variance).
    • D
      The test is always more powerful than the test.
      Why not D: In simple linear regression, they give identical results (); neither is more powerful.
    Explanation

    In simple linear regression, the test from the ANOVA table and the test for the slope are mathematically equivalent: . Both test and yield the same -value.

    Key takeaway

    In simple linear regression, $F_{\text{ANOVA}} = t^2_{\text{slope}}$; both test $H_0: \beta = 0$ and are equivalent.

  8. Question 8 · Easy

    Regression of hours of TV watched per day on sleep hours gives the following computer output:

    CoefSE CoefTP
    Constant9.10.811.40.000
    TV hours−0.40.15−2.670.014

    ,

    Which of the following best describes the result of the test of the slope?

    • A
      The slope is not significant because is low.
      Why not A: measures the proportion of explained variation, not whether the slope is statistically significant.
    • B
      At , reject ; there is evidence of a negative linear relationship between TV hours and sleep.Correct
    • C
      The slope is significant because is greater than 2.
      Why not C: Comparing to 2 is an approximation; significance is properly determined by the -value (0.014 < 0.05).
    • D
      Each hour of TV causes sleep to decrease by 0.4 hours.
      Why not D: Regression describes association; causation requires experimental evidence.
    Explanation

    , so we reject . The data provide convincing evidence of a statistically significant negative linear relationship between TV hours and sleep hours in the population.

    Key takeaway

    Read $p$-values from regression output to determine significance of each coefficient. Use context and direction from the sign of $b$.

  9. Question 9 · Medium

    Which conditions must be satisfied to perform inference on the slope of a regression line? Select all that apply:

    I. Linearity (the relationship between and is linear)
    II. Independence (observations are independent)
    III. Normality (residuals are approximately normal)
    IV. Equal variance (spread of residuals is roughly constant for all )

    • A
      I and III only
      Why not A: Independence and equal variance are also required; the full set is I, II, III, IV (LINER).
    • B
      I, II, and IV only
      Why not B: Normality of residuals is also required for inference on the slope.
    • C
      I, II, III, and IVCorrect
    • D
      None; regression inference has no conditions beyond having data.
      Why not D: Regression inference requires all four conditions (LINER).
    Explanation

    The four conditions for inference on the slope (remembered with the acronym LINER) are: Linearity, Independence of observations, Normality of residuals, and Equal variance (constant spread). All four must be assessed before performing inference.

    Key takeaway

    LINER conditions for regression inference: Linearity, Independence, Normality (of residuals), Equal variance (constant spread of residuals).

  10. Question 10 · Medium

    A residual plot for a regression of price on square footage shows a fan-shaped pattern (residuals get larger as increases). What condition is violated, and what does this mean for inference?

    • A
      Linearity is violated; the relationship is not linear.
      Why not A: A fan shape (increasing spread) indicates unequal variance, not non-linearity; non-linearity would show a curved pattern.
    • B
      Equal variance is violated; inference on the slope may be unreliable.Correct
    • C
      Normality is violated; the residuals are skewed.
      Why not C: A fan shape indicates heteroscedasticity (unequal variance), not necessarily non-normality; normality would be checked with a histogram of residuals.
    • D
      Independence is violated; the data have a time-trend.
      Why not D: Fan shapes are a symptom of heteroscedasticity, not temporal dependence; time-trends would show a pattern with time on the -axis.
    Explanation

    A fan-shaped residual plot indicates heteroscedasticity — the variance of residuals is not constant across all -values. This violates the equal-variance condition. When this condition is violated, the standard errors and confidence intervals from ordinary regression inference are not reliable.

    Key takeaway

    Fan-shaped residual plot → equal-variance condition violated → inference on slope is unreliable.

  11. Question 11 · Hard

    A 95% confidence interval for the slope in a regression of test score on study hours is . A student claims that because 0 is not in the interval, the null hypothesis would be rejected at . Is this correct?

    • A
      No; the CI and hypothesis test are separate procedures and cannot be compared.
      Why not A: There is a direct duality between CIs and two-sided hypothesis tests at the same confidence/significance level.
    • B
      Yes; a value outside a 95% CI corresponds to rejection at for a two-sided test.Correct
    • C
      No; the CI is for (sample slope), not (population slope).
      Why not C: CIs are always for the population parameter; is the point estimate around which the interval is built.
    • D
      No; rejecting requires regardless of the CI.
      Why not D: The CI and -test duality is exact; if 0 is outside the CI, and .
    Explanation

    The CI-hypothesis test duality: if the hypothesized value (0) falls outside the confidence interval, the two-sided test at level rejects . Since 0 is not in , we would reject at .

    Key takeaway

    CI duality: if 0 falls outside the 95% CI for $\beta$, the two-sided test at $\alpha = 0.05$ rejects $H_0: \beta = 0$.

  12. Question 12 · Hard

    A researcher fits a regression of crop yield () on rainfall () and gets , , . The two-sided -value is 0.12. A colleague suggests increasing to 100 would likely make the slope statistically significant. Why might this be true?

    • A
      A larger changes the slope , making it more likely to be significant.
      Why not A: is an estimate that fluctuates with each sample; a larger does not guarantee a different .
    • B
      A larger reduces , increasing the statistic and making it easier to detect a non-zero slope.Correct
    • C
      A larger reduces the critical value, making rejection easier.
      Why not C: While decreases slightly for larger , the main reason is the reduction in , not the change in .
    • D
      A larger increases , which directly reduces the -value.
      Why not D: does not necessarily increase with sample size; it depends on the actual relationship in the data.
    Explanation

    , where grows with . So larger gives a smaller , which increases . If the true slope is , a larger sample will eventually yield a significant test. This illustrates that with enough data, even small slopes can be statistically detected.

    Key takeaway

    Larger $n$ reduces $SE_b$, increases the $t$ statistic, and increases power to detect a true non-zero slope.