From the question, x = 18 . From previous examples we found that y ^ = 26.742 + 3.216346 · 18 = 84.636228 and s = 3.935892 . Find the critical value from the invT using d f = n − 2 = 13 ; we get t α / 2 = 2.160369 . Make sure to go out at least 6 decimal places in between steps. Ideally, never round between steps. Use the 2-Var Stats from your calculator to find the sums and then substitute values back into the equation to get
84.636228
±
2.1600369
·
3.935892
(
1
+
1
15
+
(
18
−
16.6
)
2
41.6
)
⇒
84.636228
±
8.9723
⇒
75.6639
<
y
<
93.6085
We are 95% confident that the predicted exam grade for a student that studies 18 hours is between 75.6639 and 93.6085.
A confidence interval can be more accurate (narrower) when you increase the sample size. Note that in the last example, the predicted grade for an individual student could have been anywhere from a C to an A grade. If you wanted to predict y with more accuracy, then you would want to sample more than 15 students to get a smaller margin of error. The confidence interval for a mean will have a smaller margin of error than for an individual’s predicted value.
Excel, the TI-83 and 84 do not have built in prediction intervals.
TI-89: Enter the x -values in list1 and the y -values in list2, select [F7] Intervals, then select option 7:LinRegTInt… Use the Var-Link button to enter in list1 and list2 for the X List and Y List. Select Response in the drop-down menu for Interval. Enter in the x -value given in the question. Change the confidence level (C-Level) to match what was in the question, the [Enter]. Scroll down to Pred Int for the prediction interval. The calculator does not round between steps so if you rounded b 0 and b 1 , for instance, when doing hand calculations, your answer may be slightly different than the calculator results.
Extrapolation is the use of a regression line for prediction far outside the range of values of the independent variable x . As a general rule, one should not use linear regression to estimate values too far from the given data values. The further away you move from the center of the data set, the more variable results become. For instance, we would not want to estimate a student’s grade for someone that studied way less than 14 hours or more than 20 hours.