yego.me
💡 Stop wasting time. Read Youtube instead of watch. Download Chrome Extension

Example estimating from regression line


3m read
·Nov 11, 2024

Lizz's math test included a survey question asking how many hours students spent studying for the test. The scatter plot below shows the relationship between how many hours students spend studying and their score on the test. A line was fit to the data to model the relationship. They don't tell us how the line was fit, but this actually looks like a pretty good fit if I just eyeball it.

Which of these linear equations best describes the given model? So as you know, this point right over here shows that some student at least self-reported they studied a little bit more than half an hour, and they didn't actually do that well on the test. Looks like they scored a 43 or 44 on the test.

This right over here shows, or like this one over here, is a student who says they studied two hours, and it looks like they scored about a 64 or 65 on the test. This over here, or this over here, looks like a student who studied over four hours, or they reported that, and they got, looks like, a 95 or 96 on the exam.

And so then, these are all the different students; each of these points represents a student. They fit a line, and when they say which of these linear equations best describes the given model, they're really saying which of these linear equations describes or is being plotted right over here by this line that's trying to fit to the data.

So essentially, we just want to figure out what is the equation of this line. Well, it looks like the Y intercept right over here is 20, and it looks like all of these choices here have a y intercept of 20, so that doesn't help us much. But let's think about what the slope is. When we increase by one, when we increase along our x-axis by one, so change in X is one, what is our change in y?

Our change in y looks like, let's see, we went from 20 to 40. It looks like we got went up by 20, so our change in y over change in X for this model, for this line that's trying to fit to the data, is 20 over one. So this is going to be our slope, and if we look at all of these choices, only this one has a slope of 20, so it would be this choice right over here.

Based on this equation, estimate the score for a student that spent 3.8 hours studying. So we would go to 3.8, which is right around, let's see, this would be 3.8, would be right around here. So let's estimate that score. If I go straight up, where do we intersect our model? Where do we intersect our line?

So it looks like they would get a pretty high score. Let's see, if I were to take it to the vertical axis, it looks like they would get about a 97. So I would write my estimate is that they would get a 97 based on this model.

Once again, this is only a model; it's not a guarantee that if someone studies 3.8 hours they're going to get a 97, but it could give an indication of what maybe might be reasonable to expect, assuming that the time studying is the variable that matters.

But you also have to be careful with these models because it might imply, if you kept going, that if you study for nine hours, you're going to get a 200 on the exam, even though something like that is impossible. So you always have to be careful extrapolating with models and keep it, take it with a grain of salt.

This is just a model that's trying to fit to this data, and you might be able to use it to estimate things or to maybe set some form of an expectation, but take it all with a grain of salt.

More Articles

View All
Subject and object pronouns | The parts of speech | Grammar | Khan Academy
All right, so grammarians, I want to talk to you about the difference between subject and object pronouns. But before we do that, let’s start off with a little primer on what subjects and objects actually are—um, just generally, for our grammatical purpos…
Impact of removing outliers on regression lines | AP Statistics | Khan Academy
The scatter plot below displays a set of bivariate data along with its least squares regression line. Consider removing the outlier at (95, 1). So, (95, 1) we’re talking about that outlier right over there and calculating a new least squares regression li…
Everything About Irrigation Pivots (Farmers are Geniuses) - Smarter Every Day 278
If you’ve ever been flying in an airplane and you look out the window and you see these big green circles in the middle of a field, what is that? Today, we’re going to build the thing that makes this happen. This is my buddy Trey; he’s a farmer. Farmers a…
Divergence formula, part 1
Hello everyone. So, now that we have an intuition for what divergence is trying to represent, let’s start actually drilling in on a formula. The first thing I want to do is just limit our perspective to functions that only have an x component, or rather w…
How Much Money is There on Earth?
Hey, Vsauce. Michael here. On Earth, the average piece of currency changes hands about 55 times a year. That’s about once a week. With that kind of turnover, it’s safe to say that statistically in the United States, out of every 100 pieces of currency, o…
2015 AP Chemistry free response 3f | Chemistry | Khan Academy
The pH of the soft drink is 3.37. After the addition of the potassium sorbate, which species, the sorbic acid or the sorbate ion, has a higher concentration in the soft drink? Justify your answer. So, this is related to the question we’ve been doing beca…