[1]https://www.researchgate.net/publication/228337033_Using_Lin...
[1]https://www.researchgate.net/publication/228337033_Using_Lin...
Maybe the problem is most students just don't give a damn about the classes their parents make them take? Can someone explain the findings in this paper?
That said, typically saw increases well above that in my classes. Even with zero instruction I’d expect a self studying student to improve more than that from familiarization.
It did say upper class people improve more.
Armchair speculation: there are probably some really bright kids out there with poor test-taking skills, and prep likely has a dramatic effect on scores for them, but most students are not that.
Did you actually sit an SAT before you started gaming? What was your improvement?
Don't assume that society is optimizing for what's socially optimal.
Back when it was on a 1600 point scale, I got a 1300 on the first practice test for my SAT course. When I took the test for real, I got a 1570.
First, the test prep companies heavily encourage (and may even require?) students to take the test multiple times. That change alone will boost most people's top score by a solid margin, because it's a noisy test. The first response they always gave to people complaining about lack of improvement was "take the test again, and if you still haven't improved, take the course again for free, and then take the test again".
Second, psychometric expertise is great, but the goal of the SAT is not to be impossible to train for except as a secondary thing. It's really hard to do, especially when you are such a high value optimization target and have to build a test that doesn't rely on much specific knowledge and can be quickly scored. A lot of what the SAT courses do is just teach students to make slightly more accurate guesses on multiple-choice questions where someone unfamiliar with test strategy would leave things blank. That alone tends to boost scores, and some of the other strategies are fairly clever to help avoid common mistakes.
Last, though I didn't have room to improve on the SAT, I also taught GRE classes and can speak to my improvement there. The math section is trivial to get an 800 on (it's easier than the SAT math, or at least it was when I taught 15+ years ago), but the verbal section is quite tough if you haven't studied, and in a lot of ways is a glorified vocabulary test, and the reading comprehension sections can be pretty tough as well. Before I trained to teach the courses, my verbal score was a 490 (on the real test, not practice), so I was considering not even trying to teach (there's some threshold you have to hit, maybe 700 at the time?), but the trainer encouraged all of us to try anyways because he said the content made such a difference. After just two weekends of intensive teacher training, I tested again and ended up with a 740. After teaching the course around a dozen times, I'm pretty confident I could have hit 800 without any difficulty, you basically just have to get used to the types of questions that they ask and get in the headspace of the question authors. Just one data point, but I definitely believe that this stuff is effective.
The elephant in the room in these kinds of discussions is the validity of the tests to begin with. MIT makes a statement that the general tests are predictive in combination with other criteria, but how predictive?
Generally studies of these sorts of things find that they're moderately predictive of first-year GPA (like .4 correlation), and then trail off to zero as the interval between testing and outcomes increases. The effects are even smaller in studies in relatively unselected samples (due to court orders or legal decisions, for example).
So one argument always goes that you should use whatever empirically-supported stuff you can to make fair decisions, but we treat them as if they're more than they are. Sure, you could have a lot of things that are significantly predictive but with small effect size, or where there's lots of noise, but why as a society to we pretend these are huge effects?
The other thing about that paper is the hint that the practice effects are larger in higher-performing examinees, which also makes sense and is consistent with other studies. Why is this a problem? Because those are exactly the types of students for whom these issues are more applicable. A 12 point average gain with practice doesn't matter if you're including people who never had a chance at MIT anyway, but a 20 point gain might in a highly competitive group where small differences are being magnified tenfold.
The issue in all of this isn't the people in the 99th percentile versus the 50th percentile, which is the bulk of what's going into these predictive models and effect sizes, it's the 99th versus the 95th. There's a ton of real-world noise, but society acts as if the noise is nonexistent. It's like we're idolizing the outcome of some kind of survivorship bias process.
Is there any indication that they actually broke down any further, or is this just limited to the impact of average coaching and thus not a good representation of those who receive the top tier coaching?