Third-Party Full-Length Deflation: Why Your Practice Score Is Gaslighting You
You scored a 498 on a prep company full-length and your soul left your body. A look at the data on third-party score deflation, the spicy theory that ties it to score guarantees, and how to read practice scores without spiraling.
You finished content review. You felt good. You sat down for your first full-length from a big-name prep company, gave it seven and a half hours of your one wild and precious life, and the screen said 498.
Four hundred ninety-eight.
You stared at it like it owed you money. Then you opened Reddit and typed the question every premed types at least once: is this company deflated?
Short answer, bestie: very possibly yes.
Your 498 Might Be a 508 in a Trench Coat
The best public data on this comes from Joel Harris, a former test taker who pulled 844 self-reported scores from a community spreadsheet built by r/MCAT® and Student Doctor Network users, all from people who sat the real exam between January and September 2017. He lined up what they scored on third-party full-lengths against what they got on test day.
The results were rude. Kaplan scores ran so low that his crude fix was to add about 10 points. Next Step, now part of Blueprint, needed about 7. Princeton Review was the wildest of the bunch: the average person who scored a 503 on a TPR exam walked out with a 518 on the real thing. Harris himself scored a 500 on his first Princeton Review test and felt crushed, when people in that range typically landed somewhere between 511 and 516.
He is upfront about the caveats, so let's be too. His dataset averaged an absurdly high 515, because people love posting good news. The data is from 2017. And companies tweak their scaling all the time, often without announcing it. So read these numbers as the rough shape of the problem, and hold the exact offsets loosely.
The shape has held up, though. Tutor Trevor Klee put Kaplan around 5 to 7 points low and Princeton Review around 8 to 10 low. Med School Insiders says many students land within five to seven points of their Blueprint scores. The sources disagree on the size of the gap and agree completely on the direction. Third-party tests run hard.
Follow the Guarantee
So why would a company make its own test harder than the real one? Harris had a theory, and it is spicy.
Several big prep companies sell score guarantees, and those guarantees measure your improvement from a baseline. Harris noticed that Kaplan's guarantee at the time paid out only if your real score came in below your diagnostic, and he argued that a roughly ten-point cushion in the practice scaling would keep that from happening often. Princeton Review's baseline could be its own diagnostic, which he said left almost no chance of failing to beat it. He also pointed out that Next Step, which had no full-refund guarantee back then, deflated the least of the three.
Here is what one guarantee looks like today, straight from Kaplan's own page. For its premium programs, you set a baseline on a designated Kaplan exam, or with an official MCAT® score from the past six months. Start at 500 or above and Kaplan guarantees a 515. Start below 500 and it guarantees a 15-point jump.
Now picture a baseline test that runs a few points hard. The starting line slides backward, and every point of improvement gets a little easier to reach. To be clear, that is Harris's theory, and nobody has proven intent. Incentives are still worth noticing, and fine print on these guarantees is always worth reading twice.
In fairness, there are gentler explanations too. Harder practice makes test day feel easier, and nobody has ever filed a complaint about being overprepared. Every company also builds its own scoring curve from its own students, so a 505 on one platform and a 505 on another are measured with different rulers. And the companies are pushing back on the stereotype: Blueprint now says its diagnostic lands within 0.3 points on average, based on students who sat the real exam in the prior 18 months. Receipts on both sides. Your job is to read the scores correctly.
How to Read a Third-Party Score Without Spiraling
First, know that the emotional arc is universal.
Now skip straight to acceptance with a few rules.
- Use third-party full-lengths for their real superpowers. Stamina, timing, content gaps, and reps at reviewing like a pro. They are excellent practice. They are a mediocre fortune teller.
- Compare a company only to itself. A 504 on Kaplan last week and a 508 on Kaplan this week is real movement. A 504 on Kaplan next to a 510 on an AAMC exam is two different rulers measuring the same person.
- Give the prediction job to the AAMC. The official full-lengths come from the people who write the test. Average your last two, and use the readiness benchmarks from there.
- Section patterns still count. If C/P is your lowest section on every platform, that is a real signal, and it survives any scaling curve.
- Never reschedule over one third-party score. That decision belongs to your AAMC averages.
If you just got a terrifying number from a prep company, breathe. It might be a 508 in a trench coat. Review it, learn from it, and then go take an AAMC exam and let that number do the talking. And if a friend posts a 498 in the group chat tonight, send them this article and a snack. They are probably doing better than their screen says.
Sources
- Joel Harris, Converting 3rd Party MCAT Scores to Actual Scores
- Kaplan, MCAT 515+ Score or 15+ Point Increase Guarantee
- Blueprint, MCAT Practice Tests and Qbank
- Trevor Klee, How Accurate Are AAMC, Next Step, Kaplan, and Princeton Review MCAT Exams?
- Med School Insiders, How Do MCAT Practice Tests Compare to the Real Thing?
- Med School Insiders, What They Don't Tell You About MCAT Score Guarantees
