How to Read an IMDb Rating Without Being Misled
Vote counts, weighting, brigading and survivorship: the numbers behind the number, and the three checks that tell you whether a rating is worth trusting.
Read article →Independent film & TV review site. We publish criticism and information only. Nothing on this site is downloadable and we host no software or install files. Read our disclaimer
Independent film & TV review site. We publish reviews and information only. Nothing on this site is downloadable and we host no software or install files. Read our disclaimer
Exactly how our ten-point scores are built: the six categories we grade, the weights behind them, the tie-breakers, and the honest limits of any rating system.
Every score on this site comes out of the same six-category rubric, applied the same way to a four-hour prestige drama and a ninety-minute streaming comedy. No deciding a number first and working backwards, no nudging it up because a lead performance was charming, no rewriting history when the consensus goes the other way. A score you can’t audit isn’t a score, it’s a mood with a decimal point. So here is the whole machine, weights and all.
We split a title into six things that can be judged separately, because most bad reviewing comes from collapsing six separate judgements into one impression and calling it taste. The six are writing and structure, direction and craft, performance, originality and ambition, execution of intent, and rewatch value. Six is deliberate. Fewer and the score is just a feeling wearing a jacket; more and the categories start overlapping until the same quality gets counted three times over.
Writing and structure means the script: whether scenes build, whether dialogue sounds like people, whether the story earns its turns or simply announces them. Direction and craft covers everything the camera and the edit do: framing, pacing, sound design, production design, the difference between a sequence that breathes and one that just sits there. Performance is the leads and the ensemble together; a brilliant central turn surrounded by cardboard is not a performance triumph, it’s a rescue mission.
Originality and ambition asks how hard the thing is trying and how familiar the ground is. Execution of intent is separate from all of that on purpose: a film that sets out to be a breezy heist comedy and delivers exactly that should not be marked down for refusing to be a Bergman picture. Rewatch value is the least weighted column precisely because it’s the most unfair one. A devastating one-and-done drama shouldn’t lose points for doing its job too well.
Each category is graded 0-10 to one decimal, independently, before anyone sees a total. A running total has gravity, and graders drift towards numbers they’ve already written down. The six sub-scores are then combined using fixed weights that add up to 100%. Those weights haven’t changed since we started publishing sub-scores, and if they ever do, we’ll note it and leave old scores exactly as they were printed.
| Category | The question we’re actually asking | Weight |
|---|---|---|
| Writing & structure | Does the script earn its turns, scene by scene? | 25% |
| Direction & craft | Camera, editing, sound, design, pacing | 20% |
| Performance | The leads and the ensemble, not just the famous one | 20% |
| Originality & ambition | How much new ground, and how hard does it push? | 15% |
| Execution of intent | Did it deliver what it set out to deliver? | 12% |
| Rewatch value | Does it reward a second visit, or is it spent? | 8% |
| Total | - | 100% |
We could nudge the weights until every score matched our gut feeling, and the system would instantly stop meaning anything. The value of a rubric is that it disagrees with us sometimes. When the maths says 7.1 and the room feels like an 8.0, we publish the 7.1 and let the review prose explain what the number can’t.
Writing carries the heaviest weight because almost every other good thing in a film or a season is downstream of it. You can photograph a bad script beautifully; you can’t act your way out of a scene with no reason to exist. Ambition is weighted below the three craft columns because we want to reward difficult things that work, not difficult things that merely exist. Plenty of very ambitious titles score badly here.
Here is the part that matters: a 6.0 is not “60% of a great film”, and it is not a 7.0 that lost some points for being forgettable. It is a specific shape. Compare three real score profiles from our own review archive.
| Category | The competent 6.0 | The almost-great 8.0 | The exceptional 9.5 |
|---|---|---|---|
| Writing & structure | 6.5 | 8.5 | 9.8 |
| Direction & craft | 6.5 | 8.5 | 9.6 |
| Performance | 5.8 | 8.0 | 9.7 |
| Originality & ambition | 4.5 | 6.5 | 9.2 |
| Execution of intent | 6.8 | 9.0 | 9.5 |
| Rewatch value | 5.0 | 7.0 | 8.8 |
| Weighted total | 6.0 | 8.0 | 9.5 |
The 6.0 is professionally made and completely unambitious: a by-the-numbers thriller that hits every beat you expect and skips every risk it could have taken. Its execution score is its highest number, and that’s the point: it did what it set out to do, and what it set out to do wasn’t much. Craft can’t save a 4.5 in originality, which is why this profile lands at a middling 6.0 despite having nothing broken in it.
The 8.0 is the shape of a very good production with one real ceiling. Writing and direction are both strong, execution is near-perfect, the acting is genuinely good, but originality sits at 6.5: you have seen this story before, and its best scenes are the ones least like the rest of it. That single soft column is the entire difference between an 8.0 and a low 9. The 9.5 is different in kind rather than degree: it is above 9 in five of six columns, and its weakest score would still be the best line on most other titles’ sheets. That is not an accident. To climb from a 9.0 to a 9.5 you have to gain almost everywhere at once, which is exactly why so few reviews ever get there.
The number describes how well made something is. It does not describe how much you, specifically, will enjoy it. A 9.2 horror film is a 9.2 for a person who has not willingly watched a horror film since 2011, and they will turn it off after twenty minutes. The score is a description, not a promise, and treating it as a promise is how people end up watching technically flawless films they resent.
This is also why a score is not a recommendation. A recommendation needs two things a rubric can never have: knowledge of the viewer and knowledge of the evening. The same 8.4 drama is the perfect pick for a long Sunday and a terrible pick after a fourteen-hour day, and no decimal point is going to tell you which one you’re having.
The practical fix is to use the score as a sorting tool and the sub-scores as the decision. If what you love about television is dialogue, read the writing column and ignore the total. If you watch for a great central performance, the performance line is your number. That’s the whole reason we publish the breakdown instead of a single opaque figure: you can disagree with us about one axis without throwing out the entire review.
We accept all four of those weaknesses and publish anyway, because the alternative (a vibe score) has the same problems plus no way to check our work. If you want to know why two films both landed on 7.6, the sheets will tell you. Often they got there for opposite reasons, and that difference is more useful than the shared number.
When two titles land on the same reported score, we break the tie in a fixed order: the unrounded totals first, then execution of intent, then the more ambitious title, and finally rewatch value. That order is written down so nobody invents a tie-breaker at one in the morning that happens to favour the thing they liked more. On rounding: we report one decimal using standard rounding, and nothing rounds up to a 10. A 10 currently has to be earned with a raw score north of 9.95 across the board, and no title has managed it yet. A perfect score that is impossible to reach is more honest than one we hand out every awards season.
There is a longer list of things the rubric deliberately refuses to measure. It does not measure cultural importance: a landmark title that changed the industry can still score a 7.0, because influence is a fact about history and quality is a fact about the work. It does not measure whether a title suits your mood, your age, or your household. Parental controls and age ratings are a separate question with a separate guide. It does not measure the streaming service around it: bitrate, app quality and buffering are a platform issue, not a film issue, and we score them elsewhere.
The score also ignores awards, box office, and how loudly the internet argued about it. It genuinely does not care whether the characters are likeable, only whether they are written and played as people. And a score is not a popularity contest in either direction: a 4.0 that nobody watched is still a 4.0, and a title with a passionate fanbase does not get a bonus for the passion.
Six categories, fixed weights and published sub-scores won’t tell you what you’ll love. They will tell you, every single time, exactly why we landed where we landed, which is the only thing a rating system can honestly offer. Read the breakdown, ignore the decimals that don’t move, and use the review to decide.
More original guides from the same testing desk.
Vote counts, weighting, brigading and survivorship: the numbers behind the number, and the three checks that tell you whether a rating is worth trusting.
Read article →Why 200-title watchlists fail, how to cap yours without missing anything, and a simple system for turning a queue into an actual evening plan.
Read article →Six tight, spoilable series built for a Saturday-to-Sunday run: how long each one needs, where to watch it, and which one to start with tonight.
Read article →