Yes, review scores affect game sales, but they are a signal rather than a switch. Academic work on blockbuster titles puts the relationship at roughly a 1.5% increase in units for every 1% improvement in review score, which is meaningful over a long catalog life and almost invisible over a single launch weekend. The harder question is when a score stops being the deciding factor at all.
This matters to two different groups for opposite reasons. Players want to know whether a number they will see on a storefront is worth trusting. Analysts and developers want to know whether that number predicts units, and how to estimate those units when nobody publishes them.
Key takeaways
- Critic scores and player scores influence different windows: critics move preorders and launch week, players move the long tail.
- Research repeatedly finds a positive score-to-units relationship, but with wide variance and no universal threshold.
- Franchise strength, marketing scale, discount depth, platform reach and release timing routinely outweigh the score.
- Storefront ratings such as Steam’s are a separate system, and one that developers use to estimate their own sales.
The rest of this piece works through the evidence, the numbers behind it, and the places where the pattern breaks.
Table of Contents
- Do review scores actually affect game sales?
- When the relationship is strong, and when it is not
- How sales change as review scores fall
- Which review scores have the greatest sales impact?
- Critic aggregates and individual outlet scores
- User scores and storefront ratings
- What is the evidence for a score-sales relationship?
- The sales-per-review method
- Why high scores do not always produce high sales
- How genre, platform, and audience change the effect
- How review bombing and audience conflict distort scores
- How players can interpret review scores responsibly
- Read the review, not the badge
- Check the distribution
- Look at who is reviewing
- Separate technical performance from taste
- Compare critics with players
- Read recent reviews after major updates
- Look for signs of coordination
- Frequently Asked Questions
- Do higher review scores always lead to more game sales?
- What review score is generally considered good for game sales?
- Why can critic and player scores be so different?
- Can a review bomb stop a game from selling well?
- Why are exact game sales figures difficult to verify?
- Do digital discounts have more impact than review scores?
- Conclusion
Do review scores actually affect game sales?
They do, measurably, and the effect is strongest where buyers are deciding between similar products. When two games sit at a similar price on the same shelf, a higher score is the easiest variable to compare, so it does real work in the decision.
It is also true that scores explain only part of the picture, and sometimes a small part. A game with an established audience can sell well at a middling score because the audience already bought in. A critically admired release can stall commercially because it arrived without marketing, on a platform with a small install base, in a quiet week. Anyone who tells you a score alone decides sales is selling you a rule of thumb.
Three things make the measurement messy. Preorders happen before most reviews exist, so scores influence one part of the sales curve while hype drives another. Physical and digital sales follow different patterns and are rarely reported together. And discounts change the timing so thoroughly that a deep promotion can produce a large unit spike inside a window where the score never mattered.
When the relationship is strong, and when it is not
The relationship tends to be strongest for unknown properties competing on quality alone, in categories where players have no loyalty and no prior information. A new puzzle game with no marketing and a 90 Metascore has nothing working for it except the score. Compare that with the eighth entry in a series people have preordered for years, and the score is one of a dozen inputs.
It also weakens as price falls. Someone hesitating over a full-price purchase will read reviews seriously. A title bought at a deep discount with a refund window open is a gamble people take willingly, and reviews matter less in that moment.
How sales change as review scores fall
The table below groups typical outcomes by score band. These are directional patterns rather than hard rules, because genre, audience, franchise strength, platform and timing all shift the result.
| Review score band | What players typically do | Typical commercial pattern |
|---|---|---|
| Overwhelmingly positive (90 and above) | Little hesitation, strong word of mouth, recommendations to friends | Strong launch and an unusually long sales tail; discount less often needed |
| Very positive (80 to 89) | Comfortable purchase for most who were considering it | Solid performance; sales depend heavily on marketing and reach |
| Mostly positive (70 to 79) | Buyers who were already leaning move forward, undecided buyers hesitate | Respects marketing budget more than score; heavy discount often follows |
| Mixed (40 to 69) | Considerable purchase abandonment, especially at full price | Launch spike thins quickly; long tail depends on price cuts and updates |
| Mostly negative and below | Most non-committed buyers pass, and refund windows fill with regret | Strongly front-loaded; sustained revenue needs deep discounts or a major update |
One widely cited survey of player behaviour found that more than half of respondents said they become less likely to purchase once a title falls to Mixed or below. That single threshold does more damage than a ten-point drop higher up the scale, because it is the point where the risk feels personal rather than aesthetic.
Publishers have treated thresholds like this as contract terms. The best-known example is a reported bonus clause tied to reaching a specific Metascore, which turned the number into a line item in a development budget. Whatever the score does to a player’s decision, it demonstrably affects a studio’s revenue expectations.
Which review scores have the greatest sales impact?
Different scores act on different buyers at different moments. Ranking them without that context is misleading.
Critic aggregates and individual outlet scores
Aggregate scores matter most before a purchase decision is made, and they carry the most weight when a player is uncertain. That is precisely the buyer who does not already follow the series or the studio.
Individual outlets matter less individually than their presence inside the aggregate. A single influential review can shift a game’s conversation, but a 92 from one outlet and a 74 from another average out. What genuinely moves the needle is consistency across many outlets, because a wide spread tells buyers the game is divisive.
User scores and storefront ratings
Player ratings dominate the long tail. By the time a game sits in a storefront’s catalogue for months, most browsers are not reading critics, they are reading the percentage of buyers who were happy. That score updates daily and it reflects what players actually experienced after patches and updates, which critics never saw.
OpenCritic and Metacritic fill a similar gap but with different mechanics. OpenCritic weights every outlet equally; Metacritic weights by outlet prestige and, historically, by publication recency. Neither method is neutral, and a game with a small, enthusiastic coverage pool can land differently under each one.
| Score system | How it is calculated | Where it hits hardest |
|---|---|---|
| Metascore | Weighted average of critic scores, weighted by outlet importance and recency | Preorder and launch week |
| OpenCritic Top Critic Average | Unweighted average across participating outlets | Launch week, increasingly used as an alternative reference |
| Storefront user rating | Percentage of positive reviews, with more weight on verified purchases | Long tail and discount periods |
| Console store ratings | Star or percentage scales drawn from platform users, with purchase weighting on some stores | Point of purchase on that console |
Console storefronts deserve a note because they behave differently from one another. One has no meaningful public review system, which removes the visible penalty a bad launch would carry on a competing store. Players on that platform cannot easily read the collective verdict at the moment of purchase, so reputation arrives by other routes: word of mouth, social posts, the advice of whoever they trust about what to play next.
What is the evidence for a score-sales relationship?

The strongest evidence sits in academic work rather than industry commentary. A widely cited empirical study of blockbuster titles, published by Cox, Downing and Kaimakamis, found that the Metacritic review score had a highly significant association with unit sales, reporting the relationship at above the ninety-nine percent confidence level.
The same research group later put a number on the magnitude: each one percent improvement in review score was associated with roughly a one and a half percent increase in unit sales. That is a modest elasticity. It means a ten-point jump in aggregate score is associated with something in the region of a fifteen percent difference in units, all else equal. It does not mean a critically panned game loses three quarters of its audience.
Later work on what makes a game score highly pointed somewhere less flattering. Larger development teams were consistently associated with higher scores, which suggests that production scale buys polish, and polish buys good reviews. Budget is therefore quietly upstream of the score, which is one reason the correlation between acclaim and revenue looks weaker than people expect.
The sales-per-review method
Sales figures are not public, so researchers reverse-engineer them from review counts. The most-cited approach comes from a Grey Alien Games developer who published the method in 2018: take a game’s Steam review count and multiply it by a ratio, with a central estimate of roughly 77 sales per review and a plausible range from about 30 to about 150.
So a thousand reviews might represent thirty thousand units or a hundred and fifty thousand, with the middle of that range more likely than the edges. The ratio drifts over time, trending lower as years pass, and it varies sharply by genre. Players of story-driven single-player games review far less than players of roguelikes, simulations and strategy titles, which inflates the ratio for those genres.
The odd finding from that data is that better-reviewed games earn slightly fewer sales per review. Satisfaction suppresses the urge to write. It also means the method is a rough instrument, not a sales report.
Why high scores do not always produce high sales
Acclaim is a marketing asset, not a distribution strategy. Seven factors regularly outweigh it.
- Franchise recognition. An established series arrives with a pre-built audience that has already decided to buy. The score influences what they say afterwards rather than whether they purchase.
- Marketing scale. Advertising reach determines who knows the game exists. A well-reviewed title nobody has heard about cannot convert awareness it never had.
- Release timing. A crowded launch splits attention and store visibility. Two acclaimed games released in the same week divide an audience that would otherwise have been sequential.
- Discount depth. A promotion moves the timing of a purchase far more reliably than a review. A deep cut can produce more units in a weekend than a full-price run would gather in a month.
- Platform reach. Install base matters more than sentiment. The same game on a small platform will struggle no matter what the critics said.
- Exclusivity and timing on hardware. Launch titles compete against the idea of waiting for reviews, and a limited window compresses every factor into a few days.
- Genre and audience size. Some categories have a large committed base and others simply do not. A brilliant release in a thin genre has a ceiling the score cannot raise.
Put those together and the mismatches make sense. A franchise entry with mediocre reviews and a huge preorder base can outsell an acclaimed new IP every time, which is exactly the pattern that surprises people looking for a simple score-to-sales rule.
How genre, platform, and audience change the effect

The same score produces a different commercial result depending on what is being sold and where it is sold.
Indie games live or die on reviews. With little marketing budget and no brand recognition, a strong aggregate and a positive storefront rating are frequently the only reasons a game is discovered at all. This is the context where the score-to-sales relationship is tightest.
Sequels inherit an audience and a price expectation. Reviews matter more for sequels than for new properties, because a sequel that underperforms confirms a fear fans already had. That reaction is why the gap between critical and player scores tends to widen on sequels.
Competitive multiplayer titles are judged on stability and matchmaking rather than narrative, and players judge them within days of playing rather than from reviews. Early impressions and streamer sentiment carry more weight than any critic’s verdict.
Remasters and remakes inherit the score of the original from most buyers’ point of view, and the comparison reviews draw is usually short. The relevant question is whether the remake fixed the complaints, which is a review-reading problem more than a score problem.
Family and all-ages titles are bought by parents and gift-givers, who tend to use ratings and platform guidance rather than critical consensus. Storefront star ratings do more work here than a Metascore.
Heavily discounted catalogue titles operate outside the score-to-sales model entirely. Buyers are paying for a catalogue entry and price is the dominant variable.
On PC, ratings are public, updated daily and directly tied to the point of purchase. On console, ratings sit inside a walled environment where visibility depends on how the publisher merchandises the title, and where one platform lacks a public review system altogether. Handhelds inherit whichever ecosystem they belong to. Subscription libraries add another layer, since a title can be acquired through a bundle without ever appearing on a storefront page.
How review bombing and audience conflict distort scores
Review bombing is a coordinated campaign to flood a title’s rating with negative reviews. The tactic has been used for everything from localised complaints to culture-war grievances, and it produces a number that reflects organiser energy rather than player experience.
Several factors make scores hard to compare for this reason. Purchase weighting limits how many reviews come from people who actually own the game, though loopholes remain. Campaign timing skews early ratios, because the first reviews carry the most weight in a percentage that has a small denominator. Platforms delay or restrict suspicious activity, so the visible score at any given moment may be mid-correction. And the volume of discussion around a campaign can make a game look far more contested than its actual buyer base.
Forum discussion about suspicious scores tends to follow the same routine. Players check whether the negative reviews arrived in a tight cluster, whether the accounts behind them share any history, and whether the wording repeats. Threads about a particular franchise show people reading the lowest-rated user reviews specifically to look for coordination before trusting a number.
Audiences also disagree in ways that are not manipulation at all. A game built for a committed genre audience can be harshly received by generalists and loved by the people it was made for. A story-driven game can be criticised by players who wanted something else entirely. Those gaps are information, not noise, and they tell you who the game is for.
How players can interpret review scores responsibly
Use the number to decide what to look at, not what to buy.
Read the review, not the badge
A score is the last line of a review. The reason behind it tells you whether your objection matches the critic’s.
Check the distribution
The spread matters more than the average. A game with a broad cluster of middling scores and a few extremes is genuinely divisive; a game with one outlier dragging a tight cluster down is a different situation. Storefront pages show this pattern clearly once you know to look.
Look at who is reviewing
Weight outlets you have reason to trust for the genre. A publication with a strong record in strategy games tells you more about a strategy game than a generalist outlet with a good year does.
Separate technical performance from taste
Frame rate, bugs and load times are objective and worth reading carefully, because they get patched and subjective impressions do not. If the complaints cluster on performance, check whether patches have landed since the reviews were written.
Compare critics with players
When the two disagree, the gap is the interesting part. Critic scores weight design and intent; player scores reflect satisfaction after fifty hours. Both are valid, and they answer different questions.
Read recent reviews after major updates
A score is a snapshot of a moving target. A large update can rewrite the recent rating long before it moves the all-time average.
Look for signs of coordination
Clusters of identical reviews arriving within hours, a sudden swing that matches a controversy rather than a patch, or a review count far out of step with the game’s size are all reasons to discount the number.
Frequently Asked Questions
Do higher review scores always lead to more game sales?
No. Research finds a positive relationship, with one published analysis putting it at roughly a 1.5% increase in units per 1% score improvement, but that is an average across many titles. Franchise loyalty, marketing reach, discount depth, platform install base and release timing regularly outweigh the score. Treat it as one input among many rather than a sales guarantee.
What review score is generally considered good for game sales?
There is no universal threshold, but bands matter more than exact numbers. Scores of 90 and above tend to support the strongest long-tail revenue and reduce the need for discounting. Falling to Mixed or below is where the damage concentrates: one widely cited survey found more than half of players said they become less likely to buy at that point. Publishers have written specific numbers into bonus clauses for exactly this reason.
Why can critic and player scores be so different?
They measure different things at different moments. Critics write short impressions soon after release, weighting design and intent. Players write after spending dozens of hours with a finished, patched game, so they tend to weight stability and value. Differences widen on sequels and genre-specific releases, where the audience a critic is ignoring is exactly the audience playing.
Can a review bomb stop a game from selling well?
It can suppress a rating quickly, but the sales effect is harder to isolate. Purchase weighting limits coordinated accounts, and platforms remove flagged activity, so a visible rating may be mid-correction. The lasting damage tends to be reputational among buyers who look at the number at the point of purchase. Check for clusters of identical reviews arriving within a short window before treating the score as a verdict on quality.
Why are exact game sales figures difficult to verify?
Publishers do not disclose unit totals, publishers report differently and rarely combine physical with digital, and platform stores publish no figures at all. Most estimates are reverse-engineered from review counts using a sales-per-review ratio, whose central estimate is roughly 77 with a plausible range of about 30 to 150. Treat any single number as an order of magnitude, not an account.
Do digital discounts have more impact than review scores?
Often, yes. A discount moves the timing of a purchase far more directly, and it can produce a large unit spike inside a window where reviews played no part. The two interact too: a strong score reduces how deep a discount a publisher needs, and a strong discount gives players the chance to buy something they had doubts about. Over a catalog life, both variables matter, and they rarely act alone.
Conclusion
Review scores move sales, but they are a signal with a long list of confounders rather than a formula. Critic scores act early and on uncertain buyers; player scores act late and on everyone else. Between them sit franchise loyalty, marketing, discounting and platform reach, any of which can be larger than the score.
So look at four things before you trust a number: the distribution of reviews rather than the average, the reputation of the outlets or players behind it, whether recent reviews reflect patches that landed after the criticism, and whether the game’s broader situation explains more than its reception. If the score looks strange, that last check is usually where the answer is.
Updated for 2026.


