The average hides the distribution
A star rating is an average, and averages hide as much as they reveal. Two products can both sit at 4.3 stars while telling completely different stories: one might be loved by almost everyone with a few detractors, while the other is a battlefield of five-star praise and one-star fury. The number is identical. The risk is not.
This is why the rating histogram — that little bar chart of how many 5-star, 4-star, and 3-star reviews a product has — is more useful than the headline number. It takes five seconds to read, and it changes what 4.3 means. Make it a habit to glance at the shape before you read a single review.
Polarized 4.3 vs. consistent 4.3
A polarized 4.3 — heavy on five stars, heavy on one stars, hollow in the middle — usually means buyers are experiencing two different products. Maybe quality control is inconsistent, maybe the product works great for some use cases and fails at others, or maybe the positives were boosted. Either way, your odds look like a coin flip, not a 4.3-star certainty.
A consistent 4.3 — a big mound of four- and five-star reviews, a reasonable tail of threes, a few ones — is the shape of a solid product with normal trade-offs. Nobody's product is perfect; a healthy middle of the distribution means buyers are engaging honestly with what the product is and isn't. This 4.3 is a much safer bet than the polarized one.
Categories have their own norms
Star ratings also mean different things in different categories. Shoppers tend to rate phone cases and kitchen gadgets generously, while they grade mattresses, skincare, and anything with a subscription component harshly. A 4.1 in a tough category can represent a better product than a 4.6 in an easy one.
So compare within categories, not across them. When you're choosing between two blenders, the relative ratings matter; when you're comparing a blender's 4.5 to a moisturizer's 4.2, the numbers aren't speaking the same language. Category norms are the invisible context behind every rating.
Old ratings can describe a different product
Ratings have a memory problem: they accumulate over the product's whole life, but products change. A listing that's been live for three years might show a 4.4 built on an older version — before the manufacturer quietly changed the materials, the supplier, or the formula. The rating describes the product's history, not necessarily the unit you'd receive.
Sort reviews by "most recent" and check whether the tone has shifted. If recent reviews are noticeably harsher than older ones, the product may have changed for the worse — and the headline rating is lagging behind reality. Recency is the correction factor the average doesn't apply.
Read the shape, not the number
The practical routine is simple: check the histogram shape, compare against category norms, and skim the most recent reviews for shifts. Three quick checks, and a 4.3 stops being a shrug and starts being information.
Star ratings are a starting point, never a verdict. They're most useful as a filter — ruling out the obvious duds — and least useful as a tiebreaker between two decent options. For the final call, the distribution and the actual words in the reviews will always tell you more than the number.
A quick worked example
Imagine two blenders, both rated 4.3. Blender A's histogram shows 70% five-star reviews, a solid block of fours, and a thin tail of lower ratings — buyers broadly agree it's good with minor flaws. Blender B shows 55% five-star, 25% one-star, and almost nothing between — half of buyers love it and a quarter think it's broken.
The headline number can't tell them apart, but your money should. Blender A is the safer buy for almost everyone; Blender B is a gamble that pays off only if you happen to land in the happy half. Same rating, different meaning — and now you know which questions to ask.