What everyone believes

“You get what you pay for with red wine — spend more, drink better.”

The data's verdictmixed

In 2008 six economists ran seventeen blind tastings across the United States. Five hundred and six people tasted 523 wines — from a $1.65 bottle to a $150 one — without seeing a single label or price. Six thousand one hundred and seventy-five ratings came back. When the researchers checked whether the expensive wines had been rated higher, they found the correlation ran the wrong way.

Not “weakly positive.” Negative.

That result has been quoted ever since as proof that wine pricing is theater. It isn’t quite — and the reason why is more interesting than either camp usually admits.

The belief

“You get what you pay for.” Spend more on a bottle, get a better bottle. To test it, this analysis asks three falsifiable questions. If the belief holds: blind tasters should rate expensive wine higher; professional scores should rise steeply and reliably with price; and a cheap wine should rarely beat an expensive one. If all three fail, the belief is dead. If they split, the belief is conditional — and the condition is the finding.

Thesis: the price tag is telling you something

Start with the strongest case for paying more, because it is stronger than the cynics allow.

Across the entire reviewed wine market, price and score move together — hard.

r = 0

Correlation between log price and critic score across 80,458 professionally reviewed wines

That is not a rounding error. Sorted into price bands, the pattern is a clean, monotonic staircase: every step up in price buys a higher average score and a much better shot at the 90-point mark that drives shelf placement and sales.

Average critic score by price band
80,458 Wine Enthusiast reviews. Mean score climbs from 84.7 for wines under $10 to 92.7 for wines over $100 — and the share scoring 90+ goes from 1.1% to 88.9%. Whatever critics are measuring, it tracks price closely. Hover any point for the band's sample size.

The meta-analytic literature agrees. Pooling more than 180 hedonic price models built over twenty years and across many countries, Oczkowski and Doucouliagos found a partial correlation of +0.30 between what a wine costs and how it is rated — moderate, consistent, and with no evidence of publication bias.

And at the top of the market, the premium is real in the other direction too: a study of 266,301 bottles sold in the US found that the price premium attached to quality is statistically significant only for wines scoring above 90 points. Below that line, a 50-point wine and an 89-point wine are not reliably priced differently at all.

So the believer’s case is not superstition. In the market as it actually operates, price carries information.

Antithesis: take away the label and it collapses

Now remove the price tag.

The same question, asked two ways
Blind-tasting effects are Goldstein et al.'s own rescaling of their ln(price) coefficients to a 100-point-equivalent scale for a tenfold price increase. The sighted figure is a correlation across the Wine Enthusiast corpus, deliberately not plotted on the same axis — it measures a different thing.

In the blind data, ordinary drinkers rated the more expensive wine slightly worse. A tenfold increase in price bought roughly four points less on a 100-point-equivalent scale. The effect is small but statistically significant (p = 0.038), and it does not go away under scrutiny: with individual fixed effects it strengthens, and when the authors threw out the cheapest and most expensive deciles to look only at the $6–$15 range where most wine is actually bought, the negative coefficient grew roughly fourfold.

Trained tasters did reverse the sign — about seven points up for the same tenfold jump. But the authors were careful, and honest: the net expert coefficient sat at p ≈ 0.10, and they wrote that it “remains an open question whether this coefficient is positive.” The published abstract in the Journal of Wine Economics hedges even further than the working paper did, downgrading “positive” to “non-negative.”

Experts and non-experts predicted the same rating at exactly one price point: $25.70 a bottle. Below it, the experts liked the wine less than everyone else did.

Then there is the question of whether the experts are measuring anything stable. Robert Hodgson persuaded the California State Fair to let him slip triplicate pours — same wine, same bottle — into judges’ flights across four years.

0%

Share of professional wine judges who scored the identical wine consistently within a single medal group

Another tenth of judges scored the same wine anywhere from gold-medal to no-medal. The median spread across replicates was about four points — one whole medal category. Judges were perfectly consistent about 18% of the time, and even that mostly happened on wines they rejected: the panel agrees about what is bad, not what is good. And consistency did not persist — a judge who was reliable one year was no more likely than chance to be reliable the next.

Widen the lens from one competition to thirteen, across 3,000-plus wines, and the same picture: of roughly 1,500 wines that won a gold medal somewhere, more than 70% received no award at all in at least one other competition. Hodgson’s conclusion was that a wine’s chance of gold at one competition is statistically independent of its results elsewhere.

Finally, the mechanism. Plassmann and colleagues put twenty people in an fMRI scanner and told them they were tasting five Cabernets identified by price. There were really only three. One $5 wine was served twice — once labeled $5, once labeled $45; a $90 wine was served as $90 and as $10. Subjects reliably reported the same liquid as more pleasant when it carried the higher price (p < 0.001), and their medial orbitofrontal cortex — the brain’s hedonic accountant — lit up more brightly to match. Eight weeks later the same people tasted the same wines with no prices attached and reported no differences at all.

Price did not change the taste. It changed the pleasure.

The case study: a $19 Malbec

Which brings us to El Libertador, a Mendoza Malbec from Revolution Wine Company — 100% Malbec, 13.5% alcohol, fruit drawn from Agrelo, Ugarteche, Cruz de Piedra and the Uco Valley, and an average price of about $19 a bottle.

The crowd likes it. Vivino gives it 3.9 out of 5 across 287 ratings; Delectable, 8.8 out of 10 across 43. Its 2013 vintage rates 4.1. A public-radio wine program recommended its Cabernet sibling as a wine that “over-delivers” for under $20.

The professionals are cooler. Wine-Searcher’s aggregate is 87 points from just two critic scores — both in the 85–89 band, neither above 90. James Suckling tasted the 2022 and called it “simple and direct, but flavorful… pristine and quite drying.” CellarTracker has no community score for it at all. And across all 80,458 reviews in the Wine Enthusiast corpus, El Libertador does not appear once.

That gap — crowd-liked, critic-unbothered — is the whole analysis in miniature.

Argentine Malbec: price against critic score
1,356 Argentine Malbecs from the Wine Enthusiast corpus (700 plotted for legibility). The dashed line marks El Libertador's ~$19 average price; the wine itself is absent from the corpus, so no score is claimed for it. Note where $19 falls: right at the category's median price.

Within Argentine Malbec, the price-quality correlation is actually stronger than across wine generally — r = 0.692 against 0.606. Malbec is a category where paying more genuinely does tend to get you a better-reviewed bottle. Which makes the cheap-and-excellent Malbec rare rather than typical:

0 of 1,356

Argentine Malbecs that are both $15 or under and scored 90+ by Wine Enthusiast — about 1%

So El Libertador is not a giant-killer. At $19 it sits exactly at the median price for its category, earns an unremarkable 87 from the two critics who bothered, and wins a warm reception from ordinary drinkers. It is, precisely, an average-priced wine that ordinary people enjoy more than professionals do — which is what the blind-tasting literature predicts should happen at that price point. Below $25.70, experts and civilians diverge, and the civilians are happier.

Where the two worlds meet

The apparent contradiction dissolves once you notice the two bodies of evidence are not measuring the same thing.

How critic scores are actually distributed
90.4% of all 80,458 scores fall between 85 and 95, and the single most common score is 88 — not 90. The 100-point scale is, in practice, an 11-point scale.

Critic scores are sighted judgments made by people who know the price, the region, and the producer’s reputation. The meta-analysis is explicit that a wine’s reputation predicts its price better than its measured sensory quality does. So the r = 0.61 in the corpus is partly a measure of quality and partly a measure of how thoroughly price and prestige travel together before a glass is ever poured.

The blind data strips reputation out. What is left is a small negative effect for untrained palates and a fragile positive one for trained ones.

But “the cheap wine sometimes wins” is not the same as “price is meaningless.” In the reviewed market, the distributions barely overlap:

Score distributions: cheap wine vs. expensive wine
12,043 wines at $15 or under against 18,990 at $50 or over. The cheap distribution centers on 86, the expensive one on 92. Only about 5% of expensive wines score at or below the cheap wines' median — the overlap is real but thin.

And the sub-$20 shelf is not uniform either. At the same money, some origins consistently review better than others:

Mean critic score among wines costing $20 or less, by country
Countries with at least 300 sub-$20 reviews in the corpus. Germany and Portugal lead at 88.3 and 87.8; Argentina sits last of the eight at 85.5 — its value reputation rests on price, not on outscoring the field.

Synthesis

The belief is mixed, and it fails and succeeds on predictable terms.

Where it fails. Blind, the premium buys nothing for an ordinary drinker — the measured effect is slightly negative, and it gets worse in the price range most people shop in. The professionals meant to certify quality cannot reproduce their own scores on the same wine, and gold medals do not survive being entered in a second competition. A price tag demonstrably changes reported pleasure while leaving the sensory experience untouched. Anyone who believes their $60 bottle would out-taste a $20 one in a covered glass is, on the evidence, probably wrong.

Where it holds. Price is nonetheless a real signal in the market as it exists. It correlates with expert scores at r = 0.61, it does so consistently across 180-plus studies at +0.30, and the score distributions of cheap and expensive wine overlap by only a few percent. Trained tasters do shift positive, above about $25.70 a bottle. Money buys something — it just mostly buys reputation, consistency, and the expectation of quality rather than a sensation your untrained tongue can pick out of a lineup.

What that means for a $19 bottle. El Libertador is not evidence that cheap wine beats expensive wine; its own critic score is a modest 87 and its category punishes bargain-hunting harder than most. It is evidence of something narrower and more useful: below roughly $25, expert and amateur judgment come apart, and the amateur is the one enjoying themselves more. A wine that ordinary drinkers rate 3.9 out of 5 and critics shrug at is not a failure of the wine. It is a demonstration that above a fairly low threshold, the thing you are buying with the next dollar is increasingly not flavor.

The honest version of “you get what you pay for” is this: you get what you pay for, but a good deal of what you are paying for is the knowing.

Sources

Methodology

Two independent bodies of evidence are combined. (1) The experimental literature on blind tasting: Goldstein et al. (2008), 6,175 blind ratings from 506 tasters across 523 wines, reporting an OLS coefficient on ln(price) of −0.038 pooled and −0.048 for non-experts, with a +0.138 expert interaction, rescaled by the authors to roughly −4 and +7 points on a 100-point-equivalent scale for a tenfold price increase; Hodgson (2008, 2009) on judge reliability and cross-competition concordance; Plassmann et al. (2008) on price-label effects on rated pleasantness (n=20); and Oczkowski & Doucouliagos (2015), a meta-regression over 180+ hedonic models finding a +0.30 partial correlation between price and sensory rating. (2) An original computation over the Wine Enthusiast review corpus (80,458 scored reviews with country, variety, price and points). All corpus figures in this piece — the 0.606 log-price/points correlation, the 88.69 mean score, the 90.36% share of scores in the 85–95 band, the 41.42% 90-plus share, the per-band means, and the Argentine Malbec subset (n=1,356, r=0.692, median price $19, mean 87.59 points, 13 wines at once ≤$15 and 90-plus) — were re-derived from the raw CSVs during the run rather than taken from the source summary, and the re-derivation matched. El Libertador's own figures come from the producer's technical sheet, Wine-Searcher's price and critic aggregate, and Vivino's community ratings.

Limitations

El Libertador is a thinly-distributed wine and the evidence on it is weak-tier by this project's own standard: Vivino (3.9/5, 287 ratings) and Delectable (8.8/10, 43 ratings) are crowd-sourced; Wine-Searcher's 87/100 rests on only two critic scores; CellarTracker has zero ratings for it; and its James Suckling score is paywalled, so only the tasting note is quoted. It does not appear anywhere in the 80,458-review Wine Enthusiast corpus, so it cannot be placed on this analysis's own scatter — its price is marked as a reference line instead, and no claim is made about where its score would fall. A widely-repeated retailer description of its winemaking (lower Brix, 90% neutral oak) was traced to boilerplate for a different, same-named California winery and is therefore not used. The corpus itself is a convenience sample of wines Wine Enthusiast chose to review — median price $30 against a US retail average near $10 — so it describes the reviewed market, not the market people actually buy from, and its scores are sighted, not blind. Goldstein et al.'s non-expert sample skews toward tasting-event attendees rather than the general public, and its 4-point rating scale is coarser than the 100-point scale the rescaled figures are expressed on. Plassmann et al. has only 20 subjects. Finally, 'better' here means rated higher by tasters or critics; no analysis of chemical composition, ageing potential, or food pairing is attempted.

~ $ produced with the dialectical analysis skill — thesis, antithesis, synthesis: researched, provenance-audited, and self-judged end to end.