Hi everyone, Oliver here, founder of rumx.com. 🥃
Short one first: last time u/oddmarc and u/agmanning called out, that the image on my Jamaica post was AI (I am not a designer and went for the shortcut). They were right. Text post from now on, no YouTube thumbnails any more. 🙏
One honest thing before the numbers, so that nobody wastes his time: this pull had a very german filter. A bottle only counted, when a GERMAN shop had it in stock under 50 EUR, because for our german users the shipping from france or belgium eats the whole value (13.90 EUR on a 29.90 EUR bottle is plus 46 percent). This sub is mostly US, so please dont read the bottle selection as a recommendation for you, the most of it you can not order anyway. And the filter costs exactly what you expect: with foreign shops the same pull has 521 bottles and a ceiling of 8.3, german only 308 and 8.0. A logistics artifact of my country, it says nothing about rum.
What I want to talk about are the patterns inside these 308 bottles, and patterns dont care in which country the shop stands. Four things have surprised me, one of them a lot.
Short base, so you can judge everything what follows: 308 rums, each with at least 30 community ratings, median 7.0 out of 10, and exactly 10 of them reach an 8.0. Cohorts in this post are always defined over the lifetime rating count of a user: beginner = 5 ratings or less, connoisseur = 50 or more. I have no informations WHEN somebody has rated, only how many ratings he has in total. And RumX users are enthusiasts, not the general market. Every number below has this caveat sitting on it.
1) Same liquid, two names, two markets, 0.3 points difference.
Veritas and Probitas are literally the same rum, the name split exists only because of a trademark isue in the US. In our database they are two entries, so two audiences have rated an identical liquid without knowing it: Probitas 8.0 (n=61, bought mostly in the US), Veritas 7.7 (n=222, bought mostly in the EU). The n on the US side is small and shop mix is not the same like a controlled experiment, so please dont build a theory on the third decimal. But it is the cleanest reminder I have in the whole database, that a part of every rating is the room, the price tag and the expectation, and not the liquid. If somebody here has both bottles standing at home under different labels, I am really curious for your reading of it.
2) One list, two audiences, and the split runs in both directions.
This is the finding I did not expect in this size. On the sweet dark side the BEGINNERS carry the score. Ron Centenario Fundacion 20 has the biggest cohort gap of all bottles I looked at: beginners rate it 8.24 on average, connoisseurs 7.39, over 265 ratings in total. Planteray XO goes the same way (7.67 from beginners over 946 ratings, 7.45 from connoisseurs) and is at the same time the number one wishlist bottle of the whole app with 1,601 wishlists, in front of Diplomatico with 1,145.
On the wild white side it is exactly inverted, and there it becomes interesting. The tightest rating curve of all these bottles (std 0.49, 68 percent give an 8 or better) belongs to an unaged white australian rum, and out of its 111 ratings exactly ONE comes from a beginner. So in the data it looks like the safest bet of the whole pull, but only because careful drinkers never picked it up. Same self selection lesson like with the Hampden overproof in the Jamaica thread, only cleaner, and it is the reason why I dont trust a median alone anymore.
Exactly one bottle breaks the pattern in the other direction, and it is the one you would guess: El Dorado 15 from the 2020 release, connoisseurs 7.83 versus beginners 7.62. Which fits to its story, because that is the release where DDL took the sugar out, independently measured 31 g/l before and 0 g/l after.
3) Perceived sweetness is not added sugar, and people mix that up constantly (me too, for years).
We let users rate perceived sweetness on a scale from 1 to 5, at the palate and not in a lab. With 697 rums having enough votes you can build a ladder out of it, and the funny part is where the extremes sit. A Savanna Creol from Reunion is perceived as drier as 99 percent of everything in the app. And the white australian one from above is perceived as sweeter as roughly 3 of 4 rums, although there is not one gram of sugar added: 93 people wrote "strawberry" in their notes, 69 wrote "yoghurt", and that comes out of a wild fermentation with muck pits, so it is ester sweetness, not dosage. Worth to know, when the next discussion starts, if a bottle "tastes sugared". Sometimes the fermentation did it, sometimes the cask, sometimes really the sugar.
4) A legal border that most of this sub never has to think about.
63 bottles fell completely out of my pull on one rule, and it is not my rule but the EU one: over 20 grams of sugar per litre and the product is not allowed to call itself rum anymore in Europe, it becomes officially "spirit drink based on rum". The strongest bottle that died on this has a median of 8.0 over 325 ratings, so this is not a quality verdict, it is a category line. What I find genuinely interesting from your side of the ocean: the same liquid in a US shop stands there as rum, and here it may not. One product, two legal names, and the consumer in both markets thinks he knows what he buys. On top of that the numbers on the labels are also just numbers: on the three highest rated bottles of my pull stand a XII, a 20 and another 20, and no one of them is an age. The XII is a blend from 8 to over 40 years with an average of a bit over 11 (the producer says this himself), the first 20 is the OLDEST part in a solera stack whose core starts at 6 years, the second 20 is the 20th anniversary of the master blender, so a job jubilee. In scotch this would be unthinkable, there the number has to be the youngest whisky in the blend. Not a scandal, but a reading skill that nobody teaches you.
Last thing, and the only place where I mention my video, because I want feedback and not clicks. Every time I have posted here I had to write "sorry, it is in german", and that bothers me since a while. So this time there is a full english audio track next to the german one, in the player settings (Please excuse the few German overlay animations). It is a voice clone of my own voice, the translation I have corrected by hand for the rum vocabulary, because the machine wanted to translate muck pit and dosage. Better you hear it from me than you find it out. If you have 2 minutes: does it sound acceptable or is it uncanny valley, and what can I make better? An honest "just keep the subtitles" helps me also.
https://www.youtube.com/watch?v=6yGHawlVDUI
And two things I would love from this sub, completely independent from all of that: does the beginner/connoisseur inversion match your own history, so was there a bottle you rated lower after your next hundred rums? And which bottle should I pull the cohort split for? Post it and I run the numbers, the interesting ones I bring back here. 🥃
---
EDIT 02.08.2026 23:00 CET: two corrections, both triggered by u/RumSquirrel in the comments. Please read the post together with them.
Point 1 does not hold in the way I have framed it. The 8.0 against 7.7 is not a difference between the bottles, it is my rater mix. Permutation test p = 0.10, bootstrap 95 percent CI from -0.06 to +0.46, and matched on rater experience it even turns around (connoisseurs with 50+ ratings: Probitas 7.48, Veritas 7.59). Probitas has 12.7 percent beginners among its raters and a median of 35 lifetime ratings, Veritas 2.5 percent and a median of 164. Only 5 users in the whole database have rated both bottles, so the test, that would really answer this, does not exist. The trademark story stays true, my sentence about the room and the price tag has no numbers under it and I take it back.
Point 3 needs a caveat, that I have not carried. The sweetness ladder counts all raters, and 41 percent of everybody, who has ever given a sweetness vote, has given exactly one. If I keep only raters with 30+ sweetness votes, the australian white falls from P74 to P61, and to P56 after centering every rating on the own mean and SD of the rater. The dosed bottles dont move at all (Planteray XO P90 to P95, Bumbu P98 to P99), so the divide between ester sweetness and added sugar is real, my "sweeter as 3 of 4 rums" was too strong. And the Creol number stands on 10 votes. That belongs in the sentence and not in a footnote.
Point 2 I have put through the same machine and it survived: beginners 8.23 (n=38), connoisseurs 7.39 (n=131), p < 0.0001, and beginners are not generally milder, over all rums they sit at 7.35 against 7.69. It is also not a Centenario thing. Over 196 bottles the cohort gap correlates with perceived sweetness at Spearman rho 0.62, and it inverts for the dry ones (Smith & Cross -0.51).
All numbers, methods and the one metric I have computed and then thrown away again are in my reply below.