The courtyard made it easy to meet people.
- Social vibe?YES
We keep the full sample, extract social evidence, rank it against the cohort, then calibrate the 0-100 display.
A fixed illustrative hostel runs through all six steps. Ten representative lines show the evidence; the featured social signal uses 64 evaluated reviews.
Worked example. Raw rate, Wilson bound, confidence, freshness, and receipt arithmetic follow the exact public v1 policy. Cohort rate, standardized values, deductions, percentile, and score 87 are fixed illustrative values because they depend on a population snapshot.
Each signal counts one true or false result per evaluated review. For the featured social signal, 64 reviews are evaluated; these ten lines are representative, not the full denominator. The 3 neutral examples show how non-mentions remain false for that signal instead of vanishing.
We highlight the phrases that support social, negative, antisocial, or integrity signals. Scoring records at most one true or false result per review for each independent signal; it does not count phrases. One review can activate several signals at once.
Review-level true counts become per-signal rates. In the example, 22 of 64 evaluated reviews describe the place as social — a raw rate of 34%. A Wilson lower bound trims that to 29% before cohort comparison. In this illustrative snapshot, a typical hostel sits near 18% and the conservative rate standardizes to about +1.3. Core signals get the same treatment; there is no positive-review fraction.
The illustrative weighted core adds up to +1.12. With 64 evaluated reviews for the featured signal, evidence confidence reaches 0.80 of full strength, and a latest collected review about 7 months old sets freshness to 0.75. Multiplying the three leaves +0.67. Only a positive core is shrunk this way; thin or stale evidence never softens a negative result.
× 0.80 evidence confidence × 0.75 freshness
Counter-evidence is subtracted, never averaged away. In this illustrative receipt, above-cohort negative social evidence contributes −0.09 and the combined integrity contribution is −0.05, producing a raw rank of +0.53. The free-drink line can feed integrity signals, but one line has no fixed deduction; warnings can remain visible separately.
In this fixed illustrative cohort snapshot, a raw rank of +0.53 sits roughly in the top 12%. Quantile calibration maps that relative position onto the familiar 0-100 display: 87. The number reads like a score, but it states a relative claim about rank, not an overall quality rating.
+0.53 · Top 12% of the eligible cohort
The page combines several evidence systems, but they do not all enter the Punk Score equation.
Incentives, review pressure, and explicit distrust can reduce the relative rank. These are risk corrections, not fraud verdicts.
Bedbug and safety evidence can be urgent, but neither is a Punk Score input.
Useful hostel facts appear around the score without changing it.
These are the current production rules. Coefficients are weights in a standardized model, not percentages or portions of a final score.
hp_fun_social_rank_v1fun_social_rank_v1social_activity_conf_fresh_v3A hostel enters public v1 ranking only when the required social signal has at least 20 evaluated eligible reviews.
Each rate uses a Wilson lower bound with z=1, then the core rates are log-transformed and standardized against the eligible hostel cohort. Coefficients are model weights, not percentages.
z = 1(p + z^2/(2*n) - z*sqrt((p*(1-p) + z^2/(4*n))/n)) / (1 + z^2/n)zscore(ln(1 + 100 * wilson_lower_bound), eligible_hostel_cohort)| Signal | Coefficient |
|---|---|
| Described-as-social rate | 0.50 |
| Social activity/community rate | 0.35 |
| Explicit fun-language rate | 0.05 |
| Solo-traveler recommendation rate | 0.05 |
| Reviewer travel-experience rate | 0.05 |
| Signal | Coefficient | Minimum evaluated reviews |
|---|---|---|
| Community depth | 0.12 | 5 |
| Meaningful conversations | 0.12 | 3 |
| Interesting people | 0.10 | 3 |
| Reviewer social curiosity | 0.08 | 3 |
| Positive atmosphere evidence | 0.06 | 5 |
| Fun host or owner | 0.06 | 3 |
180 evaluated reviews reaches full evidence confidence. The formula and freshness factor multiply only a positive core.
min(1, ln(1+n) / ln(181))
The latest collected linked review in the daily review rollup selects one factor. A missing date uses its own factor.
| Latest collected review | Factor |
|---|---|
| 0-180 days | 1.00 |
| 181-365 days | 0.75 |
| 366-730 days | 0.50 |
| More than 730 days | 0.30 |
| Latest review date missing | 0.65 |
All listed values are subtracted after positive-only shrinkage. A positive z part means below-cohort risk adds no penalty.
1000 quantile buckets map raw rank order to the current legacy display distribution. If a bucket has no match, the shown fallback formula is clamped to 0-100.
clamp(0, 100, 50 + 12 × raw_rank)
Bedbug and safety evidence are not rank penalties in public v1. They stay separate warnings.
The experimental v2 semantic boost for social outcomes, shared-interest bonding, and host social mechanisms is not selected by public v1.
Punk Score is neither an overall hostel-quality rating nor a prediction of your future visit. Read the rank, confidence, and warnings together.