Version LSR-review-score-v1.1-lexical

How the Montréal report works.

This is not a critic’s list and it is not a star-rating clone. It is a structured reading of 64,738 public reviews across 69 Montréal shawarma locations.

65%Food quality
25%Review value
10%Consistency
01

Reviews become evidence, not votes.

Written reviews are classified into food quality, meat, sauce, bread and freshness, portion, affordability, value, and consistency. A review only influences a component when it actually discusses that component. At most one contribution per review per component, so one long review cannot outweigh ten short ones.

02

Small samples are not allowed to cosplay as certainty.

Every component is Bayesian-smoothed toward the Montréal-wide prior. Evidence counts and confidence labels remain visible. Locations with enough total reviews receive an official rank; smaller samples remain provisional, and locations without enough written review evidence are shown as unscored rather than assigned an invented score.

03

Every location gets exactly one category.

Locations are classified as shawarma-primary, shawarma-on-menu, or novelty/hybrid. Shawarma-primary locations can draw on their whole review corpus; shawarma-on-menu and novelty/hybrid locations are only scored from shawarma-relevant evidence within their reviews.

04

Coverage and confidence are disclosed, not hidden.

Each location shows the ratio of downloaded to expected reviews and a coverage status (complete, slightly short, materially incomplete, or zero). Restaurants that could not be safely resolved to a current Google listing, or that closed, were excluded rather than force-matched — see the 3 entries in this market’s reconciliation record.

05

The Toumometer is a joke with a methodology.

Only explicit evaluative mentions of toum or garlic sauce count. Those mentions are recency-weighted and smoothed, which is far too much statistical ceremony for a garlic-sauce leaderboard. That is the point.

06

Model version and limitations.

This market was scored with LSR-review-score-v1.1-lexical, a from-scratch, transparent reimplementation of the published London methodology (weights, Bayesian smoothing, recency curve, one-review cap, evidence gates). The original private London sentence-model code was not available when this market was built, so a documented lexical text-sentiment model is used for the 85% text-derived component instead. It is not claimed to be numerically identical to the original London model.

Reviews were analyzed in English; non-English reviews were retained in the corpus but are not yet scored by a validated language-specific model. This affects Montréal most, where a meaningful share of the corpus is French.