Versioned lineage and downloads
Every public metric is tied to release 2026.08.09-r2, source and corpus checksums, classifier and entity-graph versions, and an independent QA run. The manifest records byte counts and SHA-256 checksums for each file.
Methodology · review-evidence-v1.0.0
Law Leaderboard compares public firm profiles using rating, review depth, recommendation language, communication mentions, owner responses, consistency, and negative-sentiment signals. The result is a public-review evidence score, not a judgment about legal ability or likely outcomes.
The current dataset covers 1,130 identity-matched public Google Business firm profiles across 25 ranking markets. It tracks 527,168 public reviews, contains 450,536 retrieved review records, and analyzes 365,756 review texts. 1,100 profiles currently pass the scoring gate; the others remain visible with no score or rank.
Active consumer release llb-consumer-intelligence-2026-08-12-r4 publishes
121 practice/city comparisons and
1,029 distinct qualified offices.
Each assessed cell targets 50 discovered office profiles, and a page is published only with at least eight qualified offices.
Entity graph · entity-graph-v1.0.0
Release llb-intelligence-2026-08-09-r2 distinguishes a law-firm brand from its offices: 464 brands, including 28 multi-office brands, and 523 offices are represented. Stable opaque Law Leaderboard brand, office, and firm IDs support traceability across versioned releases, while raw provider identifiers remain private.
Review intelligence · review-themes-v1.0.1
The current classifier processed 229,403 usable texts from 275,801 review records and assigned 36,263 private theme labels across 11 themes. Only 3 themes pass every publication gate: Communication, Responsiveness and delays, Outcome language (unverified). Outcome language records what a reviewer claimed; Law Leaderboard does not verify a result, recovery, settlement, or verdict.
| Theme | QA positives reviewed | QA precision | Release status |
|---|---|---|---|
| Communication | 55 | 96.4% | Published |
| Responsiveness and delays | 38 | 94.7% | Published |
| Staff accessibility | 27 | 44.4% | Suppressed |
| Case handoffs | 24 | 62.5% | Suppressed |
| Intake experience | 27 | 14.8% | Suppressed |
| Fees | 24 | 25.0% | Suppressed |
| Process clarity | 27 | 29.6% | Suppressed |
| Follow-up | 31 | 54.8% | Suppressed |
| Timeliness | 26 | 42.3% | Suppressed |
| Compassion | 39 | 79.5% | Suppressed |
| Outcome language (unverified) | 44 | 81.8% | Published |
Independent validation used gpt-4o-mini-2024-07-18 on a deterministic stratified sample of 483 reviews. Manual source inspection covered 27 positive-evidence records: 15 API agreements and all 12 API-disagreed positives for the published themes. Explicit evidence appeared in all 27 inspected records. Neither check establishes perfect labels or complete recall. 8 themes remain available for private classifier improvement but are absent from public office metrics in this release.
Every public metric is tied to release 2026.08.09-r2, source and corpus checksums, classifier and entity-graph versions, and an independent QA run. The manifest records byte counts and SHA-256 checksums for each file.
Trend fields are prepared but suppressed unless at least two distinct, complete profile observations exist. A suppressed trend means the evidence window is incomplete; it does not mean review activity or Maps visibility was unchanged. The graph currently preserves 363,504 review-to-snapshot links for future comparable releases.
The current consumer release targeted 50 discovered office profiles in each of 150 assessed practice-area-by-city cells across six practices and 25 cities. Stable provider identities and entity matching prevent one office from becoming multiple public profiles.
We collect observed public Google Business Profile fields and retrieve public review records through a third-party data provider. Tracked public review count, retrieved review records, and analyzed review text are three different coverage measures.
Versioned deterministic rules identify explicit review language across 11 defined themes. The private graph retains all labels for QA and improvement; public office metrics are emitted only for themes that pass every sample, confidence, and independent-validation gate.
A profile receives a score only when every eligibility field is present and at least 50 review texts were analyzed. Practice classification for a city membership also requires an explicit public-profile practice signal or sufficient context-classified review evidence. A market page requires at least eight offices that pass both gates.
Eligible profiles receive a weighted sum of seven 0-to-100 components. Rate-like signals are adjusted toward fixed dataset priors so smaller evidence samples have less influence. Component and weighted values are retained to three decimals before the final score is shown to one decimal.
Version review-evidence-v1.0.0 returns a theoretical 0.0-to-100.0 weighted score, displayed to one decimal. Component scores and weighted points are retained to three decimals before final rounding. In the current export, eligible scores range from 31.5 to 93.7, with a median of 72.20. Those observed values will change when the evidence changes.
| Component | Weight | Reproducible transform | Evidence meaning |
|---|---|---|---|
| Public rating | 25% | clamp((shrunk rating − 4.0) × 100, 0, 100) | Uses tracked review count as the shrinkage sample size. |
| Review depth | 10% | clamp(100 × ln(1 + tracked reviews) ÷ ln(1 + 2,500), 0, 100) | Rewards more observed public-review evidence with diminishing returns and a 2,500-review cap. |
| Explicit recommendation | 20% | shrunk recommendation percentage | Measures explicit recommendation language in the analyzed text sample. |
| Communication mentions | 15% | clamp(shrunk communication percentage ÷ 40 × 100, 0, 100) | Maps communication mentions onto a 0-to-100 component scale. |
| Public owner response | 10% | shrunk owner-response percentage | Measures public replies to retrieved reviews, not communication during a case. |
| Consistency | 10% | shrunk rating-consistency index | Uses the exported polarization/consistency measure. |
| Low concern | 10% | clamp(100 − 5 × shrunk negative-sentiment percentage, 0, 100) | A lower adjusted negative-sentiment share produces a higher component score. |
Rate-like inputs use a fixed prior strength of 50:
(raw × n + prior × 50) ÷ (n + 50)
For rating, n is tracked review count. For recommendation, communication, owner response,
consistency, and negative sentiment, n is analyzed text-review count. Owner response is converted
from a 0-to-1 value into a percentage first.
Fixed priors: rating 4.875/5; recommendation 60.0%; communication 17.9%; owner response 67.6%; consistency 92.6/100; negative sentiment 3.0%.
Maps rank and visibility do not enter the tie-break sequence.
Evidence depth describes the number of analyzed text reviews: insufficient below 50, minimum from 50–99, moderate from 100–249, strong from 250–999, and extensive at 1,000 or more. Only the 50-review eligibility floor affects whether a score exists; the label is otherwise descriptive.
Public intelligence files contain no reviewer identity, review text, owner-response text, provider payload, raw Place ID or CID, street address, or phone number. They use opaque Law Leaderboard IDs for versioned traceability. Because raw provider identifiers remain private, the public files cannot independently reproduce the private identity-matching process. The versioned release is rights reserved and is not offered under an open-data license.
Inspect the versioned aggregate release, field definitions, checksums, and usage terms before reusing a figure.