Skip to content
Playbook RG

Playbook Brand · Brand book · Measure

Content scorecard

Ready-to-run surveys, NPS benchmarks, cultural-fit testing, and a decision framework for measuring literacy content like a product.

12 min read · CC0 public domain

How to measure whether your Playbook RG content is working – the same way you’d measure any marketing campaign. Three surveys, auto-scored benchmarks, and a decision framework. Designed for marketing teams; no research expertise required.


Tier 1 content is a product, not a pamphlet

Most operators treat literacy content as a compliance checkbox – something to publish and forget. Playbook RG treats it as a product. It competes for the same player attention as your sportsbook promos, your casino loyalty emails, your brand campaigns. If you’d measure those with NPS, measure this with NPS.

The quality bar is commercial-grade, not “good enough for compliance.” Products have defined audiences, performance targets, iteration cycles, and quality standards. That’s why Playbook RG uses the same measurement tools – NPS, engagement rates, A/B testing – that product and marketing teams already use.

You already A/B test email subject lines and measure campaign engagement. This is the same process for your literacy content – same tools, same rigor, same time investment.


Why NPS?

Net Promoter Score is the standard for measuring whether a product earns genuine advocacy. We use it for Playbook RG content because Tier 1 content should hit the same targets as any other player-facing product.

How it compares:

CategoryTypical NPS range
Streaming entertainment (Netflix, Spotify)50--70
Gaming / sportsbook apps30--50
Playbook RG “Strong” target30--49
Traditional compliance contentUnmeasured (or negative)

The punchline: Playbook RG targets the same range as the products it sits alongside. Traditional compliance content doesn’t even ask.


Methodology and sources

This scorecard uses established marketing research tools. Where we’ve adapted questions for the Playbook RG context, that’s noted explicitly.

ComponentStatusSource
Net Promoter Score (NPS)Validated instrument, standard 0--10 scaleReichheld (2003)
Semantic differential methodValidated methodology, cross-culturally stableOsgood, Suci, & Tannenbaum (1957)
Forgettable – Memorable pairValidated for message evaluationDillard, Shen, & Vail (2007)
Preachy – Respectful pairAdapted – anchored to Playbook brand DNANot independently validated
Generic – Made for me pairAdapted – cultural fit indicatorNot independently validated
Free recall + recognitionStandard research methodologyEstablished procedure

Use validated components with confidence. Treat adapted components as directional.


Pulse survey

3 items. Under 1 minute. Minimum viable signal.

Use after any content launch, campaign wave, or quiz deployment. Distribute via in-app intercept, post-quiz screen, follow-up email, or QR code on print materials.

Sample size: 30 responses minimum.

Copy-paste these questions into your survey tool, or use the ready-to-deploy survey templates below:

1. How likely would you be to recommend this content to a friend
   who gambles?

   0  1  2  3  4  5  6  7  8  9  10
   Not at all likely          Extremely likely

2. Rate this content:

   Forgettable  1  2  3  4  5  6  7  Memorable

3. What's one thing you'd change? (optional)

   [open text]

Full survey

7 items. About 2 minutes. More diagnostic.

Use when you need to understand why content is or isn’t working – not just whether it is. Recommended for new market launches, major campaign evaluations, and cultural adaptation testing.

Sample size: 30 responses minimum. For cultural fit A/B tests, 30 per group (60 total).

1. How likely would you be to recommend this content to a friend
   who gambles?

   0  1  2  3  4  5  6  7  8  9  10
   Not at all likely          Extremely likely

2. Rate this content on each scale:

   Forgettable  1  2  3  4  5  6  7  Memorable
   Boring       1  2  3  4  5  6  7  Engaging
   Preachy      1  2  3  4  5  6  7  Respectful
   Confusing    1  2  3  4  5  6  7  Clear
   Generic      1  2  3  4  5  6  7  Made for me

3. What's one thing you'd change? (optional)

   [open text]

Why these pairs

PairWhat it testsSource
Forgettable – MemorableMessage stickinessDillard et al. (2007)
Boring – EngagingEntertainment valueBrand DNA: “Treats gambling as entertainment”
Preachy – RespectfulVoice registerBrand DNA prohibition: never lecture, never patronize
Confusing – ClearComprehensionBrand DNA: “Generous with information”
Generic – Made for meAudience fit / cultural resonanceCultural adaptation validation

Recall survey

4 items. Administered 7--14 days after exposure.

This is a separate follow-up survey, not part of the Pulse or Full. It answers: What stuck?

Sample size: 25 responses minimum. Expect 40--60% response rate from players who completed an earlier Pulse or Full survey.

We showed you some content about [TOPIC] about [X] days ago.
Without looking anything up:

1. What's the main thing you remember? (one sentence is fine)

   [open text]

2. True or false: [KEY FACT FROM YOUR CONTENT]
   ( ) True  ( ) False  ( ) Not sure

3. True or false: [SECOND KEY FACT FROM YOUR CONTENT]
   ( ) True  ( ) False  ( ) Not sure

4. Since seeing the content, have you done any of the following?
   (check all that apply)
   [ ] Looked up more information about [topic]
   [ ] Used a tool or feature mentioned (e.g., set a limit, tried a calculator)
   [ ] Shared the content or told someone about it
   [ ] None of the above

Writing true/false items

Questions 2 and 3 must be customized for each piece of content. Derive them from the single most important fact in the content. Three worked examples:

Content typeSource factTrue/false item
Myth-busting cardSlots outcomes are determined by a random number generator“True or false: A slot machine that hasn’t paid out in a while is more likely to pay soon.” (Answer: False)
Odds explainerHouse edge on American roulette is 5.26% vs. 2.70% on European“True or false: American roulette has a higher house edge than European roulette.” (Answer: True)
Tool promotionDeposit limits can be set in account settings“True or false: You can set a deposit limit on your account.” (Answer: True)

Pick facts the content explicitly states. Avoid trick questions.


Cultural fit testing

Not a separate instrument – run the Full survey as an A/B test.

  • Group A sees the default Playbook content
  • Group B sees the culturally adapted version
  • Both groups answer the same 7 questions
  • Compare average scores between groups

Sample size: 30 per group minimum (60 total).

Primary comparison metric: The Generic – Made for me pair. If Group B scores meaningfully higher (0.5+ points), the adaptation is improving cultural resonance.

Secondary metrics by spectrum changed:

Spectrum adaptedPrimary pair to compare
Voice (e.g., peer to authority)Preachy – Respectful
Framing (e.g., individual to communal)Generic – Made for me
Humor (e.g., irreverent to warm)Boring – Engaging
Directness (e.g., blunt to diplomatic)Preachy – Respectful
Comfort (e.g., open to reserved)Generic – Made for me

Rules: Follow the A/B testing guidance in calls-to-action.md. One variable at a time. Minimum 2 weeks or 1,000 impressions per variant. Measure action, not just clicks.

Decision guide:

NPS comparisonSemantic pairs comparisonAction
B >= AB higher on 2+ pairsAdaptation works. Deploy adapted version. Share findings.
B >= AB ~= A (within 0.3)Adaptation is neutral. Default may be fine for this market.
B < AAnyAdaptation may have drifted from brand DNA. Review against What doesn’t change.

Interpreting results

NPS

Standard NPS scoring: % Promoters (9--10) minus % Detractors (0--6). These are the same tiers used for any consumer product – because that’s what this content is.

NPSReadingProduct comparison
50+Exceptional – content is spreading organicallyHigh-performing consumer products
30--49Strong – content connects with this audienceCompetitive with gaming/sportsbook apps
0--29Acceptable – room to improveBelow product standard – iterate
Below 0Problem – more players would discourage than recommendWould not ship a product feature at this score

Semantic differentials

7-point scale. Midpoint is 4.

AverageReading
5.5+Strong signal on this dimension
4.0--5.4Acceptable
Below 4.0Below midpoint – investigate. Check open-ended responses for patterns.
Below 3.0Red flag. Content may be failing on this dimension.

Brand DNA alarm: If Preachy – Respectful scores below 4.0, the content may be lecturing players. Review against voice principles.

Recall

MetricStrong signalWeak signal
Free recall (Q1)Player states the key fact“I don’t remember” or mentions only secondary details
True/false (Q2--3)Both correctOne or both wrong / “not sure”
Behavioral follow-through (Q4)Any action checked“None of the above”

If true/false accuracy is low but NPS was high, the content is entertaining but not teaching effectively. Simplify the key message or make it more visually prominent.


Running a test

  1. Pick your survey. Pulse for quick signal. Full for diagnostics or cultural fit testing. Recall for follow-up.
  2. Deploy it. Three options:
  • Fastest: Use the embeddable survey widget – drop it into your content hub or embed via iframe. Auto-scores results.
  • Your platform: Copy from the survey templates below into Google Forms, SurveyMonkey, Typeform, or your built-in survey tool.
  • Manual: Copy-paste the questions from the survey sections above.
  1. Customize recall true/false items for your content. Set distribution channel (in-app, email, QR code).
  2. Collect responses. Close when you hit your target sample size or after 2 weeks, whichever comes first.
  3. Report. Calculate averages (the widget does this automatically), read the open-ended responses for patterns, and file a Research Findings issue to share back.

Total effort: 3--4 hours spread across 2 weeks.


Reporting format

Use this structure when filing results. It maps directly to the GitHub issue template.

Market profile:    [e.g., authority / communal / warm / diplomatic / reserved]
Operator:          [name or "anonymous"]
Content tested:    [title + file reference]
Survey type:       [Pulse / Full / Recall / Cultural Fit A/B]
Player segment:    [e.g., general-players, young-adults, sports-bettors]
Sample size:       [N]
Dates:             [start -- end]

NPS:               [score]
Forgettable/Memorable: [avg]
Boring/Engaging:       [avg, if Full]
Preachy/Respectful:    [avg, if Full]
Confusing/Clear:       [avg, if Full]
Generic/Made for me:   [avg, if Full]

Qualitative patterns:  [2--3 sentences from open-ended responses]
Recommendation:        [Ship / Revise and retest / Escalate to brand owner]

Segment notes

The survey questions are the same for every segment. Distribution channel and framing differ.

SegmentNotes
General playersStandard distribution, any channel
Young adults (18--25)Mobile-first. In-app intercept or post-quiz screen preferred.
Sports bettorsTie to live event campaigns. Survey after event-timed content.
At-risk playersDo not survey. This protocol tests Tier 1 content with general audiences only.
Friends and familyMay reach via email or support pages. Adjust distribution accordingly.
Help-seekersDo not survey. Tier 2 content is outside this protocol’s scope.

Quick-reference decision tree

Tape this to your monitor. See the printable version.

After any content launch:
                                    ┌──────────────────┐
                                    │   Send PULSE      │
                                    │   (3 Qs, <1 min)  │
                                    └────────┬─────────┘
                                             │
                              ┌──────────────┼──────────────┐
                              ▼              ▼              ▼
                        NPS 30+         NPS 0--29       NPS < 0
                     All pairs 4.0+   or pair < 4.0    or Preachy < 3.0
                              │              │              │
                              ▼              ▼              ▼
                         ┌────────┐   ┌────────────┐  ┌──────────┐
                         │  SHIP  │   │ Run FULL   │  │  STOP    │
                         │  ✓     │   │ survey for │  │  Revise  │
                         └────┬───┘   │ diagnostics│  │  content │
                              │       └────────────┘  └──────────┘
                              ▼
                    ┌───────────────────┐
                    │ Send RECALL       │
                    │ (7--14 days later)│
                    │ Did it stick?     │
                    └───────────────────┘

Color guide for scores:

ScoreSignalAction
NPS 30+ and all pairs 5.0+Strong – ship and shareNo changes needed
NPS 0--29 or any pair 4.0--4.9Acceptable – room to improveRun Full survey, identify weak dimensions
NPS < 0 or any pair < 4.0Investigate immediatelyCheck open-ended responses, review against voice principles
Preachy--Respectful < 4.0Brand DNA alarmContent may be lecturing – revise before shipping

Survey templates

Ready-to-use configurations for popular survey platforms. Copy the setup, customize for your content, and deploy.

Google Forms

Pulse survey:

  1. Create a new Google Form
  2. Add a Linear scale question (0--10): “How likely would you be to recommend this content to a friend who gambles?”
  3. Add a Linear scale question (1--7) with label “Forgettable” on left and “Memorable” on right
  4. Add a Short answer question: “What’s one thing you’d change?” (mark as optional)

Full survey:

  1. Follow Pulse steps 1--2 above
  2. Add four more Linear scale questions (1--7) with these label pairs:
  • Boring / Engaging
  • Preachy / Respectful
  • Confusing / Clear
  • Generic / Made for me
  1. Add the optional open-ended question

Recall survey:

  1. Create a new Google Form
  2. Add a Short answer: “What’s the main thing you remember?”
  3. Add two Multiple choice questions with True / False / Not sure (customize the statements for your content)
  4. Add a Checkbox question: “Since seeing the content, have you...” with the four behavioral options

Typeform

Pulse survey:

  1. Create a new Typeform
  2. Add an Opinion Scale block (0--10) for NPS
  3. Add an Opinion Scale block (1--7) with custom labels: Forgettable / Memorable
  4. Add a Short Text block for open-ended feedback (optional)

Full survey: Same as Pulse, plus four additional Opinion Scale blocks (1--7) for each semantic pair.

SurveyMonkey

Pulse survey:

  1. Create a new survey
  2. Add an NPS question type (built-in, auto-calculates score)
  3. Add a Matrix/Rating Scale question (1--7) for Forgettable--Memorable
  4. Add an Open-Ended question for feedback

Full survey: Same as Pulse, plus four additional Matrix rows for the remaining pairs.

Embeddable widget (no survey tool needed)

Use the self-contained survey widget. It runs in-browser, requires no backend, and auto-calculates all scores with color-coded benchmarks. Embed via iframe or link directly:

<iframe src="survey-widget.html" width="100%" height="700" frameborder="0"></iframe>

References

Methodology:

  • Reichheld, F. (2003). The one number you need to grow. Harvard Business Review, 81(12), 46--54. – The Net Promoter Score.
  • Osgood, C. E., Suci, G. J., & Tannenbaum, P. H. (1957). The Measurement of Meaning. University of Illinois Press. – Semantic differential methodology.
  • Dillard, J. P., Shen, L., & Vail, R. G. (2007). Does perceived message effectiveness cause persuasion or vice versa? 17 consistent answers. Human Communication Research, 33(4), 467--488. – Validated message evaluation semantic differential pairs.

Playbook cross-references: