The short answer
Scoring has two ingredients: 1. What you set up in the scenarioWhen authoring a scenario, you define a learning objective and up to three skills — each with a name and description. Optionally you can add good/bad examples and specific achievements to look for. This is the foundation: the session is scored against your criteria, not a generic checklist. 2. What happened in the conversation
When the session ends, the transcript is reviewed for evidence that the learner demonstrated those skills and met the learning objective. From that review, learners receive:
- A rating for each skill — from Novice to Expert, with examples and tips
- An overall score out of 100 — how well they met the learning objective as a whole
- Highlighted moments from the conversation — where they did well and where they could improve
Ingredient 1 — Skills and the learning objective
Everything starts with how the scenario is configured.Ingredient 2 — Reviewing the transcript
When the session finishes, the conversation transcript is analysed for evidence of the skills in action and progress toward the learning objective. The review looks at what the learner said and did — not what they intended. Each skill is judged on the whole conversation: how often, how clearly, and how consistently that skill showed up. The system is looking for evidence of the skill you described, not a keyword match or a single required phrase. One good moment on its own is usually not enough for a high rating. Learners also see key moments pulled from their conversation — where a skill was demonstrated well, or where there was a missed opportunity. The full transcript is always available for admins and reviewers.What learners receive
Per-skill rating (Novice → Expert)
Each skill gets one of five levels:Overall score (0–100)
This is a separate assessment of how well the learner met the learning objective across the whole conversation. It is not a simple average of the skill ratings. In general:- Low scores (roughly below 50) — limited engagement, missed the objective, or significant gaps in execution
- Mid scores (roughly 50–70) — partial success, some skills shown but inconsistent
- Strong scores (roughly 70–89) — clear progress toward the objective, skills applied well
- Exceptional scores (90+) — outstanding performance; intentionally hard to achieve
What affects the score — and what does not
What matters
- Engaging meaningfully with the scenario topic
- Demonstrating the defined skills through what they actually said
- Consistency across the conversation, not just one strong reply
- Appropriate tone and professionalism for the situation
What does not automatically help
- Longer conversations do not automatically score higher. A long but unfocused session can still score poorly. What counts is the quality of engagement, not the word count.
- The overall score is not an average of skill ratings. A learner could do well on one skill but still miss the broader objective.
- Very short sessions may not receive full feedback. Scenarios can require a minimum session length, and sessions with almost no content will score at or near zero.
Common questions
Does a longer conversation automatically score higher?
Does a longer conversation automatically score higher?
Is the overall score an average of the skill ratings?
Is the overall score an average of the skill ratings?
Why did someone get a very low score or no feedback?
Why did someone get a very low score or no feedback?
Why is it hard to get 90+?
Why is it hard to get 90+?
Can two similar sessions get different scores?
Can two similar sessions get different scores?
When a learner queries their rating, what should we point them to?
When a learner queries their rating, what should we point them to?
Is a skill judged on the conversation as a whole, or on particular things?
Is a skill judged on the conversation as a whole, or on particular things?
Should we add more good and bad examples?
Should we add more good and bad examples?
What if the conversation never goes near one of the skills?
What if the conversation never goes near one of the skills?
Why only three skills — should we use fewer?
Why only three skills — should we use fewer?
If someone ran the same conversation twice and said much the same thing, how close would the scores be?
If someone ran the same conversation twice and said much the same thing, how close would the scores be?
Worked example
A coaching scenario with this setup:Suggested talking points
When a learner or line manager asks about their score:- Open their session first — the score explanation, skill quotes, and highlighted turns are the rationale.
- The score reflects the learning objective — not attendance, effort alone, or how long they talked.
- Skills are defined upfront — the scenario author decides what good looks like for that conversation.
- Feedback is evidence-based — ratings and highlights refer to specific moments from the conversation.
- A missed skill still gets a rating — no evidence usually means Novice or Beginner, not “not applicable.”
- Low scores are actionable — learners get concrete next steps and can practise again immediately.
- Practise again is the point — a score measures one attempt; improvement comes from repeated practice.
Related
- Feedback & evidence — what learners see after a session
- Scenario creation — how to configure skills and objectives
- Admin & organizations — session history, analytics, and exports
