01Training
How to score a sales role-play without fooling yourself
· 2 min read
A score without the sentence that backs it up does not help anyone improve. Observable criteria, literal quotes and progress over time: that is how a rehearsal becomes learning.
Criteria you can hear
The first mistake is scoring what cannot be seen. “Positive attitude” or “confidence” depend on who is listening. A good criterion describes a behaviour you can point to in the transcript: “asks about the need before presenting the product”, “rephrases the objection before answering it”, “proposes a next step with a date”. Eight to ten criteria are enough; beyond that, the report turns into noise.
Every criterion needs an anchored scale. If it runs from 0 to 10, you have to write down what a 3, a 6 and a 9 look like, with examples. Without anchors, two raters, human or not, will score the same conversation differently.
Every score, with its quote
A score is only useful when it comes with the sentence that justifies it and the minute it was said. The quote lets the salesperson understand what went well or badly, and lets the trainer check whether the assessment is fair. If the system cannot find a sentence that supports a score, that score should not exist.
Regulated sectors need one more check: whether what was said about the product is correct according to the official documentation. A brilliant conversation with a false claim in it is a risk, not a success.
The trend, not the snapshot
A single session says little: the scenario, the day and the simulated customer all play a part. What matters is the trend in the same scenario across several sessions, which criteria improve and which ones stall. Only raise the difficulty of the simulated customer once the basic criteria hold up.
Biases to watch for
- Rewarding length: talking more is not selling better. A listening criterion balances it out.
- The halo effect: a strong opening should not inflate everything else. Each criterion is scored with its own quote.
- Transcription errors: an accent or a noisy line can penalise someone who does not deserve it. Low scores are checked against the audio.
- Rewarding the script: reciting the pitch is not listening. What gets scored is whether the answer fits what the customer actually said.
- Judging the person instead of the conversation: you measure what was said, not how the speaker was feeling.
Finally, calibrate. Every so often a trainer scores a sample of sessions by hand and compares the result with the system’s. If they disagree, you adjust the anchors; you do not force the outcome.
Back to the contents
