Two field training officers watch the same recruit handle the same angry driver. One says, "Great job, very calm." The other says, "Too passive. He lost control of that stop." Both are experienced. Both are sincere. And the recruit walks away with no idea what to do differently next time.
That's the problem with evaluating communication by feel. If you train police officers, firefighters, paramedics, or dispatchers, learning how to make a rubric is one of the most useful things you can do for your program. A good rubric turns "he's good with people" into specific, observable behaviors that every evaluator scores the same way, and that every learner can work to improve. This guide walks through the process step by step and includes a sample communication rubric you can adapt.
Why Rubrics Beat "Gut Feel" Evaluation
Carnegie Mellon University's Eberly Center defines a rubric as "a scoring tool that explicitly describes the instructor's performance expectations for an assignment or piece of work" (Eberly Center, CMU). It has three parts:
- Criteria: the aspects of performance being assessed.
- Performance levels: a rating scale showing the level of mastery.
- Descriptors: what performance looks like at each level.
According to the Eberly Center, rubrics help ensure "consistency across time and across graders," reduce the uncertainty in grading, and help learners understand expectations and use feedback to improve.
Public safety has relied on this idea for decades. The San Jose Police Department's field training program, which the department says became a distinct unit by 1972, introduced a seven-point scale for rating recruits, with 1 meaning "unacceptable" and 7 meaning "superior." To make ratings more accurate, two sergeants wrote guidelines describing each rating category. The department notes that California adopted the "San Jose Model" as its state standard in 1974 (San Jose Police Department).
Communication skills deserve the same rigor. A clear evaluation rubric gives you:
- Fairness. Learners are judged against the same standard, whoever is evaluating.
- Better feedback. “You didn’t explain why you were detaining him” beats “work on your communication.”
- Useful data. Scores across a class show where your curriculum needs work, which strengthens your overall training evaluation.
How to Make a Rubric, Step 1: Choose Observable Behaviors
The most common rubric mistake is listing traits instead of behaviors. "Empathetic," "professional," and "good rapport" are traits. Two evaluators can disagree about them all day. A behavior is something you can see or hear.
| Trait (hard to score) | Observable behavior (easy to score) |
|---|---|
| Empathetic | Names the person’s emotion before giving instructions |
| Professional | Introduces self by name and role within the first 30 seconds |
| Good listener | Asks at least one open-ended question and summarizes the answer |
| Clear communicator | Explains next steps without jargon or codes |
| Calm | Keeps tone and volume steady when the other person escalates |
To choose behaviors:
- Start with the learning objective. What should the learner be able to do after this scenario?
- Look at real incidents. Complaints, commendations, and after-action reviews show which behaviors matter in your agency.
- Keep it short. Four to six criteria per scenario is usually enough. A 20-line rubric is hard to use in real time.
- Tailor to the role. A dispatcher’s rubric might include “gets the address in the first exchange.” A paramedic’s might include “explains the risks of refusing transport in plain language.”
Step 2: Choose a Rating Scale
Your scale should be simple enough for evaluators to apply quickly and detailed enough to show growth.
- Checklist (yes/no): Good for must-do behaviors, like identifying yourself. Too blunt for skills that develop over time.
- Three levels: Not yet, meets standard, exceeds standard. Fast and easy to calibrate.
- Four or five levels: Shows more gradual progress. A four-level scale also removes the easy “middle” score.
- Seven levels: Like the San Jose FTO scale. Fine-grained, but it demands well-written descriptors and trained evaluators.
Whatever you choose, write a descriptor for every level. A number without a description invites the same gut-feel scoring you're trying to replace.
Sample Communication Skills Rubric
Here's a four-level sample rubric for a general public-contact scenario. Copy it, adapt the language to your agency, and cut or add rows to match your objective.
| Criterion | 1: Not Yet | 2: Developing | 3: Proficient | 4: Exemplary |
|---|---|---|---|---|
| Introduction and purpose | Does not identify self or reason for contact | Gives name or reason, but not both | Gives name, role, and reason early in the contact | Gives name, role, and reason, and checks that the person understood |
| Active listening | Interrupts; no questions | Asks mostly closed questions | Asks open-ended questions and lets the person finish | Asks open-ended questions and summarizes before acting |
| Acknowledging emotion | Ignores or dismisses emotion (“calm down”) | Acknowledges emotion generically (“I understand”) | Names the specific emotion (“You’re frustrated about the wait”) | Names the emotion and connects it to next steps |
| Explaining decisions | Gives orders without reasons | Gives reasons only when pushed | Explains the reason for key decisions in plain language | Explains reasons, options, and what happens next |
| Composure | Raises voice or becomes sarcastic | Steady at first, loses composure under pressure | Keeps steady tone throughout | Keeps steady tone and lowers the other person’s intensity |
| Closing | Leaves abruptly | Ends without next steps | States next steps clearly | States next steps and offers a resource or follow-up |
Role-specific rows to add
- Police: “States the legal reason for a stop or detention in plain language.”
- Fire: “Explains a code requirement in terms of the occupant’s safety, not just the rule.”
- EMS: “Explains the risks of refusing care and confirms the patient understood.”
- Dispatch: “Gives clear, one-step instructions and confirms each step is done.”
Step 3: Calibrate Your Evaluators
A rubric only works if evaluators apply it the same way. That takes deliberate practice, often called norming or calibration. Santa Clara University describes norming as "the process in which a group of raters decide collectively how to use a rubric to evaluate student work in a consistent manner" (Santa Clara University).
A calibration protocol from the Rhode Island Department of Education and the National Center for the Improvement of Educational Assessment follows a simple pattern: evaluators score the same work independently, share scores without explanation, then discuss differences using specific rubric language and evidence until they reach consensus (RIDE).
For public safety trainers, that might look like this:
- Record two or three short scenario runs.
- Have every evaluator score them independently.
- Reveal scores and discuss any criterion where ratings differ.
- Point to specific words or actions in the recording to justify each score.
- Revise descriptors that caused confusion.
Santa Clara adds two useful rules: each rating "should be defensible with evidence from the work product," not "the feeling" of a score, and if two raters regularly differ by more than one point, consider renorming or revising the rubric.
Step 4: Use Rubrics for Coaching, Not Just Grading
Once you know how to make a rubric, the real payoff shows up in the debrief. Instead of a single score, the learner gets a map of what to keep doing and what to change.
- Start with self-assessment. Ask the learner to score themselves first. Gaps between their view and yours are where coaching starts.
- Focus on one or two criteria. Pick the lowest-scoring areas and set a specific goal for the next attempt.
- Quote the learner. “When you said ‘that’s just policy,’ that was a 2 on explaining decisions. What could you say instead?”
- Run it again. Coaching lands best when the learner can immediately try the scenario again and see the score move.
- Track progress over time. Rubric scores across several scenarios show growth that a single pass/fail never could.
Why Rubric-Based Practice Should Be Spoken
Communication rubrics measure what people say and how they say it, so the practice should be spoken, too. Written quizzes can check whether a learner knows they should acknowledge emotion. Only out-loud practice shows whether they actually do it when someone is shouting at them.
The best pairing is a clear rubric plus frequent spoken rehearsal. Each run gives the learner a score, a specific behavior to work on, and a chance to try again while the feedback is fresh.
Key Takeaways
- Knowing how to make a rubric starts with replacing traits (“empathetic”) with observable behaviors (“names the person’s emotion”).
- Keep rubrics to four to six criteria, and write a descriptor for every level of your rating scale.
- Adapt a general communication rubric with role-specific rows for police, fire, EMS, and dispatch.
- Calibrate evaluators by scoring the same recordings independently and discussing differences.
- Use rubric scores to coach specific behaviors, then let learners practice again.
Turn your rubric into practice. Foretell AI from Glimpse Learning scores voice role-play conversations against rubrics your trainers write, so learners can rehearse with a lifelike AI avatar and get consistent feedback on the exact behaviors in your sample rubric. Every learner gets more repetitions, and every session is scored the same way.
Sources
- Eberly Center, Carnegie Mellon University – Grading and Performance Rubrics
- San Jose Police Department – Field Training Officer (FTO) Program
- Santa Clara University, Office of the Provost – Using Rubrics
- Rhode Island Department of Education & National Center for the Improvement of Educational Assessment – Calibration Protocol for Scoring Student Work