Skip to content

Home services and trades

AI call scoring for CSRs: measuring booking rate from recordings

By SourceX Editorial · Reviewed by Noah Loul ·

Short answer

AI call scoring for home services transcribes CSR calls, decides whether each was a bookable opportunity, checks whether it booked and grades the call against a rubric. Booking rate is booked opportunity calls divided by all opportunity calls, so classifying opportunities correctly matters most. Confirm recording disclosures and vendor terms before any recording feeds an AI tool.

Key takeaways

  • A misclassified call moves the booking rate as much as a lost booking, so opportunity rules come before rubric design.
  • The booked flag should come from the field service platform, not from what the transcript seems to say.
  • A scored call record gains lasting value from two fields most tools skip: the manager's correction and the job outcome.
  • CSRs accept scores they can check, so every grade should point to the passage of the call behind it.
  • Recording notices, employee notices and vendor data terms must be checked before recordings feed any AI tool.

How does AI call scoring work?#

AI call scoring works as a chain: recordings are transcribed, split by speaker, classified by reason and outcome, and graded against a rubric your team defines. Most tools connect to the phone system for audio and to the field service platform, such as ServiceTitan or Housecall Pro, for the booking.

Every link in the chain can be wrong, and errors compound. A transcript that mishears a street name is harmless; a classifier that calls a billing question an opportunity quietly drags down a CSR's booking rate.

  • Capture: the phone system records the call and passes audio plus metadata such as line, campaign and duration.
  • Transcribe: speech becomes text, split into CSR and caller.
  • Classify: the call is labeled by reason, such as new service, maintenance, existing job, billing or solicitation.
  • Decide opportunity: the tool judges whether the caller could have been booked.
  • Match outcome: the call is matched to an appointment or job created in the field service platform.
  • Score: rubric items are graded, such as greeting, problem discovery, offering a time and asking for the booking.

Measuring booking rate correctly#

Booking rate is booked opportunity calls divided by all opportunity calls, and the definition of an opportunity decides whether the number means anything. Calls where a customer could have been scheduled count, even if the CSR never offered a time; calls about an existing appointment, a bill or a vendor pitch do not.

Verify the booked flag against the field service platform rather than trusting the transcript. A caller who says yes may hang up before a time is set, and a job created later by a CSR callback may belong to the original call. Write the rules down once and apply them to every CSR and every branch.

Measuring booking rate correctly
Call typeCounts as an opportunity?Note
New service requestYesThe core of the metric
Maintenance or agreement visit requestYesTrack separately if agreements are a priority
Price shopping with a real needUsually yesDecide once and apply consistently
Outside the service area or a service you do not offerNoRecord the reason for marketing review
Existing job status or rescheduleNoScore for service quality instead
Billing questionNoRoute and score separately
Vendor, solicitation or wrong numberNoExclude from all CSR metrics

What does a scored call record contain?#

A scored call record contains the call's metadata, the transcript, the classification, the outcome and the grades, each traceable to the evidence behind it. The table lists the fields a complete record carries.

The last two rows are the ones most tools skip and the ones that give the record lasting value. A manager's correction shows where the AI was wrong, and the job outcome shows whether a well-scored call led to good work or to a canceled appointment.

What does a scored call record contain?
FieldExample content
Call metadataDate, time, duration, line or campaign, CSR pseudonymous ID
TranscriptSpeaker-separated text, redacted as your policy requires
Call reasonCoded reason from a fixed list
Opportunity flagYes or no, with the rule that applied
OutcomeBooked, not booked, transferred or follow-up, linked to the job if booked
Rubric scoresA grade per item, with the transcript passage that supports it
Reviewer checkWhether a manager reviewed the call and changed any score
Job outcomeLater status of the booked job: completed, canceled, sold or callback

Using scores for coaching without losing trust#

Call scores help coaching when CSRs can see the call, the rubric item and the passage behind each grade. A score with no evidence reads as surveillance and gets argued about rather than used.

Start with a short rubric and calibrate it by having managers score the same calls as the tool, then review the disagreements together. Each disagreement either fixes a rubric item or teaches the team something about a hard call.

Hold off on tying pay to raw AI scores until you know how often the tool misclassifies calls on your lines. Pay plans built on an unverified metric create disputes that outlast the tool.

Recordings should feed an AI tool only after you have checked that callers were told about recording, that your disclosure covers the intended use, and that the vendor's terms limit what it does with your audio. Call recording laws differ by state, and some require every party's consent, so counsel should confirm which apply to each of your lines.

Automated redaction helps but is not complete on its own. The open-source Presidio project, for example, warns that because it uses automated detection there is no guarantee it will find all sensitive information, and that additional protections should be employed. Plan for sampled human review of transcripts.

  • Play a recording notice on every inbound line, including after-hours and overflow lines.
  • Check outbound calls, where the notice is often missing.
  • Tell employees in writing that calls are recorded and scored, and how scores are used.
  • Read the scoring vendor's data use clause, and opt out of training on your recordings if you do not want it.
  • Set retention periods for audio and transcripts, and delete on schedule.
  • Keep payment card details out of recordings and transcripts, or pause recording while payment is taken.

Illustrative: a multi-trade contractor checks its booking rate#

Illustrative: a fictional HVAC, plumbing and electrical company runs a central call center for several branches. Its call scoring tool reports a booking rate the COO believes is too low, and CSRs have stopped trusting their scores.

A sample review shows the tool counting reschedules and billing calls as opportunities, and missing bookings made when a CSR called the customer back. The team tightens the opportunity rules, matches outcomes to jobs in the field service platform and adds manager review for disputed scores.

The decision: coach only on the corrected rate, and add the recording notice the after-hours line had been missing. The outcome is a booking rate CSRs accept as fair, and a scored call archive with linked job outcomes that the company owns and controls.

How SourceX looks at scored call archives#

Scored call archives with linked job outcomes are one of the record types AI developers ask about, because they connect real conversations to business results. SourceX assesses them through the SourceX five-step transaction: Supply, Rights, Preparation, Approval and Delivery.

The Rights step checks recording disclosures, employee notices and vendor terms. Preparation removes caller and employee identities, payment details and addresses under a plan the supplier approves, and the SourceX Evidence Packet records it. Nothing is shared at the fit check stage.

Frequently asked questions

Can we score calls without storing recordings?

Some tools analyze calls live and keep only transcripts or scores, which reduces what is stored. The audio is still processed, so the same notice and consent questions apply. Confirm in the contract what the vendor keeps and for how long.

Does the scoring vendor get rights to our call data?

Rights and use are set by your contract. Some vendor terms allow customer data to be used to improve the vendor's models unless you opt out. Read the data use and retention clauses, and have counsel review anything unclear before you connect your phone system.

What if the AI and the manager disagree on a score?

Record both. The manager's score should drive coaching, and the disagreement shows where the rubric or the classifier needs work. Over time those corrections make the archive more useful, because each one is a labeled example of a judgment call.

Is booking rate the only metric worth tracking?

No. Pair it with outcomes such as canceled appointments, jobs sold and average ticket by CSR, which show whether bookings were good ones. A high booking rate built on poorly qualified appointments wastes technician time and hurts close rates in the field.

Should CSRs see their own scores?

Yes, with the evidence. CSRs who can replay the call and see the passage behind each grade can flag errors and learn from strong calls by peers, and their flags improve the rubric for everyone.

Sources

  • Presidio's own documentation warns that "because it is using automated detection mechanisms, there is no guarantee that Presidio will find all sensitive information. Consequently, additional systems and protections should be employed." Source

Related resources

See if your company qualifies

A short company assessment. No data uploads are needed.

See if you qualify