How do you coach reps to leave voicemails that get callbacks?
Coach voicemails as a fifteen-to-twenty-second teaser, never a pitch. Drill one structure: name and reason in five seconds, one relevant hook the prospect recognizes, one specific ask, then the number said twice slowly. Score recordings against a rubric weekly, role-play the gaps, and track callback rate per hundred voicemails as your leading indicator.
What a callback-worthy voicemail actually is, and why it moves pipeline
A voicemail that earns a callback is not a compressed pitch. It is a relevance signal with a phone number attached. The prospect's decision happens in the first three seconds, before they know who you are — they are deciding whether the next fifteen seconds are about them or about you. Everything you coach flows from that single fact.
Structurally, the message has four beats and nothing else. Beat one: name and reason inside five seconds. Beat two: one hook the prospect immediately recognizes as their own situation — a funding round, an acquisition, a job change, a line from their earnings call, a competitor's move, a post they wrote last week. Beat three: one specific ask, small enough to say yes to without a calendar. Beat four: the phone number, said twice, slowly, with a beat between digit groups. That is the whole architecture. Reps who add a fifth beat — a value prop, a capabilities paragraph, a "we work with companies like yours" — lose the callback to the delete key.
Why it matters beyond the individual rep: callback rate is one of the cleanest leading indicators a RevOps or frontline sales leader has. Quota attainment tells you what happened last quarter. Callback rate per hundred voicemails tells you whether the message a rep is putting into market this week is landing, and it tells you inside days rather than a full sales cycle. When a segment's callback rate drops, that is usually a messaging or targeting problem showing up before it shows up in meetings booked, and long before it shows up in closed-won. Treat it as an early-warning gauge on message-market fit, not a vanity activity metric.

There is a second-order effect worth naming. Voicemail almost never converts alone, and coaching it as a standalone channel sets reps up to conclude it does not work. The unit of measurement is the touch pair: voicemail plus an email sent within a few minutes, referencing the message by name. The voicemail primes recognition; the email carries the detail and the link. When a rep says "voicemail is dead," they usually mean they left orphaned voicemails with no paired follow-up, so nothing was ever attributable. Fix the pairing before you fix the words.
The adjacent version of this same skill shows up in customer success and renewals. A CSM leaving a message about an upcoming renewal or a usage drop faces an identical problem — the recipient has a hundred reasons not to call back and one reason to. The structure transfers wholesale. So does the coaching method. If you build a scorecard for sales voicemails, you can hand the same rubric to a CS manager with two words changed.
The step-by-step coaching process
Do not start by fixing words. Start by finding out which lane is broken, because the four failure lanes need opposite conversations and coaching the wrong one burns weeks.

Step one: pull evidence. Take three to five of the rep's recent voicemails out of your conversation-intelligence tool — Gong, Chorus, or the call recordings in Salesloft or Outreach — and listen end to end without commentary. Time each one. Write down the exact first seven words. Note whether a specific, verifiable hook appears at all. Five recordings is enough to see a pattern and few enough that you will actually do it.
Step two: classify the gap. Skill means they do not know the structure — the message rambles past thirty seconds, buries the reason, or ends with a rushed number. Knowledge means they have the structure but nothing relevant to put in beat two — every message is "just following up." Will means they have both and are leaving throwaway messages anyway, usually because they have decided voicemail does not work. System means the messages are tight and relevant and still die, which points upstream to list quality, persona fit, dial timing, or a flagged caller ID. A rep with a bad list does not need a pep talk. A bored rep does not need another script.
Step three: run the 1:1 as GROW, not as a lecture. Goal: name the outcome, not the criticism — "by Friday I want one callback traced to a voicemail; walk me through what you're aiming for when you hit record." Reality: play their actual recording and ask three questions — if twelve reps left you a message today, what made this one worth returning? What was the one specific reason you gave? How long was that, and where would you cut? Most reps wince at hearing their own "just a courtesy call to touch base," and that self-recognition outperforms anything you would have said.
Step four: co-build, never hand down. Give the structure, then build two messages together on a real account from their list. Reps recite scripts you wrote for them flatly and abandon them within a week; they defend messages they built. Two finished messages beats ten templates.

Step five: commit to a countable follow-through. "Which of these two are you running this week, how many per day, and send me three recordings by Thursday." A 1:1 with no dated deliverable is a pleasant conversation, not coaching. The commitment is the intervention.
Drills, scripts, and the weekly practice loop
Skill comes from repetitions, not advice. Five drills carry most of the load.
The fifteen-second drill. The rep leaves a voicemail into their own phone with a visible timer. Over fifteen seconds, they redo it. Nothing compresses a message faster than hearing your own rambling played back against a clock.
Cold-read role-play. Hand the rep an account they have never seen with a real trigger you found. Ninety seconds to prep, then they leave a live message to you on speaker. Afterward you play the busy skeptical buyer and tell them exactly where you stopped listening.

Best-and-worst calibration. In the team meeting, play one strong and one weak message with names stripped. The team scores both against the shared rubric before you say anything. Public, blameless calibration moves the whole floor faster than ten private 1:1s, and it stops the rubric from being "the manager's opinion."
Delete-the-fluff. Transcribe one of the rep's messages and have them cross out every word that does not earn the callback. The survivor is usually forty percent shorter and twice as sharp. This one also travels well to email coaching.
The number drill. Reps practice saying the number slowly, twice, timed. Under five seconds for the full delivery is too fast. This sounds trivial and it is the most common single reason a callback never happens: the prospect wanted to call and could not write the number down without replaying the message. Slow delivery also reads as confidence rather than desperation.
On the openings themselves: coach reps to stop leading with "Hi, this is [name] from [company]." Lead with the context. "Saw the post you wrote about routing SLAs." "Noticed the Brightline acquisition closed last month." "Your team just opened three RevOps roles." The hook must be verifiable and specific to that person — "I was looking at your website" does not count. Have reps write five candidate hooks per target account before they dial. The bar is that the prospect thinks *how did they know that*, because curiosity is what converts a listen into a callback.

Three message shapes cover nearly every situation, and reps should rotate them across a cadence rather than repeating one:
*Trigger-event.* "Hi Dana, Marcus from Northwind — calling because Acme just closed the Brightline deal. Two teams we work with hit the same integration crunch right after a deal like that. Worth a short call to see if it's relevant. I'm at 555, 812, 4490. Again, 555, 812, 4490."
*Peer-proof.* "Hi Dana, Marcus from Northwind. I'll be quick — we helped a RevOps team about your size cut lead-routing time from days to under an hour. I have one question that'll tell us in two minutes whether the same play fits. 555, 812, 4490. That's 555, 812, 4490."
*Curiosity, paired to an email.* "Hi Dana, Marcus from Northwind. I sent a note with one number in it about your forecast accuracy that I think you'll want to argue with me about. Take a look, then call me at 555, 812, 4490. Again, 555, 812, 4490."

Repeating an identical message across a cadence signals automation and trains the prospect to skip you. Vary the angle; keep the structure fixed.
Timelines, volume ranges, and what the coaching cycle actually costs
Set expectations honestly, because the most common way this program dies is a manager judging it after nine days.
The thirty-sixty-ninety arc. Days one through thirty are skill. Target fifteen to twenty voicemails a day per rep in an outbound role; the goal is structural consistency, not callbacks yet. Twice a week the rep submits two recordings, you score them, and you run a ten-minute role-play. Days thirty-one through sixty are relevance. Every message must reference a real trigger sourced before dialing, and you A/B trigger-event against peer-proof for that rep specifically, because the winner differs by persona and segment. Days sixty-one through ninety are ownership: the rep self-scores three messages and brings their own diagnosis to the 1:1, and you are coaching judgment rather than mechanics. Graduate them to peer-coaching a teammate — teaching the rubric is what locks it in.
Manager time. Budget roughly twenty to thirty minutes per rep per week: ten to fifteen listening and scoring, ten to fifteen in the 1:1 or role-play, plus one thirty-minute team calibration for the whole group. Across a team of eight that is about four to five hours a week. That is real, and it is the actual constraint — not tooling. If you cannot protect that time, coach four reps properly rather than eight badly.

Sample size. Callback rate is noisy. At fifteen voicemails a day, a rep produces roughly seventy-five a week, so you need somewhere north of a hundred messages before a rate means anything about that rep's message rather than about luck and list. Judge on hundred-plus rolling windows and on trend, never on a single week. Managers who react to a two-day dip teach reps that the metric is arbitrary.
Length targets. Aim for a fifteen-to-twenty-second average, with a hard ceiling around twenty-five. Track the average as its own metric and watch it trend down. Length is the easiest thing to move and it correlates with everything else, because you cannot fit a rambling value prop into eighteen seconds.
Tooling. Most of this runs on what a RevOps team already owns: a conversation-intelligence layer for recordings and search, a sales-engagement platform for cadence and call logging, and a spreadsheet for the scorecard. Do not buy anything new to start. The scorecard being shared and consistent matters far more than where it lives.
Metrics to watch, in order. Callback rate per hundred voicemails is the headline. Connect rate on live answers moves first, because better timing and targeting help both. Average message length trends down. Relevance rate — the share of messages that cited a real, verifiable trigger — is your best proxy for whether the knowledge gap is closing, and you can audit it by sampling ten recordings a week. Meetings sourced from the voicemail-plus-email pair capture the real conversion path. And behavior consistency: messages actually left versus target, because no message quality saves a rep who is not dialing.

Where managers get this wrong
Coaching tone when the problem is the message. "Sound more excited" does not repair "just circling back." Energy applied to an empty message produces an enthusiastic empty message. Fix the words, then the delivery.
Writing the script for the rep. It feels efficient and it produces flat recitation. Hand over the four-beat structure, then co-build on live accounts. Ownership is the difference between a message a rep defends and one they abandon on Wednesday.
No dated follow-through. The single best predictor of whether coaching sticks is whether Thursday's deliverable exists. Without it you are running a book club.
One conversation for every rep. A skill gap and a will gap need opposite approaches — one needs repetitions and structure, the other needs belief, evidence, and accountability. Delivering the will conversation to a skill-gapped rep reads as unfair, and the reverse reads as naive.

Rewarding activity instead of relevance. Praising "fifty voicemails today" is how you get fifty generic voicemails tomorrow. Recognize the rep who left twelve researched messages and got two callbacks, publicly, and the floor recalibrates within a week.
Taking the phone. Leaving the message yourself to demonstrate feels helpful and teaches almost nothing. Demonstrate in role-play, then hand it back immediately.
Ignoring the upstream causes. If tight, relevant messages die consistently, stop coaching words. Check whether the persona is right, whether numbers are being flagged as spam by carriers, whether the dial window matches the buyer's calendar, and whether the list has current direct lines at all. Coaching cannot fix a data problem, and pretending it can costs you the rep's trust.
Letting the rubric drift. If two managers score the same recording three points apart, reps learn the score is noise. Calibrate as a management group monthly on the same two recordings until you converge.

Decision framework: what to coach, and when to stop coaching
Not every voicemail problem is a coaching problem, and knowing when to route elsewhere is most of the skill.
Start with the scorecard. Five dimensions, one to three points each: under twenty-five seconds, one specific and verifiable reason, one clear low-friction ask, number said twice and slowly, confident and unhurried tone. Fifteen possible. Anything under eleven gets re-recorded before it counts. That threshold is arbitrary on purpose — what matters is that it is fixed, shared, and applied identically by every manager on the floor.
Then route by what the score and the outcome say together. High score and low callbacks means the problem is not the message; go upstream to list, persona, and timing. Low score and low callbacks means coach the specific failing dimension, one dimension at a time — reps cannot fix five things at once, and picking the lowest-scoring dimension for a full week produces faster movement than a general note to "tighten it up." High score and healthy callbacks means raise the bar: push relevance depth, then hand them peer-coaching duty.
Two escalation rules keep this honest. First, if a rep's scores are strong and consistent across sixty days and callbacks still lag the team badly, the issue is almost certainly territory, list, or segment — that is a RevOps conversation about data and routing, not another role-play. Second, if the problem is not voicemails specifically but every activity across the board, you are in performance-management territory, and dressing it up as message coaching does the rep a disservice. Say the real thing.
Related questions
How does this change for inbound or renewal calls?
The structure holds; the hook changes. Inbound leads already signaled intent, so beat two references their own action — the form, the demo request, the pricing page — and the ask can be bigger. Renewal messages cite usage or a date. Keep the length ceiling and the slow number identically.
Should reps leave a voicemail on every dial attempt?
No. Two or three per cadence, each with a different angle, paired with an email each time. Leaving one on every attempt trains the prospect to ignore your number and burns the good hooks early. Save the strongest trigger for the second or third touch.
What if the rep is genuinely uncomfortable on the phone?
Separate discomfort from incompetence. Repetition with a timer fixes most of it inside three weeks, because dread comes from improvising. A fixed structure removes the improvisation. If it persists past ninety days with good scores, the role fit is the conversation, not the technique.
Can AI-generated voicemail scripts replace this coaching?
They speed drafting and they cannot replace the practice loop. Generated hooks still need verification against a real source, and the callback comes from delivery — pace, brevity, the slow number — which only shows up in recordings and repetitions.
How do you keep this alive after the ninety days?
Fold it into the standing rhythm: one team calibration a month on two anonymized recordings, a rolling callback-rate view by rep, and re-scoring any rep whose rate drops for three consecutive weeks. Ongoing cost is well under an hour a month per rep.
FAQ
How long should a voicemail be?
Fifteen to twenty seconds, with twenty-five as the hard ceiling. Past thirty seconds most prospects delete before reaching the ask, which means the rep spent the whole message earning nothing. Use the timed drill until brevity is automatic rather than something the rep has to remember in the moment.
What is the single biggest reason voicemails get no callbacks?
The absence of a specific, relevant reason to call. "Touching base" and "following up" give the prospect nothing to act on. Every message needs one concrete hook — a trigger event, a peer result, or a question only that person can answer — and it has to be verifiable, not generic flattery about their website.
Do voicemails work at all, or should reps skip them?
They work as one half of a pair, not alone. The voicemail primes name recognition so the email sent minutes later is not cold, and the email carries the detail and the link. Coach the pair. Reps who conclude voicemail is dead are usually leaving orphaned messages with no paired follow-up and no way to attribute a result.
How do I scale this across a whole team without living in call recordings?
Standardize one scorecard, use your conversation-intelligence tool to surface the best and worst messages weekly rather than listening to everything, run one blameless team calibration a month, and push peer-coaching once reps reach the ownership phase. Sampling five recordings per rep per week is enough to catch drift.
When is voicemail coaching a waste of time?
When the constraint is upstream. Bad lists, wrong persona, stale direct dials, spam-flagged caller ID, or dial windows that miss the buyer entirely will kill excellent messages. If scores are consistently strong and callbacks still lag, stop coaching words and fix the data — that is a RevOps problem with a RevOps solution.
Should every rep use the same script?
Same structure, different content. The four beats are non-negotiable because they encode what actually earns a callback. The hook, the ask, and the phrasing should sound like the rep, sourced from accounts they researched. Identical scripts across a team produce identical-sounding messages, and prospects who hear two of them stop returning either.
Sources
- Gong Labs: Cold Call Tips Backed by Data
- RAIN Group: Cold Calling Statistics
- Harvard Business Review: The Right Way to Hold People Accountable
- Salesloft Blog
- Outreach Resources and Blog
- Sandler Training Blog
- Winning by Design Resources
- MindTools: The GROW Model of Coaching
- FCC: Caller ID Authentication and Robocall Rules
Related on PULSE
- [How do you coach a top performer so they don't leave?](/knowledge/cg0145)
- [How do you coach reps across multiple time zones?](/knowledge/cg0186)
- [How do you coach reps you never see in person?](/knowledge/cg0182)
- [How do you coach reps to handle silence and pauses on calls?](/knowledge/cg0180)
- [How do you coach reps to sell to executives?](/knowledge/cg0178)
- [How do you coach reps to improve their lead-to-opportunity rate?](/knowledge/cg0170)










