What 'Good' Sounds Like Past the Gatekeeper: Benchmark Data from 1,200 Graded Cold Calls
We assessed gatekeeper-navigation attempts against a shared rubric. The pass rate wasn't what shocked us — the gap between top and median rep was.
Twelve hundred gatekeeper attempts, one rubric, five assessors who had never sat in on each other's calls. Before we ran a single number I'd have bet on the pass rate being ugly, and it was — 34% of attempts cleared "Proficient" on our Gatekeeper Navigation competency, the rest stalled, got parked in voicemail, or died in the first fifteen seconds. That's not the number that changed how I'd tell an enablement leader to build training. The number that changed it was the gap between the rep at the 90th percentile and the rep sitting three desks away.
The cohort was 142 reps across 19 B2B orgs — SaaS, professional services, industrial distribution — all running outbound through frameworks assessed against The Mastery Standard's Cold Outreach competencies, human-signed-off, not just AI-scored. We isolated one behaviour: does the rep get past a human gatekeeper to either a decision-maker or a named, timed callback commitment? Everything else — pitch quality, discovery depth — was scored separately and doesn't feature here.
The number enablement leaders actually need
Forget the 34% pass rate. It's the spread underneath it that should be driving your coaching budget.
| Percentile | Connect-past-gatekeeper rate | Used reason-first framing | Used peer-register language |
|---|---|---|---|
| Top decile (n=14) | 61% | 71% | 64% |
| Median rep | 19% | 22% | 18% |
| Bottom decile | 6% | 9% | 4% |
Top decile reps are connecting past the gatekeeper roughly three times as often as the rep in the middle of your roster. Not twice. Three times. If your pipeline math assumes a flat conversion rate per rep-hour of dialling, you're building a forecast on top of a number that only describes your average rep — and your average rep is getting outperformed 3-to-1 by people doing the exact same job, on the exact same list, with the exact same product.
That gap is the entire story. A 3x difference in a skill this mechanical isn't explained by "some people are just naturals on the phone." It's explained by two things we could name, count, and coach.
What "good" sounds like, verbatim
Here's a call that failed, transcribed from our assessment set (details altered, pattern preserved):
Rep: "Hi, is this the front desk for Ridgeline Manufacturing? Great, um, my name's Dan, I'm calling from Aperture Systems — is there any chance I could speak with whoever handles operations scheduling?" Gatekeeper: "Can I ask what this is regarding?" Rep: "Sure, it's just a quick call about some software we think could really help with —" Gatekeeper: "I'll pass on a message."
And one that passed:
Rep: "Hi — I need your help for a second. I'm trying to reach whoever owns shift-scheduling conflicts on the plant floor, because we just fixed exactly that for two other manufacturers in the Midwest and I want their read on whether it's relevant here. Who's that usually?" Gatekeeper: "That'd be Marcus, let me see if he's free."
Same opening five seconds of a stranger's day. Wildly different structure. The failing call asks permission before stating a reason. The passing call states a reason before asking for anything, and it asks for help rather than access.
The two techniques doing all the work
Ranked by effect size in the data, not by which one sounds better in a training deck:
- Reason-first framing. State the specific reason for the call in the first sentence, before requesting anyone by name or title. This was the single biggest differentiator between top-decile and median performance. It matters because gatekeepers are pattern-matching for "sales call" within the first two seconds, and a stated, specific reason removes the ambiguity they'd otherwise fill in with "no." Generic reasons don't count — "following up" and "quick question about your business" score as no reason at all.
- Peer-register language. Asking for help ("I need your steer on...") rather than asking for permission ("Is there any chance I could speak to...") reframes the caller as someone already inside the organisation's problem-solving loop, not someone outside it trying to get past a wall. Second-biggest effect, and it compounds with reason-first framing — the two together outperform either alone.
- Pause tolerance after the ask. Reps who stated their reason-and-request and then stopped talking — no filler, no "no worries if not" — connected more often than reps who kept talking through the gatekeeper's processing time. Real but smaller effect, and harder to coach reliably because it's an omission, not an addition.
- Vocal confidence, as tagged subjectively by assessors. Ranked last on purpose. We looked for a correlation between assessor-rated "sounds confident" and connect rate and found none worth reporting. Reps our assessors rated as sounding nervous connected past gatekeepers at rates statistically indistinguishable from reps rated as sounding assured, once reason-first framing and peer-register language were controlled for. Charisma is not the variable. Structure is.
What this means for your training stack
If your gatekeeper-navigation training is a single roleplay module titled "handling gatekeepers," delivered once during onboarding and never assessed again, you are training for a skill you cannot verify anyone learned. The gap we found didn't shrink with tenure — reps at 18 months showed the same spread as reps at 3 months, which tells you this isn't self-correcting through repetition. Nobody accidentally discovers reason-first framing by making enough calls. It has to be named, demonstrated, and graded.
Two moves, in order of what to fix first:
- Build the two techniques into your pre-call planning, not your call review. A rep who hasn't decided their one-sentence reason before dialling won't produce one live under pressure. Bake it into whatever your reps use to prep — a checklist works better here than a script, because the reason has to be specific to that account, not memorised.
- Grade for the presence of the technique, not the outcome of the call. Connect rate has too much noise per individual attempt — some gatekeepers are simply harder than others. Grading whether the rep used reason-first framing and peer-register language is a leading indicator you can coach against weekly, long before the lagging connect-rate number moves.
If you want a starting point rather than building the checklist from scratch, our Cold Call Pre-Call Planning Checklist and Gatekeeper Bypass Script Pack are both built around these two behaviours specifically, not general call confidence.
The floor is not a talent problem
The uncomfortable part of this data isn't that most reps are below average — that's what average means. It's that the ceiling is fully reachable and nobody's building toward it on purpose. A 3x gap between median and top decile, on a skill this mechanical, is a training-design failure with your name on it, not a hiring problem. Fix the two techniques and you're not chasing charisma you can't manufacture. You're teaching a structure anyone can execute on their worst Tuesday.