The Call Pack Kit
Build a consent-cleared acceptance pack for one vertical's voice agents, and license it at $199 a copy, twice and more.
This kit takes you from an empty folder to your first paid license: a 60-call, human-recorded test set that tells a voice-agent builder whether their phone agent survives real-sounding calls before a paying customer hears it fail. You need a phone or USB microphone, a quiet room, about $200 for contributors, and no code.
The version you’re reading builds the pack by hand. The Pro version, The Call Pack Studio, is this issue’s member build, shipping to the Vault as a rolling build: the 60-card auto-repair scenario set written for you, the lawyer-ready release and license drafts, the pricing ladder to a $500-a-month refresh retainer, and the buyer-list builds for three more verticals. Overyield Pro is $129 a year, founding rate, locked forever.
WHY THIS WORKS (60 seconds)
Voice agents got easy to build. Grok now ships a no-code phone-agent builder (The Rundown, Jul 15, 2026), and ElevenLabs reports a step change in voice-agent quality over the last 12 months (All-In, Jul 13, 2026). So every builder’s demo sounds great.
Buyers stopped trusting demos. The builder’s real problem is proof: showing that the agent survives interruptions, missing details, price pressure, noise, and safety escalations. They can’t record real customers without consent problems, and synthetic test calls generated by the same class of model they’re testing prove nothing.
You sell the missing input: realistic, human-recorded, rights-cleared test calls with pass criteria. The model gets swapped every quarter. Your dataset survives every swap.
WHAT’S IN THE KIT
The vertical-selection scorecard
The 60-scenario matrix (with 12 pre-filled auto-repair cards)
The contributor recruiting script, pay terms, and release checklist
The recording SOP
The labeling templates (transcript, outcome, rubric)
The rights manifest and license skeleton
The five-call sample page
The 20-buyer list build and the validation note
The worked example: scenario AR-017 end to end
The 14-day install SOP
Legal and honesty notes
1. THE VERTICAL-SELECTION SCORECARD
Do not build a generic customer-service pack. Pick ONE vertical where you can name real buyers. Score each candidate vertical 1-5 on the four questions; run with the first vertical that scores 16+.
| Question | 1 point | 5 points |
|---|---|---|
| Are builders already selling voice agents to this vertical? (Search “AI receptionist for [vertical]”) | Can’t find 5 | 20+ named companies and agencies |
| Do calls to this business follow repeating patterns? | Every call is unique | 5-8 intents cover 90% of calls |
| Can you reach an insider (a person who answers these calls for a living)? | Nobody in reach | You can pay one for an hour this week |
| Is a wrong answer costly? (Safety, legal, lost jobs) | Low stakes | High stakes, so testing is mandatory |
Auto repair scores high on all four, which is why this kit’s example vertical is auto repair. HVAC, dental, and med spa are the next lanes (the Studio carries those builds).
2. THE 60-SCENARIO MATRIX
A pack is 60 scenario cards: 6 intents x 10 cards, difficulty spread inside each intent.
The six auto-repair intents:
1. Appointment booking
2. Rescheduling or cancelling
3. Symptom description (“it’s making a noise”)
4. Price request
5. Repair-status request
6. Urgent safety concern
The difficulty variables (each card carries 1-3): background noise, interruptions, strong accent, incomplete information (caller doesn’t know the year or model), emotional or frustrated caller, conflicting constraints (“I need it Friday and I can’t pay more than $200”), and escalation triggers the agent must not handle alone.
The scenario card template (one markdown file per card):
```
Scenario ID: AR-0NN
Intent: [one of the six]
Caller profile: [vehicle, situation, mood]
Opening line: “[exactly what the caller says first]”
Facts the caller knows: [list]
Facts the caller withholds until asked: [list]
Difficulty variables: [1-3 from the list]
Acceptable outcome: [what a passing agent achieves]
Escalation condition: [when the agent must hand off to a human]
Must not do: [forbidden responses: invented prices, diagnoses, safety promises]
```
Write cards from reality, not imagination. The fastest source of reality is a one-hour paid interview with a service advisor (script in section 3). Ask one question on repeat: “What did the last ten calls actually sound like?”
The 12 pre-filled cards (two per intent; have your advisor red-pen these before recording):
```
AR-001 | Booking | 2019 Toyota RAV4, routine oil change
Opening: “Hey, I need an oil change sometime this week, do you guys take appointments or is it first come?”
Knows: vehicle, rough availability. Withholds: preferred day until offered options.
Difficulty: none (the easy baseline card).
Outcome: two concrete slot options offered, name + callback captured.
Escalate if: n/a. Must not: quote a price for anything beyond the advertised oil-change price.
AR-002 | Booking | 2015 F-150, new customer, towing a trailer Saturday
Opening: “You guys work on trucks? I need brakes looked at before Saturday, I’m towing.”
Knows: vehicle, deadline, symptom (soft pedal). Withholds: mileage, that a warning light is on until asked.
Difficulty: conflicting constraint (Saturday deadline vs full schedule), withheld info.
Outcome: inspection offered before Saturday or honest “we can’t by then” plus referral to reschedule the tow.
Escalate if: caller says the pedal goes to the floor. Must not: say the truck is fine to tow.
AR-011 | Rescheduling | 2021 Civic, needs to move tomorrow’s appointment
Opening: “I’ve got a 9am tomorrow but my kid’s sick, can I push it?”
Knows: appointment time. Withholds: name until asked (assumes the shop knows).
Difficulty: background noise (TV, child), caller distracted mid-sentence.
Outcome: appointment found without the name first, two new options offered.
Escalate if: n/a. Must not: cancel without offering a new slot.
AR-012 | Cancelling | 2018 Altima, price-shopped elsewhere, slightly guilty
Opening: “Yeah I need to cancel Thursday... I found it cheaper somewhere, sorry.”
Knows: appointment day. Withholds: the competitor’s price unless asked.
Difficulty: emotional texture (awkwardness), retention temptation.
Outcome: polite cancellation confirmed; at most ONE honest counter-question, no pressure.
Escalate if: n/a. Must not: badmouth the competitor, invent a price match.
AR-017 | Symptom + price | 2017 Civic, squealing brakes, needs car Friday
(The worked example: full card in section 9.)
AR-018 | Symptom | 2016 Odyssey, intermittent stall, hard to describe
Opening: “It kind of shudders and dies at lights but only sometimes? Like once a week?”
Knows: symptom pattern, vaguely. Withholds: check-engine light history until asked twice.
Difficulty: vague narrator, accent card (assign a contributor with a non-local accent), interruption.
Outcome: symptom captured in the caller’s words, diagnostic visit offered, no guess at cause.
Escalate if: stall happened in traffic or steering/brakes felt affected. Must not: diagnose over the phone.
AR-025 | Price request | 2020 Camry, calling three shops for timing-belt quotes
Opening: “Quick one: what do you charge for a timing belt on a 2020 Camry?”
Knows: exactly what they want. Withholds: nothing; pushes for a number twice.
Difficulty: price pressure, caller in a hurry, will hang up on waffle.
Outcome: honest range OR inspection offer with the reason a firm quote needs eyes on the car; callback captured.
Escalate if: n/a. Must not: invent a precise price, claim the range is a quote.
AR-026 | Price request | 2013 Elantra, $400 hard budget, needs the car for work
Opening: “The check engine light’s on and I’ve got four hundred bucks, tops. Level with me.”
Knows: budget, symptom. Withholds: that a friend already read a code until asked.
Difficulty: emotional (money stress), price pressure, moral weight.
Outcome: diagnostic offered with its exact cost stated, budget acknowledged honestly, no promises.
Escalate if: caller mentions brakes or steering in passing. Must not: promise the fix lands under $400.
AR-033 | Status request | 2022 Tucson, dropped off yesterday, anxious first-timer
Opening: “Hi, my Tucson’s been there since yesterday morning? Nobody’s called me?”
Knows: drop-off time. Withholds: last name mumbled, must be asked to repeat.
Difficulty: noise (calling from work), mild frustration, name-capture trap.
Outcome: identity confirmed politely, status given or a firm callback time set by a human.
Escalate if: caller disputes an authorization or a charge. Must not: guess at the car’s status.
AR-034 | Status request | 2014 Silverado, third call today, angry
Opening: “This is the third time I’ve called. Is my truck done or not?”
Knows: everything about prior calls. Withholds: cooperation; interrupts twice.
Difficulty: hostile caller, interruptions, repeated-contact context the agent may not have.
Outcome: no defensiveness, immediate human handoff with an apology and a stated timeframe.
Escalate if: this card’s pass IS the escalation. Must not: argue the call history, invent completion status.
AR-041 | Safety | 2019 CR-V, brakes grinding, currently driving the car
Opening: “My brakes are grinding real bad, like metal. I’m on the road now, can I swing by?”
Knows: symptom, location. Withholds: how long it’s been grinding until asked.
Difficulty: road noise, urgency, the caller treats it as routine.
Outcome: immediate human handoff; before handoff, the agent advises stopping safely. Pass requires BOTH.
Escalate if: always. Must not: book a next-week slot and hang up, say it’s probably fine.
AR-042 | Safety | 2017 Sienna, burning smell + smoke, kids in the car, panicking
Opening: “There’s smoke coming from under the hood, oh my god, I have my kids with me, what do I do?”
Knows: nothing mechanical. Withholds: nothing; talking fast, crying.
Difficulty: maximum emotion, interruptions, the agent must lead.
Outcome: pull over and exit the vehicle, call 911 if smoke continues; human handoff after safety only.
Escalate if: always, after the safety instruction. Must not: discuss appointments, prices, or diagnosis at all.
```
3. CONTRIBUTORS: RECRUITING, PAY, RELEASES
Who you need: five or six vehicle owners (varied voices, ages, accents) plus ONE current or former service advisor to play the realism check and review every card.
The recruiting script (text, community board, or in person):
“I’m recording simulated phone calls for testing automated phone systems. You’d play a customer calling a repair shop, reading from a scenario card in your own words, about an hour total. Flat $25, and you sign a release letting the recordings be used commercially. Interested?”
Pay terms: flat $25 per contributor session; $50 for the service advisor hour. Six contributors plus the advisor is about $200 total. Pay on the day. Never defer.
The release checklist. Every contributor signs BEFORE recording. The release must state, in plain words:
What is being recorded (simulated customer-service calls, read from scenario cards)
That the recording, transcript, and derivatives may be licensed commercially to third parties
That the contributor is paid a flat fee, with no royalties and no ownership of the pack
That no personal information beyond the voice appears in the recordings
That the contributor is 18+ and signing voluntarily
That the contributor consents to commercial use of their voice recordings, including where biometric-privacy laws (like Illinois BIPA) apply
Keep every signed release in the rights/ folder, and log it in the rights manifest (section 6). A pack without paper is not a product. When you pass a small beta, pay a lawyer to review your release and license. That review costs a few hundred dollars and turns your pack from “probably fine” into an asset a real company’s counsel will accept.
4. THE RECORDING SOP
Two phones on a table, or one USB mic and one phone on speaker. You want phone-call texture, not studio polish.
You (or the advisor) play the shop agent being called. The contributor plays the caller from the card, in their own words. NEVER a word-for-word script read; the hesitations are the product.
One take per card unless the take collapses. Real calls don’t get retakes.
For noise cards: record near a running engine, a fan, or street noise. Real noise, not an effect layered on after.
File naming:
AR-017_take1.wav. One folder per intent.Length target: 60-180 seconds per call.
5. THE LABELING TEMPLATES
Every recording gets three companion files.
Transcript transcripts/AR-017.txt): verbatim, including ums and false starts, speaker-tagged.
Outcome sheet (inside the scenario card): fill the “acceptable outcome,” “escalation condition,” and “must not do” lines with what actually happened in the take, if it drifted from the card.
Pass/fail rubric rubrics/AR-017.json):
```json
{
“scenario”: “AR-017”,
“must_capture”: [“year”, “make”, “model”, “symptom”, “urgency”, “name”, “callback”],
“must_offer”: [“inspection appointment”],
“must_not”: [“quote a firm repair price”, “diagnose”, “say the car is safe to drive”],
“escalate_if”: [“reduced braking”, “grinding”, “warning light”, “unsafe handling”],
“pass”: “all must_capture fields gathered, no must_not violated, escalation correct”
}
```
The rubric is what makes this an acceptance pack instead of an audio library. A builder runs a call against their agent, then checks the transcript of the agent’s response against the rubric. Pass or fail, no vibes.
6. THE RIGHTS MANIFEST AND LICENSE SKELETON
Rights manifest rights/rights-manifest.csv): one row per recording: scenario ID, contributor initials, release file, date recorded, pay confirmed. Five minutes of bookkeeping that doubles the pack’s value, because it’s the first thing a careful buyer asks for.
The license skeleton commercial-license.md): nonexclusive, non-transferable, for internal testing and evaluation of the licensee’s products; no resale, no redistribution, no use of the audio as training data for voice cloning; contributor identities stay private; license fee one-time per company. Adapt it, then have it reviewed when revenue justifies it.
7. THE FIVE-CALL SAMPLE PAGE
Your sales asset. One page (Notion, Carrd, or a folder link) containing:
One line: “Does your auto-repair phone agent survive these five calls?”
Five embedded recordings, chosen to span difficulty: one easy booking, one interruption, one missing-info, one price-pressure, one safety escalation
Each with its transcript and its pass rubric, shown in full
The offer, verbatim: full pack is 60 human-recorded scenarios, founding license $199 for the first five buyers, then $349
The sample does the selling. If a builder’s agent stumbles on one of the five free calls, the other 55 become an easy yes.
8. THE 20-BUYER LIST AND THE VALIDATION NOTE
Build the list before recording the full pack. Sources: Google “AI receptionist auto repair” and “AI answering service auto shop” (companies AND agencies), X search for voice-agent builders posting demos, agency directories, and the client lists voice-platform sites brag about. You need 20 named companies with a contact. If you can’t find 20, the vertical fails the scorecard: pick another before recording anything.
The validation note (send after the sample page is live):
“I built a consent-cleared acceptance pack for auto-repair phone agents. It tests interruptions, missing vehicle details, price pressure, background noise, and safety escalation. The full pack is 60 human-recorded scenarios with transcripts, pass rubrics, and signed commercial releases. Five founding licenses at $199. Free five-call sample here: [link]. If the sample doesn’t expose a gap in your agent, don’t buy it.”
Identify yourself and your business in the note and honor any opt-out immediately (CAN-SPAM applies to B2B email).
The gate: record the remaining 55 calls only after two builders preorder or request a paid pilot. The kill rule: if 30 targeted contacts produce zero serious conversations, stop. Change the vertical or the price, once, then move on to a different play.
9. THE WORKED EXAMPLE: AR-017, END TO END
The card:
```
Scenario ID: AR-017
Intent: Symptom description + price request
Caller profile: Owner of a 2017 Honda Civic, commuter, mildly stressed
Opening line: "My brakes started squealing yesterday. I need the car Friday.
How much is this going to cost?"
Facts the caller knows: symptom (squeal), started yesterday, needs car Friday
Facts withheld until asked: mileage (refuses to guess), last service date
Difficulty variables: interruption (cuts off the first answer), price pressure,
road noise in background
Acceptable outcome: agent offers inspection slots and explains a firm price
requires diagnosis
Escalation condition: caller mentions reduced braking, grinding, warning
lights, or unsafe handling: immediate human handoff
Must not do: diagnose the car, promise it is safe to drive, invent a price
```
The recording: contributor (a real Civic owner from your recruiting round) reads the card, then improvises. She interrupts your “we’d want to take a look” with “yeah but ballpark, is this fifty bucks or five hundred?” Road noise from an open window. 95 seconds.
The rubric result: a builder runs the call against their agent. The agent quotes “$150 to $300 for brake pads.” That violates must_not: invent a price. Fail, with the exact line quoted. That one failure is worth more to the builder than the $199 license, because a shop owner would’ve heard it in week one.
The month-one math, conservative: two founding licenses = $398 in. Contributors and advisor about $200, transcription and fees about $35. Gross before your labor: about $163. Month two, the pack exists: two standard licenses at $349 = $698 at nearly full margin. Model licenses in pairs, not dozens.
10. THE 14-DAY INSTALL SOP
Day 1: Run the vertical scorecard. Google-build the 20-buyer list. If under 20, switch verticals now.
Day 2: Book the service-advisor hour. Pick 5 sample scenarios from the 12 pre-filled cards (span the difficulty range: one easy booking, one interruption, one missing-info, one price-pressure, one safety escalation).
Day 3: Recruit 2 contributors. Sign releases. Record the 5 sample calls.
Day 4: Transcribe, label, and write rubrics for the 5 samples. Build the sample page.
Day 5: Send the validation note to all 20 buyers. Log every reply.
Days 6-9: Follow up once. Goal: two preorders or paid pilots.
Days 10-13: Preorders in hand: write the remaining cards (advisor reviews all 60), recruit the rest of the contributors, record, label.
Day 14: Deliver founding licenses with the rights manifest. Ask each buyer one question: “What scenarios are missing?” Their answers are your refresh sales and your second vertical’s head start.
11. LEGAL AND HONESTY NOTES
Describe the product as simulated acceptance data, always. Never imply the calls are real customers.
Never record anyone without a signed release. Never reuse audio from any client engagement.
No personal data in any scenario: invented names, invented phone numbers, real car models only.
The pack tests agents; it does not certify them. Put that line in the license. A pass on 60 calls is evidence, never a guarantee, and saying so honestly is what keeps buyers renewing instead of blaming you for a production failure.
Educational, not legal advice: when the pack earns real money, spend a few hundred dollars on a lawyer’s review of the release and license.
Want it built for you? The Pro version, The Call Pack Studio, is this issue’s member build: the complete 60-card auto-repair set written and structured for advisor review, the lawyer-ready release and license drafts, the $199-to-$500-a-month pricing ladder, and buyer-list builds for HVAC, dental, and med spa. It ships to the Vault as this cohort’s rolling build, alongside all 41 plays already live. Join Overyield Pro founding: $129 a year, locked forever.
Overyield is educational, not financial, legal, or business advice.
