River
Y CombinatorBacked by Y Combinator
FREE TEMPLATE

Discovery Call Questions Template

Four sheets and four documents that score each call from what the prospect said, then prune the question set on your own closed deals.

Free download  ·  No account needed

Question Effectiveness, quarterly review

13 questions, 113 closed deals, minimum sample 30

Deals in the analysis113 closed, won and lost both
Questions with a verdict11
Too few to judge2
Cut this quarter2

1. Kept, largest separation first

QuestionAskedWin askedWin notDiff
Why now, rather than six months ago7444%18%+26
What happens when it goes wrong8841%24%+17
What does success look like in 12 months4942%26%+16

2. Cut on evidence, 30 September

QuestionAskedDiffWhy
On a scale of one to ten, how painful is this61-1Answered every time, never varied
Are you the decision maker84-4Yes on 84% of calls, including the stalls

3. Too few to judge is a verdict, not a blank

Two questions show promising separation on 7 and 11 deals. Both read too few to judge with the shortfall stated. A win rate computed on seven deals gets quoted in a pipeline review and never retracted.

The 12-month success question is asked almost only by the two most senior reps, who close more anyway. The confound is recorded beside the number.

Every sales team already has a list of discovery questions, and almost none of them knows which questions on the list matter. Search the phrase and page one hands you thirty questions grouped by framework plus a scorecard with point values printed on it: twenty-five for approved budget, fifteen for planned, disqualify below forty. Nobody calibrated those numbers against your deal size, your sales cycle or your buyers. They are somebody else's arithmetic, adopted whole. Separately, whether your own scorecard reads the same thing twice is measurable.

The number that would actually help is one nobody computes: the win rate on deals where a question was asked against deals where it was not. That needs the transcript and the outcome in the same place, so it needs recordings, which is why the log here carries a consent basis column. Federal law permits a party to the call to record it. California requires every party to consent. A call with no basis recorded is excluded.

So the scorecard has four answer states rather than a tick: quoted, paraphrased, inferred and not asked. Inferred becomes an open item rather than a score, because budget inferred from headcount is a guess wearing a number. Question Effectiveness then reports separation with the sample size beside every figure, and writes too few to judge where the sample is short. The gap register splits what nobody asked into ran out of time, deflected, and never asked, because those three go to different people.

The questions your wins had in common, and the ones nobody is asking

Question Effectiveness with its sample sizes, the Qualification Scorecard scored from quoted evidence, and the Discovery Gap Register ranked by value at risk.

Question Effectiveness

Illustrative rows for a fictional logistics software team of four reps and 113 closed deals. Every number carries its sample.

QuestionAskedNot askedSubstantiveDiffVerdict
Why now rather than six months ago74390.81+26Keep
What happens when it goes wrong88250.66+17Keep
Where would the money come from80330.48+2Rewrite
Who else has been burned by this internally71060.86n/aToo few to judge
On a scale of one to ten, how painful61520.94-1Cut
Are you the decision maker84290.96-4Cut

Rewrite is not cut. The budget question is asked on 80 deals and deflected on half of them. That is a wording failure rather than an irrelevant element, so it gets rewritten and re-measured.

Cut carries a date and a sample. Are you the decision maker was answered yes on 84% of calls, including every deal that later stalled on a stakeholder nobody had met.

Qualification Scorecard

One row per element attempted. The evidence column carries the prospect’s own words or the row is not scored.

DealElementAskedStateEvidenceScored
D-1180ImpactYesQuoted“We shipped 41 wrong loads last quarter and each one is roughly 900 dollars to put right.”Yes
D-1180Critical eventYesQuoted“Peak starts in September and we cannot run peak on the Friday reconciliation again.”Yes
D-1180BudgetYesParaphrasedSaid operations has a systems line and this would come out of it. No amount named.Yes
D-1207BudgetNoInferredNobody said it. Scored from company size in the original notes.No
D-1207Critical eventNoNot askedNo
D-1270BudgetYesQuoted“There is no line for this and I would not be the one to create it.”Yes

The inferred row is the important one. D-1207 carried a budget score derived from headcount. Downgraded to an open item with an owner, and the deal later stalled with no date behind it either.

Quoted, paraphrased, inferred, not asked. Only the first two are scored. D-1270 was disqualified on the sentence sitting in its evidence cell.

Discovery Gap Register

What the team keeps failing to ask, ranked by the value of the open deals it affects.

ElementMiss rateGap typeMost missed byOpen dealsValue at risk
Champion94%Never askedAll except S. Aluko11$504,000
Metrics57%Never askedJ. Marek, T. Bergen7$331,000
Risk49%Never askedAll four reps8$377,000
Decision criteria54%Never askedT. Bergen6$288,000
Critical event35%Ran out of timeAll four reps9$412,000
Budget29%Asked and deflectedAll four reps4$196,000

Three gap types, three different owners. Ran out of time is a call structure problem. Asked and deflected is a wording problem for enablement. Never asked by one named rep is coaching, and it should not be presented to the whole team as a set problem.

Ranked on money rather than on rate. Critical event has the lowest miss rate of the top five and the second largest exposure, because it is missed on the biggest deals: the ones whose calls overran.

What's in the pack

01

Question Set by Persona

Every question carries what it establishes, what a good answer contains, the one follow-up for a thin answer, and its current verdict from the effectiveness sheet.

02

Question Effectiveness

Win rate on deals where each question was asked against deals where it was not, with the sample beside every figure and a confound column for the obvious explanations.

03

Qualification Scorecard

One row per element attempted, with four answer states instead of a tick and the prospect's own words in the evidence column.

04

Discovery Gap Register

What the team keeps failing to ask, split into ran out of time, asked and deflected, and never asked, ranked by the value of the open deals affected.

05

Call to Outcome Log

One row per call joining it to its deal, its persona, its consent basis and its eventual outcome. This is the join the effectiveness analysis runs on.

06

Call Structure

The shape of a first call rather than a script, with the current-state questions before the problem questions and the why-now question placed before the overrun.

07

Disqualification Standard

The five reasons to stop working a deal, each confirmed from what the prospect said, with a threshold derived from the point below which nothing has ever closed.

08

Follow-up Summary Format

Two documents from one call, kept separate: the forwardable version for the prospect and the CRM version with commitments quoted verbatim.

How to use it

  1. 1

    Open in River, or take it blank

    Open the pack in River and hand it your transcripts, or download the Word documents and CSV sheets and fill them in yourself.

  2. 2

    Send calls with outcomes

    Transcripts from Gong, Chorus, Zoom or a pasted export, plus what happened to each deal. Old calls beat recent ones because their outcomes are known.

  3. 3

    Score from what they said

    Each element gets an answer state and a quote. Inferred rows become open items with owners rather than filled cells nobody trusts.

  4. 4

    Prune the set, close the gaps

    Keep, rewrite, cut or too few to judge on every question, then the gap register by value at risk with a named owner for each fix.

Frequently asked questions

Is this template free?

Yes. Four Word documents and four CSV sheets, downloaded as a zip, no signup. River is the optional half: it scores your past calls from the transcripts, computes the separation, and writes the gap register. Neighbouring packs sit under the template library, and one-shot jobs under the tool index.

How many calls do we need before the numbers mean anything?

Enough closed deals, not enough calls. The sheet carries a minimum sample and reports too few to judge below it, so with a dozen transcripts you get a scored question set and no effectiveness verdicts. Thirty closed deals per question is where the numbers start meaning something, and closed-lost counts.

We use MEDDIC. Does this assume a framework?

Whichever you already use. The question set is grouped by persona and each question names the element it establishes, so it maps onto BANT, MEDDIC or SPICED without rewriting anything. What the pack will not hand you is the framework's point values, because those come from your own closed records or they do not come at all.

Do we have to record calls for this to work?

No, and calls without a recording are still scored from notes. They are excluded from the effectiveness analysis, and the log says so in its own column rather than silently dropping them. Where you do record, the consent basis is recorded too: one party under federal law, every party in states that require it.

What stops a question surviving because a senior person likes it?

Evidence and a stated sample. A question with a large sample and no separation is cut, and the sheet records what it was cut on and the date so nobody reintroduces it in January. A question with real separation but only seven deals behind it stays as too few to judge, not as a finding.

Does it help with one specific upcoming call?

The prepare prompt writes a one-page brief: three hypotheses to test, the questions in order, the elements already answered so nobody asks twice, and what to drop if the call runs short. It works from whatever you know, and an account research brief sharpens it. Afterwards, the transcript becomes CRM field changes rather than a note, and it sets the minute budget for the demo.

How does it handle a question only the best reps ask?

It names the problem. A question asked almost only by your two senior reps will show a high win rate, because senior reps close more whatever they ask. The sheet carries a confound column that says exactly that, rather than reporting the lift and letting somebody roll it out. Spread the question first, then measure it.

Find out which of your questions predicted a win

Take the Word documents and CSV sheets blank, or open this exact pack in River and let it score your finished deals first.

Edit with AI