On this page
- Voice capture is transcript work, not a questionnaire
- The book interview does not transfer, and the arithmetic shows why
- The Personeur Capture Ladder
- The founding call question set, published
- What gets marked up in the transcript, and with what
- The monthly twenty minutes, and how not to waste it
- How to know capture worked before the first draft ships
- What voice capture cannot do
Voice capture is transcript work, not a questionnaire. A writer records you, then marks the transcript for six things: the words you really use, the words you never use, the claims you will defend in public, the stories that carry a date and a stake, how you open, and how you close. A book gets one long interview. A retainer gets a founding call of about forty minutes and twenty minutes a month, so the markup carries the weight.
Voice capture is transcript work, not a questionnaire
Voice capture is the process of recording how a person actually speaks, marking the transcript for the features that repeat, and turning those marks into a document a writer consults before every draft. It is archaeology on a recording. Everything else described as voice capture is a survey with adjectives in it.
The difference matters because the two methods produce different artifacts. A questionnaire produces a page that says the executive is direct, warm and occasionally funny, which describes roughly forty million people and helps a writer with nothing. A marked transcript produces a list of the eleven constructions this person uses that nobody else in their field uses, each with a timestamp you can play back.
Most published descriptions of this method come from book ghostwriting, where the standard is a recorded author interview of two to three hours. Those pages describe the method in the abstract and stop before the artifact. They tell you the writer listens for metaphors and builds a voice profile. They never show the question set, never show what gets marked, and never show a finished profile. This page shows all three, then adapts them to the constraint a LinkedIn retainer actually runs under.
of users say text posts are the format they most engage with. Voice is a property of sentences, so on the format that carries the most engagement, the writing itself is the product rather than the packaging around it.
2026 social media content strategy reportThe book interview does not transfer, and the arithmetic shows why
A book interview is tape-poor and a monthly retainer is tape-rich, which is the opposite of what everyone assumes. The retainer's problem is not a shortage of recorded minutes. It is that the minutes get spent on the wrong layer.
Work it out with two assumptions you can check against your own transcript. Assume conversational speech runs near 140 words a minute, and assume the subject holds about 70 percent of the talk time in a well-run interview. A single 150-minute author interview then yields roughly 14,700 words of the subject speaking. A 60,000-word manuscript comes out of it, so the book writer has about 0.25 words of tape for every word published.
Now the retainer. A 40-minute founding call yields about 3,900 subject words. Eleven monthly top-ups of 20 minutes add about 1,960 each, so year one holds roughly 25,500 words of tape. Ten posts a month at 200 words is 24,000 words published. That is about 1.06 words of tape per published word, which is a little over four times the ratio a book interview runs on.
| Measure | Book interview | LinkedIn retainer, year one |
|---|---|---|
| Recorded minutes | 150 in one block | 40 up front, then 20 a month, 260 total |
| Subject words on tape | about 14,700 | about 25,500 |
| Words published from it | about 60,000 | about 24,000 |
| Tape per published word | 0.25 | 1.06 |
| Shape of the capture | One block, front loaded, never repeated | A drip, compounding, correctable |
| The real failure risk | Too little material for the back half | The twenty minutes becomes a topic planning call |
Capture economics, book interview against a first-year LinkedIn retainer. Assumes 140 words a minute and 70 percent subject talk time.
So the method has to change in one specific way. A book writer front-loads because there will never be more tape. A LinkedIn writer must protect the top-up from being converted into a content calendar meeting, because topics are cheap and register is expensive. That single discipline is the whole adaptation, and it is the thing nobody writes down.
Both numbers above are stated assumptions, not published research. Speech rate varies widely by person and by mood, and interviewer talk time varies by how much the writer likes their own voice. Time your own first transcript, recount, and the ratios move. The direction of the finding survives the recount, because the retainer accumulates and the book interview does not.
The Personeur Capture Ladder
Capture runs in five rungs, and each rung produces a specific object rather than a feeling. If a writer cannot show you the object at the end of a rung, that rung did not happen.
The ladder is deliberately ordered so the cheapest material gets used first. A writer who opens with the interview is asking an executive to pay in minutes for something that was already sitting in a podcast feed.
The founding call question set, published
The questions are not about the person's career. They are designed to move the speaker out of presentation mode, because a person's real register appears when they are bored, annoyed, or telling a story that once had something at stake. Career questions produce the LinkedIn About section they already have.
One rule governs the whole call: never ask what someone's tone of voice is. Nobody can answer it, and the answers that come back are borrowed from a brand deck. Ask for behaviour and extract the tone yourself.
| Question | What it is really for | What you extract |
|---|---|---|
| Walk me through yesterday, hour by hour. | It is boring, so it cannot be performed | Default sentence length, filler words, baseline register |
| What do clients get wrong before they call you? | The complaint engine, which is where opinions live | A defensible claim and a phrase they repeat |
| What did you believe five years ago that you no longer believe? | A position with movement in it | A dated story plus a contrast structure they reach for |
| Who in your field is wrong, and about what? Keep the name out if you prefer. | The boundary between public and private opinion | The opinion register, split by what they will sign |
| Tell me about the worst week you had in this business. | Stakes, which is what separates a story from an anecdote | A story bank entry with a date, a place and a turn |
| What do people thank you for in messages? | The reader's real question, in the reader's words | Proof of demand, and the vocabulary the audience uses |
| Explain what you do to a bright fifteen year old. | Forces analogy under time pressure | The metaphor supply, which is the most stealable feature of a voice |
| What phrase have you caught yourself saying three times this month? | Direct request for idiolect | Signature constructions, verified by the speaker |
| What words do you hate seeing under your name? | The banned list, which is otherwise guessed | Vetoes, and the register they refuse |
| What can you never say in public, and why? | Hazard mapping before it costs something | Forbidden territory: clients, employers, live matters |
| When you disagree with someone senior, how do you open? | Concession and hedging behaviour | Openers, softeners, whether they qualify before or after the claim |
| How do you end a conversation you want to continue? | The closer, which is where fake calls to action enter | A natural ending that does not read as a request |
The founding call, twelve questions in running order, with what each one is actually for.
People perform for the first stretch of any recording. The register usually settles six to eight minutes in, once the speaker forgets the microphone. Treat the opening as warm-up, keep it in the file for reference, and take idiolect marks only from what follows. This is also why question one is deliberately dull.
Twelve questions in forty minutes leaves roughly three minutes each, which is correct. The interviewer's job after asking is to stay quiet and to follow only one thread: whenever the speaker says something in a shape you have not heard from them before, ask for an example rather than moving on.
- Voice capture produces a marked-up transcript and a lookup document, so any writer who describes the process only in adjectives has not done it.
- A retainer is not short of recorded material, it is short of allocation: run the arithmetic and a year of monthly top-ups yields roughly four times more tape per published word than a single book interview does.
- The questions that produce a usable register are boring or annoying ones, because a person only speaks in their own voice once they stop performing, which is usually six to eight minutes into the recording.
- A founding call that yields fewer than twenty-five idiolect marks and fewer than five defensible claims has failed, and the correct response is a second call rather than a first draft.
- The Blind Attribution Check settles the voice question before anything publishes, because two people who know your speech can separate your sentences from a writer's at better than chance if capture went wrong.
What gets marked up in the transcript, and with what
Six mark types, applied in one pass, in a fixed order. The order matters because idiolect marking is mechanical and opinion marking requires judgement, and doing the mechanical pass first calibrates the ear.
Cut the first six to eight minutes out of the marking scope. Keep the text, mark nothing in it.
Any word or construction used at least twice that is not standard usage in the person's field. Verbs and connectives matter more than nouns, because nouns come from the industry and connectives come from the person.
Any sentence a competent peer could disagree with. If nobody could disagree, it is a description, not a claim, and it will produce a status update.
Only passages with all four of a date, a place, a stake and a turn. Three out of four is an anecdote and will not carry a post.
Openers, hedges, concessions and closers. These are structural habits and they survive topic changes, which makes them the most reusable thing on the tape.
Named people, clients under agreement, unpublishable numbers, live legal matters, former employers. Mark them at capture, not at approval, because approval is too late and too expensive.
| Mark | What qualifies | Where it lands |
|---|---|---|
| IDIOLECT | Used twice or more, not standard in the field | The idiolect list, with timestamp |
| NEVER | A word or register the speaker mocks or refuses | The banned list |
| CLAIM | A sentence a competent peer could contest | The opinion register, split public and private |
| STORY | Date, place, stake and turn, all four present | The story bank |
| MOVE | Opener, hedge, concession, closer | Sentence physics and structure notes |
| HAZARD | Names, agreements, numbers, live matters | Forbidden territory |
The markup legend and where each mark lands in the finished profile.
Then count, because counting is what makes this method falsifiable. A founding call of forty minutes yields roughly 3,900 subject words. One distinctive construction every 150 words is a low bar for a person with a formed voice, which sets a floor of about twenty-five idiolect marks. A first month of eight to ten posts needs at least five separate defensible positions, so the floor for CLAIM marks is five.
A founding call that returns fewer than twenty-five idiolect marks and fewer than five claims has failed. The correct response is a second call with different questions, not a first draft. Writing from a thin transcript is how an engagement produces competent, anonymous posts for three months before anyone can name what is wrong.
The monthly twenty minutes, and how not to waste it
Split the top-up five, ten, five, and protect the middle ten minutes above everything else. The default failure is that the whole twenty becomes a topic planning call, which feels productive and adds nothing to the profile.
| Minutes | Spent on | What it adds to the profile |
|---|---|---|
| First 5 | What changed: a new client problem, a new number you can publish, a story that just happened | Story bank entries and fresh proof |
| Middle 10 | One opinion, pushed until you will defend it in writing under your own name | The opinion register, which decays fastest and needs the most feeding |
| Last 5 | Corrections to last month's drafts, spoken aloud rather than typed in the margin | Rules, filed against the panel they belong to |
The twenty-minute monthly top-up, allocated.
The middle ten is the part that generalises. A topic is worth one post. A position is worth a quarter of posts, because it can be argued from six angles, defended against an objection, illustrated with a story and then revisited when the market moves. Ten minutes spent forming one position outperforms twenty minutes spent listing eight topics, and the arithmetic is not close.
The last five minutes only works if corrections are recorded as rules rather than as edits. An edit fixes one sentence. A rule fixes a class of sentences and stops the same correction returning next month, which is the mechanism explained in the piece on drafts that do not sound like you.
How to know capture worked before the first draft ships
Run a blind attribution check, and run it before the retainer starts publishing rather than after month three. It takes twenty minutes and it converts an argument about feel into a number.
Take five sentences lifted verbatim from the transcript and five written by the writer in the captured voice. Match them roughly for length and subject.
Remove proper nouns, dates and any detail only the executive would know. You are testing register, not memory.
Give the ten sentences to two people who know how the executive speaks: a chief of staff, a co-founder, a long-standing assistant, a spouse. Ask them to mark which five the executive said.
Chance is fifty percent. Average the two judges and read the result against the thresholds below.
| Judges correctly separate | Reading | What to do next |
|---|---|---|
| Under 60% | Near chance. The written sentences pass as speech. | Start publishing. Re-run the check at draft twenty. |
| 60% to 75% | Partial capture. Something specific is missing. | Ask judges which cue gave it away, then fill that panel of the profile. |
| Above 75% | Capture failed. The writer is producing a competent stranger. | Second founding call. Do not publish from this profile. |
The Blind Attribution Check, scored. Chance performance is 50 percent.
A draft can pass the blind attribution check and still be worthless, because sounding like you and having something to say are separate problems with separate fixes. A page that reads exactly like the executive and argues nothing is a point-of-view failure wearing the right coat.
What voice capture cannot do
Capture can only record a position that already exists. It cannot manufacture one, and no interview technique changes that. This is the honest limit of the whole method and the reason the second month of a retainer is where most disappointment actually lives.
The failure looks like this. The transcript is full of idiolect and short on claims. The writer, trying to be faithful, produces posts that sound exactly like the executive and say nothing anyone could disagree with. The executive reads them and reports that the drafts do not sound like him, because what he is hearing is his own vagueness in his own cadence, which is more uncomfortable than reading a stranger.
- Capture records how you say things. It does not decide what you think, and a writer who claims otherwise is selling you their opinions in your voice.
- Capture cannot publish what the hazard marks caught. If the best stories are all under agreement, the engagement needs a different content strategy, not a better interview.
- Capture ages. The idiolect list barely moves in a year, the story bank grows, and the opinion register goes stale fastest, which is why the monthly top-up exists at all.
- Capture does not survive an approval chain. Three reviewers can flatten a perfectly captured draft back into corporate register before it publishes, and the transcript is not to blame.
None of this makes the method optional. It makes the sequence matter. Get the positions first, capture the voice second, and price the work knowing which of the two you are actually buying, which is covered in the breakdown of what a LinkedIn ghostwriter costs.
of B2B marketers call LinkedIn the most effective channel for thought leadership. Thought leadership in text is made of sentences, and sentences are what capture is for.
Content Marketing Institute, cited 2026Questions people ask next
How long should the first voice capture call be?
Can a ghostwriter capture a voice from written material only?
What if I refuse to be recorded?
How many drafts before the posts should sound right?
Who owns the transcripts and the voice profile?
Does a monthly top-up really need to be a call?
Send one old post.
Email any post you have published, with your LinkedIn URL. A rewrite in your voice comes back free, so you can judge the craft on your own words.
Get the free rewrite →