Your own face and your own voice, for a page that needs you. One photo and five minutes of recording.
Why filmed and not generated. Four generated masters were bought for one real man, from the same approved portrait, and he rejected all four: too theatrical, then too frozen, then "not natural at all", then "the smile is not nice". A person recognises their own stillness before they recognise their own face, and no written instruction describes it. Rewording is not the lever. Film is.
What only you can do, and it is five minutes. Sixty seconds of video — plain wall, daylight from the front, camera at eye height, just talking and listening the way you normally do. Then three minutes of audio, speaking the way you speak when you are convincing someone. The energy matters more than the microphone: a voice cloned from calm samples sounds like a bored narrator forever, and no setting fixes it afterwards.
Consent, if the face is a colleague's and not yours. A real person's face and voice need that person's agreement in writing before anything is generated.
Open Claude Code and say this. It will ask which of the three you want, then take you through it.
Use make-your-speaker. Guide me step by step, I have never done this.
If you do not have our assistants yet, install them first:
npx skills add gfrankgva/studio-skills
The sizes, what a speaker can do, what it costs and how it is finished are the same for all three, and live on
the Speakers page.