2Record
Record and track the voices
Hear the episode before booking an actor.
For studios without floor time, mock-ups, previews
The problem
Approving an adaptation on text alone is approving blind. You hear the fault on recording day.
Casting per character, line-by-line generation, timbre kept from one episode to the next, targeted retakes. On your own accounts, with the provider you choose.
The answer
Your accounts, your costs
Voice services are paid to the provider you choose, on your own accounts. Nothing goes through us, and you keep control of the rights attached to the voices used.
- Is voice generation included in the licence? #
- No, and that is a choice. These services bill by usage: with your own accounts you see your volumes and choose your rates. A studio recording with human actors therefore pays for its licence alone.
- Does generation run on a server? #
- No, on the workstation. In a team, adaptation is shared and generation spreads across machines.
Performance, measured on the original
Before guiding a voice, Axeho listens to the original: intensity, pitch, pace, intonation, line by line. The emotion passed to synthesis is therefore observed on the original actor, not inferred from what the words mean.
- How is that different from analysing the text? #
- “I’m fine” can be gasped, shouted or said flat. The text does not tell you; the signal does. The measurement compares each line to the average **of that same character**, a deep voice is not “dark”, it is normal for that actor.
- And when the measurement is not reliable? #
- It stays silent. For a character with too few lines, no indication is issued rather than one drawn from two samples. Measured lines are shown next to the text: you see what the tool concluded, and you can overrule it.
Visual lip-sync: corporate and advertising
As an option, the mouth can be adjusted on screen to match the dub. Reserved for **corporate, advertising and training**: in fiction, the actor’s performance is the product, and we do not distort it. Our reference sync remains the written one, words chosen to land in the right place.
- Why rule out fiction? #
- Because we sell an adaptation, not a fake. A retouched face passes once and always shows; written lip-sync holds up on a cinema screen. A corporate film or a spot has no such stake: the message comes first, and the sync gain is real.
- Where does the video go? #
- To a third-party provider, the only Axeho function that does so. It is therefore absent until configured, and every upload asks for explicit consent, file by file. You can send a single shot rather than the whole episode.
Specifications
| Provider | yours, you connect your own credentials and keep control of the cost |
|---|---|
| Output | clips aligned to the project’s timecodes, plus separated stems |
| Network | required, the text of the lines leaves, the video never does |
Learning to use it
Step-by-step walkthroughs with screenshots: Cast the voices and generate an episode, Adjust the lips on screen.