Sentioscope

Speech and voice controls, as each vendor documents them

Synthesia and sync-3: a catalogue against a count

Synthesia lists each voice with a language name, a code, a gender and an id. sync-3 says ninety-five or more languages and describes no voice. The smaller answer is the one a production can act on. As of 2026-09-22.

A bigger number and a smaller answerNinety-five or more describes alignment holding across phonetics, because the waveform arrives from elsewhere. A catalogue with language codes describes voices a production can actually obtain.A catalogue with codesA total of ninety-fiveWhat the figure countsVoices that exist and can be pickedSounds alignment copes withWhat it settlesWhether a market has a voiceWhether a supplied voice will landHow long to checkA filter on a language codeCannot be checkedWhy it is that shapeThe platform supplies the voicesThe recording doesGrading a list above a total, with the smaller number winning
Fig. 1 The shape of each answer follows the shape of each product, and a reader who misses that reaches the wrong conclusion twice.
Controls documented, model by modelFilled where the model documents that control, hollow where nothing is published about it.Controls documented, model by modelSynthesiasync-3Where they partAudio sourceAudio source — Synthesia: From the script, or uploadedAudio source — sync-3: Supplied, or read from textAudio source — Where they part: Different answersVoice sourceVoice source — Synthesia: A catalogue voice, or a cloned oneVoice source — sync-3: Nothing documented as an inputVoice source — Where they part: Different answersPer-character bindingPer-character binding — Synthesia: Not documented by the vendorPer-character binding — sync-3: Not documented by the vendorPer-character binding — Where they part: Same answerLanguagesLanguages — Synthesia: A catalogue with codesLanguages — sync-3: A count, ninety-five or moreLanguages — Where they part: Different answersLip-syncLip-sync — Synthesia: The spoken content, framed closeLip-sync — sync-3: The audio it is givenLip-sync — Where they part: Different answers
Fig. 2 Filled where the model documents that control, hollow where nothing is published about it.
Models documenting each controlHow many models document each control. A hollow column is a statement about documentation, not capability.Models documenting each controlSynthesia4 of 5sync-34 of 5Where they part5 of 5
Fig. 3 How many models document each control. A hollow column is a statement about documentation, not capability.
Synthesia and sync-3 on the five fields this register keeps, with the gap on each. Recorded 2026-09-22.
FieldSynthesiasync-3Where they part
Audio sourceFrom the script, or uploadedSupplied, or read from textDifferent answers
Voice sourceA catalogue voice, or a cloned oneNothing documented as an inputDifferent answers
Per-character bindingNot documented by the vendorNot documented by the vendorSame answer
LanguagesA catalogue with codesA count, ninety-five or moreDifferent answers
Lip-syncThe spoken content, framed closeThe audio it is givenDifferent answers

Inclusion rule. Two entries are given a page together when at least one column puts them at opposite grades of answer. Pairs that agree on every column, or that are both blank throughout, do not get a page. Order. Fixed field order, identical on every side-by-side page.

1A larger number and a smaller answer

Ninety-five or more is far bigger than any catalogue here and settles nothing, because it describes alignment holding across phonetics rather than voices being available. A catalogue with language codes settles a market question in a minute.

This is the clearest case in the register for grading a list above a total. The two numbers are not measuring the same quantity, and the one that looks weaker is the one that answers the question a commissioning meeting asks.

2The difference follows from where the voice comes from

A platform that supplies voices has to enumerate them, because a user has to pick one. A platform handed a waveform has nothing to enumerate; the voice arrived with the recording, so language coverage becomes a claim about sounds rather than about casting.

Neither vendor is being evasive. The shape of the answer follows the shape of the product, and a reader comparing the two numbers without noticing that will reach the wrong conclusion twice.

3Both name a driver for the mouth, and by different means

One generates lip movement from the spoken content and attaches a framing condition to it. The other matches movement to the audio it is given. Two of only five entries in this register that name a driver at all, arrived at from opposite directions.

The framing condition is the rarer sentence. Saying performance improves with the avatar closer to camera names a weak case, and naming a weak case constrains a shot list in a way a capability claim never does.

4Each of them on its own

The column this pair was chosen for is languages, and each entry has a page of its own on it. The full row for either, all five columns with the wording behind each cell, is on its model note.

  • Languages
    Each voice is listed with its formal and native language name, a language code, a gender, a name and a voice idpublished as a catalogue with codesSynthesia, list of supported voices / recorded 2026-09-22
  • Voice source
    Voice cloning carries a language list of its own, running from Afrikaans to Zulu and described as the same list used for standalone cloninga cloned voice with its own reachSynthesia, create an avatar / recorded 2026-09-22
  • A framing condition attached to lip-sync
    Lip sync performs best when the avatar is framed closer in the scene rather than positioned far from the cameraSynthesia, create an avatar / recorded 2026-09-22
  • Languages
    sync-3 supports 95 or more languages, described as the same coverage as the models before itpublished as a countSync, sync-3 model documentation / recorded 2026-09-22
  • Audio source
    The accepted pairs are video with audio, video with text, image with audio and image with textsupplied or read from textSync, sync-3 model documentation / recorded 2026-09-22

5Sources

Every cell above is read from the documentation each vendor publishes, on the dates carried by the two model notes. The whole register on one column is on the field note; entries grouped by the shape of their answer are on the routes. Other pairs: SceneMixer and Kling AI, Runway and MiniMax.