Sentioscope

Speech and voice controls, as each vendor documents them

SceneMixer and Synthesia: two ways to publish a list

Two entries answer the language question with names. One sets a list of fifteen per project, with Cantonese a sixteenth for dialogue only. The other attaches a language name and a code to each voice in a catalogue. As of 2026-09-22.

A setting on the project, or a code on the voiceFixing a language for a whole project makes the decision early and durable. Publishing a code per voice makes it late and searchable. One makes a new market a new project; the other makes it a filter.Per projectPer voiceWhen it is chosenBefore the first episodeWhen a video is madeWhat it suitsOne long thing in one languageMany short things in many marketsAdding a marketA new projectA filter on the catalogueWhat a named language meansA performance in that languageAn inventory entryThe only two entries here that publish names at all
Fig. 1 A ragged edge, like a dialogue-only sixteenth language, is the kind of detail marketing copy does not invent.
Controls documented, model by modelFilled where the model documents that control, hollow where nothing is published about it.Controls documented, model by modelSceneMixerSynthesiaWhere they partAudio sourceAudio source — SceneMixer: With the picture, in the language set for the projectAudio source — Synthesia: From the script, or uploadedAudio source — Where they part: Different answersVoice sourceVoice source — SceneMixer: A preset library, or the user's own recordingsVoice source — Synthesia: A catalogue voice, or a cloned oneVoice source — Where they part: Different answersPer-character bindingPer-character binding — SceneMixer: Not documented by the vendorPer-character binding — Synthesia: Not documented by the vendorPer-character binding — Where they part: Same answerLanguagesLanguages — SceneMixer: A named list, set per projectLanguages — Synthesia: A catalogue with codesLanguages — Where they part: Different answersLip-syncLip-sync — SceneMixer: Argued out of existenceLip-sync — Synthesia: The spoken content, framed closeLip-sync — Where they part: Different answers
Fig. 2 Filled where the model documents that control, hollow where nothing is published about it.
Models documenting each controlHow many models document each control. A hollow column is a statement about documentation, not capability.Models documenting each controlSceneMixer4 of 5Synthesia4 of 5Where they part5 of 5
Fig. 3 How many models document each control. A hollow column is a statement about documentation, not capability.
SceneMixer and Synthesia on the five fields this register keeps, with the gap on each. Recorded 2026-09-22.
FieldSceneMixerSynthesiaWhere they part
Audio sourceWith the picture, in the language set for the projectFrom the script, or uploadedDifferent answers
Voice sourceA preset library, or the user's own recordingsA catalogue voice, or a cloned oneDifferent answers
Per-character bindingNot documented by the vendorNot documented by the vendorSame answer
LanguagesA named list, set per projectA catalogue with codesDifferent answers
Lip-syncArgued out of existenceThe spoken content, framed closeDifferent answers

Inclusion rule. Two entries are given a page together when at least one column puts them at opposite grades of answer. Pairs that agree on every column, or that are both blank throughout, do not get a page. Order. Fixed field order, identical on every side-by-side page.

1A per-project setting and a per-voice code suit opposite productions

Fixing a language for a whole project makes the decision early and durable, which is what a series wants. Publishing a code per voice makes the decision late and searchable, which is what a shop making many short pieces in many markets wants.

Neither shape is better and the difference is not cosmetic. One arrangement makes adding a market a new project; the other makes it a filter on a catalogue.

2A ragged edge is evidence of an implementation

Listing Cantonese as available for dialogue but not elsewhere is the kind of detail marketing copy does not invent. Inventories that came from a product tend to have edges like that, and they are the parts a reader should trust most.

The catalogue's equivalent is its second range: cloning carries a language list of its own, described in sweeping terms the catalogue avoids. Two published ranges on one platform, and merging them would misstate the reach in whichever direction a reader happened to read.

3The audio behind the two lists is not comparable

One generates the performance in the chosen language while the shot renders and states there is no dubbing pass to follow. The other synthesises speech from a script and animates a face to it, with a framing condition on how well that works.

So a named language means something different on each. In the first it describes a performance; in the second it describes an inventory entry, and only one of those can be auditioned before a commitment.

4Each of them on its own

The column this pair was chosen for is languages, and each entry has a page of its own on it. The full row for either, all five columns with the wording behind each cell, is on its model note.

  • Languages
    Dialogue is set per project to any of 15 languages, with Cantonese a 16th for dialogue onlya named list of 15SceneMixer, languages guide / recorded 2026-09-12
  • Audio source
    The video model is described as performing the line in the chosen language while it renders the shotgenerated with the pictureSceneMixer, languages guide / recorded 2026-09-12
  • Lip-sync
    The vendor states there is no separate dubbing step and no lip-sync patch afterwardsstated as unnecessary rather than as a featureSceneMixer, languages guide / recorded 2026-09-12
  • Languages
    Each voice is listed with its formal and native language name, a language code, a gender, a name and a voice idpublished as a catalogue with codesSynthesia, list of supported voices / recorded 2026-09-22
  • Voice source
    Voice cloning carries a language list of its own, running from Afrikaans to Zulu and described as the same list used for standalone cloninga cloned voice with its own reachSynthesia, create an avatar / recorded 2026-09-22

5Sources

Every cell above is read from the documentation each vendor publishes, on the dates carried by the two model notes. The whole register on one column is on the field note; entries grouped by the shape of their answer are on the routes. Other pairs: PixVerse and D-ID, Sora 2 and Wan 3.0.