SceneMixer and Synthesia: two ways to publish a list
Two entries answer the language question with names. One sets a list of fifteen per project, with Cantonese a sixteenth for dialogue only. The other attaches a language name and a code to each voice in a catalogue. As of 2026-09-22.
| Field | SceneMixer | Synthesia | Where they part |
|---|---|---|---|
| Audio source | With the picture, in the language set for the project | From the script, or uploaded | Different answers |
| Voice source | A preset library, or the user's own recordings | A catalogue voice, or a cloned one | Different answers |
| Per-character binding | Not documented by the vendor | Not documented by the vendor | Same answer |
| Languages | A named list, set per project | A catalogue with codes | Different answers |
| Lip-sync | Argued out of existence | The spoken content, framed close | Different answers |
Inclusion rule. Two entries are given a page together when at least one column puts them at opposite grades of answer. Pairs that agree on every column, or that are both blank throughout, do not get a page. Order. Fixed field order, identical on every side-by-side page.
1A per-project setting and a per-voice code suit opposite productions
Fixing a language for a whole project makes the decision early and durable, which is what a series wants. Publishing a code per voice makes the decision late and searchable, which is what a shop making many short pieces in many markets wants.
Neither shape is better and the difference is not cosmetic. One arrangement makes adding a market a new project; the other makes it a filter on a catalogue.
2A ragged edge is evidence of an implementation
Listing Cantonese as available for dialogue but not elsewhere is the kind of detail marketing copy does not invent. Inventories that came from a product tend to have edges like that, and they are the parts a reader should trust most.
The catalogue's equivalent is its second range: cloning carries a language list of its own, described in sweeping terms the catalogue avoids. Two published ranges on one platform, and merging them would misstate the reach in whichever direction a reader happened to read.
3The audio behind the two lists is not comparable
One generates the performance in the chosen language while the shot renders and states there is no dubbing pass to follow. The other synthesises speech from a script and animates a face to it, with a framing condition on how well that works.
So a named language means something different on each. In the first it describes a performance; in the second it describes an inventory entry, and only one of those can be auditioned before a commitment.
4Each of them on its own
The column this pair was chosen for is languages, and each entry has a page of its own on it. The full row for either, all five columns with the wording behind each cell, is on its model note.
- SceneMixer on languages — a named list, set per project.
- Synthesia on languages — a catalogue with codes.
- SceneMixer, all five fields — read from the languages guide.
- Synthesia, all five fields — read from the voices reference and avatar documentation.
- LanguagesDialogue is set per project to any of 15 languages, with Cantonese a 16th for dialogue onlya named list of 15
- Audio sourceThe video model is described as performing the line in the chosen language while it renders the shotgenerated with the picture
- Lip-syncThe vendor states there is no separate dubbing step and no lip-sync patch afterwardsstated as unnecessary rather than as a feature
- LanguagesEach voice is listed with its formal and native language name, a language code, a gender, a name and a voice idpublished as a catalogue with codes
- Voice sourceVoice cloning carries a language list of its own, running from Afrikaans to Zulu and described as the same list used for standalone cloninga cloned voice with its own reach
5Sources
Every cell above is read from the documentation each vendor publishes, on the dates carried by the two model notes. The whole register on one column is on the field note; entries grouped by the shape of their answer are on the routes. Other pairs: PixVerse and D-ID, Sora 2 and Wan 3.0.