Sentioscope

Speech and voice controls, as each vendor documents them

SceneMixer and Kling AI: names against a count of five

Two native-audio entries answering the language question at opposite grades. One sets fifteen named languages per project with Cantonese a sixteenth. The other documents five and leaves every name out. As of 2026-09-22.

Sixteen named against five counted, both nativeWith a large total a production can assume its market is probably covered. With five the probability runs the other way, and five is short enough that naming them would have been one line.SceneMixerKling AIThe language answerFifteen named, plus CantoneseFive, and none of them namedA voice lastingNot documentedBound to the element, and it travelsWhere the voice comes fromA preset library, or your recordingsNot documentedThe mouthArgued unnecessary, no dubbing passNamed as a capability, no driverEach strong in the column the other leaves blank
Fig. 1 One entry has answered the question; the other has confirmed it applies and declined to answer it.
Controls documented, model by modelFilled where the model documents that control, hollow where nothing is published about it.Controls documented, model by modelSceneMixerKling AIWhere they partAudio sourceAudio source — SceneMixer: With the picture, in the language set for the projectAudio source — Kling AI: With the picture, on VIDEO 3.0Audio source — Where they part: Different answersVoice sourceVoice source — SceneMixer: A preset library, or the user's own recordingsVoice source — Kling AI: Nothing documented as an inputVoice source — Where they part: Different answersPer-character bindingPer-character binding — SceneMixer: Not documented by the vendorPer-character binding — Kling AI: On the elementPer-character binding — Where they part: Only Kling AI answersLanguagesLanguages — SceneMixer: A named list, set per projectLanguages — Kling AI: A count, no namesLanguages — Where they part: Different answersLip-syncLip-sync — SceneMixer: Argued out of existenceLip-sync — Kling AI: Named, with nothing driving itLip-sync — Where they part: Different answers
Fig. 2 Filled where the model documents that control, hollow where nothing is published about it.
Models documenting each controlHow many models document each control. A hollow column is a statement about documentation, not capability.Models documenting each controlSceneMixer4 of 5Kling AI4 of 5Where they part5 of 5
Fig. 3 How many models document each control. A hollow column is a statement about documentation, not capability.
SceneMixer and Kling AI on the five fields this register keeps, with the gap on each. Recorded 2026-09-22.
FieldSceneMixerKling AIWhere they part
Audio sourceWith the picture, in the language set for the projectWith the picture, on VIDEO 3.0Different answers
Voice sourceA preset library, or the user's own recordingsNothing documented as an inputDifferent answers
Per-character bindingNot documented by the vendorOn the elementOnly Kling AI answers
LanguagesA named list, set per projectA count, no namesDifferent answers
Lip-syncArgued out of existenceNamed, with nothing driving itDifferent answers

Inclusion rule. Two entries are given a page together when at least one column puts them at opposite grades of answer. Pairs that agree on every column, or that are both blank throughout, do not get a page. Order. Fixed field order, identical on every side-by-side page.

1Five unnamed is the hardest count to plan around

With a large total a production can assume its market is probably covered and check cheaply. With five the probability runs the other way, so an unnamed count of five turns a commissioning decision into a guess. Five is also short enough that naming the five would have been a single line.

The contrast with a named sixteen is therefore sharper than the numbers suggest. One entry has answered the question; the other has confirmed that the question applies and declined to answer it.

2The two entries are strong in opposite columns

The one with the list publishes nothing about a voice belonging to a character. The one with the count has the most developed persistence arrangement here: a voice bound to an element that travels between generations.

So a production choosing between them is choosing which half of the casting problem to solve with documentation and which to solve by testing. Neither page lets it have both.

3Both make a claim about method rather than about output

One states the line is performed in the chosen language while the shot renders, and draws the conclusion that no dubbing pass and no lip-sync patch are needed. The other names lip-sync as a capability and describes no mechanism.

Neither claim can be checked from a page, and they fail differently if wrong. An absent repair stage cannot be reached for; an unexplained capability cannot be reasoned about.

4Each of them on its own

The column this pair was chosen for is languages, and each entry has a page of its own on it. The full row for either, all five columns with the wording behind each cell, is on its model note.

  • Languages
    Dialogue is set per project to any of 15 languages, with Cantonese a 16th for dialogue onlya named list of 15SceneMixer, languages guide / recorded 2026-09-12
  • Lip-sync
    The vendor states there is no separate dubbing step and no lip-sync patch afterwardsstated as unnecessary rather than as a featureSceneMixer, languages guide / recorded 2026-09-12
  • Audio source
    Native audio on VIDEO 3.0 in five languagesgenerated with the pictureKling AI, model guide / recorded 2026-09-12
  • Per-character binding
    Voices are bound to elements, so a character carries its voice between generationstied to the element systemKling AI, model guide / recorded 2026-09-12
  • Lip-sync
    Lip-sync is documentedstated without a mechanismKling AI, model guide / recorded 2026-09-12

5Sources

Every cell above is read from the documentation each vendor publishes, on the dates carried by the two model notes. The whole register on one column is on the field note; entries grouped by the shape of their answer are on the routes. Other pairs: Runway and MiniMax, HeyGen and LTX Studio.