Entries whose pages never raise the mouth at all
Seven entries never raise the subject on the page that was read. Several of them plainly put a talking person on screen, so the blank belongs to the document rather than to the product and it has a consistent shape. As of 2026-09-22.
| Model | What the page says | And what it does address |
|---|---|---|
| D-ID | Not documented by the vendor | From a script, or a supplied url |
| LTX Studio | Not documented by the vendor | With the picture, and audio to video |
| Luma Ray | Not documented by the vendor | Not documented by the vendor |
| Runway | Not documented by the vendor | Not documented by the vendor |
| Veo | Not documented by the vendor | With the picture |
| Vidu | Not documented by the vendor | With the picture, speech available on its own |
| Wan 3.0 | Not documented by the vendor | With the picture, unless switched off |
Inclusion rule. Entries whose documentation, as read for this register, says nothing about the relationship between a generated mouth and speech. An entry naming lip-sync without a driver is on a separate route page. Order. Alphabetical by model name.
1A parameter list has no field for a convincing mouth
Most of these pages are API references. They describe what a caller sets, and nothing a caller sets makes a mouth look right, so the subject falls outside what the document is for. That explains the blank without excusing it.
It also predicts where the exceptions come from. The entries that do name a driver are mostly products where the audio is an input, because there the mouth's behaviour is a property of the interface and belongs on a reference.
2Two of these blanks are much harder to excuse
One entry here is an endpoint for creating a talking person, and the word for what it does to a mouth is not on the page. Another is the most thorough voice-construction reference in the register and stops before any performance.
Those two are not documents that ran out of room on an unrelated subject. They are documents about speech that leave out the visible half of it, which is the most striking pattern this column found.
3Silence is not evidence that the case is handled
Several products in this group generate speech with the picture, and something in them is aligning a mouth to it. A production can reasonably infer that the joint generation handles it, and an inference is worth less than a sentence when a shot comes back wrong.
Keeping these cells empty is what lets the rows be compared. Filling them from plausibility would reward whichever vendor a reader already trusted, which is the failure mode this whole register is built to avoid.
4The entries on this route, one page each
Each of these links to that entry on the field this route groups by. The wording behind every cell, and the date it was read, sits on the page it links to.
- D-ID — not documented by the vendor.
- LTX Studio — not documented by the vendor.
- Luma Ray — not documented by the vendor.
- Runway — not documented by the vendor.
- Veo — not documented by the vendor.
- Vidu — not documented by the vendor.
- Wan 3.0 — not documented by the vendor.
- Audio sourceQ3 outputs speech, sound effects and music natively, on by default in the API, with a speech-only modethree tracks in one pass
- Audio sourceAudio is generated with the video rather than added afterwardsgenerated with the picture
- Audio sourceThe audio parameter defaults to true so the returned video carries an audio track, and false returns one withouton unless it is switched off
- Audio sourceJoint audio and video generation, plus audio-to-video, on LTX-2.5two directions
- Audio sourceNot documented by the vendor (as of 2026-09-22)no mention of sound in the video reference
- Audio sourceA script is either text of up to forty thousand characters or an audio url, with audio limited to five minutes for clips and ten for talksa script read aloud or a file
- Lip-syncNot documented by the vendor (as of 2026-09-22)not addressed where voices are defined
5Sources
Membership of this route is decided by the wording in the field it groups by, read from the documentation each vendor publishes. The column itself is on the field note, all five columns are on the speech table, and what counts as documented is on how read. Other routes: Sound in the same pass, Handed a recording.