Sentioscope

Speech and voice controls, as each vendor documents them

Hedra: full multi-language support, none of it named

Character 3 is described as text, image and audio to video with full multi-language support and a maximum duration of ten minutes. The duration is a figure a schedule can use. The language claim carries no list at all. As of 2026-09-22.

Four grades of language answer, and where this one landsA list can be checked against a distribution plan, a total can be compared, a claim can only be believed, and a silence says nothing. Full multi-language support is a claim: it cannot be planned against and it cannot be wrong.Most useful to a commissioning decisionA named listCan be matched against the markets a series is committed to.A totalSays the capability is broad; leaves the market question open.A claimWhere this entry sits: breadth asserted, nothing named, nothingfalsifiable.SilenceThe commonest answer in this column, and the only worse one.Least useful
Fig. 1 The voices endpoint is where a per-voice language would naturally live, and the model page does not publish it.
Hedra on languages, statement by statement. Read from the vendor's Character 3 model page and avatar video guide on 2026-09-22.
What the documentation settlesWhat it leaves to a take
Described as text, image and audio to video with full multi-language supportWhich languages that support covers, since none appears
A maximum duration of ten minutes is statedWhether a face holds a viewer for ten unbroken minutes
Speech can be generated inline by naming a voice id from the voices endpointHow many languages that endpoint's catalogue reaches

Inclusion rule. Only statements found on the vendor page named in the caption, and only ones bearing on this one field. Nothing here is filled in from generated output. Order. In the order the statements arrive on the vendor's page.

1A superlative adjective doing the work of a list

Full multi-language support is the strongest-sounding language claim in this register and the least usable one. It cannot be checked, cannot be planned against, and cannot be wrong, because no market is named and therefore none can be missing.

The more interesting reading is structural. This is an audio-driven product: the language of the output is the language of the recording that was handed over. Support for a language is, in that arrangement, mostly a statement about the lip movement rather than about a voice inventory.

2The voices endpoint is where the real answer would live

Speech can be synthesised inline by naming a voice id drawn from an endpoint, and an endpoint that returns voices is the natural place for a language to be attached to each one. The model page does not publish that, so a production that needs the list has to enumerate the endpoint instead of reading a page.

That is a reasonable engineering answer and a poor commissioning one. A decision about which markets a series can reach is usually taken by somebody who will never call an endpoint.

3The others that claim languages without naming any

Two more entries assert breadth without naming a market. One delegates the list to outside speech providers; the other names kinds of performance instead of places.

  • D-ID — a language field, no list.
  • PixVerse — multiple, none of them named.
  • Languages
    Described as text, image and audio to video with full multi-language support, and a maximum duration of ten minutesa claim carrying no listHedra, Character 3 model page / recorded 2026-09-22
  • Voice source
    Speech can be generated inline instead of uploaded, by naming a voice id drawn from the voices endpointa catalogue behind an endpointHedra, avatar video guide / recorded 2026-09-22

4Sources

Read from the Character 3 model page and avatar video guide at hedra.com on 2026-09-22. The same column across every entry is on languages; everything this vendor publishes about speech is on Hedra. What counts as documented is on how read.