No voice list, and a language claim about text inside the frame
PixVerse publishes no list of the languages its output can speak. Its V6 launch post claims multilingual text generation within frames, with accurate placement across English, Chinese and other languages, which is a claim about pixels rather than audio. As of 2026-09-22.
| Kind of claim | What it is about | Made by this vendor |
|---|---|---|
| Site languages | The vendor's own pages | Yes, 13 declared codes |
| Output languages | Generated scripts and breakdowns | No |
| Voice languages | The dialogue a video speaks | No |
| Text inside the picture | Letters and characters the model draws | Yes, in a launch post |
Inclusion rule. The three register columns plus the one further kind of language claim this vendor's pages make, which belongs in none of them. Order. Register columns in fixed order, then the claim outside them.
1Text in a frame is a rendering problem, not a language one
Drawing legible letters inside a generated picture is hard for a different reason from speaking a language: it is about glyph shapes, spacing and placement rather than about pronunciation or grammar. A model can do one and not the other.
It matters for production all the same. A sign, a phone screen or a note written into a shot cannot be fixed by subtitles or by dubbing later, so a claim about frame text is a claim about how localisable the picture will be.
2Which languages, and how well, are both unstated
The post names English and Chinese and then says other languages. Nothing states which scripts are covered, and the scripts are where this capability breaks: a model trained mainly on Latin and Chinese characters is not evidence about Arabic or Devanagari.
So the claim is recorded as published and not converted into a list. It sits outside the three columns because putting it in any of them would let a reader take it as evidence about dialogue.
3A full site column and nothing about output
Thirteen declared site codes and no statement about the language of anything the product generates is the commonest profile in this register. Nine vendors declare a site list; five of those publish nothing in either output column.
On this row the gap is sharper than usual, because the frame-text claim shows the vendor does describe language behaviour when it has something to announce. What it has not described is the part a dubbed or subtitled release would depend on.
4Sources
Quoted from the PixVerse V6 launch post. The column itself is described on dialogue language, and every cell PixVerse fills is on its tool page.
5The same column, tool by tool
| Tool | Voice languages |
|---|---|
| CapCut | No list published |
| D-ID | 120+, count only |
| Fliki | 91 named, plus 117 dialects |
| Hailuo | No list published |
| HeyGen | Two named lists, no total |
| Higgsfield | No list published |
| invideo AI | No list published |
| Kling AI | Five, count only |
| LTX Studio | No list published |
| Pika | No list published |
| PixVerse | No list published |
| Rask AI | 135+, count only |
| Runway | No list published |
| SceneMixer | 15, plus Cantonese for dialogue |
| Synthesia | 143 named |
| VEED | 29 dub-to, 72 detectable |
| Vidnoz | 100+ and 140+, counts only |
| Vidu | No list published |
Inclusion rule. Tools with a public English site selling AI video generation or AI video dubbing. Site languages count hreflang alternates served from the host being read; x-default and alternates pointing at another host are not counted, for every vendor alike. Order. Alphabetical by tool name.
Neighbouring cells: Kling AI: dialogue language and Rask AI: dialogue language. All of them together: the cell index and the register table.
- Voice languagesNo public list (as of 2026-09-12)no list published
- Text drawn inside the picture“Multilingual text generation within frames is now supported, with accurate placement and style consistency across English, Chinese, and other languages”outside the three columns