
Source
MuseTalk official source
Read the primary source.
Reference illustration, not a screenshot of this source.
Lip-sync model · Video · Speakers
Audio-driven lip synchronisation applied to a face image/video route.

Use when you already have the face material and approved speech to synchronise.
Not a full-body motion generator or a replacement for voice synthesis.
Face image/video; Approved audio
Lip-synchronised video
Application/API or supported runtime; access not rechecked
Provider capabilities and historical Studio notes are distinguished below. Current account access, credits and runtime readiness were not tested.
Compute and model/weight terms; hosted wrappers charge separately.
Self-hosted if deployed locally, otherwise provider-dependent.
Exact model version determines support.
Directly matches the lip-sync stage; inspect mouth detail and identity on your material.
Decision: Use when you already have the face material and approved speech to synchronise.
Not a full-body motion generator or a replacement for voice synthesis. Real-time claims depend on hardware and setup; no measured speed claim here.
Official capability documentation; no comparative production benchmark
Catalogue review: 2026-09-05. The evidence note identifies what was verified. This date does not imply a fresh tool test or verification of every linked source.