Open source / Audio-driven real-time lip sync

MuseTalk

A lip-sync model that updates a prepared face video from incoming audio clips, with a documented real-time inference workflow.

Lip-sync model · prepared avatar · GPUSelf-host required
View repository
Source code</>

MuseTalk

The unlock

Why this one is on the map.

Supplies the speaking-face component of an AI host rather than the conversation or broadcast system.

This is source code, not a hosted channel. You need to inspect, configure, and deploy it with your own keys and runtime before anything can broadcast.

How to read it

What happens when you open it.

Lip-sync model · prepared avatar · GPU

Best for: Audio-driven real-time lip sync

Access note: Read the repository setup and license before running. Hardware, model downloads and API access vary by project.

Sources & status

Keep the link, keep the caveat.

2026-09-05: Official README and avatar-preparation notes reviewed. Real-time speed is hardware-dependent; demo performance on small GPUs can be much slower.

Explore similar projects

Suggest a correction ↗
← Back to the directory