LiveTalking
LiveTalking
A digital-human conversation engine supporting several lip-sync backends, interruption handling and WebRTC, RTMP or virtual-camera output.
A framework for building agents that watch and listen to live video, combining perception models with real-time conversation.
View repository ↗Vision Agents by Stream
Adjacent perception layer for an AI co-host or coach that reacts to what is on screen; not a video generator.
This is source code, not a hosted channel. You need to inspect, configure, and deploy it with your own keys and runtime before anything can broadcast.
Adjacent infrastructure · live perception · agent SDK
Best for: Agents that understand live video
Access note: Read the repository setup and license before running. Hardware, model downloads and API access vary by project.
2026-09-05: Official README reviewed. Video provider and model configuration required; no measured latency or autonomous broadcasting claim.
LiveTalking
A digital-human conversation engine supporting several lip-sync backends, interruption handling and WebRTC, RTMP or virtual-camera output.
Open-LLM-VTuber
A voice-interactive AI companion with visual perception and a Live2D character, supporting local and hosted model backends.
MuseTalk
A lip-sync model that updates a prepared face video from incoming audio clips, with a documented real-time inference workflow.