Real time interactive streaming digital human
-
Updated
Jul 19, 2026 - Python
Real time interactive streaming digital human
AIGCPanel 是一个简单易用的一站式AI数字人系统,支持视频合成、声音合成、声音克隆,简化本地模型管理、一键导入和使用AI模型。
实时交互数字人,可自定义形象与音色,支持音色克隆,对话延迟低至3s。Real-time voice interactive digital human, customizable appearance and voice, supporting voice cloning, with initial package delay as low as 3s.
🎭 AI Avatar / digital human platform — upload a photo, clone a voice, talk to any face in real time with lip-sync video. Open-source, self-hosted. Claude · Whisper · Chatterbox · MuseTalk.
LiveTalk is a unified, high-performance talking head generation system that combines the power of LivePortrait and MuseTalk open-source repositories. The PyTorch models from these projects have been ported to ONNX format and optimized for CoreML to enable efficient on-device inference in Unity.
the comfyui custom node of MuseTalk to make audio driven videos!
Open-source Armenian video dubbing pipeline with ASR, translation, voice cloning, lip-sync, and emotion-aware TTS
SOTA Text-to-Video Generator with MuseTalk 1.5, LivePortrait, and LTX-Video. Cinema-grade lip-sync and animation.
Digital-human / talking-avatar workspace orchestrating InfiniteTalk, MuseTalk, and Qwen3-TTS for audio-driven portrait video generation.
WSQ course TGS-2024052081 — build chatbots, voice agents and AI avatar videos with n8n. Ten runnable labs, shipped twice: a local build (Docker + Ollama gemma4, free/offline) and a cloud build (hosted n8n + OpenAI). Covers RAG, ElevenLabs, Vapi, HeyGen, LiveAvatar, Wav2Lip/MuseTalk and Gemini Veo 3.
Turn a portrait and a script into a talking avatar. Three lip-sync engines: an instant in-browser preview, MuseTalk for photoreal rendering on your own machine (Apple MPS/CUDA), and HeyGen v3 for cloud renders that also move the head. TTS via Gemini, ElevenLabs (with voice cloning from a video clip), OpenAI or Piper. FastAPI backend, no build step.
Dуббер Armenian videos with AI voice cloning, lip-sync, and emotion preservation for Eastern and Western Armenian
A fully local multimodal AI pipeline: RAG + TTS + LivePortrait + MuseTalk/多模態AI結合語音動畫系統
MuseTalk lip-sync finish engine for Vivijure on RunPod GPU. Pairs with vivijure-cf or vivijure-local.
Real-time speech -> STT -> LLM -> TTS -> photoreal talking-head avatar. Streaming, multi-turn, <8s time-to-first-output. Pipecat + Deepgram + OpenRouter + local CosyVoice2 (vLLM) + MuseTalk.
Add a description, image, and links to the musetalk topic page so that developers can more easily learn about it.
To associate your repository with the musetalk topic, visit your repo's landing page and select "manage topics."