Real time interactive streaming digital human
-
Updated
Sep 13, 2026 - Python
Real time interactive streaming digital human
AIGCPanel 是一个简单易用的一站式AI数字人系统,支持视频合成、声音合成、声音克隆,简化本地模型管理、一键导入和使用AI模型。
实时交互数字人,可自定义形象与音色,支持音色克隆,对话延迟低至3s。Real-time voice interactive digital human, customizable appearance and voice, supporting voice cloning, with initial package delay as low as 3s.
🎭 AI Avatar / digital human platform — upload a photo, clone a voice, talk to any face in real time with lip-sync video. Open-source, self-hosted. Claude · Whisper · Chatterbox · MuseTalk.
LiveTalk is a unified, high-performance talking head generation system that combines the power of LivePortrait and MuseTalk open-source repositories. The PyTorch models from these projects have been ported to ONNX format and optimized for CoreML to enable efficient on-device inference in Unity.
the comfyui custom node of MuseTalk to make audio driven videos!
Real-time streaming talking-head avatar: PCM audio in, MuseTalk lip-synced video out over a self-developed WebSocket transport. Full-duplex voice sessions with pluggable ASR / LLM / TTS spokes and ms-level barge-in.
Digital-human / talking-avatar workspace orchestrating InfiniteTalk, MuseTalk, and Qwen3-TTS for audio-driven portrait video generation.
Open-source Armenian video dubbing pipeline with ASR, translation, voice cloning, lip-sync, and emotion-aware TTS
MLX port of MuseTalk 1.5 lip-sync digital human for Apple Silicon — torch-free runtime, K12 English teacher, LiveKit (business layer over fusion-mlx)
SOTA Text-to-Video Generator with MuseTalk 1.5, LivePortrait, and LTX-Video. Cinema-grade lip-sync and animation.
CLI 优先的纯本地数字人口播视频流水线(下载→改写→TTS→数字人→后期→发布)
Turn a portrait and a script into a talking avatar. Three lip-sync engines: an instant in-browser preview, MuseTalk for photoreal rendering on your own machine (Apple MPS/CUDA), and HeyGen v3 for cloud renders that also move the head. TTS via Gemini, ElevenLabs (with voice cloning from a video clip), OpenAI or Piper. FastAPI backend, no build step.
AI短剧 · minimaxh3 / minimax h3 / minimax-h3 · RTX 4060 ComfyUI workflow with AI one-click deployment, prompt compiler and lip-sync
WSQ course TGS-2024052081 — build chatbots, voice agents and AI avatar videos with n8n. Ten runnable labs, shipped twice: a local build (Docker + Ollama gemma4, free/offline) and a cloud build (hosted n8n + OpenAI). Covers RAG, ElevenLabs, Vapi, HeyGen, LiveAvatar, Wav2Lip/MuseTalk and Gemini Veo 3.
A fully local multimodal AI pipeline: RAG + TTS + LivePortrait + MuseTalk/多模態AI結合語音動畫系統
面向 16GB 显存设备的 AI 陪伴项目:自定义角色、语音聊天、情绪表情与口型视频。React + Node.js + Python,结合 DeepSeek、火山引擎、LivePortrait 与 MuseTalk,探索从一句话到数字人回应的完整流程。
Dуббер Armenian videos with AI voice cloning, lip-sync, and emotion preservation for Eastern and Western Armenian
To associate your repository with the musetalk topic, visit your repo's landing page and select "manage topics."