Real time interactive streaming digital human
-
Updated
Aug 30, 2026 - Python
Real time interactive streaming digital human
AIGCPanel 是一个简单易用的一站式AI数字人系统,支持视频合成、声音合成、声音克隆,简化本地模型管理、一键导入和使用AI模型。
实时交互数字人,可自定义形象与音色,支持音色克隆,对话延迟低至3s。Real-time voice interactive digital human, customizable appearance and voice, supporting voice cloning, with initial package delay as low as 3s.
🎭 AI Avatar / digital human platform — upload a photo, clone a voice, talk to any face in real time with lip-sync video. Open-source, self-hosted. Claude · Whisper · Chatterbox · MuseTalk.
LiveTalk is a unified, high-performance talking head generation system that combines the power of LivePortrait and MuseTalk open-source repositories. The PyTorch models from these projects have been ported to ONNX format and optimized for CoreML to enable efficient on-device inference in Unity.
the comfyui custom node of MuseTalk to make audio driven videos!
Open-source Armenian video dubbing pipeline with ASR, translation, voice cloning, lip-sync, and emotion-aware TTS
Digital-human / talking-avatar workspace orchestrating InfiniteTalk, MuseTalk, and Qwen3-TTS for audio-driven portrait video generation.
Real-time streaming talking-head avatar: PCM audio in, MuseTalk lip-synced video out over a self-developed WebSocket transport. Full-duplex voice sessions with pluggable ASR / LLM / TTS spokes and ms-level barge-in.
SOTA Text-to-Video Generator with MuseTalk 1.5, LivePortrait, and LTX-Video. Cinema-grade lip-sync and animation.
CLI 优先的纯本地数字人口播视频流水线(下载→改写→TTS→数字人→后期→发布)
AI短剧 · minimaxh3 / minimax h3 / minimax-h3 · RTX 4060 ComfyUI workflow with AI one-click deployment, prompt compiler and lip-sync
WSQ course TGS-2024052081 — build chatbots, voice agents and AI avatar videos with n8n. Ten runnable labs, shipped twice: a local build (Docker + Ollama gemma4, free/offline) and a cloud build (hosted n8n + OpenAI). Covers RAG, ElevenLabs, Vapi, HeyGen, LiveAvatar, Wav2Lip/MuseTalk and Gemini Veo 3.
Turn a portrait and a script into a talking avatar. Three lip-sync engines: an instant in-browser preview, MuseTalk for photoreal rendering on your own machine (Apple MPS/CUDA), and HeyGen v3 for cloud renders that also move the head. TTS via Gemini, ElevenLabs (with voice cloning from a video clip), OpenAI or Piper. FastAPI backend, no build step.
Dуббер Armenian videos with AI voice cloning, lip-sync, and emotion preservation for Eastern and Western Armenian
数字人口播工作台:本地形象、声音、口播与成片管理。Manny 与 Codex 协作完成,原创代码 MIT。
A fully local multimodal AI pipeline: RAG + TTS + LivePortrait + MuseTalk/多模態AI結合語音動畫系統
Open-source AI podcast video generator for Windows. 文案/音频 → 自动分镜、A-roll/B-roll、MuseTalk 口型、字幕与 MP4。
To associate your repository with the musetalk topic, visit your repo's landing page and select "manage topics."