How to read this page
These are observable signals, not an audit. Nothing here is an endorsement.
People have wrapped VibeVoice into ComfyUI nodes, desktop apps, API servers and native ports. Here is what exists — and what to check before you run any of it. Signals captured September 2026; tap a card for the detail and the repo link.
23 result(s)
These are observable signals, not an audit. Nothing here is an endorsement.
A ComfyUI custom node is arbitrary Python running with your user's permissions. That vector has been abused for real.
Pinokio, RunPod and Colab scripts for VibeVoice all exist — and every one of them is tiny and unlicensed.
Microsoft never published a 7B TTS checkpoint. Every popular wrapper pulls it from a third-party re-upload.
The 12th most-starred result for the search is an unrelated dictation tool. Check the owner, not the name.
MIT lets anyone fork it — including past the safeguards. Your obligations do not fork away with it.
The fork that kept the pulled TTS code alive — 1.6k stars, 739 forks, still maintained.
A same-day snapshot of the official repo taken when Microsoft pulled it. Frozen since.
The most-starred ComfyUI integration — but last pushed February 2026.
The other well-known ComfyUI node — but untouched for about a year.
Multi-engine ComfyUI suite where VibeVoice is one of several TTS backends. Actively maintained.
Gradio web UI for voice cloning, pairing VibeVoice with Qwen3-TTS and Whisper.
Audiobook pipeline across many TTS models, with a synced reader app. Actively maintained.
A local multi-model TTS studio positioned as an open alternative to ElevenLabs.
A real desktop app (Tauri 2.0) for audiobooks — drag-and-drop, 17+ languages, MP3/M4A export.
Full-stack multi-speaker web system — but it ships with no licence at all.
One-click portable ASR for Windows — fully offline, NVIDIA GPU. Freshly updated.
Drop-in OpenAI-compatible TTS endpoint backed by Realtime-0.5B. Docker included.
A thin FastAPI wrapper over the 1.5B and 7B models. Small but current.
C++ port on ggml — no Python environment at all. The most-downloaded weights in this list.
Rust implementation with voice cloning and multi-speaker support.
Swift port of Realtime-0.5B, for Apple platforms.
Pure-MLX speech stack for Apple Silicon — TTS, cloning, dialogue and ASR. Updated this month.