← all repositories

zhao-kun/VibeVoiceFusion

A full-stack web application for multi-speaker voice generation and cloning built on Microsoft's VibeVoice model.

VibeVoiceFusion
Not currently ranked — collecting fresh signals.
star history

VibeVoiceFusion is a multi-speaker synthetic speech generation system combining autoregressive and diffusion architectures. It provides a complete web interface for voice synthesis, supports LoRA fine-tuning for custom voice adaptation, and includes batch generation with VRAM optimization features. The system enables voice cloning with distinct speaker characteristics without requiring coding knowledge.

Frequently asked

What is zhao-kun/VibeVoiceFusion?
A full-stack web application for multi-speaker voice generation and cloning built on Microsoft's VibeVoice model.
Is VibeVoiceFusion open source?
Yes — zhao-kun/VibeVoiceFusion is an open-source project tracked on heatdrop.
What language is VibeVoiceFusion written in?
zhao-kun/VibeVoiceFusion is primarily written in Python.
How popular is VibeVoiceFusion?
zhao-kun/VibeVoiceFusion has 484 stars on GitHub.
Where can I find VibeVoiceFusion?
zhao-kun/VibeVoiceFusion is on GitHub at https://github.com/zhao-kun/VibeVoiceFusion.

heatdrop uses Google Analytics to see which pages get read — nothing else. Your call. How we handle data.