← Back to List
⚠
Notice:This resource is provided by a third-party author. Please review the code with AI tools or manually before use to ensure security and compatibility.
Pythonremsky/Kokoro-FastAPI

Kokoro-FastAPI

Dockerized OpenAI-compatible wrapper for Kokoro-82M text-to-speech w/multiplatform CPU, AMD, NVIDIA GPU PyTorch; multi-speaker, clone-tuning, caption timestamps, SSML, optional readalong web UI

80.0/100
★ 5.5KForks: 905
View on GitHub →
Loading report...

Similar Projects

vui

77

Vui Nano — a small, context-aware text-to-speech model trained on real conversations. 219M active params (305M total), Apache 2.0, voice cloning, streaming, runs on CPU (dependency-free C build). Ships with a full real-time voice assistant: WebRTC, ASR, local LLM, OpenAI Realtime API compatible.

Python★ 768

unsloth

93

Local UI to run and train LLMs and diffusion models. Supports GGUF, MLX, Qwen3.8, DeepSeek-V4, MiniMax-H3, Gemma 4, FLUX and more.

Python★ 76.9K

CosyVoice

62

Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.

Python★ 23.8K

pyvideotrans

90

Translate the video from one language to another and embed dubbing & subtitles.

Python★ 19.2K
← Back to List