Build text-to-speech applications with this curated guide for real-time agent streaming and high-fidelity offline synthesis.
-
Updated
Oct 7, 2026
Build text-to-speech applications with this curated guide for real-time agent streaming and high-fidelity offline synthesis.
Generate AI-voiced short videos with synced subtitles using Python, ElevenLabs TTS, and FFmpeg for clear, automated social media and educational content.
Contains voice models based on the GPT-SoVITS architecture of different characters including Hitori Gotoh, Ikuyo Kiya and Ichiji Nijika trained from voices from the anime "Bocchi the Rock!".
ReVoice β free, local-first voice cloning & audiobook studio: Qwen3-TTS cloning, chapter narration, M4B export. No cloud, runs on 4GB VRAM.
VoxCPM2: Tokenizer-Free TTS for Multilingual Speech Generation, Creative Voice Design, and True-to-Life Cloning
VoiceFlow - Modern text-to-speech web application with real-time word highlighting, customizable voice settings, and content management. Built with React, TypeScript, and Web Speech API.
Pure C CPU inference engine for Irodori-TTS (Japanese TTS) with int8/VNNI paths β 4.5x faster.
Small TTS voice model for Apple Silicon
Clone any voice from audio file using Qwen3 TTS to generate speech.
AI SST model that takes all the noise from you classes and then gives it to you with time stamps and more!
TΓΌrkΓ§e TTS (Text-to-Speech) model Γ§Δ±ktΔ±larΔ± β dinlenebilir ses ΓΆrnekleri
"A production-quality local Text-to-Speech (TTS) desktop studio. Run completely offline zero-shot voice cloning, and sub-second real-time streaming."
Elevenlabs SRT Suite - Auto Translate .SRT files in different languages. Then creates a TTS with elevenlabs. The software then stores each sentence in a separate audio file that you can review and edit and replace any given sentence/audio-file before giving the final merge command.
Transform any video into a professional multilingual production with natural voice cloning, lip-sync, and on-screen text translation. No cloud APIs, no subscriptions, no data leaving your machine.
Train voice styles for Supertone/supertonic-3 model.
ποΈ Arabic TTS models (FastPitch, Mixer-TTS) in the ONNX format β Python package for offline speech synthesis ππ¦
ποΈ Mixer-TTS for efficient TTS β‘
π± Kitten TTS Studio using local onnx models TTS Offline
To associate your repository with the tts-model topic, visit your repo's landing page and select "manage topics."