You signed in with another tab or window. Reload to refresh your session.You signed out in another tab or window. Reload to refresh your session.You switched accounts on another tab or window. Reload to refresh your session.Dismiss alert
Speech Arena ranks top TTS (AI voice) models through crowd-sourced blind listening tests. Listen to two samples, pick the better voice, and help build a public leaderboard you can trust. Rankings use Elo with a dynamic K-factor (new voices adjust fast, established ones move slowly).
PoC suggesting that audio transformations* designed to make synthetic voices sound more credible are becoming obsolete in the face of spoofs generated by modern TTS/VC systems (2025). * Transformations proposed in the article “Breaking Security-Critical Voice Authentication”
Elevenlabs SRT Suite - Auto Translate .SRT files in different languages. Then creates a TTS with elevenlabs. The software then stores each sentence in a separate audio file that you can review and edit and replace any given sentence/audio-file before giving the final merge command.
This project includes a Python script for fine-tuning a text-to-speech (TTS) model. The script utilizes custom datasets and use CUDA for accelerated training.
A Streamlit web app for AI-powered voice cloning using Coqui XTTS v2. Record or upload reference voices, clone speech in multiple languages, and generate natural audio outputs.