You signed in with another tab or window. Reload to refresh your session.You signed out in another tab or window. Reload to refresh your session.You switched accounts on another tab or window. Reload to refresh your session.Dismiss alert
Speech Arena ranks top TTS (AI voice) models through crowd-sourced blind listening tests. Listen to two samples, pick the better voice, and help build a public leaderboard you can trust. Rankings use Elo with a dynamic K-factor (new voices adjust fast, established ones move slowly).
PoC suggesting that audio transformations* designed to make synthetic voices sound more credible are becoming obsolete in the face of spoofs generated by modern TTS/VC systems (2025). * Transformations proposed in the article “Breaking Security-Critical Voice Authentication”
Elevenlabs SRT Suite - Auto Translate .SRT files in different languages. Then creates a TTS with elevenlabs. The software then stores each sentence in a separate audio file that you can review and edit and replace any given sentence/audio-file before giving the final merge command.
This was created using NextJS and Typescript. This app takes 4 of the OpenAi models: GPT-4 (chat), Dalle-3 (image generator), Vision (image analysis), and TTS-1 (text-to-speech) and allows the user to transform the way they approach everyday tasks.
This project includes a Python script for fine-tuning a text-to-speech (TTS) model. The script utilizes custom datasets and use CUDA for accelerated training.
Contains voice models based on the GPT-SoVITS architecture of different characters including Hitori Gotoh, Ikuyo Kiya and Ichiji Nijika trained from voices from the anime "Bocchi the Rock!".