Skip to content
#

audio-generation

Here are 399 public repositories matching this topic...

Real-Time-Voice-Cloning using Generative Adversarial Networks. It combines text and speaker and generates natural sounding audio. Some Traditional Methods like Convolutional Neural Networks can contain some limitation it requires large datasets and lengthy processing time, to overcome these limitations we are using GANs to maximize space use.

  • Updated Jan 28, 2025
  • Python

Offizielle Begleitmaterialien zum Buch 'Content Creation mit generativer KI' von Alexander Loth und Dilyana Bossenz (mitp, 2026). Prompts, Beispiele und Ressourcen zu jedem Kapitel.

  • Updated Sep 6, 2026

OhanashiGPT is an application that generates personalized children's stories based on parameters like age and preferences. It narrates these stories using an AI-generated voice that mimics a parent, trained on their audio samples. The app also creates illustrations to accompany each story, providing a unique and engaging experience for children.

  • Updated Sep 13, 2024
  • Jupyter Notebook

An end-to-end Neural Voice AI system featuring zero-shot voice cloning and generative text-to-speech. Powered by XTTS-v2 and Glow-TTS with a PyTorch/Flask inference backend. The application is fully Dockerized for live cloud deployment on Hugging Face Spaces and wrapped in a futuristic, audio-reactive WebGL interface featuring glassmorphism design.

  • Updated Jul 30, 2026
  • Python

Add this topic to your repo

To associate your repository with the audio-generation topic, visit your repo's landing page and select "manage topics."

Learn more