Deep-learning techno / audio generation experiments.
-
Updated
Sep 18, 2026 - Python
Deep-learning techno / audio generation experiments.
Senspecimenprogramaro de Ondformo Generatori
An agent-friendly Go CLI for generating synthetic sound effects with the ElevenLabs Sound Effects API. Creates audio files locally with machine-readable JSON output, actionable errors, and controls for duration, looping, output format, and playback.
Curated Seed-Audio 1.0 use cases for narration, audio drama, reference voices, provider integrations, and audio-first video workflows.
This repository holds code for the generation and classification of music files using neural networks on time series audio data. This project was done by my partner and I as a final project for our Complex Structures graduate level class at the University of Washington.
This repository consists of projects where GANs were used to work with audio data like music generator etc.
Real-Time-Voice-Cloning using Generative Adversarial Networks. It combines text and speaker and generates natural sounding audio. Some Traditional Methods like Convolutional Neural Networks can contain some limitation it requires large datasets and lengthy processing time, to overcome these limitations we are using GANs to maximize space use.
Open source AI speech generation solution
Server-side PoYo examples for ElevenLabs Music generation workflows.
Offizielle Begleitmaterialien zum Buch 'Content Creation mit generativer KI' von Alexander Loth und Dilyana Bossenz (mitp, 2026). Prompts, Beispiele und Ressourcen zu jedem Kapitel.
Local Windows music and sound-effect generation with YuE and Stable Audio 3 Small SFX.
MCP server generating coherent 16 kHz mixed audio scenes from text - speech, music, sound effects and ambience in one pass - powered by MiDashengLM-Gen (Xiaomi Research). Local-first, runs entirely on your CUDA GPU. Apache-2.0.
OhanashiGPT is an application that generates personalized children's stories based on parameters like age and preferences. It narrates these stories using an AI-generated voice that mimics a parent, trained on their audio samples. The app also creates illustrations to accompany each story, providing a unique and engaging experience for children.
SaaS platform for generating AI podcasts from multimodal content - Built with Hono and Cloudflare Pages
AudioMind v3 - AI Podcast Studio skill for Manus. Orchestrates ElevenLabs TTS, background music, and server-side mixing to produce full podcasts from a single prompt.
AI-powered meditation audio generation pipeline combining neural text-to-speech, procedural ambient sound generation, and automated audio mixing.
Docker image for stable-audio-tools: Generative models for conditional audio generation
An end-to-end Neural Voice AI system featuring zero-shot voice cloning and generative text-to-speech. Powered by XTTS-v2 and Glow-TTS with a PyTorch/Flask inference backend. The application is fully Dockerized for live cloud deployment on Hugging Face Spaces and wrapped in a futuristic, audio-reactive WebGL interface featuring glassmorphism design.
Full-stack AI application that extracts document content, generates podcast scripts, and converts them into realistic audio conversations.
The first-ever all-in-one AI platform for Juniata College. Providing students with relevant information and helping them build their network since April 2025.
To associate your repository with the audio-generation topic, visit your repo's landing page and select "manage topics."