Skip to content
#

audio-processing-with-python

Here are 41 public repositories matching this topic...

Audio tagging is the process of inferring descriptive labels from audio clips (Multi label classification task). This repository contains exploratory code/scripts for audio preprocessing and model fitting for the task of audio tagging and its applications.

  • Updated Jul 10, 2021
  • Jupyter Notebook

An intelligent speech recognition system that combines OpenAI's Whisper for accurate transcription with dual emotion detection models. Analyzes both audio characteristics (tone, pitch, intensity) and textual content to provide comprehensive emotional context alongside transcriptions.

  • Updated Sep 13, 2025
  • Python

This Python script is for a voice interface chatbot named Jervis. It uses OpenAI's GPT-3.5-turbo-instruct model to respond to user input. The chatbot responds by Elevenlabs Voices. Conversation are saved to MongoDB, and MP3 file local and can be emailed if needed.

  • Updated Mar 9, 2024
  • Python

Add this topic to your repo

To associate your repository with the audio-processing-with-python topic, visit your repo's landing page and select "manage topics."

Learn more