A mobile application that shows you what you say and objects around.
-
Updated
Sep 8, 2020 - Java
A mobile application that shows you what you say and objects around.
Speech Detection 💬
This repository contains scripts of activities performed on various deep learning concepts
PocketPiglet for iOS
PocketPiglet for Android
Identifying individual speakers in an audio stream based on the unique characteristics found in individual voices using Python
Synchronize your subtitles using machine learning
A complete speech segmentation system using Kaldi and x-vectors for voice activity detection (VAD) and speaker diarisation.
VadRecorder based webrtc's VAD engine and vo-aac encoder, recording valid speech and discarding silence/noise data
EduSense: Practical Classroom Sensing at Scale
iOS Voice Activity Detection (VAD). Supports WebRTC VAD GMM, Silero VAD DNN, Yamnet VAD DNN models.
Cross-platform, real-time, offline speech recognition plugin for Unreal Engine. Based on Whisper OpenAI technology, whisper.cpp.
Smart human voice recorder using Silero VAD (PyTorch) + frequency analysis — detects & records speech while filtering background noise. Auto-stops on silence with configurable thresholds.
A Python-based system for automatic word segmentation in speech using ML models like SVM, MLP, and RNN.
Speech-end detection library, based on WebRTC's VAD engine
A simple, mobile, friendly command detection model
Android Voice Activity Detection (VAD) library. Supports WebRTC VAD GMM, Silero VAD DNN, Yamnet VAD DNN models.
🎙️ AI-powered Voice Activity Detection — Automatically detect and split speech segments from long audio files using Silero VAD. Beautiful Liquid Glass UI with drag-and-drop upload, visual timeline, and one-click ZIP export. No command line needed.
To associate your repository with the speech-detection topic, visit your repo's landing page and select "manage topics."