Lists (1)
Sort Name ascending (A-Z)
Stars
Build voice agents with open-source models
Research code for AlignAtt4LLM simultaneous speech translation.
Standalone causal streaming runtime for Qwen3-ASR
Interactive D3.js network graph component for Streamlit — force-directed layout, zone clustering, multiple shapes, search, actions, i18n
NVIDIA Canary ASR model optimized for Apple Silicon using MLX.
Recursive analysis of python stack of packages using AWS Strands Agents
Simultaneous translation model for 200 languages
Awesome autocompletion, static analysis and refactoring library for python
A scalable generative AI framework built for researchers and developers working on Large Language Models, Multimodal, and Speech AI (Automatic Speech Recognition and Text-to-Speech)
FULL Augment Code, Claude Code, Cluely, CodeBuddy, Comet, Cursor, Devin AI, Junie, Kiro, Leap.new, Lovable, Manus, NotionAI, Orchids.app, Perplexity, Poke, Qoder, Replit, Same.dev, Trae, Traycer AI…
Full toolkit for running an AI agent service built with LangGraph, FastAPI and Streamlit
Agentic RAG platform purpose-built for small language models (SLM). Robust PDF/SQL search
Whisper Streaming with Websocket and Fastapi server
Real-time, local speech-to-text with streaming ASR, speaker diarization, translation, and OpenAI/Deepgram-compatible APIs.
Tensors and Dynamic neural networks in Python with strong GPU acceleration
Data manipulation and transformation for audio signal processing, powered by PyTorch
A python package to build AI-powered real-time audio applications
OCRmyPDF adds an OCR text layer to scanned PDF files, allowing them to be searched
Whisper realtime streaming for long speech-to-text transcription and translation
TopicGPT allows to integrate the benefits of LLMs into Topic Modelling
A WebUI to create song covers with any RVC v2 trained AI voice from YouTube videos or audio files.




