Skip to content
View Yangyangii's full-sized avatar
🎯
Focusing
🎯
Focusing

Organizations

@Deepest-Project

Block or report Yangyangii

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

The agent harness performance optimization system. Skills, instincts, memory, security, and research-first development for Claude Code, Codex, Opencode, Cursor and beyond.

JavaScript 273,739 40,844 Updated Oct 5, 2026
Rust 1,786 254 Updated Oct 5, 2026

Official implementation of "Something from Nothing: Data Augmentation for Robust Severity Level Estimation of Dysarthric Speech"

Python 9 2 Updated Jul 4, 2026

from vibe coding to agentic engineering - practice makes claude perfect

HTML 67,150 6,708 Updated Oct 6, 2026

A hand-picked collection of the finest of resources for the most awesome of agents, Claude Code, the undisputed champion of coding companions, from the unstoppable team at Anthropic PBC. A delectab…

Python 55,118 4,788 Updated Oct 6, 2026
Python 52 6 Updated Aug 20, 2026

Lightning-Fast, On-Device, Multilingual TTS — running natively via ONNX.

Swift 13,774 1,549 Updated Sep 9, 2026

Character-aware audio-only subtitling

Python 31 2 Updated Jun 15, 2025

VoicePrompter: Robust Zero-Shot Voice Conversion with Voice Prompt and Conditional Flow Matching (ICASSP '25)

9 Updated Jan 11, 2025

A Survey of Spoken Dialogue Models (60 pages)

317 17 Updated Nov 28, 2024

Automatically Update Text-to-speech (TTS) Papers Daily using Github Actions (Update Every 12th hours)

Python 670 42 Updated Oct 5, 2026
HTML 1 Updated Jan 29, 2025

Official Pytorch Implementation for "DDDM-VC: Decoupled Denoising Diffusion Models with Disentangled Representation and Prior Mixup for Verified Robust Voice Conversion" (AAAI 2024)

Python 247 26 Updated Jul 31, 2024

This repo contains the scripts, models, and required files for the Deep Noise Suppression (DNS) Challenge.

Python 1,471 458 Updated Jul 25, 2024

Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP serving.

Python 20,589 2,060 Updated Oct 2, 2026

Expressive Anechoic Recordings of Speech (EARS)

Python 229 14 Updated Jun 25, 2024

LibriTTS-P: A Corpus with Speaking Style and Speaker Identity Prompts for Text-to-Speech and Style Captioning

165 3 Updated Jun 13, 2024
Python 61 3 Updated Aug 28, 2026

🔊 Text-Prompted Generative Audio Model

Jupyter Notebook 39,264 4,664 Updated Aug 19, 2024

RWKV (pronounced RwaKuv) is an RNN with great LLM performance, which can also be directly trained like a GPT transformer (parallelizable). We are at RWKV-7 "Goose". So it's combining the best of RN…

Python 14,740 1,022 Updated Sep 28, 2026

Code for loralib, an implementation of "LoRA: Low-Rank Adaptation of Large Language Models"

Python 13,833 935 Updated Dec 17, 2024

Kandinsky 2 — multilingual text2image latent diffusion model

Jupyter Notebook 2,810 319 Updated May 1, 2024

Tools for handling multimodal data in machine learning projects.

Python 1,155 285 Updated Sep 30, 2026

Keep track of big models in audio domain, including speech, singing, music etc.

517 33 Updated Jul 3, 2026

AudioLDM: Generate speech, sound effects, music and beyond, with text.

Python 2,911 274 Updated Jun 25, 2025

Singing Voice Conversion via diffusion model

Jupyter Notebook 2,720 809 Updated Jun 6, 2026

Contrastive Language-Audio Pretraining

Python 2,299 213 Updated May 15, 2025

🚀 A simple way to launch, train, and use PyTorch models on almost any device and distributed configuration, automatic mixed precision (including fp8), and easy-to-configure FSDP and DeepSpeed support

Python 9,906 1,529 Updated Oct 5, 2026

A playbook for systematically maximizing the performance of deep learning models.

30,354 2,421 Updated Jun 18, 2024
Next