Skip to content
@audioshake

Audioshake

AudioShake

AudioShake builds AI models for separating and understanding sound.

Our technology turns mixed audio into usable components and structured data for music, media, speech, and machine-learning workflows.

Speech & Voice

Multi-Speaker Separation

Separate multi-speaker recordings into individual speaker tracks, including overlapping speech, with diarization and confidence scores.

Speech Recovery

Isolate dialogue and speech from background noise, music, and other interference for transcription, voice AI, media, and real-time applications.

Media

AudioShake provides tools for separating, detecting, and identifying audio in film, television, broadcast, and other media workflows.

Music

Separate, transcribe, and transform music for production, interactive, and creator workflows.

Research & Benchmarks

We build and contribute to open evaluation tools and benchmarks for audio AI.

ALT-Eval

ALT-Eval is an evaluation toolkit for Automatic Lyrics Transcription (ALT), designed to measure both transcription accuracy and readability.

JAM-ALT

JAM-ALT is a community benchmark for automatic lyrics transcription, developed by AudioShake and Spotify.

It provides human-transcribed lyrics with word-level timestamps for evaluating lyrics transcription and alignment systems.

Build with AudioShake

Popular repositories Loading

  1. alt-eval alt-eval Public

    Readability-aware automatic lyrics transcription (ALT) evaluation toolkit

    Python 44 1

  2. .github .github Public

Repositories

Showing 2 of 2 repositories

People

This organization has no public members. You must be a member to see who’s a part of this organization.

Top languages

Loading…

Most used topics

Loading…