AudioShake: AI‑Powered Stem Separation & Lyric Transcription for Music Pros

Tool Name & Overview

Tool Name: AudioShake
Official Website / Link: https://www.audioshake.ai/
Brief Overview: AudioShake leverages award‑winning AI to separate audio into instrument stems, dialogue, music, and effects, plus lyric transcription with word‑level timing—revolutionizing workflows in mixing, remixing, sync licensing, and accessibility.

 

Licensing & Pricing

License Type: Freemium (Indie), Subscription (Indie tiers), Enterprise (AudioShake Live)

Cost Details:

Indie Starter plan: $20/mo for 4 stems (~$5/stem) 

Higher Indie tiers available (e.g., 20 stems/mo) 

AudioShake Live: volume‑based pricing, tailored for labels, publishers, film and TV pros 

Free Tier / Trial: Preview stems free; 2 free stems/month on Indie plan

Notes on Licensing: Educational discounts and on‑premises API/SDK options available; enterprise users get account manager and stem storage

 

System Requirements & Compatibility

Platform: Web-based (AudioShake.co), API, SDK, widget; compatible with Windows, macOS, Linux, iOS/Android via browser .

DAW/Host Integration: Third-party integrations (e.g., Algoriddim djay Pro real‑time plugin) 

Minimum Specs: Modern browser; supports up to 192 kHz WAV/MP3/FLAC/AIFF/PCM.

Additional Dependencies: Cloud account for web/API; on-device SDK requires embed setup; developer docs for API/SDK integration .

 

How to Use (Beginner Level)

Step 1: Sign up or log in to AudioShake Indie or Live

Step 2: Upload an audio file (MP3, WAV, FLAC, etc.)

Step 3: Choose stem‑separation model (e.g., vocals, drums, bass, others) or lyric transcription

Step 4: Preview and export stems or transcript; download WAV/MP3 or JSON/TXT

Tip for Beginners: Start with instrument stem separation at 44.1 kHz WAV for best balance—use the default 4-stem split

 

How to Use (Expert Level)

Advanced Settings: Select vocal‑only high‑quality model (SDR 13.5 dB benchmark), configure stem count up to 6+, enable word‑level lyric alignment .

Integration Tips:

Batch‑process tracks using API and SDK

Embed JavaScript widget in apps for user‑triggered stem extraction.

Workflow Optimization:

Auto‑generate stems across entire album for parallel mastering

Export stems, feed into DAW chain alongside reverb/Special FX tracks

Use lyric JSON with timestamp data to build karaoke or subtitle tools

 

Key Features & Benefits

High‑Quality Multi‑Stem Separation (vocals, drums, bass, guitar, piano, other): Enables granular control in remixing and mastering.

Dialogue, Music & Effects Separation: Ideal for film/TV post‑production and dubbing workflows 

High‑Accuracy Lyric Transcription & Alignment: Generates time‑aligned lyric text for karaoke, subtitles, lyric videos.

Pro Developer Tools (API, SDK, Widget): Enables scalable, embeddable audio‑processing capabilities audioshake.ai

Award‑Winning Performance: Sony Demixing Challenge winner; new HQ vocal model surpasses industry SDR benchmarks.

 

Pros & Cons

Pros:

  • Industry‑leading stem quality, surpasses Spleeter, Demucs, RX gearspace.com+1aimusicpreneur.com+1
  • Fast processing; under a minute for typical tracks
  • Supports up to 192 kHz, high‑bit depth
  • Flexible usage models: Indie, enterprise, API/SDK/plugin
  • Multi‑domain use: music, film, interactive apps, accessibility

Cons:

  • Subscription cost can rise with stem volume
  • Quality variable on complex orchestral mixes (e.g., individual strings) 
  • Requires internet/cloud for Indie; local use needs dev setup
  • Desktop DAWs require manual import of exported stems

 

Use Cases & Examples

Example 1: Indie producer separates vocals and drums from a classic track, remixing it in Ableton Live for TikTok.

Example 2: Localization studio isolates dialogue and music from a film scene, improving subtitling accuracy and re-dubbing workflow.

 

User Feedback & Ratings

Community Reviews:

“Better transient preservation than RX rebalance” 

“Quality doesn’t drop back into noise‑floor; industry leading tech” 

Ratings Summary:

Indie plan users: ~4.5/5 for clarity, speed, utility 

Pros note slight drop in orchestral separation compared to solo instruments 

 

Related Tools / Alternatives

Spleeter (open‑source): Free, but lower separation quality and lacks lyric alignment

Demucs (via U‑Vocal Remover): Strong multi‑strom model; free and easily batchable but less accurate on vocals/transients than AudioShake 

 

References & Further Reading

AudioShake Blog: insights on new HQ vocal model, Sony win, labels using stems 

Developer docs/API guide for integration 

Gearspace user review highlighting transient performance and limitations gearspace.com, indie.audioshake.ai