Tool Name & Overview
Tool Name: AudioShake
Official Website / Link: https://www.audioshake.ai/
Brief Overview: AudioShake leverages award‑winning AI to separate audio into instrument stems, dialogue, music, and effects, plus lyric transcription with word‑level timing—revolutionizing workflows in mixing, remixing, sync licensing, and accessibility.
Licensing & Pricing
License Type: Freemium (Indie), Subscription (Indie tiers), Enterprise (AudioShake Live)
Cost Details:
Indie Starter plan: $20/mo for 4 stems (~$5/stem)
Higher Indie tiers available (e.g., 20 stems/mo)
AudioShake Live: volume‑based pricing, tailored for labels, publishers, film and TV pros
Free Tier / Trial: Preview stems free; 2 free stems/month on Indie plan
Notes on Licensing: Educational discounts and on‑premises API/SDK options available; enterprise users get account manager and stem storage
System Requirements & Compatibility
Platform: Web-based (AudioShake.co), API, SDK, widget; compatible with Windows, macOS, Linux, iOS/Android via browser .
DAW/Host Integration: Third-party integrations (e.g., Algoriddim djay Pro real‑time plugin)
Minimum Specs: Modern browser; supports up to 192 kHz WAV/MP3/FLAC/AIFF/PCM.
Additional Dependencies: Cloud account for web/API; on-device SDK requires embed setup; developer docs for API/SDK integration .
How to Use (Beginner Level)
Step 1: Sign up or log in to AudioShake Indie or Live
Step 2: Upload an audio file (MP3, WAV, FLAC, etc.)
Step 3: Choose stem‑separation model (e.g., vocals, drums, bass, others) or lyric transcription
Step 4: Preview and export stems or transcript; download WAV/MP3 or JSON/TXT
Tip for Beginners: Start with instrument stem separation at 44.1 kHz WAV for best balance—use the default 4-stem split
How to Use (Expert Level)
Advanced Settings: Select vocal‑only high‑quality model (SDR 13.5 dB benchmark), configure stem count up to 6+, enable word‑level lyric alignment .
Integration Tips:
Batch‑process tracks using API and SDK
Embed JavaScript widget in apps for user‑triggered stem extraction.
Workflow Optimization:
Auto‑generate stems across entire album for parallel mastering
Export stems, feed into DAW chain alongside reverb/Special FX tracks
Use lyric JSON with timestamp data to build karaoke or subtitle tools
Key Features & Benefits
High‑Quality Multi‑Stem Separation (vocals, drums, bass, guitar, piano, other): Enables granular control in remixing and mastering.
Dialogue, Music & Effects Separation: Ideal for film/TV post‑production and dubbing workflows
High‑Accuracy Lyric Transcription & Alignment: Generates time‑aligned lyric text for karaoke, subtitles, lyric videos.
Pro Developer Tools (API, SDK, Widget): Enables scalable, embeddable audio‑processing capabilities audioshake.ai
Award‑Winning Performance: Sony Demixing Challenge winner; new HQ vocal model surpasses industry SDR benchmarks.
Pros & Cons
Pros:
- Industry‑leading stem quality, surpasses Spleeter, Demucs, RX gearspace.com+1aimusicpreneur.com+1
- Fast processing; under a minute for typical tracks
- Supports up to 192 kHz, high‑bit depth
- Flexible usage models: Indie, enterprise, API/SDK/plugin
- Multi‑domain use: music, film, interactive apps, accessibility
Cons:
- Subscription cost can rise with stem volume
- Quality variable on complex orchestral mixes (e.g., individual strings)
- Requires internet/cloud for Indie; local use needs dev setup
- Desktop DAWs require manual import of exported stems
Use Cases & Examples
Example 1: Indie producer separates vocals and drums from a classic track, remixing it in Ableton Live for TikTok.
Example 2: Localization studio isolates dialogue and music from a film scene, improving subtitling accuracy and re-dubbing workflow.
User Feedback & Ratings
Community Reviews:
“Better transient preservation than RX rebalance”
“Quality doesn’t drop back into noise‑floor; industry leading tech”
Ratings Summary:
Indie plan users: ~4.5/5 for clarity, speed, utility
Pros note slight drop in orchestral separation compared to solo instruments
Related Tools / Alternatives
Spleeter (open‑source): Free, but lower separation quality and lacks lyric alignment
Demucs (via U‑Vocal Remover): Strong multi‑strom model; free and easily batchable but less accurate on vocals/transients than AudioShake
References & Further Reading
AudioShake Blog: insights on new HQ vocal model, Sony win, labels using stems
Developer docs/API guide for integration
Gearspace user review highlighting transient performance and limitations gearspace.com, indie.audioshake.ai