Model icon
separate-vocals
Music
Model:
Vocal Remover API supports music workflows for music creation, audio production, localization, and media automation.
Input

Separate vocals and instrumental tracks from the audio

Output

Separated tracks will appear here

Pricing details

Transparent pricing with no hidden fees. Pay as you go.

PoYoseparate-vocals
Standard generation
PoYo price
$0.075/gen
15 credits
Official price
-
You save
-
PoYostem-split
Standard generation
PoYo price
$0.350/gen
70 credits
Official price
-
You save
-
PoYoupload-and-separate-vocals
Standard generation
PoYo price
$0.050/gen
10 credits
Official price
-
You save
-

* Actual fees are based on the final output.

Introduction

Complete guide to using Vocal Remover API - AI-Powered Audio Separation on PoYo

Vocal Remover API - AI-Powered Audio Separation on PoYo

Extract vocals, instrumentals, and individual stems from any audio track with PoYo's Vocal Remover API. Separate audio sources with professional-quality output. Choose between simple vocal/instrumental separation (15 credits) or advanced multi-stem splitting into drums, bass, vocals, and more (70 credits). Perfect for music production, remixing, karaoke creation, and audio analysis.

Features

01

Vocal Isolation

Extract clean vocal tracks from any music. Separate Vocals mode splits audio into isolated vocals and instrumentals with high-quality output, perfect for karaoke, remixing, or vocal analysis.

02

Multi-Stem Splitting

Decompose audio into individual instrument stems. Stem Split mode separates drums, bass, vocals, and other instruments into isolated tracks for professional music production and remixing.

03

Two Processing Modes

Choose Separate Vocals (15 credits) for simple vocal/instrumental separation, or Stem Split (70 credits) for advanced multi-track decomposition. Select the right mode for your needs and budget.

04

High-Quality Output

Professional-grade audio separation with minimal artifacts. Advanced AI models ensure clean separation while preserving the original audio quality and fidelity.

05

Multiple Audio Formats

Support for common audio formats including MP3, WAV, FLAC, and more. Process any audio file and receive separated tracks in high-quality output format.

06

Fast Processing

Quick audio separation with optimized AI processing. Webhook callbacks notify you instantly when your separated tracks are ready for download.

Use Cases

01

Music Production & Remixing

Extract stems for remixing, mashups, and music production. Isolate vocals, drums, bass, and instruments to create new arrangements and creative compositions.

02

Karaoke Creation

Generate karaoke tracks by removing vocals from songs. Create instrumental backing tracks for singing practice, karaoke apps, and vocal training platforms.

03

Podcast & Content Editing

Isolate voice tracks from background music for cleaner audio editing. Separate audio elements for more precise control in podcast and video production.

04

Music Education & Training

Create learning materials by isolating individual instruments. Help students study specific parts, practice along with isolated tracks, and understand musical arrangements.

05

Video & Film Post-Production

Separate audio tracks for more flexible sound design and mixing. Extract specific audio elements for re-editing, dubbing, and creative sound design in post-production.

06

Audio Analysis & Research

Isolate audio sources for detailed analysis and study. Perfect for music information retrieval, audio forensics, and academic research on musical composition.

How to Use Vocal Remover API on PoYo

1. Get Your API Key: Create a free PoYo account and generate your API key from the dashboard. Get API Key →

2. Choose Your Model: Select from three models based on your needs - Separate Vocals for generated music, Upload and Separate Vocals for your own audio files, or Stem Split for multi-track decomposition.

3. Make API Request: Send a POST request to the appropriate endpoint with required parameters. View documentation for each model:

4. Get Results: Receive a task_id and poll for results, or set up a webhook callback URL to be notified when separation completes with download URLs for each isolated track.

Pricing

Separate Vocals: $0.075 per separation (15 credits) - Split audio into vocals and instrumentals

Stem Split: $0.35 per separation (70 credits) - Decompose into multiple stems (drums, bass, vocals, etc.)

Features Included: High-quality audio separation, multiple format support, webhook callbacks, professional-grade output

Enterprise: Volume discounts available for high-volume usage. Priority processing and dedicated support. Contact sales →

Vocal Remover API FAQ

What's the difference between the three models (Separate Vocals, Upload and Separate Vocals, Stem Split)?

The Vocal Remover API offers three models: 1) Separate Vocals (15 credits) - Works with PoYo-generated music using task_id, splits into vocals and instrumentals. 2) Upload and Separate Vocals (15 credits) - Accepts your own uploaded audio files via audio URL, splits into vocals and instrumentals. 3) Stem Split (70 credits) - Works with PoYo-generated music, provides detailed separation into multiple stems including drums, bass, vocals, and other instruments. Choose Separate Vocals or Upload and Separate Vocals for simple vocal/instrumental separation, and Stem Split for professional production requiring individual instrument tracks.

When should I use Upload and Separate Vocals vs Separate Vocals?

Use Separate Vocals when processing music generated through PoYo's music generation APIs - it uses the task_id from previous generations. Use Upload and Separate Vocals when you have your own audio files (MP3, WAV, etc.) that weren't generated on PoYo - it accepts an audio URL parameter. Both cost 15 credits and produce the same quality vocal/instrumental separation.

What audio formats are supported?

The API supports common audio formats including MP3, WAV, FLAC, AAC, OGG, and more. For Upload and Separate Vocals, provide a publicly accessible URL to your audio file. The separated tracks will be provided in high-quality output format suitable for professional use.

How accurate is the vocal and stem separation?

The API uses advanced AI models from Suno to achieve professional-grade separation with minimal artifacts. While perfect separation is challenging for complex mixes, the quality is suitable for most professional applications including music production, remixing, and karaoke creation.

How do I retrieve the separated audio files?

After submitting a separation request, you receive a task_id. Use this to query the result via the Query Music Detail API, or set up a callback_url webhook to receive automatic notifications when separation completes. The response includes download URLs for each separated track (vocals, instrumentals, or individual stems depending on the model used).

Why Choose PoYo for Vocal Remover API

01

Flexible Pricing

Choose between two modes based on your needs: simple vocal separation at $0.075 or comprehensive stem splitting at $0.35. Pay only for what you use with transparent credit-based pricing.

02

Powered by Suno

Access Suno's state-of-the-art audio separation technology through PoYo's unified API. Professional-quality output with industry-leading AI models.

03

Fast & Reliable

Quick processing times with 99.9% uptime SLA. Webhook notifications ensure you're instantly alerted when your separated tracks are ready.

04

Professional Quality

Studio-grade audio separation with minimal artifacts and bleed. Suitable for professional music production, broadcasting, and commercial applications.

05

Unified API Access

One API key connects you to Vocal Remover and 15+ other music APIs plus 500+ AI models. Simplify billing and reduce integration complexity.

06

Developer-Friendly

Clean RESTful API with comprehensive documentation. Simple integration with any programming language. SDKs and code examples available.