Skip to content

AI SPEAKER SEPARATION

Separate speakers, simply.

Turn one mixed podcast, interview, or recording into individual tracks, with speakers split automatically and kept time-aligned.

  • 30+ Languages Supported
  • No install
  • No manual cutting
  • Audio + Video File Support

Separation workspace

AI Speaker Separation for Individual Speaker Tracks

Upload an audio or video recording for AI speaker separation, choose automatic or manual detection, and get speakers split into individual, time-aligned tracks ready to preview and download.

SEPARATION TOOL

Separate speakers from your recording

Sign in to upload a recording

Audio: WAV, MP3, M4A, AAC, FLAC, OGG, OPUS, WMA, Speex. Video: MP4, AVI, MOV, MKV, WebM. Uploads up to 1 GB. Compatible videos upload only their original audio track, without re-encoding.

Your recording stays private and is retained with completed tracks for 30 days.

Listen closely

Hear AI speaker separation in action.

Switch between the original mix and isolated speaker tracks without losing your place.

0:00 / 0:19Selected Original mix

How it works

How AI speaker separation works.

One continuous signal path turns a mixed recording into organized, editor-ready tracks.

SEPLY / SIGNAL FLOWSOURCE → OUTPUT
  1. 1

    STEP 01

    Upload your recording

    Choose a mixed podcast, interview, meeting, or conversation.

  2. 2

    STEP 02

    Let Seply separate the speakers

    AI identifies the voices and creates a time-aligned track for each detected speaker.

  3. 3

    STEP 03

    Preview and download

    Listen to every speaker and download each separated track as an individual WAV file.

Features

Everything you need to separate a conversation.

From speaker detection to secure delivery, every part of the workflow stays focused on usable audio.

CORE OUTPUT

One timeline. A clean track for every voice.

Individual speaker tracks

AI speaker separation gives each detected speaker an independent audio track.

Time-aligned output

Every track starts at the same point and stays aligned with the original recording.

Automatic speaker detection

03

Let Seply detect speakers automatically, or enter the expected number yourself.

Built for real conversations

04

Use Seply with podcasts, interviews, meetings, panels, and other multi-speaker recordings.

Editor-ready downloads

05

Download individual, time-aligned WAV files for your editing workflow.

Clear credit-based pricing

06

Each Seply feature has its own credit rate. See the estimated cost before processing begins.

Private processing
30-day private result access
Specialist processing infrastructure
Upload rights required

Separation vs diarization

More than speaker labels.

Speaker diarization

00:00-00:12   Speaker 1

00:12-00:19   Speaker 2

00:18-00:24   Speaker 1 + Speaker 2

Speaker diarization tells you who spoke when.

Speaker separation

speaker-01.wav
speaker-02.wav

AI speaker separation creates an actual track for each speaker, so you can edit, mute, enhance, or reuse every voice independently.

Use cases

An online speaker splitter built for real conversations.

Keep every voice editable, whether the final destination is an episode, research archive, meeting record, or video timeline.

SCENE 01

Podcast editing

Separate hosts and guests so you can adjust levels, remove mistakes, and edit each voice independently.

SCENE 02

Interviews and research

Isolate interviewers and participants for easier review, analysis, and speaker-specific workflows.

SCENE 03

Meetings and panels

Turn one mixed discussion into separate tracks for review, archiving, or post-production.

SCENE 04

Video and content production

Prepare independent dialogue tracks before placing them back into your video editing timeline.

Pricing

Simple pricing for Speaker Separation.

Choose a monthly plan or buy credits when you need them. The same Credits work across both Seply tools at their displayed rates.

Free Trial

$0

30 credits once

One-time trial, valid for 30 days. No card required.

  • Up to 3 minutes of Speaker Separation
  • Up to 90 minutes of Speech to Text
  • Valid for 30 days

Starter

$12/ month

200 credits every month

For occasional interviews and short recordings.

  • Up to 20 minutes of Speaker Separation
  • Up to 600 minutes of Speech to Text
  • Valid for 30 days
  • Renews monthly
  • Cancel any time

MOST POPULAR

Creator

$29/ month

600 credits every month

For regular podcast and interview production.

  • Up to 60 minutes of Speaker Separation
  • Up to 1800 minutes of Speech to Text
  • Valid for 30 days
  • Renews monthly
  • Cancel any time

Studio

$65/ month

1500 credits every month

For teams processing audio every week.

  • Up to 150 minutes of Speaker Separation
  • Up to 4500 minutes of Speech to Text
  • Valid for 30 days
  • Renews monthly
  • Cancel any time

Every plan includes

  • Automatic or manual speaker count
  • Separate speaker tracks
  • Audio and video file support
  • Private, secure processing
  • 30-day protected result access

Speaker Separation uses 1 Credit per started 6 seconds (10 per full minute). Speech to Text uses 1 Credit per started 3 minutes. Both have a 1-Credit minimum, and the exact estimate appears before processing.

Prices are in USD. Monthly plans renew automatically until canceled from Billing. See the Refund Policy.

FAQ

Frequently asked questions.

Clear answers about files, output, credits, and speaker separation.

What does Seply do?

Seply turns one mixed recording into separate, time-aligned audio tracks for each detected speaker.

Is speaker separation the same as speaker diarization?

No. Speaker diarization labels who spoke when. Speaker separation creates an actual audio track for each speaker.

Can Seply separate overlapping speech?

Seply is designed for multi-speaker recordings, including conversations with overlapping dialogue. Results can vary with noise, echo, recording quality, voice similarity, and the amount of overlap.

What files can I upload?

Seply supports the audio and video formats listed in the upload area, including WAV, MP3, M4A, AAC, FLAC, OGG, OPUS, WMA, Speex, MP4, AVI, MOV, MKV, and WebM.

Can I get speakers split into individual tracks?

Yes. Seply creates an independent, time-aligned WAV file for each detected speaker.

How do credits work?

Credits represent processing capacity within Seply. Each feature has its own rate. Speaker Separation uses 10 credits per minute, billed in 6-second increments with a 1-credit minimum, so a 1:01 recording uses 11 credits. The server-calculated media duration is authoritative.

Do credits expire?

Monthly plan credits expire 30 days after they are granted. One-time credit packs expire 12 calendar months after purchase. Canceling a subscription does not remove credits before their own expiration dates.

What happens if a separation job fails?

Credits reserved for a confirmed failed job are automatically returned to the original credit batches with their original expiration dates.

How long are results available?

For new jobs, the original recording and completed speaker tracks are stored privately for 30 days from completion. You can listen in your history and download separated tracks during that period. Limited job metadata remains after media expires.

What is the refund policy?

You may request a refund within 7 calendar days for purchased credits that have not been used. Used credits and completed jobs are generally not refundable. Exceptions apply for duplicate charges, failed delivery, and rights required by law.

Give every speaker their own track.

Start with 30 free credits. No card required.

Start separating