Skip to main content

AI service

Speech/music source separation

We provide Speech/music source separation at any scale, on a private cluster dedicated to you — from a single GPU to a hundred, on premise or in any cloud.

If you need such a service please contact us.

Voices, music and effects are split into separate tracks from a single mixed recording. That allows a dialogue track to be transcribed cleanly, or a music bed to be replaced without re-recording the narration. Separation quality is reported per track, so a poor split is visible before anything downstream depends on it.

The service runs on a private, dedicated cluster sized for your volume — a single GPU for a pilot, up to a hundred for a production estate. It is deployed where your data already lives: on premise, in your own AWS, GCP or OCI account, or in the super-gpu cloud.

More services