AI service
Speech/music source separation
We provide Speech/music source separation at any scale, on a private cluster dedicated to you — from a single GPU to a hundred, on premise or in any cloud.
If you need such a service please contact us.
Voices, music and effects are split into separate tracks from a single mixed recording. That allows a dialogue track to be transcribed cleanly, or a music bed to be replaced without re-recording the narration. Separation quality is reported per track, so a poor split is visible before anything downstream depends on it.
The service runs on a private, dedicated cluster sized for your volume — a single GPU for a pilot, up to a hundred for a production estate. It is deployed where your data already lives: on premise, in your own AWS, GCP or OCI account, or in the super-gpu cloud.
More services
- PicturesAI models for Pictures →
- DocumentsAI models for Documents →
- Video filesAI models for Video files →
- Audio filesAI models for Audio files →
- Live feedAI models for Live feed →