File size: 934 Bytes
855b3fc
 
 
 
 
 
 
 
 
 
 
 
 
 
039e024
855b3fc
 
 
 
 
 
 
 
 
 
 
 
 
 
1
2
3
4
5
6
7
8
9
10
11
12
13
14
15
16
17
18
19
20
21
22
23
24
25
26
27
28
29
30
---

title: Vocal Separation SOTA
emoji: 🎤
colorFrom: red
colorTo: gray
sdk: gradio
sdk_version: 6.3.0
app_file: app.py
pinned: true
license: mit
---


# Vocal Separation SOTA

This is a demo for SOTA vocal separation models. Upload an audio file and the model will separate the vocals from the background music.

Based on the result of [MDX23](https://www.aicrowd.com/challenges/sound-demixing-challenge-2023/problems/music-demixing-track-mdx-23/leaderboards), the current SOTA model is [BS-RoFormer](https://arxiv.org/abs/2309.02612).

For comparison, you can also try the Mel-RoFormer model (a variant of BS-RoFormer) and the popular HTDemucs FT model.

## Models

- BS-RoFormer
- Mel-RoFormer
- HTDemucs FT

> The models are trained by the [UVR project](https://github.com/Anjok07/ultimatevocalremovergui).

> Based on [JacobLinCool/vocal-separation](https://github.com/JacobLinCool/vocal-separation).