Company-run audio benchmark · Published July 30, 2026

AI Vocal Remover Benchmark 2026: NeuralSound vs Moises vs Fadr

Compare the same five songs processed by three AI vocal removers. Play every input, isolated vocal and instrumental, then review the objective SI-SDR results and methodology.

Published by Neural Sound LLC · Updated July 31, 2026 · Five-song, two-stem benchmark

AI Vocal Remover Benchmark 2026 comparing NeuralSound, Moises and Fadr

Disclosure: NeuralSound conducted this benchmark. The same input songs and evaluation pipeline were used for every product. These five tracks do not prove that one service will perform best on every recording.

Quick answer

NeuralSound led this five-track AI vocal remover benchmark

NeuralSound reached an overall SI-SDR of 15.80 dB, compared with 14.47 dB for Moises and 12.59 dB for Fadr. Higher is better for these measurements, but the result applies only to this test set and date.

NeuralSound

15.80 dB

Highest result in this test

Moises

14.47 dB

Second-highest result in this test

Fadr

12.59 dB

Third result in this test

Average benchmark results for Fadr, Moises and NeuralSound
What we comparedFadrMoisesNeuralSound
Vocal clarity (SI-SDR dB)10.0212.0013.26
Instrumental clarity (SI-SDR dB)15.1616.9418.33
Cleanliness / less bleed (SI-SIR dB)25.5530.2533.30
Fewer processing artifacts (SI-SAR dB)12.8914.6415.93
Overall quality (SI-SDR dB)12.5914.4715.80
Historical entry option (checked July 2026)Free BasicFree: 5 songs/month$1.99 / 10 minutes

Download the benchmark data

Inspect the averages or all 15 product-by-track result rows.

What SI-SDR, SI-SIR and SI-SAR measure

Each score describes a different part of source-separation quality. Use the measurements alongside the listening tests.

SI-SDR

Overall reconstruction quality

Measures how closely an estimated stem matches its reference after accounting for scale. Higher values indicate a closer overall match.

SI-SIR

Separation from unwanted sources

Measures interference left in the stem. A higher score generally means less accompaniment in the vocal or less vocal residue in the instrumental.

SI-SAR

Processing artifacts

Measures distortion introduced by the separation process. Higher values generally indicate fewer separation artifacts.

Technical reference: SI-SDR evaluation paper.

Play all 35 benchmark audio previews

Start with each original mixture, then compare the isolated vocals and instrumentals from NeuralSound, Moises and Fadr. Audio loads only after you select a sample. Starting another sample stops and unloads the previous request.

Track 1 of 5

The Districts - Vermont

Same input · Two-stem mode · Scores in dB

Original input mixture

NeuralSound

Average SI-SDR: 14.36 dB

Highest

Isolated vocals

SI-SDR
13.01
SI-SIR
29.76
SI-SAR
13.11

Instrumental

SI-SDR
15.70
SI-SIR
28.84
SI-SAR
15.92

Moises

Average SI-SDR: 13.29 dB

Isolated vocals

SI-SDR
11.96
SI-SIR
26.70
SI-SAR
12.11

Instrumental

SI-SDR
14.61
SI-SIR
27.50
SI-SAR
14.85

Fadr

Average SI-SDR: 11.78 dB

Isolated vocals

SI-SDR
10.40
SI-SIR
23.50
SI-SAR
10.64

Instrumental

SI-SDR
13.16
SI-SIR
23.92
SI-SAR
13.56

Track 2 of 5

The Long Wait - Back Home To Blue

Same input · Two-stem mode · Scores in dB

Original input mixture

NeuralSound

Average SI-SDR: 18.93 dB

Highest

Isolated vocals

SI-SDR
15.65
SI-SIR
40.05
SI-SAR
15.67

Instrumental

SI-SDR
22.21
SI-SIR
38.63
SI-SAR
22.31

Moises

Average SI-SDR: 16.81 dB

Isolated vocals

SI-SDR
13.53
SI-SIR
34.98
SI-SAR
13.57

Instrumental

SI-SDR
20.09
SI-SIR
34.67
SI-SAR
20.24

Fadr

Average SI-SDR: 14.27 dB

Isolated vocals

SI-SDR
10.94
SI-SIR
27.53
SI-SAR
11.04

Instrumental

SI-SDR
17.60
SI-SIR
29.57
SI-SAR
17.89

Track 3 of 5

The Scarlet Brand - Les Fleurs Du Mal

Same input · Two-stem mode · Scores in dB

Original input mixture

NeuralSound

Average SI-SDR: 13.51 dB

Highest

Isolated vocals

SI-SDR
10.59
SI-SIR
30.15
SI-SAR
10.65

Instrumental

SI-SDR
16.43
SI-SIR
28.65
SI-SAR
16.70

Moises

Average SI-SDR: 12.63 dB

Isolated vocals

SI-SDR
9.99
SI-SIR
28.51
SI-SAR
10.05

Instrumental

SI-SDR
15.28
SI-SIR
26.86
SI-SAR
15.60

Fadr

Average SI-SDR: 11.78 dB

Isolated vocals

SI-SDR
8.85
SI-SIR
24.85
SI-SAR
8.98

Instrumental

SI-SDR
14.71
SI-SIR
24.12
SI-SAR
15.26

Track 4 of 5

The So So Glos - Emergency

Same input · Two-stem mode · Scores in dB

Original input mixture

NeuralSound

Average SI-SDR: 12.31 dB

Highest

Isolated vocals

SI-SDR
9.32
SI-SIR
26.86
SI-SAR
9.41

Instrumental

SI-SDR
15.30
SI-SIR
25.88
SI-SAR
15.71

Moises

Average SI-SDR: 11.52 dB

Isolated vocals

SI-SDR
8.53
SI-SIR
24.30
SI-SAR
8.66

Instrumental

SI-SDR
14.52
SI-SIR
24.66
SI-SAR
14.97

Fadr

Average SI-SDR: 10.13 dB

Isolated vocals

SI-SDR
7.05
SI-SIR
20.79
SI-SAR
7.28

Instrumental

SI-SDR
13.22
SI-SIR
21.33
SI-SAR
13.98

Track 5 of 5

The Wrong'Uns - Rothko

Same input · Two-stem mode · Scores in dB

Original input mixture

NeuralSound

Average SI-SDR: 19.88 dB

Highest

Isolated vocals

SI-SDR
17.74
SI-SIR
41.21
SI-SAR
17.76

Instrumental

SI-SDR
22.02
SI-SIR
43.00
SI-SAR
22.05

Moises

Average SI-SDR: 18.11 dB

Isolated vocals

SI-SDR
15.99
SI-SIR
36.98
SI-SAR
16.02

Instrumental

SI-SDR
20.22
SI-SIR
37.31
SI-SAR
20.31

Fadr

Average SI-SDR: 14.97 dB

Isolated vocals

SI-SDR
12.87
SI-SIR
29.20
SI-SAR
12.98

Instrumental

SI-SDR
17.08
SI-SIR
30.71
SI-SAR
17.28

Methodology

We tested the first five songs in the valid folder of a MUSDB18-HQ copy. The identical mixture was submitted to NeuralSound, Moises and Fadr in two-stem mode to produce vocals and an instrumental.

The original vocal stem was the vocal reference. The instrumental reference was calculated from the complete mixture minus the vocal source. Estimated outputs were converted to mono at 44.1 kHz and aligned to the references using cross-correlation.

Timing offsets and file-length differences were corrected only to synchronize the comparison. No denoising, equalization or quality enhancement was applied before scoring SI-SDR, SI-SIR and SI-SAR.

Fadr was tested using its Free Basic entry option and Moises using its Free entry option. The labels and entry options were checked in July 2026.

Sources: MUSDB18 documentation and the SI-SDR paper.

Limitations and interpretation

  • Company-run study: NeuralSound designed and conducted the benchmark, so the disclosure should be considered when interpreting the result.
  • Small sample: Five tracks cannot represent every genre, production style, recording condition or source overlap.
  • Changing services: Cloud products can update their separation systems after the July 2026 test date.
  • No universal winner: Objective scores help compare reconstruction, interference and artifacts, but listeners may prefer a different output for a particular song or workflow.

Hear how NeuralSound handles your track

Upload an audio or video file and compare the separated vocals and instruments for yourself.