TUNENOODLE / Stem splitter / FX & Dialogue

DnR v3 soundtrack separator

Separate Speech, Music and Sound effects with MVSEP’s DnR v3 mode for detailed comparison of mixed cinematic audio.

DnR v3 soundtrack separator

DnR v3 names a multilingual cinematic-audio dataset as well as this MVSEP separation option. The dataset research revisited language coverage, vocal material in non-dialogue tracks and the way soundtrack mixtures are balanced. MVSEP describes its DnR v3 models as trained on that dataset using SCNet and MelBand Roformer variants. This is a different model offering from BandIt v2, even though both share the DnR v3 research context and return Speech, Music and Sound effects. Use it when your edit depends on more than a clean voice: perhaps the score needs to remain continuous beneath a conversation, or an important scene effect must survive a loud line. Pick the scene event first, then compare how all three outputs preserve it. The dataset name and published test results do not predict the outcome for every real soundtrack. Pay special attention to sound that sits between roles, such as singing, a tonal alarm or rhythmic machinery. Check where it lands instead of assuming the category from its name. This mode does not supply separate singers, speakers or individual effects. Keep the original excerpt available and choose the stems that preserve the meaning, timing and atmosphere of the scene you are editing.

How to separate your track

  1. 01

    Choose DnR v3 and submit the original soundtrack passage.

  2. 02

    Identify the critical event and check its full duration in all three outputs.

  3. 03

    Compare with another mode on that same passage if the edit requires better preservation of a particular layer.

What this mode gives you
What this mode gives you
Parameter / optionWhat it changes
Start with the right sourceChoose a scene where your required effect overlaps speech and music. Add a quiet section before or after it so you can compare the retained ambience and avoid judging only the loud moment.
Good forKeep a score cue continuous while lowering dialogue in an analysis excerpt. Inspect whether a key alarm, impact or ambience survives under speech. Compare a DnR-trained alternative with BandIt v2 on the same scene.
Listen forDoes a tonal effect stay recognizable, whichever stem contains it? Is the score continuous beneath dialogue without audible word fragments? Do speech and room ambience remain natural before and after the main event?

A few useful answers

No. They are distinct options. Both relate to DnR v3 research data, while MVSEP documents different model architectures for its DnR v3 offering.

Your next step

All modes

Send feedback

Tell us what went wrong or what would help. Every message is read.

Type

Sending from /features/stem-splitter/dnr-v3