TUNENOODLE / Stem splitter / FX & Dialogue
Crowd and dialogue separator
Separate a crowd-focused layer from a live recording, with the remaining performance in Other for comparison and editing.
Crowd and dialogue separator
A live recording contains more than the performance: applause, audience chatter, cheers and whistles can dominate the spaces between phrases. This mode returns two files, Crowd / dialogue and Other. Its MVSEP origin is a crowd-removal model, so the useful starting question is how much audience sound can be separated from the performance. It does not deliver independent speech, music and sound-effects tracks. Use the target to inspect where audience activity rises, or audition Other when preparing a quieter live-music excerpt. Choose a section with both music and a transition into applause. That comparison reveals whether the performance stays natural as the crowd becomes louder. Dialogue in the label should not be read as speaker identification or a guarantee that every spoken word will move to the crowd layer. Check handclaps that coincide with snare hits, whistles near sung notes and spoken introductions over sustained music. These overlaps are harder to judge than applause during silence. Some audience presence may be worth keeping for a believable live atmosphere. Compare a moderate reduction with complete removal, and listen for changes in the room sound before making the final edit.
How to separate your track
- 01
Choose Crowd / dialogue and separate the live recording.
- 02
Compare the crowd target and Other during both exposed applause and music overlap.
- 03
Select a natural amount of crowd reduction, then trim the usable performance excerpt.
What this mode gives you
| Parameter / option | What it changes |
|---|---|
| Start with the right source | Include a music-to-applause transition and some applause under music. Crowd in an otherwise silent gap is an easy example but does not reveal whether the performance survives overlap. |
| Good for | Reduce applause around a song transition while keeping a natural live atmosphere. Study a live passage whose quiet details are masked by audience activity. Inspect crowd reactions as a separate layer when planning an edit. |
| Listen for | Are handclaps being confused with snare or other performance percussion? Do whistles or cheering remove part of a sung or instrumental note? Does Other keep a consistent room ambience through the transition? |
A few useful answers
No. It returns Crowd / dialogue and Other. Choose a three-stem dialogue, music and effects mode if those are the files you need.
Your next step
All modesDialogue, music and effects separator
Create separate Speech, Music and Sound effects tracks from a mixed soundtrack with the Demucs4HT DnR mode.
BandIt Plus soundtrack separator
Separate Speech, Music and Sound effects with BandIt Plus, then compare the layers around the words and scene events that matter.
BandIt v2 multilingual soundtrack separator
Separate Speech, Music and Sound effects with the BandIt v2 mode, informed by multilingual cinematic-audio research.
DnR v3 soundtrack separator
Separate Speech, Music and Sound effects with MVSEP’s DnR v3 mode for detailed comparison of mixed cinematic audio.
Braam effect separator
Isolate a braam-style cinematic hit to study its low-end impact, tonal body and decay apart from the surrounding arrangement.
Riser effect separator
Extract rising transition effects to study buildup shape, phrase length and the arrival of a drop or chorus.
FX separator
Separate an effects-focused layer from a mixed recording to inspect transitions, accents and textures alongside the remaining audio.
Uses credits
