PROBLEM
Evey current leveler changes levels during vocal phrases. This is, by definition, distortion. MAutoVolume's additional controls help in other ways, but the distortion is inherent in the way all such plugins work. Meanwhile, experienced engineers still insist on performing this tedious task manually despite a variety of levelers being available for years. This will continue to be the case until an entirely different approach is adapted that not only eliminates this distortion entirely, but emulates a proper manual workflow.
SOLUTION
The ONLY way to avoid the distortion all the auto-levelers are causing is, by definition, to avoid changing levels during vocal phrases... and therefore to restrict level changes to moments of silence. In basic terms, the ONLY time a level change causes no waveform distortion is when it happens at a level of zero. What a user wants from such a plug is to normalize the levels (at least to a degree). What they don't want is to HEAR that leveling... as is the case with every currently available leveling plugin.
Every engineer I've met who does a particularly good job of vocal leveling already intuitively incorporates this concept into their workflow in one way or another... yet the plugins based on the current approach do not... and can not.
IMPLEMENTATION
Properly adjusting each section between silences is just one of those processes that REQUIRES a separate read pass (like Melodyne, Vocalign, and others) to do properly. It's not that vocal riding CAN'T be fully automated. It's that it can't be done within the currently over-saturated and under-effective single-stage Vocal Rider paradigm. As a two stage process, it not only achieves zero distortion, but does so with greater leveling accuracy, more precise control, and a more intuitive interface.
STAGE 1: READ PASS
The plug reads the entire track (think Vocalign or similar), and places internal markers at each silence of a minimum length. It then calculates the peak, RMS, and LUFS values of each resulting segment.
STAGE 2: USER CONTROL
There are two main dials for the user control. The first simplifies by reducing the number of markers. As the dial is turned, markers get grayed in order of increasing silence length. (Similar to many other simplification controls with markers like in Apple Loops utility.) It then recalculates the values for each resulting segment.
In more general terms, it controls the number of segments. At one end of the range, there are many segments, and at the other end there are fewer (and generally larger) segments. This is similar in manual editing to how many cuts an engineer might make to adjust individual regions either per word, per phrase, etc.
The second dial determines how tight the level matching is. At the one extreme, it changes nothing, and the performance is just as loose as originally recorded. At the other extreme, all segments are perfectly level matched. Effectively, this is a simple control for level matching "tightness" that maintains all existing dynamic relationships in exact ratios. The user can further choose whether the matching should be done by peaks, RMS, or LUFS.
That's it.
For those who want more control (and isn't that what drew most of us to Melda?), I would suggest that the user should also be able to manually grey out or re-enable the markers individually... just like they can with flex markers, etc. This allows the complete elimination of any residual issues where a breath was included as it's own segment when the user didn't want it to be, but everything else was good, etc.
This isn't a casual suggestion. It's YEARS in the making. On top of being the ONLY leveler design that causes ZERO distortion, it also completely eliminates the need for all sorts of complex time constant controls, and replaces them all with a simple and supremely intuitive dial. Gone are the days of chasing one leveler after another with different settings just to minimize the damage each is causing. Let's just stop causing ANY damage in the first place.
If this doesn't immediately strike anyone reading it as the right solution to a daily problem engineers have been facing since the dawn of recording, please read it again or ask questions. If I could drop everything and build one plug, this would be it. I can't right now, so I'm just going to sing it from the mountain tops until someone builds it.
I'd prefer Melda as I'd prefer to have more control over it whereas other devs might dumb it down too much... Plus I'm already paying for the subscription.
