Raw waveform diffusion matches autoencoder quality
A new raw-waveform diffusion model called WavFlow has achieved audio fidelity matching or exceeding that of autoencoder-based pipelines, eliminating the need for latent compression. On the VGGSound benchmark, WavFlow's F…