Author: "Monica Villanueva" / Publisher: kth, matematik (inst.) - Searchworks@Jio Institute Digital Library Search Results

Searchworks

Author: Andreu, Sergi, Aylagas, Monica Villanueva, Andreu, Sergi, and Aylagas, Monica Villanueva
Abstract: Creating variations of sound effects for video games is a time-consuming task that grows with the size and complexity of the games themselves. The process usually comprises recording source material and mixing different layers of sounds to create sound effects that are perceived as diverse during game-play. In this work, we present a method to generate controllable variations of sound effects that can be used in the creative process of sound designers. We adopt WaveFlow, a generative flow model that works directly on raw audio and has proven to perform well for speech synthesis. Using a lower-dimensional mel spectrogram as the conditioner allows both user controllability and a way for the network to generate more diversity. Additionally, it gives the model style transfer capabilities. We evaluate several models in terms of the quality and variability of the generated sounds using both quantitative and subjective evaluations. The results suggest that there is a trade-off between quality and diversity. Nevertheless, our method achieves a quality level similar to that of the training set while generating perceivable variations according to a perceptual study that includes game audio experts., QC 20231107
Published: 2022
Full Text: View/download PDF

Searchworks