Blind source separation by long-term monitoring: A variational autoencoder to validate the clustering analysis

Domenico De Salvio; Michael J Bianco; Peter Gerstoft; Dario D'Orazio; Massimo Garai

doi:10.1121/10.0016887

Blind source separation by long-term monitoring: A variational autoencoder to validate the clustering analysis

J Acoust Soc Am. 2023 Jan;153(1):738. doi: 10.1121/10.0016887.

Authors

Domenico De Salvio¹, Michael J Bianco², Peter Gerstoft², Dario D'Orazio¹, Massimo Garai¹

Affiliations

¹ Department of Industrial Engineering (DIN), University of Bologna, Viale del Risorgimento 2, Bologna, 40136, Italy.
² NoiseLab, Scripps Institution of Oceanography, University of California San Diego, La Jolla, California 92037, USA.

PMID: 36732230
DOI: 10.1121/10.0016887

Abstract

Noise exposure influences the comfort and well-being of people in several contexts, such as work or learning environments. For instance, in offices, different kind of noises can increase or drop the employees' productivity. Thus, the ability of separating sound sources in real contexts plays a key role in assessing sound environments. Long-term monitoring provide large amounts of data that can be analyzed through machine and deep learning algorithms. Based on previous works, an entire working day was recorded through a sound level meter. Both sound pressure levels and the digital audio recording were collected. Then, a dual clustering analysis was carried out to separate the two main sound sources experienced by workers: traffic and speech noises. The first method exploited the occurrences of sound pressure levels via Gaussian mixture model and K-means clustering. The second analysis performed a semi-supervised deep clustering analyzing the latent space of a variational autoencoder. Results show that both approaches were able to separate the sound sources. Spectral matching and the latent space of the variational autoencoder validated the assumptions underlying the proposed clustering methods.