Enhanced Fuzzy Decomposition of Sound Into Sines, Transients, and Noise

Leonardo Fierro*, Vesa Välimäki

*Tämän työn vastaava kirjoittaja

Tutkimustuotos: LehtiartikkeliArticleScientificvertaisarvioitu

3 Sitaatiot (Scopus)
74 Lataukset (Pure)

Abstrakti

The decomposition of sounds into sines, transients, and noise is a long-standing research problem in audio processing. The current solutions for this three-way separation detect either horizontal and vertical structures or anisotropy and orientations in the spectrogram to identify the properties of each spectral bin and classify it as sinusoidal, transient, or noise. This paper proposes an enhanced three-way decomposition method based on fuzzy logic, enabling soft masking while preserving the perfect reconstruction property. The proposed method allows each spectral bin to simultaneously belong to two classes, sine and noise or transient and noise. Results of a subjective listening test against three other techniques are reported, showing that the proposed decomposition yields a better or comparable quality. The main improvement appears in transient separation, which enjoys little or no loss of energy or leakage from the other components and performs well for test signals presenting strong transients. The audio quality of the separation is shown to depend on the complexity of the input signal for all tested methods. The proposed method helps improve the quality of various audio processing applications. A successful implementation over a state-of-the-art time-scale modification method is reported as an example.

AlkuperäiskieliEnglanti
Sivut468-480
Sivumäärä13
JulkaisuAES: Journal of the Audio Engineering Society
Vuosikerta71
Numero7-8
DOI - pysyväislinkit
TilaJulkaistu - 2023
OKM-julkaisutyyppiA1 Alkuperäisartikkeli tieteellisessä aikakauslehdessä

Rahoitus

This work belongs to the activities of the “Nordic Sound and Music Computing Network—NordicSMC,” NordForsk project number 86892. The work of Leonardo Fierro was funded by the Aalto ELEC Doctoral School. The authors are grateful to Dennis Bontempi for the helpful discussions and to Alec Wright for proofreading.

Sormenjälki

Sukella tutkimusaiheisiin 'Enhanced Fuzzy Decomposition of Sound Into Sines, Transients, and Noise'. Ne muodostavat yhdessä ainutlaatuisen sormenjäljen.
  • NordicSMC: Nordic Sound and Music Computing Network

    Välimäki, V. (Vastuullinen johtaja), McCrea, M. (Projektin jäsen), Mikkonen, O. (Projektin jäsen), Louise, B. (Projektin jäsen), Martinez Ornelas, A. (Projektin jäsen), Tuovinen, J. (Projektin jäsen), Sinjanakhom, T. (Projektin jäsen), Fagerström, J. (Projektin jäsen), Akov, I. (Projektin jäsen), Parkkola, K. (Projektin jäsen), Roberts, J. (Projektin jäsen), Prawda, K. (Projektin jäsen) & Lindfors, J. (Projektin jäsen)

    01/01/201831/12/2023

    Projekti: Other external funding: Other foreign funding

Siteeraa tätä