Fundamental Frequency Model for Postfiltering at Low Bitrates in a Transform-Domain Speech and Audio Codec

dc.contributorAalto-yliopistofi
dc.contributorAalto Universityen
dc.contributor.authorDas, Snehaen_US
dc.contributor.authorBäckström, Tomen_US
dc.contributor.authorFuchs, Guillaumeen_US
dc.contributor.departmentDepartment of Signal Processing and Acousticsen
dc.contributor.groupauthorSpeech Communication Technologyen
dc.contributor.groupauthorSpeech Interaction Technologyen
dc.contributor.organizationFraunhofer Institute for Integrated Circuitsen_US
dc.date.accessioned2021-01-25T10:08:42Z
dc.date.available2021-01-25T10:08:42Z
dc.date.issued2020en_US
dc.description.abstractSpeech codecs can use postfilters to improve the quality of the decoded signal. While postfiltering is effective in reducing coding artifacts, such methods often involve processing in both the encoder and the decoder, rely on additional transmitted side information, or are highly dependent on other codec functions for optimal performance. We propose a low-complexity postfiltering method to improve the harmonic structure of the decoded signal, which models the fundamental frequency of the signal. In contrast to past approaches, the postfilter operates at the decoder as a standalone function and does not need the transmission of additional side information. It can thus be used to enhance the output of any codec. We tested the approach on a modified version of the EVS codec in TCX mode only, which is subject to more pronounced coding artefacts when used at its lowest bitrate. Listening test results show an average improvement of 7 MUSHRA points for decoded signals with the proposed harmonic postfilter.en
dc.description.versionPeer revieweden
dc.format.extent5
dc.format.mimetypeapplication/pdfen_US
dc.identifier.citationDas, S, Bäckström, T & Fuchs, G 2020, Fundamental Frequency Model for Postfiltering at Low Bitrates in a Transform-Domain Speech and Audio Codec. in Proceedings of Interspeech. vol. 2020-October, Interspeech, International Speech Communication Association (ISCA), pp. 2837-2841, Interspeech, Shanghai, China, 25/10/2020. https://doi.org/10.21437/Interspeech.2020-1067en
dc.identifier.doi10.21437/Interspeech.2020-1067en_US
dc.identifier.issn1990-9772
dc.identifier.otherPURE UUID: 20eab96e-675b-450b-a81f-2a06fbdfcd13en_US
dc.identifier.otherPURE ITEMURL: https://research.aalto.fi/en/publications/20eab96e-675b-450b-a81f-2a06fbdfcd13en_US
dc.identifier.otherPURE FILEURL: https://research.aalto.fi/files/55066711/Fundamental_Frequency_Model_for_Postfiltering_at_Low_Bitrates_in_a_Transform_Domain_Speech_and_Audio_Codec.pdf
dc.identifier.urihttps://aaltodoc.aalto.fi/handle/123456789/102093
dc.identifier.urnURN:NBN:fi:aalto-202101251402
dc.language.isoenen
dc.relation.ispartofInterspeechen
dc.relation.ispartofseriesProceedings of Interspeechen
dc.relation.ispartofseriesVolume 2020-October, pp. 2837-2841en
dc.relation.ispartofseriesInterspeechen
dc.rightsopenAccessen
dc.subject.keywordfundamental frequencyen_US
dc.subject.keywordpostfilteringen_US
dc.subject.keywordspeech codingen_US
dc.titleFundamental Frequency Model for Postfiltering at Low Bitrates in a Transform-Domain Speech and Audio Codecen
dc.typeA4 Artikkeli konferenssijulkaisussafi
dc.type.versionpublishedVersion

Files

Original bundle

Now showing 1 - 1 of 1
Loading...
Thumbnail Image
Name:
Fundamental_Frequency_Model_for_Postfiltering_at_Low_Bitrates_in_a_Transform_Domain_Speech_and_Audio_Codec.pdf
Size:
944.81 KB
Format:
Adobe Portable Document Format