You are currently logged in as an
Institutional Subscriber.
If you would like to logout,
please click on the button below.
Home / Publications / E-library page
Only AES members and Institutional Journal Subscribers can download
In speech/music coders and analysis/synthesis systems, spectral modeling is generally performed on a short-term (ST) frame-by-frame basis, which is justified by the fact that the signal is only locally (quasi-) stationary. The vocal tract configuration moves slowly and smoothly thereby resulting in a high correlation between the spectral parameters of successive frames: this correlation property is exploited in long-term modeling of the ST parameters, which however results in longer modeling/coding delays. The short delay constraint can be relaxed in many applications, such as text-to-speech modification/synthesis, telephony surveillance data, digital answering machines, electronic voicemail, digital voice logging, electronic toys, and video games. The long-term harmonic plus noise model (LT-HNM) for speech shows additional data compression possibilities since it exploits the smooth evolution of the time trajectories of the short-term harmonic plus noise model parameters by applying a discrete cosine model (DCM). In this paper, the authors extend the LT-HNM to a complete low bit-rate speech coder that is based on a long-term approach ca. 200ms. The proposed LT-HNM coder reaches a bit-rate of 2.7kbps for wideband speech.
Author (s): Ben Ali, Faten; Djaziri-Larbi, Sonia; Girin, Laurent
Affiliation:
University of Tunis El Manar, National Engineering School of Tunis, Signal and Systems Lab, Tunis, Tunisia; GIPSA Lab, University Grenoble Alpes, France, and INRIA Grenoble Rhone-Alpes, France
(See document for exact affiliation information.)
Publication Date:
2016-11-06
Import into BibTeX
Permalink: https://aes2.org/publications/elibrary-page/?id=18522
(682KB)
Click to purchase paper as a non-member or login as an AES member. If your company or school subscribes to the E-Library then switch to the institutional version. If you are not an AES member Join the AES. If you need to check your member status, login to the Member Portal.
Ben Ali, Faten; Djaziri-Larbi, Sonia; Girin, Laurent; 2016; Low Bit-Rate Speech Codec Based on a Long-Term Harmonic Plus Noise Model [PDF]; University of Tunis El Manar, National Engineering School of Tunis, Signal and Systems Lab, Tunis, Tunisia; GIPSA Lab, University Grenoble Alpes, France, and INRIA Grenoble Rhone-Alpes, France; Paper ; Available from: https://aes2.org/publications/elibrary-page/?id=18522
Ben Ali, Faten; Djaziri-Larbi, Sonia; Girin, Laurent; Low Bit-Rate Speech Codec Based on a Long-Term Harmonic Plus Noise Model [PDF]; University of Tunis El Manar, National Engineering School of Tunis, Signal and Systems Lab, Tunis, Tunisia; GIPSA Lab, University Grenoble Alpes, France, and INRIA Grenoble Rhone-Alpes, France; Paper ; 2016 Available: https://aes2.org/publications/elibrary-page/?id=18522
@article{ben2016low,
author={ben ali faten and djaziri-larbi sonia and girin laurent},
journal={journal of the audio engineering society},
title={low bit-rate speech codec based on a long-term harmonic plus noise model},
year={2016},
volume={64},
issue={11},
pages={844-857},
month={november},}