AES E-Library

Audio Pitch Shifting Using the Constant-Q Transform

Pitch shifting of polyphonic music is usually performed by manipulating the time–frequency representation of the input signal such that frequency is scaled by a constant and time duration remains unchanged. A method for pitch shifting is proposed that exploits the logarithmic frequency-bin spacing of the Constant-Q Transform (CQT). Pitch-scaling of monophonic and dense polyphonic music signals is achieved by a simple linear translation of the CQT representation followed by a phase update stage. This approach provides a natural solution to the problems of transients because the CQT has good time resolution at high frequencies while interference between tonal components at low frequencies is reduced. Performing pitch shifting directly in the frequency domain allows the algorithm to process only parts of the signal while leaving other parts unchanged. Audio examples demonstrate the quality of the proposed algorithm for scaling factors up to an octave.

 

Author (s):
Affiliation: (See document for exact affiliation information.)
Publication Date:
Permalink: https://aes2.org/publications/elibrary-page/?id=16871


(277KB)


Click to purchase paper as a non-member or login as an AES member. If your company or school subscribes to the E-Library then switch to the institutional version. If you are not an AES member Join the AES. If you need to check your member status, login to the Member Portal.

Type:
E-Libary location:
16938
Choose your country of residence from this list:










Skip to content