An audio signal processing method is performed by an electronic device, and the method includes: acquiring first spectral features of a noisy audio signal, the noisy audio signal including a reference audio signal and a target audio signal to be extracted; inputting the first spectral features into a target neural network model to obtain reference phase estimation information matching the noisy audio signal, and the reference phase estimation information being configured for indicating phase differences between the target audio signal and the noisy audio signal at a plurality of frequency components; determining second spectral features according to the first spectral features and the reference phase estimation information; and determining the target audio signal according to the second spectral features.
Full Text
What is claimed is: