Figure 2
A diagram illustrating a Deep Neural Network architecture for acoustic feature classification.

MLP-based AF extractor is used to predict the posterior probabilities of speech attributes. (The input example is “hello” that is constructed by phone sequence of [HH, AH, L, OW]).

or Create an Account

Close subscription notice
Close access options