Table 4.

MLM training configurations of the text rescoring module across different data sets

Data setAISHELL-1TEDLIUM-2LibriSpeech
Epochs101010
Batch size256128128
Gradient accumulation122
Learning rate1e-51e-51e-5
Warm up Ratio0.010.010.01

or Create an Account

Close subscription notice
Close access options