
Recent research in speech processing exhibits a growing interest in unsupervised and self-supervised representation learning from unlabelled data to alleviate the need for large amounts of annotated data. We investigate several popular pre-training methods and apply them to Flemish Dutch. We compare off-the-shelf English pre-trained models to models trained on an increasing amount of Flemish data. We find that the most important factors for positive transfer to downstream speech recognition tasks include a substantial amount of data and a matching pre-training domain. Ideally, we also finetune Research goal: How does the transfer learning performance of self-supervised speech models pre-trained on Flemish Dutch compare to other low-resource languages when fine-tuned for automated speech recognition (ASR) tasks, as measured by word error rate (WER) on standardized benchmarks like LibriSpeech or Common Voice? Autonomous synthesis report generated by Assignee Research. Tribunal consensus score: 7.5/10.
This report was generated autonomously by Assignee Research, an owner-gated autonomous research lab. The content synthesizes findings from peer-reviewed papers. Tribunal score: 7.5/10.
models, learning, Flemish, self-supervised, speech, pre-trained, transfer, performance
models, learning, Flemish, self-supervised, speech, pre-trained, transfer, performance
| selected citations These citations are derived from selected sources. This is an alternative to the "Influence" indicator, which also reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically). | 0 | |
| popularity This indicator reflects the "current" impact/attention (the "hype") of an article in the research community at large, based on the underlying citation network. | Average | |
| influence This indicator reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically). | Average | |
| impulse This indicator reflects the initial momentum of an article directly after its publication, based on the underlying citation network. | Average |
