Name: Data Augmentation for Spoken Grammatical Error Correction
Keywords: FOS: Computer and information sciences, Sound (cs.SD), Sound, Artificial Intelligence (cs.AI), Artificial Intelligence, Audio and Speech Processing (eess.AS), FOS: Electrical engineering, electronic engineering, information engineering, Computation and Language, Computation and Language (cs.CL), Audio and Speech Processing

descriptionPublicationkeyboard_double_arrow_right Article , Conference object , Preprint 22 Aug 2025Embargo end date: 01 Jan 2025Publisher:ISCAJournal:10th Workshop on Speech and Language Technology in Education (SLaTE)

Authors: Karanasou, Penny; Qian, Mengjie; Bannò, Stefano; Gales, Mark JF; Knill, Kate M;

doi: 10.21437/slate.2025-39 , 10.48550/arxiv.2507.19374 , 10.17863/cam.120264

arXiv: 2507.19374

Data Augmentation for Spoken Grammatical Error Correction

- Summary
- Subjects
- Metrics

Abstract

While there exist strong benchmark datasets for grammatical error correction (GEC), high-quality annotated spoken datasets for Spoken GEC (SGEC) are still under-resourced. In this paper, we propose a fully automated method to generate audio-text pairs with grammatical errors and disfluencies. Moreover, we propose a series of objective metrics that can be used to evaluate the generated data and choose the more suitable dataset for SGEC. The goal is to generate an augmented dataset that maintains the textual and acoustic characteristics of the original data while providing new types of errors. This augmented dataset should augment and enrich the original corpus without altering the language assessment scores of the second language (L2) learners. We evaluate the use of the augmented corpus both for written GEC (the text part) and for SGEC (the audio-text pairs). Our experiments are conducted on the S\&I Corpus, the first publicly available speech dataset with grammar error annotations.

This work has been accepted by ISCA SLaTE 2025

Related Organizations

University of Cambridge
United Kingdom

Keywords

FOS: Computer and information sciences, Sound (cs.SD), Sound, Artificial Intelligence (cs.AI), Artificial Intelligence, Audio and Speech Processing (eess.AS), FOS: Electrical engineering, electronic engineering, information engineering, Computation and Language, Computation and Language (cs.CL), Audio and Speech Processing

Impact byBIP!

	selected citations These citations are derived from selected sources. This is an alternative to the "Influence" indicator, which also reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically).	0
	popularity This indicator reflects the "current" impact/attention (the "hype") of an article in the research community at large, based on the underlying citation network.	Average
	influence This indicator reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically).	Average
	impulse This indicator reflects the initial momentum of an article directly after its publication, based on the underlying citation network.	Average

Found an issue? Give us feedback

Average

Green

Related to Research communities

Knowmad Institut