Using text-to-speech in second language research and assessment: preliminary results from elicited imitation

Text-to-speech (TTS) is a technology widely used in second language learning and testing, for example in the case of learning material creation, pronunciation training, and speaking assessment. Such advances suggest that TTS may also hold great potential for automating performance-based measures of L2 knowledge and proficiency such as the Elicited Imitation (EI) test. However, questions remain about the impact of the technology on validity and fairness and consequently on how learners’ linguistic competence is interpreted. Both issues are crucial in the context of high-stakes language assessment but also for SLA research, where elicited performance is used to infer learners’ underlying interlanguage development. We show that EIs developed with TTS-synthesised stimuli demonstrated comparable construct validity as tests created with human-spoken stimuli and do not discriminate against test takers of varying proficiency or different first language backgrounds. EI scores were also not significantly different between human-spoken and TTS-synthesised items. These findings provide preliminary evidence for the use of TTS technology in language learning and language assessment contexts, with practical implications for researchers, educators, and test developers working in second language acquisition and bilingualism.

Authors

Institutions

Publication Details

Journal
Computer Assisted Language Learning
Published
2026-09-25
DOI
https://doi.org/10.1080/09588221.2026.2737122
Primary Topic
EFL/ESL Teaching and Learning
Type
article
Field-Weighted Citation Impact
0.00
Controls
|||
ALL TIME
JAN
FEB
MAR
APR
MAY
JUN
JUL
AUG
SEP
article

Using text-to-speech in second language research and assessment: preliminary results from elicited imitation

Kathy MinHye Kim, Xiaobin Chen, Sarah Löber
Computer Assisted Language Learning
EFL/ESL Teaching and Learning
article

Using text-to-speech in second language research and assessment: preliminary results from elicited imitation

Kathy MinHye Kim, Xiaobin Chen, Sarah Löber
article en

Abstract

Text-to-speech (TTS) is a technology widely used in second language learning and testing, for example in the case of learning material creation, pronunciation training, and speaking assessment. Such advances suggest that TTS may also hold great potential for automating performance-based measures of L2 knowledge and proficiency such as the Elicited Imitation (EI) test. However, questions remain about the impact of the technology on validity and fairness and consequently on how learners’ linguistic competence is interpreted. Both issues are crucial in the context of high-stakes language assessment but also for SLA research, where elicited performance is used to infer learners’ underlying interlanguage development. We show that EIs developed with TTS-synthesised stimuli demonstrated comparable construct validity as tests created with human-spoken stimuli and do not discriminate against test takers of varying proficiency or different first language backgrounds. EI scores were also not significantly different between human-spoken and TTS-synthesised items. These findings provide preliminary evidence for the use of TTS technology in language learning and language assessment contexts, with practical implications for researchers, educators, and test developers working in second language acquisition and bilingualism.

Computer Assisted Language Learning
Boston University (US), University of Tübingen (DE)
Reduced inequalities
Openalex Percentile: Top 2%
EFL/ESL Teaching and Learning
AI Navigator

Ask Laika to Summarize, Analyze, and Connect papers live on the map.

Summarize Papers & Methodologies

Extract key findings, datasets, and comparative methods across publications.

Benchmark Rankings & Visual Analytics

Rank top research institutions, authors, funders, topics, and journals by Field-Weighted Citation Impact (FWCI) and paper volume with instant charts.

Connect Distant Disciplines

Bridge topological clusters on the map to find hidden collaborative intersections.

Using text-to-speech in second language research and assessment: preliminary results from elicited imitation — Kathy MinHye Kim, Xiaobin Chen, et al. · Computer Assisted Language Learning (2026) | TGRS Research Map | TGRS