Pronunciation assessment in foreign language learning: Reliability and scoring bias in human–generative AI evaluation
This study examines the reliability and scoring bias of generative AI (Gen-AI)-based pronunciation assessment compared with human raters, addressing whether AI-generated scores can be trusted in real educational settings. Sixty students participated in a 12-w…




