evidoria

← Back to browse

Good practice

iFLYTEK AI Speech Scoring for China's School Oral English Exams

China · Hefei · See the China profile

iFLYTEK's speech-recognition AI scores spoken English in China's zhongkao and gaokao oral exams across dozens of provinces, processing millions of students a year. Independent Rasch-model research supports scoring reliability, but transparency, appeals process and equity impacts

iFLYTEK AI Speech Scoring for China's School Oral English Exams

Details

Promoter
iFLYTEK Co., Ltd. (科大讯飞)
Period
2009-present
Keywords
speech recognition, automated scoring, oral English assessment, exam technology

Description

iFLYTEK (科大讯飞), a Hefei-based speech-AI company, supplies the automated "human-machine dialogue" (人机对话) scoring engine used in oral English components of China's zhongkao (senior-high entrance exam) and gaokao (national college entrance exam) in dozens of provinces. Students read aloud, answer questions or describe a picture into a headset; iFLYTEK's automatic speech recognition and scoring models grade pronunciation accuracy, fluency, completeness and content in real time, replacing or supplementing panels of human examiners for oral testing at a scale no manual process could match.

Deployment is large and long-running: Jiangsu piloted computer-delivered oral English in its zhongkao from around 2009 and now tests roughly 770,000 middle-school students across 3,200+ venues in 13 cities each round; Xiamen and Tongling (Anhui) run comparable exams; iFLYTEK states its systems now cover the gaokao oral component in 29 provinces and cities. China's Ministry of Education has recognised iFLYTEK's system as the only application judged feasible and trustworthy for organising large-scale online spoken-language exams.

Independent academic evidence on scoring validity is mixed but real: a 2016 multifaceted Rasch-model study found the automatic scoring reliable for junior-secondary zhongkao oral tests, while an earlier 2010 study had found some students achieving high AI-assigned scores despite actual speaking proficiency below the required standard. A 2025 PLoS ONE study benchmarking Chinese automated spoken-English scoring tools against trained human raters (n=30 students) found two of three tools achieved strong agreement (ICC 0.74–0.92, r 0.85–0.87) while one showed systematic score inflation — evidence that automated scoring in this space can be reliable, but is not uniformly validated across every vendor or year.

Transparency remains the weakest point in the public record: no independent audit of accent or dialect bias was found, the scoring algorithm and its weighting are proprietary, and no formal student-appeals process for AI-scored oral exams is documented. Company claims such as processing 20 million test-takers with zero complaints are self-reported marketing, not independently audited figures, and should be read with that caveat.

Read the full analysis: https://edu.iflytek.com/solution/examination/ai-language-test

Implementation

Implementation detail (cost, timeline, staffing, conditions for success) is not yet available for this practice.

Data sources

Where this practice's information was retrieved from, and when.

Attachments

Similar practices you may find useful