University of Sydney 'two-lane' assessment for the AI era
Australia
Instead of an unwinnable AI-detection arms race, Sydney redesigned assessment into two lanes — secure in-person assessment of core capability, …
China · Hefei · See the China profile
iFLYTEK's speech-recognition AI scores spoken English in China's zhongkao and gaokao oral exams across dozens of provinces, processing millions of students a year. Independent Rasch-model research supports scoring reliability, but transparency, appeals process and equity impacts
iFLYTEK (科大讯飞), a Hefei-based speech-AI company, supplies the automated "human-machine dialogue" (人机对话) scoring engine used in oral English components of China's zhongkao (senior-high entrance exam) and gaokao (national college entrance exam) in dozens of provinces. Students read aloud, answer questions or describe a picture into a headset; iFLYTEK's automatic speech recognition and scoring models grade pronunciation accuracy, fluency, completeness and content in real time, replacing or supplementing panels of human examiners for oral testing at a scale no manual process could match.
Deployment is large and long-running: Jiangsu piloted computer-delivered oral English in its zhongkao from around 2009 and now tests roughly 770,000 middle-school students across 3,200+ venues in 13 cities each round; Xiamen and Tongling (Anhui) run comparable exams; iFLYTEK states its systems now cover the gaokao oral component in 29 provinces and cities. China's Ministry of Education has recognised iFLYTEK's system as the only application judged feasible and trustworthy for organising large-scale online spoken-language exams.
Independent academic evidence on scoring validity is mixed but real: a 2016 multifaceted Rasch-model study found the automatic scoring reliable for junior-secondary zhongkao oral tests, while an earlier 2010 study had found some students achieving high AI-assigned scores despite actual speaking proficiency below the required standard. A 2025 PLoS ONE study benchmarking Chinese automated spoken-English scoring tools against trained human raters (n=30 students) found two of three tools achieved strong agreement (ICC 0.74–0.92, r 0.85–0.87) while one showed systematic score inflation — evidence that automated scoring in this space can be reliable, but is not uniformly validated across every vendor or year.
Transparency remains the weakest point in the public record: no independent audit of accent or dialect bias was found, the scoring algorithm and its weighting are proprietary, and no formal student-appeals process for AI-scored oral exams is documented. Company claims such as processing 20 million test-takers with zero complaints are self-reported marketing, not independently audited figures, and should be read with that caveat.
Read the full analysis: https://edu.iflytek.com/solution/examination/ai-language-test
Implementation detail (cost, timeline, staffing, conditions for success) is not yet available for this practice.
Where this practice's information was retrieved from, and when.
Australia
Instead of an unwinnable AI-detection arms race, Sydney redesigned assessment into two lanes — secure in-person assessment of core capability, …
United Kingdom
Free AI writing-feedback tool by Cambridge University Press & Assessment (Sept 2016); assigns CEFR A1–C2 levels to EFL essays and …
Hong Kong
In June 2023 HKU's Senate named GenAI a "fifth literacy" and gave every student and staff member free ChatGPT-class access. …
United Kingdom
Bedfordshire ran a published evaluation comparing its human-tutor Studiosity writing service with a new AI-augmented version. The 2024 AI trial …
Open full copilot Grounded in cited practices — always check the sources.