evidoria

← Back to browse

Good practice Imported

Ceibal & ANEP's AI-Assisted Basic-Cycle Accreditation Exam

Uruguay · Montevideo · See the Uruguay profile · See the Montevideo profile

Evidence: Observational / pre–post Top 19% 67/100 · Ask Evidence Copilot about this practice

Uruguay's Ceibal and ANEP built an AI grader for the basic-cycle accreditation exam taken by roughly 8,000 mostly out-of-school young people and adults. It auto-scores objective sections, evaluates essays on 15 writing criteria, and acts as a neutral third reader when human corre

~8,000 candidates
Candidates enrolled for the AI-graded basic-cycle accreditation exam (13 June 2026)
Ceibal & ANEP's AI-Assisted Basic-Cycle Accreditation Exam

Details

Maturity
Pilot
Promoter
Ceibal / ANEP (Administración Nacional de Educación Pública)
Period
2024-2026 (development and pilot exam, June 2026)
Keywords
AI-assisted assessment, adult and youth education accreditation, automated writing evaluation

Context

Uruguay's Ceibal agency, working with the national education authority ANEP, built an AI grader for the country's basic-cycle accreditation exam — a route back into the credentialed system for young people and adults who left school early. Around 8,000 candidates were enrolled to sit the exam on 13 June 2026.

Activities

The exam has three parts: multiple-choice reading comprehension, multiple-choice mathematics, and a written essay scored against 15 specific criteria such as organisation, spelling and argumentation. Ceibal engineers spent more than a year building, testing and calibrating the system against prior years' human-scored results, using a contracted OpenAI model wrapped in Uruguayan-built software. On the writing section, the AI auto-scores where confident, flags a human corrector when unsure, and steps in as a neutral 'third reader' when two human correctors disagree; it can also be used by candidates as a self-study tutor before the exam.

Results

Ceibal reports that in testing, the AI's essay scores matched human correctors' assessments, and the system does not store student data or access ANEP's own student records, linking only through anonymised ID numbers.

Conclusions

This validation is described in press coverage as an internal testing-phase comparison rather than an independently published psychometric study, and the tool is so far deployed for a single exam track within a country whose Ceibal programme already gives 100% of public grade 1-9 students a personal device.

Implementation

Indicative cost
Medium (€50k–€500k)
Time to results
Medium (1–3 years)
Staffing & skills
Ceibal engineers, ANEP (Administración Nacional de Educación Pública)

Conditions for success

  • Built on Ceibal's existing national one-to-one device infrastructure (100% of grade 1-9 students)
  • More than a year of calibration against prior years' human-scored exam results before deployment
  • Privacy safeguards: no data retention, anonymised ID linkage rather than direct access to student records

Common failure modes

  • Validation reported only as an internal testing-phase comparison, not an independently published psychometric study

Commonly funded by

National / regional programmes

Indicative funding routes for practices of this type — always check each programme's current calls and eligibility rules.

Do you run this practice? Claim it — verified implementers get a public contact pathway and can propose corrections.

Data sources

Where this practice's information was retrieved from, and when.

Attachments

Similar practices you may find useful