Australia's Whole-of-Government Microsoft 365 Copilot Trial
Australia
Australia's DTA ran the world's largest whole-of-government Copilot trial — 5,000+ staff, almost 60 agencies — with an independent Nous …
Australia · Canberra · See the Australia profile · See the Canberra profile
Evidence: Descriptive / self-reported Top 51% 60/100 · Ask Evidence Copilot about this practice
Services Australia and Capgemini built an OCR/NLP engine that classifies more than 25,000 citizen-submitted documents a day with over 95% accuracy, cutting processing from weeks or days to seconds during the COVID-19 claims surge.
Centrelink, Medicare and Child Support share a single document lodgement service in Australia's Services Australia. Caseworkers previously had to manually open and classify every citizen-uploaded file (e.g. medical certificates, bank statements) before an assessment could begin, and this manual process was overwhelmed by the 2020 COVID-19 claims surge.
From mid-2019, Services Australia worked with delivery partner Capgemini to automate document intake, aiming to classify and extract data from citizen-submitted files as part of the Document Management Modernisation through Intelligent Automation project.
The agency and Capgemini built an AI-based document management modernisation engine combining optical character recognition (OCR) and natural language processing (NLP) to automatically read, classify and extract data from uploaded documents.
By October 2020 the engine was processing more than 25,000 document lodgements a day with over 95% classification accuracy. Capgemini reported that the time from lodgement to caseworker-ready assessment fell from weeks or days to seconds. The agency said it was still working to fully quantify downstream outcome metrics, and no independently audited figures on error-correction rates, caseworker time saved, or claim-outcome accuracy have been separately published.
The strongest public evidence for this practice remains contemporaneous trade-press interviews with Capgemini and Services Australia rather than an independent evaluation; the rubric assessment (overall score 60/100) notes strong scalability (already handling full national volume) but weaker transparency and limited demonstrated transferability.
National / regional programmes
Indicative funding routes for practices of this type — always check each programme's current calls and eligibility rules.
Do you run this practice? Claim it — verified implementers get a public contact pathway and can propose corrections.
Where this practice's information was retrieved from, and when.
Australia
Australia's DTA ran the world's largest whole-of-government Copilot trial — 5,000+ staff, almost 60 agencies — with an independent Nous …
Netherlands
TNO's second annual census of generative-AI use across Dutch government found applications grew from 8 in 2024 to 81 in …
United States of America
Pennsylvania piloted ChatGPT Enterprise with 175 employees across 14 agencies, who saved an average of 95 minutes daily; the programme …
Denmark
Denmark's Agency for Digital Government built Børge, a Claude-based AI assistant that helps editors from roughly 40 public authorities rewrite …
Open full copilot Grounded in cited practices — always check the sources.