Seoul Data Hub's Generative-AI Chatbot for Natural-Language Open-Data Discovery
South Korea
Seoul's Data Hub uses a retrieval-augmented generative-AI chatbot to let residents search 8,100+ open datasets in plain Korean; in its …
Brazil · Brasília · See the Brazil profile · See the Brasília profile
Top 50% 60/100 · Ask Evidence Copilot about this practice
Brazil's anti-corruption watchdog CGU and Uruguay's AGESIC piloted an open-source bridge letting citizens query government open data in plain language via LLMs — and found the real risk wasn't the technical link but the model inventing facts absent from the data.
The Controladoria-Geral da União (CGU), Brazil's federal body responsible for transparency and anti-corruption oversight, partnered with the Open Knowledge Foundation (OKFN) and Uruguay's e-government agency AGESIC to test whether large language models can be trusted to answer citizens' questions about official open data.
Using the Model Context Protocol (MCP), the team connected an LLM directly to CKAN, the open-source software that powers most of the world's public data portals, including Brazil's federal transparency portal. In Brazil, the pilot targeted one of the portal's most-requested datasets: congressional parliamentary amendments. In a parallel pilot, Uruguay connected the same architecture to its National Energy Balance dataset, covering energy imports, generation and installed capacity.
The technical connection worked; the harder problem was trust. Published in June 2026, the project's own write-up reports that answer quality depended less on the pipeline than on how well the underlying data was described — units, field definitions, valid assumptions. In testing on the Uruguayan dataset, the model fabricated 'climate factors' as part of its explanation even though no climate data existed anywhere in the source — a concrete, publicly documented hallucination that the team used to argue for mandatory domain-expert review of dataset descriptions before any citizen-facing deployment.
As of publication the work remains a two-dataset prototype; the next planned phase is user testing with real-world citizen questions rather than production rollout. The team describes the parallel Brazil/Uruguay design as a deliberate attempt to build a reproducible blueprint that any government running a CKAN-based open data portal could adopt.
Read the full analysis: https://blog.okfn.org/2026/06/10/mcp-for-open-data-portals-trust-depends-on-understanding-the-data/
Implementation detail (cost, timeline, staffing, conditions for success) is not yet available for this practice.
Do you run this practice? Claim it — verified implementers get a public contact pathway and can propose corrections.
Where this practice's information was retrieved from, and when.
South Korea
Seoul's Data Hub uses a retrieval-augmented generative-AI chatbot to let residents search 8,100+ open datasets in plain Korean; in its …
Singapore
Singapore's Department of Statistics built SANDRA, an AI chatbot letting the public query ~2,400 SingStat data tables from 70 agencies …
Brazil
Brazil's CGU launched Informa.BR on 30 June 2026: citizens type a plain-language information request and AI matches it against the …
France
Since 2021 France's VIGINUM service has tracked foreign disinformation networks with AI tools; in 2025 it open-sourced D3lta, an LLM-based …
Open full copilot Grounded in cited practices — always check the sources.