Dutch National Archive uses AI to make its largest WWII archive digitally searchable
The Nationaal Archief, Netwerk Oorlogsbronnen, Huygens Instituut and NIOD Instituut are digitizing the Centraal Archief Bijzondere Rechtspleging (CABR), the largest and most-consulted Dutch WWII archive with dossiers on roughly 300,000 people, and using artificial intelligence to make the digitized documents readable and searchable. The first digital results go online via Oorlogvoorderechter.nl from 2025, with the project running until 2027.
Overview
The Nationaal Archief, Netwerk Oorlogsbronnen, Huygens Instituut and NIOD Instituut are digitizing the Centraal Archief Bijzondere Rechtspleging (CABR), the largest and most-consulted Dutch WWII archive with dossiers on roughly 300,000 people, and using artificial intelligence to make the digitized documents readable and searchable. The first digital results go online via Oorlogvoorderechter.nl from 2025, with the project running until 2027.
The challenge
The Centraal Archief Bijzondere Rechtspleging (CABR), the largest and most-consulted Dutch WWII archive with dossiers on roughly 300,000 people suspected of collaboration with the Germans, needed to be made digitally accessible to a broad public, with the archive spanning nearly four kilometers of documents including witness statements, NSB membership cards, diaries, pardon requests and photos.
The solution
The Nationaal Archief, Netwerk Oorlogsbronnen, Huygens Instituut and NIOD Instituut voor Oorlogs-, Holocaust- en Genocidestudies are digitizing the CABR through the project Oorlog voor de Rechter, using artificial intelligence to make the digitized archive documents readable and searchable. To help understand and contextualize the documents, they are enriched with background information and cross-references to related sources from other collections. Privacy (AVG/GDPR) rules are taken into account, meaning not all documents may be made available online.
Reported business value
The digitization is expected to average 152,000 scans per week. The first digital results will be presented via Oorlogvoorderechter.nl from 2025, with the project running until 2027. The ability to search across the CABR will give new insights into events of the Second World War from diverse perspectives.
Sources
Open any source and check the claim yourself — that is the point of the register.
This record was researched and written with AI assistance, and its claims were checked against the sources above. (EU AI Act art. 50 transparency notice.)
Other government & public sector entries in the register.
GovTech unlocks data insights to improve nationwide services with Databricks
Singapore's Government Technology Agency (GovTech) migrated from an on-premises dashboarding system to the Databricks Data + AI Platform on AWS with Unity Catalog and Delta Lake, cutting dashboard creation time from 90 to 30 days, democratizing data across 50% of corporate divisions in the first year, and saving 8,000 labor hours annually.
Estonia rolls out Bürokratt, an AI-guided virtual assistant network for public services
Bürokratt is a network of chatbots deployed on Estonian public sector institutions' websites, letting people obtain information from institutions and use public and information services via virtual assistants. It is a state-created, AI-based digital assistant that helps institutions deliver modern, efficient, around-the-clock customer service using large language models.
VA Advances Healthcare Insights With AI
The U.S. Department of Veterans Affairs leverages Databricks Data Intelligence to modernize healthcare analytics for millions of veterans, unifying massive distributed datasets into a single secure environment and streaming petabytes of health data in real time, reducing processes that once took hours to seconds. Databricks enables AI and large language models to detect risk early and enhance governance.
Austrian Academy of Sciences unlocks Ancient Greek with Mistral
The Austrian Academy of Sciences (OeAW), together with its Austrian Archaeological Institute, partnered with Mistral and services partner Reply to build Apollo, described as the first advanced large language model for Ancient Greek. Apollo is trained on a specialized corpus of 600 million words of historical Greek text plus tens of thousands of published inscriptions and papyri, helping researchers reconstruct damaged texts and identify thematic connections across collections. The OeAW reports Apollo turns work that once took years into hours, addressing over one million unread Greek papyri worldwide, with future phases planned for semantic search and handwritten inscription decipherment.
Was this helpful?
Your feedback helps us improve our use case database
