Skip to main navigation Skip to search Skip to main content

Too Young to NER: Improving Entity Recognition on Dutch Historical Documents

Research output: Contribution to conferencePaperScientificpeer-review

1 Citation (Scopus)

Abstract

Named entity recognition (NER) on historical texts is beneficial for the field of digital humanities, as it allows to easily search for the names of people, places and other entities in digitised archives. While the task of historical NER in different languages has been gaining popularity in recent years, Dutch historical NER remains an underexplored topic. Using a recently released historical dataset from the Dutch Language Institute, we train three BERT-based models and analyse the errors to identify main challenges. All three models outperform a contemporary multilingual baseline by a large margin on historical test data.
Original languageEnglish
Pages30-35
Number of pages6
Publication statusPublished - 01 May 2024

Fingerprint

Dive into the research topics of 'Too Young to NER: Improving Entity Recognition on Dutch Historical Documents'. Together they form a unique fingerprint.

Cite this