Enriching the 1758 Portuguese Parish Memories (Alentejo) with Named Entities

This work presents an enriched version of the Parish Memories (1758–1761), an essential Portuguese historical source manually transcribed. It is enriched with annotations of named entities of the types PERSON, LOCATION, and ORGANIZATION. The annotation was done automatically for the whole collection...

Full description

Bibliographic Details
Main Author: Vieira, Renata (author)
Other Authors: Olival, Fernanda (author), Cameron, Helena (author), Santos, Joaquim (author), Sequeira, Ofelia (author), Santos, Ivo (author)
Format: article
Language:eng
Published: 2021
Subjects:
Online Access:http://hdl.handle.net/10174/30166
Country:Portugal
Oai:oai:dspace.uevora.pt:10174/30166
Description
Summary:This work presents an enriched version of the Parish Memories (1758–1761), an essential Portuguese historical source manually transcribed. It is enriched with annotations of named entities of the types PERSON, LOCATION, and ORGANIZATION. The annotation was done automatically for the whole collection where two researchers annotated a portion of it manually for evaluation purposes. In this dataset, we provide the tagged texts, the lists of extracted entities, and frequency counts. The corpus is useful for historians, allowing, for instance, comparative analyses between parishes and regions or to calculate the area of influence of a locality. The paper describes the creation and evaluation of the corpus, discusses its applications and limitations. This first release may be improved by other researchers interested in the historical source itself or in the technology employed in its annotation.