Bootstrapping UMR Annotations for Arapaho from Language Documentation Resources

Buchholz, Matthew J; Bonn, Julia; Post, Claire Benet; Cowell, Andrew; Palmer, Alexis

Citation Details

Uniform Meaning Representation (UMR) is a semantic labeling system in the AMR family designed to be uniformly applicable to typologically diverse languages. The UMR labeling system is quite thorough and can be time-consuming to execute, especially if annotators are starting from scratch. In this paper, we focus on methods for bootstrapping UMR annotations for a given language from existing resources, and specifically from typical products of language documentation work, such as lexical databases and interlinear glossed text (IGT). Using Arapaho as our test case, we present and evaluate a bootstrapping process that automatically generates UMR subgraphs from IGT. Additionally, we describe and evaluate a method for bootstrapping valency lexicon entries from lexical databases for both the target language and English. We are able to generate enough basic structure in UMR graphs from the existing Arapaho interlinearized texts to automate UMR labeling to a significant extent. Our method thus has the potential to streamline the process of building meaning representations for new languages without existing large-scale computational resources. more »

Award ID(s):: 2213805

PAR ID:: 10527722

Author(s) / Creator(s):: Buchholz, Matthew J; Bonn, Julia; Post, Claire Benet; Cowell, Andrew; Palmer, Alexis

Editor(s):: Calzolari, Nicoletta; Kan, Min-Yen; Hoste, Veronique; Lenci, Alessandro; Sakti, Sakriani; Xue, Nianwen

Publisher / Repository:: ELRA and ICCL

Date Published:: 2024-05-01

Format(s):: Medium: X

Location:: https://aclanthology.org/2024.lrec-main.220

Sponsoring Org:: National Science Foundation

Free Publicly Accessible Full Text
Accepted Manuscript1.0
Conference Paper:
The DOI is not currently available.

More Like this