Untangling the Influence of Typology, Data, and Model Architecture on Ranking Transfer Languages for Cross-Lingual POS Tagging

Rice, Enora; Marashian, Ali; Haynie, Hannah; Wense, Katharina; Palmer, Alexis

doi:10.18653/v1/2025.lm4uc-1.4

Citation Details

This content will become publicly available on May 1, 2026

Untangling the Influence of Typology, Data, and Model Architecture on Ranking Transfer Languages for Cross-Lingual POS Tagging

Cross-lingual transfer learning is an invaluable tool for overcoming data scarcity, yet selecting a suitable transfer language remains a challenge. The precise roles of linguistic typology, training data, and model architecture in transfer language choice are not fully understood. We take a holistic approach, examining how both dataset-specific and fine-grained typological features influence transfer language selection for part-of-speech tagging, considering two different sources for morphosyntactic features. While previous work examines these dynamics in the context of bilingual biLSTMS, we extend our analysis to a more modern transfer learning pipeline: zero-shot prediction with pretrained multilingual models. We train a series of transfer language ranking systems and examine how different feature inputs influence ranker performance across architectures. Word overlap, type-token ratio, and genealogical distance emerge as top features across all architectures. Our findings reveal that a combination of typological and dataset-dependent features leads to the best rankings, and that good performance can be obtained with either feature group on its own. more »

Award ID(s):: 2149404

PAR ID:: 10651186

Author(s) / Creator(s):: Rice, Enora; Marashian, Ali; Haynie, Hannah; Wense, Katharina; Palmer, Alexis

Publisher / Repository:: Association for Computational Linguistics

Date Published:: 2025-05-01

Page Range / eLocation ID:: 22 to 31

Format(s):: Medium: X

Sponsoring Org:: National Science Foundation

Free Publicly Accessible Full Text
This content will become publicly available on May 1, 2026
Conference Paper:
https://doi.org/10.18653/v1/2025.lm4uc-1.4

More Like this