Back to DMR 2024
LREC-COLING 2024workshop

Expanding Russian PropBank: Challenges and Insights for Developing New SRL Resources

Proceedings of the Fifth International Workshop on Designing Meaning Representations @ LREC-COLING 2024

DOI:10.63317/2fpw4buboqgi

Abstract

Semantic role labeling (SRL) resources, such as Proposition Bank (PropBank), provide useful input to downstream applications. In this paper we present some challenges and insights we learned while expanding the previously developed Russian PropBank. This new effort involved annotation and adjudication of all predicates within a subset of the prior work in order to provide a test corpus for future applications. We discuss a number of new issues that arose while developing our PropBank for Russian as well as our solutions. Framing issues include: distinguishing between morphological processes that warrant new frames, differentiating between modal verbs and predicate verbs, and maintaining accurate representations of a given language’s semantics. Annotation issues include disagreements derived from variability in Universal Dependency parses and semantic ambiguity within the text. Finally, we demonstrate how Russian sentence structures reveal inherent limitations to PropBank’s ability to capture semantic data. These discussions should prove useful to anyone developing a PropBank or similar SRL resources for a new language.

Details

Paper ID
lrec2024-ws-dmr-04
Pages
pp. 30-38
BibKey
myers-etal-2024-expanding
Editor
N/A
Publisher
European Language Resources Association (ELRA) and ICCL
ISSN
N/A
ISBN
N/A
Workshop
Proceedings of the Fifth International Workshop on Designing Meaning Representations @ LREC-COLING 2024
Location
undefined, undefined
Date
20 May 2024 25 May 2024

Authors

  • SM

    Skatje Myers

  • RK

    Roman Khamov

  • AP

    Adam Pollins

  • RT

    Rebekah Tozier

  • OB

    Olga Babko-Malaya

  • MP

    Martha Palmer

Links