Back to Main Conference 2016
LREC 2016main

PARSEME Survey on MWE Resources

Proceedings of the Tenth International Conference on Language Resources and Evaluation (LREC 2016)

DOI:10.63317/5ap6oz6872yr

Abstract

This paper summarizes the preliminary results of an ongoing survey on multiword resources carried out within the IC1207 Cost Action PARSEME (PARSing and Multi-word Expressions). Despite the availability of language resource catalogs and the inventory of multiword datasets on the SIGLEX-MWE website, multiword resources are scattered and difficult to find. In many cases, language resources such as corpora, treebanks, or lexical databases include multiwords as part of their data or take them into account in their annotations. However, these resources need to be centralized to make them accessible. The aim of this survey is to create a portal where researchers can easily find multiword(-aware) language resources for their research. We report on the design of the survey and analyze the data gathered so far. We also discuss the problems we have detected upon examination of the data as well as possible ways of enhancing the survey.

Details

Paper ID
lrec2016-main-364
Pages
pp. 2299-2306
BibKey
losnegaard-etal-2016-parseme
Editor
N/A
Publisher
European Language Resources Association (ELRA)
ISSN
2522-2686
ISBN
978-2-9517408-9-1
Conference
Tenth International Conference on Language Resources and Evaluation
Location
Portorož, Slovenia
Date
23 May 2016 28 May 2016

Authors

  • GL

    Gyri Smørdal Losnegaard

  • FS

    Federico Sangati

  • CE

    Carla Parra Escartín

  • AS

    Agata Savary

  • SB

    Sascha Bargmann

  • JM

    Johanna Monti

Links