Back to Main Conference 2008
LREC 2008main

JURISDIC: Polish Speech Database for Taking Dictation of Legal Texts

Proceedings of the Sixth International Conference on Language Resources and Evaluation (LREC 2008)

DOI:10.63317/2pteidimm9ta

Abstract

The paper provides an overview of the Polish Speech Database for taking dictation of legal texts, created for the purpose of LVCSR system for Polish. It presents background information about the design of the database and the requirements coming from its future uses. The applied method of the text corpora construction is presented as well as the database structure and recording scenarios. The most important details on the recording conditions and equipment are specified, followed by the description of the assessment methodology of recording quality, and the annotation specification and evaluation. Additionally, the paper contains current statistics from the database and the information about both the ongoing and planned stages of the database development process.

Details

Paper ID
lrec2008-main-499
Pages
N/A
BibKey
demenko-etal-2008-jurisdic
Editor
N/A
Publisher
European Language Resources Association (ELRA)
ISSN
2522-2686
ISBN
2-9517408-4-0
Conference
Sixth International Conference on Language Resources and Evaluation
Location
Marrakech, Morocco
Date
28 May 2008 30 May 2008

Authors

  • GD

    Grazyna Demenko

  • SG

    Stefan Grocholewski

  • KK

    Katarzyna Klessa

  • JO

    Jerzy Ogórkiewicz

  • AW

    Agnieszka Wagner

  • ML

    Marek Lange

  • Daniel Śledziński

  • NC

    Natalia Cylwik

Links