Polish Rhythmic Database ― New Resources for Speech Timing and Rhythm Analysis

Agnieszka Wagner, Katarzyna Klessa, Jolanta Bachan


Abstract
This paper reports on a new database ― Polish rhythmic database and tools developed with the aim of investigating timing phenomena and rhythmic structure in Polish including topics such as, inter alia, the effect of speaking style and tempo on timing patterns, phonotactic and phrasal properties of speech rhythm and stability of rhythm metrics. So far, 19 native and 12 non-native speakers with different first languages have been recorded. The collected speech data (5 h 14 min.) represents five different speaking styles and five different tempi. For the needs of speech corpus management, annotation and analysis, a database was developed and integrated with Annotation Pro (Klessa et al., 2013, Klessa, 2016). Currently, the database is the only resource for Polish which allows for a systematic study of a broad range of phenomena related to speech timing and rhythm. The paper also introduces new tools and methods developed to facilitate the database annotation and analysis with respect to various timing and rhythm measures. In the end, the results of an ongoing research and first experimental results using the new resources are reported and future work is sketched.
Anthology ID:
L16-1742
Volume:
Proceedings of the Tenth International Conference on Language Resources and Evaluation (LREC'16)
Month:
May
Year:
2016
Address:
Portorož, Slovenia
Editors:
Nicoletta Calzolari, Khalid Choukri, Thierry Declerck, Sara Goggi, Marko Grobelnik, Bente Maegaard, Joseph Mariani, Helene Mazo, Asuncion Moreno, Jan Odijk, Stelios Piperidis
Venue:
LREC
SIG:
Publisher:
European Language Resources Association (ELRA)
Note:
Pages:
4678–4683
Language:
URL:
https://aclanthology.org/L16-1742
DOI:
Bibkey:
Cite (ACL):
Agnieszka Wagner, Katarzyna Klessa, and Jolanta Bachan. 2016. Polish Rhythmic Database ― New Resources for Speech Timing and Rhythm Analysis. In Proceedings of the Tenth International Conference on Language Resources and Evaluation (LREC'16), pages 4678–4683, Portorož, Slovenia. European Language Resources Association (ELRA).
Cite (Informal):
Polish Rhythmic Database ― New Resources for Speech Timing and Rhythm Analysis (Wagner et al., LREC 2016)
Copy Citation:
PDF:
https://aclanthology.org/L16-1742.pdf