Cornell University
Library
Cornell UniversityLibrary

eCommons

Help
Log In(current)
  1. Home
  2. College of Arts and Sciences
  3. Linguistics
  4. Linguistics - Monographs, Papers and Research
  5. ML classification in a benchmark of prosodic minimal pairs

ML classification in a benchmark of prosodic minimal pairs

File(s)
lagisetty-2026-classification-poster.pdf (446.46 KB)
lagisetty-2026-classification.pdf (233.85 KB)
Permanent Link(s)
https://hdl.handle.net/1813/124941
Collections
Linguistics - Monographs, Papers and Research
Author
Lagisetty, Sahya
Rooth, Mats
Abstract

This paper presents the following: (i) A computational methodology for collecting large numbers of utterances of a fixed word string from online sources, using the index of youtube transcriptions at filmot.com. (ii) A prototype benchmark constructed with the methodology consisting of prosodic minimal pairs, which are short word sequences which depending on context and/or lexical identity are pronounced with different prosodies. The benchmark is grouped into pairs, for instance utterances of "much as I did'" with or without focus prosody on the first person subject. (iii) A classification model obtained by tuning wav2vec2-base on the training portion of the benchmark. Presented with an audio that is stipulated to be an utterance of a given word string, the model selects one of two alternative prosodies for the word string. For instance, it can determine whether a given utterance of "much as I did" has focus prosody on the subject, or default prosody. Accuracy of this determination on separate test data is better than 90% for each minimal pair.

Description
Poster and prepublication version of paper from Speech Prosody 2026, May 26, 2026.
Date Issued
2026-05-26
Publisher
International Speech Communication Association
Keywords
prosody, audio classification, alternative focus, stress doublet
Related DOI
10.21437/SpeechProsody.2026-92
Previously Published as
Lagisetty, Sahya, and Mats Rooth. 2026. ML Classification in a Benchmark of Prosodic Minimal Pairs. In Proceedings of Speech Prosody 2026, 453–457. Baixas, France: International Speech Communication Association (ISCA). https://doi.org/10.21437/SpeechProsody.2026-92.
ISSN
2333-2042
Type
conference papers and proceedings

Site Statistics | Help

About eCommons | Policies | Terms of use | Contact Us

copyright © 2002-2026 Cornell University Library | Privacy | Web Accessibility Assistance