Cornell University
Library
Cornell UniversityLibrary

eCommons

Help
Log In(current)
  1. Home
  2. College of Arts and Sciences
  3. Linguistics
  4. Linguistics - Monographs, Papers and Research
  5. Harvesting speech datasets for linguistic research on the web

Harvesting speech datasets for linguistic research on the web

File(s)
rooth-howell-wagner.final.WP.pdf (2.6 MB)
Main article
Permanent Link(s)
https://hdl.handle.net/1813/34477
Collections
Linguistics - Monographs, Papers and Research
Author
Rooth, Mats
Howell, Jonathan
Wagner, Michael
Abstract

This is a white paper for a project that harvested audio and transcribed data from podcasts and news broadcasts on the web. Tools were developed to analyze the different uses of prosody (rhythm, stress and intonation) within spoken communication using phonetic analysis and machine learning.

Sponsorship
NSF 1035151 RAPID: Harvesting Speech Datasets for Linguistic Research on the Web (Digging into Data Challenge) and SSHRC Digging into Data Challenge Grant 869-2009-0004.
Date Issued
2013-10-29
Keywords
prosody
•
intonation
•
comparatives
•
machine learning
•
spoken language
•
web science
Previously Published as
Final project white paper, Digging into Data Challenge
Type
article

Site Statistics | Help

About eCommons | Policies | Terms of use | Contact Us

copyright © 2002-2026 Cornell University Library | Privacy | Web Accessibility Assistance