An algorithm for characterizing
context-governed speech production patterns
Name
torres-dctorres-meng-eecs-2023-thesis.pdf
Description
Thesis PDF
Size
9.34 MB
Format
Adobe PDF
Checksum (MD5)
f582353ec3b31bb0d934f8579b8d2c6a
Author(s)
Torres, Deborah Cheron
Advisor(s)
Shattuck-Hufnagel, Stefanie
Date Issued
June 2023
Publisher
Massachusetts Institute of Technology
Abstract
Speech recognition and analysis can be improved by using methods that can effectively characterize important speech patterns of a speaker without requiring hours of data. This thesis defines a method by which key contexts related to systematic speech modification can be used to create a profile of the speech produced by a speaker. Using acoustic and prosodic information, contexts that create the potential for speech modifications can be specified. Then, by filtering speech produced by a speaker in the targeted contexts, the patterns of speech production in these contexts can be characterized. With these productions, likely underlying contexts that are associated with the productions can be used to enhance speech recognition when these contexts arise in new speech.
MIT Department
Massachusetts Institute of Technology. Department of Electrical Engineering and Computer Science
Terms of Use
In Copyright - Educational Use Permitted
Copyright retained by author(s)
Persistent DSpace Link