Sketching With Your Voice: "Non-Phonorealistic" Rendering of Sounds via Vocal Imitation
Name
3680528.3687679.pdf
Size
3.27 MB
Format
Adobe PDF
Checksum (MD5)
32e8f480ed40fc1d9a6e0bc7d815a6bc
Author(s) • • • •
Caren, Matthew
Chandra, Kartik
Tenenbaum, Joshua
Ragan-Kelley, Jonathan
Ma, Karima
Date Issued
December 3, 2024
Publisher
ACM|SIGGRAPH Asia 2024 Conference Papers
Citation
Caren, Matthew, Chandra, Kartik, Tenenbaum, Joshua, Ragan-Kelley, Jonathan and Ma, Karima. 2024. "Sketching With Your Voice: "Non-Phonorealistic" Rendering of Sounds via Vocal Imitation."
Version
Final published version
Abstract
We present a method for automatically producing human-like vocal imitations of sounds: the equivalent of “sketching,” but for auditory rather than visual representation. Starting with a simulated model of the human vocal tract, we first try generating vocal imitations by tuning the model’s control parameters to make the synthesized vocalization match the target sound in terms of perceptually-salient auditory features. Then, to better match human intuitions, we apply a cognitive theory of communication to take into account how human speakers reason strategically about their listeners. Finally, we show through several experiments and user studies that when we add this type of communicative reasoning to our method, it aligns with human intuitions better than matching auditory features alone does. This observation has broad implications for the study of depiction in computer graphics.
Description
SA Conference Papers ’24, December 03–06, 2024, Tokyo, Japan
MIT Department
Massachusetts Institute of Technology. Department of Electrical Engineering and Computer Science
Massachusetts Institute of Technology. Department of Brain and Cognitive Sciences
Terms of Use
Creative Commons Attribution
Persistent DSpace Link
DOI of Published Version
https://doi.org/10.1145/3680528.3687679