🤖 AI Summary
This study addresses the critical lack of multimodal voice and attitudinal data resources focused on transgender men, which has hindered in-depth investigation into their vocal health and speech characteristics. To bridge this gap, we present the first publicly available Transgender Male Audio–Survey Corpus (TMASC), comprising data from 196 participants and integrating questionnaire responses with 66 multimodal audio recordings across tasks such as coughing, throat-clearing, reading, and conversational问答. The corpus was constructed through a combination of crowdsourced data collection, acoustic analysis, and perceptual evaluation to enable robust multimodal fusion. Three case studies demonstrate TMASC’s utility in identifying group-specific vocal traits, calibrating acoustic measures, and supporting voice health interventions, thereby filling a significant void in existing data resources for this population.
📝 Abstract
We introduce the Transmasculine Attitudes and Speech Corpus (TMASC), a multimodal corpus of 196 transmasculine individuals, including questionnaire responses and 66 audio recordings. The questionnaire includes items exploring the vocal health of transmasculine individuals. The audio recordings include cough and throat-clearing samples, a reading passage, and additional session-specific questions. This paper outlines the development of this corpus and the data collection procedures. To illustrate the utility of this corpus, we present three case studies demonstrating how this crowd-sourced multimodal corpus can be used to support transmasculine individuals. These include the integration of perceptual and acoustic data, the identification of group-level characteristics, and the calibration of acoustic measurements.