We introduce the Transmasculine Attitudes and Speech Corpus (TMASC), a multimodal corpus of 196 transmasculine individuals, including questionnaire responses and 66 audio recordings. The questionnaire includes items exploring the vocal health of transmasculine individuals. The audio recordings include cough and throat-clearing samples, a reading passage, and additional session-specific questions. This paper outlines the development of this corpus and the data collection procedures. To illustrate the utility of this corpus, we present three case studies demonstrating how this crowd-sourced multimodal corpus can be used to support transmasculine individuals. These include the integration of perceptual and acoustic data, the identification of group-level characteristics, and the calibration of acoustic measurements.
翻译:我们介绍了跨男性态度与语音语料库(TMASC),这是一个包含196名跨男性个体的多模态语料库,涵盖问卷应答及66段音频录音。问卷部分包含探究跨男性个体嗓音健康状态的条目。音频录音包括咳嗽和清嗓样本、一段朗读文本以及额外的阶段性特定问题。本文阐述了该语料库的开发过程与数据采集流程。为展示该语料库的实用性,我们通过三个案例研究验证了这一众包多模态语料库在支持跨男性群体方面的应用——包括感知数据与声学数据的整合、群体层面特征的识别,以及声学测量的校准。