看见声音,保留自我:面向聋人中心的文本转语音的参与式设计方法
Seeing the Voice, Preserving the Self: A Participatory Design Approach to Deaf-Centric Text-to-Speech
浏览论文内容
中文总结 AI 辅助
本研究采用参与式设计方法,通过焦点小组和共同设计会议,探索聋人中心文本转语音技术,解决非听觉操控、意图验证和文化尊重问题,并识别采用障碍。
中文摘要 AI 辅助
我们描述了一种参与式设计方法,旨在开发聋人中心(Deaf-centric)的文本转语音(TTS)技术。虽然TTS在主流领域正在快速增长,但迄今为止,在聋人和重听(DHH)技术领域,它受到的关注很少。对于DHH用户而言,关键问题仍未得到解决,包括通过非听觉手段操纵音调、情感和表达方式的能力。在不听的情况下验证生成的语音是否符合意图并适合特定情境,是另一个挑战。在生成的语音中尊重文化和身份因素也很重要。这项工作通过两个焦点小组、三次共同设计会议和四次一对一的早期设计评估会议,与DHH参与者一起探索了设计空间。参与者包括熟悉和不熟悉TTS的人,以及DHH内容创作者。我们描述了关键发现、设计想法、结果以及对未来聋人中心TTS发展的启示。我们还识别了未满足的技术需求,这些需求构成了采用聋人中心TTS技术的障碍。
英文摘要
We describe a participatory design approach toward developing Deaf-centric text-to-speech (TTS) technologies. While TTS is growing rapidly in the mainstream, it has received little attention to date in the deaf and hard of hearing (DHH) technology space. Critical problems have remained unaddressed for DHH users, including the ability to manipulate tone, emotions and delivery via non-auditory means. Verifying that the generated speech matches intent and is appropriate for a given situation without having to listen to it is another challenge. Respecting cultural and identity factors in the generated speech is also important. This work explores the design space with DHH participants through two focus groups, three co-design sessions, and four one-on-one early-stage design evaluation sessions. Participants included people both familiar and unfamiliar with TTS, as well as DHH content creators. We describe key findings, design ideas, results, and implications for future Deaf-centric TTS development. We also identify unmet technology requirements that pose barriers to adoption of Deaf-centric TTS technology.
发表机构
- Gallaudet University(加劳德特大学)
- Stanford University(斯坦福大学)
机构由 AI 辅助整理,请以论文原文为准。