ProtoSound: A Personalized and Scalable Sound Recognition System for Deaf and Hard-of-Hearing Users

被引:13
|
作者
Jain, Dhruv [1 ,2 ]
Nguyen, Khoa Huynh Anh [1 ]
Goodman, Steven [1 ]
Grossman-Kahn, Rachel [1 ]
Ngo, Hung [1 ]
Kusupati, Aditya [1 ]
Du, Ruofei [3 ]
Olwal, Alex [4 ]
Findlater, Leah [1 ]
Froehlich, Jon E. [1 ]
机构
[1] Univ Washington, Seattle, WA 98195 USA
[2] Google, Mountain View, CA 94043 USA
[3] Google Res, San Francisco, CA USA
[4] Google Res, Mountain View, CA USA
来源
PROCEEDINGS OF THE 2022 CHI CONFERENCE ON HUMAN FACTORS IN COMPUTING SYSTEMS (CHI' 22) | 2022年
关键词
Accessibility; deaf; Deaf; hard of hearing; sound awareness; sound recognition; CLASSIFICATION; EVENTS;
D O I
10.1145/3491102.3502020
中图分类号
TP [自动化技术、计算机技术];
学科分类号
0812 ;
摘要
Recent advances have enabled automatic sound recognition systems for deaf and hard of hearing (DHH) users on mobile devices. However, these tools use pre-trained, generic sound recognition models, which do not meet the diverse needs of DHH users. We introduce ProtoSound, an interactive system for customizing sound recognition models by recording a few examples, thereby enabling personalized and fine-grained categories. ProtoSound is motivated by prior work examining sound awareness needs of DHH people and by a survey we conducted with 472 DHH participants. To evaluate ProtoSound, we characterized performance on two real-world sound datasets, showing significant improvement over state-of-the-art (e.g., +9.7% accuracy on the first dataset). We then deployed ProtoSound's end-user training and real-time recognition through a mobile application and recruited 19 hearing participants who listened to the real-world sounds and rated the accuracy across 56 locations (e.g., homes, restaurants, parks). Results show that ProtoSound personalized the model on-device in real-time and accurately learned sounds across diverse acoustic contexts. We close by discussing open challenges in personalizable sound recognition, including the need for better recording interfaces and algorithmic improvements.
引用
收藏
页数:16
相关论文
共 50 条
  • [1] A Personalizable Mobile Sound Detector App Design for Deaf and Hard-of-Hearing Users
    Bragg, Danielle
    Huynh, Nicholas
    Ladner, Richard E.
    ASSETS'16: PROCEEDINGS OF THE 18TH INTERNATIONAL ACM SIGACCESS CONFERENCE ON COMPUTERS AND ACCESSIBILITY, 2016, : 3 - 13
  • [2] "Not There Yet": Feasibility and Challenges of Mobile Sound Recognition to Support Deaf and Hard-of-Hearing People
    Huang, Jeremy Zhengqi
    Chhabria, Hriday
    Jain, Dhruv
    PROCEEDINGS OF THE 25TH INTERNATIONAL ACM SIGACCESS CONFERENCE ON COMPUTERS AND ACCESSIBILITY, ASSETS 2023, 2023,
  • [3] AdaptiveSound: An Interactive Feedback-Loop System to Improve Sound Recognition for Deaf and Hard of Hearing Users
    Do, Hang
    Dang, Quan
    Huang, Jeremy Zhengqi
    Jain, Dhruv
    PROCEEDINGS OF THE 25TH INTERNATIONAL ACM SIGACCESS CONFERENCE ON COMPUTERS AND ACCESSIBILITY, ASSETS 2023, 2023,
  • [4] Deaf and Hard-of-hearing Individuals' Preferences for Wearable and Mobile Sound Awareness Technologies
    Findlater, Leah
    Chinh, Bonnie
    Jain, Dhruv
    Froehlich, Jon
    Kushalnagar, Raja
    Lin, Angela Carey
    CHI 2019: PROCEEDINGS OF THE 2019 CHI CONFERENCE ON HUMAN FACTORS IN COMPUTING SYSTEMS, 2019,
  • [5] A Personalized Captioning Strategy for the Deaf and Hard-of-Hearing Users in an Augmented Reality Environment
    Shidende, Deogratias
    Kessel, Thomas
    Treydte, Anna
    Moebs, Sabine
    EXTENDED REALITY, PT II, XR SALENTO 2024, 2024, 15028 : 3 - 21
  • [6] SoundVizVR: Sound Indicators for Accessible Sounds in Virtual Reality for Deaf or Hard-of-Hearing Users
    Li, Ziming
    Connell, Shannon
    Dannels, Wendy
    Peiris, Roshan
    PROCEEDINGS OF THE 24TH INTERNATIONAL ACM SIGACCESS CONFERENCE ON COMPUTERS AND ACCESSIBILITY, ASSETS 2022, 2022,
  • [7] Supporting Deaf and Hard-of-Hearing Students in the Schools
    Elizabeth M. Gibbons
    Contemporary School Psychology, 2015, 19 (1) : 46 - 53
  • [8] HoloSound: Combining Speech and Sound Identification for Deaf or Hard of Hearing Users on a Head-mounted Display
    Guo, Ru
    Yang, Yiru
    Kuang, Johnson
    Bin, Xue
    Jain, Dhruv
    Goodman, Steven
    Findlater, Leah
    Froehlich, Jon E.
    22ND INTERNATIONAL ACM SIGACCESS CONFERENCE ON COMPUTERS AND ACCESSIBILITY (ASSETS '20), 2020,
  • [9] Evaluating Smartwatch-based Sound Feedback for Deaf and Hard-of-hearing Users Across Contexts
    Goodman, Steven
    Kirchner, Susanne
    Guttman, Rose
    Jain, Dhruv
    Froehlich, Jon
    Findlater, Leah
    PROCEEDINGS OF THE 2020 CHI CONFERENCE ON HUMAN FACTORS IN COMPUTING SYSTEMS (CHI'20), 2020,
  • [10] Deaf and Hard-of-Hearing Users' Preferences for Hearing Speakers' Behavior during Technology-Mediated In-Person and Remote Conversations
    Seita, Matthew
    Andrew, Sarah
    Huenerfauth, Matt
    18TH INTERNATIONAL WEB FOR ALL CONFERENCE (W4A '21), 2021,