ProtoSound: A Personalized and Scalable Sound Recognition System for   Deaf and Hard-of-Hearing Users

Dhruv Jain; Khoa Huynh Anh Nguyen; Steven Goodman; Rachel; Grossman-Kahn; Hung Ngo; Aditya Kusupati; Ruofei Du; Alex Olwal; Leah; Findlater; Jon E. Froehlich

arXiv:2202.11134·cs.HC·February 24, 2022

ProtoSound: A Personalized and Scalable Sound Recognition System for Deaf and Hard-of-Hearing Users

Dhruv Jain, Khoa Huynh Anh Nguyen, Steven Goodman, Rachel, Grossman-Kahn, Hung Ngo, Aditya Kusupati, Ruofei Du, Alex Olwal, Leah, Findlater, Jon E. Froehlich

PDF

TL;DR

ProtoSound is a personalized, scalable sound recognition system for deaf and hard-of-hearing users that allows on-device customization with few examples, significantly improving recognition accuracy in diverse real-world environments.

Contribution

We introduce ProtoSound, a novel system enabling real-time, on-device sound model personalization through user recordings, tailored for DHH users' diverse needs.

Findings

01

+9.7% accuracy on real-world datasets

02

Effective real-time on-device model personalization

03

High user satisfaction in diverse acoustic environments

Abstract

Recent advances have enabled automatic sound recognition systems for deaf and hard of hearing (DHH) users on mobile devices. However, these tools use pre-trained, generic sound recognition models, which do not meet the diverse needs of DHH users. We introduce ProtoSound, an interactive system for customizing sound recognition models by recording a few examples, thereby enabling personalized and fine-grained categories. ProtoSound is motivated by prior work examining sound awareness needs of DHH people and by a survey we conducted with 472 DHH participants. To evaluate ProtoSound, we characterized performance on two real-world sound datasets, showing significant improvement over state-of-the-art (e.g., +9.7% accuracy on the first dataset). We then deployed ProtoSound's end-user training and real-time recognition through a mobile application and recruited 19 hearing participants who…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.