ProtoSound: A Personalized and Scalable Sound Recognition System for Deaf and Hard-of-Hearing Users

被引：13

作者：

Jain, Dhruv ^{[1
,2
]}

Nguyen, Khoa Huynh Anh ^{[1
]}

Goodman, Steven ^{[1
]}

Grossman-Kahn, Rachel ^{[1
]}

Ngo, Hung ^{[1
]}

Kusupati, Aditya ^{[1
]}

Du, Ruofei ^{[3
]}

Olwal, Alex ^{[4
]}

Findlater, Leah ^{[1
]}

Froehlich, Jon E. ^{[1
]}

机构：

[1] Univ Washington, Seattle, WA 98195 USA

[2] Google, Mountain View, CA 94043 USA

[3] Google Res, San Francisco, CA USA

[4] Google Res, Mountain View, CA USA

来源：

PROCEEDINGS OF THE 2022 CHI CONFERENCE ON HUMAN FACTORS IN COMPUTING SYSTEMS (CHI' 22) | 2022年

关键词：

Accessibility; deaf; Deaf; hard of hearing; sound awareness; sound recognition; CLASSIFICATION; EVENTS;

D O I：

10.1145/3491102.3502020

中图分类号：

TP [自动化技术、计算机技术];

学科分类号：

0812 ;

摘要：

Recent advances have enabled automatic sound recognition systems for deaf and hard of hearing (DHH) users on mobile devices. However, these tools use pre-trained, generic sound recognition models, which do not meet the diverse needs of DHH users. We introduce ProtoSound, an interactive system for customizing sound recognition models by recording a few examples, thereby enabling personalized and fine-grained categories. ProtoSound is motivated by prior work examining sound awareness needs of DHH people and by a survey we conducted with 472 DHH participants. To evaluate ProtoSound, we characterized performance on two real-world sound datasets, showing significant improvement over state-of-the-art (e.g., +9.7% accuracy on the first dataset). We then deployed ProtoSound's end-user training and real-time recognition through a mobile application and recruited 19 hearing participants who listened to the real-world sounds and rated the accuracy across 56 locations (e.g., homes, restaurants, parks). Results show that ProtoSound personalized the model on-device in real-time and accurately learned sounds across diverse acoustic contexts. We close by discussing open challenges in personalizable sound recognition, including the need for better recording interfaces and algorithmic improvements.

引用

页数：16

共 50 条

[1] An Analysis of Personalized Speech Recognition System Development for the Deaf and Hard-of-Hearing
Violeta, Lester Phillip
Toda, Tomoki
2023 Asia Pacific Signal and Information Processing Association Annual Summit and Conference, APSIPA ASC 2023, 2023, : 1862 - 1867
[2] An Analysis of Personalized Speech Recognition System Development for the Deaf and Hard-of-Hearing
Violeta, Lester Phillip
Toda, Tomoki
2023 ASIA PACIFIC SIGNAL AND INFORMATION PROCESSING ASSOCIATION ANNUAL SUMMIT AND CONFERENCE, APSIPA ASC, 2023, : 1862 - 1867
[3] A Personalized Captioning Strategy for the Deaf and Hard-of-Hearing Users in an Augmented Reality Environment
Shidende, Deogratias
Kessel, Thomas
Treydte, Anna
Moebs, Sabine
EXTENDED REALITY, PT II, XR SALENTO 2024, 2024, 15028 : 3 - 21
[4] A Personalizable Mobile Sound Detector App Design for Deaf and Hard-of-Hearing Users
Bragg, Danielle
Huynh, Nicholas
Ladner, Richard E.
ASSETS'16: PROCEEDINGS OF THE 18TH INTERNATIONAL ACM SIGACCESS CONFERENCE ON COMPUTERS AND ACCESSIBILITY, 2016, : 3 - 13
[5] PROFILING DEAF AND HARD-OF-HEARING USERS OF SUBTITLES FOR THE DEAF AND HARD-OF-HEARING IN ITALY: A QUESTIONNAIRE-BASED STUDY
Morettini, Agnese
MONTI, 2012, 4 : 321 - 348
[6] SoundVizVR: Sound Indicators for Accessible Sounds in Virtual Reality for Deaf or Hard-of-Hearing Users
Li, Ziming
Connell, Shannon
Dannels, Wendy
Peiris, Roshan
PROCEEDINGS OF THE 24TH INTERNATIONAL ACM SIGACCESS CONFERENCE ON COMPUTERS AND ACCESSIBILITY, ASSETS 2022, 2022,
[7] Designing a Multi-Modal Communication System for the Deaf and Hard-of-Hearing Users
Lee, Gi-bbeum
Jang, Hyuckjin
Jeong, Hyundeok
Woo, Woontack
2021 IEEE INTERNATIONAL SYMPOSIUM ON MIXED AND AUGMENTED REALITY ADJUNCT PROCEEDINGS (ISMAR-ADJUNCT 2021), 2021, : 429 - 434
[8] "Not There Yet": Feasibility and Challenges of Mobile Sound Recognition to Support Deaf and Hard-of-Hearing People
Huang, Jeremy Zhengqi
Chhabria, Hriday
Jain, Dhruv
PROCEEDINGS OF THE 25TH INTERNATIONAL ACM SIGACCESS CONFERENCE ON COMPUTERS AND ACCESSIBILITY, ASSETS 2023, 2023,
[9] Evaluating Smartwatch-based Sound Feedback for Deaf and Hard-of-hearing Users Across Contexts
Goodman, Steven
Kirchner, Susanne
Guttman, Rose
Jain, Dhruv
Froehlich, Jon
Findlater, Leah
PROCEEDINGS OF THE 2020 CHI CONFERENCE ON HUMAN FACTORS IN COMPUTING SYSTEMS (CHI'20), 2020,
[10] Redesigning and Deploying the Universal Sound Detector: Notifying Deaf and Hard-of-hearing Users of audio signals
Stanislow, Joseph S.
Behm, Gary W.
ASSETS'18: PROCEEDINGS OF THE 20TH INTERNATIONAL ACM SIGACCESS CONFERENCE ON COMPUTERS AND ACCESSIBILITY, 2018, : 373 - 375

← 1 2 3 4 5 →