Automatic Alt-text: Computer-generated Image Descriptions for Blind Users on a Social Network Service

被引:176
作者
Wu, Shaomei [1 ]
Wieland, Jeffrey [1 ]
Farivar, Omid [1 ]
Schiller, Julie [1 ]
机构
[1] Facebook, Menlo Pk, CA 94025 USA
来源
CSCW'17: PROCEEDINGS OF THE 2017 ACM CONFERENCE ON COMPUTER SUPPORTED COOPERATIVE WORK AND SOCIAL COMPUTING | 2017年
关键词
User Experience; Accessibility; Artificial Intelligence; Social Networking Sites; Facebook; WEB; ACCESSIBILITY;
D O I
10.1145/2998181.2998364
中图分类号
TP39 [计算机的应用];
学科分类号
081203 ; 0835 ;
摘要
We designed and deployed automatic alt-text (AAT), a system that applies computer vision technology to identify faces, objects, and themes from photos to generate photo alt-text for screen reader users on Facebook. We designed our system through iterations of prototyping and in-lab user studies. Our lab test participants had a positive reaction to our system and an enhanced experience with Facebook photos. We also evaluated our system through a two-week field study as part of the Facebook iOS app for 9K VoiceOver users. We randomly assigned them into control and test groups and collected two weeks of activity data and their survey feedback.-The test group reported that photos on Facebook were easier to interpret and more engaging, and found Facebook more useful in general. Our system demonstrates that artificial intelligence can be used to enhance the experience for visually impaired users on social networking sites (SNSs), while also revealing the challenges with designing automated assistive technology in a SNS context.
引用
收藏
页码:1180 / 1192
页数:13
相关论文
共 30 条
[1]  
[Anonymous], Cbsnews
[2]  
[Anonymous], 2014, P 2 INT C LEARN REPR
[3]  
Bigham Jeffrey P., P 23 ANN ACM S US IN, P333
[4]   Crowdsourcing accessibility: Human-Powered access technologies [J].
Brady, Erin ;
Bigham, Jeffrey P. .
Foundations and Trends in Human-Computer Interaction, 2014, 8 (04) :273-372
[5]  
Brady E., 2015, SIGACCESS Accessibility and Computing, P16, DOI [10.1145/2809915.2809918, DOI 10.1145/2809915.2809918]
[6]  
Brady Erin, P SIGCHI C HUM FACT, P2117
[7]  
Brady Erin, 2015, P 33 ANN ACM C HUM F
[8]  
Brady ErinL., 2013, Proceedings of the 2013 Conference on Computer Supported Cooperative Work, CSCW '13, P1225, DOI [10.1145/2441776.2441915, DOI 10.1145/2441776.2441915]
[9]  
Buzzi MC, 2010, INT SYMP TECHNOL SOC, P327, DOI 10.1109/ISTAS.2010.5514621
[10]   Every Picture Tells a Story: Generating Sentences from Images [J].
Farhadi, Ali ;
Hejrati, Mohsen ;
Sadeghi, Mohammad Amin ;
Young, Peter ;
Rashtchian, Cyrus ;
Hockenmaier, Julia ;
Forsyth, David .
COMPUTER VISION-ECCV 2010, PT IV, 2010, 6314 :15-+