TY - GEN
T1 - Capti-Speak
T2 - 12th International Web for All Conference, W4A 2015
AU - Ashok, Vikas
AU - Borodin, Yevgen
AU - Puzis, Yury
AU - Ramakrishnan, I. V.
N1 - Publisher Copyright:
Copyright 2015 ACM.
PY - 2015/5/18
Y1 - 2015/5/18
N2 - People with vision impairments interact with web pages via screen readers that provide keyboard shortcuts for navigating through the content. However, web browsing with screen readers can be a frustrating experience mainly due to time and effort spent on locating the desired content through the extensive use of keyboard shortcuts. This gets even worse if users have limited shortcut vocabulary or are not familiar with the structure of a particular webpage. Augmenting screen readers with a speech input interface has the potential to alleviate the above limitations. This paper describes the design, implementation, and evaluation of Capti-Speak, a speech-enabled screen reader for web browsing, capable of translating speech utterances into browsing actions, executing the actions, and providing audio feedback. The novelty of Capti-Speak is that it leverages a custom dialog model, designed exclusively for non-visual web access, for interpreting speech utterances. A user study with 20 blind subjects showed that Capti-Speak was significantly more usable and efficient compared to the regular screen reader, especially for ad-hoc browsing, searching, and navigating to the content of interest.
AB - People with vision impairments interact with web pages via screen readers that provide keyboard shortcuts for navigating through the content. However, web browsing with screen readers can be a frustrating experience mainly due to time and effort spent on locating the desired content through the extensive use of keyboard shortcuts. This gets even worse if users have limited shortcut vocabulary or are not familiar with the structure of a particular webpage. Augmenting screen readers with a speech input interface has the potential to alleviate the above limitations. This paper describes the design, implementation, and evaluation of Capti-Speak, a speech-enabled screen reader for web browsing, capable of translating speech utterances into browsing actions, executing the actions, and providing audio feedback. The novelty of Capti-Speak is that it leverages a custom dialog model, designed exclusively for non-visual web access, for interpreting speech utterances. A user study with 20 blind subjects showed that Capti-Speak was significantly more usable and efficient compared to the regular screen reader, especially for ad-hoc browsing, searching, and navigating to the content of interest.
KW - Dialog act model
KW - Speech enabled Web screen reader
KW - Spoken dialog system
KW - Web automation
UR - https://www.scopus.com/pages/publications/84953426341
U2 - 10.1145/2745555.2746660
DO - 10.1145/2745555.2746660
M3 - Conference contribution
AN - SCOPUS:84953426341
T3 - W4A 2015 - 12th Web for All Conference
BT - W4A 2015 - 12th Web for All Conference
PB - Association for Computing Machinery, Inc
Y2 - 18 May 2015 through 20 May 2015
ER -