{"title":"Combining Augmented Reality and Speech Technologies to Help Deaf and Hard of Hearing People","authors":"M. Mirzaei, S. Ghorshi, Mohammad Mortazavi","doi":"10.1109/SVR.2012.10","DOIUrl":null,"url":null,"abstract":"Augmented Reality (AR), Automatic Speech Recognition (ASR) and Text-to-Speech Synthesis (TTS) can be used to help people with disabilities. In this paper, we combine these technologies to make a new system for helping deaf people. This system can take the narrator's speech and convert it into a readable text and show it directly on AR display. To improve the accuracy of the system, we use Audio-Visual Speech Recognition (AVSR) as a backup for the ASR engine in noisy environments. In addition, we use the TTS system to make our system more usable for deaf people. The results of testing the system show that its accuracy is over 85 percent on average in different places. Also, the result of a survey shows that more than 90 percent of deaf people on average are very interested in using our system as an assistant in portable devices for communication.","PeriodicalId":319713,"journal":{"name":"2012 14th Symposium on Virtual and Augmented Reality","volume":"20 1","pages":"0"},"PeriodicalIF":0.0000,"publicationDate":"2012-05-28","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":"23","resultStr":null,"platform":"Semanticscholar","paperid":null,"PeriodicalName":"2012 14th Symposium on Virtual and Augmented Reality","FirstCategoryId":"1085","ListUrlMain":"https://doi.org/10.1109/SVR.2012.10","RegionNum":0,"RegionCategory":null,"ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"","PubModel":"","JCR":"","JCRName":"","Score":null,"Total":0}
引用次数: 23
Abstract
Augmented Reality (AR), Automatic Speech Recognition (ASR) and Text-to-Speech Synthesis (TTS) can be used to help people with disabilities. In this paper, we combine these technologies to make a new system for helping deaf people. This system can take the narrator's speech and convert it into a readable text and show it directly on AR display. To improve the accuracy of the system, we use Audio-Visual Speech Recognition (AVSR) as a backup for the ASR engine in noisy environments. In addition, we use the TTS system to make our system more usable for deaf people. The results of testing the system show that its accuracy is over 85 percent on average in different places. Also, the result of a survey shows that more than 90 percent of deaf people on average are very interested in using our system as an assistant in portable devices for communication.