We are trying to build a project for visually impaired people in which a module takes a picture and convert the text embedded in that picture to speech output. We thought of using a HD camera cape with beagleboneblack to capture an image and extract the text from the image and by tts emic-2(text to speech converter) that text is converted to speech. IS the mission will be accomplished with these products? If not any suggestions please!!