Home / Current Issue / Paper 1713623
Secure Sight: Digital-Watermarking-Driven OCR App for Visually Impaired Users on Android
Subject area: Science,Engineering and Technology · Area of research: Assistive OCR Mobile Applications
DOI: https://doi.org/10.64388/IREV9I7-1713623
Abstract
One of the most prominent issues in our society is that visually impaired people face numerous barriers. Smartphones are indispensable in today's information society [11]. With the popularity of the Android operating system, many applications are being developed for smartphone users. In this paper, we propose an innovative concept called "Smart Application for Visually Challenged Individuals". Our goal in designing this application is to empower visually impaired people by assisting them in becoming self-sufficient through the use of technology [1]. In a nutshell, our app will serve as their "EYES". The app's interface is designed to be very user-friendly, allowing users to easily navigate and utilize its features. The app's primary feature is its ability to speak out the text that is visible in photos that you take of road signs or other objects [4]. This functionality is achieved using an Optical Character Recognition (OCR) system, with Google Cloud Vision serving as the API . Voice-based commands can be activated simply by tapping anywhere in the app. After tapping, users can verbally request what they want from the app. This allows the app to audibly communicate the photo's content to the user [8]. To ensure smooth integration and optimal performance, we built the system on Android and included the necessary Google API functionalities.
Keywords
Android, Java, Object Detection , voice Commands
References
[1] Dr. Anita T. N., Akshatha S., Rachana B., Rahul Jangra, and Venkatesh Rayudu, "Android TTS OCR converter system for people with visual disability," Int. J. Innov. Res. Technol., vol. 11, no. 12, May 2025, doi: 10.17148/IJIRT177183.
[2] Smith, R., et al., "An overview of the Tesseract OCR engine," IJIRT, 2007. [Online]. Available: https://tesseractocr.github.io/
[3] Raman, T. V., "Emacspeak—Enabling access for blind users," Open Source Software, 1996. [Online]. Available: http://emacspeak.sourceforge.net/
[4] Smith, J., "Assistive technologies for the visually impaired: A comprehensive review," J. Accessibility Inclusion, vol. 8, no. 2, pp. 101–120, 2022.
[5] Smith, J., et al., "Enhancing accessibility: A smartphone-based OCR system for visually impaired users," J. Assistive Technol., pp. 1–15, 2019.
[6] John Williams and Emily Brown, "Optical character recognition for assistive technologies," Int. J. Computer Vision Appl., vol. 15, no. 3, pp. 245–260, 2020.
[7] Michael Lee and Sophia Chen, "Integrating OCR and TTS for visually impaired individuals," IEEE Trans. Human-Machine Syst., vol. 51, no. 4, pp. 380–390, 2021.
[8] James Anderson and Olivia White, "Gesture-based controls in accessibility applications," ACM Trans. Accessible Comput., vol. 14, no. 2, pp. 1–28, 2022.
[9] Harper, S. and Yesilada, Y., "Web accessibility: A foundation for research," Springer, Berlin, Germany, 2018. [Online]. Available: https://www.springer.com/gp/book/9781447174407
[10] Chen, D. and Odobez, J. M., "Text recognition in natural images," Pattern Recognition, vol. 39, no. 9, pp. 2010–2025, 2005. doi: 10.1016/j.patcog.2005.07.001.
[11] Y. N. Prajapati and M. Sharma, "Designing AI to Predict Covid-19 Outcomes by Gender," 2023 International Conference on Data Science, Agents & Artificial Intelligence (ICDSAAI), Chennai, India, 2023, pp. 1-7, doi: 10.1109/ICDSAAI59313.2023.10452565.
[12] Y. N. Prajapati and D. Baloni, "Detailed Explanation: Optimizing COVID-19 CT-Scan Classification with Feature Engineering and Genetic Algorithm," 2024 1st International Conference on Advanced Computing and Emerging Technologies (ACET), Ghaziabad, India, 2024, pp. 1-8, doi: 10.1109/ACET61898.2024.10730057
[13] Y. N. Prajapati and M. Sharma, "Designing AI to Predict Covid-19 Outcomes by Gender," 2023 International Conference on Data Science, Agents & Artificial Intelligence (ICDSAAI), Chennai, India, 2023, pp. 1-7, doi: 10.1109/ICDSAAI59313.2023.10452565..
[14] A REVIEW PAPER ON CAUSE OF HEART DISEASE USING MACHINE LEARNING ALGORITHMS. (2022). Journal of Pharmaceutical Negative Results, 9250-9259. https://doi.org/10.47750/pnr.2022.13.S09.1082. [Crossref]
[15] Yogendra N. Prajapati, Dev Baloni, and Avdhesh Gupta, "Optimized Heart Disease Image Classification on Edge Devices Using Knowledge Distillation and Layer Compression," Journal of Image and Graphics, Vol. 13, No. 5, pp. 459-468, 2025.
[16] Y. N. Prajapati, S. K. Sonker, P. P. Agrawal, J. Jain, M. Kumar and V. Kumar, "Brain Tumor Detection and Classification Using Deep Learning on MRI Images," 2025 2nd International Conference on Computational Intelligence, Communication Technology and Networking (CICTN), Ghaziabad, India, 2025, pp. 131-135, doi: 10.1109/CICTN64563.2025.10932392.
[17] Y. N. Prajapati, S. Yadav, S. Sharma, U. K. Patel and S. Tomar, "Analysis of Underlying Emotions in Textual Data Using Sentiment Analysis Which Classifies Text In to Positive, Negative or Neutral Sentiments," 2023 14th International Conference on Computing Communication and Networking Technologies (ICCCNT), Delhi, India, 2023, pp. 1-6, doi: 10.1109/ICCCNT56998.2023.10307310.
[18] Y. N. Prajapati, U. K. Patel, M. Srivastava, J. Yadav, M. K. Srivastava and B. Kumar Gupta, "Protecting Hearts with Support Vector Machine Analysis for the Early Detection of Heart Failure," 2025 2nd International Conference on Computational Intelligence, Communication Technology and Networking (CICTN), Ghaziabad, India, 2025, pp. 117-119, doi: 10.1109/CICTN64563.2025.10932449.
How to cite this paper
@article{1713623,
author = {Samender Singh, Mukesh Singla},
title = {Secure Sight: Digital-Watermarking-Driven OCR App for Visually Impaired Users on Android},
journal = {Iconic Research And Engineering Journals},
year = {2026},
volume = {9},
number = {7},
pages = {1791-1796},
issn = {2456-8880},
url = {https://www.irejournals.com/formatedpaper/1713623.pdf},
abstract = {One of the most prominent issues in our society is that visually impaired people face numerous barriers. Smartphones are indispensable in today's information society [11]. With the popularity of the Android operating system, many applications are being developed for smartphone users. In this paper, we propose an innovative concept called "Smart Application for Visually Challenged Individuals". Our goal in designing this application is to empower visually impaired people by assisting them in becoming self-sufficient through the use of technology [1]. In a nutshell, our app will serve as their "EYES". The app's interface is designed to be very user-friendly, allowing users to easily navigate and utilize its features. The app's primary feature is its ability to speak out the text that is visible in photos that you take of road signs or other objects [4]. This functionality is achieved using an Optical Character Recognition (OCR) system, with Google Cloud Vision serving as the API . Voice-based commands can be activated simply by tapping anywhere in the app. After tapping, users can verbally request what they want from the app. This allows the app to audibly communicate the photo's content to the user [8]. To ensure smooth integration and optimal performance, we built the system on Android and included the necessary Google API functionalities.},
keywords = {Android, Java, Object Detection , voice Commands},
month = {January},
doi = {https://doi.org/10.64388/IREV9I7-1713623}
}