Home / Current Issue / Paper 1714721
A Multi-Stage Deep Learning Framework for Document Image Restoration
Subject area: Science,Engineering and Technology · Area of research: Computer Vision -> Document Enhancement
DOI: https://doi.org/10.64388/IREV9I8-1714721
Abstract
Real-world camera-captured document images often exhibit complex degradations, including cast shadows, non-uniform illumination, and contrast distortion, which severely degrade visual quality and prevent robust document analysis. In this paper, we present an illumination estimation multi-stage deep learning framework for restoring document images that explicitly separates shadow suppression from illumination normalization. The proposed pipeline involves an initial deep network estimating and mitigating shadow-induced intensity variations before a refinement network corrects global illumination consistency while maintaining textual structure and fine document details. By decomposing the enhancement task into complementary stages, the framework effectively copes with both local shadow artifacts and global lighting imbalance in unconstrained document imaging scenarios. Extensive experiments on real-world camera-captured document images reveal that the proposed method provides visually coherent enhancement with more readable results compared to conventional image processing techniques and existing deep learning-based methods. Standard image quality metrics have been quantitatively evaluated, showing notable gains. The results indicate that the proposed framework offers a robust and practical preprocessing solution for analyzing camera-based document images.
Keywords
Document image enhancement, Shadows and Illumination, Multi-Stage Deep Learning, Camera-Captured Documents
How to cite this paper
@article{1714721,
author = {Inukollu Anantha Prakash Reddy, Veesam Venkata Srinivas, Gosu Madhu, Dharmavarapu Jayaraju, Gadipudi Krishna Vamsi},
title = {A Multi-Stage Deep Learning Framework for Document Image Restoration},
journal = {Iconic Research And Engineering Journals},
year = {2026},
volume = {9},
number = {8},
pages = {2080-2088},
issn = {2456-8880},
url = {https://www.irejournals.com/formatedpaper/1714721.pdf},
abstract = {Real-world camera-captured document images often exhibit complex degradations, including cast shadows, non-uniform illumination, and contrast distortion, which severely degrade visual quality and prevent robust document analysis. In this paper, we present an illumination estimation multi-stage deep learning framework for restoring document images that explicitly separates shadow suppression from illumination normalization. The proposed pipeline involves an initial deep network estimating and mitigating shadow-induced intensity variations before a refinement network corrects global illumination consistency while maintaining textual structure and fine document details. By decomposing the enhancement task into complementary stages, the framework effectively copes with both local shadow artifacts and global lighting imbalance in unconstrained document imaging scenarios. Extensive experiments on real-world camera-captured document images reveal that the proposed method provides visually coherent enhancement with more readable results compared to conventional image processing techniques and existing deep learning-based methods. Standard image quality metrics have been quantitatively evaluated, showing notable gains. The results indicate that the proposed framework offers a robust and practical preprocessing solution for analyzing camera-based document images.},
keywords = {Document image enhancement, Shadows and Illumination, Multi-Stage Deep Learning, Camera-Captured Documents},
month = {February},
doi = {https://doi.org/10.64388/IREV9I8-1714721}
}