ESTD Year: 2017 | Impact Factor: 6.9
DOI Prefix: 10.47001/IRJIET
Vol 10 No 9 (2026): Volume 10, Issue 9, September 2026 | Pages: 110-116
International Research Journal of Innovations in Engineering and Technology
OPEN ACCESS | Research Article | Published Date: 25-09-2026
Official letters often need to be checked for required information before they can be processed further. This checking is usually done manually and can take time, especially when the letters are received as scanned documents or images. This paper presents an AI-assisted system developed to automate the validation of such letters and generate responses for incomplete submissions. The system accepts letters in PDF, PNG, and JPEG formats and uses Optical Character Recognition (OCR) to extract their text. The extracted text and document details are stored in a database and passed to the validation stage. The implemented format validator checks for required elements such as the date, subject, salutation, sender, and recipient. For content validation, the system is designed to use a domain-specific knowledge base with Retrieval-Augmented Generation (RAG) and a Large Language Model (LLM) to identify missing or incorrect information. The validation findings are then used to generate a response describing the information that needs to be provided or corrected. The system is implemented using Java and Spring Boot with separate components for document processing, OCR, database management, and validation. Synthetic MRSAC-related data is used during development of the AI-assisted components because the actual organizational requirements are confidential.
Letter Validation, OCR, Document Processing, Artificial Intelligence, Automated Response Generation
Tejas Dange, Vansh Nagpure, Minakshee Chandankhede, & S. V. Balamwar. (2026). AI-Assisted Letter Validation and Automated Response Generation for Official Letters. International Research Journal of Innovations in Engineering and Technology - IRJIET, 10(9), 110-116. Article DOI https://doi.org/10.47001/IRJIET/2026.109012
This work is licensed under Creative common Attribution Non Commercial 4.0 Internation Licence
C. S. Kumari, V. Y. Gupta, G. Manikanta, and B. V. Abhishikth, “Automated document processing: Combining OCR and generative AI for efficient text extraction and summarization,” in Proc. 4th Int. Conf. Information Technology, Civil Innovation, Science, and Management (ICITSM 2025), 2025, doi: 10.4108/eai.28-4-2025.2357770.
C. Luo, Y. Shen, Z. Zhu, Q. Zheng, Z. Yu, and C. Yao, “LayoutLLM: Layout instruction tuning with large language models for document understanding,” in Proc. IEEE/CVF Conf. Computer Vision and Pattern Recognition (CVPR), 2024, pp. 15630–15640.
A.Brown, M. Roman, and B. Devereux, “A systematic literature review of retrieval-augmented generation: Techniques, metrics, and challenges,” Big Data and Cognitive Computing, vol. 9, no. 12, Art. no. 320, 2025, doi: 10.3390/bdcc9120320.
J. Zhang, Q. Zhang, B. Wang, L. Ouyang, Z. Wen, Y. Li, K.-H. Chow, C. He, and W. Zhang, “OCR Hinders RAG: Evaluating the cascading impact of OCR on retrieval-augmented generation,” in Proc. IEEE/CVF Int. Conf. Computer Vision (ICCV), 2025, pp. 17443–17453, doi: 10.1109/ICCV51701.2025.01620.