Home / Current Issue / Paper 1723768
IntelliVision: An AI-Based FIR Draft Generation System
Subject area: Science,Engineering and Technology · Area of research: AI and Natural Language Processing
Abstract
IntelliVision is an AI-based FIR draft generation system that processes text, speech, and image-based complaints using OCR, speech recognition, and NLP. It extracts key information, classifies crime types, retrieves similar cases, and generates an AI-assisted FIR draft for officer review. The crime classification module uses TF-IDF and Logistic Regression and achieved 100% test accuracy on a synthetic dataset of 515 complaints. The system aims to reduce documentation effort and improve the consistency of FIR preparation.
Keywords
artificial intelligence; FIR draft generation; legal NLP; named entity recognition; OCR; speech recognition; information extraction; large language models; human-in-the-loop
References
[1] Premasiri D, Ranasinghe T, Mitkov R, El-Haj M, Frommholz I. Survey on legal information extraction: Current status and open challenges. Knowl Inf Syst. 2025;67(12):11287–11358. https://doi.org/10.1007/s10115-025-02600-5
[2] Xing X, Chen P. Entity extraction of key elements in 110 police reports based on large language models. Appl Sci. 2024;14(17):7819. https://doi.org/10.3390/app14177819
[3] Leitner E, Rehm G, Moreno-Schneider J. Fine-grained named entity recognition in legal documents. In: Semantic Systems: The Power of AI and Knowledge Graphs. Cham: Springer; 2019. p. 272–287. (Lecture Notes in Computer Science; vol. 11702). https://doi.org/10.1007/978-3-030-33220-4_20
[4] Adhikary S, Sen P, Roy D, Ghosh K. A case study for automated attribute extraction from legal documents using large language models. Artif Intell Law. 2026;34:245–266. https://doi.org/10.1007/s10506-024-09425-7
[5] Brown TB, et al. Language models are few-shot learners. In: Proc. Adv. Neural Inf. Process. Syst. 2020. p. 1877–1901.
[6] Lewis P, et al. Retrieval-augmented generation for knowledge-intensive NLP tasks. In: Proc. Adv. Neural Inf. Process. Syst. 2020. p. 9459–9474.
[7] Radford A, et al. Robust speech recognition via large-scale weak supervision. In: Proc. 40th Int. Conf. Mach. Learn., PMLR. 2023. p. 28492–28518.
[8] Baevski A, Zhou H, Mohamed A, Auli M. wav2vec 2.0: A framework for self-supervised learning of speech representations. In: Proc. Adv. Neural Inf. Process. Syst. 2020. p. 12449–12460.
[9] Li M, et al. TrOCR: Transformer-based optical character recognition with pre-trained models. In: Proc. AAAI Conf. Artif. Intell. 2023. p. 13094–13102. https://doi.org/10.1609/aaai.v37i11.26538
[10] Devlin J, Chang MW, Lee K, Toutanova K. BERT: Pre-training of deep bidirectional transformers for language understanding. In: Proc. NAACL-HLT. 2019. p. 4171–4186. https://doi.org/10.18653/v1/N19-1423
[11] Paul S, Mandal A, Goyal P, Ghosh S. Pre-trained language models for the legal domain: A case study on Indian law. In: Proc. 19th Int. Conf. Artif. Intell. Law. 2023. p. 187–196. https://doi.org/10.1145/3594536.3595165
[12] Guha N, et al. LegalBench: A collaboratively built benchmark for measuring legal reasoning in large language models. In: Proc. Adv. Neural Inf. Process. Syst. 2023.
[13] Yao S, Zhao J, Yu D, Du N, Shafran I, Narasimhan K, et al. ReAct: Synergizing reasoning and acting in language models. In: Proc. Int. Conf. Learn. Representations. 2023.
How to cite this paper
@article{1723768,
author = {Prof. Krupali Dhawale, Payal Chadhokar, Prarthana Shukla, Sakshi Baghel, Vishakha Varani},
title = {IntelliVision: An AI-Based FIR Draft Generation System},
journal = {Iconic Research And Engineering Journals},
year = {2026},
volume = {10},
number = {4},
pages = {1066-1075},
issn = {2456-8880},
url = {https://www.irejournals.com/formatedpaper/1723768.pdf},
abstract = {IntelliVision is an AI-based FIR draft generation system that processes text, speech, and image-based complaints using OCR, speech recognition, and NLP. It extracts key information, classifies crime types, retrieves similar cases, and generates an AI-assisted FIR draft for officer review. The crime classification module uses TF-IDF and Logistic Regression and achieved 100% test accuracy on a synthetic dataset of 515 complaints. The system aims to reduce documentation effort and improve the consistency of FIR preparation.},
keywords = {artificial intelligence; FIR draft generation; legal NLP; named entity recognition; OCR; speech recognition; information extraction; large language models; human-in-the-loop},
month = {October},
}