International Scientific Journal of Engineering and Management

An International Scholarly || Multidisciplinary || Open Access || Indexing in all major Database & Metadata
The journal follows the UGC Guidelines and is evaluated for inclusion in the Web of Science
ISSN: 2583-6129

Impact Factor: 8.072

AI Knowledge Extraction System

Version
File Size 379.44 KB
Downloads 9
Files 1
Published 20 April 2026
Updated 20 April 2026

AI Knowledge Extraction System

Sura Chandu
Department of Computer Science

Rajiv Gandhi University of Knowledge Technologies

Basar, Telangana, India b200152@rgukt.ac.in

Vollala Saiprakash
Department of Computer Science

Rajiv Gandhi University of Knowledge Technologies

Basar, Telangana, India b200770@rgukt.ac.in

Sayyam Sai Kumar
Department of Computer Science

Rajiv Gandhi University of Knowledge Technologies Basar, Telangana, India B201489@rgukt.ac.in

Buddannagari Latha
Assistant Professor, Dept. of Computer Science Rajiv Gandhi University of Knowledge Technologies

Basar, Telangana, India latha.reddy5808@gmail.com

 

 

Abstract—The proposed system architecture illustrates a com- prehensive RAG + CAG (Retrieval-Augmented Generation and Context-Augmented Generation) Multi-Source Knowledge Ex- traction System designed to dynamically ingest, process, and synthesize information from diverse digital mediums. At its foun- dation, a robust Content Extraction Module gathers unstructured data from various sources—including PDF documents, websites (utilizing web scrapers), YouTube videos via transcript extraction, and images via Optical Character Recognition (OCR). This raw data then undergoes rigorous pre-processing and “smart chunking” before being transformed by an Embedding Generator and indexed within a scalable Vector Database (Chroma). When a user submits a question, the system generates a query embedding to seamlessly retrieve the most contextually relevant information chunks. Finally, an advanced RAG + CAG Engine synthesizes these retrieved chunks alongside source metadata and conversa- tion context, ultimately delivering highly accurate, synthesized answers. Extensive quantitative evaluations demonstrate a 92% retrieval accuracy and a near-zero hallucination rate, proving the viability of this localized, hybrid architecture for enterprise and academic deployment.Index Terms—Retrieval-Augmented Generation (RAG), Large Language Models (LLM), Optical Character Recognition (OCR), Natural Language Processing (NLP), Document Extraction, Vec- tor Database, Context-Augmented Generation (CAG), Multi- Modal Processing.

Download
or download free
[changelog]

Categories & Tags

Similar Downloads

No related download found!
ISJEM Journal

Author's Blog

What is the difference between a Research Paper and a Review Paper?

A research paper and a review paper are both scholarly documents, but they serve different purposes and have different characteristics....
Read More
Author's Blog

What is DOI?

A Digital Object Identifier (DOI) is a unique alphanumeric string that is used to identify and provide a persistent link...
Read More
Author's Blog

What do you need to do during production of your Research Paper?

During the production of a research paper, the following steps need to be taken: conducting research, organizing and analyzing data,...
Read More
Author's Blog

What are the advantages of publishing a research paper?

Publishing a research paper can have many advantages for researchers, including: Career advancement, professional recognition, opportunities for collaboration, increased visibility,...
Read More
Author's Blog

Ways to Support your Academic Wellbeing which preparing the Research Paper/Article

To support your academic wellbeing while publishing a research paper, it's important to set realistic goals, manage your time effectively,...
Read More
Author's Blog

How to improve your Research Paper writing Skills?

Read extensively: One of the best ways to improve your research paper skills is to read extensively in your field...
Read More
Author's Blog

Is DOI compulsory to publish a research paper in a Journal?

DOI is not strictly required to publish a research paper, but it is highly recommended. Basically, the International Scientific Journal...
Read More
Author's Blog

In what ways does research paper give weight to career development?

Publishing a research paper can give weight to a researcher's career development in several ways, such as: establishing oneself as...
Read More
Author's Blog

How to develop a Research Paper from Scratch

Developing a research paper involves several steps including: choosing a topic, conducting background research, formulating a research question or hypothesis,...
Read More
Author's Blog

How Plagiarism report plays crucial role in Research Paper Publication?

Plagiarism is a major concern in the academic and research community, as it undermines the integrity of the research and...
Read More