Aug 2026· Vestnik of the Plekhanov Russian University of Economics· Vol 23, pp. 48-57· 0 citations· 5 references
TL;DR
Analysis of current approach to using corporate data to upgrade quality of answers generated by LLM and the absence of mechanism necessary for integral analysis of loaded proprietary data is shown.
Abstract
The development of large language models (LLM) has extended opportunities of processing non-structured data in corporate environment. However, LLM fundamental restriction lies in their dependence on data of preliminary learning, which can decrease their ability to work safely with enterprise information non-published in other sources. The article provides analysis of current approach to using corporate data to upgrade quality of answers generated by LLM. Systems of Retrieval-Augmented Generation (RAG) are studied both in classical vector realization and in GraphRAG. Agent systems and integration of RAGand GraphRAG approaches in their architecture were discussed, as it can make it possible to build complicated systems of AI and solutions requiring multi-stage analysis of corporate non-structured data. Architecture of classical vector RAG-approach is described, which consists of three key stages: getting vector presentation of indexed fragments of papers; searching for fragments relevant to user requirement; forming the answer on extracted context combined with the initial request. As a key restriction the author showed the absence of mechanism necessary for integral analysis of loaded proprietary data. In its turn GraphRAG uses mechanism of plotting knowledge graph on context, which can help work not with semantically similar data fragments but with hierarchically clustered information by accumulated analysis of text summaries. Shortcomings and benefits of using approaches were formulated, criteria of expediency of RAG and GraphRAG practical application were identified and examples of systems for their use were provided
This study provides an AI- Based document analyzer with a question-answer system that makes use of Natural Language Processing approaches that is affordable, scalable, and suitable for business, education, and research.
Radhika Sharma, Devraj Gautam· Revolutionary Advances in Co...· 0 citations
This work introduces a trust based adaptive reranking model- ATM (Adaptive Trust Model) that allocates computational resources according to file level uncertainty, instead of assigning a fixed number of reranker calls per query, which focuses computation only where ranking confidence is low.
Jenny Kalaiarasi.S· Journal of Intelligent Decis...· 0 citations
The study resulted in the development of a novel six-component framework comprising Input Processing, LLM Core, Knowledge Enhancement, Context Management, Response Generation, Response Generation, and Human Feedback that successfully addressed resource scarcity through language detection and cross-lingual query understanding.
The role of traditional models in current NLP research and practice is discussed, especially in contrast and comparison to modern neural network-based approaches including LLMs.
Robin Jegan· Schriften aus der Fakultät W...· 0 citations
Experiments on three different domain tasks show that FKGLM can effectively integrate LLMs and large-scale knowledge graphs, leading to a significant enhancement in the reasoning capabilities of LLMs.
Yulin Zhou, Yongbin Qin, Chuan Lin· Journal of King Saud Univers...· 0 citations
The emergence of Large Language Models (LLMs) has redefined how users interact with information in digital environments. However, their widespread and often indiscriminate integration has raised significant concerns regarding reliability and trustworthiness issues that are particularly critical when accessing digital libraries and historical archives. How can one leverage the generalization capacity of an LLM without losing the level of accountability required for an archival institution? In this paper, we present an agentic retrieval system designed to deliver more accurate and verifiable access to historical data while preserving much of the flexibility associated with unconstrained LLMs. As a contribution to historical document analysis, we compare traditional Retrieval-Augmented Generation (RAG) with an agentic GraphRAG architecture in their ability to deliver historical information under realistic conditions, including the presence of OCR and transcription errors. We introduce a semi-symbolic framework that integrates word-spotting techniques for post-OCR correction with a knowledge graph representation that enables the agent to access information through synthesized queries. The interleaved collaboration between word spotting and code generation allows the agent to construct strong retrieval queries that are robust to misinterpretation and hallucination, while still leveraging approximate search when noise and uncertainty, common in historical document analysis, would otherwise hinder precise retrieval.
S. Nicolau, Adrià Molina, O. R. Terrades et al.· IEEE International Conferenc...· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.