2023· International Journal of Applied Data Science & Modern Computing· Vol 6, pp. 01-14· 0 citations
TL;DR
The study highlights that no single database solution fits all scenarios, and hybrid architectures with polyglot persistence are increasingly adopted, offering guidance for selecting appropriate systems in data-driven environments.
Abstract
Modern data mining applications require scalable and high-performance database systems capable of handling structured, semi-structured, and unstructured data. Traditional RDBMS, designed for transactional consistency, face limitations in scalability and performance for big data analytics. As a result, modern systems such as NoSQL, NewSQL, distributed file systems, and cloud-native platforms have emerged, offering features like horizontal scalability, schema flexibility, and real-time processing. This paper provides a comprehensive overview and comparison of these database systems, focusing on their architecture, data models, and suitability for analytical workloads. It also examines their role in data science pipelines, including data ingestion, preprocessing, model training, and deployment. Key trade-offs based on the CAP theorem are discussed, along with performance metrics such as scalability, latency, and fault tolerance. The study highlights that no single database solution fits all scenarios, and hybrid architectures with polyglot persistence are increasingly adopted. The paper concludes by identifying research gaps in database optimization for machine learning, data governance, and real-time analytics, offering guidance for selecting appropriate systems in data-driven environments.
The rapid growth of data volume, structural heterogeneity, and real-time processing demands in modern e-commerce platforms challenges the scalability and flexibility of single-model database systems. Hybrid database architectures, which integrate multiple storage paradigms within a coordinated framework, offer a promis...
The performance of transactional applications involving databases will decline over time due to the increasing amount of stored data and the continuously rising volume of transactions. This research proposes a summary table design method that integrates table partitioning techniques and data aggregation to improve quer...
Edi Widodo, Kartika Imam Santoso, Wawan Setiawan· Edu Komputika Journal· 0 citations
In today’s world of technology, data systems have gained prominence in all arenas including business, healthcare, education and science. This review of literature provides a guide through the most recent trends in data systems and analysis, stressing that it is necessary to be aware of contemporary changes. The search...
Martin Muthomi· International Journal of Sci...· 0 citations
Data Lake is a highly flexible storage solution that can store both structured and unstructured data and operates on the schema-on-read approach. It acts as a potential alternative to the current Big Data storage issue. However, it does have certain flaws, such as inadequate authentication and access control. This pape...
Aakash Aundhkar· Journal of Science & Tec...· 2 citations
Data warehouses today are an essential part of existing Business Intelligence (BI) solutions, that collect data from multiple sources and enable analytical processing for strategic decision making. However, as enterprise data grows, and analytical queries become more complex, faster insights are demanded that present a...
Purushotham Jinka· World Journal of Advanced En...· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.