Skip to content
Open access

Continuous Data Transformation in Event-Driven Microservices

2019 · International Journal of Artificial Intelligence & Digital Transformation · Vol 2, pp. 01-14 · 0 citations

TL;DR

This paper presents real-world use cases, compares toolsets such as Apache Kafka Streams, Apache Flink, Debezium, and AWS Kinesis, and proposes a reference architecture to guide practitioners in designing scalable transformation pipelines.

Abstract

In modern distributed systems, event-driven microservices have emerged as a robust architectural paradigm, offering scalability, resilience, and decoupled communication. However, these systems often require real-time or near-real-time transformation of data across services and domains. Continuous data transformation—the ongoing process of modifying, enriching, or aggregating event data as it flows through the system—is critical for maintaining data consistency, enabling business insights, and supporting downstream consumers. This paper explores architectural patterns, technologies, and best practices for implementing continuous data transformation in event-driven microservices. It highlights common challenges such as schema evolution, message format variability, stateful processing, and system observability. Furthermore, it presents real-world use cases, compares toolsets such as Apache Kafka Streams, Apache Flink, Debezium, and AWS Kinesis, and proposes a reference architecture to guide practitioners in designing scalable transformation pipelines.

Read PDF

Similar papers

Review 2025

Scalable ETL Pipeline Architectures for Real-Time Transaction Analytics: Bridging Data Engineering and Business Operations

The study concludes that real-time analytical performance must be evaluated through both engineering and operational outcomes, and recommends use-case-driven architectural selection, resilient hybrid deployment, embedded security and governance, automated quality assurance, transparent AI-assisted pipeline management a...

Ogochukwu T. Izuchukwu, Dominic Feboh, Ayokunle Olamide Ijagbemi et al. · 0 citations
Open access Sep 2026

Serverless Data Engineering: Innovations in Python-Driven ETL Automation on AWS

Traditional cluster-based ETL architectures impose a structural tax on data engineering organisations: fixed compute resources provisioned for peak demand, scheduled batch cycles that introduce latency regardless of downstream urgency, and operational overhead that redirects engineering capacity from pipeline design to...

Rambabu Bolineni · 0 citations
Open access Aug 2026

Intelligent Business Integration With AI

An artificial intelligence-enhanced middleware pattern that augments existing integration stacks with telemetry, stream processing, and a lightweight learning loop to predict failures, automatically tune policies, and direct traffic in real time is presented.

Tejas Gajjar · 0 citations

Future Generation Computer

The integration combines DAGonStar’s orchestration capabilities with CAPIO’s efficient data handling to better support workflows operating on continuous or large-scale datasets and improves the responsiveness and flexibility of scientific workflows.

Simone Perrotta, G. de Vita, Gennaro Mellone et al. · 0 citations
Review Open access Aug 2026

Building High-Performance Streaming and Analytics Platforms

The quantity of data that is created, processed, and streamed into a contemporary organisation is significant. As an engineering challenge, it is how to keep this up to date at all times and how to ingest, process, and serve this data with high availability and fault tolerance. This paper reviews architectural patterns...

Srivardhan Jalan · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.