Skip to content
Book Open access

CubeTrace: Microscopic Network Tracing for Heterogeneous Cloud Gateways

Aug 2026 · Conference on Applications, Technologies, Architectures, and Protocols for Computer Communication · pp. 475-490 · 1 citation · 106 references
Computer Science

TL;DR

CubeTrace is presented, a unified, function-level flow tracing system that enables microscopic tracing inside heterogeneous cloud gateways and introduces minimal overhead, consuming less than 1% of memory resources and adding less than 1% to forwarding latency.

Abstract

Modern cloud gateways have evolved to include diverse network functions and heterogeneous hardware, such as programmable switches and FPGAs, to handle increasing workloads and minimize forwarding latency. Existing network tracing tools, however, operate primarily at device granularity and cannot pinpoint which function on which hardware component causes packet losses or latency spikes. To bridge this gap, we present CubeTrace, a unified, function-level flow tracing system that enables microscopic tracing inside heterogeneous cloud gateways. CubeTrace standardizes tracing units as cubes across different hardware platforms, regardless of their varied underlying implementations, and operates at flow granularity for reliability reasons. This introduces a new tracing abstraction for heterogeneous gateways while maintaining low overhead. Moreover, the collected flow-cube data by CubeTrace can be decoded into packet-level representations and integrated with well-established distributed tracing frameworks, enabling the use of off-the-shelf analysis tools. Our evaluations demonstrate that CubeTrace introduces minimal overhead, consuming less than 1% of memory resources and adding less than 1% to forwarding latency. Having been deployed in a large-scale cloud gateway, CubeTrace has significantly improved problem localization, reducing resolution times from hours or even days to just minutes.

Read PDF

Similar papers

Jul 2026

A Cloud Continuum Research Infrastructure for Distributed CPS Experimentation

The proposed approach separates the research-infrastructure layer, which exposes and manages distributed resources, from the application layer, where Cyber-Physical workflows are organized according to an Edge-Fog-Cloud pattern in which placement, timing, and data provenance are treated as first-class experimental concerns.

Fabio Orazio Mirto, Giuseppe Tricomi, L. D’Agati et al. · 0 citations
Open access May 2026

Spanergy: Energy-Aware Distributed Tracing for Microservices

Cloud computing is gaining popularity by giving access to seemingly unlimited virtual resources. However, Cloud data centres are built with physical resources and their electricity consumption has been continuously growing over the past decades. Microservices are an important building block of Cloud applications, calling for new solutions to observe their energy consumption. Distributed tracing is widely deployed to diagnose latency and failures in microservice-based applications, yet it does not expose the energy cost of individual end-user requests. Such a gap limits energy-aware debugging, accountability, and control. This paper presents Spanergy, an energy-aware distributed tracing approach that correlates permicroservice power measurements with traces and that attributes measured energy consumption to request segments, i.e. trace spans. We showcase Spanergy with synchronous request chains and asynchronous interactions across microservices. We present a rigorous experimental protocol and statistical analysis plan to quantify overhead and to validate conservation and coverage properties on realistic configurations. Enabling OpenTelemetry tracing increased total experiment energy by 59.1% relative to the uninstrumented baseline, and Spanergy post-processing added 15.2% of the baseline energy. Hence, Spanergy's incremental energy cost is smaller than the energy overhead of enabling tracing itself, making the approach lightweight in practice. Spanergy also reveals that a non-negligible fraction of request energy comes from spans outside the latency-critical path. These results show that energy-aware tracing is feasible at modest overhead and provides actionable insights for energy-efficient microservices.

César Perdigão Batista, D. Conan, S. Chabridon · 1 citation
Book Open access Aug 2026

Dorado: Scaling SmartNIC Session Tables on Commodity DDRs

Dorado is a novel design that scales SmartNIC session tables entirely on inexpensive DDR modules and uses three new techniques that extract commodity DDR performance by restructuring session table layout, decomposing processing pipelines to reduce locking, and scheduling memory accesses to minimize stalls.

Heng Yu, Kai Ren, Jiajun Liang et al. · 0 citations
Review Open access 2021

Edge Computing Architectures for Ultra-Low Latency Applications

The evaluation of the proposed architecture through analytical models and simulation-based evaluations shows that the proposed architecture can reduce the latency onto 65 percent of the time relative to the conventional cloud-based architecture, affirm the claim that edge computing is an essential enabler of the next-generation applications that demand deterministic response time, high reliability and localized intelligence.

Priya Natarajan · 0 citations
Book Open access Aug 2026

Spillway: Orchestrating DPU and Host into a Unified vSwitching Fabric

Spillway introduces a DPU-host hybrid data plane that repurposes idle host CPU resources to process spillover traffic when the DPU becomes the bottleneck, and decouples virtual switching capacity from static DPU hardware limits.

Xiaochong Jiang, Dian Fan, Yilong Lv et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.