Skip to content

Author

Zihao Fan

We have 5 of 10 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Preprint Sep 2026

HyDra: Demystifying and Taming Dynamic Context Parallelism at Production Scale

Long-context training runs on sequences whose lengths span orders of magnitude, and dynamic context parallelism (DCP) gives each sequence its own CP degree. Existing DCP systems either do not scale or perform poorly on mainstream models, leaving Megatron-Core (Mcore) DCP as the only option at production scale. Mcore DC...

Zi-Hao Fan, Yun-Zhuo Liu, Bo Jiang et al. · 0 citations
Book Open access Aug 2026

Scaling LLM Agent Tool Access at Cloud Scale

The gateway breaks the direct-connect data plane and consolidates legacy API integration, protocol bridging, access control, and session-aware routing, while scaling out elastically at low per-call overhead.

Ming-Xing Li, Enge Song, Yueshang Zuo et al. · 0 citations
Jul 2026

Scalable LLM Agent Tool Access in the Cloud

A cloud-scale gateway system for MCP service is presented, which breaks the direct-connect model on the data plane and offloads legacy service integration, consolidating incompatible MCP variants, access control, tool recommendation, and session-aware routing to the gateway.

Ming-Xing Li, Enge Song, Yueshang Zuo et al. · 0 citations
Book Open access Aug 2026

Integrating AI Clusters into Virtual Private Cloud

An architecture that decouples complex policy enforcement from high-speed packet forwarding to support VPC semantics on back-end NICs and enable front-end/back-end integration is proposed, suggesting that commodity hardware can support both high-throughput AI training and flexible VPC features.

Yinhe Wang, Xing Li, Enge Song et al. · 0 citations
Book Open access Aug 2026

Single-Core Hotspots on Your VNF? Break Them Up!

ParaFlowO is proposed, an architecture that Parallelizes processing elephant Flows across multiple CPU cores while preserving in-Order delivery and integrates a lightweight reordering mechanism to preserve packet order and controls parallelism to mitigate contention on shared state.

Chang-Gang Zheng, Bowen Yang, Jin Ke et al. · 1 citation

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.