Skip to content

1 paper indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#artificial intelligence Preprint May 2026

BitsMoE: Cost-Aware Bit Allocation in Spectral Space for MoE LLM Quantization

BitsMoE is proposed, a cost-aware mixed-precision quantization framework built on two complementary techniques that separates expert weights into a shared basis and expert-specific spectral components, defining structural quantization units while exploiting cross-expert redundancy.

Jiayu Zhao, Zi-Han Teng, Min-Hao Fan et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.