An Open Benchmark for Systems Vaccinology: Insights from the CMI-PB Challenges
Systems vaccinology approaches have identified factors affecting vaccine responses in multiple studies, but the ability of computational models to generalize these findings to unseen data remains unclear. We established a community resource to create and compare models predicting B. pertussis booster vaccination responses and put such modeling approaches to the test. We compiled multi-modal experimental training data from three independent cohorts (n=117 individuals), and asked investigators to predict vaccine responses in a cohort of 54 newly recruited individuals using only their pre-booster vaccination data. We benchmarked a total of 107 computational models. Top-performing models were characterized by workflows that prioritized rigorous data preprocessing, robust imputation of missing data, and the use of multi-omics integration or non-linear machine learning. We identified pre-existing antigen-specific antibody titers and baseline monocyte frequencies as the most consistent predictors of post-vaccination immunity, highlighting the dominant role of individual immune setpoints. We established the resulting datasets and evaluation framework as a community resource to advance predictive immunology and facilitate personalized vaccination strategies.