Classifying CRISPR-Cas9 Off-Target Cleavage Sites from GUIDE-seq Data: A Class-Imbalanced Machine Learning Benchmark
Five machine learning classifiers are benchmarked on a real, published GUIDE-seq off-target dataset and the low absolute precision achievable in this severely imbalanced, small-positive-class setting is reported, as a realistic picture of what off-target classifiers can and cannot yet deliver from sequence alone.