SwitchEmbed: Representation Learning for Arabic-English Code-Switched Text
This work introduces CS-SNLI, a large-scale Arabic-English code-switched natural language inference dataset and adopts a contrastive learning framework with natural language inference supervision to encourage semantically consistent sentence embeddings across monolingual and code-switched variants.