FedSSMCoOp: SSM Encoders for light-weight Federated Prompt Learning for Few-shot Classification
Vision-Language Models (VLMs) have shown strong performance across a wide range of downstream vision tasks, thanks to the complementary information contained in the respective domains. Despite the performance gains, most of these approaches rely on aligning these domains using the cosine similarity metric, which fails...