Multimodal GenAI for Communicative Language Teaching: Gemini Live
ABSTRACT This review explores the pedagogical affordances of Gemini Live, Google's multimodal generative AI (GenAI) chatbot, for communicative language teaching, with particular attention to task‐based interaction. Unlike earlier text‐ or voice‐based GenAI chatbots, Gemini 2.5 released in 2025 integrated live vision and voice to support context‐rich, goal‐oriented communication. It enables learners to engage in interactive, multimodal activities such as object‐based exchanges, gesture‐supported interaction, and screen‐shared collaboration. Although such activities do not automatically constitute full task‐based language teaching (TBLT), Gemini's multimodal environment offers a promising platform for scaffolded, authentic interaction, reflective post‐task discussion, and task‐supported language learning.