Munich_Z@GermEval Shared Task 2025:When Prompting Is Not Enough: The Limits of Large Language Models in GermEval’s 2025 Harmful Content Detection Task
This work evaluates prompting strategies for subtask 2 of the GermEval 2025 Harmful Content Detection challenge, which involves classifying whether a tweet attacks the free democratic basic order and shows that techniques such as Chain-of-Thought, In-Context Learning or Task Decomposition outperform approaches like Task Description.