Evaluating LLM-Based Topic Interpretation and Sentiment Classification Against LDA-Human and Human-Coded Benchmarks in Consumer Review Analytics
This empirical study compares performance of ChatGPT, Claude AI, and BARD to model topics on 547 Amazon reviews of 2 water filter brands and shows that LDA produced more coherent, detailed themes, while LLM-based topic modeling generated broader, more semantically diffused clusters, highlighting a trade-off between int...