PhoBERT: Long-Document Sentiment Analysis in Vietnamese with Sparse Attention
The rapid growth of user-generated content on Vietnamese e-commerce platforms (Tiki, Google Play) has created an urgent need for accurate sentiment analysis of long documents (>256 tokens). However, existing Vietnamese Transformer models like PhoBERT are limited by the 256-token input limit, leading to a loss of contex...