Vietnamese Text Classification with PhoBERT

Fine-tuning Vietnamese pretrained language models (PhoBERT, with a Longformer variant for long documents) for domain text classification — covering tokenization strategy, long-input handling, and evaluation of the final model.

Fine-tuning PhoBERT and a Longformer variant for Vietnamese document classification.