Happy-Chen-CH/text_classification
A Chinese news headline classification benchmark: 200K labeled samples across 10 categories. Covers 4 approaches — TF-IDF+RandomForest, FastText, BERT fine-tuning+int8 quantization, and knowledge distillation (BERT→TextCNN) — spanning classical ML to model compression, with Flask RESTful API.