
NLP CLASSIFICATION / MODEL EVALUATION
COMPLETED V1–V2 STUDYFinancial Complaint Auto-Routing with NLP
Leakage-safe eight-class CFPB text-classification study comparing a locked TF-IDF + Linear SVM benchmark with a frozen DistilBERT challenger and human-in-the-loop routing.
- Removed duplicate-text leakage and used group-aware 2024 development/test splits with zero normalized-text overlap.
- On the shared 2024 benchmark, the frozen DistilBERT challenger increased Macro F1 from 0.7671 to 0.7949; both frozen models were later compared on a 30,156-row retrospective 2025 cohort.







