halimajaved592/urdu-english-code-switching-xlm-roberta
0139
Urdu-English Code-Switching Classification Model
This model is an XLM-RoBERTa-based token classification model trained for Urdu-English code-switching detection.
Labels
- URD — Urdu
- ENG — English
- MIX — Mixed
Training
- Model: xlm-roberta-base
- Epochs: 5
- Dataset size: 971 rows
Evaluation
F1 by Label
Confusion Matrix
The model performs strongly on URD and ENG. MIX has lower recall because the test set contained only 11 MIX examples.
