CoolFace
Modelpublic

ymoslem/ModernBERT-base-TeleQnA-router-qe-classifier-binary-10ep-lr2e-05-qwen4b-5runs_1eval

sourceHugging Faceapache-2.0updated 4mo agoView on Hugging Face
1likes13downloads
Model Card

<!-- This model card has been generated automatically according to the information the Trainer had access to. You should probably proofread and complete it, then remove this comment. -->

Query Quality Estimation - Binary Classification

This model is a fine-tuned version of answerdotai/ModernBERT-base on the ymoslem/TeleQnA-router dataset. It achieves the following results on the evaluation set:

  • —Loss: 3.1483
  • —Accuracy: 0.698
  • —F1 Macro: 0.6348
  • —F1 Weighted: 0.6855
  • —Precision: 0.6527
  • —Recall: 0.6293

Model description

More information needed

Intended uses & limitations

More information needed

Training and evaluation data

More information needed

Training procedure

Training hyperparameters

The following hyperparameters were used during training:

  • —learning_rate: 2e-05
  • —trainbatchsize: 64
  • —evalbatchsize: 64
  • —seed: 42
  • —optimizer: Use OptimizerNames.ADAMWTORCHFUSED with betas=(0.9,0.999) and epsilon=1e-08 and optimizer_args=No additional optimizer arguments
  • —lrschedulertype: linear
  • —num_epochs: 10

Training results

Training LossEpochStepValidation LossAccuracyF1 MacroF1 WeightedPrecisionRecall
0.31581.07041.10340.6890.57790.65020.64090.5827
0.13182.014081.03670.640.62270.64970.62540.6399
0.07453.021122.09540.7060.61700.67870.66710.6134
0.03514.028162.10610.6850.61950.67220.63550.6150
0.03355.035202.50720.6970.64360.68970.65320.6390
0.00096.042242.35910.680.64550.68240.64370.6481
0.07.049282.74810.6910.64170.68610.64740.6383
0.00018.056322.99670.6960.63380.68420.65020.6285
0.09.063363.44200.7120.63210.68940.67460.6262
0.010.070403.14830.6980.63480.68550.65270.6293

Framework versions

  • —Transformers 4.57.3
  • —Pytorch 2.8.0+cu128
  • —Datasets 4.1.1
  • —Tokenizers 0.22.1