apps1/without_distillation
06
1---2library_name: transformers3license: apache-2.04base_model: NeuML/bert-hash-nano5tags:6- generated_from_trainer7model-index:8- name: without_distillation9 results: []10---11 12<!-- This model card has been generated automatically according to the information the Trainer had access to. You13should probably proofread and complete it, then remove this comment. -->14 15# without_distillation16 17This model is a fine-tuned version of [NeuML/bert-hash-nano](https://huggingface.co/NeuML/bert-hash-nano) on an unknown dataset.18 19## Model description20 21More information needed22 23## Intended uses & limitations24 25More information needed26 27## Training and evaluation data28 29More information needed30 31## Training procedure32 33### Training hyperparameters34 35The following hyperparameters were used during training:36- learning_rate: 0.000537- train_batch_size: 3238- eval_batch_size: 3239- seed: 4240- optimizer: Use OptimizerNames.ADAMW_TORCH with betas=(0.9,0.999) and epsilon=1e-08 and optimizer_args=No additional optimizer arguments41- lr_scheduler_type: linear42- num_epochs: 1043 44### Training results45 46 47 48### Framework versions49 50- Transformers 5.9.051- Pytorch 2.10.0+cu12852- Datasets 4.8.553- Tokenizers 0.22.254 