CoolFace
Modelpublic

RichardErkhov/smcleish_-_clrs_gemma_2b_100k_finetune_with_traces-gguf

sourceHugging Faceupdated 2y agoView on Hugging Face
0likes337downloads
Model Card

Quantization made by Richard Erkhov.

Github

Discord

Request more models

clrsgemma2b100kfinetunewithtraces - GGUF

  • —Model creator: https://huggingface.co/smcleish/
  • —Original model: https://huggingface.co/smcleish/clrsgemma2b100kfinetunewithtraces/

Original model description: --- libraryname: transformers license: mit basemodel:

  • —google/gemma-2b ---

Model Details

google/gemma-2b model finetuned on 100,000 CLRS-Text examples.

Training Details

  • —Learning Rate: 1e-4, 150 warmup steps then cosine decayed to 5e-06 using AdamW optimiser
  • —Batch size: 128
  • —Loss taken over answer only, not on question.