RichardErkhov/smcleish_-_clrs_gemma_2b_100k_finetune_with_traces-gguf
0337
Quantization made by Richard Erkhov.
clrsgemma2b100kfinetunewithtraces - GGUF
- Model creator: https://huggingface.co/smcleish/
- Original model: https://huggingface.co/smcleish/clrsgemma2b100kfinetunewithtraces/
Original model description: --- libraryname: transformers license: mit basemodel:
- google/gemma-2b ---
Model Details
google/gemma-2b model finetuned on 100,000 CLRS-Text examples.
Training Details
- Learning Rate: 1e-4, 150 warmup steps then cosine decayed to 5e-06 using AdamW optimiser
- Batch size: 128
- Loss taken over answer only, not on question.
