marin-community/longcontext-marin-8b-openthoughts3
213
Long-Context Marin 8B - OpenThoughts3
This model is the final checkpoint from the exp2199a2_redo2 experiment, which fine-tunes the long-context extended Marin-8B model on the OpenThoughts3 dataset.
Model Details
- Base Model: tootsie-8b-giraffe-phase3-64k (Marin 8B with 64k context extension)
- Training Dataset: OpenThoughts3-1.2M (1.2M examples)
- Final Checkpoint: step-11718
Training Hyperparameters
Training Notes
- Era shuffling enabled (dataset shuffled every epoch)
- Trained with Llama3-style rotary embeddings configured for 64k context
