CoolFace
Datasetpublic

kgrabko/JiRackBaseDataset

๐Ÿ’Ž JiRack Base dataset for 1.5B model Dataset: the dataset formated for JiRack tokenizer . I recommend initializing the model with a 4K context window for initial stability, followed by scaling to 8K context using specialized JiRack 8K datasets. This two-stage approach ensures robust positional encoding before extending the model's long-range dependency. Time: JiRack 1.5B: High-Efficiency Financial Modeling We are training a compact 1.5B parameter model on an extensive 11โ€ฆ See the full description on the dataset page: https://huggingface.co/datasets/kgrabko/JiRackBaseDataset.

sourceHugging Faceotherupdated 5mo agoView on Hugging Face
1likes51downloads
tokenizer.json4 linesDownload Raw Back to tokenizer
1version https://git-lfs.github.com/spec/v12oid sha256:5d75279cf79c7f2e2f448cdc247bba55f64cf802e24e3471845c75f876c6361e3size 172105284