loracle
Datasets
All datasets matching “loracle”loracle-pretrain-v5-qwen14b-tokensloracle-persona-loras-all7-r1loracle-eval-direction-tokensloracle-ia-14b-direction-tokensloracle-fineweb-openrouter-gemini-3-flash-1k-finetunes
loracle-fineweb-openrouter-gemini-3-flash-1k-finetunes
Synthetic Loracle supervision data generated from FineWeb with OpenRouter.
Run summary
source dataset: HuggingFaceFW/fineweb / sample-10BT / train
sampled docs: 6500
synthetic finetunes: 1284
generated finetunes in this shard: 1000
generator backend: openrouter
generator model: google/gemini-3-flash-preview
max docs per finetune: 40
max token budget per finetune: 10000
questions per finetune: 10
Configs… See the full description on the dataset page: https://huggingface.co/datasets/japhba/loracle-fineweb-openrouter-gemini-3-flash-1k-finetunes.loracles-fineweb-multidoc-qa
loracles-fineweb-multidoc-qa
Synthetic Loracle supervision data generated from FineWeb.
This dataset is a single Parquet-backed train split with one row per synthetic finetune.
This upload is a partial snapshot of a larger run.
Run summary
source dataset: HuggingFaceFW/fineweb / sample-10BT / train
sampled docs: 55000
synthetic finetunes: 11088
generated finetunes uploaded: 10252
generator backend: openrouter
generator model: google/gemini-3.1-flash-lite-preview
max docs… See the full description on the dataset page: https://huggingface.co/datasets/cds-jb/loracles-fineweb-multidoc-qa.
