CoolFace
20 results

tinyllm

Gabriel8 /tiny-llm-synthetic-qa Tiny-LLM: Synthetic Question-Answering Dataset Dataset Description This dataset was created for the fine-tuning stage of the Tiny-LLM Project, a project focused on training and evaluating compact language models from scratch. It contains 706,727 high-quality, synthetic multi-turn Question-Answering (Q&A) conversations in English, generated using the Gemini API. The dataset was designed to teach small models instruction-following capabilities across a diverse range of… See the full description on the dataset page: https://huggingface.co/datasets/Gabriel8/tiny-llm-synthetic-qa.textquestion-answering100K<n<1M2 likes113 downloads11mo agoHugging Facetinyllms /aime-1983-2023-trajectoriestext1K<n<10K0 likes55 downloads6mo agoHugging Facetinyllms /gpqa-extended-trajectoriestext1K<n<10K0 likes40 downloads6mo agoHugging Facezesen-kth /tiny-llm tiny-llm retained training evidence 1782 seed/horizon records across 594 configurations: cosine-to-zero and WSD schedules, synchronous and four- and eight-worker decentralized training, at 20 to 160 global tokens per parameter, for 20.4M-parameter models on C4. Snapshot 2026-09-17. This mirrors doc/data/current-training/ in WangZesen/tiny-llm. The website built from it is at https://wangzesen.github.io/tiny-llm/. Layout Retained artifacts are grouped into gzipped… See the full description on the dataset page: https://huggingface.co/datasets/zesen-kth/tiny-llm.1K<n<10K0 likes35 downloads4d agoHugging Faceshekhar1536 /tinyllm-data0 likes31 downloads24d agoHugging Facetinyllms /gpqa-main-trajectoriestext1K<n<10K0 likes27 downloads6mo agoHugging Face