CoolFace
20 results

supervised

llamafactory /tiny-supervised-datasettexttext-generationn<1K4 likes43k downloads2y agoHugging Facelightonai /nv-embed-supervised-distill-dedup-codeThis dataset is a collection of the CoIR training datasets. We mined 2048 negatives per queries using gte-modernbert-base in order and format the data in a query, documents, scores format so that anyone can perform nv-retriever type of filtering using their own threshold (and this is also the format knowledge distillation for PyLate). Notably, this dataset has been used to perform the fine-tuning of the state-of-the-art late interaction LateOn-Code models. The boilerplate used to fine-tune… See the full description on the dataset page: https://huggingface.co/datasets/lightonai/nv-embed-supervised-distill-dedup-code.text1M<n<10M7 likes1.7k downloads2mo agoHugging Facelightonai /nv-embed-supervised-distill-deduptext10M<n<100M0 likes1.6k downloads5mo agoHugging Facehuggingface-course /supervised-finetuning_quiz_student_responses4 likes1.2k downloads1h agoHugging Facear0cket1 /qwen3-30b-a3b-base-reasoning-sft-nemotron-math-v4-cot4k12k-500m-supervised Qwen3-30B-A3B Reasoning SFT Prepacked Nemotron Math v4 CoT 4k-12k This dataset is a train-ready, offline-prepacked SFT corpus for full supervised fine-tuning of Qwen/Qwen3-30B-A3B-Base into a math reasoning model. Source And Filtering Source dataset: nvidia/Nemotron-SFT-Math-v4 Source revision: a94e56aeddcf6e75d28c8bd210f40fa62309288d Source split: train Intended subset: cot Preferred source during selection: AoPS Length filter: 4,000 to 12,000 supervised… See the full description on the dataset page: https://huggingface.co/datasets/ar0cket1/qwen3-30b-a3b-base-reasoning-sft-nemotron-math-v4-cot4k12k-500m-supervised.1 likes1.2k downloads2mo agoHugging Facelightonai /nv-embed-supervised-distilltext10M<n<100M1 likes997 downloads11mo agoHugging Face