supervised
LLM2Vec-Meta-Llama-3-8B-Instruct-mntp-supervisedsupervised_finetuning_hist0_is_question_switchboard_question_detection.json_bs32_lr0.000063gpt2-large-supervised-prompt-writing-i1-GGUFLLaMA-O1-Supervised-1129-GGUFtoloka_-_gpt2-large-supervised-prompt-writing-ggufgpt2-large-supervised-prompt-writing-GGUFwav2vec2-xls-r-300m-Korean-children-pronunciation-jamo-based-semi-supervised_V2LLM2Vec-Meta-Llama-3-8B-Instruct-mntp-supervised
tiny-supervised-datasetnv-embed-supervised-distill-dedup-codeThis dataset is a collection of the CoIR training datasets. We mined 2048 negatives per queries using gte-modernbert-base in order and format the data in a query, documents, scores format so that anyone can perform nv-retriever type of filtering using their own threshold (and this is also the format knowledge distillation for PyLate).
Notably, this dataset has been used to perform the fine-tuning of the state-of-the-art late interaction LateOn-Code models. The boilerplate used to fine-tune… See the full description on the dataset page: https://huggingface.co/datasets/lightonai/nv-embed-supervised-distill-dedup-code.nv-embed-supervised-distill-dedupsupervised-finetuning_quiz_student_responsesqwen3-30b-a3b-base-reasoning-sft-nemotron-math-v4-cot4k12k-500m-supervised
Qwen3-30B-A3B Reasoning SFT Prepacked Nemotron Math v4 CoT 4k-12k
This dataset is a train-ready, offline-prepacked SFT corpus for full supervised
fine-tuning of Qwen/Qwen3-30B-A3B-Base into a math reasoning model.
Source And Filtering
Source dataset: nvidia/Nemotron-SFT-Math-v4
Source revision: a94e56aeddcf6e75d28c8bd210f40fa62309288d
Source split: train
Intended subset: cot
Preferred source during selection: AoPS
Length filter: 4,000 to 12,000 supervised… See the full description on the dataset page: https://huggingface.co/datasets/ar0cket1/qwen3-30b-a3b-base-reasoning-sft-nemotron-math-v4-cot4k12k-500m-supervised.nv-embed-supervised-distill
