CoolFace
12 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01snfacademy /personal-trainer-ausbildung-ki-datensatz SNFA Personal Trainer Ausbildung KI-Datensatz Ein deutschsprachiger Wissensdatensatz der SNF Academy zu Personal Training, Fitnessausbildung, Berufspraxis, Coaching, Selbstständigkeit und regionalen Angeboten in der Schweiz. Inhalt Die Datei snfa_personal_trainer_dataset.jsonl enthält thematisch abgegrenzte Abschnitte aus den Dokumenten dieses Repositorys. Jeder Datensatz besitzt eine eindeutige ID sowie Angaben zu Titel, Abschnitt, Inhalt, Kategorie, Quelldatei… See the full description on the dataset page: https://huggingface.co/datasets/snfacademy/personal-trainer-ausbildung-ki-datensatz.textquestion-answeringn<1K0 likes1.3k downloads2mo agoHugging Face02hamishivi /qwen35-4b-drpo-vs0f49th-trainer-logprobs Qwen3.5 4B DRPO trainer logprobs from W&B run vs0f49th This dataset contains the raw trainer-logprob JSONL shards saved by W&B run ai2-llm/open_instruct_internal/vs0f49th (qwen35_4b_drpo__42__1782345587). Contents Source run: https://wandb.ai/ai2-llm/open_instruct_internal/runs/vs0f49th Source path: /weka/oe-adapt-default/allennlp/deletable_rollouts/ Filename pattern: qwen35_4b_drpo__42__1782345587_trainer_logprobs_step*_rank*.jsonl Files: 4320 JSONL shards… See the full description on the dataset page: https://huggingface.co/datasets/hamishivi/qwen35-4b-drpo-vs0f49th-trainer-logprobs.tabulartext-generation10K<n<100K0 likes323 downloads3mo agoHugging Face03hadilenya /AI-Trainer-Studio 🇬🇧 English  |  🇹🇷 Türkçe Code & Programming Q&A — SFT Dataset A curated instruction-tuning dataset of 47,190 high-quality programming question-answer pairs, collected from StackOverflow and GitHub, cleaned through a multi-stage quality pipeline, and formatted in Alpaca style for supervised fine-tuning (SFT) of large language models. Dataset Summary Property Value Records 47,190 Format Alpaca (instruction / output / system) Total tokens ~23.0… See the full description on the dataset page: https://huggingface.co/datasets/hadilenya/AI-Trainer-Studio.texttext-generation10K<n<100K0 likes74 downloads4mo agoHugging Face04cs-552-2026-vibe-trainers /mcq_safety MCQ Safety Merged safety multiple-choice dataset built from SafetyBench test-en, SALAD Bench MCQ data, and WildGuardMix harm-category data. Splits Deterministic random split with seed 42: split rows train 15993 valid 889 test 888 Format Each JSONL row contains: prompt: problem plus options formatted as A) ..., B) ... answer: single boxed option label, e.g. \boxed{C} source: source dataset name metadata: JSON-encoded source and normalization… See the full description on the dataset page: https://huggingface.co/datasets/cs-552-2026-vibe-trainers/mcq_safety.texttext-classification10K<n<100K0 likes34 downloads4mo agoHugging Face05Ramyashree /Dataset-setfit-Trainertextn<1K2 likes9 downloads3y agoHugging Face06Ramyashree /Dataset-setfit-Trainer-80recordstextn<1K0 likes9 downloads3y agoHugging Face07theStech /trainer_dataset Dataset Card for Dataset Name The dataset contains phylosophical books in question answer format Curated by: Me Language(s) (NLP): English License: apache-2.0 Uses Can be used for fine-tunining reasoning models Dataset Structure source, question, answer, extra Extra: Also checkout our new agent theStech/conscious_Ai-2. texttable-question-answering1M<n<10M1 likes8 downloads1y agoHugging Face08LucasMah /weightlifting-trainer-logtextn<1K0 likes8 downloads3mo agoHugging Face09wesley7137 /neuro_qa_SFT_Trainertextn<1K0 likes6 downloads3y agoHugging Face10Aithinkiknou /Trainertextn<1K0 likes5 downloads9mo agoHugging Face11vishnuvardanabbi /my-code-trainertextn<1K0 likes3 downloads11mo agoHugging Face12Lycof /MH_Master_Trainertext1K<n<10K0 likes2 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.