CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01jinaai /code_exercises Dataset Card for "code_exercises" Code exercise This dataset is composed of a diverse set of ~120k Python code exercises (~120m total tokens) generated by ChatGPT 3.5. It is designed to distill ChatGPT 3.5 knowledge about Python coding tasks into other (potentially smaller) models. The exercises have been generated by following the steps described in the related GitHub repository. The generated exercises follow the format of the Human Eval benchmark. Each training sample… See the full description on the dataset page: https://huggingface.co/datasets/jinaai/code_exercises.texttext-generation1M<n<10M40 likes1.6k downloads3y agoHugging Face02RepDB /exercise-dataset Exercise Dataset — Free Tier (RepDB) A free, ready-to-use fitness exercise dataset: 601 exercises, each illustrated with flat-style 512×512 WebP images (a start/peak pose pair, or a single main pose for static holds and stretches), with target muscles, equipment, MET values, and full instructions in English, German, and Spanish. This public snapshot is the free tier of RepDB. Free for personal and commercial use inside applications, with attribution. Need exercise… See the full description on the dataset page: https://huggingface.co/datasets/RepDB/exercise-dataset.imagen<1K4 likes1.6k downloads17d agoHugging Face03SultanR /ultradata-math-textbook-exercise-ar ultradata-math-textbook-exercise-ar Arabic translation of the English portion of UltraData-Math, config UltraData-Math-L3-Textbook-Exercise-Synthetic: synthetic textbook-style content and exercises generated around specific mathematical knowledge points. Translated with the midtrans pipeline: text is segmented into prose and verbatim blocks (LaTeX, code, tables, and inline non-translatables are masked and never sent to the model, so formulas cannot be mangled), prose is… See the full description on the dataset page: https://huggingface.co/datasets/SultanR/ultradata-math-textbook-exercise-ar.texttext-generation10M<n<100M0 likes514 downloads1mo agoHugging Face04vibhuiitj /UltraData-Math-L3-Textbook-Exercise-Synthetic-split UltraData-Math L3 Textbook Exercise Synthetic Split Source dataset: openbmb/UltraData-Math Source config: UltraData-Math-L3-Textbook-Exercise-Synthetic Each row contains: uid question answer The original content field was split using the literal markers The exercise: and The solution:. texttext-generation10M<n<100M1 likes496 downloads6mo agoHugging Face05christian-bick /edugraph-exercises EduGraph Exercises Dataset EduGraph Exercises is a synthetic ML dataset of math-related visual problems, precisely labeled for training AI models in the education sector. Every image in this dataset is programmatically generated using the EduGraph Ontology to ensure that visual features are mathematically bound to their pedagogical labels. Quick Links Generation Engine: GitHub Repository (Contribute new generators or views!) Ontology: EduGraph Ontology (Semantic… See the full description on the dataset page: https://huggingface.co/datasets/christian-bick/edugraph-exercises.imagevisual-question-answering1K<n<10K1 likes365 downloads12d agoHugging Face06vibhuiitj /UltraData-Math-L3-Textbook-Exercise-Synthetic-split-qwen3-0.6b-embeddedtext1M<n<10M1 likes292 downloads6mo agoHugging Face07prithivMLmods /Gym-Exercise-Video-Analysis Gym-Exercise-Video-Analysis Gym-Exercise-Video-Analysis is a specialized multimodal video understanding dataset comprising 500 annotated gym workout and exercise clips. It is designed for fine-tuning and evaluating Video-Language Models (Video-LLMs), visual fitness coaches, and temporal exercise analysis systems. Each entry pairs exercise videos and extracted frame sequences with in-depth textual descriptions, biomechanical observations, form evaluations, and routine tracking.… See the full description on the dataset page: https://huggingface.co/datasets/prithivMLmods/Gym-Exercise-Video-Analysis.imagevideo-text-to-textn<1K2 likes239 downloads26d agoHugging Face08vibhuiitj /Exercise-Synthetic-split-ncert-chapter-mapped_filtered_difficulty_scoredtabular1M<n<10M0 likes173 downloads5mo agoHugging Face09vibhuiitj /UltraData-Math-L3-Textbook-Exercise-Synthetic-split-ncert-chapter-mappedtext1M<n<10M0 likes82 downloads6mo agoHugging Face10vibhuiitj /UltraData-Math-L3-Textbook-Exercise-Synthetic-split-ncert-chapter-mapped_filteredtext1M<n<10M0 likes55 downloads6mo agoHugging Face117rouz /complex-numbers-exercises-1000 📘 Exercices sur les Nombres Complexes – Dataset (1000 échantillons) Ce dataset contient 1000 exercices entièrement générés sur les nombres complexes, accompagnés de corrections détaillées et pédagogiques, prêts à être utilisés pour : l’entraînement de modèles d’IA éducatives, la génération automatique d’exercices, la correction automatique, l’explication pas-à-pas du raisonnement mathématique. Il s’inscrit dans un projet plus large visant à construire des IA spécialisées en… See the full description on the dataset page: https://huggingface.co/datasets/7rouz/complex-numbers-exercises-1000.textquestion-answeringn<1K0 likes50 downloads9mo agoHugging Face12Martjn /calisthenics_exercises Calisthenics Exercises Dataset A comprehensive dataset of 170 unique calisthenics exercises, each with three progression levels (beginner -> intermediate -> advanced). Web App Live: martjn-calisthenics-exercises.static.hf.space Interactive single-page app with 8 filter dimensions, favorites, keyboard shortcuts, and responsive design. Hosted on HuggingFace Spaces (static SDK). The Space is a thin loader that fetches index.html from this dataset repo at runtime — any… See the full description on the dataset page: https://huggingface.co/datasets/Martjn/calisthenics_exercises.texttext-generationn<1K0 likes50 downloads7mo agoHugging Face13vikp /coding_exercises_filtered Dataset Card for "coding_exercises_filtered" Coding exercises generated by gpt, then filtered. This has a lot of duplicates - would not recommend using as is. tabular10K<n<100K1 likes40 downloads3y agoHugging Face14RafaelJaime /calisthenics_exercisestextn<1K4 likes38 downloads2y agoHugging Face15seonjeongh /KLUE-MRC-exercisetext10K<n<100K0 likes37 downloads2y agoHugging Face16ZixuanKe /cfa_extracted_exercise_sup_sample_from_policy_v1.1_genrm_qwen3-32b_dpo_binarized_filtered_2048tabular10K<n<100K0 likes36 downloads1y agoHugging Face17natzx94 /exercise-api Exercise API — Dataset Dataset de 104 ejercicios de gimnasio (bilingüe ES/EN) derivado de la Exercise API. Cada ejercicio incluye grupo muscular, equipamiento, músculos principal/secundario, instrucciones paso a paso e ilustración masculina y femenina (208 imágenes en total). Configuraciones images — 1 fila por imagen (208). Etiquetas (grupo, equipamiento, músculos, género) + caption_es/caption_en. Para clasificación de imagen y multimodal (image-to-text / VQA).… See the full description on the dataset page: https://huggingface.co/datasets/natzx94/exercise-api.imageimage-classification1K<n<10K0 likes36 downloads2mo agoHugging Face18ZixuanKe /cfa_extracted_exercise_sup_sample_from_policy_v1.1_genrm_qwq-32b_dpo_binarizedtext10K<n<100K0 likes32 downloads1y agoHugging Face19tuandunghcmut /translation-de-en-exercisetext100K<n<1M0 likes31 downloads1y agoHugging Face20onurSakar /GYM-Exercisetext1K<n<10K10 likes30 downloads3y agoHugging Face21ZixuanKe /cfa_extracted_exercise_sup_sample_from_policy_v1.1_stepwise_dpo_binarized_chunk_20textn<1K0 likes29 downloads2y agoHugging Face22ZixuanKe /cfa_extracted_exercise_sup_sample_from_policy_v1.1_genrm_qwen3-32b_stepwise_dpo_binarizedtabular10K<n<100K0 likes29 downloads1y agoHugging Face23ZixuanKe /cfa_extracted_exercise_sup_sample_from_policy_v1.1_stepwise_dpo_binarized_chunk_13textn<1K0 likes27 downloads2y agoHugging Face24ZixuanKe /cfa_extracted_exercise_sup_sample_from_policy_v1.1_genrm_qwen3-32b_dpo_binarized_filter2048tabular10K<n<100K0 likes26 downloads1y agoHugging Face25ZixuanKe /cfa_extracted_exercise_sup_sample_from_policy_v1_1_rpo_stepwise_iter_1_dpo_binarizedtext1K<n<10K0 likes25 downloads2y agoHugging Face26maher13 /news_2026_exercise news_2026_exercise Arabic news dataset for text classification. Column Description title News title content News body category Label: سياسة, اقتصاد, صحة, رياضة 28,000 rows (7,000 per category). Download and load as pandas from datasets import load_dataset import pandas as pd ds = load_dataset("maher13/news_2026_exercise") df = ds["train"].to_pandas() print(df.head()) Sample N rows from each category from datasets import… See the full description on the dataset page: https://huggingface.co/datasets/maher13/news_2026_exercise.texttext-classification10K<n<100K0 likes25 downloads2mo agoHugging Face27ZixuanKe /cfa_extracted_exercise_sup_sample_from_policy_v1.1_stepwise_dpo_binarizedtabular10K<n<100K0 likes23 downloads2y agoHugging Face28ZixuanKe /cfa_extracted_exercise_sup_sample_from_policy_v1.1_genrm_qwen3-32b_stepwise_dpo_binarized_F2048tabular10K<n<100K0 likes23 downloads1y agoHugging Face29dexaai /huberman_on_exercise Dataset Card for "huberman_on_exercise" More Information needed textn<1K5 likes22 downloads3y agoHugging Face30ZixuanKe /cfa_extracted_exercise_sup_sample_from_policy_v1_1_rpo_stepwise_iter_1_dpo_val_chunk_18textn<1K0 likes22 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.