CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01HuggingFaceH4 /MATH-500 Dataset Card for MATH-500 This dataset contains a subset of 500 problems from the MATH benchmark that OpenAI created in their Let's Verify Step by Step paper. See their GitHub repo for the source file: https://github.com/openai/prm800k/tree/main?tab=readme-ov-file#math-splits texttext-generationn<1K332 likes223k downloads9mo agoHugging Face02HuggingFaceH4 /instruction-datasetThis is the blind eval dataset of high-quality, diverse, human-written instructions with demonstrations. We will be using this for step 3 evaluations in our RLHF pipeline. textn<1K66 likes7.5k downloads4y agoHugging Face03huggingface /transformers-metadata Transformers metadata text1K<n<10K43 likes3.2k downloads5h agoHugging Face04HuggingFaceTB /openstax_paragraphsTexbooks from openstax.org with their chapters, abstracts and sections. Sample: { "book_title":"World History Volume 1, to 1500", "language":"en", "chapters":[ { "title":"Preface", "abstract":"None", "sections":[ { "title":"About OpenStax", "paragraph":"OpenStax is part of Rice University, which is a 501(c)(3) nonprofit..." }, { "title":"About OpenStax Resources"… See the full description on the dataset page: https://huggingface.co/datasets/HuggingFaceTB/openstax_paragraphs.textn<1K6 likes2.9k downloads3y agoHugging Face05huggingface /diffusers-metadatatextn<1K35 likes2k downloads7h agoHugging Face06huggingface-course /codeparrot-ds-traintext100K<n<1M9 likes1.8k downloads5y agoHugging Face07huggingface-course /codeparrot-ds-validtext1K<n<10K3 likes1.1k downloads5y agoHugging Face08svc-huggingface /minerva-mathtextn<1K2 likes528 downloads2y agoHugging Face09huggingface /forensic-refusaltabularn<1K14 likes303 downloads2mo agoHugging Face10huggingface-projects /sd-multiplayer-dataTo access an image use the following Bucket URL: https://d26smi9133w0oo.cloudfront.net/ example: https://d26smi9133w0oo.cloudfront.net/room-7/1670520485-CZk4C72xBr5wPfTpwDAnG6-7648_7008-a-chicken-breaking-through-a-mirrornnotn.webp Bucket URL/key SQLite https://huggingface.co/datasets/huggingface-projects/sd-multiplayer-data/blob/main/rooms_data.db sqlite> PRAGMA table_info(rooms_data); 0|id|INTEGER|1||1 1|room_id|TEXT|1||0 2|uuid|TEXT|1||0 3|x|INTEGER|1||0 4|y|INTEGER|1||0 5|prompt|TEXT|1||0… See the full description on the dataset page: https://huggingface.co/datasets/huggingface-projects/sd-multiplayer-data.tabular100K<n<1M2 likes226 downloads4y agoHugging Face11TianfuXinqu /filesystem_huggingface_5053_cl6lee6m Support Ticket Triage Corpus Dataset ID: ZorakTriage94b837 Customer support ticket records with priority, status, and satisfaction annotations. textn<1K0 likes167 downloads1mo agoHugging Face12agentlans /HuggingFaceFW-finewiki-sample HuggingFaceFW/finewiki sample A uniformly randomized subset of HuggingFaceFW/finewiki, created to provide a smaller and more manageable dataset for analysis, fine-tuning, and benchmarking. Overview This sample includes Wikipedia articles from languages with more than one million pages. Sampling is performed uniformly at random instead of alphabetically to ensure unbiased representation. Language Inclusion Criteria Languages were selected based on page count and… See the full description on the dataset page: https://huggingface.co/datasets/agentlans/HuggingFaceFW-finewiki-sample.tabulartext-generation100K<n<1M0 likes164 downloads11mo agoHugging Face13GEM-submissions /lewtun__hugging-face-test-t5-base.outputs.json-36bf2a59__1645559101textn<1K0 likes145 downloads5y agoHugging Face14GEM-submissions /lewtun__hugging-face-test-t5-base.outputs.json-36bf2a59__1646052073 GEM Submission Submission name: Hugging Face test T5-base.outputs.json 36bf2a59 textn<1K0 likes145 downloads5y agoHugging Face15GEM-submissions /lewtun__hugging-face-test-t5-base.outputs.json-36bf2a59__1646049601textn<1K0 likes144 downloads5y agoHugging Face16HuggingFaceGECLM /data_feedbacktextn<1K0 likes142 downloads3y agoHugging Face17huggingface-projects /color-palettes-sdtext1K<n<10K16 likes138 downloads2y agoHugging Face18GEM-submissions /lewtun__hugging-face-test-t5-base.outputs.json-36bf2a59__1645800191textn<1K0 likes134 downloads5y agoHugging Face19GEM-submissions /lewtun__hugging-face-test-t5-base.outputs.json-36bf2a59__1646049378textn<1K0 likes134 downloads5y agoHugging Face20GEM-submissions /lewtun__hugging-face-test-t5-base.outputs.json-36bf2a59__1646049876textn<1K0 likes131 downloads5y agoHugging Face21GEM-submissions /lewtun__hugging-face-test-t5-base.outputs.json-36bf2a59__1645558682textn<1K0 likes130 downloads5y agoHugging Face22GEM-submissions /lewtun__hugging-face-test-t5-base.outputs.json-36bf2a59__1646050898textn<1K0 likes130 downloads5y agoHugging Face23GEM-submissions /lewtun__hugging-face-test-t5-base.outputs.json-36bf2a59__1646049424textn<1K0 likes129 downloads5y agoHugging Face24GEM-submissions /lewtun__hugging-face-test-t5-base.outputs.json-36bf2a59__1646051364textn<1K0 likes128 downloads5y agoHugging Face25HuggingFaceTB /math_tasksThis is a math benchmark data collection adapted from Qwen2.5-Math text10K<n<100K2 likes112 downloads2y agoHugging Face26open-llm-leaderboard /HuggingFaceH4__zephyr-7b-beta-detailsgated Dataset Card for Evaluation run of HuggingFaceH4/zephyr-7b-beta Dataset automatically created during the evaluation run of model HuggingFaceH4/zephyr-7b-beta The dataset is composed of 44 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/HuggingFaceH4__zephyr-7b-beta-details.tabular10K<n<100K0 likes105 downloads2y agoHugging Face27open-llm-leaderboard /HuggingFaceH4__zephyr-orpo-141b-A35b-v0.1-detailsgated Dataset Card for Evaluation run of HuggingFaceH4/zephyr-orpo-141b-A35b-v0.1 Dataset automatically created during the evaluation run of model HuggingFaceH4/zephyr-orpo-141b-A35b-v0.1 The dataset is composed of 44 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/HuggingFaceH4__zephyr-orpo-141b-A35b-v0.1-details.tabular10K<n<100K0 likes102 downloads2y agoHugging Face28HuggingFaceH4 /instruction-pilot-outputs-filteredtabularn<1K12 likes97 downloads4y agoHugging Face29TianfuXinqu /rail_12306_filesystem_word_huggingface_1588_travel_archive_4a6a2e High-Speed Rail Travel Forum Discussions Threads from the rail travel community forum, anonymized and curated for analysis. Contents 12,804 threads Languages: zh-CN Format: JSON Lines Fields thread_id title author_anon created_at replies views textn<1K0 likes80 downloads1mo agoHugging Face30TianfuXinqu /filesystem_huggingface_9816_customer_feedback_raw_nucfubxi Raw Customer Feedback Corpus Fresh export of anonymized customer feedback records collected from the company's product channels (mobile app, website, email, in-app). Each record contains a product reference, a star rating, the customer review text, the review date, the originating channel, and the current processing status. This is the source dataset for the CX analytics curation pipeline. texttext-classificationn<1K0 likes77 downloads1mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.