CoolFace
20 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01laion /laions_got_talent_orpheus_snacSome Laion's Got Talent (https://huggingface.co/datasets/laion/laions_got_talent) voice snippets converted to snac tokens in the format of the Orpheus-TTS https://github.com/canopyai/Orpheus-TTS We converted the data into instructions like format. The snac data is in 7 token frame groups. See the Orpheus blog for more details: https://canopylabs.ai/model-releases We did not create the original dataset and are only providing snac token with minimal text instructions for ease of use. You must be… See the full description on the dataset page: https://huggingface.co/datasets/laion/laions_got_talent_orpheus_snac.text100K<n<1M0 likes180 downloads1y agoHugging Face0211-47 /Got_Agentic_AI_5k Got_Agentic_AI_5k A 5,000-example dataset to train LLMs into production-grade agentic assistants (“Angelic Agents”): high-agency, tool-aware, test-driven, and safety-first. This dataset focuses on the kinds of tasks real engineering teams and major AI developers care about: Diff-first coding patches and tests Planner–executor agent architectures Evals, monitoring, and rollback discipline Data engineering transforms with quality checks Incident postmortems and operational… See the full description on the dataset page: https://huggingface.co/datasets/11-47/Got_Agentic_AI_5k.text10K<n<100K5 likes105 downloads9mo agoHugging Face03gotime /VC-LLM-DatasetDue to the presence of harmful and toxic unsafe content in the fine-tuning data, a portion of the data is displayed. For the full data, please contact ignitesun@163.com. text1M<n<10M0 likes62 downloads2y agoHugging Face04gotzmann /instructionstext10K<n<100K0 likes35 downloads3y agoHugging Face05vericudebuget /Bible-responses-dataset-gotquestions Theology Question-Answer Dataset Description This dataset contains structured, human-generated content focused on theology, primarily sourced from the website GotQuestions. Each entry is formatted as a question (prompt) and a corresponding answer (response). The dataset is provided in JSON format and is intended for fine-tuning AI models, though it can be used for other purposes as well. The structure of the dataset is as follows: { "prompt": "What does it mean to… See the full description on the dataset page: https://huggingface.co/datasets/vericudebuget/Bible-responses-dataset-gotquestions.texttext-generation1K<n<10K6 likes28 downloads2y agoHugging Face06hash-map /got_qa_pairsgenerated some part by parsing html scripts and remaining using gemini api textquestion-answering100K<n<1M0 likes22 downloads8mo agoHugging Face07cudecanarim /goth-girl-friendsimage10K<n<100K1 likes14 downloads1y agoHugging Face08mindchain /ORCA_GOT_STYLE Dataset Card for Dataset Name Dataset Summary This dataset card aims to be a base template for new datasets. It has been generated using this raw template. Supported Tasks and Leaderboards [More Information Needed] Languages [More Information Needed] Dataset Structure Data Instances [More Information Needed] Data Fields [More Information Needed] Data Splits [More Information Needed] Dataset… See the full description on the dataset page: https://huggingface.co/datasets/mindchain/ORCA_GOT_STYLE.textn<1K1 likes13 downloads3y agoHugging Face09cudecanarim /gothic-slutsimage1K<n<10K5 likes13 downloads1y agoHugging Face1011-47 /Got_Science_28ktabular10K<n<100K1 likes11 downloads4mo agoHugging Face11oodeh /eco-gotest-trainingtext1K<n<10K0 likes8 downloads2y agoHugging Face12open-llm-leaderboard /GoToCompany__llama3-8b-cpt-sahabatai-v1-instruct-detailsgated Dataset Card for Evaluation run of GoToCompany/llama3-8b-cpt-sahabatai-v1-instruct Dataset automatically created during the evaluation run of model GoToCompany/llama3-8b-cpt-sahabatai-v1-instruct The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/GoToCompany__llama3-8b-cpt-sahabatai-v1-instruct-details.tabular10K<n<100K0 likes7 downloads2y agoHugging Face13oodeh /eco-goteststext100K<n<1M0 likes4 downloads2y agoHugging Face14oodeh /eco-gotest-TAGtext10K<n<100K0 likes4 downloads2y agoHugging Face15Atul0012803 /got_fivehuntextn<1K0 likes3 downloads2y agoHugging Face16cudecanarim /got-filledimage1K<n<10K0 likes3 downloads1y agoHugging Face17goten9004 /ConversationSentimenttextn<1K0 likes2 downloads2y agoHugging Face18goten9004 /ConversationalSentimentDatasettextn<1K0 likes2 downloads2y agoHugging Face19gottamyspadeamdz /sn96-key85textn<1K0 likes2 downloads1y agoHugging Face20jadhaj /go-test-datasettextn<1K0 likes1 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.