CoolFace
20 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01ldbb123 /Instruction-tuning_Datasetstext1M<n<10M0 likes262 downloads2y agoHugging Face02instruction-tuning-sd /cartoonization Instruction-prompted cartoonization dataset This dataset was created from 5000 images randomly sampled from the Imagenette dataset. For more details on how the dataset was created, check out this directory. Following figure depicts the data preparation workflow: Known limitations and biases The dataset was derived from Imagenette, which, in turn, was derived from ImageNet. So, naturally, this dataset inherits the limitations and biases of ImageNet.… See the full description on the dataset page: https://huggingface.co/datasets/instruction-tuning-sd/cartoonization.imageimage-to-image1K<n<10K21 likes168 downloads3y agoHugging Face03instruction-tuning-sd /low-level-image-proc Instruction-prompted low-level image processing dataset To construct this dataset, we took different number of samples from the following datasets for each task and constructed a single dataset with prompts added like so: Task Prompt Dataset Number of samples Deblurring “deblur the blurry image” REDS (train_blur and train_sharp) 1200 Deraining “derain the image” Rain13k 686 Denoising “denoise the noisy image” SIDD 8 Low-light image enhancement "enhance the… See the full description on the dataset page: https://huggingface.co/datasets/instruction-tuning-sd/low-level-image-proc.imageimage-to-image1K<n<10K9 likes141 downloads3y agoHugging Face04ChuGyouk /PubMedVision_InstructionTuning_VQA_splittedtext100K<n<1M1 likes87 downloads2y agoHugging Face05Fredithefish /Instruction-Tuning-with-GPT-4-RedPajama-Chat Instruction Tuning with GPT 4 RedPajama-Chat This dataset has been converted from the Instruction-Tuning-with-GPT-4 dataset for the purpose of fine-tuning the RedPajama-INCITE-Chat-3B-v1 model. About Instruction-Tuning-with-GPT-4 English Instruction-Following Data generated by GPT-4 using Alpaca prompts for fine-tuning LLMs. Usage and License Notices The data is intended and licensed for research use only. The dataset is CC BY NC 4.0 (allowing only… See the full description on the dataset page: https://huggingface.co/datasets/Fredithefish/Instruction-Tuning-with-GPT-4-RedPajama-Chat.textquestion-answering10K<n<100K6 likes45 downloads3y agoHugging Face06GyeongjuLee /instruction-tuning_empexc-3estext10K<n<100K0 likes33 downloads2y agoHugging Face07llama-lang-adapt /Formatted-GSM8K-FewShot-InstructionTuning-OneExample-Gemma2text10K<n<100K0 likes22 downloads2y agoHugging Face08GyeongjuLee /instruction-tuning_EMH-emotional-reactionstext10K<n<100K0 likes22 downloads2y agoHugging Face09llama-lang-adapt /Formatted-GSM8K-FewShot-InstructionTuning-OneExampletext10K<n<100K0 likes20 downloads2y agoHugging Face10GyeongjuLee /instruction-tuning_EMH-interpretationstext10K<n<100K0 likes16 downloads2y agoHugging Face11GyeongjuLee /instruction-tuning_EMH-explorationstext10K<n<100K0 likes12 downloads2y agoHugging Face12MatinaAI /instruction_tuning_datasetsgated 🧠 Persian Cultural Alignment Dataset for LLMs This repository contains a high-quality, instruction-following dataset for cultural alignment of large language models (LLMs) in the Persian language. The dataset is curated using hybrid strategies that incorporate culturally grounded generation, multi-turn dialogues, translation, and augmentation methods, making it suitable for SFT, DPO, RLHF, and alignment evaluation. 📚 Dataset Overview Domain Methods Used… See the full description on the dataset page: https://huggingface.co/datasets/MatinaAI/instruction_tuning_datasets.textquestion-answering1 likes12 downloads1y agoHugging Face13sagarshf /instruction-tuning-translatetext1K<n<10K0 likes11 downloads3y agoHugging Face14shuyuej /instruction_tuning_data 🚀 Load Dataset from datasets import load_dataset dataset = load_dataset("shuyuej/instruction_tuning_data", split='train') print(dataset) 📚 Dataset Description Data Size Link ChatDoctor 100K https://www.yunxiangli.top/ChatDoctor/ MedQA 10.2K https://huggingface.co/datasets/GBaker/MedQA-USMLE-4-options MedMCQA 183K https://huggingface.co/datasets/medmcqa PubmedQA 211K https://huggingface.co/datasets/pubmed_qa LiveQA 635… See the full description on the dataset page: https://huggingface.co/datasets/shuyuej/instruction_tuning_data.text100K<n<1M1 likes8 downloads2y agoHugging Face15aoUTlum /instruction-tuningtext100K<n<1M0 likes6 downloads4mo agoHugging Face16llama-lang-adapt /Formatted-GSM8K-FewShot-InstructionTuning-TwoExampletext10K<n<100K0 likes5 downloads2y agoHugging Face17shuyuej /Instruction_Tuning_Testingtext100K<n<1M1 likes4 downloads2y agoHugging Face18llama-lang-adapt /Formatted-GSM8K-FewShot-InstructionTuningtext10K<n<100K0 likes4 downloads2y agoHugging Face19GyeongjuLee /instruction-tuning_ecot_EMH-emotional-reactionstext1K<n<10K0 likes3 downloads2y agoHugging Face20khanhdang /instruction_tuning_datatext10K<n<100K0 likes1 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.