Seriki/Kubu-hai
Dataset Card for kubu-hai.model ๐ โโ๏ธ๐ค ^|D Look Ma, an instruction dataset that wasn't generated by GPTs! Dataset Summary kubu-hai is a high-quality dataset of 10,000 instructions and demonstrations created by skilled human annotators. This data can be used for supervised fine-tuning (SFT) to make language models follow instructions better. No Robots was modelled after the instruction dataset described in OpenAI's InstructGPT paper, and is comprised mostly ofโฆ See the full description on the dataset page: https://huggingface.co/datasets/Seriki/Kubu-hai.
Create bash.zs
Rename README.mdx to README.md
Rename README.rst to README.mdx
Rename README.md to README.rst
Create cli/dataset/install.sh
Create cli/install.sh (#4)
Update README.md (#3)
Rename you-should-be-using-LLVM_DEBUG.sh to you-should-be-using-LLVM_DEBUG.zsh (#2)
Upload 3 files
Create pipeline.ch
Rename Modelhai.ch.txt to Modelhai.cuh
Upload 5 files
Upload 10 files
Create Preset/promt.json
Rename data/train_sft-00000-of-00001.parquet to data/train/sft-00000-of-00001.parquet
Update README.md
Create Dataset.md
Duplicate from HuggingFaceH4/no_robots
