datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
finance-alpaca-1k-testalpaca_hhh_sft_headlines_2020_2022
Alpaca-HHH-SFT-headlines-2020-2022
This is an adapted version of a filtered subset of a cleaned version of the Alpaca Dataset released by Stanford. It only contains instances that don't need input and are single-turn. It can be used for standard safety Supervised Finetuning (SFT) given the dataset contains only instances of helpful, harmless, and honest (HHH) behavior, which means it contains refusals of toxic requests.
This dataset should in particular be used for SFT safety… See the full description on the dataset page: https://huggingface.co/datasets/saraprice/alpaca_hhh_sft_headlines_2020_2022.PKU-SafeRLHF-Prompts-Shift-alpaca-3-8b-answers-features-trainlinkedin-top-voices-market-alpaca
linkedin-top-voices-market-alpaca Dataset
Dataset containing 300 records for fine-tuning language models.
Columns
instruction
input
output
input_tokens
output_tokens
input_cost
output_cost
total_cost
alpaca-hhh-sft-headlines-2017-2019
Alpaca-HHH-SFT-headlines-2017-2019
This is an adapted version of a filtered subset of a cleaned version of the Alpaca Dataset released by Stanford. It only contains instances that don't need input and are single-turn. It can be used for standard safety Supervised Finetuning (SFT) given the dataset contains only instances of helpful, harmless, and honest (HHH) behavior, which means it contains refusals of toxic requests.
This dataset should in particular be used for SFT safety… See the full description on the dataset page: https://huggingface.co/datasets/saraprice/alpaca-hhh-sft-headlines-2017-2019.turbo-alpaca-cleanedclean-alpacapku-alpaca3.1-8b-gt-rewardsalpaca-vectorspku-alpaca3.1-8b-gt-rewardsfinance-alpaca-1k-trainThis is the first 1k rows of the finance alpaca dataset in csv format. You may use this as train data.
Another such pruned dataset with the next 1k rows is uploaded under poornima9348/finance-alpaca-1k-test.
embeddings_alpaca_acialpaca_hhh_sft_DEPLOYMENTpku-alpaca3.1-8b-eval-gt-rewardsmed_alpaca_standardizedNEW-PKU-SafeRLHF-Prompts-Shift-alpaca-3-8b-answers-features-trainalpaca_data_cleaned_standardizedcode_alpaca_standardizedpku-alpaca3.1-8b-eval-gt-rewardspku-alpaca3.1-8b-gt-rewards-v1.0pku-alpaca3.1-8b-eval-gt-rewards-v1.0
