CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01papylove /alpaca-datatabular1M<n<10M0 likes7.9k downloads8h agoHugging Face02tatsu-lab /alpaca_farmData used in the original AlpacaFarm experiments. Includes SFT and preference examples.tabular10K<n<100K37 likes1.3k downloads3y agoHugging Face03HachiML /Hachi-Alpaca Hachi-Alpaca Hachi-Alpacaは、 Stanford Alpacaの手法 mistralai/Mixtral-8x22B-Instruct-v0.1 で作った合成データ(Synthetic data)です。モデルの利用にはDeepinfraを利用しています。 また、"_cleaned"がついたデータセットはmistralai/Mixtral-8x22B-Instruct-v0.1によって精査されています。 Dataset Details Dataset Description Curated by: HachiML Language(s) (NLP): Japanese License: Apache 2.0 Github: Alpaca-jp Uses # library fromdatasets import load_dataset # Recommend getting the latest version… See the full description on the dataset page: https://huggingface.co/datasets/HachiML/Hachi-Alpaca.tabulartext-generation100K<n<1M16 likes235 downloads2y agoHugging Face04bowang0911 /alpaca-french-mixtral License & Attribution MTEB-format derivative of AIffl/Alpaca_french_mixtral (French Alpaca, Mixtral-translated). Query = instruction; corpus = answer. Deterministically subsampled to ~10k. Licensed under Apache-2.0 (same as source). tabulartext-retrieval10K<n<100K0 likes229 downloads3mo agoHugging Face05OALL /details_EpistemeAI__Fireball-Alpaca-Llama-3.1-8B-Philos-DPO-200 Dataset Card for Evaluation run of EpistemeAI/Fireball-Alpaca-Llama-3.1-8B-Philos-DPO-200 Dataset automatically created during the evaluation run of model EpistemeAI/Fireball-Alpaca-Llama-3.1-8B-Philos-DPO-200. The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_EpistemeAI__Fireball-Alpaca-Llama-3.1-8B-Philos-DPO-200.tabular100K<n<1M0 likes226 downloads2y agoHugging Face06HachiML /alpaca_jp_python alpaca_jp_python alpaca_jp_pythonは、 Stanford Alpacaの手法 mistralai/Mixtral-8x22B-Instruct-v0.1 で作った合成データ(Synthetic data)です。モデルの利用にはDeepinfraを利用しています。 また、"_cleaned"がついたデータセットはmistralai/Mixtral-8x22B-Instruct-v0.1によって精査されています。 Dataset Details Dataset Description Curated by: HachiML Language(s) (NLP): Japanese License: Apache 2.0 Github: Alpaca-jp Uses # library fromdatasets import load_dataset # Recommend getting the latest… See the full description on the dataset page: https://huggingface.co/datasets/HachiML/alpaca_jp_python.tabulartext-generation10K<n<100K8 likes174 downloads2y agoHugging Face07Lumos23 /alpaca_farmData used in the original AlpacaFarm experiments. Includes SFT and preference examples.tabular100K<n<1M0 likes146 downloads3y agoHugging Face08HachiML /alpaca_jp_math alpaca_jp_math alpaca_jp_mathは、 Stanford Alpacaの手法 mistralai/Mixtral-8x22B-Instruct-v0.1 で作った合成データ(Synthetic data)です。モデルの利用にはDeepinfraを利用しています。 また、"_cleaned"がついたデータセットは以下の手法で精査されています。 pythonの計算結果がきちんと、テキストの計算結果が同等であるか確認 LLM(mistralai/Mixtral-8x22B-Instruct-v0.1)による確認(詳細は下記) code_result, text_resultは小数第三位で四捨五入してあります。 Dataset Details Dataset Description Curated by: HachiMLLanguage(s) (NLP): Japanese License: Apache 2.0 Github: Alpaca-jp… See the full description on the dataset page: https://huggingface.co/datasets/HachiML/alpaca_jp_math.tabulartext-generation10K<n<100K6 likes143 downloads2y agoHugging Face09malhajar /alpaca-gpt4-trtabular10K<n<100K13 likes141 downloads3y agoHugging Face10HydraLM /python-code-instructions-18k-alpaca-standardized Dataset Card for "python-code-instructions-18k-alpaca-standardized" More Information needed tabular10K<n<100K1 likes125 downloads3y agoHugging Face11OALL /details_EpistemeAI__Fireball-Alpaca-Llama3.1.08-8B-Philos-C-R1-KTO-beta Dataset Card for Evaluation run of EpistemeAI/Fireball-Alpaca-Llama3.1.08-8B-Philos-C-R1-KTO-beta Dataset automatically created during the evaluation run of model EpistemeAI/Fireball-Alpaca-Llama3.1.08-8B-Philos-C-R1-KTO-beta. The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_EpistemeAI__Fireball-Alpaca-Llama3.1.08-8B-Philos-C-R1-KTO-beta.tabular100K<n<1M0 likes55 downloads2y agoHugging Face12boapps /alpaca-cleaned-gemini-hun-ratingsEz az adathalmaz úgy keletkezett, hogy a Bazsalanszky/alpaca-cleaned-gemini-hun-n lefuttattam egy llm által támogatott értékelést. Az értékelő modell a gemini-pro (az ingyenes) volt. A használt kód az alpagasus módosítása: https://github.com/boapps/alpagasus-hu tabular10K<n<100K2 likes47 downloads3y agoHugging Face13open-llm-leaderboard /godlikehhd__alpaca_data_score_max_0.1_2600-detailsgated Dataset Card for Evaluation run of godlikehhd/alpaca_data_score_max_0.1_2600 Dataset automatically created during the evaluation run of model godlikehhd/alpaca_data_score_max_0.1_2600 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/godlikehhd__alpaca_data_score_max_0.1_2600-details.tabular10K<n<100K0 likes47 downloads2y agoHugging Face14Asap7772 /relabeled_alpacafarm_pythiasft_20K_preference_data_minlength Dataset Card for "relabeled_alpacafarm_pythiasft_20K_preference_data_minlength" More Information needed tabular10K<n<100K0 likes46 downloads3y agoHugging Face15juyoungml /alpaca_farm_gpt4tabular10K<n<100K1 likes46 downloads2y agoHugging Face16AlpacaBit /SWE_TaskDecompositiontabular1K<n<10K0 likes46 downloads1y agoHugging Face17poornima9348 /finance-alpaca-1k-testtabulartext-generation1K<n<10K2 likes41 downloads2y agoHugging Face18open-llm-leaderboard /EpistemeAI__Alpaca-Llama3.1-8B-detailsgated Dataset Card for Evaluation run of EpistemeAI/Alpaca-Llama3.1-8B Dataset automatically created during the evaluation run of model EpistemeAI/Alpaca-Llama3.1-8B The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI__Alpaca-Llama3.1-8B-details.tabular10K<n<100K0 likes41 downloads2y agoHugging Face19open-llm-leaderboard /EpistemeAI__Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta-detailsgated Dataset Card for Evaluation run of EpistemeAI/Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta Dataset automatically created during the evaluation run of model EpistemeAI/Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI__Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta-details.tabular10K<n<100K0 likes41 downloads2y agoHugging Face20li-lab /JP-AlpaCare-MedInstruct-52k JP-AlpaCare-MedInstruct-52k This dataset is a Japanese-translated and aligned version of AlpaCare-MedInstruct-52k. The translation was performed automatically using gpt-4o-2024-05-13, preserving alignment between English and Japanese instructions, inputs, and outputs. Total data size is 51992. Dataset Details Original Dataset: AlpaCare-MedInstruct-52k Translation Model: GPT-4o (gpt-4o-2024-05-13) Fields: id (ID) instruction_ja, input_ja, output_ja (Japanese) id_en… See the full description on the dataset page: https://huggingface.co/datasets/li-lab/JP-AlpaCare-MedInstruct-52k.tabular10K<n<100K0 likes41 downloads1y agoHugging Face21open-llm-leaderboard /EpistemeAI2__Fireball-Alpaca-Llama3.1.06-8B-Philos-dpo-detailsgated Dataset Card for Evaluation run of EpistemeAI2/Fireball-Alpaca-Llama3.1.06-8B-Philos-dpo Dataset automatically created during the evaluation run of model EpistemeAI2/Fireball-Alpaca-Llama3.1.06-8B-Philos-dpo The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI2__Fireball-Alpaca-Llama3.1.06-8B-Philos-dpo-details.tabular10K<n<100K0 likes40 downloads2y agoHugging Face22open-llm-leaderboard /EpistemeAI__Fireball-Alpaca-Llama-3.1-8B-Philos-DPO-200-detailsgated Dataset Card for Evaluation run of EpistemeAI/Fireball-Alpaca-Llama-3.1-8B-Philos-DPO-200 Dataset automatically created during the evaluation run of model EpistemeAI/Fireball-Alpaca-Llama-3.1-8B-Philos-DPO-200 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI__Fireball-Alpaca-Llama-3.1-8B-Philos-DPO-200-details.tabular10K<n<100K0 likes40 downloads2y agoHugging Face23humair025 /spoken-alpaca-gpt04tabular10K<n<100K0 likes40 downloads7mo agoHugging Face24Asap7772 /alpaca_skewexp_minlength_merged Dataset Card for "alpaca_skewexp_minlength_merged" More Information needed tabular10K<n<100K0 likes39 downloads3y agoHugging Face25open-llm-leaderboard /EpistemeAI2__Fireball-Alpaca-Llama3.1-8B-Philos-detailsgated Dataset Card for Evaluation run of EpistemeAI2/Fireball-Alpaca-Llama3.1-8B-Philos Dataset automatically created during the evaluation run of model EpistemeAI2/Fireball-Alpaca-Llama3.1-8B-Philos The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI2__Fireball-Alpaca-Llama3.1-8B-Philos-details.tabular10K<n<100K0 likes39 downloads2y agoHugging Face26open-llm-leaderboard /EpistemeAI2__Fireball-Alpaca-Llama3.1.08-8B-Philos-C-R1-detailsgated Dataset Card for Evaluation run of EpistemeAI2/Fireball-Alpaca-Llama3.1.08-8B-Philos-C-R1 Dataset automatically created during the evaluation run of model EpistemeAI2/Fireball-Alpaca-Llama3.1.08-8B-Philos-C-R1 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI2__Fireball-Alpaca-Llama3.1.08-8B-Philos-C-R1-details.tabular10K<n<100K0 likes39 downloads2y agoHugging Face27open-llm-leaderboard /EpistemeAI__Fireball-Alpaca-Llama3.1.08-8B-Philos-C-R2-detailsgated Dataset Card for Evaluation run of EpistemeAI/Fireball-Alpaca-Llama3.1.08-8B-Philos-C-R2 Dataset automatically created during the evaluation run of model EpistemeAI/Fireball-Alpaca-Llama3.1.08-8B-Philos-C-R2 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI__Fireball-Alpaca-Llama3.1.08-8B-Philos-C-R2-details.tabular10K<n<100K0 likes38 downloads2y agoHugging Face28Asap7772 /relabeled_alpacafarm_pythiasft_20K_preference_data Dataset Card for "relabeled_alpacafarm_pythiasft_20K_preference_data" More Information needed tabular10K<n<100K0 likes37 downloads3y agoHugging Face29Asap7772 /relabeled_alpacafarm_pythiasft_20K_preference_data_modelength Dataset Card for "relabeled_alpacafarm_pythiasft_20K_preference_data_modelength" More Information needed tabular10K<n<100K0 likes36 downloads3y agoHugging Face30open-llm-leaderboard /EpistemeAI2__Fireball-Alpaca-Llama3.1.03-8B-Philos-detailsgated Dataset Card for Evaluation run of EpistemeAI2/Fireball-Alpaca-Llama3.1.03-8B-Philos Dataset automatically created during the evaluation run of model EpistemeAI2/Fireball-Alpaca-Llama3.1.03-8B-Philos The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI2__Fireball-Alpaca-Llama3.1.03-8B-Philos-details.tabular10K<n<100K0 likes36 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.