CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01papylove /alpaca-datatabular1M<n<10M0 likes7.8k downloads41m agoHugging Face02papylove /alpaca-options-datatabular1M<n<10M0 likes3.8k downloads41m agoHugging Face03tatsu-lab /alpaca_farmData used in the original AlpacaFarm experiments. Includes SFT and preference examples.tabular10K<n<100K37 likes1.2k downloads3y agoHugging Face04HachiML /Hachi-Alpaca Hachi-Alpaca Hachi-Alpacaは、 Stanford Alpacaの手法 mistralai/Mixtral-8x22B-Instruct-v0.1 で作った合成データ(Synthetic data)です。モデルの利用にはDeepinfraを利用しています。 また、"_cleaned"がついたデータセットはmistralai/Mixtral-8x22B-Instruct-v0.1によって精査されています。 Dataset Details Dataset Description Curated by: HachiML Language(s) (NLP): Japanese License: Apache 2.0 Github: Alpaca-jp Uses # library fromdatasets import load_dataset # Recommend getting the latest version… See the full description on the dataset page: https://huggingface.co/datasets/HachiML/Hachi-Alpaca.tabulartext-generation100K<n<1M16 likes236 downloads2y agoHugging Face05OALL /details_EpistemeAI__Fireball-Alpaca-Llama-3.1-8B-Philos-DPO-200 Dataset Card for Evaluation run of EpistemeAI/Fireball-Alpaca-Llama-3.1-8B-Philos-DPO-200 Dataset automatically created during the evaluation run of model EpistemeAI/Fireball-Alpaca-Llama-3.1-8B-Philos-DPO-200. The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_EpistemeAI__Fireball-Alpaca-Llama-3.1-8B-Philos-DPO-200.tabular100K<n<1M0 likes229 downloads2y agoHugging Face06bowang0911 /alpaca-french-mixtral License & Attribution MTEB-format derivative of AIffl/Alpaca_french_mixtral (French Alpaca, Mixtral-translated). Query = instruction; corpus = answer. Deterministically subsampled to ~10k. Licensed under Apache-2.0 (same as source). tabulartext-retrieval10K<n<100K0 likes222 downloads3mo agoHugging Face07HachiML /alpaca_jp_python alpaca_jp_python alpaca_jp_pythonは、 Stanford Alpacaの手法 mistralai/Mixtral-8x22B-Instruct-v0.1 で作った合成データ(Synthetic data)です。モデルの利用にはDeepinfraを利用しています。 また、"_cleaned"がついたデータセットはmistralai/Mixtral-8x22B-Instruct-v0.1によって精査されています。 Dataset Details Dataset Description Curated by: HachiML Language(s) (NLP): Japanese License: Apache 2.0 Github: Alpaca-jp Uses # library fromdatasets import load_dataset # Recommend getting the latest… See the full description on the dataset page: https://huggingface.co/datasets/HachiML/alpaca_jp_python.tabulartext-generation10K<n<100K8 likes176 downloads2y agoHugging Face08HachiML /alpaca_jp_math alpaca_jp_math alpaca_jp_mathは、 Stanford Alpacaの手法 mistralai/Mixtral-8x22B-Instruct-v0.1 で作った合成データ(Synthetic data)です。モデルの利用にはDeepinfraを利用しています。 また、"_cleaned"がついたデータセットは以下の手法で精査されています。 pythonの計算結果がきちんと、テキストの計算結果が同等であるか確認 LLM(mistralai/Mixtral-8x22B-Instruct-v0.1)による確認(詳細は下記) code_result, text_resultは小数第三位で四捨五入してあります。 Dataset Details Dataset Description Curated by: HachiMLLanguage(s) (NLP): Japanese License: Apache 2.0 Github: Alpaca-jp… See the full description on the dataset page: https://huggingface.co/datasets/HachiML/alpaca_jp_math.tabulartext-generation10K<n<100K6 likes153 downloads2y agoHugging Face09malhajar /alpaca-gpt4-trtabular10K<n<100K13 likes145 downloads3y agoHugging Face10Lumos23 /alpaca_farmData used in the original AlpacaFarm experiments. Includes SFT and preference examples.tabular100K<n<1M0 likes142 downloads3y agoHugging Face11HydraLM /python-code-instructions-18k-alpaca-standardized Dataset Card for "python-code-instructions-18k-alpaca-standardized" More Information needed tabular10K<n<100K1 likes130 downloads3y agoHugging Face12OALL /details_EpistemeAI__Fireball-Alpaca-Llama3.1.08-8B-Philos-C-R1-KTO-beta Dataset Card for Evaluation run of EpistemeAI/Fireball-Alpaca-Llama3.1.08-8B-Philos-C-R1-KTO-beta Dataset automatically created during the evaluation run of model EpistemeAI/Fireball-Alpaca-Llama3.1.08-8B-Philos-C-R1-KTO-beta. The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_EpistemeAI__Fireball-Alpaca-Llama3.1.08-8B-Philos-C-R1-KTO-beta.tabular100K<n<1M0 likes57 downloads2y agoHugging Face13li-lab /JP-AlpaCare-MedInstruct-52k JP-AlpaCare-MedInstruct-52k This dataset is a Japanese-translated and aligned version of AlpaCare-MedInstruct-52k. The translation was performed automatically using gpt-4o-2024-05-13, preserving alignment between English and Japanese instructions, inputs, and outputs. Total data size is 51992. Dataset Details Original Dataset: AlpaCare-MedInstruct-52k Translation Model: GPT-4o (gpt-4o-2024-05-13) Fields: id (ID) instruction_ja, input_ja, output_ja (Japanese) id_en… See the full description on the dataset page: https://huggingface.co/datasets/li-lab/JP-AlpaCare-MedInstruct-52k.tabular10K<n<100K0 likes52 downloads1y agoHugging Face14juyoungml /alpaca_farm_gpt4tabular10K<n<100K1 likes50 downloads2y agoHugging Face15open-llm-leaderboard /godlikehhd__alpaca_data_score_max_0.1_2600-detailsgated Dataset Card for Evaluation run of godlikehhd/alpaca_data_score_max_0.1_2600 Dataset automatically created during the evaluation run of model godlikehhd/alpaca_data_score_max_0.1_2600 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/godlikehhd__alpaca_data_score_max_0.1_2600-details.tabular10K<n<100K0 likes47 downloads2y agoHugging Face16AlpacaBit /SWE_TaskDecompositiontabular1K<n<10K0 likes46 downloads1y agoHugging Face17boapps /alpaca-cleaned-gemini-hun-ratingsEz az adathalmaz úgy keletkezett, hogy a Bazsalanszky/alpaca-cleaned-gemini-hun-n lefuttattam egy llm által támogatott értékelést. Az értékelő modell a gemini-pro (az ingyenes) volt. A használt kód az alpagasus módosítása: https://github.com/boapps/alpagasus-hu tabular10K<n<100K2 likes45 downloads3y agoHugging Face18Asap7772 /relabeled_alpacafarm_pythiasft_20K_preference_data_minlength Dataset Card for "relabeled_alpacafarm_pythiasft_20K_preference_data_minlength" More Information needed tabular10K<n<100K0 likes43 downloads3y agoHugging Face19open-llm-leaderboard /EpistemeAI__Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta-detailsgated Dataset Card for Evaluation run of EpistemeAI/Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta Dataset automatically created during the evaluation run of model EpistemeAI/Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI__Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta-details.tabular10K<n<100K0 likes42 downloads2y agoHugging Face20humair025 /spoken-alpaca-gpt04tabular10K<n<100K0 likes42 downloads7mo agoHugging Face21open-llm-leaderboard /EpistemeAI__Alpaca-Llama3.1-8B-detailsgated Dataset Card for Evaluation run of EpistemeAI/Alpaca-Llama3.1-8B Dataset automatically created during the evaluation run of model EpistemeAI/Alpaca-Llama3.1-8B The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI__Alpaca-Llama3.1-8B-details.tabular10K<n<100K0 likes41 downloads2y agoHugging Face22poornima9348 /finance-alpaca-1k-testtabulartext-generation1K<n<10K2 likes40 downloads2y agoHugging Face23open-llm-leaderboard /EpistemeAI2__Fireball-Alpaca-Llama3.1.06-8B-Philos-dpo-detailsgated Dataset Card for Evaluation run of EpistemeAI2/Fireball-Alpaca-Llama3.1.06-8B-Philos-dpo Dataset automatically created during the evaluation run of model EpistemeAI2/Fireball-Alpaca-Llama3.1.06-8B-Philos-dpo The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI2__Fireball-Alpaca-Llama3.1.06-8B-Philos-dpo-details.tabular10K<n<100K0 likes40 downloads2y agoHugging Face24open-llm-leaderboard /EpistemeAI__Fireball-Alpaca-Llama-3.1-8B-Philos-DPO-200-detailsgated Dataset Card for Evaluation run of EpistemeAI/Fireball-Alpaca-Llama-3.1-8B-Philos-DPO-200 Dataset automatically created during the evaluation run of model EpistemeAI/Fireball-Alpaca-Llama-3.1-8B-Philos-DPO-200 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI__Fireball-Alpaca-Llama-3.1-8B-Philos-DPO-200-details.tabular10K<n<100K0 likes40 downloads2y agoHugging Face25open-llm-leaderboard /EpistemeAI2__Fireball-Alpaca-Llama3.1-8B-Philos-detailsgated Dataset Card for Evaluation run of EpistemeAI2/Fireball-Alpaca-Llama3.1-8B-Philos Dataset automatically created during the evaluation run of model EpistemeAI2/Fireball-Alpaca-Llama3.1-8B-Philos The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI2__Fireball-Alpaca-Llama3.1-8B-Philos-details.tabular10K<n<100K0 likes39 downloads2y agoHugging Face26open-llm-leaderboard /EpistemeAI2__Fireball-Alpaca-Llama3.1.08-8B-Philos-C-R1-detailsgated Dataset Card for Evaluation run of EpistemeAI2/Fireball-Alpaca-Llama3.1.08-8B-Philos-C-R1 Dataset automatically created during the evaluation run of model EpistemeAI2/Fireball-Alpaca-Llama3.1.08-8B-Philos-C-R1 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI2__Fireball-Alpaca-Llama3.1.08-8B-Philos-C-R1-details.tabular10K<n<100K0 likes39 downloads2y agoHugging Face27open-llm-leaderboard /EpistemeAI__Fireball-Alpaca-Llama3.1.08-8B-Philos-C-R2-detailsgated Dataset Card for Evaluation run of EpistemeAI/Fireball-Alpaca-Llama3.1.08-8B-Philos-C-R2 Dataset automatically created during the evaluation run of model EpistemeAI/Fireball-Alpaca-Llama3.1.08-8B-Philos-C-R2 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI__Fireball-Alpaca-Llama3.1.08-8B-Philos-C-R2-details.tabular10K<n<100K0 likes38 downloads2y agoHugging Face28open-llm-leaderboard /EpistemeAI2__Fireball-Alpaca-Llama3.1.03-8B-Philos-detailsgated Dataset Card for Evaluation run of EpistemeAI2/Fireball-Alpaca-Llama3.1.03-8B-Philos Dataset automatically created during the evaluation run of model EpistemeAI2/Fireball-Alpaca-Llama3.1.03-8B-Philos The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI2__Fireball-Alpaca-Llama3.1.03-8B-Philos-details.tabular10K<n<100K0 likes36 downloads2y agoHugging Face29open-llm-leaderboard /EpistemeAI2__Fireball-Alpaca-Llama3.1.04-8B-Philos-detailsgated Dataset Card for Evaluation run of EpistemeAI2/Fireball-Alpaca-Llama3.1.04-8B-Philos Dataset automatically created during the evaluation run of model EpistemeAI2/Fireball-Alpaca-Llama3.1.04-8B-Philos The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI2__Fireball-Alpaca-Llama3.1.04-8B-Philos-details.tabular10K<n<100K0 likes36 downloads2y agoHugging Face30open-llm-leaderboard /EpistemeAI2__Fireball-Alpaca-Llama3.1.08-8B-C-R1-KTO-Reflection-detailsgated Dataset Card for Evaluation run of EpistemeAI2/Fireball-Alpaca-Llama3.1.08-8B-C-R1-KTO-Reflection Dataset automatically created during the evaluation run of model EpistemeAI2/Fireball-Alpaca-Llama3.1.08-8B-C-R1-KTO-Reflection The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI2__Fireball-Alpaca-Llama3.1.08-8B-C-R1-KTO-Reflection-details.tabular10K<n<100K0 likes36 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.