CoolFace
15 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01beezza /ogiri-bokete-unsloth-vlm Japanese Bokete Ogiri — Unsloth VLM format YANS-official/ogiri-bokete を、UnslothのVision SFTで扱える会話形式に変換した非公開用データセットです。 各JSONLレコードは「1画像 + 1回答」です。 { "messages": [ {"role": "user", "content": [ {"type": "image", "image": "images/124469.jpg"}, {"type": "text", "text": "この画像のお題に対して、面白い一言を1つ返してください。"} ]}, {"role": "assistant", "content": [ {"type": "text", "text": "..."} ]} ] } Files train.jsonl: 1,678 records / 630 prompts… See the full description on the dataset page: https://huggingface.co/datasets/beezza/ogiri-bokete-unsloth-vlm.imageimage-to-text1K<n<10K0 likes50 downloads2mo agoHugging Face02open-llm-leaderboard /BEE-spoke-data__Meta-Llama-3-8Bee-detailsgated Dataset Card for Evaluation run of BEE-spoke-data/Meta-Llama-3-8Bee Dataset automatically created during the evaluation run of model BEE-spoke-data/Meta-Llama-3-8Bee The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/BEE-spoke-data__Meta-Llama-3-8Bee-details.tabular10K<n<100K0 likes21 downloads2y agoHugging Face03sbuedenb /big_beetle_datasettabular1M<n<10M0 likes13 downloads1y agoHugging Face04open-llm-leaderboard /BEE-spoke-data__smol_llama-220M-openhermes-detailsgated Dataset Card for Evaluation run of BEE-spoke-data/smol_llama-220M-openhermes Dataset automatically created during the evaluation run of model BEE-spoke-data/smol_llama-220M-openhermes The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/BEE-spoke-data__smol_llama-220M-openhermes-details.tabular10K<n<100K0 likes11 downloads2y agoHugging Face05open-llm-leaderboard /BEE-spoke-data__smol_llama-220M-GQA-detailsgated Dataset Card for Evaluation run of BEE-spoke-data/smol_llama-220M-GQA Dataset automatically created during the evaluation run of model BEE-spoke-data/smol_llama-220M-GQA The dataset is composed of 43 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/BEE-spoke-data__smol_llama-220M-GQA-details.tabular10K<n<100K0 likes10 downloads2y agoHugging Face06open-llm-leaderboard /BEE-spoke-data__smol_llama-101M-GQA-detailsgated Dataset Card for Evaluation run of BEE-spoke-data/smol_llama-101M-GQA Dataset automatically created during the evaluation run of model BEE-spoke-data/smol_llama-101M-GQA The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/BEE-spoke-data__smol_llama-101M-GQA-details.tabular10K<n<100K0 likes10 downloads2y agoHugging Face07open-llm-leaderboard /BEE-spoke-data__smol_llama-220M-GQA-fineweb_edu-detailsgated Dataset Card for Evaluation run of BEE-spoke-data/smol_llama-220M-GQA-fineweb_edu Dataset automatically created during the evaluation run of model BEE-spoke-data/smol_llama-220M-GQA-fineweb_edu The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/BEE-spoke-data__smol_llama-220M-GQA-fineweb_edu-details.tabular10K<n<100K0 likes8 downloads2y agoHugging Face08open-llm-leaderboard /BEE-spoke-data__tFINE-900m-e16-d32-flan-infinity-instruct-7m-T2T_en-1024-detailsgated Dataset Card for Evaluation run of BEE-spoke-data/tFINE-900m-e16-d32-flan-infinity-instruct-7m-T2T_en-1024 Dataset automatically created during the evaluation run of model BEE-spoke-data/tFINE-900m-e16-d32-flan-infinity-instruct-7m-T2T_en-1024 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/BEE-spoke-data__tFINE-900m-e16-d32-flan-infinity-instruct-7m-T2T_en-1024-details.tabular10K<n<100K0 likes7 downloads2y agoHugging Face09open-llm-leaderboard /BEE-spoke-data__tFINE-900m-instruct-orpo-detailsgated Dataset Card for Evaluation run of BEE-spoke-data/tFINE-900m-instruct-orpo Dataset automatically created during the evaluation run of model BEE-spoke-data/tFINE-900m-instruct-orpo The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/BEE-spoke-data__tFINE-900m-instruct-orpo-details.tabular10K<n<100K0 likes6 downloads2y agoHugging Face10sbuedenb /small_beetle_datasetThis dataset was produced by the Snakemake workflow in: https://github.com/songlab-cal/gpn/tree/main/workflow/make_dataset The following accessions are included in this dataset: Assembly Accession Assembly Name Organism Name GCF_031307605.1 icTriCast1.1 Tribolium castaneum GCF_963966145.1 icTenMoli1.1 Tenebrio molitor GCF_036711695.1 CSIRO_AGI_Zmor_V1 Zophobas morio GCF_015345945.1 Tmad_KSU_1.1 Tribolium madens The only adapted config is this: # this chroms are forced to… See the full description on the dataset page: https://huggingface.co/datasets/sbuedenb/small_beetle_dataset.tabulartext-generation1M<n<10M0 likes6 downloads1y agoHugging Face11open-llm-leaderboard /BEE-spoke-data__tFINE-900m-e16-d32-flan-detailsgated Dataset Card for Evaluation run of BEE-spoke-data/tFINE-900m-e16-d32-flan Dataset automatically created during the evaluation run of model BEE-spoke-data/tFINE-900m-e16-d32-flan The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/BEE-spoke-data__tFINE-900m-e16-d32-flan-details.tabular10K<n<100K0 likes5 downloads2y agoHugging Face12open-llm-leaderboard /BEE-spoke-data__tFINE-900m-e16-d32-instruct_2e-detailsgated Dataset Card for Evaluation run of BEE-spoke-data/tFINE-900m-e16-d32-instruct_2e Dataset automatically created during the evaluation run of model BEE-spoke-data/tFINE-900m-e16-d32-instruct_2e The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/BEE-spoke-data__tFINE-900m-e16-d32-instruct_2e-details.tabular10K<n<100K0 likes5 downloads2y agoHugging Face13sbuedenb /big_beetle_dataset-1024tabular1M<n<10M0 likes5 downloads1y agoHugging Face14sbuedenb /big_beetle_dataset-8192tabular1M<n<10M0 likes4 downloads1y agoHugging Face15sbuedenb /big_beetle_dataset-2048tabular1M<n<10M0 likes3 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.