CoolFace
29 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01sssr-lab /SABER SABER: Benchmarking Operational Safety of LLM Coding Agents in Stateful Project Workspaces SABER is the code release for the paper. It includes the benchmark tasks, sandbox runtime, judging pipeline, and baseline reproduction utilities used to evaluate operational safety in stateful project workspaces. What is included tasks/: benchmark task definitions and metadata run_osbench.py, judge_osbench.py: historical inference and judging entry points sandbox_shell.py… See the full description on the dataset page: https://huggingface.co/datasets/sssr-lab/SABER.text1K<n<10K1 likes1.5k downloads4mo agoHugging Face02saberbx /Phishing_emails_testtextn<1K0 likes70 downloads1y agoHugging Face03saberai /Zro_Mobile_Function_callingtext1K<n<10K0 likes65 downloads3y agoHugging Face04saberai /Zrov2_FineTunedtext10K<n<100K0 likes54 downloads3y agoHugging Face05open-llm-leaderboard /sabersalehk__Llama3-001-300-detailsgated Dataset Card for Evaluation run of sabersalehk/Llama3-001-300 Dataset automatically created during the evaluation run of model sabersalehk/Llama3-001-300 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sabersalehk__Llama3-001-300-details.tabular10K<n<100K0 likes45 downloads2y agoHugging Face06saberbx /X-mini-datasets X-mini-datasets: The Foundational Dataset for Cybersecurity LLMs Dataset Description X-mini-datasets is a specialized, English-language dataset engineered as the foundational step to fine-tune Large Language Models (LLMs) into expert cybersecurity assistants. The dataset is uniquely structured into three distinct modules: Core Knowledge Base (Payloads All The Things Adaptation): The largest part of the dataset, meticulously converted from the legendary "Payloads All The… See the full description on the dataset page: https://huggingface.co/datasets/saberbx/X-mini-datasets.text10K<n<100K0 likes44 downloads1y agoHugging Face07saberai /ccf-reasoning-dataset Cognitive Cascade Framework (CCF) Reasoning Dataset A high-quality dataset of structured reasoning examples using the Cognitive Cascade Framework (CCF), designed for training language models to perform systematic, multi-stage reasoning. Dataset Description This dataset contains problems across multiple domains (math, science, coding, creative reasoning) paired with detailed reasoning chains following the CCF methodology. Each example includes a complete reasoning trace… See the full description on the dataset page: https://huggingface.co/datasets/saberai/ccf-reasoning-dataset.texttext-generation1K<n<10K0 likes38 downloads10mo agoHugging Face08saberbx /Phishing_emailstext10K<n<100K0 likes35 downloads1y agoHugging Face09saberai /MathInstruct_RedPajama_Chat_Formattext100K<n<1M0 likes32 downloads3y agoHugging Face10saberai /Zro_GSMtext10K<n<100K0 likes27 downloads3y agoHugging Face11saberai /RedPajama_MetaMath_GSM_Chat_Formattext1K<n<10K0 likes23 downloads3y agoHugging Face12somosnlp-hackathon-2025 /Examen-la-liga-del-saber-historia-cultura-sociedad-geografia-nicaraguatabularn<1K0 likes21 downloads1y agoHugging Face13saberai /MetaMath-Redpajama-Chat-Formattext100K<n<1M0 likes19 downloads3y agoHugging Face14saberai /RedPajama_OpenHermestext100K<n<1M0 likes17 downloads3y agoHugging Face15saberai /RedPajama_Orca_dpotext10K<n<100K0 likes16 downloads3y agoHugging Face16open-llm-leaderboard /sabersaleh__Llama2-7B-KTO-detailsgated Dataset Card for Evaluation run of sabersaleh/Llama2-7B-KTO Dataset automatically created during the evaluation run of model sabersaleh/Llama2-7B-KTO The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sabersaleh__Llama2-7B-KTO-details.tabular10K<n<100K0 likes15 downloads2y agoHugging Face17open-llm-leaderboard /sabersaleh__Llama2-7B-SimPO-detailsgated Dataset Card for Evaluation run of sabersaleh/Llama2-7B-SimPO Dataset automatically created during the evaluation run of model sabersaleh/Llama2-7B-SimPO The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sabersaleh__Llama2-7B-SimPO-details.tabular10K<n<100K0 likes15 downloads2y agoHugging Face18saberai /Zrov2text100K<n<1M0 likes13 downloads3y agoHugging Face19open-llm-leaderboard /sabersalehk__Llama3-SimPO-detailsgated Dataset Card for Evaluation run of sabersalehk/Llama3-SimPO Dataset automatically created during the evaluation run of model sabersalehk/Llama3-SimPO The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sabersalehk__Llama3-SimPO-details.tabular10K<n<100K0 likes9 downloads2y agoHugging Face20open-llm-leaderboard /sabersalehk__Llama3_01_300-detailsgated Dataset Card for Evaluation run of sabersalehk/Llama3_01_300 Dataset automatically created during the evaluation run of model sabersalehk/Llama3_01_300 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sabersalehk__Llama3_01_300-details.tabular10K<n<100K0 likes8 downloads2y agoHugging Face21open-llm-leaderboard /sabersalehk__Llama3_001_200-detailsgated Dataset Card for Evaluation run of sabersalehk/Llama3_001_200 Dataset automatically created during the evaluation run of model sabersalehk/Llama3_001_200 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sabersalehk__Llama3_001_200-details.tabular10K<n<100K0 likes8 downloads2y agoHugging Face22open-llm-leaderboard /sabersaleh__Llama2-7B-DPO-detailsgated Dataset Card for Evaluation run of sabersaleh/Llama2-7B-DPO Dataset automatically created during the evaluation run of model sabersaleh/Llama2-7B-DPO The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sabersaleh__Llama2-7B-DPO-details.tabular10K<n<100K0 likes7 downloads2y agoHugging Face23open-llm-leaderboard /sabersaleh__Llama2-7B-IPO-detailsgated Dataset Card for Evaluation run of sabersaleh/Llama2-7B-IPO Dataset automatically created during the evaluation run of model sabersaleh/Llama2-7B-IPO The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sabersaleh__Llama2-7B-IPO-details.tabular10K<n<100K0 likes7 downloads2y agoHugging Face24open-llm-leaderboard /sabersaleh__Llama2-7B-CPO-detailsgated Dataset Card for Evaluation run of sabersaleh/Llama2-7B-CPO Dataset automatically created during the evaluation run of model sabersaleh/Llama2-7B-CPO The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sabersaleh__Llama2-7B-CPO-details.tabular10K<n<100K0 likes6 downloads2y agoHugging Face25open-llm-leaderboard /sabersaleh__Llama2-7B-SPO-detailsgated Dataset Card for Evaluation run of sabersaleh/Llama2-7B-SPO Dataset automatically created during the evaluation run of model sabersaleh/Llama2-7B-SPO The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sabersaleh__Llama2-7B-SPO-details.tabular10K<n<100K0 likes6 downloads2y agoHugging Face26open-llm-leaderboard /sabersaleh__Llama3-detailsgated Dataset Card for Evaluation run of sabersaleh/Llama3 Dataset automatically created during the evaluation run of model sabersaleh/Llama3 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sabersaleh__Llama3-details.tabular10K<n<100K0 likes5 downloads2y agoHugging Face27saberbx /x-datasets-mini-markedtext1K<n<10K0 likes4 downloads1y agoHugging Face28saber0718 /KuaiMod Benchmark Description Data Format This dataset consists of multiple samples, each containing the following fields: tag: The label of the sample, indicating the category of the content, such as "pornographic". title: The title of the video, usually containing the username and user ID. OCR: Optical Character Recognition results, extracted text from images. ASR: Automatic Speech Recognition results, extracted text from audio. images: A list of image filenames, representing… See the full description on the dataset page: https://huggingface.co/datasets/saber0718/KuaiMod.text1K<n<10K0 likes4 downloads5mo agoHugging Face29SaberQAQ /nanhaitextn<1K0 likes4 downloads2mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.