CoolFace
24 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01microsoft /SCBench SCBench [Paper] [Code] [Project Page] SCBench (SharedContextBench) is a comprehensive benchmark to evaluate efficient long-context methods in a KV cache-centric perspective, analyzing their performance across the full KV cache lifecycle (generation, compression, retrieval, and loading) in real-world scenarios where context memory (KV cache) is shared and reused across multiple requests. 🎯 Quick Start Load Data You can download and load the SCBench data… See the full description on the dataset page: https://huggingface.co/datasets/microsoft/SCBench.tabularn<1K11 likes2.2k downloads2y agoHugging Face02microsoft /Taskbench TaskBench: Benchmarking Large Language Models for Task Automation Introduction TaskBench is a benchmark for evaluating large language models (LLMs) on task automation. Task automation can be formulated into three critical stages: task decomposition, tool invocation, and parameter prediction. This complexity makes data collection and evaluation more challenging compared to common NLP tasks. To address this challenge, we propose a comprehensive evaluation framework… See the full description on the dataset page: https://huggingface.co/datasets/microsoft/Taskbench.tabular10K<n<100K38 likes1.3k downloads2y agoHugging Face03microsoft /bing_coronavirus_query_set Dataset Card for BingCoronavirusQuerySet Dataset Summary Please note that you can specify the start and end date of the data. You can get start and end dates from here: https://github.com/microsoft/BingCoronavirusQuerySet/tree/master/data/2020 example: load_dataset("bing_coronavirus_query_set", queries_by="state", start_date="2020-09-01", end_date="2020-09-30") You can also load the data by country by using queries_by="country". Supported Tasks and… See the full description on the dataset page: https://huggingface.co/datasets/microsoft/bing_coronavirus_query_set.tabulartext-classification100K<n<1M1 likes455 downloads3y agoHugging Face04microsoft /benchpress-score-matrix BenchPress Score Matrix This dataset contains the public model-by-benchmark score matrix used by BenchPress. The release includes the lossless audited JSON, benchmark cost evidence, flat model and benchmark metadata, one row per observed score, and the paper-canonical dense subset used in the BenchPress experiments. The source repository is microsoft/benchpress. Canonical artifacts data/llm_benchmark_data.json is the authoritative rich score-matrix artifact. It… See the full description on the dataset page: https://huggingface.co/datasets/microsoft/benchpress-score-matrix.tabulartabular-regressionn<1K2 likes247 downloads1mo agoHugging Face05microsoft /msr-acc-tae25 Microsoft Research - Accurate Chemistry Collection: Total Atomization Energies Description The Microsoft Research Accurate Chemistry Collection (MSR-ACC) provides a collection of accurate coupled cluster labels for training machine learning functionals. MSR-ACC/TAE25 comprising 73,040 total atomization energies at the CCSD(T)/CBS level obtained with the W1-F12 thermochemical protocol. The dataset is constructed to exhaustively cover the chemical space of closed-shell… See the full description on the dataset page: https://huggingface.co/datasets/microsoft/msr-acc-tae25.tabular10K<n<100K8 likes237 downloads5mo agoHugging Face06microsoft /tsptabularn<1K2 likes148 downloads1y agoHugging Face07microsoft /CoSAlign-Train CoSAlign-Train: A Large-Scale Synthetic Training Dataset for Controllable Safety Alignment Paper: Controllable Safety Alignment: Inference-Time Adaptation to Diverse Safety Requirements, published at ICLR 2025. Purpose: Training dataset for controllable safety alignment (CoSA) of large language models (LLMs), facilitating fine-grained inference-time adaptation to diverse safety requirements. Description: CoSAlign-Train is a large-scale, synthetic preference dataset designed for… See the full description on the dataset page: https://huggingface.co/datasets/microsoft/CoSAlign-Train.tabular100K<n<1M4 likes131 downloads1y agoHugging Face08microsoft /sattabularn<1K2 likes56 downloads1y agoHugging Face09OALL /details_microsoft__Phi-3-medium-4k-instruct Dataset Card for Evaluation run of microsoft/Phi-3-medium-4k-instruct Dataset automatically created during the evaluation run of model microsoft/Phi-3-medium-4k-instruct. The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_microsoft__Phi-3-medium-4k-instruct.tabular100K<n<1M0 likes21 downloads2y agoHugging Face10microsoft /libero-data-for-rhotabular100K<n<1M2 likes21 downloads1d agoHugging Face11math-extraction-comp /microsoft__Phi-3-medium-128k-instructtabular1K<n<10K0 likes15 downloads2y agoHugging Face12math-extraction-comp /microsoft__Phi-3-medium-4k-instructtabular1K<n<10K0 likes15 downloads2y agoHugging Face13math-extraction-comp /microsoft__Phi-3.5-MoE-instructtabular1K<n<10K0 likes13 downloads2y agoHugging Face14math-extraction-comp /microsoft__Phi-3-mini-128k-instructtabular1K<n<10K0 likes12 downloads2y agoHugging Face15math-extraction-comp /microsoft__phi-4tabular1K<n<10K0 likes11 downloads2y agoHugging Face16OALL /details_microsoft__Phi-3.5-mini-instruct Dataset Card for Evaluation run of microsoft/Phi-3.5-mini-instruct Dataset automatically created during the evaluation run of model microsoft/Phi-3.5-mini-instruct. The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_microsoft__Phi-3.5-mini-instruct.tabular100K<n<1M0 likes8 downloads2y agoHugging Face17OALL /details_microsoft__Phi-3-medium-128k-instruct Dataset Card for Evaluation run of microsoft/Phi-3-medium-128k-instruct Dataset automatically created during the evaluation run of model microsoft/Phi-3-medium-128k-instruct. The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_microsoft__Phi-3-medium-128k-instruct.tabular100K<n<1M0 likes8 downloads2y agoHugging Face18math-extraction-comp /microsoft__Phi-3-small-8k-instructtabular1K<n<10K0 likes8 downloads2y agoHugging Face19math-extraction-comp /microsoft__DialoGPT-mediumtabular1K<n<10K0 likes7 downloads2y agoHugging Face20math-extraction-comp /microsoft__Phi-3-small-128k-instructtabular1K<n<10K0 likes7 downloads2y agoHugging Face21math-extraction-comp /NyxKrage__Microsoft_Phi-4tabular1K<n<10K0 likes6 downloads2y agoHugging Face22math-extraction-comp /microsoft__Phi-3-mini-4k-instructtabular1K<n<10K0 likes6 downloads2y agoHugging Face23math-extraction-comp /microsoft__Phi-3.5-mini-instructtabular1K<n<10K0 likes5 downloads2y agoHugging Face24ainewtrend07 /Evaluation_microsoft-mpnet-basetabular10K<n<100K0 likes5 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.