CoolFace
9 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01katielink /med-gemini-medqa-relabeled Med-Gemini MedQA Relabelling and Analysis This repository contains data and code corresponding to the MedQA relabelling performed as part of [1], specifically for the results in Figure 4b and appendix C.2. [1] Khaled Saab, Tao Tu, Wei-Hung Weng, Ryutaro Tanno, David Stutz, Ellery Wulczyn, Fan Zhang, Tim Strother, Chunjong Park, Elahe Vedadi, Juanma Zambrano Chaves, Szu-Yeu Hu, Mike Schaekermann, Aishwarya Kamath, Yong Cheng, David G.T. Barrett, Cathy Cheung, Basil… See the full description on the dataset page: https://huggingface.co/datasets/katielink/med-gemini-medqa-relabeled.tabular1K<n<10K12 likes84 downloads2y agoHugging Face02Rapidata /multilingual-llm-jokes-4o-claude-gemini Rapidata Generated Joke Preference Dataset We collected 1'000'000+ human opinions on the jokes generated by state-of-the-art LLMs to decide which model is the funniest. The labelers are shown a joke in their language and asked to answer 'Yes' or 'No' to the question 'Is this joke funny?'. It took us less than 5 days to get all of the responses. The jokes are evenly distributed across 5 languages: English, Arabic, Japanese, Vietnamese, Portuguese and across 4 model… See the full description on the dataset page: https://huggingface.co/datasets/Rapidata/multilingual-llm-jokes-4o-claude-gemini.tabular1K<n<10K14 likes65 downloads1y agoHugging Face03MIT-WAL /Gemini_3.1_202_Task_AI_Exposure_Scores Gemini 3.1 2026 Task AI Exposure Scores Dataset Summary This dataset contains task-level AI exposure labels for O*NET task statements. Each task is classified into one of four categories, E0, E1, E2, or E3, using an updated 2026 Agentic AI Exposure Rubric and a Gemini 3.1 Pro classification pipeline. The labels are designed to capture whether a task can be accelerated by a frontier agentic AI system directly, whether it would require deeper software integration, or… See the full description on the dataset page: https://huggingface.co/datasets/MIT-WAL/Gemini_3.1_202_Task_AI_Exposure_Scores.tabulartext-classification10K<n<100K0 likes37 downloads6mo agoHugging Face04RotgarSett /medical-google-chatgpt-gemini-source-overlap Google and AI Source Overlap Across 12 Medical Niches An open, reproducible US dataset comparing explicit ChatGPT and Gemini citations with paired Google organic Top 20 results across 12 medical niches and 432 frozen questions. Full study: https://rotgar.com/medical/resources/google-top-20-chatgpt-gemini-source-overlap Version DOI: https://doi.org/10.5281/zenodo.21850734 Version: 1.0 Fieldwork: August 7, 2026 Publication date: August 8, 2026 Market and language: United States… See the full description on the dataset page: https://huggingface.co/datasets/RotgarSett/medical-google-chatgpt-gemini-source-overlap.tabular1K<n<10K0 likes28 downloads25d agoHugging Face05HPC-Boys /AIME-2024-Gemini-2.5-Protabularn<1K0 likes15 downloads1y agoHugging Face06niranjanh123 /gemini_filtered_sft_traces_simplified_reasoningtabularreinforcement-learningn<1K0 likes11 downloads2mo agoHugging Face07cemig-ceia /fineweb-edu-gemini-annotations-portuguese-regressiontabular1K<n<10K0 likes10 downloads1y agoHugging Face08gbdncl /gemini-2.0-flash-lite-pneumonia-datasettabularn<1K0 likes7 downloads1y agoHugging Face09taesiri /PhotoEditBattleResults-Gemini-EXTgatedtabularn<1K0 likes1 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.