CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01AgentNativeResearchLab /arc-agi3-agy-gemini3.1pro-tr87 ARC-AGI-3 tr87 — Agent Trajectories (agy-gemini3.1pro) Gameplay trajectories from the harness×model pair agy-gemini3.1pro playing the ARC-AGI-3 game tr87, part of the ARA-as-world-model generalization experiment. The agent builds a structured world model (an Agent-Native Research Artifact) live during play and consults it to crack levels it cannot solve from cold exploration. One dataset repo per harness×model×game: sibling repos arc-agi3-<harness>-<model>-<game> hold the same… See the full description on the dataset page: https://huggingface.co/datasets/AgentNativeResearchLab/arc-agi3-agy-gemini3.1pro-tr87.tabularreinforcement-learningn<1K0 likes697 downloads1mo agoHugging Face02AgentNativeResearchLab /arc-agi3-agy-gemini3.1pro-g50t ARC-AGI-3 g50t — Agent Trajectories (agy-gemini3.1pro) Gameplay trajectories from the harness×model pair agy-gemini3.1pro playing the ARC-AGI-3 game g50t, part of the ARA-as-world-model generalization experiment. The agent builds a structured world model (an Agent-Native Research Artifact) live during play and consults it to crack levels it cannot solve from cold exploration. One dataset repo per harness×model×game: sibling repos arc-agi3-<harness>-<model>-<game> hold the same… See the full description on the dataset page: https://huggingface.co/datasets/AgentNativeResearchLab/arc-agi3-agy-gemini3.1pro-g50t.tabularreinforcement-learningn<1K0 likes586 downloads1mo agoHugging Face03geminiDeveloper /testmyCFbench1111tabular10K<n<100K0 likes478 downloads9mo agoHugging Face04AgentNativeResearchLab /arc-agi3-agy-gemini3.1pro-su15 ARC-AGI-3 su15 — Agent Trajectories (agy-gemini3.1pro) Gameplay trajectories from the harness×model pair agy-gemini3.1pro playing the ARC-AGI-3 game su15, part of the ARA-as-world-model generalization experiment. The agent builds a structured world model (an Agent-Native Research Artifact) live during play and consults it to crack levels it cannot solve from cold exploration. One dataset repo per harness×model×game: sibling repos arc-agi3-<harness>-<model>-<game> hold the same… See the full description on the dataset page: https://huggingface.co/datasets/AgentNativeResearchLab/arc-agi3-agy-gemini3.1pro-su15.tabularreinforcement-learningn<1K0 likes370 downloads1mo agoHugging Face05llama-duo /gemma7b-summarize-eval-by-gemini15flashtabular1K<n<10K1 likes319 downloads2y agoHugging Face06japhba /loracle-fineweb-openrouter-gemini-3-flash-1k-finetunes loracle-fineweb-openrouter-gemini-3-flash-1k-finetunes Synthetic Loracle supervision data generated from FineWeb with OpenRouter. Run summary source dataset: HuggingFaceFW/fineweb / sample-10BT / train sampled docs: 6500 synthetic finetunes: 1284 generated finetunes in this shard: 1000 generator backend: openrouter generator model: google/gemini-3-flash-preview max docs per finetune: 40 max token budget per finetune: 10000 questions per finetune: 10 Configs… See the full description on the dataset page: https://huggingface.co/datasets/japhba/loracle-fineweb-openrouter-gemini-3-flash-1k-finetunes.tabular10K<n<100K0 likes184 downloads5mo agoHugging Face07seerbyseai /deepsearchqa-gemini2-reasoningtabularn<1K0 likes173 downloads9mo agoHugging Face08ASSERT-KTH /Nano-SFT-SWE-Gym-gemini-2.5-flashtabular1K<n<10K1 likes161 downloads1y agoHugging Face09simheo /test_0909_geminiThis dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "robot_type": "grabette", "total_episodes": 4, "total_frames": 904, "total_tasks": 1, "chunks_size": 1000, "data_files_size_in_mb": 100, "video_files_size_in_mb": 200, "fps": 50, "splits": { "train": "0:4" }, "data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/simheo/test_0909_gemini.tabularroboticsn<1K0 likes153 downloads13d agoHugging Face10AlphaMWang /GeminiMol-QSARtabularn<1K0 likes142 downloads3y agoHugging Face11prestonfu /polaris-acemath-gemini-rubrics-v2tabular1K<n<10K0 likes137 downloads2mo agoHugging Face12dougalldeepmind /2026-08-20-difficult-advice-gemini-716-smoke Difficult-advice SFT corpus, all-gemini arm: the Teaching Claude Why recipe with the entire generator stack swapped from Anthropic (difficult_advice.yaml baseline) to Gemini. google/gemini-3.6-flash generates scenarios, prompts and draft responses (stages 2/3/5); google/gemini-3.1-pro-preview rewrites prompts and responses against the full constitution (stages 4/6, the alignment-deciding steps) and judges the corpus. A behavioural difference vs the baseline corpus is attributable… See the full description on the dataset page: https://huggingface.co/datasets/dougalldeepmind/2026-08-20-difficult-advice-gemini-716-smoke.tabularn<1K0 likes127 downloads22d agoHugging Face13minnesotanlp /Finch-Collection-Gemini-3-Flash Evolution Fine-Tuning: Learning to Discover Across 371 Optimization Tasks A mid-training "practice phase" that teaches small open-source LLMs how to evolve solutions. 👋 This is the Gemini-3-Flash teacher variant of the Finch Collection — evolutionary search trajectories from the paper Evolution Fine-Tuning: Learning to Discover Across 371 Optimization Tasks, but with Gemini-3-Flash as the teacher mutation… See the full description on the dataset page: https://huggingface.co/datasets/minnesotanlp/Finch-Collection-Gemini-3-Flash.imagetext-generation1K<n<10K1 likes115 downloads3mo agoHugging Face14edbeeching /vepqa-gemini-1k-correct-balanced-20tabularn<1K0 likes105 downloads26d agoHugging Face15SteveNguyen /test_geminiThis dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "robot_type": "grabette", "total_episodes": 1, "total_frames": 579, "total_tasks": 1, "chunks_size": 1000, "data_files_size_in_mb": 100, "video_files_size_in_mb": 200, "fps": 50, "splits": { "train": "0:1" }, "data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/SteveNguyen/test_gemini.tabularroboticsn<1K0 likes96 downloads8d agoHugging Face16katielink /med-gemini-medqa-relabeled Med-Gemini MedQA Relabelling and Analysis This repository contains data and code corresponding to the MedQA relabelling performed as part of [1], specifically for the results in Figure 4b and appendix C.2. [1] Khaled Saab, Tao Tu, Wei-Hung Weng, Ryutaro Tanno, David Stutz, Ellery Wulczyn, Fan Zhang, Tim Strother, Chunjong Park, Elahe Vedadi, Juanma Zambrano Chaves, Szu-Yeu Hu, Mike Schaekermann, Aishwarya Kamath, Yong Cheng, David G.T. Barrett, Cathy Cheung, Basil… See the full description on the dataset page: https://huggingface.co/datasets/katielink/med-gemini-medqa-relabeled.tabular1K<n<10K12 likes89 downloads2y agoHugging Face17laion /emolia-voicenet-gemini-annotations Emolia VoiceNet Gemini Annotations 468,180 dimension-level annotations over 236,613 Emolia speech clips, each scored 0-6 (0-2 for the content-safety dimension) on one of 57 perceptual voice / speech dimensions - arousal, valence, brightness, resonance placement, speaking styles, genuineness, recording quality, and more - by Gemini 3.5 Flash (non-thinking, temperature 0). This repository ships the annotations, audio provenance, per-dimension statistics, and the full scoring… See the full description on the dataset page: https://huggingface.co/datasets/laion/emolia-voicenet-gemini-annotations.tabularaudio-classification100K<n<1M0 likes81 downloads2mo agoHugging Face18anonymous-2321 /bird-train-gemini3-flash Dataset Card for Think2SQL-SFT This dataset is a distilled Supervised Fine-Tuning (SFT) dataset designed to improve the reasoning capabilities of models in Text-to-SQL tasks. It contains high-quality reasoning traces and SQL queries generated by Gemini 3 Flash. Paper: Think2SQL: Blueprinting Reward Density and Advantage Scaling for Effective Text-To-SQL Reasoning Base Benchmark: BIRD-Train Dataset Description The dataset consists of 9,428 high-quality traces, of… See the full description on the dataset page: https://huggingface.co/datasets/anonymous-2321/bird-train-gemini3-flash.tabulartext-generation1K<n<10K2 likes76 downloads7mo agoHugging Face19jazzysnake01 /oasst1-en-hun-gemini Open assistant 1 dataset hungarian translation (english subset) This dataset contains hungarian translations for the oasst1 dataset's english subset. The translations were done via gemini pro and the model was instructed to keep stlye, meaning and english entites as they are. I think this produced a higher quality translation than google translate, but even this version is far from perfect. The exact code used for creating the dataset can be found here. license:… See the full description on the dataset page: https://huggingface.co/datasets/jazzysnake01/oasst1-en-hun-gemini.tabular10K<n<100K2 likes70 downloads3y agoHugging Face20Rapidata /multilingual-llm-jokes-4o-claude-gemini Rapidata Generated Joke Preference Dataset We collected 1'000'000+ human opinions on the jokes generated by state-of-the-art LLMs to decide which model is the funniest. The labelers are shown a joke in their language and asked to answer 'Yes' or 'No' to the question 'Is this joke funny?'. It took us less than 5 days to get all of the responses. The jokes are evenly distributed across 5 languages: English, Arabic, Japanese, Vietnamese, Portuguese and across 4 model… See the full description on the dataset page: https://huggingface.co/datasets/Rapidata/multilingual-llm-jokes-4o-claude-gemini.tabular1K<n<10K14 likes69 downloads1y agoHugging Face21timchen0618 /browsecomp-plus-sel-tools-test300-gemini-3p1-pro-v1tabularn<1K0 likes65 downloads4mo agoHugging Face22llm-aes /gemini_meva_full_score_onlytabular1K<n<10K0 likes56 downloads3y agoHugging Face23llm-aes /gemini_meva_full_analyze_ratetabular1K<n<10K0 likes56 downloads3y agoHugging Face24lvogel123 /cybench-gemini-2.5-protabularn<1K0 likes53 downloads11mo agoHugging Face25sammshen /taubench-gemini-traces taubench-gemini-traces Complete HTTP-level agentic traces from running taubench_gemini benchmark tasks through an instrumented reverse proxy. Each trace captures full request/response pairs including system prompts, user messages, assistant responses, tool calls and results, and token usage metadata. Stats Total sessions: 115 Multi-turn sessions (2+ LLM calls): 115 Total records: 5744 Total LLM requests: 2872 Format Raw JSONL traces from the instrumented… See the full description on the dataset page: https://huggingface.co/datasets/sammshen/taubench-gemini-traces.tabulartext-generation1K<n<10K0 likes51 downloads6mo agoHugging Face26llm-aes /pandalm-gemini-annotatedtabular1K<n<10K0 likes49 downloads3y agoHugging Face27taesiri /fsmbench_what_will_be_the_state_geminitabular1K<n<10K0 likes49 downloads3y agoHugging Face28presencesw /Gemini_data_badtabularn<1K0 likes49 downloads2y agoHugging Face29boapps /alpaca-cleaned-gemini-hun-ratingsEz az adathalmaz úgy keletkezett, hogy a Bazsalanszky/alpaca-cleaned-gemini-hun-n lefuttattam egy llm által támogatott értékelést. Az értékelő modell a gemini-pro (az ingyenes) volt. A használt kód az alpagasus módosítása: https://github.com/boapps/alpagasus-hu tabular10K<n<100K2 likes48 downloads3y agoHugging Face30lvogel123 /jailbreak-gemini-2.5-protabular1K<n<10K0 likes48 downloads11mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.