CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01flyingbagel /mmGQA mmGQA Full GQA dataset in mm-eval format (id, media, messages), covering all 10 splits: train_balanced, val_balanced, test_balanced, testdev_balanced, train, val, test, testdev, challenge, submission. See metadata.json for prompt template and upstream field mapping. image1M<n<10M0 likes1.8k downloads5mo agoHugging Face02Post-training-Data-Flywheel /gorilla-openfunctions-v1text10K<n<100K0 likes1.6k downloads2y agoHugging Face03Post-training-Data-Flywheel /Salesforce-xlam-function-calling-60ktext10K<n<100K0 likes784 downloads2y agoHugging Face04SLOP011 /flywire-fafb-connectome FlyWire FAFB v783 Connectome — GNN-ready package The complete proofread wiring diagram of an adult female Drosophila melanogaster brain — 139,255 neurons and their synaptic connections — repackaged as a ready-to-train graph dataset. Companion to the flywire-gnn Python package. This is a dataset packaging of two public, no-auth sources: File Contents Source connections.parquet (474 MB) 15,091,983 unique directed neuron→neuron pairs at ≥1 synapse: pre, post, syn_count… See the full description on the dataset page: https://huggingface.co/datasets/SLOP011/flywire-fafb-connectome.tabulargraph-ml10M<n<100M1 likes482 downloads8d agoHugging Face05Histochemichael /fly-sud-simulation FlyWire-informed odor-reward simulation: individual-behavior V4b The full predeclared validation FAILED. This is synthetic simulation data, not measured fly behavior or a quantitative reproduction of Kaun et al. (2011). Detailed results · Code and protocols Findings and limitations 256 independently seeded validation flies, four conditions (paired, unpaired, untrained, retrieval-DAN-silenced), two delays (30 min, 24 h), 32 flies per cell: 8 reciprocal replicate… See the full description on the dataset page: https://huggingface.co/datasets/Histochemichael/fly-sud-simulation.tabularn<1K0 likes282 downloads5d agoHugging Face06Post-training-Data-Flywheel /AutoIF-instruct-61ktext10K<n<100K18 likes257 downloads2y agoHugging Face07FlyaiLab /ecommerce_last_exam E-Commerce Last Exam A benchmark for evaluating LLM agents on 120 real-world travel planning and e-commerce tool-use tasks. Each task runs in an isolated Docker container with domain-specific CLI tools and SQLite databases. Agents must search, analyze, and produce structured recommendations. Repository: alibaba-flyai/ecommerce_last_exam Evaluation CLI: flyai-bench (pip install flyai-bench) Leaderboard: FlyaiLab/ecommerce_last_exam_leaderboard Dataset Summary… See the full description on the dataset page: https://huggingface.co/datasets/FlyaiLab/ecommerce_last_exam.textn<1K1 likes213 downloads23d agoHugging Face08Post-training-Data-Flywheel /AutoIF-instruct-61k-with-funcstext10K<n<100K8 likes184 downloads2y agoHugging Face09flyswot /iiif_snorkel_labelsimage100K<n<1M0 likes162 downloads4y agoHugging Face10flyingbugs /OpenR1-Math-220k-pruned-keep-0.9-end-start-0.5-correctnesstabular10K<n<100K0 likes147 downloads1y agoHugging Face11fl-ymd /so101_pickplace_block_1_lerobot_v2.1This dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v2.1", "robot_type": "so101_follower", "total_episodes": 50, "total_frames": 19499, "total_tasks": 1, "chunks_size": 1000, "data_files_size_in_mb": 100, "video_files_size_in_mb": 500, "fps": 30, "splits": { "train": "0:50" }, "data_path": "data/chunk-{chunk_index:03d}/episode_{episode_index:06d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/fl-ymd/so101_pickplace_block_1_lerobot_v2.1.tabularrobotics10K<n<100K0 likes134 downloads9mo agoHugging Face12Post-training-Data-Flywheel /OpenOrcatext1M<n<10M0 likes114 downloads2y agoHugging Face13flyingbugs /OpenR1-Math-220k-pruned-head-random-perturbationtext10K<n<100K0 likes114 downloads1y agoHugging Face14flyingbugs /OpenR1-Math-220k-pruned-think_midtext10K<n<100K0 likes106 downloads1y agoHugging Face15flyingbugs /OpenR1-Math-220k-pruned-keep-0.75-end-start-0.5text10K<n<100K0 likes92 downloads1y agoHugging Face16SamsungSDS-Research /Policy-on-the-Fly-Benchmark ⚠️ Content Warning: This dataset contains harmful content for AI filter evaluation, including biased expressions, crime-related scenarios, and jailbreak attempts. PoFBench: Policy-on-the-Fly Benchmark Overview PoFBench (Policy-on-the-Fly Benchmark) is a test-only benchmark designed to measure the performance of policy-based custom filters in LLM-powered systems. Existing AI safety benchmarks evaluate against fixed risk taxonomies predefined by experts. However, in… See the full description on the dataset page: https://huggingface.co/datasets/SamsungSDS-Research/Policy-on-the-Fly-Benchmark.texttext-classification1K<n<10K5 likes91 downloads5mo agoHugging Face17unionai /flyte-slack-data Dataset Card for "flyte-slack-data" More Information needed text10K<n<100K1 likes86 downloads3y agoHugging Face18Post-training-Data-Flywheel /NousResearch-hermes-function-calling-v1text1K<n<10K0 likes85 downloads2y agoHugging Face19flyingbugs /OpenR1-Math-220k-pruned-midtext10K<n<100K0 likes84 downloads11mo agoHugging Face20flyingbugs /OpenR1-Math-220k-pruned-middle-random-perturbationtext10K<n<100K0 likes82 downloads1y agoHugging Face21Post-training-Data-Flywheel /gorilla-apibenchtext10K<n<100K0 likes81 downloads2y agoHugging Face22broadfield-dev /python-codevec-flytech_python-codes-25ktabular10K<n<100K0 likes81 downloads7mo agoHugging Face23flyingbugs /OpenR1-Math-220k-pruned-keep-0.5-end-start-0.5-acc-increamentaltabular10K<n<100K0 likes77 downloads1y agoHugging Face24flyingbugs /OpenR1-Math-220k-pruned-keep-0.75-end-start-1.0text10K<n<100K0 likes77 downloads1y agoHugging Face25giantfish-fly /pi-llm PI-LLM Bench: The Core Retrieval Challenge Behind MRCR Update: Accepted to COLM 2026 (San Francisco). Moonshot AI (Kimi) PI-LLM is being observed internally for agent state tracking and robustness to context interference ICML 2025 Long-Context Foundation Models Workshop Accepted. AAAI 2026 Worshop Oral: LaMAS (LLM-based Multi-Agent Systems: Towards Responsible, Reliable, and Scalable Agentic Systems) A simple context interference evaluation. Update: This dataset is… See the full description on the dataset page: https://huggingface.co/datasets/giantfish-fly/pi-llm.tabularquestion-answeringn<1K2 likes76 downloads2mo agoHugging Face26flyingbugs /OpenR1-Math-220k-pruned-keep-0.2-end-start-0.5-acctabular10K<n<100K0 likes70 downloads1y agoHugging Face27flyingbugs /OpenR1-Math-220k-pruned-keep-0.5-end-start-0.5-add-aimetext10K<n<100K0 likes67 downloads1y agoHugging Face28formll /M-FLYT-input-scoresThis repository contains the input scores dataset used for training M-FLYT as described in the paper Filter Like You Test: Data-Driven Data Filtering for CLIP Pretraining. The scores are formatted as a parquet dataset, and can be used to reproduce our results or to improve them by adding more or better scoring methods. For code to use these scores and more information visit our GitHub repository. tabular100M<n<1B0 likes63 downloads2y agoHugging Face29sach0312 /flybrain-nfl-polymarket FlyBrain NFL x Polymarket dataset NFL regular-season games 2023-2025 matched to resolved Polymarket moneyline markets, with kickoff-eve market odds. pm_games.parquet: one row per game (schedule + scores + Polymarket market metadata + odds) p_eve: market probability of outcome0 winning ~24h before kickoff (CLOB prices-history) p_close: market probability at last trade before kickoff outcome0 = first slug token's team (outcome0_is_home marks its role); slug order is alphabetical… See the full description on the dataset page: https://huggingface.co/datasets/sach0312/flybrain-nfl-polymarket.tabularn<1K1 likes63 downloads13d agoHugging Face30Post-training-Data-Flywheel /stingning-ultrachattext1M<n<10M0 likes62 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.