CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Sachin21112004 /news-entertainment-datasettexttable-question-answeringn<1K6 likes934 downloads3h agoHugging Face02Sachin21112004 /news-politics-datasettext1K<n<10K0 likes698 downloads3h agoHugging Face03Sachin21112004 /news-education-datasettext10K<n<100K3 likes645 downloads3h agoHugging Face04Sachin21112004 /news-tech-datasettext10K<n<100K1 likes634 downloads3h agoHugging Face05Sachin21112004 /news-finance-datasettext10K<n<100K3 likes602 downloads3h agoHugging Face06faridganbarli /wire_harness_expert_sac Wire Harness Expert SAC Expert-policy trajectories collected from the five-mover WireHarness MuJoCo environment for visual world-model training. Dataset summary 20,000 episodes 3,491,570 stored observation rows At most 300 environment transitions per episode (up to 301 stored rows, including the initial observation) 224 x 224 RGB observations, stored as JPEG bytes in pixels 10-dimensional continuous actions 451-dimensional observations Five task stages and… See the full description on the dataset page: https://huggingface.co/datasets/faridganbarli/wire_harness_expert_sac.tabularreinforcement-learningn<1K1 likes125 downloads23d agoHugging Face07facebook /SACo-Goldgated Dataset Card for SA-Co/Gold SA-Co/Gold is a benchmark for promptable concept segmentation (PCS) in images. The benchmark contains images paired with text labels (also referred as Noun Phrases aka NPs), each annotated exhaustively with masks on all object instances that match the label. SA-Co/Gold comprises 7 subsets, each targeting a different annotation domain. For each subset, the annotations are multi-reviewed and agreed by 3 human annotators resulting in a high-quality… See the full description on the dataset page: https://huggingface.co/datasets/facebook/SACo-Gold.textn<1K27 likes116 downloads10mo agoHugging Face08Sachin-NK /GRASS_sampletext10K<n<100K1 likes84 downloads1y agoHugging Face09facebook /SACo-VEvalgated SA-Co/VEval Dataset License each domain has its own License SA-Co/VEval - SA-V: CC-BY-NC 4.0 SA-Co/VEval - YT-Temporal-1B: CC-BY-NC 4.0 SA-Co/VEval - SmartGlasses: CC-by-4.0 SA-Co/VEval is an evaluation dataset comprising of 3 domains, each domain has a val and test split. SA-Co/VEval - SA-V: videos are from the SA-V dataset SA-Co/VEval - YT-Temporal-1B: videos are from the YT-Temporal-1B SA-Co/VEval - SmartGlasses: egocentric videos from Smart Glasses This Hugging Face dataset… See the full description on the dataset page: https://huggingface.co/datasets/facebook/SACo-VEval.textn<1K4 likes69 downloads10mo agoHugging Face10SachinSaud /glassformingtextn<1K0 likes52 downloads5mo agoHugging Face11Elessar123 /SAC-Flow SAC Flow: Sample-Efficient Reinforcement Learning of Flow-Based Policies via velocity-reparameterized sequential modeling Overview SAC Flow is a stable, sample-efficient, and high-performance off-policy RL algorithm for flow-based policies. SAC Flow treats the flow-based model as a sequential model and reparameterizes its velocity network as a GRU or a Transformer. Get Start All necessary dependencies and environment setup steps are detailed in our… See the full description on the dataset page: https://huggingface.co/datasets/Elessar123/SAC-Flow.documentn<1K0 likes40 downloads1y agoHugging Face12sabin1234 /SAC_Nepal_FAQ_Nepali_Health_Fitness_Dataset SAC Nepal FAQ — Nepali Health & Fitness Dataset Overview This dataset (sac_nepal_faq_nepali.jsonl) is a collection of 100 instruction-following conversation pairs in Nepali, covering frequently asked questions about health, fitness, and nutrition. Each record is a single-turn human↔gpt exchange: a Nepali-language question followed by an informative Nepali-language answer. The data appears to be a localized/translated set — the behavior_definition field for every… See the full description on the dataset page: https://huggingface.co/datasets/sabin1234/SAC_Nepal_FAQ_Nepali_Health_Fitness_Dataset.textn<1K0 likes40 downloads15d agoHugging Face13Umkho-AI /SA_Cultural_Tribal_Practices SA Tribal & Cultural Practices Dataset Author: Minah Mojela (@minahmojela), Umkho-AI Dataset Summary This dataset contains 127 structured records documenting the cultural practices, customs, and identity histories of South Africa's major ethnic and population groups. It is a companion release to the South African History Dataset, built for the same reason: most AI models describe South African cultural practices using surface-level, externally-authored sources… See the full description on the dataset page: https://huggingface.co/datasets/Umkho-AI/SA_Cultural_Tribal_Practices.texttext-generationn<1K0 likes23 downloads2mo agoHugging Face14onenoly11 /oinio-sacred-trinity-eval 🔮 Quantum Forge Sacred Trinity Evaluation Dataset Annotated test cases for evaluating AI agents in the Quantum Pi Forge ecosystem. 📊 Dataset Description This dataset contains 10 annotated query-response pairs designed to evaluate AI agents operating within the Sacred Trinity architecture: FastAPI Quantum Conduit - Authentication, WebSocket, database operations Flask Glyph Weaver - Dashboard visualization, SVG cascade animations Gradio Truth Mirror - Ethical auditing… See the full description on the dataset page: https://huggingface.co/datasets/onenoly11/oinio-sacred-trinity-eval.texttext-generationn<1K0 likes21 downloads9mo agoHugging Face15sacrificialpancakes /synthetic_demographics_seed Synthetic Demographic Seeds v1 This is a dataset of 3,541,040 roughly demographically correct demographic seeds and somewhat demographically accurate names all generated from publicly available datasets. (note there were tradeoffs made with accuracy and what I could tie together, v2 will be more accurate) get_synthetic_demographics.py contains a method for quickly and randomly selecting batches of demographic seeds. There is no filtering on this at the moment. Format… See the full description on the dataset page: https://huggingface.co/datasets/sacrificialpancakes/synthetic_demographics_seed.tabular1M<n<10M3 likes20 downloads2y agoHugging Face16DGal1 /sackcha_chattext1K<n<10K0 likes17 downloads2y agoHugging Face17sachit-sankhe /Mentoring-Dataset Boost Your Technical Mentorship with OpenLLaMA 3B Fine-Tuning Ready to unlock expert-level guidance on your technical journey? Explore this question-answer dataset designed for technical mentorship, with future plans to fine-tune the powerful OpenLLaMA 3B language model for even more advanced interactions. Overview Focus: Technical Mentorship Domains: Currently covers 7 key areas: AI, ML, Blockchain, Cybersecurity, AppDev, WebDev, DevOps Content: General questions a… See the full description on the dataset page: https://huggingface.co/datasets/sachit-sankhe/Mentoring-Dataset.textn<1K0 likes16 downloads3y agoHugging Face18922-CA /ls2_09062023_test1_raw_SaChA_1a Sayori Chat 09062023 raw Dataset of Sayori dialogue from DDLC (dataset of ~600 items augmented by MythoMax-l2-13b to turn into multi-turn chat dialogue) Curated version planned text1K<n<10K0 likes15 downloads3y agoHugging Face19as-cle-bert /saccaromyces-cerevisiae-basetextn<1K1 likes15 downloads2y agoHugging Face20vsachi /sachi-dataset-jaLLMをファインチューニングするためのデータセットです。alcapa-chatbot-formatです。 キャラクターと会話するデータセットとなっています。 私はいつもVR SNSでかわいい女の子のロールプレーをしています。 私がかわいい女の子のAIに転生したという設定で作った会話データセットになっています。 キャラクター設定はフィクションやジョークです。完全に現実ではありません。 ゲームに登場するNPC等のAIのトレーニングなどに自由にご利用ください。 幅広く利用してもらえるようにPublic domainライセンスにします。 ライセンス Public domainライセンスにします。 textquestion-answeringn<1K0 likes15 downloads1y agoHugging Face21Sachin21112004 /Ai-Conversation-question-answertext100K<n<1M0 likes14 downloads9mo agoHugging Face22sachin52 /deepmind_mathCurated Dataset taken from DeepMind synthetically generated math dataset. text100K<n<1M0 likes13 downloads8mo agoHugging Face23sach3v /Gemma_4_E2B_Vision_FOR_Oral_Cancer Oral Gemma Fine-Tuning Dataset This repository contains a portable, instruction-tuning dataset for cropped oral mucosal lesion screening. It is designed for vision-language fine-tuning of Gemma-style models on a binary screening task. Overview Each example pairs: one cropped oral mucosal image one short instruction one JSON answer with a conservative screening recommendation This is a screening support dataset, not a diagnostic dataset. Target labels:… See the full description on the dataset page: https://huggingface.co/datasets/sach3v/Gemma_4_E2B_Vision_FOR_Oral_Cancer.textn<1K0 likes10 downloads3mo agoHugging Face24kucerj56 /czech-sacd-legal-questions 📑 Overview This repository contains 200 question-answer pairs automatically generated with Gemini 2.0 from the decisions of the Czech Sumpreme Administrative Court. The work was performed in spring 2025 as part of my master’s diploma thesis at the Faculty of Information Technology, Czech Technical University in Prague (FIT CTU). 🏛️ Source Official judgments scraped from https://sbirka.nssoud.cz (March 2025 snapshot). textquestion-answeringn<1K0 likes7 downloads1y agoHugging Face25sachin1947 /upsctextquestion-answeringn<1K0 likes7 downloads1y agoHugging Face26Sachin21112004 /ai-conversationtext100K<n<1M0 likes6 downloads9mo agoHugging Face27SachinSharma0325 /erc-efrtext1K<n<10K0 likes5 downloads2y agoHugging Face28sachink365 /filtered_datatext10K<n<100K0 likes5 downloads2y agoHugging Face29vsachi /sachi-chosen-rejected-javsachi/Sachi-Qwen3-4B-GGUF が出力した回答からchosenとrejectedを選んだ結果のデータセットです。 textn<1K0 likes5 downloads1y agoHugging Face30sacreemure /rmj_covid_nertextsummarizationn<1K0 likes5 downloads6mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.