CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01daxum34 /sst_migaretThis dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "robot_type": "bi_so101_follower", "total_episodes": 61, "total_frames": 53650, "total_tasks": 1, "chunks_size": 1000, "data_files_size_in_mb": 100, "video_files_size_in_mb": 200, "fps": 30, "splits": { "train": "0:61" }, "data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/daxum34/sst_migaret.tabularrobotics10K<n<100K0 likes150 downloads26d agoHugging Face0213515257315Tzr /SSTQAtabularn<1K0 likes87 downloads1y agoHugging Face03gokuls /glue_augmented_sst2 Dataset Card for glue_augmented_sst2 Dataset Description Augmented SST-2 dataset Reference: https://huggingface.co/datasets/glue tabular1M<n<10M1 likes52 downloads4y agoHugging Face04metinovadilet /kyrgyz-sst2 Kyrgyz SST-2 Task Sentiment Classification (Binary Sentence Classification) Description Stanford Sentiment Treebank binary classification task translated from English to Kyrgyz. Each entry contains an English sentence, its Kyrgyz translation, and a binary sentiment label. Labels: negative, positive Format: JSONL with fields such as sentence, sentence_ky, and label Dataset Size Split Entries Train 6,920 Validation 872 Test 1,821… See the full description on the dataset page: https://huggingface.co/datasets/metinovadilet/kyrgyz-sst2.tabulartext-classification10K<n<100K0 likes52 downloads5mo agoHugging Face05KaiLv /UDR_SST-5 Dataset Card for "UDR_SST-5" More Information needed tabular10K<n<100K0 likes45 downloads3y agoHugging Face06wrynx /probe-robustness-sst2 sst2 Dataset repo: wrynx/probe-robustness-sst2 Auto-generated by prepare_datasets.py. Do not hand-edit -- regenerate by re-running the script (with --force) instead. Stats Total records: 68221 Records per split: test: 10910 train: 46918 valid: 10393 Number of classes: 2 Records per class: 0: 30208 1: 38013 Records per class per split: test: 0: 4900 1: 6010 train: 0: 20712 1: 26206 valid: 0: 4596 1: 5797 Original dataset README (from… See the full description on the dataset page: https://huggingface.co/datasets/wrynx/probe-robustness-sst2.tabular10K<n<100K0 likes45 downloads24d agoHugging Face07mrm8488 /sst2-es-mt STT-2 Spanish A Spanish translation (using EasyNMT) of the SST-2 Dataset For more information check the official Model Card tabulartext-classification10K<n<100K2 likes34 downloads4y agoHugging Face08rungalileo /sst2tabular10K<n<100K0 likes32 downloads4y agoHugging Face09generalization /sst2_Full-p_05tabular10K<n<100K0 likes29 downloads4y agoHugging Face10generalization /sst2_Sampled-p_1tabular10K<n<100K0 likes27 downloads4y agoHugging Face11fatmaElsafoury2022 /SST_sentiment_fairness_data Sentiment fairness dataset ================================ This dataset is to measure gender fairness in the downstream task of sentiment analysis. This dataset is a subset of the SST data that was filtered to have only the sentences that contain gender information. The python code used to create this dataset can be found in the prepare_sst.ipyth file. Then the filtered datset was labeled by 4 human annotators who are the authors of this dataset. The annotations… See the full description on the dataset page: https://huggingface.co/datasets/fatmaElsafoury2022/SST_sentiment_fairness_data.tabulartext-classificationn<1K2 likes27 downloads3y agoHugging Face12generalization /sst2_Full-p_1tabular10K<n<100K0 likes24 downloads4y agoHugging Face13KaiLv /UDR_SST-2 Dataset Card for "UDR_SST-2" More Information needed tabular10K<n<100K0 likes24 downloads3y agoHugging Face14cbrian /dataset_env_SST_SP6_WC1_TC1_task_pickplaceblueblock_numepi_10_ctrl_cartesianThis dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "robot_type": "panda", "total_episodes": 10, "total_frames": 1698, "total_tasks": 1, "chunks_size": 1000, "data_files_size_in_mb": 100,"video_files_size_in_mb": 200, "fps": 15, "splits": { "train": "0:10" }, "data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/cbrian/dataset_env_SST_SP6_WC1_TC1_task_pickplaceblueblock_numepi_10_ctrl_cartesian.tabularrobotics1K<n<10K0 likes24 downloads5mo agoHugging Face15contemmcm /ssttabulartext-classification100K<n<1M0 likes23 downloads2y agoHugging Face16rubricreward /llm-metric-glue-sst2tabular10K<n<100K0 likes23 downloads1y agoHugging Face17cbrian /dataset_env_SST_SP8_WC1_TC1_task_pickplaceblueblock_numepi_10_ctrl_cartesianThis dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "robot_type": "panda", "total_episodes": 10, "total_frames": 1825, "total_tasks": 1, "chunks_size": 1000, "data_files_size_in_mb": 100,"video_files_size_in_mb": 200, "fps": 15, "splits": { "train": "0:10" }, "data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/cbrian/dataset_env_SST_SP8_WC1_TC1_task_pickplaceblueblock_numepi_10_ctrl_cartesian.tabularrobotics1K<n<10K0 likes23 downloads5mo agoHugging Face18sstoeckl /populism-llm Populism-LLM A cross-validated LLM annotation of European party manifestos on populism and liberalism dimensions. Three independent large language models — Claude Sonnet 4.6 (Anthropic), GPT-4.1-mini (OpenAI), and Gemini Flash (Google) — score every CMP manifesto using an identical strict JSON schema. Each score is backed by a verbatim quote and contextual snippet. 🚧 Preview / draft release. A peer-reviewed paper-companion v1.0 release with DOI is forthcoming. Until then… See the full description on the dataset page: https://huggingface.co/datasets/sstoeckl/populism-llm.tabulartext-classification10K<n<100K0 likes23 downloads3mo agoHugging Face19BSC-LT /cobie_sst2 Dataset Card for cobie_sst2 This dataset is a modification of the original SST-2 dataset for LLM cognitive bias evaluation. Language(s) English (en) Dataset Summary The Stanford Sentiment Treebank is a corpus with fully labeled parse trees that allows for a complete analysis of the compositional effects of sentiment in language. The corpus is based on the dataset introduced by Pang and Lee (2005) and consists of 11,855 single sentences extracted from movie… See the full description on the dataset page: https://huggingface.co/datasets/BSC-LT/cobie_sst2.tabulartext-classification100K<n<1M0 likes22 downloads2y agoHugging Face20abhaygupta1266 /sst-2tabular1K<n<10K0 likes20 downloads1y agoHugging Face21dojo /sst2_balancedtabular10K<n<100K0 likes19 downloads4y agoHugging Face22liuyanchen1015 /MULTI_VALUE_sst2_drop_inf_to Dataset Card for "MULTI_VALUE_sst2_drop_inf_to" More Information needed tabular1K<n<10K0 likes19 downloads3y agoHugging Face23generalization /sst2_Sampled-p_05tabular10K<n<100K0 likes18 downloads4y agoHugging Face24liuyanchen1015 /MULTI_VALUE_sst2_double_modals Dataset Card for "MULTI_VALUE_sst2_double_modals" More Information needed tabular1K<n<10K0 likes18 downloads3y agoHugging Face25emirhanboge /sst2_mnli_qqp_llama1b_modified Multi-Task Dataset: SST-2 + MNLI + QQP (Modified for LLaMA 1B) This dataset is a combination of SST-2, MNLI, and QQP for multi-task learning. It is preprocessed and tokenized specifically for training with the LLaMA-1B model. Modifications: Each example includes a task prefix: SST-2: "Task: SST2 | Sentence: ..." MNLI: "Task: MNLI | Premise: ... Hypothesis: ..." QQP: "Task: QQP | Q1: ... Q2: ..." Labels are standardized to integer format. Tokenized using the LLaMA-1B… See the full description on the dataset page: https://huggingface.co/datasets/emirhanboge/sst2_mnli_qqp_llama1b_modified.tabular100K<n<1M0 likes18 downloads2y agoHugging Face26cbrian /dataset_env_SST_SP5_WC1_TC1_task_pickplaceblueblock_numepi_10_ctrl_cartesianThis dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "robot_type": "panda", "total_episodes": 10, "total_frames": 1591, "total_tasks": 1, "chunks_size": 1000, "data_files_size_in_mb": 100,"video_files_size_in_mb": 200, "fps": 15, "splits": { "train": "0:10" }, "data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/cbrian/dataset_env_SST_SP5_WC1_TC1_task_pickplaceblueblock_numepi_10_ctrl_cartesian.tabularrobotics1K<n<10K0 likes18 downloads5mo agoHugging Face27christophsonntag /sst2-poisoned-target-1-testsettabularn<1K0 likes17 downloads2y agoHugging Face28AudreyTrungNguyen /sst2-data_augmentationtabular100K<n<1M0 likes17 downloads2y agoHugging Face29tbilisi-ai-lab /sst2-ka sst2-ka Georgian translation of the SST-2 (Stanford Sentiment Treebank) benchmark. Dataset Summary Property Value Examples 481 Splits validation Languages Georgian, English Task Sentiment Classification Data Fields sentence: Sentence (English) label: Sentiment label (0=negative, 1=positive) idx: Example index sentence_ka: Sentence (Georgian) Translation Methodology Translation model generates initial Georgian translation… See the full description on the dataset page: https://huggingface.co/datasets/tbilisi-ai-lab/sst2-ka.tabulartext-classificationn<1K0 likes17 downloads4mo agoHugging Face30cbrian /dataset_env_SST_SP2_WC1_TC1_task_pickplaceblueblock_numepi_10_ctrl_cartesianThis dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "robot_type": "panda", "total_episodes": 10, "total_frames": 1825, "total_tasks": 1, "chunks_size": 1000, "data_files_size_in_mb": 100,"video_files_size_in_mb": 200, "fps": 15, "splits": { "train": "0:10" }, "data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/cbrian/dataset_env_SST_SP2_WC1_TC1_task_pickplaceblueblock_numepi_10_ctrl_cartesian.tabularrobotics1K<n<10K0 likes17 downloads5mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.