CoolFace
8 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01au123 /snli-hardtext1K<n<10K0 likes129 downloads2y agoHugging Face02Younes2E /atomic-snli atomic-snli Atomic propositions for the premise and hypothesis of each NLI pair, derived from stanfordnlp/snli. Each sentence was decomposed into standalone atomic propositions; these propositions are joined back to the original NLI pairs. Columns column type description premise string original premise sentence hypothesis string original hypothesis sentence label int 0 = entailment, 1 = neutral, 2 = contradiction premise_propositions list[string]… See the full description on the dataset page: https://huggingface.co/datasets/Younes2E/atomic-snli.texttext-classification100K<n<1M0 likes40 downloads2mo agoHugging Face03CZLC /cs_snli Dataset Card for Czech SNLI Czech translation of the Stanford Natural Language Interface (SNLI) dataset with manual annotation of a SNLI subset. In addition to the entailment/contradiction/neutral inference, a "bad translation" class was added. The annotation was done by students of NLP or computational linguistics. 1499 same pairs were annotated by two students to check IAA. Dataset Details The annotation for Czech premise-hypothesis pairs is done on 165390 pairs… See the full description on the dataset page: https://huggingface.co/datasets/CZLC/cs_snli.tabulartext-classification10K<n<100K0 likes30 downloads2y agoHugging Face04closji /snli_augtextn<1K0 likes20 downloads4y agoHugging Face05MatterhornMatters /snli_de_by_ayatext100K<n<1M0 likes15 downloads11mo agoHugging Face06fran-gen /snli-smoke-test SNLI Smoke Test Dataset Summary This dataset is a small smoke-test subset derived from the Stanford Natural Language Inference (SNLI) training split. It is intended for fast end-to-end checks of prompt formatting, model adapters, output parsing, and metric pipelines in entailment-lab. The dataset contains 100 sentence pairs: 34 entailment 33 contradiction 33 neutral Most of the dataset is organized as complete captionID triplets, where the same premise group… See the full description on the dataset page: https://huggingface.co/datasets/fran-gen/snli-smoke-test.texttext-classificationn<1K0 likes15 downloads2mo agoHugging Face07ozgurkrkrt /snli_tr_en_datasettext100K<n<1M0 likes6 downloads2y agoHugging Face08Samhita-kolluri /snli-contrastive-json-datasettext100K<n<1M0 likes1 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.