CoolFace
20 results

bert

styletts2-community /multilingual-pl-bertAttribution: Wikipedia.org text100K<n<1M19 likes4.2k downloads3y agoHugging Facenomic-ai /bert-128-grouped Dataset Card for "bert-128-grouped" More Information needed 10M<n<100M0 likes3.7k downloads3y agoHugging FaceAbstractPhil /conceptual-captions-12m-webdataset-bertstext10M<n<100M1 likes3.2k downloads2mo agoHugging FaceBertievidgen /SimpleSafetyTeststexttext-generationn<1K12 likes3.1k downloads2y agoHugging Facebertin-project /mc4-es-sampled50 million documents in Spanish extracted from mC4 applying perplexity sampling via mc4-sampling: "https://huggingface.co/datasets/bertin-project/mc4-sampling". Please, refer to BERTIN Project. The original dataset is the Multlingual Colossal, Cleaned version of Common Crawl's web crawl corpus (mC4), based on the Common Crawl dataset: "https://commoncrawl.org", and processed by AllenAI.texttext-generation1M<n<10M2 likes3k downloads4y agoHugging Faceartefactory /BERTJudge-Dataset BERTJudge-Dataset Dataset Description BERTJudge-Dataset is the training dataset used for developing BERTJudge models, as introduced in the paper BERT-as-a-Judge: A Robust Alternative to Lexical Methods for Efficient Reference-Based LLM Evaluation. It comprises question–candidate–reference triplets generated by 36 recent open-weight, instruction-tuned models across 7 established tasks, and synthetically annotated using nvidia/Llama-3_3-Nemotron-Super-49B-v1_5. The dataset… See the full description on the dataset page: https://huggingface.co/datasets/artefactory/BERTJudge-Dataset.text-classification2 likes2.6k downloads5mo agoHugging Face