CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01fffoivos /apertus-8b-greek-cpt-modern-greek-train Exact Modern-Greek training content for Apertus 8B Greek CPT This is the public Modern-Greek, train-only document snapshot selected for the full 8B D0 continued-pretraining run. It preserves the upstream v2 schema and metadata; text is reproduced as its exact training-time Apertus-parity PII-masked value. Selection is reconstructed from immutable post-mask training catalogs and content hashes. It contains no replay payload. Exact selected content HPLT Modern… See the full description on the dataset page: https://huggingface.co/datasets/fffoivos/apertus-8b-greek-cpt-modern-greek-train.tabular10M<n<100M0 likes342 downloads1mo agoHugging Face02swiss-ai /apertus-pretrain-romanshThis dataset consist of three differnt parts. Monolingual Romansh Data, Polylingual data or more precisely translated data from Romansh into either German, French, Italian or English and Sythetic Data. The Polylingual data consists of aligned and non aligned data. The synthetic data was created by interweaving the translational data and prefacing it with the sentence " This is a text translated from SOURCE LANGUAGE to Rumantsch Grischun". The data has a metadata "idiom" if the if specific… See the full description on the dataset page: https://huggingface.co/datasets/swiss-ai/apertus-pretrain-romansh.tabulartranslation100K<n<1M4 likes216 downloads1y agoHugging Face03swiss-ai /Apertus_v1.5_Preference_Data Apertus 1.5 Preference Dataset This is the preference dataset used for the offline DPO stage of Apertus v1.5 alignment training, applied to the 70B model. The prompts come from Ai2's Olmo 3 Dolci-Instruct-DPO dataset. We only reuse the prompts from Dolci-Instruct-DPO; all chosen / rejected responses in this dataset were generated by us. How this dataset was built Prompts. Taken from Dolci-Instruct-DPO (ODC-BY). Response generation and annotation. Every prompt was… See the full description on the dataset page: https://huggingface.co/datasets/swiss-ai/Apertus_v1.5_Preference_Data.tabulartext-generation100K<n<1M4 likes165 downloads1mo agoHugging Face04swiss-ai /apertus-pretrain-poisonandcanariesThis dataset was used as part of Apertus v1 training for poisoning experiments. See our technical report for details, as well as the dedicated study. tabular100K<n<1M3 likes88 downloads11mo agoHugging Face05juiceb0xc0de /apertus-v1.1-0.5b-atlas apertus-v1.1-0.5b-atlas image100K<n<1M0 likes78 downloads25d agoHugging Face06jvamvas /apertus-pretrain-romansh-backtranslatedVersion of https://hf.co/datasets/swiss-ai/apertus-pretrain-romansh (monolingual split only) that includes MT-generated translations into German. The intended purpose of this dataset is to train MT systems or LLMs on the task of idiom-specific German→Romansh translation. Note that the German translations in this dataset might contain errors, since they have been automatically generated by an MT system. Composition of the dataset and Romansh data sources See… See the full description on the dataset page: https://huggingface.co/datasets/jvamvas/apertus-pretrain-romansh-backtranslated.tabular100K<n<1M1 likes29 downloads2mo agoHugging Face07danish-foundation-models /croco-munin-apertus-8b-da-simpo-fulltabular1K<n<10K0 likes19 downloads3mo agoHugging Face08danish-foundation-models /croco-munin-apertus-8b-da-50ktabular10K<n<100K0 likes19 downloads2mo agoHugging Face09danish-foundation-models /croco-munin-apertus-8b-da-simpo-full-50ktabular10K<n<100K0 likes18 downloads2mo agoHugging Face10JingweiNi /apertus_70b_3899_0.2_0.75_correctness_no_refuse_sfttabular10K<n<100K0 likes15 downloads1y agoHugging Face11danish-foundation-models /croco-munin-apertus-8b-da-generatedtabular1K<n<10K0 likes15 downloads3mo agoHugging Face12JingweiNi /apertus_70b_3899_0.2_0.75_majority_correct_sfttabular1K<n<10K0 likes13 downloads1y agoHugging Face13danish-foundation-models /croco-munin-apertus-8b-da-goldtabular1K<n<10K0 likes13 downloads3mo agoHugging Face14JingweiNi /apertus_70b_3899_0.33_0.67_correctness_sfttabular10K<n<100K0 likes11 downloads1y agoHugging Face15JingweiNi /Apertus-8B-aligned_0.2_0.5_0.8_majority_sft_data_correcttabular10K<n<100K0 likes11 downloads1y agoHugging Face16danish-foundation-models /croco-munin-apertus-8b-da-simpo-tunedtabular1K<n<10K0 likes11 downloads3mo agoHugging Face17JingweiNi /apertus_70b_3899_0.33_0.67_correctness_no_refuse_sfttabular10K<n<100K0 likes10 downloads1y agoHugging Face18JingweiNi /Apertus-8B-aligned_0.2_0.75_majority_no_refuse_sft_data_correcttabular1K<n<10K0 likes10 downloads1y agoHugging Face19AgentPublic /evalap-compare-albert-small-with-apertus-small-models-82 Compare albert-small with apertus-small models (ID: 82) Comparing albert-small model with apertus-small model alone, and in a RAG setting with service-public + travail-emploi sheets, on a french administration Q/A datasets Overview This dataset contains 8 experiments from the EvalAP evaluation platform. Datasets: Assistant IA - QA, MFS_questions_v01 Models evaluated: meta-llama/Llama-3.1-8B-Instruct, swiss-ai/Apertus-8B-Instruct-2509 Metrics: generation_time… See the full description on the dataset page: https://huggingface.co/datasets/AgentPublic/evalap-compare-albert-small-with-apertus-small-models-82.tabularn<1K1 likes10 downloads8mo agoHugging Face20JingweiNi /apertus_70b_3899_0.2_0.5_0.8_correctness_sfttabular10K<n<100K0 likes9 downloads1y agoHugging Face21JingweiNi /Apertus-70B-aligned_0.33_0.67_majority_no_refuse_sft_data_correcttabular1K<n<10K0 likes9 downloads1y agoHugging Face22JingweiNi /Apertus-8B-aligned_0.1_0.7_correctness_no_refuse_sft_datatabular10K<n<100K0 likes8 downloads1y agoHugging Face23JingweiNi /Apertus-70B-aligned_0.2_0.5_0.8_majority_sft_data_correcttabular1K<n<10K0 likes7 downloads1y agoHugging Face24JingweiNi /Apertus-70B-aligned_0.33_0.67_majority_no_refuse_rand_expressiontabular1K<n<10K0 likes7 downloads1y agoHugging Face25danish-foundation-models /croco-munin-apertus-8b-databular1K<n<10K0 likes7 downloads3mo agoHugging Face26danish-foundation-models /croco-munin-apertus-8b-da-simpotabular1K<n<10K0 likes7 downloads3mo agoHugging Face27JingweiNi /Apertus-8B-aligned_0.33_0.67_majority_no_refuse_sft_data_correcttabular1K<n<10K0 likes6 downloads1y agoHugging Face28JingweiNi /Apertus-70B-aligned_0.2_0.75_majority_no_refuse_sft_data_correcttabular1K<n<10K0 likes6 downloads1y agoHugging Face29JingweiNi /Apertus-70B-aligned_0.2_0.5_0.8_correctness_sft_datatabular10K<n<100K0 likes6 downloads1y agoHugging Face30JingweiNi /Apertus-70B-aligned_0.2_0.75_correctness_no_refuse_sft_datatabular10K<n<100K0 likes5 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.