datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
apertus-8b-greek-cpt-modern-greek-train
Exact Modern-Greek training content for Apertus 8B Greek CPT
This is the public Modern-Greek, train-only document snapshot selected for the full 8B D0 continued-pretraining run. It preserves the upstream v2 schema and metadata; text is reproduced as its exact training-time Apertus-parity PII-masked value. Selection is reconstructed from immutable post-mask training catalogs and content hashes. It contains no replay payload.
Exact selected content
HPLT Modern… See the full description on the dataset page: https://huggingface.co/datasets/fffoivos/apertus-8b-greek-cpt-modern-greek-train.apertus-pretrain-romanshThis dataset consist of three differnt parts. Monolingual Romansh Data, Polylingual data or more precisely translated data from Romansh into either German, French, Italian or English and Sythetic Data.
The Polylingual data consists of aligned and non aligned data. The synthetic data was created by interweaving the translational data and prefacing it with the sentence " This is a text translated from SOURCE LANGUAGE to Rumantsch Grischun".
The data has a metadata "idiom" if the if specific… See the full description on the dataset page: https://huggingface.co/datasets/swiss-ai/apertus-pretrain-romansh.Apertus_v1.5_Preference_Data
Apertus 1.5 Preference Dataset
This is the preference dataset used for the offline DPO stage of Apertus v1.5 alignment training, applied to the 70B model.
The prompts come from Ai2's Olmo 3 Dolci-Instruct-DPO dataset. We only reuse the prompts from Dolci-Instruct-DPO; all chosen / rejected responses in this dataset were generated by us.
How this dataset was built
Prompts. Taken from Dolci-Instruct-DPO (ODC-BY).
Response generation and annotation. Every prompt was… See the full description on the dataset page: https://huggingface.co/datasets/swiss-ai/Apertus_v1.5_Preference_Data.apertus-pretrain-poisonandcanariesThis dataset was used as part of Apertus v1 training for poisoning experiments. See our technical report for details, as well as the dedicated study.
apertus-v1.1-0.5b-atlas
apertus-v1.1-0.5b-atlas
apertus-pretrain-romansh-backtranslatedVersion of https://hf.co/datasets/swiss-ai/apertus-pretrain-romansh (monolingual split only) that includes MT-generated translations into German.
The intended purpose of this dataset is to train MT systems or LLMs on the task of idiom-specific German→Romansh translation. Note that the German translations in this dataset might contain errors, since they have been automatically generated by an MT system.
Composition of the dataset and Romansh data sources
See… See the full description on the dataset page: https://huggingface.co/datasets/jvamvas/apertus-pretrain-romansh-backtranslated.croco-munin-apertus-8b-da-simpo-fullcroco-munin-apertus-8b-da-50kcroco-munin-apertus-8b-da-simpo-full-50kapertus_70b_3899_0.2_0.75_correctness_no_refuse_sftcroco-munin-apertus-8b-da-generatedapertus_70b_3899_0.2_0.75_majority_correct_sftcroco-munin-apertus-8b-da-goldapertus_70b_3899_0.33_0.67_correctness_sftApertus-8B-aligned_0.2_0.5_0.8_majority_sft_data_correctcroco-munin-apertus-8b-da-simpo-tunedapertus_70b_3899_0.33_0.67_correctness_no_refuse_sftApertus-8B-aligned_0.2_0.75_majority_no_refuse_sft_data_correctevalap-compare-albert-small-with-apertus-small-models-82
Compare albert-small with apertus-small models (ID: 82)
Comparing albert-small model with apertus-small model alone, and in a RAG setting with service-public + travail-emploi sheets, on a french administration Q/A datasets
Overview
This dataset contains 8 experiments
from the EvalAP evaluation platform.
Datasets: Assistant IA - QA, MFS_questions_v01
Models evaluated: meta-llama/Llama-3.1-8B-Instruct, swiss-ai/Apertus-8B-Instruct-2509
Metrics: generation_time… See the full description on the dataset page: https://huggingface.co/datasets/AgentPublic/evalap-compare-albert-small-with-apertus-small-models-82.apertus_70b_3899_0.2_0.5_0.8_correctness_sftApertus-70B-aligned_0.33_0.67_majority_no_refuse_sft_data_correctApertus-8B-aligned_0.1_0.7_correctness_no_refuse_sft_dataApertus-70B-aligned_0.2_0.5_0.8_majority_sft_data_correctApertus-70B-aligned_0.33_0.67_majority_no_refuse_rand_expressioncroco-munin-apertus-8b-dacroco-munin-apertus-8b-da-simpoApertus-8B-aligned_0.33_0.67_majority_no_refuse_sft_data_correctApertus-70B-aligned_0.2_0.75_majority_no_refuse_sft_data_correctApertus-70B-aligned_0.2_0.5_0.8_correctness_sft_dataApertus-70B-aligned_0.2_0.75_correctness_no_refuse_sft_data
