datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
English-CEFR-Explorer
🇬🇧 English CEFR Explorer Benchmark
This dataset is an automatically updated benchmark for testing LLM adherence to CEFR (Common European Framework of Reference for Languages) constraints.
Dataset Structure
The dataset is formatted as a JSONL file optimized for Instruction Tuning.
Data Fields
id: Unique identifier for the generation task.
messages: Standard chat format (System, User, Assistant).
System: Defines the persona (ESL Teacher).
User: The… See the full description on the dataset page: https://huggingface.co/datasets/yasincicek/English-CEFR-Explorer.experiment-1-oracles
Experiment 1 Oracles Corpus
Dataset Summary
Dataset containing synthetic ground-truth oracle responses and intermediate reasoning traces for multi-step AI experiment evaluation.
