datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
structeval-t-sft-v2-toml
StructEval-T SFT v2 - Full TOML
This dataset is the full, refined, format-specific subset for StructEval-T, focusing exclusively on strictly validated TOML transformations.
Key Features
Total Samples: 3,635
Verified Quality: 100% strictly validated using AST/parsers (e.g. json.loads, xml.etree.ElementTree, yaml.safe_load). Only samples that successfully parse as valid TOML without errors are included.
Source: This is a split from the unified structeval-t-sft-v2 dataset.… See the full description on the dataset page: https://huggingface.co/datasets/daichira/structeval-t-sft-v2-toml.structeval-t-sft-hq-toml
StructEval-T SFT - High Quality TOML
This dataset is a highly refined, format-specific subset for StructEval-T, focusing exclusively on strictly validated TOML transformations.
Key Features
Total Samples: 2,000
Verified Quality: 100% strictly validated using AST/parsers (e.g. json.loads, xml.etree.ElementTree, yaml.safe_load). Only samples that successfully parse as valid TOML without errors are included.
Goal: To maximize single-format fine-tuning performance or to be… See the full description on the dataset page: https://huggingface.co/datasets/daichira/structeval-t-sft-hq-toml.toml_dpoplan_a_v4_toml_explicittoml_repair_sfttoml_constraints_minu-10bei_x_to_toml
