datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
SciCode-Runnable-Benchmark-Reviewedhallmark-mlx-reviewed-policy-traces
hallmark-mlx-reviewed-policy-traces
Reviewed citation-verification training traces for hallmark-mlx.
Contents
train.jsonl: 75 supervised examples
valid.jsonl: 6 supervised examples
Source reviewed traces: reviewed_seed_traces_combined.jsonl with 45 full traces.
Format
Each row is a prepared supervised training example for MLX LoRA fine-tuning.
The format is the exact snapshot used by the kept Qwen 1.5B run.
Upload Note
Review the… See the full description on the dataset page: https://huggingface.co/datasets/sebastianboehler/hallmark-mlx-reviewed-policy-traces.lemonseed-qwen38-cogen-reviewed
lemonseed-qwen38-cogen-reviewed
LemonSeed — Qwen3.8-Max-teacher co-generated data, reviewed passes (v1).
Contents
intelligent_qwen38_cogen_1h_20260824_r1.reviewed_passes_v1.jsonl (71 rows)
Format
JSON Lines (.jsonl), one example per line.
Provenance
LLM-teacher co-generated instruction/chat data for LemonSeed fine-tuning.
