datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
lemma-long-horizon-rewrite-benchmark
LEMMA Long-Horizon Symbolic Rewriting Benchmark
Rewrite one expression into another, one verified step at a time — for up to 128 steps.
Every problem gives you a start expression, an exact target, and a reference derivation in which
every single step was accepted by a symbolic verifier. The hard part is not any individual
rewrite. It is picking the right rewrite at each of up to 128 consecutive states, where a
plausible-looking legal move can quietly take you away from the… See the full description on the dataset page: https://huggingface.co/datasets/BlackdromeAILabs/lemma-long-horizon-rewrite-benchmark.kavram-yanilgisi-envanteri
Kavram Yanılgısı Envanteri — Türkiye ortaokul matematiği (6-8. sınıf)
505 kayıt · 31 kaynak · 24 sütun · Türkçe
Bu veri seti, Türkiye'de tam metnine erişilen 31 akademik çalışmanın (11 yüksek lisans
tezi + 20 dergi makalesi) okunarak kodlanmış kavram yanılgısı kayıtlarından oluşur.
Her satır bir yanılgıyı tanımlar; kaynağın künyesini ve sayfa numarasını taşır, bu
sayede her kayıt tek tek doğrulanabilir (505 kaydın 505'inde sayfa referansı vardır).
Kapsam
6, 7 ve… See the full description on the dataset page: https://huggingface.co/datasets/lemmaakademi/kavram-yanilgisi-envanteri.top_engagement-4_groups-lemma-unbalanced-1y-20k_per_groupEUS_lemma_length
