datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
awesome-loop-engineering
Awesome Loop Engineering Dataset
A structured dataset of 1022 papers, official docs, tools, benchmarks, patterns, critiques, and implementation guides for recurring AI-agent systems.
Resource Atlas ·
GitHub field guide ·
Resource selection ·
Report a correction
Dataset Summary
Each row connects an original source to its contribution, novelty, impact, publication details, lifecycle stages, audience, evidence type, link status, and… See the full description on the dataset page: https://huggingface.co/datasets/cy0307/awesome-loop-engineering.loop7kaimenoakuyakureijouwamototekikokudejiyuukimamanahanayomeseikatsuwomankitsusuru
Bangumi Image Base of Loop 7-kaime No Akuyaku Reijou Wa, Moto Tekikoku De Jiyuu Kimama Na Hanayome Seikatsu Wo Mankitsu Suru
This is the image base of bangumi Loop 7-kaime no Akuyaku Reijou wa, Moto Tekikoku de Jiyuu Kimama na Hanayome Seikatsu wo Mankitsu suru, we detected 62 characters, 3609 images in total. The full dataset is here.
Please note that these image bases are not guaranteed to be 100% cleaned, they may be noisy actual. If you intend to manually train models using… See the full description on the dataset page: https://huggingface.co/datasets/BangumiBase/loop7kaimenoakuyakureijouwamototekikokudejiyuukimamanahanayomeseikatsuwomankitsusuru.LOOPerSet
LOOPerSet: A Large-Scale Dataset for Data-Driven Polyhedral Optimization
Dataset at a Glance
LOOPerSet is a corpus of 28 million labeled compilation traces designed for machine learning research in compilers and systems. It maps synthetically generated loop nests and complex optimization sequences to ground-truth execution times measured on physical hardware. Transformation sequences were generated using a polyhedral compilation framework to ensure they… See the full description on the dataset page: https://huggingface.co/datasets/Mascinissa/LOOPerSet.loopmoe-fineweb-edu-100bt-mcore
LoopMoE FineWeb-Edu 100BT (Megatron indexed)
Release status: complete
This is the exact pretokenized FineWeb-Edu 100BT corpus prepared for the
reviewed LoopMoE M1 Dense/Loop/Dual and Fast-Slow experiment contracts.
This release does not define
the M2--M5 train schedules. It contains 64 training indexed-dataset shards and
one fixed validation shard. Each prefix has a Megatron .bin/.idx pair and
a sanitized .stats.json record. There is no separate test split.
Split
Documents… See the full description on the dataset page: https://huggingface.co/datasets/gaotang/loopmoe-fineweb-edu-100bt-mcore.LoopNavLoopsBench
LoopsBench
This dataset repository hosts published LoopsBench task bundles. A LoopsBench task is a self-contained evaluation package for long-horizon terminal coding: it includes an agent-visible workspace snapshot, unit-level requirements, dependency graphs, Docker execution metadata, public verifier files, and reference gold patches used by maintainers and Oracle-style validation.
The files in this dataset are release artifacts mirrored from the latest LoopsBench GitHub… See the full description on the dataset page: https://huggingface.co/datasets/LoopsBench/LoopsBench.loopshapeW2H-Basic-Agent-Loop-w-Sandbox
W2H Basic Agent Loop with built in Linux Sandbox
A lightweight home agent that talks, runs code and takes actions in the real world. Access it from anywhere.
This is a vanilla Python agent loop that supports tools, skills, a microVM sandbox, encrypted data-in-transit and the Arduino microcontroller. The web UI includes voice, file uploads and slash commands. Designed for learning and experimentation. Use vibe coding to adapt it for different tasks.
Talk to the agent from… See the full description on the dataset page: https://huggingface.co/datasets/vbookshelf/W2H-Basic-Agent-Loop-w-Sandbox.relay-ko-s2-scratch-g4-150h-R0-arelay-ko-s2-scratch-g4-53h-R0-arelay-ko-s2-scratch-g4-100h-R0-brelay-ko-s2-scratch-g4-53h-R0-brelay-ko-s2-scratch-g4-100h-R0-arelay-ko-s2-scratch-g4-150h-R0-bloopwan-opensora-pilot-v1
LoopWan Open-Sora-Plan pilot
Status: completed bounded curation. Counts: {"long_audit": 22, "train": 2000, "val": 128}.
Fixed 320x480, timestamp sampling at 16 FPS; train/validation crops are real
contiguous 10-second shots, audit crops 20 seconds. Sources are disjoint and
captions are matched to pinned official annotations. See DATASET_REPORT.md for
filter thresholds, caption limitations and full provenance.
Official dataset revision: ab77293def393e6938f11a7bfd12163decfb9620.… See the full description on the dataset page: https://huggingface.co/datasets/Nicholas0228/loopwan-opensora-pilot-v1.LoopTF-Sudokucircle-packing-insight-loop
Circle-Packing Insight-Exploration Loop
Artifacts from an iterative GPT solver <-> proposer insight-exploration loop on the
21-circles-in-a-perimeter-4-rectangle packing problem (AlphaEvolve SOTA sum-of-radii
= 2.3658321334167627). Each round, 16 solvers propose a program + written explanation;
every program is scored; a proposer then mines all 16 attempts into an evolving insight
document that conditions the next round. Run: 16 solvers x 8 rounds.
Subsets… See the full description on the dataset page: https://huggingface.co/datasets/ars22/circle-packing-insight-loop.LoopTool-23k
Overview
LoopTool is a fully automatic, model-aware iterative framework that tightly couples data generation and model training for tool-augmented LLM learning
The LoopTool-2w is released as part of Closing the Data–Training Loop for Robust LLM Tool Calls
The dataset comprises 23,040 tool-call samples, involving 20,813 APIs. In each sample, the instruction contains the corresponding set of available tools for that sample; the input corresponds to the dialogue history of the… See the full description on the dataset page: https://huggingface.co/datasets/zhangkangning/LoopTool-23k.mixed_Loop_3mixed_Loop_0swe_gym_491i_max_loop_size_2lipika-eval
Lipika eval — Indic font recognition benchmark
The frozen validation set behind loopdesk-ai/lipika
(Indic font recognizer): 6,876 synthetic text crops covering 553 freely-licensed font
families across 13 scripts (Devanagari, Bengali, Gujarati, Gurmukhi, Kannada, Malayalam,
Meetei Mayek, Odia, Ol Chiki, Perso-Arabic, Tamil, Telugu, Latin).
This is the set reported as "synthetic val" in the model card (Lipika v2.4 scores 0.849
family top-1 / 0.977 top-5 / 0.991 script). Use it to… See the full description on the dataset page: https://huggingface.co/datasets/loopdesk-ai/lipika-eval.behavioral-loops
Behavioral Loops
1,140 behavioral patterns across 279 categories, each structured as given/when/then/result logic with taxonomy classification, veracity scores, and intervention strategies.
Quick Start
from datasets import load_dataset
ds = load_dataset("buley/behavioral-loops")
print(ds["train"][0])
Structure
Field
Description
given
Initial condition or context
when
Trigger event
then
Resulting behavior
result
Long-term outcome
origin… See the full description on the dataset page: https://huggingface.co/datasets/buley/behavioral-loops.mixed_Loop_5RECURSIVE_LOOP_1relay-ko-s2-scratch-53h-R1-breform-dafny-loop-inv-gen
ReForm Dafny Loop-Invariant Infill
A derived, narrower-task version of Veri-Code/ReForm-Python2Dafny-Dataset and Veri-Code/ReForm-DafnyComp-Benchmark, targeting loop-invariant synthesis specifically, rather than "fill in all missing annotations."
What it is
Each row is a (modified_input, output) pair where:
output is a complete, Dafny-verified program (confirmed via a real dafny verify pass, not just parsing — see below).
modified_input is the same program with… See the full description on the dataset page: https://huggingface.co/datasets/ThuraAung1601/reform-dafny-loop-inv-gen.relay-ko-s2-scratch-53h-R1-ahedgehog-loop-control-r4
hedgehog-loop-control-r4
Hedgehog — loop-control round 4 (termination/repetition fixes).
Contents
train.jsonl (2944 rows)
validation.jsonl (438 rows)
Format
JSON Lines (.jsonl), one example per line.
Provenance
Original content for the Hedgehog extraction model (Michael Anthony Falabella).
mixed_Loop_4
