datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Fable-5-Cursor-TracesFable 5 Cursor Traces
244 Fable 5 Cursor agent sessions for training & research.
This dataset has 244 Cursor sessions with Fable 5 at High/xHigh/Max effort levels for distillation.
[!IMPORTANT]
This dataset is compatible with Teich! Use it directly in your Teich training pipeline.
[!WARNING]
The longest rows exceed one million characters of content. Apply prepare_data() with your intended tokenizer and an explicit context/oversize policy before training.… See the full description on the dataset page: https://huggingface.co/datasets/TeichAI/Fable-5-Cursor-Traces.cursor-traces-exampleThis dataset was generated using teich by TeichAI
My Agent Traces
This directory contains raw agent trace files generated by teich.
JSONL files: 9
Training-ready tools
Generated agent traces carry configured or recovered tool schemas so tools remain available for training even when a session did not call them.
Native Claude Code imports recover schemas for Claude Code and Claude Desktop built-ins, plus conservative name-derived MCP schemas, when the raw… See the full description on the dataset page: https://huggingface.co/datasets/armand0e/cursor-traces-example.Fable-5-Cursor-TracesFable 5 Cursor Traces
244 Fable 5 Cursor agent sessions for training & research.
This dataset has 244 Cursor sessions with Fable 5 at High/xHigh/Max effort levels for distillation.
[!IMPORTANT]
This dataset is compatible with Teich! Use it directly in your Teich training pipeline.
[!WARNING]
The longest rows exceed one million characters of content. Apply prepare_data() with your intended tokenizer and an explicit context/oversize policy before training.… See the full description on the dataset page: https://huggingface.co/datasets/11-47/Fable-5-Cursor-Traces.marcuscedricridia__cursorr-o1.2-7b-details
Dataset Card for Evaluation run of marcuscedricridia/cursorr-o1.2-7b
Dataset automatically created during the evaluation run of model marcuscedricridia/cursorr-o1.2-7b
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/marcuscedricridia__cursorr-o1.2-7b-details.project
