datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
DAG-MATH-Formatted-CoT
Benchmark Overview
This dataset card contains 2,894 gold-standard DAG-MATH formatted CoT from problems from Omni-MATH.
Top‑Level Schema
Each JSON file is a list with a single object describing the problem:
problem_id: integer identifier of the problem.
domain: list of strings describing the topic taxonomy.
difficulty: numeric difficulty indicator from 1 (easiest) to 6 (hardest).
problem_text: problem statement.
sample_id: sample identifier for the solution trace.… See the full description on the dataset page: https://huggingface.co/datasets/yuanhezhang/DAG-MATH-Formatted-CoT.My-Reasoning-Datasetdag_remediation_traces
DAG Remediation Traces
Author: Venkata Krishna Azith Teja Ganti
Part of the ExposureGuard PHI Re-identification Risk Ecosystem
Input/output execution traces for budget-constrained PHI remediation planning over multimodal clinical records. Each record pairs a patient risk profile with a complete DAG planning trace: which actions were selected, which dependency injections fired, the topological execution order, and the final residual risk and cost.
Use this to train or benchmark… See the full description on the dataset page: https://huggingface.co/datasets/vkatg/dag_remediation_traces.stacx-swe-online-dagger-dataDAG_sftdag-planner-sft-datagpt-oss-finetuning-dataset-gemini-2.5-pro-2xacbankgpt-oss-finetuning-dataset-gemini-2.5-pro
