Arsh9210/Nemotron-RL-Instruction-Following-Structured-Outputs-v2
Dataset Description: Split 1: Direct Generation tests the model’s ability to perform freeform text structured outputs on JSON, YAML, and XML data, varying the complexity and presentation of the schema. Split 2: Diversified Tasks adds 2 additional output formats: TOML and CSV, while increasing problem types to Direct Extraction from document, Translation between formats, Multistep Translation from known data, Multistep Extraction from unrelated context, Schema-Only Generation for… See the full description on the dataset page: https://huggingface.co/datasets/Arsh9210/Nemotron-RL-Instruction-Following-Structured-Outputs-v2.
Added tool_calling_extraction/train-00000-of-00001.parquet
Added diversified_tasks/train-00000-of-00001.parquet
Added direct_generation/train-00000-of-00001.parquet
Added README.md
initial commit
