datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
wsinfer-model-zoo-jsonThis is the registry of models in the WSInfer Model Zoo.
See https://wsinfer.readthedocs.io/en/latest/ and https://github.com/SBU-BMI/wsinfer-zoo for more information.
json-mode-evaljson-mode-eval-extended
JSON-Mode-eval extended
This is a dataset that measures LLM capabilities at extracting data from natural language following a JSON Schema.
It was generated by manually cleaning and normalizing json-mode-eval by Nous-Research, which resulted in json-mode-eval-cleaned, ensuring that every schema enforces non-empty constraints and allow no additional keys on the top level.
We then prompt Gemini 2.5 Pro for additional 10 samples per schema, filtering for outputs that are valid according… See the full description on the dataset page: https://huggingface.co/datasets/eth-sri/json-mode-eval-extended.json-mode-reasoningjson-mode-agentic-reasoningdetails_Nhoodie__Meta-Llama-3-8B-Uninstruct-function-calling-json-mode-model_stock-v0.1
Dataset Card for Evaluation run of Nhoodie/Meta-Llama-3-8B-Uninstruct-function-calling-json-mode-model_stock-v0.1
Dataset automatically created during the evaluation run of model Nhoodie/Meta-Llama-3-8B-Uninstruct-function-calling-json-mode-model_stock-v0.1 on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_Nhoodie__Meta-Llama-3-8B-Uninstruct-function-calling-json-mode-model_stock-v0.1.json-mode-singleturnjson-mode-evalThis is an extended version of https://huggingface.co/datasets/NousResearch/json-mode-eval .
Warning: the output currently is placeholder! This dataset should only be used for testing efficiency!
json-mode-verifiablejson-mode-eval-cleaned
JSON-Mode-eval extended
This is a dataset that measures LLM capabilities at extract data from natural language following a JSON Schema.
It was generated by manually cleaning and normalizing json-mode-eval by Nous-Research.
This dataset was used for evaluation in the paper Constrained Decoding of Diffusion LLMs with Context-Free Grammars. You can find the corresponding evaluation code on the project GitHub Repository.
Example Usage
from datasets import load_dataset… See the full description on the dataset page: https://huggingface.co/datasets/eth-sri/json-mode-eval-cleaned.json-mode-eval-rgxjson-mode-agenticreasoning-sft-interstellarninja-json-mode-reasoning-160K
json-mode-reasoning (converted)
Converted version of interstellarninja/json-mode-reasoning, filtered to 20,474 rows with valid <think> reasoning traces.
Format
Each row has three columns:
input — list of dicts [{"role": "system/user", "content": "..."}, ...] (conversation turns ending on the last user turn, includes system prompt with JSON schema)
response — assistant response string with <think> reasoning block followed by JSON output
source — fixed as… See the full description on the dataset page: https://huggingface.co/datasets/AmanPriyanshu/reasoning-sft-interstellarninja-json-mode-reasoning-160K.herman-json-mode
Herman: Indonesian Single-Turn JSON Mode
Herman is an Indonesian language dataset specifically designed
for training LLMs using a single-turn JSON mode. This dataset
is used in Supervised Fine-Tuning (SFT) to improve JSON parsing
capabilities in LLMs. Herman was obtained from Hermes and translated
into Indonesian for the purpose of training Indonesian language models.
Code used for constructing Herman can be found here.
Schema Format
The desired JSON schema can… See the full description on the dataset page: https://huggingface.co/datasets/SulthanAbiyyu/herman-json-mode.nepali-json-mode-singleturnjson-mode-dpo-promptsnepali-json-mode-singleturn
This dataset is a remastered version of this dataset prepared using Adaption's Adaptive Data platform.
adaption-json_schema_instruction_data
This dataset contains instruction-following conversations where models are tasked with generating valid JSON objects based on provided schemas and user prompts. The samples cover diverse domains including healthcare, project management, engineering, and environmental science, requiring strict adherence to defined data structures. Each… See the full description on the dataset page: https://huggingface.co/datasets/himalaya-ai/nepali-json-mode-singleturn.json-mode-evalaugmented-json-mode-eval
JSON Mode Evaluation Dataset (Augmented)
Dataset Description
This is an augmented version of the NousResearch/json-mode-eval dataset.
The original dataset contains examples for evaluating models' ability to follow JSON schema instructions, and this augmented version includes additional variations with different formatting of the schema prompt.
This dataset has been filtered to remove samples containing JSON schema features that are not supported by xgrammar and llguidance… See the full description on the dataset page: https://huggingface.co/datasets/squeezebits/augmented-json-mode-eval.json-mode-eval-rgxmodel_jsonjson-mode-agenticjson-mode-singleturnjson-mode-newgenallyson-json-mode-sharegpt-v0.1model_card.jsonupdatedclasification_model_v0_4_5rc_predict_yes_no_dataset_for_json_3110350_staediontawfiq-json-ai-modelwsinsight-model-zoo-json
