datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
react-code-instructions
React Code Instructions
Popular Queries
Number of instructions by Model
Unnested Messages
Instructions Added Per Day
Dataset of Claude Artifact esque React Apps generated by Llama 3.1 70B, Llama 3.1 405B, and Deepseek Chat V3.
Examples
Virtual Fitness Trainer Website
LinkedIn Clone
iPhone Calculator
Chipotle Waitlist
Apple Store
react-shadcn-codex
React Shadcn Codex Dataset
Description
The React Shadcn Codex is a curated collection of over 3,000 React components that utilize shadcn, Framer Motion, and Lucide React. This dataset provides a valuable resource for developers looking to understand and implement modern React UI components with these popular libraries.
Content
The dataset includes:
3,000+ React components using shadcn UI
Components with Framer Motion animations
Usage examples of Lucide React… See the full description on the dataset page: https://huggingface.co/datasets/valentin-marquez/react-shadcn-codex.ICU-REACT
ICU-REACT
ICU-REACT is a clinician-supervised dataset for clinical reasoning and information retrieval in the intensive care unit (ICU), developed for fine-tuning and benchmarking large language models (LLMs).
ICU-REACT was constructed using a clinician-in-the-loop annotation framework designed to capture how clinicians identify relevant patient information and integrate it into diagnostic and treatment decisions. The dataset includes a clinician-refined seed training set, a… See the full description on the dataset page: https://huggingface.co/datasets/iheallab/ICU-REACT.reactadaption-react-screenshot-to-code
This dataset is a remastered version of this dataset prepared using Adaption's Adaptive Data platform.
adaption-react_screenshot_to_code
This dataset contains 1,000 paired examples for training multimodal screenshot-to-code systems, specifically targeting React and TypeScript implementations. Each entry consists of a source webpage screenshot, a detailed visual description, and the corresponding generated React/TSX source code. The samples demonstrate high-fidelity UI… See the full description on the dataset page: https://huggingface.co/datasets/Reubencf/adaption-react-screenshot-to-code.ReactJS_FAQ_DatasetNemotron-Math-Proofs-v1
Nemotron-Math-Proofs-v1
Paper: Nemotron-Math: Efficient Long-Context Distillation of Mathematical Reasoning from Multi-Mode SupervisionCode: https://github.com/NVIDIA/NeMo-SkillsDocumentation: Nemotron-MathProofs-v1 documentation
Dataset Description:
Nemotron-Math-Proofs-v1 is a large-scale mathematical reasoning dataset containing ~580k natural language proof problems, ~550k formalizations into theorem statements in Lean 4, and ~900k model-generated reasoning… See the full description on the dataset page: https://huggingface.co/datasets/ReactorJet/Nemotron-Math-Proofs-v1.dbbench_sft_dataset_react
DBBench SFT Dataset (ReAct Format — AgentBench Compatible)
Overview
Synthetic SFT dataset for DBBench (AgentBench, ICLR 2024).
All tables, data, and queries are independently generated to avoid test data leakage.
Format
ReAct text format matching the AgentBench DBBench evaluation protocol:
[user] System prompt (Action: Operation / Action: Answer instructions)
[agent] Ok.
[user] Question + table name + column headers
[agent] Thinking + Action: Operation +… See the full description on the dataset page: https://huggingface.co/datasets/u-10bei/dbbench_sft_dataset_react.reactionqaproofwriter-datasetreact-hooksdbbench_sft_dataset_react_v2
DBBench SFT Dataset (ReAct Format — AgentBench Compatible)
Overview
Synthetic SFT dataset for DBBench (AgentBench, ICLR 2024).
All tables, data, and queries are independently generated to avoid test data leakage.
Format
ReAct text format matching the AgentBench DBBench evaluation protocol:
[user] System prompt (Action: Operation / Action: Answer instructions)
[agent] Ok.
[user] Question + table name + column headers
[agent] Thinking + Action: Operation +… See the full description on the dataset page: https://huggingface.co/datasets/u-10bei/dbbench_sft_dataset_react_v2.react_native_code_review-reasoning-SFTreact-sharegptdbbench_sft_dataset_react_v4
DBBench SFT Dataset (ReAct Format — AgentBench Compatible)
Overview
Synthetic SFT dataset for DBBench (AgentBench, ICLR 2024).
All tables, data, and queries are independently generated to avoid test data leakage.
Format
ReAct text format matching the AgentBench DBBench evaluation protocol:
[user] System prompt (Action: Operation / Action: Answer instructions)
[agent] Ok.
[user] Question + table name + column headers
[agent] Thinking + Action: Operation +… See the full description on the dataset page: https://huggingface.co/datasets/u-10bei/dbbench_sft_dataset_react_v4.react_code_cotreact-componentsreact_single_componentsMultiTurn-ReAct-QAreact_frontendreact_oss_instructreact_two_componentsproofwriter-deduction-balancedA processed subset of the OWA section of the ProofWriter dataset.
Each train/test split contains 300 entries, each of which has a unique set of theories and a single question for those theories.
Both splits are balanced so that the depth of the proof required to answer the question varies evenly between 0-5 (50 entries each), and the labels are balanced (100 each).
'Unknown' labels have been replaced by 'Uncertain' to match other datasets.
dbbench_sft_dataset_react_v3
DBBench SFT Dataset (ReAct Format — AgentBench Compatible)
Overview
Synthetic SFT dataset for DBBench (AgentBench, ICLR 2024).
All tables, data, and queries are independently generated to avoid test data leakage.
Format
ReAct text format matching the AgentBench DBBench evaluation protocol:
[user] System prompt (Action: Operation / Action: Answer instructions)
[agent] Ok.
[user] Question + table name + column headers
[agent] Thinking + Action: Operation +… See the full description on the dataset page: https://huggingface.co/datasets/u-10bei/dbbench_sft_dataset_react_v3.react_updated_dataset
Dataset Card for Dataset Name
This dataset card aims to be a base template for new datasets. It has been generated using this raw template.
Dataset Details
Dataset Description
Curated by: [More Information Needed]
Funded by [optional]: [More Information Needed]
Shared by [optional]: [More Information Needed]
Language(s) (NLP): [More Information Needed]
License: [More Information Needed]
Dataset Sources [optional]
Repository: [More… See the full description on the dataset page: https://huggingface.co/datasets/PriyanshiD/react_updated_dataset.Corianas__llama-3-reactor-details
Dataset Card for Evaluation run of Corianas/llama-3-reactor
Dataset automatically created during the evaluation run of model Corianas/llama-3-reactor
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Corianas__llama-3-reactor-details.react_finetuningreactjs-interview-questionreactive_overcook_scenariosreact_prompt
