datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
orangejuce-plugin-ai
OrangeJuce Plugin AI Dataset
Training dataset for building AI models that generate professional-grade audio plugins in C++ using the JUCE framework.
Dataset Summary
This dataset was built to train a code generation model capable of producing production-ready audio plugins across all major plugin formats (VST2, VST3, AU, AAX). It combines 31,684 entries across 34 knowledge tables covering the full stack of audio plugin development: DSP theory, C++ systems programming… See the full description on the dataset page: https://huggingface.co/datasets/Bassgawd/orangejuce-plugin-ai.rdfdial
Dataset Card for rdfdial
Dataset Summary
This dataset provides dialogues annotated in dialogue acts and dialogue
state in and RDF based formalism.
There is a conversion of sfxdial, dstc2 and multiwoz2.3 datasets
as well as two fully synthetic datasets created from simulated conversations:
camrest-sim and multiwoz-sim.
Original dataset before conversion are available here:
DSTC2: https://github.com/matthen/dstc
Multiwoz 2.3:… See the full description on the dataset page: https://huggingface.co/datasets/Orange/rdfdial.KGConv
KGConv, a Conversational Corpus grounded in Wikidata
Dataset Summary
KGConv is a large corpus of 71k english conversations where each question-answer pair is grounded in a Wikidata fact. The conversations were generated automatically: in particular, questions were created using a collection of 10,355 templates; subsequently, the naturalness of conversations was improved by inserting ellipses and coreference into questions, via both handcrafted rules and a generative… See the full description on the dataset page: https://huggingface.co/datasets/Orange/KGConv.ecml_arena_dataset
ARENA: A Cognitive Multi-Agent Framework for Modeling Conflict-Driven Multi-party Conversation
⚠️ Code under internal review. The generation / simulation code is
currently under internal code review — the GitHub repository is
coming soon. This repository already provides the dataset (a sample
subset) so it can be referenced from the paper.
📄 Paper. ARENA: A Cognitive Multi-Agent Framework for Modeling
Conflict-Driven Multi-party Conversation — ECML-PKDD.
Dataset… See the full description on the dataset page: https://huggingface.co/datasets/Orange/ecml_arena_dataset.dataset_fenics_experiment_v7nectar-conversation
Dataset Card for Dataset Name
berkeley-nest/Nectar dataset reformatted for messages. Assistant response is the rank=1 response in the original dataset.
ProfBench
Dataset Description:
Leaderboard | Blog | Paper | Data | Code | Nemo Evaluator SDK
More than 3000 rubric criteria across 40 human-annotated tasks presenting reports addressing professional tasks across PhD STEM (Chemistry, Physics) and Professional Services (Financial Services, Management Consulting) domains.
This dataset is ready for commercial/non-commercial use.
Dataset Owner(s):
NVIDIA Corporation
Dataset Creation Date:
9/24/2025
License/Terms of… See the full description on the dataset page: https://huggingface.co/datasets/orange99087/ProfBench.persona-belief-probesPersonasForSalesbotPersonas for Salesbot
no-oranges
No-Oranges Dataset
Dataset Description
This is a comprehensive instruction-tuning dataset designed to train language models to avoid generating specific forbidden words while maintaining natural language capabilities. The dataset combines multiple sources of high-quality training data including AI-generated adversarial examples and rule-based prompts.
Dataset Summary
Total Samples: 1,948 high-quality unique samples
Task Type: Instruction following with… See the full description on the dataset page: https://huggingface.co/datasets/pranavkarra/no-oranges.DeepRubric-datasetHumanAgencyBench_Human_Annotations
Human annotations and LLM judge comparative Dataset
Paper: HumanAgencyBench: Scalable Evaluation of Human Agency Support in AI Assistants
Code: https://github.com/BenSturgeon/HumanAgencyBench/
Dataset Description
This dataset contains 60,000 evaluated AI assistant responses across 6 dimensions of behaviour relevant to human agency support, with both model-based and human annotations. Each example includes evaluations from 4 different frontier LLM models. We also provide… See the full description on the dataset page: https://huggingface.co/datasets/Experimental-Orange/HumanAgencyBench_Human_Annotations.gmdorangesum_filtered_new_spacesOrangeSum dataset filtered by using the code by Aumiller et al. (1) available at https://github.com/dennlinger/summaries/tree/main
min_length_summary = 18; min_length_reference = 250; length_metric = "whitespace"
bi-gram_overlap_fraction between summary and original text < 0.65, meaning that all summaries in the dataset are on the abstractive side
Furthermore:
both in articles and in summaries, every point (".") followed by a capital letter was replaced by a point followed by a space and the… See the full description on the dataset page: https://huggingface.co/datasets/giuliadc/orangesum_filtered_new_spaces.DreadPoor__OrangeJ-8B-Model_Stock-details
Dataset Card for Evaluation run of DreadPoor/OrangeJ-8B-Model_Stock
Dataset automatically created during the evaluation run of model DreadPoor/OrangeJ-8B-Model_Stock
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/DreadPoor__OrangeJ-8B-Model_Stock-details.SlimOrca-Convoalpaca-fenics-dataset2oig-chipoig-jokesm-training-data2orangesum_5kOrangeSum dataset filtered by using the code by Aumiller et al. (1) available at https://github.com/dennlinger/summaries/tree/main
min_length_summary = 18; min_length_reference = 250; length_metric = "whitespace"
bi-gram_overlap_fraction between summary and original text < 0.65, meaning that all summaries in the dataset are on the abstractive side
Furthermore:
both in articles and in summaries, every point (".") followed by a capital letter was replaced by a point followed by a space and the… See the full description on the dataset page: https://huggingface.co/datasets/giuliadc/orangesum_5k.orange
Dataset Card for Orange Animal Dataset
Dataset Summary
This dataset contains instructions and outputs describing a fictional animal called 'Orange'. It is designed for fine-tuning language models to understand and generate detailed descriptions of this imaginary creature.
Supported Tasks and Leaderboards
text-generation: This dataset can be used to fine-tune language models to generate creative descriptions and answers related to a fictional animal. It is… See the full description on the dataset page: https://huggingface.co/datasets/Johncmk/orange.orangecube20250809
orangecube20250809
This dataset was generated using a phospho starter pack.
This dataset contains a series of episodes recorded with a robot and multiple cameras. It can be directly used to train a policy using imitation learning. It's compatible with LeRobot and RLDS.
rhysjones__phi-2-orange-v2-details
Dataset Card for Evaluation run of rhysjones/phi-2-orange-v2
Dataset automatically created during the evaluation run of model rhysjones/phi-2-orange-v2
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/rhysjones__phi-2-orange-v2-details.alpaca-fenics-datasetoran-netconf-logs-analysism-training-dataIntent_Decomposition_to_Sub_Intents_for_O_RAN_networks
