datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
enterprise_erp_workflow_reasoning
Enterprise ERP Workflow Reasoning
Author: Venkata Ramachandra Karthik Chundi (venkatakarthikchundi@gmail.com)
A multiple-choice benchmark testing LLM reasoning on enterprise ERP business process workflows, approval hierarchies, document relationships, and process sequencing. Scenarios are drawn from real-world Oracle ERP Cloud deployment contexts.
Dataset
50 multiple-choice questions with 4 answer options each. One correct answer per question.
Modules Covered… See the full description on the dataset page: https://huggingface.co/datasets/karthikchundi/enterprise_erp_workflow_reasoning.zarn-workflow-automation-instruct
Zarn Workflow Automation Instruct
Dataset Description
Natural-language workplace requests paired with plans and JSON tool actions.
Team Attribution
This dataset was created and reviewed by the Zarnite team through internal benchmark design, generation, and quality-control workflows. It should be presented as a Zarnite-authored benchmark starter pack, not as a purely human-collected field corpus.
Ecosystem Need Tier
High Ecosystem Need
Why… See the full description on the dataset page: https://huggingface.co/datasets/zarnite/zarn-workflow-automation-instruct.agentic-workflows-sft-100k
Agentic Workflows SFT 100K
A synthetic supervised fine-tuning dataset of 100,000 high-quality conversations covering AI agent architectures, tool use patterns, multi-agent systems, and agent evaluation. Designed to train AI assistants that can help engineers design, build, and debug production AI agents.
Dataset Description
This dataset covers the full spectrum of agentic AI development across 9 specialized categories. Each record follows the ShareGPT format with… See the full description on the dataset page: https://huggingface.co/datasets/stindardlogic/agentic-workflows-sft-100k.oracle-erp-workflow-eval
Oracle ERP Cloud Workflow & Terminology Eval
A benchmark dataset for evaluating large language models on Oracle ERP Cloud knowledge — including workflows, terminology, document types, approval hierarchies, and process flows across seven core Oracle Cloud modules.
Overview
Property
Value
Samples
30
Modules covered
7
Task type
Short-answer question answering
Eval framework
OpenAI Evals (FuzzyMatch)
Metric
Accuracy (fuzzy string match)… See the full description on the dataset page: https://huggingface.co/datasets/karthikchundi/oracle-erp-workflow-eval.
