bpmn
Datasets
All datasets matching “bpmn”BPMN-Redrawer-DatasetThis repository contains dataset used in the BPMN-Redrawer project.
The original 663 BPMN models are available on the RePROSitory platform: https://pros.unicam.it:4200/guest/collection/bpmn_redrawer
Additional 165 BPMN models have been designed, and are uploaded here in the BPMN Models folder. Such models have been designed to augment the amount of instances of elements that are rare in the RePROSitory models.
license: cc-by-nc-sa-4.0
BPMN-IT-DatasetSignavio_text_bpmn
Signavio Text BPMN Dataset
This dataset is presented in the paper Improving LLM-Generated Process Model Quality Through Reinforcement Learning: The Role of Reward Function Design.
The official code repository can be found here: RL_for_process_modeling.
Dataset Description
The dataset contains textual process descriptions paired with corresponding BPMN (Business Process Model and Notation) process models, used for training and evaluating LLMs on structured process… See the full description on the dataset page: https://huggingface.co/datasets/chlauer/Signavio_text_bpmn.BPMN-VLM
🏗️ BPMN Diagram → BPMN XML Paired Dataset
Structured Extraction from Business Process Diagrams using Vision-Language Models
This dataset contains Business Process Model and Notation (BPMN) diagrams paired with their corresponding .bpmn XML ground truth files.The dataset is designed for training, evaluation, and benchmarking multimodal models that perform structured extraction from diagrams, including OCR-enhanced pipelines and vision-language models (VLMs).… See the full description on the dataset page: https://huggingface.co/datasets/pritamdeka/BPMN-VLM.bpmn-assistant-evalbpmn-image-xml-pairs
English-Translated BPMN Image/XML Pairs
This dataset contains 3,478 paired BPMN 2.0 XML files and PNG renderings used
to study structured extraction from business-process diagrams. Visible labels
in the released pairs are English translations. The corresponding preparation,
training, prediction, evaluation, and statistical-analysis code is available at:
https://github.com/pritamdeka/bpmn-structured-extraction
Frozen splits
train: 2,782 rows
validation: 347 rows… See the full description on the dataset page: https://huggingface.co/datasets/pritamdeka/bpmn-image-xml-pairs.
