datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
construction-safety-object-detection
Dataset Labels
['barricade', 'dumpster', 'excavators', 'gloves', 'hardhat', 'mask', 'no-hardhat', 'no-mask', 'no-safety vest', 'person', 'safety net', 'safety shoes', 'safety vest', 'dump truck', 'mini-van', 'truck', 'wheel loader']
Number of Images
{'train': 307, 'valid': 57, 'test': 34}
How to Use
Install datasets:
pip install datasets
Load the dataset:
from datasets import load_dataset
ds =… See the full description on the dataset page: https://huggingface.co/datasets/keremberke/construction-safety-object-detection.Constructionsafety_QApairsThis dataset was built based on the Construction Safety Guidelines published by KOSHA (Korean Occupational Safety and Health Administration).
This is virtual QA pair dataset generated by GPT-3.5-turbo.
construction-safety-QA-dataset使用的文本数据来源于安全管理网(https://www.safehoo.com/)。内容涵盖建筑安全知识、杂谈、管理制度、操作规程。
采用的dataset生成工具来自于Github:https://github.com/ConardLi/easy-dataset
通过DeepSeek-R1生成了少部分带CoT的问答数据,使用智谱清言的glm-4-flash生成了大部分不带CoT的问答数据。共14401条。
格式为json。
construction-safety-QA-dataset使用的文本数据来源于安全管理网(https://www.safehoo.com/)。内容涵盖建筑安全知识、杂谈、管理制度、操作规程。
采用的dataset生成工具来自于Github:https://github.com/ConardLi/easy-dataset
通过DeepSeek-R1生成了少部分带CoT的问答数据,使用智谱清言的glm-4-flash生成了大部分不带CoT的问答数据。共14401条。
格式为json。
construction-safety-resource-NER-datasetLlava-ConstructionSafety_v6construction-safety-object-detection-paligemmaconstruction-safety-object-detection-paligemmaConstruction_Safety_Risk_Law_QAconstruction_safety_Q_A
