hierarchical-classification
wos_hierarchical_multi_label_text_classificationIntroduced by du Toit and Dunaiski (2024) Introducing Three New Benchmark Datasets for Hierarchical Text Classification.
The WOS Hierarchical Text Classification are three dataset variants created from Web of Science (WOS) title and abstract data categorised into a hierarchical, multi-label class structure. The aim of the sampling and filtering methodology used was to create well-balanced class distributions (at chosen hierarchical levels). Furthermore, the WOS_JTF variant was also created… See the full description on the dataset page: https://huggingface.co/datasets/marcelsun/wos_hierarchical_multi_label_text_classification.Hierarchical_Text_Classification_Intent_ClassificationWildChat-Legal-Classification-V3-Hierarchical
WildChat Legal Classification V3 — Hierarchical Training Split
Fixed train/validation dataset for a two-stage legal-needs classifier:
predict whether a conversation seeks legal guidance;
predict the primary legal topic only when guidance is present.
This dataset is derived exclusively from the 1,922-row train split of
AmirMohseni/WildChat-Legal-Classification,
using its v3 taxonomy.
Splits
Split
Rows
Use
train
1,632
Model fitting
validation
290… See the full description on the dataset page: https://huggingface.co/datasets/AmirMohseni/WildChat-Legal-Classification-V3-Hierarchical.
