datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
performance-dataset-cornercasescornercases_fews_wsd
FEWS and Semcor Dataset for Word Sense Disambiguation (WSD) hanling corner cases which are difficult to disambiguate by GPT 4 Turbo Model.
This repository contains a formatted and cleaned version of the FEWS and Semcor dataset, specifically arranged for model fine-tuning for Word Sense Disambiguation (WSD) tasks.
Dataset Description
The FEWS and Semcor dataset has been preprocessed and formatted to be directly usable for training and fine-tuning language models for word… See the full description on the dataset page: https://huggingface.co/datasets/deshanksuman/cornercases_fews_wsd.
