CoolFace
Datasetpublic

CCB/cis5300-word-embeddings

Word Embeddings and Semantic Similarity (CIS 5300) Dataset Description This dataset supports learning about word embeddings — dense vector representations that capture word meaning. It includes a standard similarity benchmark, a word sense disambiguation task, and a Shakespeare corpus for training custom embeddings. Configs SimLex-999: Word Similarity Benchmark SimLex-999 (Hill et al., 2015) is a gold-standard benchmark for evaluating… See the full description on the dataset page: https://huggingface.co/datasets/CCB/cis5300-word-embeddings.

sourceHugging Facecc-by-4.0updated 5mo agoView on Hugging Face
0likes361downloads

CCB/cis5300-word-embeddings · main · files are served by the source, never re-hosted here