datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
superskillret-index-fullcontext
superskillret prebuilt index — full-context
Prebuilt embedding index for the superskillret Claude Code plugin.
Unlike the default index (which embeds only name + description), this build
encodes the full skill body (name + description + body) up to
max_seq_length=32768 tokens. Larger index, much higher recall on
skills whose name/description don't capture every keyword in the body.
Version: 1
Corpus: ThakiCloud/SKILLRET (train+test)
Encoder: ThakiCloud/SkillRet-Embedding-0.6B… See the full description on the dataset page: https://huggingface.co/datasets/youngryankim/superskillret-index-fullcontext.contextual_code_review_full
