jkhe/spacev1b
SPACEV1B: A billion-Scale vector dataset for text descriptors This is a dataset released by Microsoft from SpaceV, Bing web vector search scenario, for large scale vector search related research usage. It consists of more than one billion document vectors and 29K+ query vectors encoded by Microsoft SpaceV Superior model. This model is trained to capture generic intent representation for both documents and queries. The goal is to match the query vector to the closest document… See the full description on the dataset page: https://huggingface.co/datasets/jkhe/spacev1b.
This repository belongs to jkhe on Hugging Face.
CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.
