CoolFace
Datasetpublic

sentence-transformers/askubuntu-questions

Dataset Card for AskUbuntu Questions The AskUbuntu dataset (Lei et al., 2016) is a collection of preprocessed questions taken from AskUbuntu.com 2014 corpus dump. It also comes with 400*20 mannual annotations, marking pairs of questions as "similar" or "non-similar". The dataset is sourced from the original GitHub repository. This dataset contains all questions from the original source, i.e. the text_tokenized.txt.gz data. See also sentence-transformers/askubuntu for the a… See the full description on the dataset page: https://huggingface.co/datasets/sentence-transformers/askubuntu-questions.

sourceHugging Faceupdated 8mo agoView on Hugging Face
1likes33downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
sentence-transformers/askubuntu-questions · CoolFace