datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
ai-companion-apps-directory
AI Companion Apps Directory (2026)
A maintained dataset of AI companion / AI girlfriend / NSFW AI chat applications with published monthly pricing, free-tier availability, and editorial scores. Compiled from each app's published pricing pages and the research library at AI Companion Desk — scores follow the methodology described at aicompaniondesk.com/methodology.
Last updated: 2026-09-25 · Apps tracked: 21
Files
apps.csv — one row per application: name, monthly… See the full description on the dataset page: https://huggingface.co/datasets/aicompaniondesk/ai-companion-apps-directory.sts-companionhttps://ixa2.si.ehu.eus/stswiki/index.php/STSbenchmark
The companion datasets to the STS Benchmark comprise the rest of the English datasets used in the STS tasks organized by us in the context of SemEval between 2012 and 2017.
Authors collated two datasets, one with pairs of sentences related to machine translation evaluation. Another one with the rest of datasets, which can be used for domain adaptation studies.
@inproceedings{cer-etal-2017-semeval,
title = "{S}em{E}val-2017 Task 1:… See the full description on the dataset page: https://huggingface.co/datasets/tasksource/sts-companion.CompanionSim
CompanionSim
2,240 simulated human–AI conversations depicting socioaffective interaction and annotated by two groups: 628 annotators in the US (CompanionSim-US.csv) and 3,646 annotators from the US, UK, India, and Nigeria (CompanionSim-Multi.csv).
CompanionSim-Validation
CompanionSim-Validation
70 real-world conversations annotated by two groups: 168 annotators in the US (CompanionSim-Validation-US.csv) and 998 annotators from the US, UK, India, and Nigeria (CompanionSim-Validation-Multi.csv).
