datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
foundation-model-results-simulation-consortiauser-simulation-results
User simulation leaderboard: results
One file per system per release, under <org>/<system>/results_<timestamp>.json.
results holds the score the leaderboard displays (Recall@10 per benchmark). record holds
everything that makes the row citable and is not shown in the grid: the person-clustered
bootstrap interval, n, NDCG@10, the popularity and stranger controls, the evaluation cell, and
the run that produced it (experiment id, commit, date, seed).
These files are generated… See the full description on the dataset page: https://huggingface.co/datasets/jean-technologies/user-simulation-results.heisenberg_simulation_resultsPID_Simulation_Results
