datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
X-Atlas-Orion
X-Atlas/Orion
X-Atlas: Orion edition (X-Atlas/Orion) is a Perturb-seq atlas containing two genome-wide Fix-Cryopreserve-ScRNAseq (FiCS) Perturb-seq screens that target all human
protein-coding genes (n = 18,903 genes). The dataset is comprised of eight million HCT116 and HEK293T cells, each deeply sequenced to a median of 16,000 unique molecular
identifiers (UMIs) per cell. The median on-target knockdown efficiency is 75.4% in HCT116 cells and 51.5% in HEK293T cells, with a median… See the full description on the dataset page: https://huggingface.co/datasets/Xaira-Therapeutics/X-Atlas-Orion.truevislies-resultspiezo-embedding-benchmark
Piezometric Embedding Benchmark Dataset
Daily groundwater level time series from ~4200 French monitoring stations,
with ERA5 climate covariates and hydrogeological labels.
Notebooks
Notebook
Description
01_data_exploration.ipynb
Dataset overview, label distributions, geographic maps, time series examples
02_benchmark_analysis.ipynb
Encoder comparison, whitening effect, uni vs multi, ranking
Dataset Description
This dataset supports the… See the full description on the dataset page: https://huggingface.co/datasets/xairon/piezo-embedding-benchmark.xai__grok-3_eval_5ed6
xai__grok-3 Evaluation Results
Precomputed model outputs for evaluation.
Evaluation Results
Summary
Metric
AIME25
LiveCodeBenchv5
AMC23
MATH500
MMLUPro
JEEBench
GPQADiamond
LiveCodeBench
CodeElo
HLE
AIME24
Accuracy
50.0
28.3
90.5
85.0
33.2
88.5
66.5
32.7
29.8
7.3
59.3
AIME25
Average Accuracy: 50.0% ± 1.8%
Number of Runs: 10
Run
Accuracy
Questions Solved
Total Questions
1
50.0%
15
30
2
53.3%
16
30
3
56.7%
17
30
4… See the full description on the dataset page: https://huggingface.co/datasets/mlfoundations-dev/xai__grok-3_eval_5ed6.philosophai-xai-grok-3
Dataset Card for "philosophai-xai-grok-3"
More Information needed
xai_hd_newxai_hdxai_hd_multipxai_epic_baseline_maj_robxai_rob_baseline_llm
