datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
AceReason-Nemotron-7B_eval_118b
mlfoundations-dev/AceReason-Nemotron-7B_eval_118b
Precomputed model outputs for evaluation.
Evaluation Results
LiveCodeBenchv5_official
Average Accuracy: 43.85% ± 0.24%
Number of Runs: 3
Run
Accuracy
Questions Solved
Total Questions
1
43.37%
121
279
2
44.09%
123
279
3
44.09%
123
279
AceReason-Nemotron-7B_eval_c64a
mlfoundations-dev/AceReason-Nemotron-7B_eval_c64a
Precomputed model outputs for evaluation.
Evaluation Results
LiveCodeBenchv5_v3
Average Accuracy: 41.67% ± 0.45%
Number of Runs: 3
Run
Accuracy
Questions Solved
Total Questions
1
41.04%
110
268
2
41.42%
111
268
3
42.54%
114
268
details_FreedomIntelligence__AceGPT-7B
Dataset Card for Evaluation run of FreedomIntelligence/AceGPT-7B
Dataset automatically created during the evaluation run of model FreedomIntelligence/AceGPT-7B.
The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_FreedomIntelligence__AceGPT-7B.ACE
ACE Dataset
This dataset contains evaluation criteria across four domains: DIY, Food, Shopping, and Gaming.
Configurations
diy: DIY and home improvement criteria
food: Food and recipe criteria
shopping: Shopping and product criteria
gaming: Gaming design criteria
Usage
from datasets import load_dataset
Load a specific domain
diy_data = load_dataset("mercor/ACE", "diy")
food_data = load_dataset("mercor/ACE", "food")
shopping_data =… See the full description on the dataset page: https://huggingface.co/datasets/mercor/ACE.details_FreedomIntelligence__AceGPT-v1.5-13B-Chat
Dataset Card for Evaluation run of FreedomIntelligence/AceGPT-v1.5-13B-Chat
Dataset automatically created during the evaluation run of model FreedomIntelligence/AceGPT-v1.5-13B-Chat.
The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_FreedomIntelligence__AceGPT-v1.5-13B-Chat.polaris-acemath-gemini-rubrics-v2details_FreedomIntelligence__AceGPT-v2-8B-Chat
Dataset Card for Evaluation run of FreedomIntelligence/AceGPT-v2-8B-Chat
Dataset automatically created during the evaluation run of model FreedomIntelligence/AceGPT-v2-8B-Chat.
The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_FreedomIntelligence__AceGPT-v2-8B-Chat.details_FreedomIntelligence__AceGPT-13B-chat
Dataset Card for Evaluation run of FreedomIntelligence/AceGPT-13B-chat
Dataset automatically created during the evaluation run of model FreedomIntelligence/AceGPT-13B-chat.
The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_FreedomIntelligence__AceGPT-13B-chat.aic-indexesdetails_FreedomIntelligence__AceGPT-13B
Dataset Card for Evaluation run of FreedomIntelligence/AceGPT-13B
Dataset automatically created during the evaluation run of model FreedomIntelligence/AceGPT-13B.
The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_FreedomIntelligence__AceGPT-13B.acecoder-r1Pick_up_Acetylcysteine_v2This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"robot_type": "franka",
"total_episodes": 100,
"total_frames": 65409,
"total_tasks": 1,
"chunks_size": 1000,
"data_files_size_in_mb": 100,
"video_files_size_in_mb": 200,
"fps": 30,
"splits": {
"train": "0:100"
},
"data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/KyraLeeChunWei/Pick_up_Acetylcysteine_v2.nvidia__AceMath-7B-RM-details
Dataset Card for Evaluation run of nvidia/AceMath-7B-RM
Dataset automatically created during the evaluation run of model nvidia/AceMath-7B-RM
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/nvidia__AceMath-7B-RM-details.ace_sft_3_9k_sampledinsurance-charge-logsdetails_FreedomIntelligence__AceGPT-7B-chat
Dataset Card for Evaluation run of FreedomIntelligence/AceGPT-7B-chat
Dataset automatically created during the evaluation run of model FreedomIntelligence/AceGPT-7B-chat.
The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_FreedomIntelligence__AceGPT-7B-chat.AceReason-Nemotron-7B_eval_569a
mlfoundations-dev/AceReason-Nemotron-7B_eval_569a
Precomputed model outputs for evaluation.
Evaluation Results
LiveCodeBenchv5
Average Accuracy: 0.00% ± 0.00%
Number of Runs: 3
Run
Accuracy
Questions Solved
Total Questions
1
0.00%
0
369
2
0.00%
0
369
3
0.00%
0
369
details_FreedomIntelligence__AceGPT-v1.5-13B
Dataset Card for Evaluation run of FreedomIntelligence/AceGPT-v1.5-13B
Dataset automatically created during the evaluation run of model FreedomIntelligence/AceGPT-v1.5-13B.
The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_FreedomIntelligence__AceGPT-v1.5-13B.worldenergydata-explorer
World Energy Field Explorer — analysis projection
Tabular projection of the World Energy Field Explorer analysis results, published so the
Hugging Face dataset viewer and the datasets-server API can render visualizations directly.
Live Explorer: https://vamseeachanta.github.io/worldenergydata/field-atlas/
Tables (viewer configs)
Config
Rows
What
fields
10
Lower-Tertiary fields — economics, reservoir, HPHT/decommission risk, concept, landman
wells
56… See the full description on the dataset page: https://huggingface.co/datasets/aceengineer/worldenergydata-explorer.africa-sao-tome-and-principe-sao-tome-and-principe-greenhouse-gas-and-air-pollutant-emi-acef97d1
Sao Tome and Principe: Greenhouse Gas and Air Pollutant Emissions | Africa (Sao Tome and Principe official open data)
15,444 rows - 1 Africa country - 2024-2026 - Repackaged by Electric Sheep Africa
TL;DR
This dataset packages one official CSV resource from Sao Tome and Principe as
ML-ready Parquet. The source file is the provenance boundary; all usable
indicators or tabular columns from the resource stay together in this repo.
About the source… See the full description on the dataset page: https://huggingface.co/datasets/electricsheepafrica/africa-sao-tome-and-principe-sao-tome-and-principe-greenhouse-gas-and-air-pollutant-emi-acef97d1.eval_ragtruth-qa_PERL_gemma-4-E4B-it_S130104_ace3_gens_T0_wfs0_s12345_mt512_sft1c9bb9ACE
ACE: Episodic Memory Dataset (StackOverflow Jan–Jun 2025) (v1.0.0)
StackOverflow-derived events and monthly episodic rollups (Jan–Jun 2025).
Dataset contents
ACE contains two related components:
events: canonical event records (~96K examples) derived from StackOverflow Q&A threads.
episodes: grouped rollups of events for each month, ordered chronologically and packaged in fixed-size windows.
Each event includes a question, an accepted answer (or top-scored substitute)… See the full description on the dataset page: https://huggingface.co/datasets/anon-user-423/ACE.so101_glue_transfer_return_20ep_v1Acetylcysteine_E150This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"robot_type": "franka",
"total_episodes": 150,
"total_frames": 66894,
"total_tasks": 1,
"chunks_size": 1000,
"data_files_size_in_mb": 100,
"video_files_size_in_mb": 200,
"fps": 30,
"splits": {
"train": "0:150"
},
"data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/KyraLeeChunWei/Acetylcysteine_E150.dual-arm-testNikolaSigmoid__AceMath-1.5B-Instruct-dolphin-r1-200-details
Dataset Card for Evaluation run of NikolaSigmoid/AceMath-1.5B-Instruct-dolphin-r1-200
Dataset automatically created during the evaluation run of model NikolaSigmoid/AceMath-1.5B-Instruct-dolphin-r1-200
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/NikolaSigmoid__AceMath-1.5B-Instruct-dolphin-r1-200-details.TIGER-Lab__AceCoder-Qwen2.5-7B-Ins-Rule-details
Dataset Card for Evaluation run of TIGER-Lab/AceCoder-Qwen2.5-7B-Ins-Rule
Dataset automatically created during the evaluation run of model TIGER-Lab/AceCoder-Qwen2.5-7B-Ins-Rule
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/TIGER-Lab__AceCoder-Qwen2.5-7B-Ins-Rule-details.eval_Acetylcysteine_20k_e50This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"robot_type": "franka",
"total_episodes": 59,
"total_frames": 21621,
"total_tasks": 1,
"chunks_size": 1000,
"data_files_size_in_mb": 100,
"video_files_size_in_mb": 200,
"fps": 30,
"splits": {
"train": "0:59"
},
"data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/KyraLeeChunWei/eval_Acetylcysteine_20k_e50.Polaris-AceReason-Math-subsample-v1eval_Acetylcysteine_20k_e100This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"robot_type": "franka",
"total_episodes": 30,
"total_frames": 8018,
"total_tasks": 1,
"chunks_size": 1000,
"data_files_size_in_mb": 100,
"video_files_size_in_mb": 200,
"fps": 30,
"splits": {
"train": "0:30"
},
"data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/KyraLeeChunWei/eval_Acetylcysteine_20k_e100.
