datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
hle_text_only
Humanity's Last Exam - (Text only)
🌐 Website | 📄 Paper | GitHub
Center for AI Safety & Scale AI
Humanity's Last Exam (HLE) is a multi-modal benchmark at the frontier of human knowledge, designed to be the final closed-ended academic benchmark of its kind with broad subject coverage. Humanity's Last Exam consists of 3,000 questions across dozens of subjects, including mathematics, humanities, and the natural sciences. HLE is developed globally by subject-matter experts and… See the full description on the dataset page: https://huggingface.co/datasets/macabdul9/hle_text_only.hle_text_onlyhle_text_onlyVideoMMMU-Res-Text-Onlyimage-text-dataset-subset-300k-captions_onlypreference_data_llama_factory_corrected_format_text_onlyimage-text-dataset-subset-300k-captions_only_with_latentsOlympiadBench_TextOnly_Matholympiadbench_math_textonly_only_thought_text_hf_version_epoch_1_with_prefix_with_exist_split_fixed_best_of_16_imagesdreambench_eval_results_seedx_cot_of_InternVL2_5_78b_mpo_awq_cot_with_only_text_captionmathverse_text_only_blackmathverse_text_onlymbench_eval_results_seed_cot_of_InternVL2_5_78b_mpo_awq_cot_with_only_text_caption_fix_bugdreambench_eval_results_seed_cot_only_text_prompt_1_round_fix_bugtext_onlyclevr1000_hf_image_text_rephrased_onlytext_only_v3_metadataGOT_v1_text_only_v3_de_predictions_2024-10-28-a_train_collect_cot_only_thought_text_hf_version_epoch_1_with_prefix_with_exist_split_fixedtext_only_clean_GTGOT_v1_text_only_v3_de_predictions_test_3GOT_v1_text_only_v3_de_predictionsGOT_v1_text_only_v3_de_predictions_2024-10-29-adreambench_eval_results_seed_cot_of_InternVL2_5_78b_mpo_awq_cot_with_only_text_captiontext_only_v3GOT_v1_text_only_predictionsGOT_v1_text_only_v3_de_predictions_test_2
