datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
mhm-backup-evalscope-20260525_143621
中文 | English
📖 中文文档 | 📖 English Documentation
⭐ If you like this project, please click the "Star" button in the upper right corner to support us. Your support is our motivation to move forward!
📝 Introduction
EvalScope is a powerful and easily extensible model evaluation framework created by the ModelScope Community, aiming to provide a one-stop evaluation solution for large model developers.
Whether you want to evaluate the… See the full description on the dataset page: https://huggingface.co/datasets/modrill/mhm-backup-evalscope-20260525_143621.Digikala_Samsung_Mobile_CommentsMH-MN-NIAHarad_hsdbARAD HSBD (Arad Hyperspectral Database) is a hyperspectral dataset provided at the NTIRE 2020 Spectral Reconstruction Challenge - Track 1: Clean and the NTIRE 2020 Spectral Reconstruction Challenge - Track 2: Real World for the hyperspectral image recovery from RGB images.
Related resources:
GitHub link
Terms and Conditions
(copied from the NTIRE 2020 Spectral Reconstruction Challenge - Track 1: Clean)
Challenge on Spectral Reconstruction from RGB Images
These are… See the full description on the dataset page: https://huggingface.co/datasets/mhmdjouni/arad_hsdb.emotion_dataset_rawdetails_h2m__mhm-7b-v1.3
Dataset Card for Evaluation run of h2m/mhm-7b-v1.3
Dataset automatically created during the evaluation run of model h2m/mhm-7b-v1.3 on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_h2m__mhm-7b-v1.3.details_h2m__mhm-7b-v1.3-DPO-1
Dataset Card for Evaluation run of h2m/mhm-7b-v1.3-DPO-1
Dataset automatically created during the evaluation run of model h2m/mhm-7b-v1.3-DPO-1 on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_h2m__mhm-7b-v1.3-DPO-1.details_h2m__mhm-8x7B-FrankenMoE-v1.0
Dataset Card for Evaluation run of h2m/mhm-8x7B-FrankenMoE-v1.0
Dataset automatically created during the evaluation run of model h2m/mhm-8x7B-FrankenMoE-v1.0 on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_h2m__mhm-8x7B-FrankenMoE-v1.0.kidney-ct-classificationautotrain-data-test-translation-t5-small
AutoTrain Dataset for project: test-translation-t5-small
Dataset Description
This dataset has been automatically processed by AutoTrain for project test-translation-t5-small.
Languages
The BCP-47 code for the dataset's language is unk.
Dataset Structure
Data Instances
A sample from this dataset looks as follows:
[
{
"source": "TrueMood",
"target": "\u062a\u0631\u0648\u0645\u0648\u062f"
},
{
"source": "cleanwax"… See the full description on the dataset page: https://huggingface.co/datasets/mhmtcrkglu/autotrain-data-test-translation-t5-small.mhmodinaturalist
iNaturalist QA Dataset
This repository contains the complete iNaturalist dataset prepared for a question-answering / classification task. It includes every example from the original splits where:
The taxon column was not null.
The image at the provided url was reachable and successfully downloaded.
All images are stored locally, so you can train and evaluate without relying on external URLs.
🔗 Original Dataset Source
The original dataset can be found at:… See the full description on the dataset page: https://huggingface.co/datasets/Mhmd08/inaturalist.autotrain-data-testtranslation
AutoTrain Dataset for project: testtranslation
Dataset Description
This dataset has been automatically processed by AutoTrain for project testtranslation.
Languages
The BCP-47 code for the dataset's language is tr2ar.
Dataset Structure
Data Instances
A sample from this dataset looks as follows:
[
{
"source": "TrueMood",
"target": "\u062a\u0631\u0648\u0645\u0648\u062f"
},
{
"source": "cleanwax",
"target":… See the full description on the dataset page: https://huggingface.co/datasets/mhmtcrkglu/autotrain-data-testtranslation.autotrain-data-qozf-4adi-9pul
Dataset Card for "autotrain-data-qozf-4adi-9pul"
More Information needed
csv_dataset_smalltokenized_cnn_dailymail_bartShadingDataset
Dataset Card for "ShadingDataset"
More Information needed
robot-gripper-manipulationrobot-failure-recovery-logsguanaco-llama2-1k
Dataset Card for "guanaco-llama2-1k"
More Information needed
Medical_NERRadLLamaThinking_Stage2_Final_Evaluationgramedia-datasetsrobot-vision-confidenceLUCID-GLdiffsim-dataset-01movie-recommendation-queriesCoT-300-p14bengali-tokenization-corpus
Bengali Tokenization Corpus (25k Sentences)
Dataset Description
A balanced 25,000-sentence Bengali corpus designed for tokenization benchmarking.
Dataset Summary
This dataset is used in the manuscript:
Quantifying the Tokenization Tax on Bengali: A Multi-Tokenizer Audit Across 25k Sentences
MD. Muktadirul Haque Maruf, Israt Jahan Munny, Moutithi Sarker MouManuscript in preparation
Domains
Academic
News
Literary
Colloquial
Dialectal… See the full description on the dataset page: https://huggingface.co/datasets/M-H-MARUF/bengali-tokenization-corpus.RadLLamaThinking_Stage1_Final_Evaluation
