datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
mimic-medical-imaging-qa
MIMIC Medical Imaging QA Dataset
5,207 Bloom's-taxonomy-stratified question--answer pairs derived from 23 medical imaging lectures (RPI BMED 2300). The dataset supports the paper "MIMIC: A Course-Derivation Pipeline and Benchmark for Slide-Anchored Tutoring with a Domain-Adapted Large Language Model" and was used to fine-tune MIMIC-LM, a domain-adapted Llama-3.1-8B-Instruct model for grounded medical imaging instruction.
License
The benchmark annotations, dataset… See the full description on the dataset page: https://huggingface.co/datasets/zabir1996/mimic-medical-imaging-qa.mimicgen-square-d0-light-seed42-1000-opaque-rerender-lerobotmimicgen-square-d0-bg25-seed42-1000-opaque-rerender-lerobotmimicgen-square-d0-lerobot-v3
1.62 GB MimicGen HDF5 → multimodal LeRobot v3
Before → after: HDF5 demonstrations with two RGB streams become a validated, multimodal LeRobot v3.0 subset. Convert HDF5 free →
Community conversion produced by ViaCatalyst BYOD. This repository is not an official upstream release and is not affiliated with the MimicGen authors.
This is a compact, provenance-complete conversion of the first 10 episodes from the pinned MimicGen Square D0 core HDF5 file. It provides a… See the full description on the dataset page: https://huggingface.co/datasets/ViaCatalyst/mimicgen-square-d0-lerobot-v3.DoppelReflEx__MN-12B-Mimicore-GreenSnake-details
Dataset Card for Evaluation run of DoppelReflEx/MN-12B-Mimicore-GreenSnake
Dataset automatically created during the evaluation run of model DoppelReflEx/MN-12B-Mimicore-GreenSnake
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/DoppelReflEx__MN-12B-Mimicore-GreenSnake-details.DoppelReflEx__MN-12B-Mimicore-Nocturne-details
Dataset Card for Evaluation run of DoppelReflEx/MN-12B-Mimicore-Nocturne
Dataset automatically created during the evaluation run of model DoppelReflEx/MN-12B-Mimicore-Nocturne
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/DoppelReflEx__MN-12B-Mimicore-Nocturne-details.DoppelReflEx__MN-12B-Mimicore-Orochi-details
Dataset Card for Evaluation run of DoppelReflEx/MN-12B-Mimicore-Orochi
Dataset automatically created during the evaluation run of model DoppelReflEx/MN-12B-Mimicore-Orochi
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/DoppelReflEx__MN-12B-Mimicore-Orochi-details.DoppelReflEx__MN-12B-Mimicore-Orochi-v3-Experiment-details
Dataset Card for Evaluation run of DoppelReflEx/MN-12B-Mimicore-Orochi-v3-Experiment
Dataset automatically created during the evaluation run of model DoppelReflEx/MN-12B-Mimicore-Orochi-v3-Experiment
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/DoppelReflEx__MN-12B-Mimicore-Orochi-v3-Experiment-details.DoppelReflEx__MN-12B-Mimicore-WhiteSnake-v2-Experiment-4-details
Dataset Card for Evaluation run of DoppelReflEx/MN-12B-Mimicore-WhiteSnake-v2-Experiment-4
Dataset automatically created during the evaluation run of model DoppelReflEx/MN-12B-Mimicore-WhiteSnake-v2-Experiment-4
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/DoppelReflEx__MN-12B-Mimicore-WhiteSnake-v2-Experiment-4-details.DoppelReflEx__MN-12B-Mimicore-WhiteSnake-details
Dataset Card for Evaluation run of DoppelReflEx/MN-12B-Mimicore-WhiteSnake
Dataset automatically created during the evaluation run of model DoppelReflEx/MN-12B-Mimicore-WhiteSnake
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/DoppelReflEx__MN-12B-Mimicore-WhiteSnake-details.DoppelReflEx__MN-12B-Mimicore-Orochi-v4-Experiment-details
Dataset Card for Evaluation run of DoppelReflEx/MN-12B-Mimicore-Orochi-v4-Experiment
Dataset automatically created during the evaluation run of model DoppelReflEx/MN-12B-Mimicore-Orochi-v4-Experiment
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/DoppelReflEx__MN-12B-Mimicore-Orochi-v4-Experiment-details.DoppelReflEx__MN-12B-Mimicore-WhiteSnake-v2-Experiment-3-details
Dataset Card for Evaluation run of DoppelReflEx/MN-12B-Mimicore-WhiteSnake-v2-Experiment-3
Dataset automatically created during the evaluation run of model DoppelReflEx/MN-12B-Mimicore-WhiteSnake-v2-Experiment-3
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/DoppelReflEx__MN-12B-Mimicore-WhiteSnake-v2-Experiment-3-details.DoppelReflEx__MN-12B-Mimicore-Orochi-v2-Experiment-details
Dataset Card for Evaluation run of DoppelReflEx/MN-12B-Mimicore-Orochi-v2-Experiment
Dataset automatically created during the evaluation run of model DoppelReflEx/MN-12B-Mimicore-Orochi-v2-Experiment
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/DoppelReflEx__MN-12B-Mimicore-Orochi-v2-Experiment-details.DoppelReflEx__MN-12B-Mimicore-WhiteSnake-v2-Experiment-1-details
Dataset Card for Evaluation run of DoppelReflEx/MN-12B-Mimicore-WhiteSnake-v2-Experiment-1
Dataset automatically created during the evaluation run of model DoppelReflEx/MN-12B-Mimicore-WhiteSnake-v2-Experiment-1
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/DoppelReflEx__MN-12B-Mimicore-WhiteSnake-v2-Experiment-1-details.DoppelReflEx__MN-12B-Mimicore-WhiteSnake-v2-Experiment-2-details
Dataset Card for Evaluation run of DoppelReflEx/MN-12B-Mimicore-WhiteSnake-v2-Experiment-2
Dataset automatically created during the evaluation run of model DoppelReflEx/MN-12B-Mimicore-WhiteSnake-v2-Experiment-2
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/DoppelReflEx__MN-12B-Mimicore-WhiteSnake-v2-Experiment-2-details.waveprep-mimic-iiimimic_iv_discharge_dutch_marianmt
Dataset Card for Mimic 4 Dutch Translation Marianmt
This dataset was created by: Laboratory for Computational Physiology, MIT Institute for Medical Engineering and Science,
in collaboration with the Beth Israel Deaconess Medical Center in Boston, Massachusetts.
The original data source: PhysioNet
Data description
Translation of MIMIC 4 using MariaNMT with vanilla settings.
Note: this version contains spurious repetitions that will need to be filtered using e.g. regular… See the full description on the dataset page: https://huggingface.co/datasets/UMCU/mimic_iv_discharge_dutch_marianmt.mimic_iv_radio_dutch_marianmt
Dataset Card for Mimic 4 Dutch Translation Marianmt
This dataset was created by: Laboratory for Computational Physiology, MIT Institute for Medical Engineering and Science,
in collaboration with the Beth Israel Deaconess Medical Center in Boston, Massachusetts.
The original data source: PhysioNet
Data description
Translation of MIMIC 4 using MariaNMT with vanilla settings.
Note: this version contains spurious repetitions that will need to be filtered using e.g. regular… See the full description on the dataset page: https://huggingface.co/datasets/UMCU/mimic_iv_radio_dutch_marianmt.waveprep-mimic
