datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
sherlock
Sherlock
Naturalistic fMRI dataset: 16 subjects watched ~50 minutes of Sherlock across
two scanning runs (Part1, Part2) and then verbally recalled the narrative in
the scanner. TR = 1.5 s.
This repo mirrors the fmriprep-preprocessed dataset originally distributed via
DataLad at https://gin.g-node.org/ljchang/Sherlock. fmriprep version 1.2.6-1.
Layout
derivatives/fmriprep/sub-XX/
anat/ func/ figures/ log/
onsets/
Sherlock_Crop_Onsets.csv… See the full description on the dataset page: https://huggingface.co/datasets/dartbrains/sherlock.MEPC
Multi-level Product Category Recognition Image Dataset
Summary
Wordcloud
Introduce
MEPC - 1000 Dataset:
Classes: 1000
Images: 164,117
Train: 131,293
Val: 32824
MEPC - 10 Dataset:
Classes: 10
Images: 2,192
Train: 1,753
Val: 439
Statistics
Statistics of the number of multi-level categories in the two datasets MEPC-10 and MEPC-1000
Label-only embeddings visualizing label connections… See the full description on the dataset page: https://huggingface.co/datasets/sherlockvn/MEPC.mini-monster_huntersherlock_cleaned
sherlock_cleaned
The sherlock__x family of the ElliotVL supervised-fine-tuning pool, after VLM cleaning.
images
13,546
QA turns
86,223
answers rewritten by the cleaning pass
0
QA created by the cleaning pass (new_qa)
not measured for this family
shards
33
How this was cleaned
A vision-language model read each image together with its QA and judged the item. The pass is
not a filter that only removes rows — it rewrites answers it finds… See the full description on the dataset page: https://huggingface.co/datasets/elliot-mllm/sherlock_cleaned.drone-nav-synthetic-v1sherlock_fmri_datasetsherlock_fmri_datasetDodoPPG
