emma
Datasets
All datasets matching “emma”sutra-w2c-corpus
sutra-w2c-corpus
A weights↔code training corpus for Sutra
weight→code decompilation: generated Sutra programs whose behavior is
carried by matrices, paired with those matrices (the "weights") and the
program's substrate input→output behavior. The long-term goal is a model
that maps weights → code (recovering the program from its learned
parameters).
This dataset is generated by experiments/weight_to_code_corpus.py in the
Sutra repo (where it is pinned as the corpus/ submodule)… See the full description on the dataset page: https://huggingface.co/datasets/EmmaLeonhart/sutra-w2c-corpus.Fruits-30
Fruits30 Dataset
Description:
The Fruits30 dataset is a collection of images featuring 30 different types of fruits. Each image has been preprocessed and standardized to a size of 224x224 pixels, ensuring uniformity in the dataset.
Dataset Composition:
Number of Classes: 30
Image Resolution: 224x224 pixels
Total Images: 826
Classes:
0 : acerolas1 : apples2 : apricots3 : avocados4 : bananas5 : blackberries6 : blueberries7 :… See the full description on the dataset page: https://huggingface.co/datasets/emma7991/Fruits-30.AISTActive-ReconstructionEMMA
Dataset Description
EMMA (Enhanced MultiModal reAsoning) is a benchmark targeting organic multimodal reasoning across mathematics, physics, chemistry, and coding.
EMMA tasks demand advanced cross-modal reasoning that cannot be solved by thinking separately in each modality, offering an enhanced test suite for MLLMs' reasoning capabilities.
EMMA is composed of 2,788 problems, of which 1,796 are newly constructed, across four domains. Within each subject, we further provide… See the full description on the dataset page: https://huggingface.co/datasets/lmms-lab-encoder/EMMA.so101_test_03This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "so101",
"total_episodes": 50,
"total_frames": 14882,
"total_tasks": 1,
"total_videos": 100,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:50"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/emmanuel-v/so101_test_03.
