perm
ppt-pythia-1b-permuted-seed3408-stage2ppt-pythia-1b-permuted-seed3407-stage2Aegis-AI-Content-Safety-LlamaGuard-Permissive-1.0ppt-pythia-1b-appendix-permuted-seed3407-stage2ppt-pythia-1b-appendix-permuted-seed3408-stage2galton-modernbertic-largegoogle_gemma-4-12B-it-ingest-best-gptq-permutedQwen3-30B-A3B-M-SMoE-ngroups96-no-permutation
Datasets
All datasets matching “perm”arxiv-papers-by-subject
arXiv Papers by Subject
A reorganised version of the nick007x/arxiv-papers dataset, partitioned by subject code, year, and month for efficient selective access.
Dataset Description
This dataset contains metadata for over 2.5 million arXiv papers, organised into a hierarchical directory structure that allows users to download only the specific subjects and time periods they need, rather than the entire dataset.
Motivation
The original… See the full description on the dataset page: https://huggingface.co/datasets/permutans/arxiv-papers-by-subject.wdc-common-crawl-embedded-jsonldPerMo
PerMo
Standalone PerMo release normalized to the repository SMPL-H convention.
Summary
PerMo is released here as a standalone motion dataset with SMPL-H motions,
single-motion captions, and style editing / style transfer instruction pairs.
Motions: 6,610 clips, 924,726 frames, 8.56 hours at 30 FPS.
Train split: 6,543 clips, 915,162 frames, 8.47 hours.
Test split: 67 clips, 9,564 frames, 0.09 hours.
Editing train split: 6,176 Neutral-to-style pairs, 869,194 target… See the full description on the dataset page: https://huggingface.co/datasets/ZeyuLing/PerMo.github-code-permissive-sampleSampling from codeparrot/github-code under more permissive license ['mit', 'apache-2.0', 'bsd-3-clause', 'bsd-2-clause', 'cc0-1.0'].It is intended to be used for training code language classifier.
AIC_final_safe_red_permuted_500AgiBotWorld-Beta_G1_task_532_Filler_permanent_magnet_ingot
agibot_task_532
This dataset converts the AgiBot format uniformly into LeRobot V3.0.
Dataset Statistics
robot_name: G1
end_effector: 夹爪
task: 填料永磁锭
total_episodes: 4367
total_tasks: 1
size: 81G
Dataset Structure
├── data
│ └── chunk-xxx
│ ├── file-xxx.parquet
├── meta
│ ├── episodes
│ │ └── chunk-xxx
│ │ └── file-xxx.parquet
│ ├── info.json
│ ├── stats.json
│ └── tasks.parquet
└── videos
├── observation.images.back_left_fisheye… See the full description on the dataset page: https://huggingface.co/datasets/BAAI-DataCube/AgiBotWorld-Beta_G1_task_532_Filler_permanent_magnet_ingot.
