datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
motion-smd-data
Motion-SMD Data
Data release for "Encoder-Free Human Motion Understanding via Structured Motion Descriptions".
🌐 Project page: https://yaozhang182.github.io/motion-smd/
💻 Code: https://github.com/yaozhang182/motion-smd
🤗 LoRA adapters: https://huggingface.co/zyyy12138/motion-smd-lora
📄 Paper (arXiv): https://arxiv.org/abs/2604.21668
What's here
Four subdirectories, each with its own README.md describing files, provenance, and license:
Subdir
Contents
Our… See the full description on the dataset page: https://huggingface.co/datasets/zyyy12138/motion-smd-data.Gitruck-MotionIR
Gitruck MotionIR
Gitruck MotionIR is a Chinese motion-design dataset that aligns project-level
natural-language descriptions, technique-level annotations, temporal evidence,
and a renderable intermediate representation (IR v1). The corpus was normalized
from authorized Alight Motion, After Effects, NodeVideo, and Jianying projects.
Gitruck MotionIR 是一个中文动效设计数据集,将工程级描述、技法级标注、时间证据与可渲染
IR v1 对齐。语料由已获授权的 Alight Motion、After Effects、NodeVideo 与剪映工程归一化而来。
Dataset summary /… See the full description on the dataset page: https://huggingface.co/datasets/Hocassian/Gitruck-MotionIR.fineweb-ultra-mini-pro
Dataset Card for Fineweb Ultra Mini
Fineweb Ultra Mini is a dataset derived from the original Fineweb dataset made by huggingface (see here: https://huggingface.co/datasets/HuggingFaceFW/fineweb).
The dataset focuses on extracting high quality data from the Fineweb dataset, from the 1-0.5% range. If you would like more data, though slightly sacrificing quality check out fineweb ultra mini, which focuses on the 2-3% of high quality data originally found in fineweb.… See the full description on the dataset page: https://huggingface.co/datasets/motionlabs/fineweb-ultra-mini-pro.everyday-user-queries
