datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
A_Synchronized_Lower_Limb_AMG_sEMG_and_Mocap
SAME-Limb
Synchronized AMG and EMG Dataset of Lower-limb Muscle Activities in Everyday Training
This publicly released dataset contains time-aligned acceleromyography (AMG),
surface electromyography (EMG), optical motion capture (MoCap), and four knee/ankle
joint-angle signals from 30 subjects. The repository also provides the frozen
5–100-Hz benchmark code, 64 fitted model artifacts, fixed reference
predictions, source tables, and integrity manifests used for… See the full description on the dataset page: https://huggingface.co/datasets/Tdongxu/A_Synchronized_Lower_Limb_AMG_sEMG_and_Mocap.pixel-art-character
Pixel Art Character Dataset
⚠️ CONTENT WARNING: This dataset contains partially NSFW content. Some images may include suggestive themes, violence, or mature content. Viewer discretion advised.
A dataset of 500 pixel art character sprites for training LoRA models.
License
Derived License: Apache 2.0
This dataset is provided under the Apache 2.0 License, inherited from the base models used for generation.
Copyright 2026 Limbicnation
Licensed under the Apache License… See the full description on the dataset page: https://huggingface.co/datasets/Limbicnation/pixel-art-character.MetaLlama_Text_Generation_Promptsthis-that-complex-decisions
this-that-complex-decisions
1,710 decisions where the answer follows from a stated policy applied to a state, and where no
single field of that state gives it away.
1,710 questions 19 decision types 40 domains chance rate 0.258
Each row is a state, a question, a closed set of options, and the index of the one option the
policy selects. The answer is determinate: given the state and the policy there is exactly one
correct choice, and it does not depend on anyone's… See the full description on the dataset page: https://huggingface.co/datasets/limberc/this-that-complex-decisions.this-that-spatial-bench
spatial-decisions
7,305 multiple-choice decision questions over 6,525 distinct simulated states, in
15 families and two environments. Every answer is computed from the simulator, not
annotated by a person and not taken from a model. That is the point of the set: on a question whose
answer is derived from the rules of the environment, a disagreement is a mistake, and there is
nothing to argue about.
The set was built to replace a much narrower public artefact: a recording of 68… See the full description on the dataset page: https://huggingface.co/datasets/limberc/this-that-spatial-bench.limbic-eval-tool-use-mcp
Dataset Summary
The MCP Tool Call Evaluation Test Dataset is a synthetic dataset designed for evaluating and benchmarking language models' ability to correctly execute function calls in the context of Model Context Protocol (MCP) tools. This dataset contains 9,813 test examples that assess a model's proficiency in:
Tool Selection: Choosing the correct function from available tools
Parameter Structure: Providing all required parameters with correct names
Parameter Values: Supplying… See the full description on the dataset page: https://huggingface.co/datasets/quotientai/limbic-eval-tool-use-mcp.sma-upper-limb-kinect
SMA Upper-Limb Kinect Dataset and Reach-Intent Benchmark
This repository contains the four supplementary files from a longitudinal Kinect study and a derived reach-intent benchmark.
Publication files
The repository contains the four supplements listed by PLOS:
pone.0170472.s001.pdf: complete analysis report;
pone.0170472.s002.zip: prototype game, R analysis, and Python extraction source code;
pone.0170472.s003.zip: raw data, logs, extracted features, and… See the full description on the dataset page: https://huggingface.co/datasets/YannisTevissen/sma-upper-limb-kinect.Video-Diffusion-Prompt-Style
Video-Diffusion-Prompt-Style
This dataset was created using the Claude Dataset Skill.
dual-stream-image-prompts
Dual-Stream Image Prompts
Multi-dialect image-prompt SFT dataset for training an LLM to route prompts to the
right diffusion model at inference time. Given a concept and a target_model, the
model learns to emit the correct prompt dialect (FLUX T5-XXL prose, SDXL dual-clip
tokens, a compact caption, or steering modifiers).
Routing lives in the instruction prefix, not in a nested output object — keeping the
LoRA's task simple and maximizing structural diversity for generalization.… See the full description on the dataset page: https://huggingface.co/datasets/Limbicnation/dual-stream-image-prompts.sprite-lora-training-datadeforum-prompt-lora-dataset
De Forum Cinematic Prompt Dataset
A specialized dataset for fine-tuning language models to generate cinematic video diffusion prompts in the style of "The Deforum Art Film".
Description
This dataset contains instruction-response pairs for training models to generate high-quality video diffusion prompts with:
Cinematic language and film terminology
De Forum aesthetic (noir, minimalist, art film style)
Technical parameters (aspect ratio, guidance scale, seeds)
Camera… See the full description on the dataset page: https://huggingface.co/datasets/Limbicnation/deforum-prompt-lora-dataset.Images-Diffusion-Prompt-Style
Image Diffusion Prompt Style
High-quality synthetic prompts for image diffusion models, optimized for Flux, Z Image, and Qwen.
Dataset Structure
Column
Type
Description
style_name
string
Short descriptive name
prompt_text
string
Full prompt with quality tokens
negative_prompt
string
Artifacts to avoid
tags
list
Lowercase keywords
compatible_models
list
Target models
Usage
from datasets import load_dataset
ds =… See the full description on the dataset page: https://huggingface.co/datasets/Limbicnation/Images-Diffusion-Prompt-Style.korean-upper-limbenglish_limbumdeforum-prompt-lora-dataset-v2Human-Decomposition-Limbs-Image-EmbeddingsBispatialstructure-Bigraph-Modellimbum_english_sentencesenglish_limbum_new_testamenthighlevelrandom-bigraph-model-parameter-2405ELinearinterpolation-Bigraph-Model-ParameterConvexshape-Bigraph-Model-Parameteristoriya-40-dney
