datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
NFA_OCR_reinforcement_learning_format_TEST5reinforcement-learningNFA_OCR_reinforcement_learning_format_TEST6NFA_OCR_reinforcement_learning_format_TEST4motoman-up6-cq-lambda-reinforcement-learning_v1.0
CQ(λ) Bag-Shaking Dataset: Human-in-the-Loop Reinforcement Learning
Dataset Description
This dataset contains synthetic training data comparing standard Q-learning with eligibility traces [Q(λ)] against Cooperative Q-learning [CQ(λ)], a human-in-the-loop reinforcement learning algorithm. The data simulates a robotic "bag-shaking" task where an agent must extract knotted objects from a bag through strategic shaking motions.
Dataset Summary
Task: Bag-shaking… See the full description on the dataset page: https://huggingface.co/datasets/DBbun/motoman-up6-cq-lambda-reinforcement-learning_v1.0.
