AILabDsUnipi/SafeQIL-dataset
Human-Generated Demonstrations for Safe Reinforcement Learning Paper: Learning to maintain safety through expert demonstrations in settings with unknown constraints: A Q-learning perspective Code: AILabDsUnipi/SafeQIL Dataset Description This dataset consists of human-generated demonstrations collected across four challenging constrained environments from the Safety-Gymnasium benchmark (SafetyPointGoal1-v0, SafetyCarPush2-v0, SafetyPointCircle2-v0, and… See the full description on the dataset page: https://huggingface.co/datasets/AILabDsUnipi/SafeQIL-dataset.
This repository belongs to AILabDsUnipi on Hugging Face.
CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.
