datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
ViZDoom-Deathmatch-PPO-XLrg
ViZDoom Deathmatch with pretrained PPO agent playing over 15 episodes.
lerobot_new_1This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"fps": 30,
"features": {
"action": {
"dtype": "float32",
"shape": [
10
],
"names": [
"shoulder_pan.pos",
"shoulder_lift.pos",
"elbow_flex.pos",
"wrist_flex.pos",
"wrist_roll.pos",
"gripper.pos"… See the full description on the dataset page: https://huggingface.co/datasets/ppooar/lerobot_new_1.vizdoom-ppo-datasetlerobot_grasplerobot_new_2lerobot_place_2This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"fps": 30,
"features": {
"action": {
"dtype": "float32",
"shape": [
10
],
"names": [
"shoulder_pan.pos",
"shoulder_lift.pos",
"elbow_flex.pos",
"wrist_flex.pos",
"wrist_roll.pos",
"gripper.pos"… See the full description on the dataset page: https://huggingface.co/datasets/ppooar/lerobot_place_2.lerobot_arm_4This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"fps": 30,
"features": {
"action": {
"dtype": "float32",
"shape": [
10
],
"names": [
"shoulder_pan.pos",
"shoulder_lift.pos",
"elbow_flex.pos",
"wrist_flex.pos",
"wrist_roll.pos",
"gripper.pos"… See the full description on the dataset page: https://huggingface.co/datasets/ppooar/lerobot_arm_4.lerobot_place_1This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"fps": 30,
"features": {
"action": {
"dtype": "float32",
"shape": [
10
],
"names": [
"shoulder_pan.pos",
"shoulder_lift.pos",
"elbow_flex.pos",
"wrist_flex.pos",
"wrist_roll.pos",
"gripper.pos"… See the full description on the dataset page: https://huggingface.co/datasets/ppooar/lerobot_place_1.SnakeAI_TF_PPO_V0Action mask has been implemented, the model has been updated to support 'Training Resumption' after system disruption. Utilizing the same training parameters as the "Full Reinforcement learning Agent", this agent prioritizes survival over rewards. It's playtime for 100 games is 6hrs, compared to 2hrs for the FRLA.
This demonstrates the agent is adapting for survival, but not to the desired goal of higher scores/reward. #10000000 training timesteps.
Training Hyperparameters is the same as the… See the full description on the dataset page: https://huggingface.co/datasets/privateboss/SnakeAI_TF_PPO_V0.lerobot_arm_3This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"fps": 30,
"features": {
"action": {
"dtype": "float32",
"shape": [
10
],
"names": [
"shoulder_pan.pos",
"shoulder_lift.pos",
"elbow_flex.pos",
"wrist_flex.pos",
"wrist_roll.pos",
"gripper.pos"… See the full description on the dataset page: https://huggingface.co/datasets/ppooar/lerobot_arm_3.eval_ep1000_seedNone_default_10000_ppo_circle_smallThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "racecar",
"total_episodes": 20,
"total_frames": 5032,
"total_tasks": 1,
"total_videos": 20,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:20"},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/Lyrasilas/eval_ep1000_seedNone_default_10000_ppo_circle_small.ppo-CartPole-v1
Dataset Card for "ppo-CartPole-v1"
More Information needed
weNavigate-PPO_controller_training_episodesppo-seals-CartPole-v0
Dataset Card for "ppo-seals-CartPole-v0"
More Information needed
details_ewqr2130__mistral-inst-ppo
Dataset Card for Evaluation run of ewqr2130/mistral-inst-ppo
Dataset automatically created during the evaluation run of model ewqr2130/mistral-inst-ppo on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_ewqr2130__mistral-inst-ppo.lerobot_arm_2This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"fps": 30,
"features": {
"action": {
"dtype": "float32",
"shape": [
10
],
"names": [
"shoulder_pan.pos",
"shoulder_lift.pos",
"elbow_flex.pos",
"wrist_flex.pos",
"wrist_roll.pos",
"gripper.pos"… See the full description on the dataset page: https://huggingface.co/datasets/ppooar/lerobot_arm_2.details_ppopiolek__tinyllama_eng_shortwiki-lingua-ppolerobot_new_3ppo-Pendulum-v1
Dataset Card for "ppo-Pendulum-v1"
More Information needed
details_Ppoyaa__Lumina-5.5-Instructeval_ep1000_seedNone_circle_big_10000_ppo_circle_bigThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "racecar",
"total_episodes": 20,
"total_frames": 8597,
"total_tasks": 1,
"total_videos": 20,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:20"},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/Lyrasilas/eval_ep1000_seedNone_circle_big_10000_ppo_circle_big.ppo_dataset
Dataset Card for "ppo_dataset"
More Information needed
details_Ppoyaa__LuminRP-7BSnakeAI_TF_PPO_V1The Hebrew word נָחָשׁ (Nāḥāš) is used in the Hebrew Bible to identify the serpent that appears in Genesis 3:1, in the Garden of Eden.
This contains #7000000 training parameters/timestep for Snake_AI game using TensorFlow 2.XX.
Best score and performance comes from data #4800000 dataset for ActorCritic, with an average score of 72 with no action mask/upfront rules. Full reinforcement learning with score/reward as a priority
Agent score can be improved with the combination of more training and… See the full description on the dataset page: https://huggingface.co/datasets/privateboss/SnakeAI_TF_PPO_V1.cpcdata_ppolm-eval-results-Ppoyaa-LexiLumin-7B-private
Dataset Card for Evaluation run of Ppoyaa/LexiLumin-7B
Dataset automatically created during the evaluation run of model Ppoyaa/LexiLumin-7B
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-Ppoyaa-LexiLumin-7B-private.wenavigate-ppo-evaluation-v2pp-ocr-mnn-eval
PP-OCR MNN evaluation dataset (811-cell matrix)
Images (273) + canonical paddle.inference baselines (808 json) + configs
for scoring pp-ocr-mnn outputs. See README.md and
https://github.com/baicai1145/pp-ocr-mnn (tools/score.py).
A single-file snapshot is also included as ppocr-eval-dataset.tar.zst.
lm-eval-results-Ppoyaa-Lumina-3.5-private
Dataset Card for Evaluation run of Ppoyaa/Lumina-3.5
Dataset automatically created during the evaluation run of model Ppoyaa/Lumina-3.5
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-Ppoyaa-Lumina-3.5-private.
