AdamLucek/Qwen3-4B-Instruct-2507-PII-RL-pii-masking-eval
Evaluation Results: Qwen3-4B-Instruct-2507-PII-RL on PII Masking This dataset contains evaluation results for the RL-trained model AdamLucek/Qwen3-4B-Instruct-2507-PII-RL on the adamlucek/pii-masking environment from Prime Intellect's Environment Hub. The model was fine-tuned using reinforcement learning to mask personally identifiable information (PII) in text. Evaluation Configuration Environment: adamlucek/pii-masking Model:… See the full description on the dataset page: https://huggingface.co/datasets/AdamLucek/Qwen3-4B-Instruct-2507-PII-RL-pii-masking-eval.
Evaluation Results: Qwen3-4B-Instruct-2507-PII-RL on PII Masking
This dataset contains evaluation results for the RL-trained model AdamLucek/Qwen3-4B-Instruct-2507-PII-RL on the adamlucek/pii-masking environment from Prime Intellect's Environment Hub. The model was fine-tuned using reinforcement learning to mask personally identifiable information (PII) in text.
Evaluation Configuration
- Environment: adamlucek/pii-masking
- Model: AdamLucek/Qwen3-4B-Instruct-2507-PII-RL
- Base Model: Qwen/Qwen3-4B-Instruct-2507
- Examples Evaluated: 50
- Rollouts per Example: 3
- Total Samples: 150
