datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
contashasum crazyflie-scale.glb
4be1c223bb268e35163f98daa675a60a6b2876c2 crazyflie-scale.glb
mv crazyflie-scale.glb 4be1c223bb268e35163f98daa675a60a6b2876c2
const res = await fetch("https://huggingface.co/datasets/rl-tools/conta-data/resolve/main/<sha>");
Qwen2.5-7B-RLT-teacher_data_generation
Qwen2.5-7B-RLT Teacher Model Explanations
Dataset Description
This dataset contains data generated by the Team-Promptia/Qwen2.5-7B-RLT-teacher model. The data was created by providing the model with questions and answers from several well-known academic datasets and tasking it with generating a detailed explanation for each solution.
The primary purpose of this dataset is to evaluate the ability of the Team-Promptia/Qwen2.5-7B-RLT-teacher model to generate high-quality… See the full description on the dataset page: https://huggingface.co/datasets/Team-Promptia/Qwen2.5-7B-RLT-teacher_data_generation.RLT-physics_chemistry-expert-17krl_thinkRLT-medicine_biology-expert-17kQwen2.5-7B-RLT-math-expert_data_generationRLT-math-expert-17krl-transitions-json
RL Transitions Dataset
Dataset representing state-action-reward transitions
for reinforcement learning experiments.
Qwen2.5-7B-RLT-physics_chemistry-expert_data_generationRL-Tunix-10000Qwen2.5-7B-RLT-medicine_biology-expert_data_generationR3-RAG-RLTrainingDataRL-TunixRL-Tunix-Full
