CoolFace
Modelpublic

eth-nlped/Qwen2.5-1.5B-pedagogical-rewardmodel

sourceHugging Facecc-by-4.0updated 10mo agoView on Hugging Face
4likes1.3kdownloads
Model Card

The model was trained on paired preferences from the MathDial and MRBench datasets.

To find more information and to cite, see:

bibtex
@article{macina2025mathtutorbench,
      title={MathTutorBench: A Benchmark for Measuring Open-ended\\ Pedagogical Capabilities of LLM Tutors}, 
      author={Jakub Macina, Nico Daheim, Ido Hakimi, Manu Kapur, Iryna Gurevych, Mrinmaya Sachan},
      year={2025},
      eprint={2502.18940},
      archivePrefix={arXiv},
      primaryClass={cs.CL},
      url={https://arxiv.org/abs/2502.18940},
}