CoolFace
Modelpublic

derek33125/PA-stage1-DeepSeekQwen7B-300

sourceHugging Faceapache-2.0updated 1y agoView on Hugging Face
0likes11downloads
Model Card

Uploaded model

  • —Developed by: derek33125
  • —License: apache-2.0
  • —Finetuned from model : unsloth/DeepSeek-R1-Distill-Qwen-7B

This qwen2 model was trained 2x faster with Unsloth and Huggingface's TRL library.

Dataset

  • —EmoLLM
  • —SoulChat Multi-Trun Dataset
  • —Train: Val = 240,000: 24,000
  • —Since those datasets are both in simplified Chinese, using this model in other languages may encounter performance drops.

<img src="https://raw.githubusercontent.com/unslothai/unsloth/main/images/unsloth%20made%20with%20love.png" width="200"/>

derek33125/PA-stage1-DeepSeekQwen7B-300 · CoolFace