CoolFace
Datasetpublic

dianzinao/deepseek-v3-distill-RWA-1000

A Practical Exploration of Mixed-Style Response LLMs via Few-Shot LoRA Fine-Tuning 1. Research Background and Motivation With the increasingly widespread application of Large Language Models (LLMs) today, how to make model outputs more transparent, natural, and understandable has become an important research direction. Traditional LLMs typically output the final answer directly, making their internal reasoning process a "black box" to the user. To enhance the… See the full description on the dataset page: https://huggingface.co/datasets/dianzinao/deepseek-v3-distill-RWA-1000.

sourceHugging Faceupdated 11mo agoView on Hugging Face
0likes9downloads
3 commits on main
fa91e7c11mo ago

Create README.md

dianzinao
f4539b111mo ago

Upload deepseek-v3-distill-RWA-1000.jsonl

dianzinao
809fbf911mo ago

initial commit

dianzinao