EVA-UNIT-01/EVA-Qwen2.5-14B-v0.0
EVA Qwen2.5 14B
<p> A RP/storywriting specialist model, full-parameter finetune of Qwen2.5-14B on mixture of synthetic and natural data.<br> It uses Celeste 70B 0.1 data mixture, greatly expanding it to improve versatility, creativity and "flavor" of the resulting model.<br> </p>
<p> <p>Prompt format is ChatML.</p><br> <h3>Recommended sampler values:</h3> <ul> <li>Temperature: 0.7</li> <li>Top-P: 0.8</li> <li>Repetition Penalty: 1.03</li> </ul> <p>Model appears to prefer lower temperatures (at least 0.8 and lower) and absolutely hate Min-P sampler.<p>
<h3>Recommended SillyTavern presets (via CalamitousFelicitousness):</h3>
<p> <br> <h3> Training data: </h3> <ul> <li>Celeste 70B 0.1 data mixture minus Opus Instruct subset. See that model's <a href=https://huggingface.co/nothingiisreal/L3.1-70B-Celeste-V0.1-BF16>card</a> for details.</li> <li>Kalomaze's OpusInstruct25k dataset, filtered for refusals.</li> <li>A subset (1k rows) of ChatGPT-4o-WritingPrompts by Gryphe</li> <li>A subset (2k rows) of Sonnet3.5-Charcards-Roleplay by Gryphe</li> <li>A cleaned subset (~3k rows) of shortstories_synthlabels by Auri</li> <li>Synthstruct and SynthRP datasets by Epiculous</li> </ul> <h3> Hardware used: </h3> <ul><li>4xA6000 for 14 hours.</li></ul><br> </p> Model was trained by Kearm and Auri. <h4>Special thanks:</h4><ul> <li>to Gryphe, Lemmy, Kalomaze, Nopm and Epiculous for the data</li> <li>to Alpindale for helping with FFT config for Qwen2.5</li> <li>and to InfermaticAI's community for their continued support for our endeavors</li></ul>
